TAPoR 2.0

Discover Research Tools for Textual Study

  • Browse Tools by Type or Tag
  • Search and Use Tools
  • Read and Create Tool Reviews
  • Contribute and Advertise Tools

TAPoR 2.5 is scheduled for decommissioning.
Please visit TAPoR 3

Popular Tools
User Recommended Tools
Random Tools

Voyant Links

Links finds collocates for words and displays links between them using a force directed graph. It shows term frequencies in proximity to keyword. It is a visualization and shows a web of terms.
Voyant Links
Voyant Links

TextArc

TextArc is a free visualization tool that represents an entire text on a single page. It has elements of an index, concordance and summary all in one place, encouraging the viewer to use its juxtapositions to uncover meaning. The web-based applet is ...
TextArc
TextArc

XTRACT

XTRACT was a tool for lexical collocation developed by Frank Smadja, then of Columbia University. It was designed to use statistical techniques to identify collocations of aribitrary length, and to generate syntactic relationships between words. This ...
XTRACT
XTRACT

Voyant Cirrus

Cirrus is a visualization tool that displays a word cloud relating to the frequency of words appearing in one or more documents. One can click on any word appearing in the cloud to obtain detailed information about its relativity.
Voyant Cirrus
Voyant Cirrus

TUSTEP

TUSTEP (Tubingen System of Text Processing Tools) is a free, open source, widely-used toolbox for text processing. It is aimed at scholarly audiences, can work with texts in both latin and non-latin scripts, and is primarily designed for humanites applications. ...
TUSTEP
TUSTEP

Neatline

Neatline is a free, open-source geotemporal exhibit-builder for creating complex maps and narrative sequences from collections of archives and artifacts. It is first and foremost a suite of plugins for the Omeka framework, but can also be accessed as ...
Neatline
Neatline

DocuBurst

DocuBurst is a free web-based visualization tool for exploring the contents of a text.  Visitors can upload their own text or view those provided by others. DocuBurst presents an interactive chart called a ‘radial sunburst’ diagram which organizes ...
DocuBurst
DocuBurst

Wordle

Wordle is an online toy for generating word clouds using the text you provide.  Text can be submitted by providing an URL or by pasting raw text into an input.  The most frequent words from the text are then used as the source for the resulting visualization, ...
Wordle
Wordle

RATS (Random-Accessible Text Systems)

RATS (Random-Accessible Text Systems) was a PL/I utility for text analysis developed by John B. Smith of Pennsylvania State Univeristy in the early 1970s. It was aimed at researchers with some experience with programming and aimed to streamline the ...
RATS (Random-Accessible Text Systems)
RATS (Random-Accessible Text Systems)

TextArc

TextArc is a free visualization tool that represents an entire text on a single page. It has elements of an index, concordance and summary all in one place, encouraging the viewer to use its juxtapositions to uncover meaning. The web-based applet is ...
TextArc
TextArc

R

R is an open source programing language designed for statistical analysis and parallel computing. R began its life as a research project at the University of Aukland, but has since expanded to become a collaborativly run open source project run by the ...
R
R

Stanford NLP Group: CoreNLP

Stanford CoreNLP is a free Natural Language Processing tool. It processes English language text and provides the base forms of words, parts of speech, indicates whether they are proper names, normalizes dates, times and numeric quantities, and marks ...
Stanford NLP Group: CoreNLP
Stanford NLP Group: CoreNLP

Wordle

Wordle is an online toy for generating word clouds using the text you provide.  Text can be submitted by providing an URL or by pasting raw text into an input.  The most frequent words from the text are then used as the source for the resulting visualization, ...
Wordle
Wordle

BookLamp

BookLamp, part of the Book Genome Project, is a tool and a resource for finding books. It offers an alternative to social recommendation engines reliant on author popularity by treating its books as equal regardless of number of copies sold. BookLamp's ...
BookLamp
BookLamp

RSiena

RSiena is a free, open source social network analysis package for R. It replaces Siena (Simulation Investigation for Empirical Network Analysis), which was a stand-alone program for Windows. Like its predecessor, RSiena is optimized for social network ...
RSiena
RSiena

Paper Machines

Paper Machines is a topic modelling and visualization tool available as a plugin for Zotero. It analyzes Zotero bibliographic collections based on a selection of text mining processes, and enables users to export a variety of visualizations, such as ...
Paper Machines
Paper Machines

TextGrid

TextGrid is a virtual research environment for text-based humanities scholarship. It offers a variety of tools and services for collaboratively creating, analyzing, editing and publishing texts. The TextGrid environment is split into two components, ...
TextGrid
TextGrid

Textometrica

Textometrica is a free, web based text analysis tool offered by HUMlab at Umeå University. Users can upload a plain-text file and examine its word frequencies, see co-occurrences, and generate visualizations and graphs.
Textometrica
Textometrica

List Words - HTML (TAPoRware)

This tool lists words in an HTML document, either uploaded by the user or from a web address. List Words works with relatively small texts of under a megabyte in size. It is part of the TAPoRware collection of tools; there are XML and plain text versions ...
List Words - HTML (TAPoRware)
List Words - HTML (TAPoRware)

List Words - HTML (TAPoRware)

This tool lists words in an HTML document, either uploaded by the user or from a web address. List Words works with relatively small texts of under a megabyte in size. It is part of the TAPoRware collection of tools; there are XML and plain text versions ...
List Words - HTML (TAPoRware)
List Words - HTML (TAPoRware)

INL BlackLab

From the official BlackLab site: "BlackLab is a corpus retrieval engine built on top of Apache Lucene. It allows fast, complex searches with accurate hit highlighting on large, tagged and annotated, bodies of text. It was developed at the Institute ...
INL BlackLab
INL BlackLab
View tools by tag:
1960s 1970s 1980s 1990s 2000s 2010s American Annotation Canadian Comparator English English (language) French (language) German Historic Java Metadata Multilingual Natural language processing Social media
All Tags: