TAPoR 2.0

Discover Research Tools for Textual Study

  • Browse Tools by Type or Tag
  • Search and Use Tools
  • Read and Create Tool Reviews
  • Contribute and Advertise Tools

Popular Tools
User Recommended Tools
Random Tools

Voyant Links

Links finds collocates for words and displays links between them using a force directed graph. It shows term frequencies in proximity to keyword. It is a visualization and shows a web of terms.
Voyant Links
Voyant Links

TextArc

TextArc is a free visualization tool that represents an entire text on a single page. It has elements of an index, concordance and summary all in one place, encouraging the viewer to use its juxtapositions to uncover meaning. The web-based applet is ...
TextArc
TextArc

WebLicht

WebLicht is an architecture for creating annotated text corpora. It offers a fully-functional virtual research environment with chains of RESTful web services, each providing a linguistic tool such as format conversion, tokenizing, tagging or parsing. ...
WebLicht
WebLicht

DocuBurst

DocuBurst is a free web-based visualization tool for exploring the contents of a text.  Visitors can upload their own text or view those provided by others. DocuBurst presents an interactive chart called a ‘radial sunburst’ diagram which organizes ...
DocuBurst
DocuBurst

TUSTEP

TUSTEP (Tubingen System of Text Processing Tools) is a free, open source, widely-used toolbox for text processing. It is aimed at scholarly audiences, can work with texts in both latin and non-latin scripts, and is primarily designed for humanites applications. ...
TUSTEP
TUSTEP

OpenRefine

OpenRefine (formerly Google Refine) is a free, open-source tool for working with messy data. It enables users to clean data, transform it between a variety of formats, extend it with web services, and link it to databases. This tool is available for ...
OpenRefine
OpenRefine

TextArc

TextArc is a free visualization tool that represents an entire text on a single page. It has elements of an index, concordance and summary all in one place, encouraging the viewer to use its juxtapositions to uncover meaning. The web-based applet is ...
TextArc
TextArc

R

R is an open source programing language designed for statistical analysis and parallel computing. R began its life as a research project at the University of Aukland, but has since expanded to become a collaborativly run open source project run by the ...
R
R

Tokenize - Plain Text (TAPoR)

This tool splits an HTML document at specified points into 'tokens' - words, lines, sentences, paragraphs or characters. The user can specify characters, patterns, or tags upon which to separate tokens, and choose to have the results listed separator ...
Tokenize - Plain Text (TAPoR)
Tokenize - Plain Text (TAPoR)

Wordle

Wordle is an online toy for generating word clouds using the text you provide.  Text can be submitted by providing an URL or by pasting raw text into an input.  The most frequent words from the text are then used as the source for the resulting visualization, ...
Wordle
Wordle

BookLamp

BookLamp, part of the Book Genome Project, is a tool and a resource for finding books. It offers an alternative to social recommendation engines reliant on author popularity by treating its books as equal regardless of number of copies sold. BookLamp's ...
BookLamp
BookLamp

WordSmith

WordSmith Tools is a commercial integrated suite of programs designed to analyze word behaviour in a text. It can be used to generate a list of all words or word clusters, concord, find keywords and more. This tool is recommended for publishers, language ...
WordSmith
WordSmith

Voyant Cirrus

Cirrus is a visualization tool that displays a word cloud relating to the frequency of words appearing in one or more documents. One can click on any word appearing in the cloud to obtain detailed information about its relativity.
Voyant Cirrus
Voyant Cirrus

Wordle

Wordle is an online toy for generating word clouds using the text you provide.  Text can be submitted by providing an URL or by pasting raw text into an input.  The most frequent words from the text are then used as the source for the resulting visualization, ...
Wordle
Wordle

DocuScope

DocuScope is a text analysis environment first developed in 1998. It contains a suite of interactive visualization tools for corpus-based rhetorical analysis. At present, DocuScope is not available outside the originating research group. However, the ...
DocuScope
DocuScope

Paper Machines

Paper Machines is a topic modelling and visualization tool available as a plugin for Zotero. It analyzes Zotero bibliographic collections based on a selection of text mining processes, and enables users to export a variety of visualizations, such as ...
Paper Machines
Paper Machines

TextGrid

TextGrid is a virtual research environment for text-based humanities scholarship. It offers a variety of tools and services for collaboratively creating, analyzing, editing and publishing texts. The TextGrid environment is split into two components, ...
TextGrid
TextGrid

SOLAR (A Semantically Oriented Lexical ARchive)

SOLAR (A Semantically Oriented Lexical ARchive) was a historically important program originally available only over ARPAnet. It facilitated the semantic and conceptual analysis of English-language texts.
SOLAR (A Semantically Oriented Lexical ARchive)
SOLAR (A Semantically Oriented Lexical ARchive)

List Words - HTML (TAPoRware)

This tool lists words in an HTML document, either uploaded by the user or from a web address. List Words works with relatively small texts of under a megabyte in size. It is part of the TAPoRware collection of tools; there are XML and plain text versions ...
List Words - HTML (TAPoRware)
List Words - HTML (TAPoRware)

List Words - HTML (TAPoRware)

This tool lists words in an HTML document, either uploaded by the user or from a web address. List Words works with relatively small texts of under a megabyte in size. It is part of the TAPoRware collection of tools; there are XML and plain text versions ...
List Words - HTML (TAPoRware)
List Words - HTML (TAPoRware)

Co-Occurrence - HTML (TAPoRware)

This tool looks for two words a certain distance apart from one another in an HTML document, within the user-specified limits of words, sentences or lines. The results can be narrowed to only include words found within certain tags. XML and plain text ...
Co-Occurrence - HTML (TAPoRware)
Co-Occurrence - HTML (TAPoRware)
View tools by tag:
1960s 1970s 1980s 1990s 2000s 2010s American Annotation Canadian Comparator English (language) European French (language) German Historic Java Metadata Multilingual Natural language processing Social media
All Tags: