Difference between revisions of "TextAnalysis"
Jump to navigation
Jump to search
Line 32: | Line 32: | ||
* [http://www.laurenceanthony.net/software/antconc/ Laurence Anthony's AntConc], a GUI concordancing and text analysis toolkit | * [http://www.laurenceanthony.net/software/antconc/ Laurence Anthony's AntConc], a GUI concordancing and text analysis toolkit | ||
* [https://sites.google.com/site/casualconc/ CasualConc], a Mac OSX-native toolkit (AntConc's Mac version is ported) | * [https://sites.google.com/site/casualconc/ CasualConc], a Mac OSX-native toolkit (AntConc's Mac version is ported) | ||
− | * David McClure's TextPlot (a | + | * David McClure's TextPlot (a Python package that produces force-directed network of words in a text, based on estimated kernel densities) |
** [http://dclure.org/essays/mental-maps-of-texts/ Blog post explaining concept] | ** [http://dclure.org/essays/mental-maps-of-texts/ Blog post explaining concept] | ||
** [http://dclure.org/tutorials/textplot-refresh/ Blog post to download and set up] | ** [http://dclure.org/tutorials/textplot-refresh/ Blog post to download and set up] |
Revision as of 04:31, 24 February 2016
Resources for Exploring Text Analysis
- "Where to Start," courtesy of Ted Underwood
- Stanford's Introduction, from Tooling Up for Digital Humanities
- For the social scientists
R
- Matthew Jockers, Text Analysis With R for Students of Literature (PDF available for download via the NEU Library)
- Download and install R
- Download and install RStudio
- RSeek (search tool for finding resources on R)
- Simple data types in R
Topic Modeling
- Megan R. Brett's "Basic Introduction" (conceptual)
- Scott Weingart's "Guided Tour" (comprehensive, lots of links)
- Ben Schmidt's article about Latent Dirichlet allocation's (LDA's) limitations
Tools
- MALLET (An open-source, Java-based LDA package)
- GUI Tools that use MALLET
- Stanford Topic Modeling Toolbox
word2vec
- Ben Schmidt's Blog Post on Vector Space Models
- which links to his R wrapper package for word2vec
Miscellaneous text analysis tools
- Voyant Tools
- Laurence Anthony's AntConc, a GUI concordancing and text analysis toolkit
- CasualConc, a Mac OSX-native toolkit (AntConc's Mac version is ported)
- David McClure's TextPlot (a Python package that produces force-directed network of words in a text, based on estimated kernel densities)
- Bookworm
Corpus building
Some places to get text
Plain text
- Project Gutenberg
- Early English Books Online (EEBO) (some texts TEI-encoded)
- Early Caribbean Digital Archive (ECDA)