RDocumentation
Moon
Learn R
Search all packages and functions
⚠️
There's a newer version (0.7-12) of this package.
Take me there.
tm (version 0.7-4)
Text Mining Package
Description
A framework for text mining applications within R.
Copy Link
Copy
Link to current version
Version
Version
0.7-12
0.7-11
0.7-10
0.7-9
0.7-8
0.7-7
0.7-6
0.7-5
0.7-4
0.7-3
0.7-2
0.7-1
0.6-2
0.6-1
0.5-10
0.5-9.1
0.5-8.3
0.5-8.1
0.5-7.1
0.5-6
0.5-5
0.5-4.1
0.5-3
0.5-2
0.5-1
0.4
0.3-4.1
0.3-3
0.3-2
0.3-1
0.2-3.7
0.2-1
0.1-1
Down Chevron
Install
install.packages('tm')
Monthly Downloads
65,954
Version
0.7-4
License
GPL-3
Maintainer
Ingo Feinerer
Last Published
June 19th, 2018
Functions in tm (0.7-4)
Search functions
VectorSource
Vector Source
findMostFreqTerms
Find Most Frequent Terms
Zipf_n_Heaps
Explore Corpus Term Frequency Characteristics
content_transformer
Content Transformers
readReut21578XML
Read In a Reuters-21578 XML Document
getTransformations
Transformations
getTokenizers
Tokenizers
readTagged
Read In a POS-Tagged Word Text Document
removePunctuation
Remove Punctuation Marks from a Text Document
tokenizer
Tokenizers
removeSparseTerms
Remove Sparse Terms from a Term-Document Matrix
readDataframe
Read In a Text Document from a Data Frame
weightBin
Weight Binary
crude
20 Exemplary News Articles from the Reuters-21578 Data Set of Topic crude
readPDF
Read In a PDF Document
findFreqTerms
Find Frequent Terms
findAssocs
Find Associations in a Term-Document Matrix
readPlain
Read In a Text Document
stemDocument
Stem Words
readRCV1
Read In a Reuters Corpus Volume 1 Document
TermDocumentMatrix
Term-Document Matrix
stripWhitespace
Strip Whitespace from a Text Document
tm_filter
Filter and Index Functions on Corpora
tm_map
Transformations on Corpora
termFreq
Term Frequency Vector
stemCompletion
Complete Stems
meta
Metadata Management
removeWords
Remove Words from a Text Document
stopwords
Stopwords
writeCorpus
Write a Corpus to Disk
weightTfIdf
Weight by Term Frequency - Inverse Document Frequency
hpc
Parallelized ‘lapply’
inspect
Inspect Objects
plot
Visualize a Term-Document Matrix
readDOC
Read In a MS Word Document
readXML
Read In an XML Document
removeNumbers
Remove Numbers from a Text Document
tm_term_score
Compute Score for Matching Terms
tm_reduce
Combine Transformations
weightSMART
SMART Weightings
weightTf
Weight by Term Frequency
foreign
Read Document-Term Matrices
DirSource
Directory Source
Docs
Access Document IDs and Terms
ZipSource
ZIP File Source
PCorpus
Permanent Corpora
PlainTextDocument
Plain Text Documents
XMLSource
XML Source
XMLTextDocument
XML Text Documents
Source
Sources
TextDocument
Text Documents
Reader
Readers
SimpleCorpus
Simple Corpora
acq
50 Exemplary News Articles from the Reuters-21578 Data Set of Topic acq
tm_combine
Combine Corpora, Documents, Term-Document Matrices, and Term Frequency Vectors
URISource
Uniform Resource Identifier Source
Corpus
Corpora
DataframeSource
Data Frame Source
VCorpus
Volatile Corpora
WeightFunction
Weighting Function