how to find most common words in text by matlab
4 次查看(过去 30 天)
显示 更早的评论
how to tag POS on nouns and verbs in MATLAB, Is it related to regular expressions? I know that regular expressions find a pattern in a text, but I want to find the most common words in texts and tag POS on them( I mean the words are nouns or verbs) and then exchange that POS and make an unfamiliar pair of words. how can I find the most common words in texts by MATLAB?is there any solution for that or I should use another software?
0 个评论
采纳的回答
Christopher Creutzig
2017-11-2
编辑:Christopher Creutzig
2018-11-26
Finding the most common words is easy with Text Analytics Toolbox:
>> sonnets = extractFileText("sonnets.txt");
>> sonnets = erasePunctuation(sonnets);
>> tokenizedSonnets = tokenizedDocument(lower(sonnets));
>> bag = bagOfWords(tokenizedSonnets);
>> topkwords(bag, 10)
ans =
10×2 table
Word Count
______ _____
"and" 490
"the" 436
"to" 409
"my" 371
"of" 370
"i" 344
"in" 321
"that" 320
"thy" 281
"thou" 234
You probably want to remove some words (check out removeWords and stopWords). POS tagging is supported in release R2018b and later, see addPartOfSpeechDetails.
2 个评论
Christopher Creutzig
2018-5-2
What command(s) did you try to read that file? The error message looks like you tried to read it as a table; try using the commands listed above instead.
更多回答(2 个)
Sarah Palfreyman
2018-4-30
编辑:Sarah Palfreyman
2018-4-30
2 个评论
IORUNDU GABRIEL
2018-5-16
Which version of matlab is the least that supports the Text analytic toolbox?
Charmaine Tan
2018-11-26
Hi, after finding my topkwords (most frequent words), how can I plot a histogram of these?
2 个评论
Christopher Creutzig
2018-11-26
txt = extractFileText('sonnets.txt');
td = tokenizedDocument(lower(txt));
td = erasePunctuation(td);
bow = bagOfWords(td);
top = topkwords(bow,20);
bar(top.Count)
set(gca,'XTick',1:size(top,1),'XTickLabel',top.Word,'XTickLabelRotation',45)
(In general, it's a good idea not to ask a new question as an “answer,” but to open a new question instead. It helps other people searching MATLAB Answers in the future.)
另请参阅
类别
在 Help Center 和 File Exchange 中查找有关 Language Support 的更多信息
Community Treasure Hunt
Find the treasures in MATLAB Central and discover how the community can help you!
Start Hunting!