Community & Support
Learn
Marketplace
Discussions
Categories
Discussions
General
Platform
Academic
Partner
Regional
User Groups
Documentation
Events
Altair Exchange
Share or Download Projects
Resources
News & Instructions
Programs
YouTube
Employee Resources
This tab can be seen by employees only. Please do not share these resources externally.
Groups
Join a User Group
Support
Altair RISE
A program to recognize and reward our most engaged community members
Nominate Yourself Now!
Home
Discussions
Altair RapidMiner
"Truncated words when using WVTool (Text tool)"
drstevekramer
I have been using the default WVToolConfiguration (with no stemmer requested explicitly). Unfortunately, when I create a WVTWordList from a number of input documents using createWordList(WVTInputList input, WVTConfiguration config, java.util.List initialWords, boolean addWords), quite a few of the words end up being truncated when I iterate through the WVTWordList. For example, "time" -> "tim" and "country" -> "countr".
Is there a reason that this is happening with the sample, standard configuration? I did not explicitly set any stemmer or tokenizer options.
Thanks in advance for your help.
Cheers,
Steve
Find more posts tagged with
AI Studio
Text Mining + NLP
Comments
There are no comments yet
Quick Links
All Categories
Recent Discussions
Activity
My Discussions
Unanswered
日本語 (Japanese)
한국어(Korean)
Groups