Stemming

« Back to Glossary Index

Introduction to Stemming
– Stemming is the process of reducing inflected words to their word stem.
– Stemming is used in linguistic morphology and information retrieval.
– Stemming algorithms have been studied since the 1960s.
– Stemming helps search engines treat words with the same stem as synonyms.
– The first published stemmer was written by Julie Beth Lovins in 1968.

Types of Stemming Algorithms
– Simple stemmers use a lookup table to map inflected forms to their stems.
– Lookup approach may use part-of-speech tagging to avoid overstemming.
– Suffix-stripping algorithms find root forms using a set of rules.
– Prefix stripping can also be implemented in some languages.
– Suffix stripping algorithms may differ in results and performance.

Production Technique of Stemming Algorithms
– The lookup table used by a stemmer is produced semi-automatically.
– Inverted algorithms generate inflected forms from a given root form.
– The generation of unlikely forms can be avoided in the production technique.
– The Paice-Husk Stemmer features an externally stored set of stemming rules.
– Chris D Paice developed a direct measurement for comparing stemmers.

Lemmatisation Algorithms
– Lemmatisation involves determining the part of speech of a word.
– Different normalization rules are applied based on the part of speech.
– Correct identification of the lexical category is crucial for accurate lemmatisation.
– Lemmatisation provides more accurate normalization than suffix stripping.
– Lemmatisation algorithms can modify the stem based on additional information.

Stochastic algorithms and other techniques
– Stochastic algorithms use probability to identify the root form of a word.
– Gram analysis uses the n-gram context of a word to determine the correct stem.
– Hybrid approaches combine two or more stemming techniques.
– Affix stemmers deal with both prefixes and suffixes.
– Matching algorithms use a stem database to identify stems.

Note: The references provided in the content are not included in the groups as they are not directly related to the concepts being organized.

Stemming (Wikipedia)

In linguistic morphology and information retrieval, stemming is the process of reducing inflected (or sometimes derived) words to their word stem, base or root form—generally a written word form. The stem need not be identical to the morphological root of the word; it is usually sufficient that related words map to the same stem, even if this stem is not in itself a valid root. Algorithms for stemming have been studied in computer science since the 1960s. Many search engines treat words with the same stem as synonyms as a kind of query expansion, a process called conflation.

Illustration of word stemming that is similar to tree pruning
Illustration of word stemming that is similar to tree pruning

A computer program or subroutine that stems word may be called a stemming program, stemming algorithm, or stemmer.

« Back to Glossary Index

Submit your RFP

We can't wait to read about your project. Use the form below to submit your RFP!

Gabrielle Buff
Gabrielle Buff

Just left us a 5 star review

google

Great customer service and was able to walk us through the various options available to us in a way that made sense. Would definitely recommend!

google

Stoute Web Solutions has been a valuable resource for our business. Their attention to detail, expertise, and willingness to help at a moment's notice make them an essential support system for us.

google

Paul and the team are very professional, courteous, and efficient. They always respond immediately even to my minute concerns. Also, their SEO consultation is superb. These are good people!

google

Paul Stoute & his team are top notch! You will not find a more honest, hard working group whose focus is the success of your business. If you’re ready to work with the best to create the best for your business, go Stoute Web Solutions; you’ll definitely be glad you did!

google

Wonderful people that understand our needs and make it happen!

google

Paul is the absolute best! Always there with solutions in high pressure situations. A steady hand; always there when needed; I would recommend Paul to anyone!

facebook
Vince Fogliani
recommends

The team over at Stoute web solutions set my business up with a fantastic new website, could not be happier

facebook
Steve Sacre
recommends

If You are looking for Website design & creativity look no further. Paul & his team are the epitome of excellence.Don't take my word just refer to my website "stevestours.net"that Stoute Web Solutions created.This should convince anyone that You have finally found Your perfect fit

facebook
Jamie Hill
recommends

Paul and the team at Stoute Web are amazing. They are super fast to answer questions. Super easy to work with, and knows their stuff. 10,000 stars.

facebook

Paul and the team from Stoute Web solutions are awesome to work with. They're super intuitive on what best suits your needs and the end product is even better. We will be using them exclusively for our web design and hosting.

facebook
Dean Eardley
recommends

Beautifully functional websites from professional, knowledgeable team.

google

Along with hosting most of my url's Paul's business has helped me with website development, graphic design and even a really cool back end database app! I highly recommend him as your 360 solution to making your business more visible in today's social media driven marketplace.

yelp

I hate dealing with domain/site hosts. After terrible service for over a decade from Dreamhost, I was desperate to find a new one. I was lucky enough to win...

google

Paul Stoute has been extremely helpful in helping me choose the best package to suite my needs. Any time I had a technical issue he was there to help me through it. Superb customer service at a great value. I would recommend his services to anyone that wants a hassle free and quality experience for their website needs.

google

Paul is the BEST! I am a current customer and happy to say he has never let me down. Always responds quickly and if he cant fix the issue right away, if available, he provides you a temporary work around while researching the correct fix! Thanks for being an honest and great company!!

google

Paul Stoute is absolutely wonderful. Paul always responds to my calls and emails right away. He is truly the backbone of my business. From my fantastic website to popping right up on Google when people search for me and designing my business cards, Paul has been there every step of the way. I would recommend this company to anyone.

yelp

I can't say enough great things about Green Tie Hosting. Paul was wonderful in helping me get my website up and running quickly. I have stayed with Green...