Text mining is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Read the full entry →Text, language and qualitative analytics
Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence.
Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence.
Use methods only after defining the estimand, data structure, assumptions, validation plan and decision consequence.
Sentiment analysis classifies or scores expressed polarity in text. It is sensitive to domain language, sarcasm, multilingual context, sampling bias and the difference between online expression and population opinion.
Read the full entry →Sentiment analysis classifies or scores expressed polarity in text. It is sensitive to domain language, sarcasm, multilingual context, sampling bias and the difference between online expression and population opinion.
Read the full entry →Emotion detection is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Read the full entry →Topic modelling identifies recurring patterns of word co-occurrence to organise large text collections. Topics are statistical structures that require human interpretation, validation and sensitivity checks.
Read the full entry →Latent Dirichlet allocation is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Read the full entry →Non-negative matrix factorisation is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Read the full entry →BERTopic is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Read the full entry →Keyword extraction is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Named-entity recognition is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Intent classification is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Semantic similarity is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Embedding analysis is a method within text, language and qualitative analytics. Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence. Its usefulness depends on data structure, assumptions, validation and whether the output answers the intended decision.
Document clustering is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Text classification is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Automated verbatim coding is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Thematic coding is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Co-occurrence analysis is a method within text, language and qualitative analytics. Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence. Its usefulness depends on data structure, assumptions, validation and whether the output answers the intended decision.
Lexical analysis is a method within text, language and qualitative analytics. Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence. Its usefulness depends on data structure, assumptions, validation and whether the output answers the intended decision.
Discourse analysis is a method within text, language and qualitative analytics. Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence. Its usefulness depends on data structure, assumptions, validation and whether the output answers the intended decision.
Conversation analysis is a method within text, language and qualitative analytics. Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence. Its usefulness depends on data structure, assumptions, validation and whether the output answers the intended decision.
Readability analysis is a method within text, language and qualitative analytics. Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence. Its usefulness depends on data structure, assumptions, validation and whether the output answers the intended decision.
Summarisation is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Multilingual text analysis is a method within text, language and qualitative analytics. Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence. Its usefulness depends on data structure, assumptions, validation and whether the output answers the intended decision.
Sarcasm detection is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Stance detection is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Toxicity detection is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Speech-to-text analytics is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Call-centre conversation analytics is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Large-language-model-assisted coding is a method within text, language and qualitative analytics. Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence. Its usefulness depends on data structure, assumptions, validation and whether the output answers the intended decision.
Human validation of AI coding is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Browse linked method entries
Text mining
Text mining is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodSentiment analysis
Sentiment analysis classifies or scores expressed polarity in text. It is sensitive to domain language, sarcasm, multilingual context, sampling bias and the difference between online expression and population opinion.
Open methodAspect-based sentiment analysis
Sentiment analysis classifies or scores expressed polarity in text. It is sensitive to domain language, sarcasm, multilingual context, sampling bias and the difference between online expression and population opinion.
Open methodEmotion detection
Emotion detection is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodTopic modelling
Topic modelling identifies recurring patterns of word co-occurrence to organise large text collections. Topics are statistical structures that require human interpretation, validation and sensitivity checks.
Open methodLatent Dirichlet allocation
Latent Dirichlet allocation is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodNon-negative matrix factorisation
Non-negative matrix factorisation is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodBERTopic
BERTopic is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodKeyword extraction
Keyword extraction is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodNamed-entity recognition
Named-entity recognition is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodIntent classification
Intent classification is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodSemantic similarity
Semantic similarity is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodEmbedding analysis
Embedding analysis is a method within text, language and qualitative analytics. Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence. Its usefulness depends on data structure, assumptions, validation and whether the output answers the intended decision.
Open methodDocument clustering
Document clustering is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodText classification
Text classification is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodAutomated verbatim coding
Automated verbatim coding is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodThematic coding
Thematic coding is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodCo-occurrence analysis
Co-occurrence analysis is a method within text, language and qualitative analytics. Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence. Its usefulness depends on data structure, assumptions, validation and whether the output answers the intended decision.
Open methodLexical analysis
Lexical analysis is a method within text, language and qualitative analytics. Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence. Its usefulness depends on data structure, assumptions, validation and whether the output answers the intended decision.
Open methodDiscourse analysis
Discourse analysis is a method within text, language and qualitative analytics. Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence. Its usefulness depends on data structure, assumptions, validation and whether the output answers the intended decision.
Open methodConversation analysis
Conversation analysis is a method within text, language and qualitative analytics. Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence. Its usefulness depends on data structure, assumptions, validation and whether the output answers the intended decision.
Open methodReadability analysis
Readability analysis is a method within text, language and qualitative analytics. Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence. Its usefulness depends on data structure, assumptions, validation and whether the output answers the intended decision.
Open methodSummarisation
Summarisation is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodMultilingual text analysis
Multilingual text analysis is a method within text, language and qualitative analytics. Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence. Its usefulness depends on data structure, assumptions, validation and whether the output answers the intended decision.
Open methodSarcasm detection
Sarcasm detection is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodStance detection
Stance detection is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodToxicity detection
Toxicity detection is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodSpeech-to-text analytics
Speech-to-text analytics is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodCall-centre conversation analytics
Call-centre conversation analytics is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodLarge-language-model-assisted coding
Large-language-model-assisted coding is a method within text, language and qualitative analytics. Methods that convert unstructured language into codes, themes, entities, sentiment, topics and interpretable evidence. Its usefulness depends on data structure, assumptions, validation and whether the output answers the intended decision.
Open methodHuman validation of AI coding
Human validation of AI coding is a statistical or analytical concept within text, language and qualitative analytics. It should be selected for the data-generating process and decision question rather than because software makes it available.
Open methodNo entries match this search.