Latent semantic analysis
Technique in natural language processing
Latent semantic analysis (LSA) is a technique in natural language processing, in particular distributional semantics, of analyzing relationships between a set of documents and the terms they contain by producing a set of concepts related to the documents and terms. LSA assumes that words that are close in meaning will occur in similar pieces of text (the distributional hypothesis).
Nº Q1806883 ★
Common · Knowledge
Latent semantic analysis
Technique in natural language processing
Latent semantic analysis (LSA) is a technique in natural language processing, in particular distributional semantics, of analyzing relationships between a set of documents and the terms they contain by producing a set of concepts related to the documents and terms. LSA assumes that words that are close in meaning will occur in similar pieces of text (the distributional hypothesis).
Last price
—
Floor price
—
7-day median
—
30-day sales
0
30-day range
—
In circulation
0
Price history
median
low – high
sales
No sales in this period
Show table
| Date | median | Low | High | sales |
|---|
Sales history
- Last sale
- —
- 30-day average
- —
- 30-day low
- —
- 30-day high
- —
- Sales 7d
- 0
- Sales 30d
- 0
No sales yet.
Anonymous sales: no buyer or seller shown. Figures count player-to-player sales only.
From Wikipedia
Latent semantic analysis (LSA) is a technique in natural language processing, in particular distributional semantics, of analyzing relationships between a set of documents and the terms they contain by producing a set of concepts related to the documents and terms. LSA assumes that words that are close in meaning will occur in similar pieces of text (the distributional hypothesis). A matrix containing word counts per document (rows represent unique words and columns represent each document) is constructed from a large piece of text and a mathematical technique called singular value decomposition (SVD) is used to reduce the number of rows while preserving the similarity structure among columns. Documents are then compared by cosine similarity between any two columns. Values close to 1 represent very similar documents while values close to 0 represent very dissimilar documents. An information retrieval technique using latent semantic structure was patented in 1988 by Scott Deerwester, Susan Dumais, George Furnas, Richard Harshman, Thomas Landauer, Karen Lochbaum and Lynn Streeter. In the context of its application to information retrieval, it is sometimes called latent semantic indexing (LSI).
Text: Wikipédia, CC BY-SA 4.0. ·
Related cards
Latent Dirichlet allocation
Generative statistical model that allows sets of observations to be explained by unobserved groups that explain why some parts of the data are similar
Nº Q269236 ★
Linear discriminant analysis
Method used in statistics, pattern recognition and machine learning
Nº Q1228929 ★★
Semantics
Study of meaning in language
Nº Q39645 ★★★★
Parsing
Process of analyzing a string of symbols, either in natural language, computer languages or data structures, conforming to the rules of a formal grammar
Nº Q194152 ★★
Semantic satiation
Psychological phenomenon in which repetition causes a word to temporarily lose meaning for the listener
Nº Q226007 ★★★
LLM-as-a-Judge
Use of large language models as automated evaluators
Nº Q134468621 ★