LabSE
LabSE is a language model that specializes in providing language-agnostic embeddings for cross-lingual tasks.
379 plain-language definitions from the TiorAI glossary, filed under Natural Language Processing (NLP). Every entry opens with a one-sentence definition, then explains where the term is used.
LabSE is a language model that specializes in providing language-agnostic embeddings for cross-lingual tasks.
LangChain is a framework designed for building applications powered by large language models (LLMs) that can handle complex workflows and logic.
Language detection is the process of identifying the natural language of a given text or speech input.
A language family is a group of languages that share a common ancestral language.
Language Identification is the process of determining the natural language of a given piece of text.
Language modeling is a computational process that predicts the probability of a sequence of words in natural language.
A Large Language Model (LLM) is an AI system trained on massive amounts of text data to understand, generate, and reason with human language.
A laser is a device that emits a concentrated beam of coherent light through the process of optical amplification.
Latent Dirichlet Allocation (LDA) is a generative statistical model used to classify text in a document to a particular topic.
Latent Semantic Analysis is a natural language processing technique used to analyze relationships between a set of documents and the terms they contain by producing a set of concepts related to the documents and…
A lemma is a form of a word that serves as a base for its various inflected forms, often used as a reference point in dictionaries.
Lemmatization is the process of reducing words to their base or root form, known as a lemma.
Levenshtein Distance is a measure of the difference between two strings, quantified by counting the minimum number of single-character edits required to change one string into the other.
Linformer is a transformer model variant designed to improve computational efficiency by approximating self-attention mechanisms.
LlamaIndex is a framework designed for integrating and managing various data sources to create efficient, scalable applications.
A locutionary act is the basic act of producing meaningful language, involving the actual utterance and its literal significance.
Long-range dependency is a characteristic in sequences where elements separated by large distances influence each other.
The Longest Common Subsequence (LCS) is a sequence that appears in the same order within two strings but not necessarily consecutively.
Longformer is a transformer-based model designed to handle long document sequences by using a combination of local and global attention mechanisms.
LoRA is a technique used in machine learning to efficiently fine-tune models by reducing the number of trainable parameters.
Low-Rank Adaptation is a technique used in machine learning to optimize models by reducing their complexity while maintaining performance.
A Low-Resource Language is a language for which there is limited digital data available for computational processing.