22 terms

LabSE

LabSE is a language model that specializes in providing language-agnostic embeddings for cross-lingual tasks.

LangChain

LangChain is a framework designed for building applications powered by large language models (LLMs) that can handle complex workflows and logic.

Language Detection

Language detection is the process of identifying the natural language of a given text or speech input.

Language Family

A language family is a group of languages that share a common ancestral language.

Language Identification

Language Identification is the process of determining the natural language of a given piece of text.

Language Modeling

Language modeling is a computational process that predicts the probability of a sequence of words in natural language.

Large Language Model (LLM)

A Large Language Model (LLM) is an AI system trained on massive amounts of text data to understand, generate, and reason with human language.

Laser

A laser is a device that emits a concentrated beam of coherent light through the process of optical amplification.

Latent Dirichlet Allocation

Latent Dirichlet Allocation (LDA) is a generative statistical model used to classify text in a document to a particular topic.

Latent Semantic Analysis

Latent Semantic Analysis is a natural language processing technique used to analyze relationships between a set of documents and the terms they contain by producing a set of concepts related to the documents and…

Lemma

A lemma is a form of a word that serves as a base for its various inflected forms, often used as a reference point in dictionaries.

Lemmatization

Lemmatization is the process of reducing words to their base or root form, known as a lemma.

Levenshtein Distance

Levenshtein Distance is a measure of the difference between two strings, quantified by counting the minimum number of single-character edits required to change one string into the other.

Linformer

Linformer is a transformer model variant designed to improve computational efficiency by approximating self-attention mechanisms.

LlamaIndex

LlamaIndex is a framework designed for integrating and managing various data sources to create efficient, scalable applications.

Locutionary Act

A locutionary act is the basic act of producing meaningful language, involving the actual utterance and its literal significance.

Long-Range Dependency

Long-range dependency is a characteristic in sequences where elements separated by large distances influence each other.

Longest Common Subsequence

The Longest Common Subsequence (LCS) is a sequence that appears in the same order within two strings but not necessarily consecutively.

Longformer

Longformer is a transformer-based model designed to handle long document sequences by using a combination of local and global attention mechanisms.

LoRA

LoRA is a technique used in machine learning to efficiently fine-tune models by reducing the number of trainable parameters.

Low-Rank Adaptation

Low-Rank Adaptation is a technique used in machine learning to optimize models by reducing their complexity while maintaining performance.

Low-Resource Language

A Low-Resource Language is a language for which there is limited digital data available for computational processing.