Glossary

Every AI term, in plain words. 805 terms; search, pick a category, or browse A to Z.

TermIn one lineRead
3D Gaussian splatting3D Gaussian splatting rebuilds a scene from ordinary photos as many small, soft, coloured blobs that can be drawn quickly from new viewpoints.Quick
3D generation modelsThree-dimensional generation models create digital shapes, scenes, assets, or representations from text, images, geometry, or other conditioning inputs.Quick
A/B testing LLMsA/B testing LLMs compares two model, prompt, or workflow variants on real traffic to measure differences in quality and product outcomes.Quick
A2A protocolA2A is an open protocol for AI agents: one agent can find another, hand it a task and track the result, even when different companies built them.4 min
AccuracyAccuracy is the share of a model's predictions that were right, the correct answers divided by all answers.4 min
Activation functionAn activation function is the non-linear step a neuron applies to its weighted sum, so a network of layers can learn curved patterns, not only straight lines.4 min
ActivepiecesActivepieces is a workflow automation platform with visual flows, integrations and AI-oriented components that can be self-hosted. in production workflows.Quick
Adam and AdamWAdam is an optimizer that adjusts each weight's step size using running averages of recent gradients; AdamW is a version that handles weight decay separately and is widely used to train large models.Quick
Adobe FireflyAdobe Firefly is Adobe’s family of generative creative tools for images, video, audio and design, integrated across Adobe applications.Quick
Adversarial examplesAdversarial examples are inputs deliberately altered to cause model errors, often through changes that appear small or irrelevant to people.Quick
Agent evaluationAgent evaluation tests an AI agent on set tasks, checking both its final result and the steps it took, over several tries.4 min
Agent memoryAgent memory is how an AI agent saves useful information outside the model and loads the right pieces back into its prompt later.5 min
Agent orchestrationAgent orchestration is how an app with several AI agents decides which agent works, in what order, and who picks what happens next.4 min
Agent planningAgent planning is how an AI agent splits a big task into smaller steps first, then carries out those steps one by one.4 min
Agent skillsAgent skills are reusable packages of instructions, knowledge, or tool procedures that help an AI agent perform a defined class of tasks consistently.Quick
Agent stateAgent state is the running record an AI agent keeps while it works, so a task can pause, resume, or recover from a crash.5 min
Agentic modelsAn agent plans, acts, observes the result, adjusts its approach, and repeats until the task is complete.3 min
Agentic RAGAgentic RAG gives a model control over retrieval, allowing it to plan searches, choose sources, refine queries, and verify evidence before answering.Quick
Agentic workflowsAn agentic workflow runs a task as several language model and tool steps, joined by a path that your code fixes in advance.4 min
AGENTS.mdAGENTS.md is a convention for placing repository-specific instructions where coding agents can discover guidance on structure, commands, style, and verification.Quick
AGIAGI is a proposed category for AI at least as capable as a person across most thinking tasks. Definitions differ; a 2023 Google DeepMind framework marked the level matching many definitions as not yet achieved.4 min
AgnoAgno is open-source software for building AI agents and running them as a live service. You can build single agents, teams of agents and step-by-step workflows.4 min
AI acceleratorsAI accelerators are chips built to speed up the maths inside AI models, which is largely multiply-add sums.5 min
AI agent with toolsAn AI agent with tools lets a model choose defined functions, inspect their results and continue until it reaches a controlled stopping point.Quick
AI agentsAn AI agent is a language model that chooses its own next steps and tools while it works through a task.4 min
AI artAI art is visual, musical, literary, or other creative work generated or transformed with artificial-intelligence tools, often under human direction.Quick
AI engineeringAI engineering is the discipline of building dependable products around models, covering data, prompts, retrieval, evaluation, infrastructure, safety, and user experience.Quick
AI ethicsAI ethics examines moral questions about designing and using AI, including fairness, autonomy, accountability, privacy, labor, power, and potential harm.Quick
AI governanceAI governance comprises the policies, roles, processes, standards, and oversight used to direct and control AI development and deployment.Quick
AI in educationAI in education supports tutoring, feedback, accessibility, assessment, content creation, and administration, while raising questions about accuracy, privacy, and academic integrity.Quick
AI in financeAI in finance supports forecasting, fraud detection, trading, risk assessment, customer service, compliance, and operational automation within regulated financial systems.Quick
AI in healthcareAI in healthcare supports tasks such as diagnosis, imaging, monitoring, administration, research, and treatment planning, generally requiring careful clinical validation and oversight.Quick
AI incidentsAI incidents are events where an AI system causes, contributes to, or nearly causes harm, failure, misuse, or unexpected disruption.Quick
AI modelAn AI model is the part of an AI system that has learned from data. It is a structure plus a set of learned numbers that turns an input into a prediction or new content.4 min
AI safetyAI safety studies and applies methods to prevent, detect, and reduce harms arising from the development, deployment, or misuse of AI systems.Quick
AI searchIn consumer applications, AI search combines information retrieval with language models or other AI methods to interpret questions, rank evidence, and synthesize answers.Quick
AI weather forecastingAI weather forecasting uses learned models to predict atmospheric conditions from historical observations, simulations, and current measurements, complementing numerical forecasting methods.Quick
AI winterAn AI winter is a period when disappointment with artificial intelligence leads to reduced funding, investment, public interest, and research activity.Quick
AiderAider is an open-source AI pair-programming tool that edits code in your local git repository from the terminal.Quick
Alexa+Alexa+ is Amazon’s generative assistant for conversation, smart-home control, entertainment and supported tasks across Echo devices and connected services.Quick
AlexNetAlexNet is an influential convolutional neural network that demonstrated the effectiveness of deep learning on large-scale image classification.Quick
AlgorithmAn algorithm is a precise, step-by-step procedure for getting a result. In AI, learning algorithms are the procedures that turn data into a trained model.4 min
AlignmentAI alignment is the effort to make systems reliably pursue intended goals and behave consistently with relevant human values, instructions, and constraints.Quick
Allen Institute for AIThe Allen Institute for AI is a nonprofit research institute that publishes AI research, models, datasets and open software.Quick
AlphaCodeAlphaCode is Google DeepMind research on generating competitive-programming solutions, using large language models and large-scale candidate filtering to solve coding challenges.Quick
AlphaFoldAlphaFold is Google DeepMind’s system for predicting three-dimensional protein structures from amino-acid sequences, accelerating parts of biological research.Quick
AlphaGoAlphaGo is Google DeepMind’s Go-playing system, combining deep neural networks and tree search to defeat leading professional players.Quick
AlphaZeroAlphaZero is a DeepMind reinforcement-learning system that learned chess, shogi, and Go from self-play without human game examples.Quick
Amazon BedrockAmazon Bedrock is an AWS service that lets developers call AI models from several companies through one set of APIs, with extra tools for search and safety.5 min
Amazon Bedrock AgentCoreA set of AWS services, called Amazon Bedrock AgentCore, for running AI agents in the cloud, with hosting, memory, sign-in, tool access and monitoring.5 min
Amazon NovaAmazon Nova is AWS’s family of foundation models for text, multimodal understanding, image generation, video generation, and related application workflows.Quick
Amazon Q BusinessAmazon Q Business is an enterprise assistant that searches connected organisational data, answers questions and supports authorised workplace tasks.Quick
Amazon Q DeveloperAmazon Q Developer assists with coding, cloud operations and AWS development through IDE, command-line and AWS service integrations.Quick
Amazon SageMaker AISageMaker AI, from AWS, lets you build, train and run machine learning models without managing your own servers.4 min
AMD InstinctAMD Instinct is AMD’s family of data-centre GPU accelerators for high-performance computing and machine-learning training or inference. in practical systems.Quick
Anomaly detectionAnomaly detection finds the few items or events that do not fit the usual pattern in data, such as a fraudulent card payment or an odd burst of network traffic.4 min
Anthropic APIThe Anthropic API gives developers hosted access to Claude models for text, vision, tool use and agentic application workflows.Quick
AnythingLLMAnythingLLM is a self-hosted AI workspace that combines model chat, document retrieval, agents and team-oriented knowledge features. under user control.Quick
Apple Foundation ModelsApple Foundation Models are the on-device and server models underlying Apple Intelligence features and exposed to developers through Apple frameworks.Quick
Apple IntelligenceApple Intelligence is Apple’s system-level collection of generative features for writing, images, notifications, Siri and personal context across supported devices.Quick
Apple M-series chipsApple M-series chips are system-on-chip processors for Macs and selected Apple devices, combining CPU, GPU, media and neural-processing components.Quick
Apple Neural EngineApple Neural Engine is dedicated hardware within Apple chips for accelerating machine-learning operations efficiently on supported devices. in practical systems.Quick
Approximate nearest neighbor searchApproximate nearest neighbor search finds stored vectors that are very close to a query quickly, by checking only a promising part of the collection and accepting a few misses.4 min
ARC-AGIARC-AGI presents demonstration grid pairs to a system, then asks it to construct outputs for new test grids.3 min
Arize PhoenixArize Phoenix is an open-source AI observability and evaluation platform for tracing applications, inspecting retrieval, and analyzing model performance.Quick
Artificial intelligenceAn AI system infers from received inputs how to generate predictions, content, recommendations or decisions for explicit or implicit objectives.5 min
AssemblyAIAssemblyAI provides hosted speech-recognition APIs with transcription and audio-intelligence features such as speaker labels, chapters and content analysis.Quick
AttentionAttention lets a model build an output by assigning weights to available pieces of information and mixing their values.3 min
AutoencodersAn autoencoder learns an encoder that maps an input to a code and a decoder that reconstructs the input from that code.3 min
AutoGenAutoGen is an open-source Microsoft framework for building apps where several AI agents talk to each other, and sometimes to people, to finish a task.4 min
AutoGPTAutoGPT is an experimental agent project that helped popularise autonomous, tool-using task loops built around language models. in production workflows.Quick
AutoMLAutoML automates parts of building a machine learning model, such as choosing the algorithm, the features and the settings.Quick
Autonomous vehiclesAutonomous vehicles use sensors, maps, perception, prediction, planning, and control systems to navigate with reduced or no direct human driving.Quick
Autoregressive modelsAn autoregressive model factorizes a sequence probability into conditional probabilities.3 min
AWS InferentiaAWS Inferentia is a family of computer chips that Amazon designed to run trained AI models cheaply and quickly in its cloud.4 min
AWS TrainiumAWS Trainium is a computer chip Amazon designed for training and running AI models, rented through special Amazon EC2 cloud servers.4 min
AxolotlAxolotl is an open-source tool for fine-tuning language models where you describe the model, data and training method in one configuration file instead of writing code.Quick
Azure Machine LearningAzure Machine Learning is Microsoft’s managed service for training, tracking, deploying and governing machine-learning models and pipelines. in production.Quick
Azure OpenAIAzure OpenAI lets companies use OpenAI's models through Microsoft's Azure cloud, with Azure billing, safety filters and data controls.5 min
BackpropagationBackpropagation computes each parameter's gradient by applying the chain rule from the last layer back to the first.4 min
Base modelsA base language model is a pretrained model before further instruction or conversational fine-tuning.3 min
Batch inferenceBatch inference sends many AI requests as one job that runs in the background, trading an instant reply for a lower price.3 min
Batch normalizationBatch normalization rescales a layer's outputs using statistics from the current batch of examples, which helps many networks train faster and more steadily.Quick
Batch sizeBatch size is how many training examples are grouped before the weights change once.3 min
Bayesian inferenceBayesian inference treats an unknown number as uncertain, starts from a belief about it, and uses Bayes' theorem to update that belief as data arrives.5 min
Beam searchBeam search keeps a fixed number of high-scoring partial sequences, expands them, and retains the best candidates at each step.3 min
Benchmark contaminationBenchmark contamination happens when test questions, or close copies of them, end up in a model's training data, so its score can look better than its real skill.4 min
BenchmarksA benchmark is a fixed set of test tasks with a scoring rule, so different AI models can be compared on the same exam.3 min
BERTBERT is Google’s bidirectional transformer encoder, influential for understanding word meaning in context and fine-tuning on language classification tasks.Quick
bfloat16Bfloat16 is a 16-bit floating-point format with a wide exponent range but reduced precision, commonly used for efficient neural-network computation.Quick
BGEBGE is BAAI’s family of open embedding and reranking models for semantic search, retrieval, and multilingual representation learning.Quick
Bias in AIBias in AI is systematic skew in data, design, or outputs that can produce inaccurate or unfair results across groups or situations.Quick
Bias–variance tradeoffBias is average error across training sets, while variance indicates sensitivity to varying training sets.4 min
BIG-benchBIG-bench is a collaborative collection of diverse language-model tasks designed to probe capabilities, limitations, and behaviors beyond standard benchmarks.Quick
bitsandbytesbitsandbytes is an open-source library for 8-bit and 4-bit quantization, used to load and fine-tune large models with less GPU memory.Quick
BLEU and ROUGEBLEU and ROUGE score generated text by counting how many words and phrases overlap with reference texts; they are common in translation and summarisation.Quick
BM25BM25 is a widely used keyword-ranking algorithm that scores documents using term frequency, term rarity, and document length.Quick
Bolt.newBolt.new is StackBlitz’s browser-based AI app builder, combining code generation with a live development environment, package installation and deployment.Quick
BraintrustBraintrust is an AI evaluation platform for running experiments, managing datasets and prompts, tracing applications, and monitoring production quality.Quick
Browser agentsBrowser agents are AI systems that navigate websites, extract information, fill forms, and complete web tasks through browser controls or automation APIs.Quick
Browser automation agentA browser automation agent observes web pages and performs bounded actions such as clicking, typing and collecting results.Quick
Browser UseBrowser Use is an open-source framework for connecting AI agents to web browsers so they can observe pages and perform actions.Quick
Byte-pair encodingByte-pair encoding builds a subword vocabulary by repeatedly merging the most frequent adjacent symbol pair.3 min
C4C4, or Colossal Clean Crawled Corpus, is a cleaned English web-text dataset derived from Common Crawl for language-model training.Quick
Canva AICanva AI brings generative writing, image, design and editing features into Canva’s visual communication and template-based creation platform.Quick
Catastrophic forgettingCatastrophic forgetting occurs when learning new information substantially degrades a neural network's performance on knowledge or tasks learned earlier.Quick
CatBoostCatBoost is Yandex’s gradient-boosting library with built-in handling for categorical features and tools for tabular prediction problems. in practice.Quick
CerebrasCerebras provides AI computing systems and hosted inference services built around its wafer-scale processors. in production. in production.Quick
Chain of thoughtChain-of-thought prompting asks for or demonstrates intermediate reasoning text before a final answer, but that text is not a guaranteed view of hidden model internals.4 min
Character.AICharacter.AI is a conversational platform centred on user-created AI characters, role-play and interactive storytelling rather than one general-purpose assistant.Quick
Chat templatesA chat template converts a structured list of messages into a formatted sequence.3 min
Chat with PDFA Chat with PDF application extracts and indexes a PDF so users can ask questions and receive answers grounded in its pages.Quick
Chat with your databaseA Chat with your database application translates natural-language requests into controlled queries, then explains results returned from structured data.Quick
ChatbotsChatbots are software interfaces that conduct text or voice conversations using rules, retrieval, generative models, or combinations of these methods.Quick
ChatGPTChatGPT is OpenAI’s conversational product for asking questions, creating things and completing work with models and tools.5 min
ChatGPT WorkChatGPT Work is OpenAI’s agent for longer projects, able to work across connected apps and files and create finished documents, spreadsheets, presentations, reports, and sites.Quick
CheckpointsCheckpoints are saved snapshots of model parameters and training state, enabling recovery, comparison, sharing, or continued training from an earlier point.Quick
Chinchilla scalingChinchilla scaling allocates a fixed training-compute budget by growing model parameters and training tokens in roughly equal proportions.3 min
ChromaChroma is a developer-oriented vector database for storing embeddings and adding semantic retrieval to AI applications. in production workflows.Quick
ChunkingChunking splits a long document into smaller pieces that can be embedded, searched, or fitted into a model prompt.3 min
ClassificationClassification predicts which discrete category or categories an input belongs to, such as identifying an email as spam or legitimate.Quick
Classifier-free guidanceClassifier-free guidance is a setting in diffusion image generators that controls how strongly the result follows the text prompt.Quick
ClaudeClaude is Anthropic’s family of conversational AI models, designed for writing, analysis, coding, multimodal understanding, and tool-assisted workflows.Quick
Claude Agent SDKAnthropic's Claude Agent SDK is a Python and TypeScript library for building agents that run on the same loop, tools and context handling as Claude Code.5 min
Claude appClaude is Anthropic’s AI product for conversation, analysis and completing tasks with files, connected tools and editable outputs.5 min
Claude CodeClaude Code is Anthropic's coding assistant that reads a project, edits files and runs commands for you, asking before risky steps.4 min
Claude Fable 5.1Claude Fable 5.1 is Anthropic’s generally available model for ambitious coding, knowledge work, research, and long-running agents, with configurable effort and broad platform availability.Quick
Claude HaikuClaude Haiku refers to Anthropic’s fastest, smallest model tier, designed for responsive and cost-sensitive workloads such as classification and support automation.Quick
Claude Opus 5.5Claude Opus 5.5 is Anthropic’s 2026 Opus model, positioned for demanding coding, research, and professional work with stronger efficiency than the previous Opus generation.Quick
ClineCline is an open-source AI coding agent that works inside the VS Code editor and can edit files and run commands with your approval.Quick
CLIPCLIP is an OpenAI model that learns shared representations of images and text, enabling zero-shot image classification and cross-modal retrieval.Quick
Closed modelsA fully closed model keeps its weights and code proprietary for internal use.3 min
Cloudflare Workers AICloudflare Workers AI lets developers run supported AI models from Cloudflare’s serverless platform close to applications and users.Quick
ClusteringClustering sorts data that has no labels into groups of items that resemble each other.4 min
COCOCOCO is a large image dataset annotated for object detection, segmentation, captioning, and keypoint estimation in everyday scenes.Quick
Code generationCode generation uses AI to produce software from natural-language instructions, partial code, examples, specifications, or surrounding repository context.Quick
Code interpretersA code interpreter lets an AI model write a small program, run it in a sealed-off sandbox, and use the result in its answer.4 min
Code modelsCode Llama is a family of language models for code with infilling and instruction-following capabilities.3 min
Code review botA code review bot examines proposed changes for defects, style issues or policy violations and leaves review comments for developers.Quick
CodestralCodestral is Mistral AI’s code-focused model family, supporting code generation, completion, fill-in-the-middle editing, and programming assistance across common languages.Quick
Coding assistantA coding assistant uses repository context to explain code, propose changes, generate tests and help debug software. It remains subject to human review.Quick
Cohere EmbedCohere Embed is Cohere’s commercial embedding model family for semantic search, classification, clustering, and retrieval across multiple languages and content types.Quick
Cohere PlatformCohere’s platform provides APIs for enterprise language models, embeddings, reranking and retrieval-oriented generative applications. in production. in production.Quick
Cohere RerankCohere Rerank is a model service that reorders retrieved documents by their relevance to a query, improving search and RAG results.Quick
ColBERTColBERT is a retrieval method that compares a query and a document word by word using token-level embeddings, sitting between fast vector search and slower reranking.Quick
Collaborative filteringCollaborative filtering recommends items by learning from what many people rated, clicked or watched, then predicting the gaps in your own history.4 min
ColPaliColPali is a document retrieval model that searches page images directly, using a vision-language model with late-interaction matching.Quick
ComfyUIComfyUI is a free, open-source app where you build AI image, video and audio generators by wiring boxes called nodes together.4 min
Command ACommand A is Cohere’s enterprise language model for retrieval-augmented generation, multilingual work, tool use, and business-focused agentic applications.Quick
Common CrawlCommon Crawl is a nonprofit repository of web-page snapshots, widely used as raw training data for search, language, and web research.Quick
Computer useComputer use lets an AI interpret a graphical interface and take actions such as clicking, typing, scrolling, or reading screen content.Quick
Computer visionComputer vision enables machines to extract information from images and video, including objects, text, motion, depth, scenes, and spatial relationships.Quick
Computer-use modelsA computer-use model reads a screen and emits mouse or keyboard actions that an application can execute.4 min
Confusion matrixA confusion matrix is a table that counts a classifier's predictions by what it predicted and what was really true, so you see which mistakes it makes.5 min
Constitutional AIConstitutional AI trains and evaluates model behavior against an explicit written set of principles.4 min
Content moderation pipelineA content moderation pipeline classifies submitted material against written policies, applies thresholds and sends ambiguous or high-risk cases to reviewers.Quick
Content provenanceContent provenance records information about media's origin and editing history, helping people and systems assess how a digital artifact was created.Quick
Context compactionContext compaction swaps the older part of a long AI conversation for a short summary, so the chat can keep going inside the model's memory limit.4 min
Context engineeringContext engineering is choosing, trimming and updating the tokens a model sees at each step, so an AI agent works from a small, useful context.4 min
Context windowA context window is all the text a language model can reference while generating a response, including the response itself.3 min
Contextual retrievalContextual retrieval has a model write a short note that places each chunk in its document, then indexes the note with the chunk so search can find it.5 min
Continual learningContinual learning trains a model on a stream of changing data while measuring whether later learning damages or helps earlier tasks.4 min
Continuous batchingContinuous batching dynamically adds and removes inference requests as sequences finish, improving accelerator utilization compared with waiting for an entire fixed batch.Quick
ControlNetControlNet is an open neural-network approach and tooling ecosystem for guiding diffusion image generation with edges, poses, depth and other conditions.Quick
Convolutional neural networksA convolutional neural network uses sparse convolutions that reuse the same weights at multiple locations in an ordered grid.3 min
Copilot+ PCsCopilot+ PCs are Windows computers that meet Microsoft’s hardware requirements for on-device AI features, including a sufficiently capable neural processor.Quick
Copyright and AICopyright and AI concerns how protected works may be used in training and how authorship, ownership, licensing, and infringement apply to generated outputs.Quick
Coqui TTSCoqui TTS is a deep-learning toolkit for training and running text-to-speech, voice-cloning and vocoder models across many languages.Quick
Cosine similarityCosine similarity scores how closely two vectors point the same way, ignoring how long they are. The score runs from −1 to 1.4 min
Cost optimizationAI cost optimization reduces spending through model selection, caching, batching, shorter contexts, efficient retrieval, usage controls, and infrastructure tuning.Quick
CPUsA CPU is the chip that runs a computer's software, carrying out instructions with just a few cores backed by lots of cache memory.5 min
Crawl4AICrawl4AI is a web crawler designed to extract configurable, structured and AI-ready content from sites in self-hosted Python workflows.Quick
CrewAICrewAI is an open-source toolkit, written in Python, for building teams of AI agents, each with a role, that work through a list of tasks together.4 min
Cross-validationCross-validation repeatedly trains and evaluates a method on different data partitions, providing a more reliable performance estimate when data is limited.Quick
Curriculum learningCurriculum learning presents training examples in a deliberate order, often from easier to harder, to improve learning efficiency or final performance.Quick
CursorCursor is a coding tool with an AI agent that can read your project, edit files and run commands for you.4 min
Custom GPTsCustom GPTs are user-configured ChatGPT experiences with tailored instructions, knowledge files, capabilities and optional actions for a particular purpose.Quick
Customer support botA customer support bot answers common questions, retrieves policy information and hands uncertain or sensitive cases to people for review.Quick
DALL·EDALL·E is OpenAI’s family of generative image models, creating original images from natural-language descriptions and supporting image editing in some versions.Quick
Data analysis agentA data analysis agent inspects datasets, writes or runs analysis code and explains results while preserving traceable calculations.Quick
Data and concept driftData drift and concept drift are changes after launch in the data a model sees, or in what the right answer is, that make its predictions worse over time.Quick
Data augmentationData augmentation makes extra training examples by changing existing ones in small ways that keep the answer the same, such as flipping or cropping a photo.4 min
Data curationData curation selects, cleans, balances, documents, and organizes examples so a dataset better supports its intended training or evaluation purpose.Quick
Data labelingData labeling attaches the answer a supervised model should learn to raw examples such as images, text, audio or table rows.3 min
Data leakageData leakage is when a model learns from information it will not have when making real predictions, so its test score looks better than it really is.4 min
Data parallelismData parallelism runs replicas of one model on different slices of a batch, then combines their gradients before a synchronized update.4 min
Data poisoningData poisoning deliberately corrupts training data to damage a model, introduce hidden behavior, or manipulate its future predictions.Quick
Data privacyData privacy is about what an AI company may do with your chats, such as training models, and the settings that let you limit it.5 min
Databricks Mosaic AIDatabricks Mosaic AI brings model development, evaluation, retrieval, agents and governance into the Databricks data and lakehouse environment.Quick
Decision treesDecision trees make predictions through a sequence of feature-based splits, forming interpretable branches that end in class or value estimates.Quick
Decoder-only modelsA decoder-only model predicts each next token from the tokens that came before it.3 min
Deep learningDeep learning is machine learning with neural networks that stack many layers, so each layer can build a more abstract picture of the data than the one before.4 min
Deep research agentsDeep research agents plan multi-step investigations, search and read many sources, synthesize evidence, and produce cited reports with limited supervision.Quick
DeepEvalDeepEval is an open-source framework for testing language-model applications with configurable metrics, datasets, model-based judges, regression checks, and continuous-integration support.Quick
DeepfakesDeepfakes are synthetic or manipulated media that realistically depict people saying or doing things that did not occur.Quick
DeepgramDeepgram provides speech-to-text, text-to-speech and voice-agent APIs for developers building real-time audio applications. for practical use. for practical use.Quick
DeepSeek APIThe DeepSeek API provides hosted access to DeepSeek language and reasoning models through developer-compatible text interfaces. in production.Quick
DeepSeek-R1DeepSeek-R1 is an open-weight reasoning model family from DeepSeek, trained to solve mathematics, coding, and other multi-step problems.Quick
DeepSpeedDeepSpeed is Microsoft's open-source library for training and running very large models across many GPUs, known for its ZeRO memory-saving technique.Quick
DescriptDescript is an audio and video editor built around editable transcripts, with AI tools for cleanup, clips, captions and synthetic voice correction.Quick
DevinDevin is Cognition’s software-engineering agent for taking on scoped coding tasks, using development tools and returning changes for review.Quick
DevstralDevstral is Mistral AI’s model family specialized for software-engineering agents that inspect repositories, edit files, and solve coding tasks.Quick
Differential privacyDifferential privacy adds carefully calibrated randomness so aggregate analysis reveals useful patterns while limiting what can be inferred about any individual record.Quick
DiffusersDiffusers is Hugging Face’s library for running and training diffusion models for images, video, audio and related generative tasks.Quick
Diffusion modelsA diffusion model learns to reverse a process that gradually adds noise to data.3 min
DifyDify is an open platform for building, testing and operating AI applications, agents and retrieval workflows through visual and API tools.Quick
Dimensionality reductionDimensionality reduction represents data using fewer variables while preserving useful structure, aiding visualization, compression, denoising, or downstream modeling.Quick
DINOv2DINOv2 is Meta’s self-supervised vision model, trained to produce general-purpose image features without relying on manually labeled datasets.Quick
Discord botA Discord bot listens for commands or events and provides focused AI features within authorised servers and channels.Quick
DistillationKnowledge distillation trains a smaller student model to imitate a larger teacher's outputs or internal representations, often reducing deployment cost.Quick
DoclingDocling parses formats such as PDF and Word into structured representations that preserve text, tables, layout and document hierarchy.Quick
Document parsingDocument parsing extracts structure and content from files such as PDFs, forms, and slides so downstream systems can search or analyze them.Quick
Document Q&AA document Q&A application retrieves relevant sections from one or more files and answers questions using that supplied context.Quick
DolmaDolma is Ai2’s open corpus of web text, books, code, papers, and other sources used to train OLMo models.Quick
Dot productThe dot product multiplies two equal-length lists of numbers position by position and adds the results into one number set by the angle between them and their lengths.4 min
DPODPO teaches a language model to prefer the better of two answers by learning from the comparison itself, with no separate reward model trained first.4 min
DropoutDuring training, dropout randomly replaces some input tensor elements with zero.3 min
Drug discoveryAI-assisted drug discovery uses computational models to identify targets, predict molecular properties, design candidates, and prioritize experiments during medicine development.Quick
DSPyDSPy is a Python framework where you describe an AI job by naming its inputs and outputs, and it tunes the prompts for you.4 min
DVCDVC adds versioning and reproducible pipelines for datasets, models and machine-learning experiments alongside Git-managed code. in practice. in practice.Quick
E5E5 is a family of text-embedding models trained for retrieval and semantic similarity using a unified text-to-vector approach.Quick
ElasticsearchElasticsearch is a distributed search and analytics engine that supports keyword, vector and hybrid retrieval over indexed data.Quick
EleutherAIEleutherAI is a nonprofit open research collective known for language models, training datasets and evaluation tooling released for public use.Quick
ElevenLabsElevenLabs is an AI audio platform for speech synthesis, voice cloning, dubbing, transcription and conversational voice applications in practical workflows.Quick
Email assistantAn email assistant drafts, summarises, classifies or routes messages using the user’s instructions and authorised mailbox context. It remains subject to human review.Quick
Embedding layerAn embedding layer turns integer indexes into dense vectors of fixed size.3 min
Embedding modelsAn embedding model maps an input to a numeric vector whose distance from another vector can measure relatedness.3 min
EmbeddingsAn embedding maps a discrete item such as a token ID to a dense vector of numbers.3 min
Emergent abilitiesEmergent abilities are task capabilities that appear absent in smaller models but present at a larger scale.3 min
Encoder-decoderAn encoder-decoder model combines an encoder with an autoregressive decoder for sequence generation.3 min
Encoder-only modelsAn encoder-only model uses bidirectional self-attention to represent a supplied sequence.3 min
Ensemble learningEnsemble learning combines the predictions of several models, usually by a vote or an average, so that their separate mistakes partly cancel out.4 min
EpochAn epoch is one complete pass through every example used to train a model.4 min
EU AI ActThe EU AI Act is a European Union legal framework that regulates AI systems according to risk, with differing obligations for providers and deployers.Quick
EvalsAn eval checks an AI system's work. You feed it an input and use a set of rules to score what comes back.5 min
Evaluation harnessAn evaluation harness runs models or prompts against a fixed test set, records outputs and calculates repeatable quality, safety or performance measures.Quick
Existential riskExistential risk from AI refers to scenarios where AI contributes to human extinction or permanently and drastically curtails humanity's future potential.Quick
Expert systemsAn expert system stores a specialist's know-how as if-then rules and uses a separate inference engine to apply those rules to a new case and explain its advice.5 min
F1 scoreThe F1 score rolls a classifier's precision and recall into one number from 0 to 1, and it stays low if either of them is low.4 min
FAISSFAISS is Meta’s open library for efficient similarity search and clustering over dense vectors, including very large collections.Quick
faster-whisperfaster-whisper is an open reimplementation of Whisper inference using CTranslate2 for faster, more memory-efficient speech transcription in practical workflows.Quick
FastMCPFastMCP is a Python framework for building Model Context Protocol servers and clients, so an AI app can reach your own tools and data through one standard interface.Quick
Feature engineeringFeature engineering turns raw data, like prices and colour names, into the lists of numbers a model can learn from.4 min
Feature importanceFeature importance gives each input column a score for how much a trained model relies on it, so you can see what drives its predictions.4 min
Feature scalingFeature scaling puts numeric variables on comparable ranges, limiting disproportionate influence from large-magnitude features in methods sensitive to distances or gradient sizes.Quick
FeaturesFeatures are the input facts a machine learning model reads about each example, such as a car's mileage or colour, turned into a list of numbers.5 min
Federated learningFederated learning trains a shared model across distributed devices or organizations while keeping raw training data at its original location.Quick
Few-shot learningFew-shot learning is the ability to perform a task from only a small number of labelled examples.3 min
Few-shot promptingFew-shot prompting places a small set of example inputs and outputs in the prompt so the model can imitate the task on a new input.3 min
Fine-tuningFine-tuning takes a model that is already trained and keeps training it on a smaller set of examples for one task.4 min
Fine-tuning pipelineA fine-tuning pipeline prepares examples, runs an adaptation job, records configuration and evaluates the resulting model against a held-out set.Quick
FineWebFineWeb is Hugging Face’s large, open, filtered web-text dataset created from Common Crawl for training and researching language models.Quick
FirecrawlFirecrawl crawls websites and converts pages into cleaner Markdown or structured data for search, retrieval and agent applications.Quick
Fireworks AIFireworks AI provides managed inference and model customisation for generative applications, with APIs optimised for speed and production operation.Quick
FlashAttentionFlashAttention computes exact attention with tiling that reduces reads and writes between GPU memory levels.3 min
FLOPsFLOPs count the small arithmetic steps, like one multiply or one add on decimal numbers, that a computer does to train an AI model.5 min
Flow matchingFlow matching trains a generative model to learn a smooth path that turns random noise into data, an approach closely related to diffusion.Quick
FlowiseFlowise is a visual builder for creating model, retrieval and agent workflows from connected nodes and integrations. in production workflows.Quick
FLUXFLUX is Black Forest Labs’ family of generative image models, offered in open and hosted variants for text-to-image creation and editing.Quick
Foundation modelsA foundation model is trained on broad data and can be adapted for many downstream tasks.3 min
FP8FP8 refers to eight-bit floating-point formats that reduce memory and increase accelerator throughput, while requiring careful scaling to manage limited precision and range.Quick
Fraud detectionAI fraud detection identifies suspicious transactions or behavior by learning patterns associated with legitimate activity, known abuse, and unusual events.Quick
Frequency and presence penaltiesFrequency and presence penalties are generation settings that make a model less likely to repeat tokens it has already used.Quick
Frontier modelsFrontier models are highly capable general-purpose AI models at or beyond the capabilities of the most advanced current models.3 min
Full Self-Driving (Supervised)Full Self-Driving (Supervised) is Tesla’s driver-assistance software for navigation and vehicle control; the human driver must remain attentive and responsible.Quick
Function callingFunction calling lets a model request a named function by returning structured arguments that an application, or the provider, then runs.4 min
GammaGamma generates and edits presentations, documents and web-style pages from prompts using structured layouts rather than traditional slide-by-slide authoring.Quick
GANsA GAN trains a generator to make samples and a discriminator to distinguish generated samples from training data.3 min
GeminiGemini is Google’s family of multimodal AI models, spanning cloud and on-device systems for text, images, audio, video, code, and reasoning.Quick
Gemini 3.5Gemini 3.5 is Google’s 2026 multimodal model generation, built for stronger reasoning, coding, agentic workflows, and understanding across text, images, audio, video, and tools.Quick
Gemini APIThe Gemini API gives developers hosted access to Google’s Gemini models for text, image, audio, video and tool-using applications.Quick
Gemini appGemini is Google’s AI app: you ask questions, upload files for answers and summaries, and can have it research a topic across many sources.5 min
Gemini CLIGemini CLI is Google’s open-source terminal agent for coding, research and automation with Gemini models and tool integrations.Quick
Gemini EnterpriseGemini Enterprise is Google Cloud’s workplace AI platform for searching company information, using enterprise data and applications, and building governed agents for business workflows.Quick
Gemini for Google WorkspaceGemini for Google Workspace adds writing, summarisation, analysis and meeting assistance across Gmail, Docs, Sheets, Slides, Drive and Meet.Quick
Gemini NanoGemini Nano is Google’s compact model family for on-device AI, enabling selected generative features without sending every request to the cloud.Quick
Gemini RoboticsGemini Robotics is Google DeepMind’s model family for connecting multimodal reasoning with physical robot actions, spatial understanding, and instruction following.Quick
GemmaGemma is Google’s family of lightweight open models, derived from Gemini research and intended for developers to run, fine-tune, and deploy themselves.Quick
Gemma 4Gemma 4 is Google’s open-weight, natively multimodal model family, designed for advanced reasoning and agentic workflows across edge devices and personal computers.Quick
Generative AIGenerative AI learns patterns from existing data, then uses those patterns to create new text, images, audio, video, code or other content.5 min
Generative engine optimizationGenerative engine optimization adapts content so AI-powered answer systems can discover, understand, cite, or accurately represent it to users.Quick
Genetic algorithmsA genetic algorithm searches for good answers the way breeding does. It keeps a population of candidates, mixes the fitter ones and repeats over many generations.4 min
GenkitGenkit is Google’s open-source framework for building and observing AI applications with models, tools, retrieval, flows, and evaluation.Quick
GGUFGGUF is a binary format for storing quantized models and metadata, widely used by llama.cpp and related local-inference tools.Quick
GitHub CopilotGitHub Copilot is an AI helper for programmers that offers the next lines of code while you write, explains code when you ask, and can take on jobs you give it.4 min
GleanGlean is an enterprise search and AI platform for finding organisational knowledge and building assistants or agents over connected company systems.Quick
GLM-4.5GLM-4.5 is a Zhipu AI open-weight model generation designed for reasoning, coding, tool use, and agentic application workflows.Quick
GloVeGloVe is a word-embedding method that learns vector representations from global word co-occurrence statistics across a text corpus.Quick
Google ADKGoogle ADK is an open-source toolkit from Google for writing AI agents in code, giving them tools, a record of each chat and helper agents.5 min
Google AI OverviewsGoogle AI Overviews are generated summaries shown for some searches, combining information from web results with links for further reading.Quick
Google AI StudioGoogle AI Studio is a browser-based workspace for prototyping prompts, testing Gemini capabilities and obtaining API integration code.Quick
Google ColabGoogle Colab is a hosted Jupyter notebook environment with shareable documents and optional access to accelerated computing resources.Quick
Google FlowGoogle Flow is an AI filmmaking tool that combines Google’s video, image and language models for creating shots, scenes and cinematic sequences.Quick
Google TPUGoogle Tensor Processing Units are specialised accelerators designed for machine-learning workloads across Google’s cloud and internal infrastructure. in practical systems.Quick
GPQAGPQA is a 448-question multiple-choice benchmark written by experts in biology, physics and chemistry.3 min
GPTGPT is OpenAI’s family of generative transformer models, designed to understand prompts and produce text, code, structured data, and other outputs.Quick
GPT ImageGPT Image is OpenAI’s image generation family for creating and editing visuals from conversational instructions, including text rendering and iterative refinements.Quick
GPT-2GPT-2 is an early OpenAI transformer language model that helped demonstrate coherent long-form text generation from a simple textual prompt.Quick
GPT-3GPT-3 is OpenAI’s influential 175-billion-parameter language model, which demonstrated that large-scale pretraining can support many tasks through prompting alone.Quick
GPT-4GPT-4 is an OpenAI large language model generation known for stronger reasoning and instruction following than the earlier GPT-3.5 family.Quick
GPT-5GPT-5 is an OpenAI model generation built for general-purpose reasoning, coding, writing, analysis, and tool-assisted work across consumer and developer products.Quick
GPT-6 AstraGPT-6 Astra is OpenAI’s 2026 frontier model for advanced reasoning, coding, computer use, scientific work, and long-running agentic tasks across ChatGPT and the API.Quick
gpt-ossgpt-oss is OpenAI’s family of open-weight language models, intended for developers who need locally deployable reasoning models they can inspect and adapt.Quick
GPUsA GPU is a chip built to run many similar calculations at the same time, which suits the matrix maths inside neural networks.5 min
Gradient boostingGradient boosting builds one strong model from many small decision trees, added one at a time, each trained to fix the errors the trees before it still make.4 min
Gradient checkpointingGradient checkpointing stores fewer activations by recomputing them when needed, exchanging extra computation for lower memory use.4 min
Gradient descentGradient descent repeatedly measures how loss changes with each parameter and moves the parameters a small step toward lower loss.3 min
GradioGradio is a Python library for quickly building web interfaces and demos around machine-learning models or functions. in practice.Quick
GrammarlyGrammarly provides writing assistance for grammar, clarity, tone and rewriting across browser, desktop and workplace applications. in daily work.Quick
GraniteGranite is IBM’s open model family for enterprise language, code, retrieval, safety, document understanding, and time-series applications across business domains.Quick
Graph neural networksA graph neural network updates each node's hidden state based on messages.3 min
GraphRAGGraphRAG turns your documents into a map of people, places and links, groups and summarises it, then answers questions from that map.4 min
Greedy decodingGreedy decoding generates a sequence by choosing the most likely next token at every step.3 min
GrokGrok is xAI’s family of conversational models, built for general knowledge, reasoning, coding, multimodal understanding, and tool use.Quick
Grok 4.7Grok 4.7 is xAI’s 2026 model release for reasoning, coding, information synthesis, and tool-using tasks across the Grok product and developer platform.Quick
Grok appGrok is the assistant from SpaceXAI (formerly xAI) for conversation, web and X search, file analysis, voice and media creation.5 min
GroqGroq provides hosted model inference on its Language Processing Unit hardware, focusing on low-latency execution for supported workloads.Quick
Ground truthGround truth is the best available account of what really happened, the reality that a model's predictions are checked against.5 min
GroundingGrounding gives a model relevant source material and asks it to base its response on that evidence.3 min
Grouped-query attentionGrouped-query attention lets each group of query heads share one key head and one value head.3 min
GRPOGroup relative policy optimization trains a policy by comparing rewards among multiple responses to the same prompt, avoiding a separate learned value model.Quick
GRUA GRU is a recurrent unit with reset and update gates.3 min
GSM8KGSM8K is a dataset of grade-school mathematics word problems used to evaluate multi-step arithmetic reasoning in language models.Quick
GuardrailsGuardrails are checks placed around an AI model that inspect what goes in and what comes out, and can block, change or check inputs and replies that are unsafe, off-topic or break the app's rules.5 min
Guardrails AIGuardrails AI is an open-source framework for checking and correcting language-model inputs and outputs against rules you define.Quick
GuidanceGuidance is a language for controlling model generation with templates, constraints, tool calls, and programmatic branching around generated text.Quick
HailuoHailuo is MiniMax’s generative video product and model family for creating, extending, and transforming clips from text and image prompts.Quick
HallucinationA hallucination is generated content that sounds plausible but is false, unsupported by evidence, or inconsistent with the given input.3 min
HarveyHarvey is a legal AI platform for research, drafting, analysis and professional workflows within law firms and legal departments.Quick
HaystackAn open-source Python toolkit from deepset that builds AI apps (search, RAG, agents) by linking small parts into a pipeline.4 min
HBM memoryHigh-bandwidth memory stacks memory near processors to deliver very high data transfer rates, helping keep AI accelerators supplied with model data.Quick
HeliconeHelicone is an observability gateway for language-model applications, providing request logging, cost tracking, caching, experiments, and production monitoring.Quick
HellaSwagHellaSwag tests commonsense reasoning by asking a model to select the most plausible continuation of a described everyday situation.Quick
HeyGenHeyGen creates avatar-led and translated videos from scripts, with synthetic presenters, voice tools and lip-synchronised dubbing. for creative production.Quick
Hidden Markov modelsA hidden Markov model infers a sequence of states you cannot see from a sequence of observations you can, using the odds of each state following another and of each state producing each observation.4 min
HNSWHNSW finds the stored vectors closest to a query quickly by searching a stack of linked graphs, from a sparse top layer down to a dense bottom one.4 min
Hugging FaceHugging Face runs the Hub, a website where people share AI models, datasets and small demo apps so others can find, download and reuse them.4 min
Hugging Face AccelerateHugging Face Accelerate simplifies running PyTorch training and inference across CPUs, GPUs, mixed precision and distributed environments. in practice.Quick
Hugging Face DatasetsHugging Face Datasets provides consistent tools for finding, loading, processing and streaming datasets used in machine-learning workflows. in practice.Quick
Hugging Face TransformersTransformers is a free, open-source Python library that lets you download a ready-trained AI model by name and run or train it with a short piece of code.4 min
Human evaluationHuman evaluation asks people to assess model outputs for qualities that automated metrics may miss, such as usefulness, clarity, correctness, or preference.Quick
Human in the loopWith this setup, a person checks an AI system's work at chosen points and can approve it, change it, reject it or step in.5 min
HumanEvalHumanEval measures functional correctness for programs synthesized from docstrings.3 min
Humanity's Last ExamHumanity's Last Exam is a broad benchmark of difficult expert-level questions intended to probe advanced academic reasoning across many disciplines.Quick
HunyuanVideoHunyuanVideo is Tencent’s generative video model family for text-to-video and image-conditioned video creation, with open model releases available.Quick
Hybrid searchHybrid search runs a keyword ranker and a vector ranker on one query, then merges the two lists into a single order.4 min
HyDEHyDE has a language model write a made-up answer to your question, then searches for real documents that look like that answer.4 min
HyperparametersHyperparameters are the settings people choose to control training, such as the learning rate, rather than values the model learns itself.5 min
IdeogramIdeogram is a generative image product known for prompt-based visual creation and comparatively strong rendering of text inside images.Quick
Image captioning appAn image captioning application analyses visual content and produces descriptive text for search, accessibility or downstream processing. It remains subject to human review.Quick
Image classificationImage classification assigns one or more predefined category labels to an entire image based on its overall visual content.Quick
Image editing modelsAn image editing model changes a picture you already have.4 min
Image generator appAn image generator application turns prompts and optional references into generated visuals with controls for style, size and iteration.Quick
Image segmentationImage segmentation assigns labels to individual pixels or regions, separating objects, surfaces, or semantic categories within an image.Quick
ImagenImagen is Google’s family of text-to-image diffusion models, built to generate and edit high-quality images from natural-language instructions.Quick
ImageNetImageNet is a large labeled image dataset and benchmark that helped drive modern deep-learning progress in visual recognition.Quick
Imbalanced dataImbalanced data is a dataset where one class is far rarer than another, such as a few fraud cases among many ordinary card payments.4 min
Imitation learningImitation learning trains an agent to act by copying examples of an expert's behaviour instead of learning only from rewards.Quick
In-context learningIn-context learning is a model's ability to infer a task or pattern from instructions and examples in its prompt without updating parameters.Quick
Indirect prompt injectionIndirect prompt injection is when an attacker hides instructions in content an AI app reads later, such as an email or web page, instead of typing them in.4 min
InferenceInference is using a trained AI model on new input to get an answer, such as a label, a number or a written reply, without changing what the model learned.5 min
Inference optimizationInference optimization is a set of methods, like caching, quantization and speculative decoding, that let a trained language model answer with less work or less memory.3 min
Inspect AIInspect AI is the UK AI Security Institute’s open-source framework for building reproducible model evaluations with tasks, tools, solvers, and scoring.Quick
Instruct modelsAn instruct model is a pretrained base version further fine-tuned on instructions and conversational data.3 min
Instruction followingInstruction following is how well a model does what a written request asks, meeting each rule it sets, such as rules on content, style, format or numbers.3 min
Instruction tuningInstruction tuning fine-tunes a pretrained model on many tasks written as instructions paired with desired responses.3 min
InstructorInstructor is a library for extracting validated structured data from model responses using schemas, retries, and provider-specific integrations.Quick
InterpretabilityInterpretability seeks to understand how a model processes information and reaches outputs, using analyses of behavior, representations, parameters, or internal computations.Quick
Invoice processingAn invoice processing workflow extracts supplier, line-item, tax and payment fields, validates them and routes exceptions for review.Quick
JailbreaksA jailbreak is a prompt written to trick an AI model into ignoring its safety training and producing something it would normally refuse.4 min
JanJan is an open desktop application for running local language models and connecting to compatible hosted providers through one interface.Quick
JAXJAX is a Python numerical-computing library that adds automatic differentiation, compilation, vectorization, and accelerator support to NumPy-style programs.Quick
JevJev is TypeSafe AI’s decision model, returning typed choices, scores, or probabilities for software automation instead of generating open-ended natural-language responses.Quick
JSON modeJSON mode is an API setting that asks a language model to reply in JSON, text a program can parse (read into data), and on some providers, such as OpenAI, enforces valid syntax, without fixing which keys or types appear.4 min
JSON SchemaJSON Schema is a standard vocabulary for describing and validating the structure, types, constraints, and required fields of JSON data.Quick
JulesJules is Google’s asynchronous coding agent, designed to work on repository tasks in the background, propose changes, and return results for developer review.Quick
JupyterJupyter is an open interactive computing environment built around notebooks that combine executable code, narrative text, data and visual output.Quick
K-means clusteringK-means splits unlabelled data into k groups by putting each point with its nearest centre, then moving each centre to the average of its group, over and over.4 min
K-nearest neighborsk-nearest neighbours labels a new example by finding the k most similar stored examples and taking their majority vote, or their average for numbers.4 min
KaggleKaggle is a data-science platform for datasets, notebooks, models, competitions and community learning with hosted compute. in production.Quick
KerasKeras is a high-level deep-learning API for defining and training neural networks with a concise interface across supported computation backends.Quick
Kimi K2Kimi K2 is Moonshot AI’s open-weight mixture-of-experts language model, designed for coding, tool use, software development, and agentic application tasks.Quick
KiroKiro is AWS’s agentic development environment, combining code assistance with specification-driven planning, tasks and automated implementation workflows. in software teams.Quick
KlingKling is Kuaishou’s generative video model family, producing and editing video from text or image prompts through a commercial platform.Quick
Knowledge cutoffA knowledge cutoff is the latest period a model can reliably know from its training alone.3 min
Knowledge graph builderA knowledge graph builder extracts entities and relationships from source material, resolves duplicates and stores linked facts with provenance.Quick
Knowledge graphsA knowledge graph stores facts as labelled links between real-world things, so software can follow the links to answer questions.4 min
KokoroKokoro is a lightweight open text-to-speech model designed to produce natural-sounding speech across multiple voices with relatively modest computational requirements.Quick
KV cacheA KV cache lets a language model save work from earlier tokens, so each new token does not redo the same maths.3 min
LabelsA label is the answer attached to a training example, such as a category, a number or a ranking, that a supervised model learns to predict.5 min
LAION-5BLAION-5B is a large open dataset of image-text pairs collected from the web and filtered using CLIP representations.Quick
LanceDBLanceDB is an embedded and serverless-oriented vector database built on the Lance columnar data format for multimodal AI data.Quick
LangChainLangChain is an open-source framework for building apps and agents on top of large language models, with one standard way to talk to many AI providers.5 min
LangflowLangflow is a visual framework for composing, testing and serving AI workflows, agents and retrieval pipelines. in production workflows.Quick
LangfuseLangfuse is an open-source platform for tracing, evaluating, prompt-managing, experimenting with, and monitoring language-model applications in development and production.Quick
LangGraphLangGraph is an open-source framework for building AI agents as graphs, where steps share one state that can be saved, paused and resumed.5 min
LangSmithLangSmith is LangChain’s platform for tracing, testing, evaluating, debugging, and monitoring language-model applications and agent workflows throughout development and production.Quick
Language learning tutorA language-learning tutor provides conversation practice, corrections, vocabulary support and level-appropriate exercises with immediate feedback. It remains subject to human review.Quick
LatencyRequest latency is the elapsed time between sending an inference request and receiving its final response.3 min
LayaLaya is ConvAI Innovations’ open-weight decision-model family, designed to return typed classifications, scores, and calibrated probabilities rather than open-ended generated text.Quick
Layer normalizationLayer normalization normalizes selected activations independently for each example in a batch.3 min
Lead qualification agentA lead qualification agent evaluates prospective customers against explicit criteria, enriches records and routes qualified opportunities for human follow-up.Quick
Learning rateThe learning rate multiplies a gradient, and that product is how far the weight moves.3 min
Learning-rate schedulesA learning-rate schedule changes how big each training step is over time, often starting with a short warmup and then lowering the rate.Quick
LettaLetta is a framework for building stateful agents with explicit memory management, tools and long-running interactions. in production workflows.Quick
LibreChatLibreChat is a self-hosted multi-provider chat interface with model switching, agents, tools, retrieval and user-management features. under user control.Quick
LightGBMLightGBM is Microsoft’s gradient-boosting library, optimised for efficient training on large tabular datasets using histogram-based tree methods. in practice.Quick
Linear algebra for MLThe small part of linear algebra that models use every day, vectors, matrices, their products and their sizes, to hold data, run layers and learn.4 min
Linear attentionLinear attention rewrites attention with feature maps so key-value summaries are formed before they are combined with queries.3 min
Linear regressionLinear regression predicts a number by multiplying each feature by a learned weight, adding the results and adding a bias.5 min
LiteLLMLiteLLM provides a unified API and proxy for calling many model providers, with routing, spend tracking, retries, and observability features.Quick
LlamaLlama is Meta’s family of openly available large language models, used for research, fine-tuning, and self-hosted generative AI applications.Quick
Llama 4Llama 4 is a generation of Meta’s open-weight model family, using mixture-of-experts designs and supporting multimodal inputs in selected variants.Quick
Llama GuardLlama Guard is a family of Meta safety models that classify prompts and responses as safe or unsafe.Quick
LLaMA-FactoryLLaMA-Factory is an open-source toolkit, with a web interface and a command line, for fine-tuning many open language models.Quick
llama.cppllama.cpp is an open-source program that runs large language models on your own computer.4 min
llamafilellamafile packages model weights and a portable inference runtime into executable files intended to run across common operating systems.Quick
LlamaIndexLlamaIndex is an open-source toolkit that brings your own files and data to a language model when you ask a question, so apps and agents can answer from them.4 min
LlamaParseLlamaParse is LlamaIndex's document parsing service that turns PDFs and other files into text ready for LLM apps.Quick
LLMAn LLM predicts a token or sequence of tokens, sometimes many paragraphs long.3 min
LLM evaluationLLM evaluation means testing an AI feature with structured tests, so you can check how accurate and reliable it is even though its answers vary.5 min
LLM gatewaysAn LLM gateway is one server that sits between your apps and AI model providers, adding limits, caching, fallbacks and logs to model calls.5 min
LLM observabilityLLM observability means recording what an AI app does on each request, such as the prompt, each step, the time taken and the tokens used.4 min
LLM-as-a-judgeLLM-as-a-judge means asking a strong language model to grade answers to open-ended questions.3 min
LLMOpsThe day-to-day work of running an app built on a language model, from managing prompts and testing answers to watching speed and cost.5 min
llms.txtllms.txt is a proposed website file format that gives language models a concise, curated map of important documentation and resources.Quick
LM StudioLM Studio is a desktop application for discovering, downloading and running compatible language models locally through chat and developer endpoints.Quick
LMArenaLMArena is a platform that compares language models through blinded user votes on responses to the same prompts.Quick
LobeChatLobeChat is an open self-hosted chat and agent interface supporting multiple model providers, plugins, knowledge bases and multimodal interaction.Quick
Local LLM chat appA local LLM chat app runs a compatible model on user-controlled hardware and presents it through a conversational interface.Quick
LocalAILocalAI is a self-hosted API layer for running multiple local model types behind interfaces compatible with common hosted AI services.Quick
Logistic regressionLogistic regression weighs and adds up an example's features, squeezes the total into a probability between 0 and 1, and uses a threshold to pick a class.4 min
LogitsLogits are a model's raw, unnormalized prediction scores.3 min
Long contextTransformer-XL reuses hidden states from previous segments instead of computing them from scratch for each new segment.4 min
LoRALoRA adapts a large model to a new task by freezing its weights and training two thin matrices whose product is added to them.5 min
Loss functionA loss function turns the difference between a model's prediction and its target into a number that training tries to reduce.3 min
Lost in the middleLanguage models often use facts near the opening or closing of a long input better than facts buried in the middle.5 min
LovableLovable turns natural-language product requests into editable full-stack web applications, with visual iteration, code ownership and integrated deployment workflows.Quick
LSTMAn LSTM is a recurrent unit whose multiplicative gates regulate access to its memory path.3 min
Luma Dream MachineLuma Dream Machine is a generative video service that creates and modifies clips from text prompts, images and visual direction.Quick
LyriaLyria is Google DeepMind’s generative music model family, designed to create instrumental and vocal music from textual or musical guidance.Quick
Machine learningMachine learning trains software on data so it can find patterns and make useful predictions or generate content for new inputs.5 min
Machine translationMachine translation automatically converts text or speech from one language into another while attempting to preserve meaning, tone, and context.Quick
MagistralMagistral is Mistral AI’s reasoning-model family, designed to work through multi-step mathematics, coding, analysis, and decision problems with deliberate inference.Quick
MambaMamba is a sequence architecture whose state space parameters depend on the input, letting each token control what information is kept or forgotten.4 min
ManusManus is a general AI agent product that works through multi-step tasks using web research, files, code and other tools in a managed environment.Quick
MarkerMarker converts PDFs and other documents into Markdown, JSON or structured output for search, analysis and model-based workflows.Quick
Markov chainsA Markov chain models transitions among states where the next state's probability depends on the current state under the standard Markov assumption.Quick
Markov decision processA Markov decision process describes a decision problem as states, actions, rewards and the chances of moving between states; it is the standard frame for reinforcement learning.Quick
Masked language modelingMasked language modeling trains a model to guess words hidden in a sentence from the words around them; it is the training task used by BERT.Quick
MastraMastra is an open-source TypeScript toolkit for making AI agents, step-by-step workflows and the apps that use them.3 min
MATH benchmarkMATH is a benchmark of competition-style mathematics problems that tests multi-step reasoning across algebra, geometry, calculus, probability, and related subjects.Quick
MatricesA matrix is a rectangular grid of numbers with a fixed number of rows and columns. Machine learning keeps data and layer weights in matrices and multiplies them.4 min
Max tokensMax tokens sets an upper bound on how many tokens a model may generate for one response.3 min
MCPMCP standardizes how AI applications connect to external context and tools.4 min
MCP serverAn MCP server exposes tools, resources or prompts through the Model Context Protocol so compatible AI clients can discover and use them.Quick
Mechanistic interpretabilityMechanistic interpretability investigates neural networks by identifying internal components, representations, and computations that causally produce particular observed behaviors.Quick
Meeting summarizerA meeting summarizer turns a transcript into concise decisions, discussion themes and follow-up actions that participants can verify.Quick
Mem0Mem0 is a memory layer for AI applications that extracts, stores and retrieves useful information across user or agent interactions.Quick
Meta AIMeta AI is Meta’s assistant across its apps, the web and a standalone app for conversation, search, voice and image creation.5 min
Metadata filteringMetadata filtering limits a vector search to records whose labels, such as category or year, match rules you set.4 min
Microsoft 365 CopilotMicrosoft 365 Copilot embeds AI assistance across Word, Excel, PowerPoint, Outlook, Teams and organisational data for drafting, analysis and collaboration.Quick
Microsoft Agent FrameworkMicrosoft Agent Framework is Microsoft's open-source framework for building AI agents and multi-agent workflows, the successor to AutoGen and Semantic Kernel.Quick
Microsoft CopilotMicrosoft Copilot is an AI helper from Microsoft that you can type to, talk to or show a picture, and it will reply, write something new or carry out a job for you.5 min
Microsoft Copilot StudioMicrosoft Copilot Studio is a managed environment for designing, connecting, governing and publishing agents across Microsoft and external systems.Quick
Microsoft FoundryMicrosoft Foundry is Microsoft's Azure platform for building AI apps and agents, with a large model catalogue, tools, testing and safety controls in one place.5 min
MidjourneyMidjourney is a generative image service for prompt-driven visual creation and iterative editing through its web and community interfaces.Quick
MilvusMilvus is a distributed vector database built for large-scale similarity search across high-dimensional embedding collections. in production workflows.Quick
MiniMaxMiniMax is a family of language and multimodal models from the company MiniMax, used for text, speech, music, image, and video applications.Quick
MinistralMinistral is Mistral AI’s compact model family, designed for edge, on-device, and latency-sensitive language applications with constrained computing resources.Quick
Missing dataMissing data are empty cells in a dataset. You can drop them, fill them with reasoned guesses or use a model that accepts gaps, and why they are missing decides which is safe.5 min
Mistral La PlateformeMistral La Plateforme provides APIs, model deployment, fine-tuning and agent-building services around Mistral’s model family. in production. in production.Quick
Mistral LargeMistral Large is Mistral AI’s higher-capability commercial model tier for complex reasoning, multilingual work, code, and enterprise applications.Quick
Mistral MediumMistral Medium is a mid-tier Mistral AI model positioned between small, efficient models and the company’s highest-capability offerings.Quick
Mistral SmallMistral Small is Mistral AI’s efficiency-focused model tier for responsive conversational, coding, and business automation workloads at production scale.Quick
Mixed precisionMixed precision trains with both 16-bit and 32-bit numbers so the run is faster and uses less memory.4 min
MixtralMixtral is Mistral AI’s open mixture-of-experts model family, activating only part of the network for each token to improve efficiency.Quick
Mixture of expertsA mixture-of-experts layer uses a learned gate to select a sparse combination of expert subnetworks for each input.3 min
MLC LLMMLC LLM compiles and deploys language models across GPUs, CPUs, browsers and mobile devices using machine-learning compilation techniques.Quick
MLflowMLflow is an open platform for tracking experiments, packaging models, managing versions and supporting machine-learning deployment workflows in practical workflows.Quick
MLOpsMLOps is a way of working that makes building and releasing machine learning models simpler and more automatic.5 min
MLXMLX is Apple's open-source array framework for machine learning on Apple silicon, used to run and fine-tune models on Macs.Quick
MMLUMMLU is a 57-task test of a text model's multitask accuracy.3 min
MNISTMNIST is a classic dataset of handwritten digit images, commonly used to teach and benchmark basic image-classification systems.Quick
Model cardsModel cards are structured documents describing a model's intended uses, evaluation results, limitations, training context, and important ethical or safety considerations.Quick
Model collapseModel collapse is degradation that can occur when successive models train heavily on generated data, losing diversity or fidelity to the original distribution.Quick
Model mergingModel merging combines parameters or learned changes from multiple related models to blend capabilities without conventional joint retraining.Quick
Model parallelismModel parallelism divides a model's computation or parameters across multiple devices when one device cannot efficiently hold or run it.Quick
Model routingModel routing sends each prompt to the model that suits it, so simple requests go to cheaper models and hard ones to stronger models.4 min
Model servingModel serving means running a trained model on a server so apps can send it requests over the network and get answers back.5 min
Monte Carlo methodsMonte Carlo methods estimate a number by running many random simulations and averaging what comes out, instead of working the answer out exactly.4 min
MoshiMoshi is Kyutai’s open speech-language model for real-time, full-duplex spoken conversation with low-latency audio input, understanding, generation, and output.Quick
Movie GenMovie Gen is Meta research on generative media models for creating and editing video and producing synchronized audio from prompts.Quick
Multi-agent systemsA multi-agent system is a group of AI agents, each a language model using tools in its own loop, that work together on one job.5 min
Multi-agent teamA multi-agent team divides work among specialised model-driven roles, with explicit handoffs, shared state and a final integration step.Quick
Multi-armed banditsA multi-armed bandit applies an action, observes its reward, then continues the process with another action.3 min
Multi-head attentionMulti-head attention consists of several attention layers running in parallel.3 min
Multilayer perceptronA multilayer perceptron can fit a nonlinear model to training data.3 min
Multimodal modelsA multimodal model can process more than one type of input, such as text, images, audio or video.4 min
MuseMuse is Meta’s personal AI agent for carrying out tasks and longer-term goals across connected apps, using a dedicated secure virtual machine and asking for approval before sensitive actions.Quick
Muse Spark 1.1Muse Spark 1.1 is Meta’s multimodal reasoning model for agentic tasks, with tool and computer use, coding, long-context work, and multi-agent orchestration through Meta AI and the Meta Model API.Quick
Music generation modelsA music generation model can generate music in the raw-audio domain.3 min
n8nn8n is a workflow-automation platform with visual and code-based building blocks, including AI nodes, agents and integrations in practical workflows.Quick
Naive BayesNaive Bayes sorts things into classes by multiplying how common each class is by how likely each feature is in that class, then picking the top score.4 min
Named entity recognitionNamed entity recognition finds and categorizes mentions such as people, organizations, locations, dates, products, or quantities within text.Quick
Nano BananaNano Banana is a widely used nickname associated with a Google Gemini image-generation and editing model, especially for conversational visual transformations.Quick
Narrow AINarrow AI is designed for a limited task or domain, even when its performance there appears sophisticated or exceeds human ability.Quick
Natural language processingNatural language processing develops computational methods for understanding, generating, translating, searching, and analyzing human language in text or speech.Quick
NeMo GuardrailsNeMo Guardrails is NVIDIA's open-source toolkit for adding programmable rules around a conversational AI app, such as blocking unsafe topics or keeping answers on subject.Quick
NemotronNemotron is NVIDIA’s family of language models and training resources, often used to create synthetic data and develop customized enterprise models.Quick
Neural networksA neural network is a model built from layers of simple units that each weigh their inputs, add them up and bend the result, so that together they can learn curved, complex patterns.4 min
Neural radiance fieldsA neural radiance field maps a 3D position and viewing direction to density and view-dependent color.3 min
Next-token predictionNext-token prediction estimates which token should follow the tokens already present in a sequence.3 min
Next.js AI chatbotA Next.js AI chatbot combines a React interface, server routes, model streaming and storage into a deployable conversational web application.Quick
NLLBNLLB, or No Language Left Behind, is Meta’s machine-translation research program and model family covering many low-resource languages.Quick
Nomic EmbedNomic Embed is Nomic AI’s open text-embedding model family, designed for retrieval, clustering, classification, and long-context semantic representation across documents.Quick
NotebookLMNotebookLM is Google’s source-grounded research and learning tool for asking questions, creating summaries and generating audio or study material from selected sources.Quick
Notion AINotion AI adds writing, summarisation, search and workflow assistance to Notion pages, databases and connected workspace knowledge in practical workflows.Quick
Nous ResearchNous Research is an open AI research community and company known for publishing language models, datasets and training work.Quick
NPUsAn NPU is a part of a chip built to speed up AI models at low power.5 min
NVIDIA BlackwellNVIDIA Blackwell is a GPU architecture and computing platform designed for large-scale AI training and inference in data centres.Quick
NVIDIA CUDACUDA is NVIDIA’s parallel computing platform and programming model for running general-purpose workloads, including AI, on NVIDIA GPUs.Quick
NVIDIA H100NVIDIA H100 is a data-centre GPU based on the Hopper architecture, widely used for training and serving large AI models.Quick
NVIDIA NIMNVIDIA NIM packages optimised model inference as deployable microservices for NVIDIA-accelerated infrastructure and enterprise AI applications. in production.Quick
o3o3 is an OpenAI reasoning model designed for complex analysis, coding, mathematics, visual reasoning, and tasks that benefit from deliberate computation.Quick
Object detectionObject detection identifies and locates instances of specified object categories within an image or video, usually using bounding boxes.Quick
OCROptical character recognition converts text visible in images or scanned documents into machine-readable characters for searching, editing, or analysis.Quick
OllamaOllama is an open-source program that downloads open AI models and runs them on your own computer, with a simple command and a local API.4 min
OLMoOLMo is Ai2’s fully open language-model family, publishing model weights, training data, code, and research artifacts for transparent study.Quick
Omni modelsAn omni model processes several input and output modalities within one multimodal model rather than routing everything through a text-only core.4 min
On-device AIOn-device AI runs the model on your own phone, laptop or browser, so your data does not have to go to a server.4 min
On-device assistantAn on-device assistant runs some or all inference locally to reduce latency, limit data transfer and work with device context.Quick
One-hot encodingOne-hot encoding turns a category, such as a colour or a type of transport, into a list of 0s with a single 1 marking which category it is.4 min
ONNXONNX is an open format for representing machine-learning models, enabling models to move between frameworks and inference runtimes.Quick
ONNX RuntimeONNX Runtime is an open inference engine for running models exported in the ONNX format across different hardware and operating systems.Quick
Open WebUIOpen WebUI, a free chat app you host yourself, lets you talk to AI models running on your own computer or in the cloud, from one web page.4 min
Open-weight modelsAn open-weight model makes its trained weights publicly available for download.3 min
OpenAI Agents SDKOpenAI's Agents SDK is a small open-source toolkit from OpenAI for building AI agents that use tools, pass work to each other and can record each run as a trace.5 min
OpenAI APIThe OpenAI API gives developers hosted access to OpenAI models and tools for text, audio, vision, images and agentic applications.Quick
OpenAI CodexOpenAI Codex is an agentic software-development product that can inspect repositories, edit code, run commands and collaborate on engineering tasks.Quick
OpenAI embeddingsOpenAI embeddings are API models that convert text into numerical vectors for semantic search, clustering, recommendations, classification, and retrieval systems.Quick
OpenAI EvalsOpenAI Evals is an open-source framework and registry for evaluating model behavior on structured test cases and custom tasks.Quick
OpenAI-compatible APIsOpenAI-compatible APIs imitate common OpenAI request and response formats, letting existing clients connect to other model providers with fewer changes.Quick
OpenAPIOpenAPI is a machine-readable standard for describing HTTP APIs, including endpoints, inputs, outputs, authentication methods, errors, and reusable data schemas.Quick
OpenCVOpenCV is a computer-vision library for image and video processing, feature extraction, camera workflows and model integration. in practice.Quick
OpenRouterOpenRouter provides one API and routing layer for accessing models from multiple providers, with unified billing, usage and availability controls.Quick
OpenSearchOpenSearch is a search and analytics suite with keyword, vector and hybrid search capabilities derived from the Elasticsearch ecosystem.Quick
OpenVINOOpenVINO is Intel’s toolkit for optimising and running AI inference across supported Intel CPUs, GPUs and accelerators. in practice.Quick
OpikOpik is Comet’s open-source platform for tracing, evaluating, testing, debugging, and monitoring language-model applications and agent workflows across development and production.Quick
OptunaOptuna is a hyperparameter-optimisation framework that searches configuration spaces, prunes weak trials and records experiment results. in practice.Quick
Otter.aiOtter.ai records, transcribes and summarises meetings, with collaboration features for reviewing conversations, identifying speakers and tracking follow-up items.Quick
OutliersAn outlier is a data point that sits an unusually long way from the rest. It may be a mistake to fix or a rare real event worth keeping.4 min
OutlinesOutlines is a library for constrained generation, guiding language models to produce outputs that follow regex patterns, grammars, or typed schemas.Quick
OverfittingOverfitting happens when a model learns its training examples so specifically that it performs poorly on new examples.3 min
PagedAttentionPagedAttention stores a model's attention memory in small fixed-size blocks, like pages in computer memory, so a server can fit and share more requests at once; it was introduced with vLLM.Quick
Parameter-efficient fine-tuningPEFT adapts a pretrained model by training a small set of added or selected parameters while keeping most base weights frozen.4 min
ParametersParameters are the model values that training can change to improve how inputs map to outputs.3 min
PEFTPEFT is a free Hugging Face library that teaches a big AI model a new task by training a small set of extra weights instead of the whole model.4 min
PerceptronA perceptron is a linear classifier.3 min
PerplexityPerplexity is an AI search product that searches sources, synthesizes an answer and links the evidence used.5 min
Perplexity (metric)Perplexity is the exponentiated average negative log-likelihood that a language model assigns to a token sequence.4 min
Perplexity CometComet is Perplexity’s AI browser, combining conventional browsing with page-aware assistance, search and agent-like actions across open tabs and websites.Quick
Personal knowledge baseA personal knowledge base indexes a user’s notes and files for private search, summarisation and question answering. It remains subject to human review.Quick
PersonalizationPersonalization adapts content, recommendations, interfaces, or services to an individual's observed behavior, stated preferences, context, or predicted needs.Quick
pgvectorpgvector is an open PostgreSQL extension that adds vector storage, distance operators and indexes for similarity search in practical workflows.Quick
PhiPhi is Microsoft’s family of small language models, designed to deliver useful reasoning and language capability with modest compute requirements.Quick
Phi-4Phi-4 is a Microsoft small-language-model generation focused on efficient reasoning, mathematics, coding, instruction following, and selected multimodal application workloads.Quick
PineconePinecone is a managed vector database for semantic search, recommendation and retrieval workloads in AI applications. in production.Quick
Pipeline parallelismPipeline parallelism places consecutive model stages on different devices and overlaps their work using multiple batches moving through the pipeline.Quick
PixtralPixtral is Mistral AI’s vision-language model family, capable of understanding images alongside text in conversational and document workflows.Quick
Podcast transcriberA podcast transcriber converts long-form audio into timestamped text, often adding speaker labels, chapters or searchable excerpts. It remains subject to human review.Quick
Positional encodingPositional encoding adds token-order information to representations before attention reads them.3 min
PPOProximal policy optimization is a reinforcement-learning algorithm that limits the size of policy updates to improve training stability.Quick
Precision and recallPrecision asks how many of the things a model flagged were right. Recall asks how many of the real cases it managed to find.5 min
Predictive maintenancePredictive maintenance uses sensor data and models to estimate equipment failures or degradation, enabling service before costly breakdowns occur.Quick
Prefill and decodePrefill processes all prompt tokens in parallel to initialize model state; decode then generates new tokens sequentially using that cached state.Quick
PretrainingIn one transfer-learning setup, a model is pretrained on a data-rich task before fine-tuning on a downstream task.3 min
Principal component analysisPCA finds the directions in which data spreads out most, then keeps only the first few, so many features become a few new ones.4 min
ProbabilityA probability is a number from 0 to 1 that says how likely something is. Many machine learning problems need one as the answer.4 min
Prompt cachingPrompt caching saves the work a model did on the unchanged start of a prompt, so the next request with the same start is faster and cheaper.4 min
Prompt chainingIn prompt chaining, you give a model one small job at a time in a set order, and each answer becomes the starting material for the next job.5 min
Prompt engineeringPrompt engineering is writing and testing the instructions and examples you give a model so its answer meets a goal you can check.4 min
Prompt injectionPrompt injection is text that sneaks new instructions into what an AI model reads, so the model behaves in unintended ways.5 min
Prompt playgroundA prompt playground lets teams compare instructions, models and settings on shared test cases before changes reach production.Quick
Prompt templatesPrompt templates are reusable prompt structures with variable fields, helping applications apply consistent instructions and context across many requests.Quick
Prompt versioningPrompt versioning tracks changes to prompts alongside metadata and results, making experiments reproducible and production behavior easier to audit.Quick
promptfoopromptfoo is an open-source tool for testing prompts, models, and RAG systems with configurable assertions, comparisons, and security evaluations.Quick
Protein structure predictionProtein structure prediction estimates a protein's three-dimensional shape from its amino-acid sequence or related evidence, helping researchers investigate function and interactions.Quick
PruningPruning removes weights, units, or connections judged less important, shrinking computation or storage while aiming to preserve model quality.Quick
Pydantic AIPydantic AI is a Python framework for building AI agents whose tool inputs and final answers are checked against types you define.4 min
PyTorchPyTorch is a free Python library for building and training neural networks, with fast maths on GPUs and automatic gradients.4 min
PyTorch GeometricPyTorch Geometric is a library of data structures, layers and utilities for graph neural networks built on PyTorch.Quick
PyTorch LightningPyTorch Lightning is a framework that organizes PyTorch training code, reducing boilerplate while supporting distributed training, logging, and reproducible experiments.Quick
Q-learningQ-learning improves an estimate of each state-action value from sampled rewards and the best estimated value at the next state.3 min
QdrantQdrant is a vector database and search engine with filtering, payload storage and APIs for production semantic retrieval.Quick
QLoRAQLoRA fine-tunes a large language model by storing its frozen weights in 4 bits and training only a small add-on that sits beside them.5 min
Qualcomm Snapdragon XQualcomm Snapdragon X is a family of Arm-based PC processors with integrated neural-processing hardware for Windows laptops. in practical systems.Quick
QuantizationQuantization stores a model's numbers with fewer bits, such as 8-bit integers instead of 32-bit decimals, so it needs less memory.4 min
Query rewritingQuery rewriting transforms a user’s request into one or more clearer search queries, improving retrieval when the original wording is vague or conversational.Quick
Question answeringQuestion-answering systems produce answers to natural-language questions using learned knowledge, provided context, retrieved sources, databases, tools, or combinations of these.Quick
QwenQwen is Alibaba Cloud’s family of language and multimodal models, covering general conversation, coding, vision, audio, and specialized tasks.Quick
Qwen3Qwen3 is a generation of Alibaba’s open-weight model family, offering multiple sizes and modes for both direct responses and extended reasoning.Quick
Qwen3-CoderQwen3-Coder is an Alibaba model family specialized for software development, repository-scale understanding, tool use, and agentic coding workflows.Quick
RAGRAG lets an AI model look up relevant documents first, then answer from what it found.4 min
RAG chatbotA RAG chatbot retrieves relevant passages from a chosen knowledge source and gives them to a language model before it answers.Quick
RagasRagas is an open-source evaluation framework for retrieval-augmented generation, measuring retrieved-context relevance, answer faithfulness, and end-to-end response quality.Quick
RAGFlowRAGFlow is an open-source retrieval-augmented generation engine that focuses on reading complex documents, such as tables and scanned pages, before answering questions about them.Quick
Random forestA random forest combines many decision trees trained on varied samples and features, reducing individual-tree instability through aggregated predictions.Quick
Rate limitsA rate limit caps how much you can use an AI service in a set stretch of time, such as how many requests you send each minute.4 min
RayRay is a distributed computing framework for scaling Python and machine-learning workloads from one machine to a cluster.Quick
Ray-Ban Meta glassesRay-Ban Meta glasses combine cameras, microphones, speakers, and Meta AI so wearers can capture media, ask questions, and use voice-controlled assistance.Quick
ReActReAct is a prompting method in which a language model alternates written thoughts with tool actions, reading each result before it picks the next step.5 min
Reasoning modelsReasoning models use intermediate reasoning tokens before producing a final response.4 min
Recommendation engineA recommendation engine ranks items using behavioural, content or business signals and evaluates whether its suggestions are useful.Quick
Recommender systemsA recommender system picks the few items, out of a catalogue too big to browse, that each person sees first, by predicting what that person is likely to want.4 min
Recurrent neural networksAt each time step, a recurrent neural network updates its hidden state from the previous hidden state and the current input.3 min
Red teamingRed teaming means deliberately testing an AI system the way an attacker would, to find its weak spots.4 min
RegressionRegression predicts a number, such as a house price, a tree's life span or a rainfall total, from an example's features.3 min
RegularizationRegularization discourages overly complex model behavior, often by penalizing large weights or adding noise, to improve performance on unseen data.Quick
Reinforcement learningReinforcement learning trains a program, called an agent, to choose actions by trial and error, guided by a number called a reward that says how well it is doing.5 min
Reinforcement learning with verifiable rewardsReinforcement learning with verifiable rewards trains models on tasks whose answers can be checked automatically, such as by tests, proofs, or exact calculations.Quick
ReLUReLU is an activation function that replaces every negative input with zero and leaves every positive input unchanged.3 min
ReplicateReplicate is a hosted platform and API for running published machine-learning models without managing their serving infrastructure directly.Quick
Replit AgentReplit Agent builds, tests and deploys applications inside Replit from natural-language requests, while exposing the project for further editing.Quick
Reranker modelsA reranker sorts existing text candidates by semantic relevance to a specified query.3 min
RerankingReranking takes the short list a first search returned and scores each item again with a slower, more careful model.4 min
Research agentA research agent breaks a question into searches, gathers evidence, compares sources and produces a cited synthesis for review.Quick
Residual connectionsA residual connection combines a residual mapping F(x) with the block input x as F(x) + x.4 min
Residual networksA residual block adds a learned residual function F(x) to a shortcut carrying the block input x.3 min
Resume screenerA resume screener compares application material with explicit job criteria and presents evidence for human review rather than making an opaque hiring decision.Quick
Reward hackingReward hacking occurs when an AI system finds an unintended way to maximize its measured reward without achieving the outcome designers actually wanted.Quick
Reward modelsReward models assign scores to candidate behaviors or outputs, approximating preferences or objectives that reinforcement learning can optimize.Quick
RLAIFRLAIF runs the usual RLHF training loop, but another AI model supplies some of the preference labels that people would normally give.4 min
RLHFRLHF uses human comparisons to learn a reward signal, then optimizes a model to produce responses that score better under that signal.4 min
RoBERTaRoBERTa is a robustly optimized BERT-style encoder that changed training choices and data scale to improve language-understanding performance.Quick
RoboticsRobotics combines sensing, computation, planning, mechanical design, and physical control to build machines that act in the real world.Quick
Robotics foundation modelsA robotics foundation model is pretrained across many tasks, environments or robot bodies so it can serve as a starting point for new robot policies.4 min
ROC curveAn ROC curve shows, for every threshold a classifier could use, how many real positives it catches against how many false alarms it raises.4 min
Rotary position embeddingsRoPE encodes absolute position with a rotation matrix and adds explicit relative-position dependence to self-attention.3 min
RunwayRunway is a creative AI platform for generating and editing video, images and other media through models and production tools.Quick
RWKVRWKV is an open neural architecture combining transformer-like training with recurrent inference, allowing language models to process sequences without a key-value cache.Quick
SafetensorsSafetensors is a secure, fast tensor-storage format designed to load and share model weights without executing arbitrary serialized program code.Quick
Sales email writerA sales email writer drafts personalised outreach from approved product facts and recipient context, leaving claims and final sending under human control.Quick
Salesforce AgentforceSalesforce Agentforce is a platform for building and deploying AI agents that use Salesforce data, workflows and authorised business actions.Quick
SAM 2SAM 2 is Meta’s Segment Anything model for images and video, supporting prompted object segmentation and tracking across video frames.Quick
Sandboxing agentsSandboxing an agent means running its commands inside a fence, set up by the operating system, that limits which files and websites they can reach.4 min
Scalable oversightScalable oversight develops ways for people to supervise systems on tasks too numerous or difficult for unaided humans to evaluate directly.Quick
Scaling lawsScaling laws are empirical relationships describing how model performance changes predictably with factors such as parameter count, training data, and computation.Quick
scikit-learnscikit-learn is an open Python library for classical machine learning, offering consistent tools for preprocessing, modelling, evaluation and pipelines.Quick
SeedanceSeedance is ByteDance’s generative video model family for producing coherent, multi-shot video sequences from text or image instructions and references.Quick
SeedreamSeedream is ByteDance’s image-generation model family, designed for prompt-based image creation and editing with strong text and visual fidelity.Quick
Segment AnythingSegment Anything is Meta’s foundation model for image segmentation, producing object masks from points, boxes, or other visual prompts.Quick
Self-attentionSelf-attention updates each position by comparing it with positions in the same sequence and mixing their information.3 min
Self-consistencySelf-consistency generates several independent reasoning paths for the same problem and selects the answer that appears most consistently across them.Quick
Self-hosting LLMsSelf-hosting means running a language model on a computer you control, so your prompts go to your own machine instead of a company's service.4 min
Self-reflectionSelf-reflection is when an AI model checks its own first answer, writes feedback on it, and then tries again using that feedback.3 min
Self-supervised learningSelf-supervised learning trains a model on raw data by hiding part of each example and asking the model to predict it, so the data supplies its own labels.5 min
Semantic cachingSemantic caching reuses earlier responses when a new request is sufficiently similar in meaning, reducing latency and model cost.Quick
Semantic KernelSemantic Kernel is an open-source Microsoft toolkit that connects AI models to your own code, so a model can ask your functions to do real work.4 min
Semantic searchSemantic search finds text by meaning, so a passage can match a question even when the two share no words.3 min
Semantic search engineA semantic search engine embeds queries and indexed content so it can retrieve material by meaning, not only matching words.Quick
Semi-supervised learningSemi-supervised learning trains a model on a few labelled examples plus many unlabelled ones, so the unlabelled data can help when answers are scarce.5 min
Sentence TransformersSentence Transformers is a Python library for creating and using text or image embeddings for similarity, retrieval and clustering.Quick
Sentence-BERTSentence-BERT adapts BERT into a model that produces useful sentence embeddings, making semantic similarity and retrieval substantially more efficient.Quick
Sentiment analysisSentiment analysis estimates attitudes or emotional polarity in text, such as positive, negative, or neutral, often for a specific target.Quick
Sentiment dashboardA sentiment dashboard classifies feedback, aggregates trends and preserves links to source material so teams can inspect what drove the summary.Quick
SEO content assistantAn SEO content assistant organises research, search intent and page structure into drafts that still require factual and editorial review.Quick
SGLangSGLang is open-source server software that runs large language models on GPUs and answers many requests quickly, partly by reusing work shared between prompts.4 min
SigLIPSigLIP is Google’s vision-language representation model, aligning images and text with a sigmoid-based training objective for retrieval and classification.Quick
SiriSiri is Apple’s voice assistant for device controls, personal requests and questions across Apple hardware and connected services.Quick
Slack botA Slack bot receives messages or events, applies defined logic or model calls and responds through authorised workspace actions.Quick
Small language modelsSmall-model research includes sub-billion-parameter language models for mobile deployment.3 min
SmolagentsSmolagents is a small open-source Python library from Hugging Face for building AI agents, including ones that act by writing short pieces of Python code.4 min
Social media assistantA social media assistant adapts approved messages into platform-specific drafts, schedules and response suggestions for human review. It remains subject to human review.Quick
SoftmaxSoftmax determines a probability for each possible class, and those probabilities add up to exactly 1.3 min
SoraSora is OpenAI’s generative video model, creating and transforming video from text, images, or existing footage through a prompt-driven workflow.Quick
spaCyspaCy is a production-oriented Python library for natural-language processing, including tokenisation, tagging, parsing and entity recognition. in practice.Quick
Sparse attentionSparse attention computes selected query-key connections instead of filling the entire attention matrix.3 min
SparsitySparsity means that many possible model values or computation paths are zero, absent, or inactive for a given input.4 min
Spec-driven developmentSpec-driven development begins with an explicit description of behavior, constraints, and acceptance criteria that guides human and AI implementation work.Quick
Speculative decodingSpeculative decoding uses a target model to evaluate guesses in parallel.4 min
Speech-to-text appA speech-to-text application accepts audio, transcribes speech and presents time-aligned text for review or further processing. It remains subject to human review.Quick
Speech-to-text modelsA speech-to-text model can transcribe a speech utterance into written characters.3 min
Spreadsheet assistantA spreadsheet assistant helps create formulas, clean tables, analyse values and explain calculations within bounded workbook context. It remains subject to human review.Quick
Spring AISpring AI brings model clients, vector stores, tools, structured outputs, and retrieval patterns into the Spring application ecosystem.Quick
SQuADSQuAD is a widely used reading-comprehension dataset whose human-written questions are answered from passages drawn from English Wikipedia articles.Quick
Stable AudioStable Audio is Stability AI’s generative audio model family for creating music, sound effects, and other audio from text prompts.Quick
Stable DiffusionStable Diffusion is Stability AI’s open latent-diffusion model family for generating and editing images from text and image prompts.Quick
Stable Diffusion 3Stable Diffusion 3 is a Stability AI image-model generation using a multimodal diffusion-transformer architecture to improve prompt adherence and typography.Quick
StagehandStagehand is an open-source framework from Browserbase for building AI agents that control a web browser, mixing plain-language instructions with regular browser automation code.Quick
StarCoderStarCoder is the BigCode collaboration’s open code-model family, trained on permissively licensed source code for generation and completion.Quick
State space modelsA continuous-time linear state space model uses x-dot = Ax + Bu for state evolution and y = Cx + Du for readout.4 min
Statistics for MLStatistics describes data and measures how far to trust a model's score, since every test set is only a sample of the cases the model will meet.4 min
Stochastic gradient descentStochastic gradient descent steps the weights using a slope estimated from one randomly picked example.3 min
Stop sequencesA stop sequence is a configured character string that stops output generation.3 min
Strands AgentsStrands Agents is a free AWS toolkit, released as open source under Apache 2.0, for making AI agents with Python or TypeScript, where the model itself plans the steps and picks the tools.5 min
Streaming responsesStreaming responses deliver a model’s output incrementally as it is generated, improving perceived speed and enabling responsive conversational interfaces.Quick
StreamlitStreamlit is a Python framework for turning data scripts and model workflows into interactive web applications. in practice.Quick
Streamlit chatbotA Streamlit chatbot uses Streamlit’s Python interface components to wrap a model call in a simple interactive web application.Quick
Structured data extractionA structured data extraction workflow converts unstructured documents into a defined schema, then validates required fields and uncertain values.Quick
Structured outputsStructured outputs constrain a model response to a declared schema so software receives expected fields and types instead of free-form prose.3 min
Study tutorA study tutor explains material, asks diagnostic questions and adjusts practice to the learner while avoiding simply completing assessed work.Quick
Sub-agentsA sub-agent is a helper AI that a main agent sends off to do one side task in its own workspace, then report back a short result.3 min
SummarizationSummarization is when a model turns a long text into a shorter version that keeps the important information.3 min
SunoSuno generates complete songs from text descriptions or lyrics, combining composition, vocals and production in a consumer music-creation interface.Quick
SuperintelligenceSuperintelligence is a hypothetical form of intelligence that substantially exceeds the best human abilities across most cognitively demanding fields.Quick
Supervised fine-tuningSupervised fine-tuning keeps training an already-trained model on example questions paired with good answers, so it learns to answer that way.4 min
Supervised learningSupervised learning trains a model on examples that already carry the right answer, called a label, so it can predict that answer for new examples that do not.4 min
Support vector machinesA support vector machine separates two classes with the boundary that leaves the widest possible gap to the nearest training points of each class.5 min
SWE-benchSWE-bench asks a system to generate a repository patch for a real GitHub issue.3 min
SycophancySycophancy is when an assistant favors matching a user's beliefs over a truthful response.3 min
Symbolic AISymbolic AI writes knowledge down as readable symbols and if-then rules, then reaches answers by applying logic and search to them.4 min
SynthesiaSynthesia creates presenter-led videos from scripts using synthetic avatars, voices, templates and enterprise publishing controls. for creative production.Quick
Synthetic dataSynthetic data is artificial data built from seed data so some patterns of that seed can be reused.4 min
Synthetic data generatorA synthetic data generator creates artificial examples under defined constraints, then checks coverage, realism, privacy and downstream usefulness.Quick
System promptA system prompt is high-priority context that sets an AI assistant's role, boundaries, style, and response rules before the user's request is handled.3 min
t-SNEt-SNE is a nonlinear visualization method that places similar high-dimensional examples nearby in two or three dimensions, but global distances may be misleading.Quick
T5T5 is Google’s text-to-text transformer framework, recasting many language tasks so both inputs and outputs use a unified text format.Quick
TemperatureTemperature enters softmax by dividing each logit by T.3 min
Tensor parallelismTensor parallelism splits individual tensor operations and parameter matrices across devices, allowing them to collaborate within the same model layer.Quick
TensorFlowTensorFlow is Google’s open-source machine-learning framework for building, training, and deploying models across servers, browsers, mobile devices, and specialized hardware.Quick
TensorRT-LLMTensorRT-LLM is NVIDIA's open-source library for compiling and running large language models efficiently on NVIDIA GPUs, aimed at fast, low-cost serving in production.Quick
Test setA test set is a reserved group of labeled examples used for a final check of a model on data it did not train on.3 min
Test-time computeTest-time compute lets a language model improve its output by using more computation at test time.3 min
Text classificationText classification assigns predefined labels to documents or passages, supporting tasks such as topic detection, moderation, routing, and intent recognition.Quick
Text Generation InferenceText Generation Inference (TGI) is Hugging Face's open-source toolkit for deploying and serving large language models behind an API, with batching and streaming built in.Quick
Text-to-image modelsA text-to-image model can use text as a conditioning input to an image generator.3 min
Text-to-speech modelsA text-to-speech model synthesizes speech directly from text.3 min
Text-to-SQLText-to-SQL systems translate natural-language questions into database queries, often adding database schema context, validation, permissions, and execution safeguards.Quick
Text-to-video modelsGiven a text prompt, a text-to-video model generates a video.3 min
The StackThe Stack is BigCode’s large source-code dataset, assembled for training code models with tools for filtering and attribution.Quick
ThroughputThroughput is the amount of completed inference work divided by the time spent measuring it.3 min
Time series forecastingTime series forecasting predicts future values of something measured over time, such as sales or electricity demand, from the patterns in its past.5 min
Time to first tokenTime to first token is how long a model takes to start its reply after receiving a request, a key measure of how responsive an AI app feels.Quick
timmtimm is a PyTorch library of image models, pretrained weights, training utilities and reproducible computer-vision recipes. in practice.Quick
Together AITogether AI provides hosted inference, fine-tuning and GPU infrastructure for open and custom generative AI models. in production.Quick
Token costsAI model APIs charge by the token, with separate prices for the text you send in and the text the model writes back.4 min
TokenizationTokenization turns text into a sequence of vocabulary IDs that a model can process, then reverses generated IDs back into readable text.3 min
TokensTokens are chunks of text processed by text-generation and embedding models.3 min
Tokens per secondTokens per second reports how many tokens a server returns in a set amount of time.3 min
Tool useTool use is a loop where a model proposes a call, a program outside the model runs it, and the result is written back.5 min
Top-k samplingTop-k sampling keeps the k highest-probability next-token candidates, rescales their probabilities and samples one of them.3 min
Top-p samplingTop-p keeps the smallest set of most probable tokens whose probabilities add up to at least p.3 min
TPUsTPUs are chips Google designed to speed up the maths of machine learning, rented out through Google Cloud.5 min
TrainingTraining is the repeated loop that nudges a model's internal numbers so its predictions get closer to the right answers in its examples.5 min
Training dataTraining data is the set of examples a machine learning model studies to learn its patterns, kept apart from the examples used to test it.5 min
Transfer learningTransfer learning reuses representations learned on a source task to improve learning on a different target task.3 min
TransformersA transformer is a neural network that uses attention to decide which parts of a sequence matter to one another.3 min
Translation appA translation application converts text or speech between languages while preserving meaning, terminology and formatting where possible. It remains subject to human review.Quick
Tree of thoughtsTree of Thoughts is a prompting and search approach that explores several candidate reasoning paths, evaluates them, and selects promising branches.Quick
Triton Inference ServerNVIDIA Triton Inference Server serves models from multiple frameworks with batching, model management and HTTP or gRPC endpoints.Quick
TRLTRL is Hugging Face’s library for post-training language models with supervised fine-tuning, preference optimisation and reinforcement-learning methods. for practical reuse.Quick
TruLensTruLens is an open-source framework for evaluating and tracing language-model applications, especially retrieval quality, groundedness, and feedback signals.Quick
TruthfulQATruthfulQA evaluates whether language models avoid producing common human misconceptions when answering questions designed to elicit plausible falsehoods.Quick
Turing testThe Turing test asks whether a judge, chatting by text with a hidden person and a hidden machine, can reliably tell which one is the machine.5 min
txtaitxtai is an open-source Python framework for semantic search, embeddings databases, retrieval workflows, agents, and language-model pipelines over unstructured data.Quick
U-NetU-Net combines upsampled decoder features with matching high-resolution features from a contracting path.4 min
UdioUdio is a generative music service for creating and extending songs from prompts, lyrics and uploaded audio within an editing workflow.Quick
UltralyticsUltralytics provides computer-vision tooling best known for training, evaluating and deploying YOLO object-detection and segmentation models. in practice.Quick
UMAPUMAP reduces high-dimensional data using a graph of local relationships, often producing useful visualizations whose apparent spacing should not be treated as exact geometry.Quick
UnderfittingUnderfitting occurs when a model is too simple or insufficiently trained to capture important patterns, producing weak results even on training data.Quick
UnslothUnsloth is an open-source tool for training and running AI language models on your own computer, built to fine-tune them faster and with less graphics-card memory.4 min
UnstructuredUnstructured partitions PDFs, office files, HTML and other documents into typed elements for downstream search, retrieval and AI pipelines.Quick
Unsupervised learningUnsupervised learning trains a model on examples that have no answers attached, so it finds structure on its own, such as groups of similar items or points that do not fit.5 min
V-JEPAV-JEPA is Meta research on video representation learning by predicting abstract visual features rather than reconstructing every image pixel.Quick
v0v0 is Vercel’s AI interface and application builder for generating web UI, code and deployable projects through conversation.Quick
Validation setA validation set is held-out data used during development to compare model choices without training on the same examples.3 min
Vanishing and exploding gradientsVanishing and exploding gradients happen when the error signal shrinks toward zero or grows very large as it passes back through many layers, which makes deep networks hard to train.Quick
Variational autoencodersA variational autoencoder pairs a deep latent-variable model with a corresponding approximate inference model.4 min
Vector databasesA vector database stores vectors and returns the saved records closest to a query vector, often with a score and any extra fields you ask for.4 min
VectorsA vector is an ordered list of numbers that you can also picture as an arrow in space. Machine learning stores examples, words and images this way.4 min
VeoVeo is Google’s family of generative video models, producing video from text or image prompts with controls for style, composition, and motion.Quick
Veo 3Veo 3 is a generation of Google’s video model family that creates prompted video and can generate synchronized audio alongside visuals.Quick
Vercel AI SDKThe Vercel AI SDK is an open-source TypeScript toolkit, free to use, that lets apps talk to AI models from many companies through the same code.4 min
Vertex AIVertex AI is Google Cloud’s platform for choosing AI models, tuning them on your own examples and running models and agents for real users.5 min
VespaVespa is a serving engine for large-scale search, recommendation and ranking that combines text, structured fields and vector retrieval.Quick
Vibe codingVibe coding is an informal development style in which a person describes desired software behavior while an AI generates and revises much of the code.Quick
Video summarizerA video summarizer combines transcript and visual analysis to identify key moments, topics and a shorter account of the recording.Quick
Virtual assistantsVirtual assistants help users complete tasks through conversational commands, often combining language understanding with search, applications, devices, or external services.Quick
Vision transformersA vision transformer represents an image as a sequence of patches and processes that sequence with a transformer.3 min
Vision-language modelsA vision-language model can take visual data and text as input and produce text as output.4 min
Vision-language-action modelsA vision-language-action model maps visual observations and a language instruction to robot actions.4 min
vLLMvLLM is open-source software that runs large language models on a server and answers many requests at once, using GPU memory carefully so fewer bytes go to waste.4 min
VocabularyA model vocabulary supplies the mapping used to convert text into an ID sequence and back.3 min
Voice agentsA voice agent is an AI app you talk to out loud; it listens, works out what you want, can use tools, and answers in speech.5 min
Voice assistantA voice assistant combines speech recognition, a reasoning or dialogue layer and speech synthesis for real-time spoken interaction.Quick
Voice cloningVoice cloning uses a speech model to produce new speech that sounds like a particular person, usually learned from a short recording of their voice.Quick
Voyage embeddingsVoyage embeddings are commercial embedding models optimized for retrieval across general text, code, finance, law, and other specialized domains.Quick
WanWan is Alibaba’s generative video model family, designed to create and edit video from text, images, and other conditioning inputs.Quick
WatermarkingAI watermarking embeds detectable signals in generated content or model outputs to help identify origin, though robustness and reliability vary by method.Quick
WaymoWaymo operates autonomous ride-hailing vehicles using cameras, lidar, radar, maps and driving software within selected service areas. in practical systems.Quick
WeaviateWeaviate is a vector database for semantic and hybrid search, with schema, filtering and optional integrated model capabilities.Quick
Web scraping agentA web scraping agent navigates selected sites, extracts defined fields and returns structured data while respecting access and rate limits.Quick
WebGPUWebGPU is a web standard that exposes modern GPU computation and graphics capabilities to browsers, including acceleration for local AI inference.Quick
WeightsTrainable weights are model parameters that can be learned from data.3 min
Weights & BiasesWeights & Biases supports experiment tracking, model and application evaluation, artifact management and collaboration across machine-learning teams. in production.Quick
WhisperWhisper is OpenAI’s open-source speech-recognition model, trained for multilingual transcription, translation into English, and robust audio understanding across varied recordings.Quick
Word2VecWord2Vec is a technique for learning vector representations of words from nearby context, making semantic relationships measurable through geometry.Quick
Workflow automation with LLMsWorkflow automation with LLMs adds language understanding or generation to a deterministic process while keeping triggers, permissions and actions explicit.Quick
World modelsA world model predicts how an environment may change after an action, so an agent can evaluate possible futures before acting.4 min
xAI APIThe xAI API gives developers hosted access to Grok models for text, vision, tool use and structured application workflows.Quick
XGBoostXGBoost is an open gradient-boosted tree library designed for efficient, accurate supervised learning on structured data. in practice.Quick
YOLOYOLO, or You Only Look Once, is a family of real-time computer-vision models that detect and classify objects in a single pass.Quick
Zapier AgentsZapier Agents lets users configure AI agents that work with Zapier-connected applications to research information and perform approved actions.Quick
Zero-shot learningZero-shot learning applies a model to a task without task-specific examples, relying on prior training and a natural-language instruction or label description.Quick