graphics processing unit communication bottleneck
A graphics processing unit communication bottleneck occurs when moving or synchronizing data between processors takes longer than the useful computation it supports.
Search clear AI terminology definitions, acronyms, related concepts and the reviewed videos that explain them.
Showing 441–460 of 705 terms
Clear filtersA graphics processing unit communication bottleneck occurs when moving or synchronizing data between processors takes longer than the useful computation it supports.
A grounded AI agent bases its answers and actions on relevant, verifiable information rather than relying only on model memory or unsupported inference.
Hard red-list language model watermarking forbids one keyed subset of tokens and allows sampling only from the complementary preferred set.
AI hardware memory capacity is the amount of fast memory available to hold model weights, runtime state and active context during model execution.
Hardware-aware code optimization adapts algorithms, memory access, scheduling, and communication to the real characteristics of the target processor and system.
A headless application programming interface exposes an application's data or operations independently of its graphical user interface so another program can use them directly.
A heterogeneous robot team combines different robot types so their distinct physical or sensing capabilities can contribute to one task.
High-stakes AI is AI used where errors can materially affect health, rights, safety, finances or other important outcomes.
Human-directed AI research keeps people responsible for hypotheses and priorities while AI systems execute bounded experimental work.
Human-out-of-the-loop AI is a workflow in which an AI system executes defined actions without human approval at each operational step.
AI hypothesis generation uses AI to propose testable explanations or strategies from observations and data.
Idempotency means that repeating the same operation with the same identity has the same intended effect as performing it once.
Image-to-code uses a multimodal model to turn a screenshot or visual reference into a software implementation.
Image-to-three-dimensional reconstruction estimates a three-dimensional scene or object from one or more two-dimensional images.
Imitation learning trains an agent to perform a task by learning from demonstrations of behavior rather than relying only on an explicitly written reward function.
An in-application AI assistant is an assistant embedded inside a particular product and designed to help users operate that product with access to its approved context and actions.
AI incident response is the organized process for detecting, containing, investigating, recovering from, and learning from harmful model or agent behavior.
Independent AI evaluation tests an AI system using evidence and methods not controlled solely by the model provider.
AI inference autoscaling automatically adjusts model-serving capacity in response to workload demand and performance targets.
AI inference cost is the expense of running a trained model to process requests and produce outputs.
Follow general terms into specialised sub-terms. Select any node to open its definition.