Inference
Technical processes and infrastructure for running AI models.
Definitions (2)
The technical execution and infrastructure required to operate AI models (particularly large language models), including model hosting, compute resources, and serving mechanisms; the strategy discusses feasibility of providing inference services on state-owned servers (in cooperation with HPC.nrw) to support academia and administration.
The process of applying a trained model to new data to produce outputs or predictions.
Related Terms
Inference (Inferenz)
The runtime process of executing AI models to produce outputs....
High-performance computing (HPC)
Advanced computational infrastructure supporting AI research and deployment....
Infrastructure
Compute and storage resources like HPC and cloud services....
Infrastructure (AI layered model)
Compute, data centers and national platforms supporting AI....
Digital Sovereignty (Digitale Souveränität)
Autonomy and control over data, infrastructure, and AI tools....