Zum Inhalt springen
//
AI RESOURCE
HUB
Ressourcen
APIs
Prompts
Tutorials
Glossar
English
简体中文
日本語
Español
Français
Deutsch
한국어
[MENÜ]
Startseite
/
Glossar
/
Inference
Inference
The process of running a trained model to produce outputs, as opposed to training it.
Verwandte Ressourcen
vLLM High-Performance Inference Framework
A high-throughput inference and serving framework for large language models that significantly improves deployment efficiency via PagedAttention.
Groq High-Speed Inference API
An inference service built on LPU hardware that delivers ultra-low-latency API access for open-source large language models.