Aller au contenu
//
AI RESOURCE
HUB
Ressources
APIs
Prompts
Tutoriels
Glossaire
English
简体中文
日本語
Español
Français
Deutsch
한국어
[MENU]
Accueil
/
Glossaire
/
Inference
Inference
The process of running a trained model to produce outputs, as opposed to training it.
Ressources associées
vLLM High-Performance Inference Framework
A high-throughput inference and serving framework for large language models that significantly improves deployment efficiency via PagedAttention.
Groq High-Speed Inference API
An inference service built on LPU hardware that delivers ultra-low-latency API access for open-source large language models.