Saltar al contenido
//
AI RESOURCE
HUB
Recursos
APIs
Prompts
Tutoriales
Glosario
English
简体中文
日本語
Español
Français
Deutsch
한국어
[MENÚ]
Inicio
/
Glosario
/
Inference
Inference
The process of running a trained model to produce outputs, as opposed to training it.
Recursos relacionados
vLLM High-Performance Inference Framework
A high-throughput inference and serving framework for large language models that significantly improves deployment efficiency via PagedAttention.
Groq High-Speed Inference API
An inference service built on LPU hardware that delivers ultra-low-latency API access for open-source large language models.