본문으로 건너뛰기
//
AI RESOURCE
HUB
리소스
API
프롬프트
튜토리얼
용어집
English
简体中文
日本語
Español
Français
Deutsch
한국어
[메뉴]
홈
/
용어집
/
Inference
Inference
The process of running a trained model to produce outputs, as opposed to training it.
관련 리소스
vLLM High-Performance Inference Framework
A high-throughput inference and serving framework for large language models that significantly improves deployment efficiency via PagedAttention.
Groq High-Speed Inference API
An inference service built on LPU hardware that delivers ultra-low-latency API access for open-source large language models.