Groq
An AI infrastructure company and model platform that provides ultra-low-latency inference services for large language models and other AI models, optimized through its custom Language Processing Unit (LPU) architecture.
Explore more about Model Providers
Related terms
A software development kit provided by Groq that enables developers to integrate, manage, and interact with Groq's AI models and high-speed inference services through programmatic APIs.
Inference APIAn API that enables applications to send input data to a deployed AI model and receive generated predictions or outputs, providing programmatic access to inference capabilities without managing the underlying model infrastructure.
Model ServingThe process of making an AI model available for inference by deploying it behind an interface that accepts requests and returns predictions or generated outputs.