On-device AI
Also called: On-device Artificial Intelligence
AI capabilities that run directly on local devices, reducing reliance on remote servers and enabling lower-latency, privacy-focused inference.
Explore more about Deployment
Related terms
The process of reducing the precision of a model's weights and activations to lower memory footprint and speed up inference.
Model ServingThe process of making an AI model available for inference by deploying it behind an interface that accepts requests and returns predictions or generated outputs.