Cyber Security
Generative AI
Data Science
Blog About
WhatsApp Free Demo
Generative AI · Glossary

Inference

The process of running a trained AI model to generate a prediction or output, as opposed to training the model itself.

When you send a prompt to an LLM API and get a response back, that's inference. Inference cost and speed are major practical considerations when building AI-powered products.

Want to actually work with Inference? This concept is covered hands-on in our AI Engineering Foundations program — not just defined, but practiced.
View AI Engineering Foundations
← Back to full glossary
Chat with us
WhatsApp Call Free Demo