Fundamentals
What is Inference?
Running a trained model to get an output — the part that happens when you use an AI tool.
Inference is using a trained model to produce an output — what happens each time you send a prompt and get a response.
Inference cost and speed (latency) are the main factors in how much an AI product costs to run.
Related terms
Find fundamentals tools
Browse hand-reviewed AI tools — compare on pricing, features and real ratings.