Inference
Running a trained model to produce an output, as opposed to training it.
Last updated
Related terms
Latency
The time between sending a request and receiving a model's response.
Parameters
The internal numbers a model learns during training; their count is a rough measure of model size.
Large Language Model
A neural network trained on massive text corpora to predict the next token, enabling it to generate and understand langu...
Keep exploring
Browse the full AI glossary or compare AI tools that use this technology.