Glossary · Infrastructure
Inference endpoint
An inference endpoint is a network address, usually an HTTPS URL, where a deployed model accepts requests and returns predictions or generated outputs to calling applications.
Inference endpoint sits in the Infrastructure part of the Agentik {OS} glossary, which defines the words used to build and run AI agent systems.