Glossary · Models
Multimodal model
A multimodal model is an AI model that can process and often generate more than one type of data, such as text, images, audio, and video, within a single system.
Multimodal model sits in the Models part of the Agentik {OS} glossary, which defines the words used to build and run AI agent systems.
Also called Multimodal LLM, MLLM.