Multimodal AI
Multimodal AI refers to systems that can process or generate information across more than one type of data, such as text, images, audio, video or computer interfaces.
A multimodal model can combine information across these modalities rather than treating each as an isolated task.
← Back to Index