Multimodal AI

Multimodal AI refers to systems that can process or generate information across more than one type of data, such as text, images, audio, video or computer interfaces.

A multimodal model can combine information across these modalities rather than treating each as an isolated task.

← Back to Index