Multimodal AI

Multimodal AI handles more than one type of input or output - text, image, audio, video - in a single model, so you can prompt with an image and get text, or vice versa.

In practice, Multimodal AI shows up across many AI tools. Below are 8 tools where the idea is useful - open any to see it applied.

← Back to AI Glossary