Release Note

Last verified 28 Apr 2026

We now support multimodal models for serverless inference. Multimodal models process and generate content across multiple data types, including images, audio, video, and text, thus enabling a much broader range of real-world applications, including document intelligence, voice agents, content generation, and accessibility tools. For more information, see Use Multimodal Inference.

We can't find any results for your search.

Try using different keywords or simplifying your search terms.