> ## Documentation Index
> Fetch the complete documentation index at: https://portkey-docs-docs-prisma-airs-updates.mintlify.site/llms.txt
> Use this file to discover all available pages before exploring further.

# Multimodal Capabilities

<Info>
  This feature is available on all Prisma AIRS AI Gateway plans.
</Info>

The Gateway is your unified interface for **multimodal models**, along with chat, text, and embedding models.

Using the Gateway, you can call `vision`, `audio (text-to-speech & speech-to-text)`, `image generation` and other multimodal models from multiple providers (like `OpenAI`, `Anthropic`, `Stability AI`, etc.) — all using the familiar OpenAI signature.

#### Explore the AI Gateway's Multimodal capabilities below:

<Card title="Vision" href="/aigw/product/ai-gateway/multimodal-capabilities/vision" />

<Card title="Image Generation" href="/aigw/product/ai-gateway/multimodal-capabilities/image-generation" />

<Card title="Function Calling" href="/aigw/product/ai-gateway/multimodal-capabilities/function-calling" />

<Card title="Speech-to-Text" href="/aigw/product/ai-gateway/multimodal-capabilities/speech-to-text" />

<Card title="Text-to-Speech" href="/aigw/product/ai-gateway/multimodal-capabilities/text-to-speech" />
