Skip to main content
Helicone OSS LLM Observability

Get Multimodal Models

Returns all available multimodal models supported by Helicone AI Gateway (OpenAI-compatible endpoint)
1 min read
get/v1/models/multimodal
Request example
Response
get/v1/models/multimodal

This endpoint returns a list of all multimodal AI models supported by the Helicone AI Gateway. Multimodal models are those that support more than one input modality (e.g., text + images) or more than one output modality. This is an OpenAI-compatible endpoint that follows the same response format as OpenAI's /v1/models endpoint.

Use this endpoint to discover which multimodal models are available for routing through the AI Gateway.

Endpoint URL#

What Makes a Model Multimodal?#

A model is considered multimodal if it meets either of these criteria:

  • Multiple Input Modalities: Accepts more than one type of input (e.g., text, images, audio)
  • Multiple Output Modalities: Produces more than one type of output (e.g., text, images, audio)

Example Request#

Example Response#

Use Cases#

  • OpenAI Compatibility: Use this endpoint as a drop-in replacement for OpenAI's /v1/models endpoint with multimodal filtering
  • Multimodal Model Discovery: Discover which multimodal models are available through Helicone AI Gateway
  • Vision/Audio Applications: Find models that support image or audio inputs for your applications
  • Integration Testing: Verify multimodal model availability for your applications

Responses

application/json
200
Ok
objectenum<string>required#
Options:list
dataobject[]required#
Show child attributes
idstringrequired#
objectenum<string>required#
Options:model
creatednumberrequired#
owned_bystringrequired#