Skip to content
DocsTry Aspire
DocsTry

OpenAI

Classstaticnet10.0
📦 Aspire.Hosting.Foundry v13.5.3-preview.1.26425.3
Models published by OpenAI.
namespace Aspire.Hosting.Foundry;
public static class OpenAI
{
// ...
}
CodexMinistatic
codex-mini is a fine-tuned variant of the o4-mini model, designed to deliver rapid, instruction-following performance for developers working in CLI workflows. Whether you're automating shell commands, editing scripts, or refactoring repositories, Codex-Min
ComputerUsePreviewstatic
computer-use-preview is the model for Computer Use Agent for use in Responses API. You can use computer-use-preview model to get instructions to control a browser on your computer screen and take action on a user's behalf.
Davinci002static

Azure Direct Models

Direct from Azure models are a select portfolio curated for their market-differentiated capabilities:

  • Secure and managed by Microsoft: Purchase and manage models directly through Azure with a single license, consistent support, and no third-party dependencies, backed by Azure's enterprise-grade infrastructure.

  • Streamlined operations: Benefit from unified billing, governance, and seamless PTU portability across models hosted on Azure - all as part of one Azure AI Foundry platform.

  • Future-ready flexibility: Access the latest models as they become available, and easily test, deploy, or switch between them within Azure AI Foundry; reducing integration effort.

  • Cost control and optimization: Scale on demand with pay-as-you-go flexibility or reserve PTUs for predictable performance and savings.

Learn more about Direct from Azure models.

Key capabilities

About this model

The provider has not supplied this information.

Key model capabilities

Davinci-002 supports fine-tuning, allowing developers and businesses to customize the model for specific applications.

Use cases

See Responsible AI for additional considerations for responsible use.

Key use cases

The provider has not supplied this information.

Out of scope use cases

The provider has not supplied this information.

Pricing

Pricing is based on a number of factors, including deployment type and tokens used. See pricing details here.

Technical specs

The provider has not supplied this information.

Training cut-off date

This model supports 16384 max input tokens and training data is up to Sep 2021.

Training time

The provider has not supplied this information.

Input formats

Your training data and validation data sets consist of input and output examples for how you would like the model to perform. The training and validation data you use must be formatted as a JSON Lines (JSONL) document in which each line represents a single prompt-completion pair.

Output formats

The provider has not supplied this information.

Supported languages

The provider has not supplied this information.

Sample JSON response

The provider has not supplied this information.

Model architecture

Davinci-002 is the latest version of Davinci, a gpt-3 based model.

Long context

This model supports 16384 max input tokens.

Optimizing model performance

The provider has not supplied this information.

Additional assets

Learn more at https://learn.microsoft.com/azure/cognitive-services/openai/concepts/models

Training disclosure

Training, testing and validation

The provider has not supplied this information.

Distribution

Distribution channels

The provider has not supplied this information.

More information

The provider has not supplied this information.

Gpt41static
gpt-4.1 outperforms gpt-4o across the board, with major gains in coding, instruction following, and long-context understanding
Gpt41Ministatic
gpt-4.1-mini outperform gpt-4o-mini across the board, with major gains in coding, instruction following, and long-context handling
Gpt41Nanostatic
gpt-4.1-nano provides gains in coding, instruction following, and long-context handling along with lower latency and cost
Gpt4ostatic
OpenAI's most advanced multimodal model in the gpt-4o family. Can handle both text and image inputs.
Gpt4oMinistatic
An affordable, efficient AI solution for diverse text and image tasks.
Gpt4oMiniTranscribestatic
A highly efficient and cost effective speech-to-text solution that deliverables reliable and accurate transcripts.
Gpt4oMiniTtsstatic
An advanced text-to-speech solution designed to convert written text into natural-sounding speech.
Gpt4oTranscribestatic
A cutting-edge speech-to-text solution that deliverables reliable and accurate transcripts.
Gpt4oTranscribeDiarizestatic
A cutting-edge speech-to-text solution that deliverables reliable and accurate transcripts; now equipped with diarization support aka identifying different speakers through the transcription.
Gpt5static
gpt-5 is designed for logic-heavy and multi-step tasks.
Gpt51static
gpt-5.1 is designed for logic-heavy and multi-step tasks.
Gpt51Chatstatic
gpt-5.1-chat (preview) is an advanced, natural, multimodal, and context-aware conversations for enterprise applications.
Gpt51Codexstatic
gpt-5.1-codex is designed for steerability, front end development, and interactivity.
Gpt51CodexMaxstatic
gpt-5.1-codex-max is agentic coding model designed to streamline complex development workflows with advanced efficiency
Gpt51CodexMinistatic
gpt-5.1-codex-mini is designed for steerability, front end development, and interactivity.
Gpt52static
GPT-5.2 is engineered for enterprise agent scenarios—delivering structured, auditable outputs, reliable tool use, and governed integrations.
Gpt52Chatstatic
gpt-5.2-chat (preview) is an advanced, natural, multimodal, and context-aware conversations for enterprise applications.
Gpt52Codexstatic
gpt-5.2-codex is designed for steerability, front end development, and interactivity.
Gpt53Chatstatic
gpt-5.3-chat (preview) is an advanced, natural, multimodal, and context-aware conversations for enterprise applications.
Gpt53Codexstatic
gpt-5.3-codex is designed for steerability, front end development, and interactivity.
Gpt54static
GPT‑5.4 is OpenAI’s most capable frontier model, built to deliver faster, more reliable results for complex professional work.
Gpt54Ministatic
GPT‑5.4‑mini is a compact, cost‑efficient model designed for reliable performance across high‑volume, everyday AI workloads.
Gpt54Nanostatic
GPT‑5.4‑nano is a lightweight, ultra‑efficient model designed for low‑latency, cost‑effective tasks at massive scale.
Gpt54Prostatic
GPT‑5.4-Pro is OpenAI's most capable frontier model, built to deliver faster, more reliable results for complex professional work.
Gpt55static
GPT‑5.5 is OpenAI’s most capable frontier model, built to deliver faster, more reliable results for complex professional work.
Gpt56Lunastatic
GPT‑5.6-luna is OpenAI's most capable frontier model, built to deliver faster, more reliable results for complex professional work.
Gpt56Solstatic
GPT‑5.6-sol is OpenAI's most capable frontier model, built to deliver faster, more reliable results for complex professional work.
Gpt56Terrastatic
GPT‑5.6-terra is OpenAI's most capable frontier model, built to deliver faster, more reliable results for complex professional work.
Gpt5Chatstatic
gpt-5-chat (preview) is an advanced, natural, multimodal, and context-aware conversations for enterprise applications.
Gpt5Codexstatic
gpt-5-codex is designed for steerability, front end development, and interactivity.
Gpt5Ministatic
gpt-5-mini is a lightweight version for cost-sensitive applications.
Gpt5Nanostatic
gpt-5-nano is optimized for speed, ideal for applications requiring low latency.
Gpt5Prostatic
gpt-5-pro uses more compute to think harder and provide consistently better answers.
GptAudiostatic
Best suited for rich, asynchronous audio input/output interactions, such as creating spoken summaries from text.
GptAudio15static
A new S2S (speech to speech) model with improved instruction following.
GptAudioMinistatic
Best suited for rich, asynchronous audio input/output interactions, such as creating spoken summaries from text.
GptChatLateststatic
gpt-chat-latest (preview) is an advanced, natural, multimodal, and context-aware conversations for enterprise applications.
GptImage15static
An efficient AI solution for diverse text and image tasks, including high quality, and editing scenarios
GptImage1Ministatic
An efficient AI solution for diverse text and image tasks, including high quality, cheap text to image generation
GptImage2static
An efficient AI solution for diverse text and image tasks, including high quality, and editing scenarios
GptLiveTranscribestatic
A new real-time speech-to-text (STT) model with enhanced transcription accuracy and low-latency streaming capabilities.
GptOss120bstatic
Push the open model frontier with GPT-OSS models, released under the permissive Apache 2.0 license, allowing anyone to use, modify, and deploy them freely.
GptOss20bstatic
Push the open model frontier with GPT-OSS models, released under the permissive Apache 2.0 license, allowing anyone to use, modify, and deploy them freely.
GptRealtimestatic
A new S2S (speech to speech) model with improved instruction following.
GptRealtime15static
A new S2S (speech to speech) model with improved instruction following.
GptRealtime2static
Gpt‑realtime‑2 is a next‑generation speech‑to‑speech reasoning model that processes live audio input and generates audio responses with built‑in reasoning, enabling low‑latency conversational voice interactions.
GptRealtime21static
Gpt‑realtime‑2.1 is a next‑generation speech‑to‑speech reasoning model that processes live audio input and generates audio responses with built‑in reasoning, enabling low‑latency conversational voice interactions.
GptRealtime21Ministatic
Gpt‑realtime‑2.1‑mini is a next‑generation speech‑to‑speech reasoning model that processes live audio input and generates audio responses with built‑in reasoning, enabling low‑latency conversational voice interactions.
GptRealtimeMinistatic
gpt-realtime-mini is a smaller version of gpt-realtime S2S (speech to speech) model built on chive architecture. This model excels at instruction following and is optimized for cost efficiency.
GptRealtimeTranslatestatic
Gpt‑realtime‑translate is a low‑latency streaming model that converts spoken audio into translated output in real time, enabling live cross‑language communication within voice applications.
GptRealtimeWhisperstatic
A new STT (speech to text) model with realtime capability.
GptTranscribestatic
A new real-time speech-to-text (STT) model with enhanced transcription accuracy and low-latency streaming capabilities.
O1static
Focused on advanced reasoning and solving complex problems, including math and science tasks. Ideal for applications that require deep contextual understanding and agentic workflows.
O3static
o3 includes significant improvements on quality and safety while supporting the existing features of o1 and delivering comparable or better performance.
O3DeepResearchstatic
The o3 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o1-pro model uses more compute to think harder and provide consistently better answers.
O3Ministatic
o3-mini includes the o1 features with significant cost-efficiencies for scenarios requiring high performance.
O3Prostatic
The o3 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o1-pro model uses more compute to think harder and provide consistently better answers.
O4Ministatic
o4-mini includes significant improvements on quality and safety while supporting the existing features of o3-mini and delivering comparable or better performance.
TextEmbedding3Largestatic
Text-embedding-3 series models are the latest and most capable embedding model from OpenAI.
TextEmbedding3Smallstatic
Text-embedding-3 series models are the latest and most capable embedding model from OpenAI.
TextEmbeddingAda002static

Direct from Azure models

Direct from Azure models are a select portfolio curated for their market-differentiated capabilities:

  • Secure and managed by Microsoft: Purchase and manage models directly through Azure with a single license, consistent support, and no third-party dependencies, backed by Azure's enterprise-grade infrastructure.

  • Streamlined operations: Benefit from unified billing, governance, and seamless PTU portability across models hosted on Azure - all part of Microsoft Foundry.

  • Future-ready flexibility: Access the latest models as they become available, and easily test, deploy, or switch between them within Microsoft Foundry; reducing integration effort.

  • Cost control and optimization: Scale on demand with pay-as-you-go flexibility or reserve PTUs for predictable performance and savings.

Learn more about Direct from Azure models.

Key capabilities

About this model

text-embedding-ada-002 outperforms all the earlier embedding models on text search, code search, and sentence similarity tasks and gets comparable performance on text classification.

Key model capabilities

  • Text search

  • Code search

  • Sentence similarity tasks

  • Text classification

Note: this model can be deployed for inference, specifically for embeddings, but cannot be finetuned.

Use cases

See Responsible AI for additional considerations for responsible use.

Key use cases

The provider has not supplied this information.

Out of scope use cases

The provider has not supplied this information.

Pricing

Pricing is based on a number of factors, including deployment type and tokens used. See pricing details here.

Technical specs

The provider has not supplied this information.

Training cut-off date

The provider has not supplied this information.

Training time

The provider has not supplied this information.

Input formats

The provider has not supplied this information.

Output formats

The provider has not supplied this information.

Supported languages

The provider has not supplied this information.

Sample JSON response

The provider has not supplied this information.

Model architecture

The provider has not supplied this information.

Long context

The provider has not supplied this information.

Optimizing model performance

The provider has not supplied this information.

Additional assets

The provider has not supplied this information.

Training disclosure

Training, testing and validation

The provider has not supplied this information.

Distribution

Distribution channels

The provider has not supplied this information.

More information

The provider has not supplied this information.

Ttsstatic

Direct from Azure models

Direct from Azure models are a select portfolio curated for their market-differentiated capabilities:

  • Secure and managed by Microsoft: Purchase and manage models directly through Azure with a single license, consistent support, and no third-party dependencies, backed by Azure's enterprise-grade infrastructure.

  • Streamlined operations: Benefit from unified billing, governance, and seamless PTU portability across models hosted on Azure - all part of Microsoft Foundry.

  • Future-ready flexibility: Access the latest models as they become available, and easily test, deploy, or switch between them within Microsoft Foundry; reducing integration effort.

  • Cost control and optimization: Scale on demand with pay-as-you-go flexibility or reserve PTUs for predictable performance and savings.

Learn more about Direct from Azure models.

Key capabilities

About this model

TTS is a model that converts text to natural sounding speech. TTS is optimized for realtime or interactive scenarios. For offline scenarios, TTS-HD provides higher quality. The API supports six different voices.

Key model capabilities

  • TTS: optimized for speed.

  • TTS-HD: optimized for quality.

Use cases

See Responsible AI for additional considerations for responsible use.

Key use cases

The provider has not supplied this information.

Out of scope use cases

The provider has not supplied this information.

Pricing

Pricing is based on a number of factors, including deployment type and tokens used. See pricing details here.

Technical specs

The provider has not supplied this information.

Training cut-off date

The provider has not supplied this information.

Training time

The provider has not supplied this information.

Input formats

Max request data size: 4,096 chars can be converted from text to speech per API request.

Output formats

The provider has not supplied this information.

Supported languages

The provider has not supplied this information.

Sample JSON response

The provider has not supplied this information.

Model architecture

The provider has not supplied this information.

Long context

The provider has not supplied this information.

Optimizing model performance

The provider has not supplied this information.

Additional assets

The provider has not supplied this information.

Training disclosure

Training, testing and validation

The provider has not supplied this information.

Distribution

Distribution channels

The provider has not supplied this information.

More information

The provider has not supplied this information.

TtsHdstatic

Direct from Azure models

Direct from Azure models are a select portfolio curated for their market-differentiated capabilities:

  • Secure and managed by Microsoft: Purchase and manage models directly through Azure with a single license, consistent support, and no third-party dependencies, backed by Azure's enterprise-grade infrastructure.

  • Streamlined operations: Benefit from unified billing, governance, and seamless PTU portability across models hosted on Azure - all part of Microsoft Foundry.

  • Future-ready flexibility: Access the latest models as they become available, and easily test, deploy, or switch between them within Microsoft Foundry; reducing integration effort.

  • Cost control and optimization: Scale on demand with pay-as-you-go flexibility or reserve PTUs for predictable performance and savings.

Learn more about Direct from Azure models.

Key capabilities

About this model

TTS-HD is a model that converts text to natural sounding speech.

Key model capabilities

  • TTS: optimized for speed.

  • TTS-HD: optimized for quality.

Use cases

See Responsible AI for additional considerations for responsible use.

Key use cases

TTS is optimized for realtime or interactive scenarios. For offline scenarios, TTS-HD provides higher quality.

Out of scope use cases

The provider has not supplied this information.

Pricing

Pricing is based on a number of factors, including deployment type and tokens used. See pricing details here.

Technical specs

The provider has not supplied this information.

Training cut-off date

The provider has not supplied this information.

Training time

The provider has not supplied this information.

Input formats

Max request data size: 4,096 chars can be converted from text to speech per API request.

Output formats

The provider has not supplied this information.

Supported languages

The provider has not supplied this information.

Sample JSON response

The provider has not supplied this information.

Model architecture

The provider has not supplied this information.

Long context

The provider has not supplied this information.

Optimizing model performance

The provider has not supplied this information.

Additional assets

The provider has not supplied this information.

Training disclosure

Training, testing and validation

The provider has not supplied this information.

Distribution

Distribution channels

The provider has not supplied this information.

More information

The provider has not supplied this information.

Whisperstatic

Direct from Azure models

Direct from Azure models are a select portfolio curated for their market-differentiated capabilities:

  • Secure and managed by Microsoft: Purchase and manage models directly through Azure with a single license, consistent support, and no third-party dependencies, backed by Azure's enterprise-grade infrastructure.

  • Streamlined operations: Benefit from unified billing, governance, and seamless PTU portability across models hosted on Azure - all part of Microsoft Foundry.

  • Future-ready flexibility: Access the latest models as they become available, and easily test, deploy, or switch between them within Microsoft Foundry; reducing integration effort.

  • Cost control and optimization: Scale on demand with pay-as-you-go flexibility or reserve PTUs for predictable performance and savings.

Learn more about Direct from Azure models.

Key capabilities

About this model

The Whisper models are trained for speech recognition and translation tasks, capable of transcribing speech audio into the text in the language it is spoken (automatic speech recognition) as well as translated into English (speech translation).

Key model capabilities

  • Speech recognition (automatic speech recognition)

  • Speech translation into English

  • Processing of audio up to 25mb per API request

Use cases

See Responsible AI for additional considerations for responsible use.

Key use cases

The provider has not supplied this information.

Out of scope use cases

The provider has not supplied this information.

Pricing

Pricing is based on a number of factors, including deployment type and tokens used. See pricing details here.

Technical specs

The provider has not supplied this information.

Training cut-off date

The provider has not supplied this information.

Training time

The provider has not supplied this information.

Input formats

Max request data size: 25mb of audio can be converted from speech to text per API request.

Output formats

The provider has not supplied this information.

Supported languages

The provider has not supplied this information.

Sample JSON response

The provider has not supplied this information.

Model architecture

The provider has not supplied this information.

Long context

The provider has not supplied this information.

Optimizing model performance

The provider has not supplied this information.

Additional assets

The provider has not supplied this information.

Training disclosure

Training, testing and validation

Researchers at OpenAI developed the models to study the robustness of speech processing systems trained under large-scale weak supervision.

Distribution

Distribution channels

The provider has not supplied this information.

More information

The provider has not supplied this information.

View all fields