LLM Reference
Azure OpenAI

Azure OpenAI

Researched 129d agoHyperscalerTier 1

Microsoft

CodingRAGAgentsLong contextVisionClassificationJSON / Tool useHighlightHyperscaler

Azure OpenAI offers 15 tracked models (10 with output token pricing). This catalog covers coding, rag, and agents; open any model detail page for benchmarks, batch tiers, and migration prompts.

Covers 7 workload areas across 15 tracked models; last verified 2026-05-10.

Use it for

  • Teams comparing token and batch pricing across this provider's models
  • Operators routing coding, rag, and agents workloads through this API
  • Batch buyers auditing discount coverage model-by-model

Do not use it for

  • Final benchmark picks without opening the relevant model detail page

Tracked models

15

Models available through this provider

Priced output routes

10

Models with output token pricing tracked

Cheapest output

$0.400

babbage on this route

Batch-ready models

3

Models with discounted batch pricing

Latest model release

2025-04-01

533d since newest release

Freshness

2026-05-10

Researched 129d ago

stale

Routes available via routers & gateways

Browse routers ->

Information

TypeHyperscaler
TierTier 1
Models15
CompanyMicrosoft
Founded1975
Redmond, Washington, United States

Azure OpenAI Service hosts OpenAI's GPT-4o, GPT-4, GPT-3.5, and embedding models on Microsoft Azure with enterprise SLAs. Microsoft's broader AI platform spans Azure Machine Learning, Azure Cognitive Services, and Azure AI Search, plus productivity tools like Microsoft 365 Copilot and developer tooling in Visual Studio and the .NET framework. The platform emphasizes responsible AI, security, and regional compliance.

Where this host wins

  • Coding: 2 tracked models with SWE-bench / HumanEval-style scores.
  • RAG: 3 tracked models with ruler / needle retrieval benchmarks.
  • Agentic: 2 tracked models with BFCL, tau-bench, and SWE-bench tool-use coverage.
  • Long-context: 3 tracked models with context-token or InfiniteBench-class signal.

Getting started

Verify: quotas and regions in the linked vendor documentation.

SDKs & libraries

Platform Overview

Azure OpenAI Service hosts OpenAI's GPT-4o, GPT-4, GPT-3.5, and embedding models on Microsoft Azure with enterprise SLAs. Deployments run in customer-selected regions with private networking, role-based access control, and capacity options spanning Standard pay-per-token, Provisioned Throughput Units (PTUs) for reserved capacity, Global Standard shared capacity, and Batch processing. Azure OpenAI sits inside the wider Microsoft Foundry / Azure AI Studio control plane, which adds an evaluation, monitoring, and Agent Service layer on top of the base model APIs.

Available Models(15)

View all →
ModelInput (per 1M)Output (per 1M)Batch input (per 1M)Batch output (per 1M)Type
GPT-4.1$1.00$4.00
Serverless
GPT-4.5
Serverless
o3 Mini$0.15$0.60
Serverless
GPT-4o-mini
Serverless
GPT-4o$1.25$5.00
Serverless
GPT-4o (05-13)$5$15
Serverless
GPT-4 Turbo$10$30
Serverless
GPT-4 Turbo Preview$10$30
Serverless
GPT-4 Vision Preview$10$40.00
Serverless
GPT-3.5 Turbo (Instruct)$1.5$2
Serverless
View full catalog →

Where else to run this