CompareLLM.ai
Live
Leaderboard
Models
Compare
Best of & Stacks
Research & News
…
CompareLLM.ai
Precision Benchmarks

Programmatic, dated AI model benchmarks, head-to-head comparisons, and Stack Engine presets.

Daily ingest · 06:00 UTC

Analytics & Benchmarks

  • AI Model Leaderboard
  • Head-to-Head Compare Hub
  • Models Directory
  • Stack Engine Presets
  • Frontier Models
  • Open Weights Catalog

Guides & Intent Lists

  • Best LLM Lists (2026)
  • Best Coding LLM
  • Best Cheap LLM
  • Fastest Low-Latency LLM
  • Claude vs GPT Benchmark
  • What is Elo?
  • Methodology Guides
  • News & Dispatches

Transparency & API

  • Evaluation Methodology
  • Benchmark Changelog
  • Public JSON API
  • llms.txt Specification
  • Privacy Policy
  • Sign In / Account

© 2026 CompareLLM. Public benchmark data aggregated from Arena Elo, LiveBench, SWE-bench & OpenRouter.

Every score has a dated snapshot.

Theme:
Currency:
  1. Home
  2. Providers
  3. Microsoft AI
AI Provider Portfolio
4 active models in catalog

Microsoft AI Models

Comprehensive benchmark scores, coding performance, latency metrics, and API pricing for all models developed by Microsoft AI.

Flagship ModelMAI-Image-2.5
Portfolio Avg Elo—
Browse Other Providers
AnthropicOpenAIGoogle

Available Models & Verified Benchmarks

#1MAI-Image-2.5
Proprietary

Microsoft AI proprietary foundation image model for Copilot Studio and enterprise creative workflows.

Elo—
SWE-bench—
Out $/1M—
Model Fact Sheetvs all hub
#2MAI-Image-2.5-Flash
Proprietary

Cost-optimized Microsoft AI image model delivering sub-2 second responses at $20/1k images.

Elo—
SWE-bench—
Out $/1M—
Model Fact Sheetvs all hub
#3Microsoft: Phi 4
Proprietary

Auto-discovered from OpenRouter (microsoft/phi-4). Preview until a second source matches.

Elo—
SWE-bench—
Out $/1M$0.14
Model Fact Sheetvs all hub
#4WizardLM-2 8x22B
Proprietary

Auto-discovered from OpenRouter (microsoft/wizardlm-2-8x22b). Preview until a second source matches.

Elo—
SWE-bench—
Out $/1M$0.62
Model Fact Sheetvs all hub