Frontier Model Intelligence

AI Models Directory

Comprehensive directory of 91 foundation models, LLMs, and multimodal engines from top research labs. Real pricing, benchmark scores, and architecture cards.

OpenAI

OpenAI GPT-6 model for focused, high-volume workloads where cost efficiency is the priority.

Context :1M ctx
Pricing :USD 0.10 / 1M input tok; USD 0.50 / 1M output tok
Type :reasoning-llm
Anthropic

Anthropic Claude model for long-running agentic coding and knowledge work.

Context :1M ctx
Pricing :USD 4 / 1M input tok; USD 20 / 1M output tok
Type :reasoning-llm
OpenAI

OpenAI GPT-6 model built for complex coding and agentic workflows, balancing intelligence and cost.

Context :1M ctx
Pricing :USD 2 / 1M input tok; USD 10 / 1M output tok
Type :reasoning-llm
TypeSafe AI

TypeSafe current System One model for fast, structured decisions with typed outputs, probabilities, and confidence.

Context :64K ctx
Pricing :USD 0.042 / 1M input tok; output free
Type :structured-decision
DeepSeek

DeepSeek current native-multimodal Flash model with an asymmetric 552B-parameter MoE architecture.

Context :1M ctx
Pricing :Peak: USD 0.30 / 1M input tok; USD 1.20 / 1M output tok
Type :multimodal-llm
Google

The latest evolution in the Gemini 3 family, delivering state-of-the-art software engineering (73.7% DeepSWE) and agentic enterprise knowledge workflows.

Context :1M ctx
Pricing :$0.75 / 1M tok ($1.50 reg)
Type :multimodal-realtime
Anthropic

The ultimate model for demanding reasoning and long-horizon agentic work, featuring 0.025x cache reads ($0.25/MTok).

Context :1M ctx
Pricing :$10.00 / 1M tok (0.025x Cache Reads)
Type :reasoning-llm
Anthropic

The most capable autonomous intelligence system engineered by Anthropic, restricted to Project Glasswing participants.

Context :1M ctx
Pricing :$10.00 / 1M tok (Glasswing Exclusive)
Type :reasoning-llm
Google

Universal omnimodal generation model that accepts any combination of text, audio, image, and video to generate any combination of outputs.

Context :1M ctx
Pricing :$0.60 / 1M tok (Universal Omni)
Type :multimodal-realtime
Z.AI

Ultra-low latency version of GLM-5.3-Flash engineered for high-concurrency real-time IDE code completion and interactive agents.

Context :200K ctx
Pricing :$0.37 / 1M tok (Ultra-Low Latency MoE)
Type :multimodal-realtime
Z.AI

The first native multimodal model in the GLM-5 series: 320B parameters (18B active) combining linear and sparse attention for low-cost visual coding.

Context :200K ctx
Pricing :$0.15 / 1M tok (320B MoE Visual Coding)
Type :multimodal-realtime
OpenAI

OpenAI’s premier sixth-generation frontier model, built for the most demanding end-to-end software engineering and scientific research tasks.

Context :1M ctx
Pricing :$10.00 / 1M tok (GPT-6 Frontier)
Type :reasoning-llm
Z.AI

Z.AI’s premier flagship model, delivering a 50% performance gain on Code Bench and matching Claude Mythos 5 in cybersecurity and vulnerability discovery.

Context :1M ctx
Pricing :$1.40 / 1M tok (Flagship Coding & Cyber)
Type :reasoning-llm
Google

High-performance multimodal foundation model featuring algorithmic reasoning enhancements and agentic long-form video understanding.

Context :1M ctx
Pricing :$0.50 / 1M tok (Agentic Video)
Type :multimodal-realtime
DeepSeek

DeepSeek V4 Pro API model, version DeepSeek-V4-Pro-0813, with thinking and non-thinking modes.

Context :1M ctx
Pricing :Peak: USD 1.32 / 1M input tok; USD 3.96 / 1M output tok
Type :reasoning-llm
Google

Cloud embodied foundation model capable of synthesizing novel robotic tool use and executing complex mechanical assemblies.

Context :524K ctx
Pricing :Enterprise Robotics Cloud 2.0
Type :multimodal-realtime
Google

Second-generation on-device physical AI foundation model incorporating tactile sensor feedback for delicate robotic manipulation.

Context :33K ctx
Pricing :On-Device Edge Embedded 2.0
Type :multimodal-realtime
Anthropic

The premier enterprise flagship model for multi-hour autonomous coding, complex system engineering, and vision-heavy agent tasks.

Context :1M ctx
Pricing :$5.00 / 1M tok (Opus 5 Flagship)
Type :reasoning-llm
Google

Iterative performance leap in the Flash family featuring major upgrades in code refactoring and autonomous agent tool loops.

Context :1M ctx
Pricing :$0.35 / 1M tok
Type :multimodal-realtime
Google

High-speed, high-density lightweight model providing strong analytical accuracy for 8 cents per million tokens.

Context :1M ctx
Pricing :$0.08 / 1M tok
Type :text-to-text
Moonshot AI

Moonshot AI’s premier flagship model with 2.8 trillion parameters, Kimi Delta Attention, 1M context, and open weights.

Context :1M ctx
Pricing :$3.00 / 1M tok (2.8T Open Weights)
Type :reasoning-llm
Google

Sub-second, sub-cent visual generation model designed for high-scale gaming asset generation, thumbnails, and preview workflows.

Context :262K ctx
Pricing :$0.008 / image
Type :text-to-image
Anthropic

Anthropic’s fifth-generation workhorse, combining 1M context with adaptive thinking on by default at a lower $2/$10 price point.

Context :1M ctx
Pricing :$2.00 / 1M tok (1M Context Sonnet 5)
Type :reasoning-llm
OpenAI

Flagship model family of mid-2026 featuring Sol (flagship), Terra (balanced), Luna (high-speed), and Cyber (authorized vulnerability research).

Context :1M ctx
Pricing :$4.00 / 1M tok (Flagship Sol)
Type :reasoning-llm
Z.AI

Flagship model built for project-scale context, supporting truly usable 1M-token context with 128k output.

Context :1M ctx
Pricing :$1.40 / 1M tok (1M Lossless Context)
Type :reasoning-llm
Moonshot AI

High-speed variant of Kimi K2.7 Code with generation speeds of 180 to 260 tokens per second for instantaneous coding feedback.

Context :262K ctx
Pricing :$1.90 / 1M tok (180-260 tok/s)
Type :reasoning-llm
Moonshot AI

Dedicated coding model delivering higher task success rates, tighter instruction following, and a 30% reduction in overthinking.

Context :262K ctx
Pricing :$0.95 / 1M tok (Dedicated Coding)
Type :reasoning-llm
Anthropic

Fifth-generation Claude architecture engineered specifically for multi-hour autonomous agent workflows and deep research.

Context :1M ctx
Pricing :$10.00 / 1M tok (Long-Horizon Agents)
Type :reasoning-llm
Anthropic

Limited-availability autonomous intelligence model delivering breakthrough problem-solving across scientific and technical disciplines.

Context :1M ctx
Pricing :$10.00 / 1M tok (Limited Availability)
Type :reasoning-llm
Anthropic

The pinnacle of the Claude 4 generation, serving as the universal recommended replacement for legacy Claude 2 and Claude 3 Opus.

Context :1M ctx
Pricing :$5.00 / 1M tok (Default Thinking)
Type :reasoning-llm
Google

Mid-generation multimodal speed model with configurable thinking levels, allowing developers to balance latency and reasoning depth.

Context :1M ctx
Pricing :$0.25 / 1M tok (with Thinking Levels)
Type :multimodal-realtime
OpenAI

OpenAI’s premier full-duplex voice model for natural, expressive spoken dialogues with instantaneous interruption handling.

Context :131K ctx
Pricing :$0.05 / min (Full-Duplex Voice)
Type :multimodal-realtime
Google

Extended Reasoning (ER) foundation model enabling fine-motor tool use and dynamic physics adaptation for autonomous robotic arms.

Context :262K ctx
Pricing :Enterprise Robotics Cloud
Type :multimodal-realtime
OpenAI

A new class of intelligence for advanced coding and professional knowledge work with a 1.05M context window and 128K output capacity.

Context :1M ctx
Pricing :$5.00 / 1M tok (Professional Class)
Type :reasoning-llm
Anthropic

The first Opus model with 1,000,000 tokens of context and 128,000 tokens of maximum output generation.

Context :1M ctx
Pricing :$5.00 / 1M tok (1M Context Opus)
Type :reasoning-llm
Moonshot AI

General-purpose model supporting text, image, and video inputs, switchable thinking modes, and autonomous agent tasks over 256k context.

Context :262K ctx
Pricing :$0.95 / 1M tok (Vision & Agentic)
Type :reasoning-llm
Z.AI

Engineered for long-horizon tasks, able to work independently for up to 8 hours in a single run, aligned with Claude Opus 4.6.

Context :200K ctx
Pricing :$1.40 / 1M tok (8h Autonomous Agent)
Type :reasoning-llm
OpenAI

OpenAI’s premier image synthesis and conversational in-painting model, supporting pinpoint region edits and variable quality modes.

Context :262K ctx
Pricing :$0.045 / image (Precision Editing)
Type :text-to-image
Anthropic

A preview of Anthropic’s next-generation autonomous reasoning system with 1M context, deployed for frontier safety evaluation.

Context :1M ctx
Pricing :$10.00 / 1M tok (Mythos Preview)
Type :reasoning-llm
OpenAI

First OpenAI model family introducing a 1.05-million-token context window with configurable reasoning effort and affordable pricing.

Context :1M ctx
Pricing :$2.50 / 1M tok (1.05M Context)
Type :reasoning-llm
Google

The Gemini 3.1 family’s budget workhorse, delivering solid reasoning and vision parsing at 6 cents per million tokens.

Context :1M ctx
Pricing :$0.06 / 1M tok
Type :text-to-text
Google

High-speed image generation and editing model engineered for conversational chats and real-time interactive design.

Context :524K ctx
Pricing :$0.02 / image
Type :text-to-image
Google

Google’s most advanced model for complex coding, scientific discovery, and long-horizon multimodal reasoning with a 64K output token buffer.

Context :1M ctx
Pricing :$1.50 / 1M tok (64k output)
Type :text-to-text
Anthropic

The designated LTS upgrade target for all Claude 3 and early Claude 4 applications, offering 500k context.

Context :500K ctx
Pricing :$3.00 / 1M tok (Sonnet 4.6)
Type :reasoning-llm
Z.AI

Fifth-generation foundation model shifting from coding to complex systems engineering, benchmarked against Claude Opus 4.5.

Context :200K ctx
Pricing :$1.00 / 1M tok (DeepSeek Sparse Attention)
Type :reasoning-llm
Anthropic

Refined reasoning engine delivering over 80% on SWE-bench Verified with advanced tool choice control.

Context :500K ctx
Pricing :$5.00 / 1M tok (Opus 4.6)
Type :reasoning-llm
Z.AI

Compact optical character recognition model combining CogViT with GLM-0.5B for fast, highly accurate document and table extraction.

Context :33K ctx
Pricing :$0.03 / 1M tok (CogViT + GLM-0.5B)
Type :multimodal-realtime
Z.AI

Free-tier version of GLM-4.7 delivering fast inference, 200k context, and strong coding for high-frequency applications.

Context :200K ctx
Pricing :Free API (200k Context 4.7)
Type :text-to-text
Z.AI

State-of-the-art image generation model combining autoregressive semantic understanding with diffusion decoding for accurate text rendering.

Context :4K ctx
Pricing :$0.015 / image (Text Rendering SOTA)
Type :text-to-image
Z.AI

Optimized agentic coding model setting open-source SOTA performance on major coding and reasoning benchmarks with 200k context.

Context :200K ctx
Pricing :$0.60 / 1M tok (Agentic Coding SOTA)
Type :reasoning-llm
Google

The third-generation speed-optimized foundation model combining sub-second latency with near-Pro intelligence.

Context :1M ctx
Pricing :$0.15 / 1M tok
Type :multimodal-realtime
Z.AI

Multimodal mobile automation framework that understands smartphone screen content and executes real user actions through ADB across 50+ mainstream apps.

Context :64K ctx
Pricing :Mobile OS Automation Agent
Type :autonomous-agent
Z.AI

High-accuracy automatic speech recognition model delivering a Character Error Rate as low as 0.0717 with user-defined vocabulary support.

Context :33K ctx
Pricing :$0.0024 / min audio (0.0717 CER)
Type :speech-to-text
Z.AI

Multimodal vision model with native function calling and controllable thinking mode switch over 128k tokens.

Context :128K ctx
Pricing :$0.30 / 1M tok (Native Tool Calling Vision)
Type :multimodal-realtime
Google

Professional-grade visual synthesis model featuring flawless text rendering, spatial consistency, and multi-turn iterative image editing.

Context :524K ctx
Pricing :$0.04 / image (HD)
Type :text-to-image
Google

The inaugural model in the Gemini 3 family, establishing a new frontier in native multimodal reasoning and autonomous coding.

Context :1M ctx
Pricing :$1.50 / 1M tok
Type :text-to-text
Anthropic

A major architectural refinement reducing Opus pricing by 66% while increasing context to 500,000 tokens.

Context :500K ctx
Pricing :$5.00 / 1M tok (Opus 4.5)
Type :reasoning-llm