200+ models supported

Every model.
One memory.

Switch between GPT-5.6 Sol, Claude Opus 5, Gemini 3.6 Flash, Llama 4, and 200+ more — your memory follows seamlessly.

swap models, keep memory
// Monday: Use GPT-5.6 Sol
model: "gpt-5.6-sol"

// Tuesday: Try Claude
model: "anthropic/claude-opus-5"

// Wednesday: Test Gemini
model: "google/gemini-3.6-flash"

// Memory persists across all of them ✓

No lock-in. No migration. Just change the model parameter.

OpenAIFlagship
Context
1.05M

GPT-5.6 Sol

OpenAI flagship for complex professional reasoning, coding, research, and computer use. 128K max output.

ReasoningComputer UseCoding
OpenAIPopular
Context
1.05M

GPT-5.6 Terra

Balanced GPT-5.6 model for everyday production work, with frontier tools and lower cost than Sol.

AgenticBest Value128K Output
OpenAIFast
Context
1.05M

GPT-5.6 Luna

Fast, cost-efficient GPT-5.6 model for high-volume workflows and multi-step tool use.

High VolumeLow CostTools
OpenAINew
Context
Realtime

GPT-Live-1

OpenAI voice model for natural, continuous conversation with frontier-model delegation for deeper work.

VoiceRealtimeMultimodal
AnthropicFlagship
Context
1M

Claude Opus 5

Anthropic premium model for serious coding, long-running agents, and complex professional work.

CodingAgentsFast Mode
AnthropicPopular
Context
1M

Claude Sonnet 5

Anthropic's most agentic Sonnet: strong reasoning, browser and terminal use, coding, and knowledge work.

Computer UseCodingBest Value
Anthropic
Context
200K

Claude Haiku 4.5

Fast, economical Claude model for responsive assistants and simple tasks at scale.

VisionUltra FastLow Cost
GoogleFlagship
Context
1.05M

Gemini 3.6 Flash

Google's stable agentic workhorse for coding, multimodal tasks, spatial reasoning, and rapid tool loops.

AgenticMultimodal65K Output
Google
Context
1.05M

Gemini 3.5 Flash

Frontier-speed model for subagents, multi-step workflows, and long-horizon coding tasks.

AgenticFastCoding
GoogleFast
Context
1.05M

Gemini 3.5 Flash-Lite

Low-latency, cost-efficient multimodal model for document parsing and high-throughput subagents.

High VolumeLow CostMultimodal
GooglePreview
Context
1M

Gemini 3.1 Pro

Google Pro-class reasoning model for advanced knowledge work and multimodal workflows.

VisionReasoningAgentic
Meta via OpenRouterFlagship
Context
1M

Llama 4 Maverick

Open-weight, natively multimodal 128-expert MoE model with 17B active parameters.

VisionMoE 128EOpen Weights
Meta via OpenRouter10M Context
Context
10M

Llama 4 Scout

Open-weight 16-expert MoE built for massive documents, personalization, and codebases.

VisionLong ContextOpen Weights
Meta via OpenRouterNew
Context
Long

Muse Spark 1.1

Meta multimodal reasoning model for agentic tasks, coding, tool use, and computer use.

MultimodalComputer UseAgentic
Mistral via OpenRouterFlagship
Context
256K

Mistral Medium 3.5

128B open-weight multimodal flagship for long-horizon coding, reasoning, and productivity agents.

ReasoningAgenticOpen Weights
Mistral via OpenRouterPopular
Context
256K

Mistral Small 4

Apache 2.0 hybrid model unifying instruct, reasoning, multimodal, and agentic coding capabilities.

MultimodalCodingApache 2.0
Mistral via OpenRouterNew
Context
Documents

Mistral OCR 4

Document-intelligence model with bounding boxes, confidence scores, and support for 170 languages.

OCRDocumentsMultilingual
xAI via OpenRouterFlagship
Context
256K

Grok 4.3

General Grok model for fast chat, reasoning, and tool-heavy agents.

ReasoningRealtimeAgentic
xAI via OpenRouter
Context
256K

Grok 4.20

High-capability Grok 4 line model available through OpenRouter.

AnalysisChatTools
xAI via OpenRouter
Context
256K

Grok Build 0.1

xAI build-focused model for implementation-heavy agent workflows.

BuildCodeAgentic
DeepSeek via OpenRouterFlagship
Context
1M

DeepSeek V4 Pro

1.6T-total, 49B-active open model for agentic coding, reasoning, and rich world knowledge.

ReasoningCodingOpen Weights
DeepSeek via OpenRouterNew
Context
1M

DeepSeek V4 Flash 0731

Public-beta V4 Flash update with thinking modes, native Responses API support, and Codex optimization.

FastResponses API384K Output
DeepSeek via OpenRouter
Context
1M

DeepSeek V4 Flash

284B-total, 13B-active high-throughput model with thinking and non-thinking modes.

FastLow CostAgentic
Perplexity via OpenRouterSearch
Context
128K

Sonar Pro Search

Perplexity search model for grounded answers and web-aware agents.

SearchCitationsResearch
Perplexity via OpenRouterReasoning
Context
128K

Sonar Reasoning Pro

Perplexity reasoning model for grounded, multi-step research tasks.

ReasoningSearchResearch
Perplexity via OpenRouter
Context
128K

Sonar Deep Research

Deep research model for long-form investigations and synthesis.

ResearchCitationsSynthesis
And More

200+ More Models

Cohere, Perplexity, Together, and dozens more. If it's on OpenRouter, it works with MemoryRouter.

CoherePerplexityTogether+ more

Ready to Add Memory?

One line change. Same code. Persistent memory across every model.

Give your AI a memory that lasts.

Start free with a 14-day trial. No credit card required.