← Back to Guides

Choosing the Right AI Model

Every companion on Eidolon can use a different AI model. This guide covers what's available, what each model is best at, and how to choose the right one for the experience you want.

Quick Recommendations

Model Best For… Tier
DeepSeek Chat V3 0324 Low-cost capable model. Included
Gemma 4 26B (MoE) Very cheap multimodal MoE model with image support. Included
Gemma 4 31B Compact dense model with strong instruction following and vision. Included
Gemini 2.5 Flash Fast, creative all-rounder with strong instruction following. Included
Mistral Medium Expressive creative writing and consistent persona maintenance. Included
Mistral Small 4 Low-cost fast model balancing reasoning and speed. Included
Kimi K2 Excellent multi-lingual native comprehension. Included
Qwen Models Vision-language tasks and low-cost reasoning capabilities. Included
DeepSeek V3.2 Strong general reasoning and instruction following at very low cost. Included
DeepSeek Chat V3.1 Ultra-low cost conversational model, great for budget-conscious users. Included
GPT-4.1 Mini Fast, capable everyday model from OpenAI with a 1M token context window. Included
Claude Sonnet 4.6 High emotional range and nuanced relationship dynamics. Premium
Gemini 2.5 Pro Large-scale context and deep analytical reasoning. Premium
Kimi K2.5 Moonshot's powerful reasoning variant. Premium
GPT-5.4 Mini Balanced cost and quality from OpenAI's GPT-5.4 generation. Premium
GPT-4.1 OpenAI's previous-gen flagship — strong instruction following and reasoning. Premium
GPT-5.4 High-capability GPT-5.4 generation model for complex chat and roleplay. Premium
Claude Opus 4.6 Maximum logical depth and structural reasoning. Premium

How to Change Your Model

You can set a different model for each companion individually:

On Android & Web

  1. Open your companion's Profile.
  2. The AI Model selector is right on the profile page — choose a model.
  3. Your companion will use the new model starting with your next message.

On iPhone

  1. Open your companion's Profile.
  2. Tap Edit (pencil icon).
  3. Find the AI Model selector and choose a model.
  4. Save. Your companion will use the new model starting with your next message.

Models marked Premium require a BYOK (Bring Your Own Key) connection to OpenRouter. The current Included catalog is available without a key; new models added in the future will be Premium and require BYOK.

Included Models

These models are included for all users. They cover a wide range of styles and strengths. New models added to Eidolon will be Premium and available through BYOK.

DeepSeek V4 Flash (DeepSeek V4 Flash)

Included

by DeepSeek

Our default model for new companions. Fast, capable, and economical, with strong instruction following and tool use for everyday companion conversations.

⭐ Recommended Fast Default

Google Gemini 2.5 Flash (Gemini 2.5 Flash)

Included

by Google

Fast, creative, and excellent at following complex persona instructions. One of the best all-around included options — great for everyday conversations and companions with detailed personalities.

⭐ Recommended Fast Creative

Google Gemma 4 26B A4B (Gemma 4 26B MoE)

Included

by Google DeepMind

A Mixture-of-Experts model with 26B total parameters and only 4B active per token — delivering strong quality at very low cost. Supports image and video input with a 262K context window. A great budget pick with multimodal capability.

💰 Low Cost Vision MoE Efficient

Google Gemma 4 31B (Gemma 4 31B)

Included

by Google DeepMind

A compact 31B dense multimodal model with native function calling, configurable reasoning mode, and a 256K context window. Supports text, image, and video input. Strong instruction following at a very competitive price point.

💰 Low Cost Vision Reasoning

OpenAI GPT-5 Mini (GPT-5 Mini)

Included

by OpenAI

A fast, cost-efficient version of GPT-5. Great for snappy interactions without sacrificing the architectural improvements of the GPT-5 generation.

Speed Efficiency

OpenAI GPT-4.1 Mini (GPT-4.1 Mini)

Included

by OpenAI

A fast, capable model from OpenAI's GPT-4.1 generation with an impressive 1M token context window. Great for companions that handle long conversations or need strong general instruction following without premium pricing.

Long Context General Purpose

Moonshot Kimi K2 (Kimi K2)

Included

by Moonshot AI

A highly capable model known for excellent adherence to instructions and strong multilingual support.

Multilingual General Purpose

Mistral AI Mistral Small 4 (Mistral Small 4)

Included

by Mistral AI

A fast, cost-effective model that excels in reasoning and daily conversational tasks while maintaining Mistral's characteristic stylistic flair.

Fast Budget-Friendly

Mistral AI Mistral Large (Mistral Large)

Included

by Mistral AI

Mistral's flagship model. More analytical and structured than Medium — better for companions with complex, highly detailed personas or those that need to stay closely aligned with specific companion traits.

Detailed Precise

Alibaba Qwen 3.5 Flash (Qwen 3.5 Flash)

Included

by Alibaba Cloud

The most affordable model in our lineup. Despite the name, Qwen 3.5 Flash is a reasoning model — it "thinks" before responding, which means it's not the fastest in terms of latency. However, its extremely low cost makes it an excellent choice for users who prioritize budget over speed.

💰 Cheapest Budget-Friendly Reasoning

Alibaba Qwen 3.5 35B (Qwen 3.5 35B)

Included

by Alibaba Cloud

A compact but capable reasoning model. Good general-purpose performance at an extremely low cost. Works well for everyday conversation and companions that don't need heavy tool usage.

Budget-Friendly Compact

Alibaba Qwen 3 VL (Qwen 3 VL)

Included

by Alibaba Cloud

A vision-language model with strong image understanding. Among the best included options for companions that frequently interact with images and visual content. Not a reasoning model — responses are direct and fast.

Vision Budget-Friendly

DeepSeek V3.2 (DeepSeek V3.2)

Included

by DeepSeek

DeepSeek's latest flagship non-reasoning model. Strong general instruction following, creative writing, and tool use at an extremely low cost. A great budget pick for companions that need capable, well-rounded performance without premium pricing.

💰 Great Value General Purpose Budget-Friendly

DeepSeek Chat V3 0324 (DeepSeek Chat V3 0324)

Included

by DeepSeek

A capable model from DeepSeek that balances performance with affordability, serving as an excellent budget-conscious alternative.

Budget-Friendly Reasoning

DeepSeek Chat V3.1 (DeepSeek Chat V3.1)

Included

by DeepSeek

The most cost-effective model in our lineup — ideal for users who want to maximize conversation volume at the lowest possible cost. Solid conversational quality for everyday chat companions. Note that image inputs require a description workaround rather than direct vision.

💰 Lowest Cost Budget-Friendly Conversational

Premium Models

These models require a BYOK key connected to OpenRouter. You pay OpenRouter directly — we don't add any markup. All new models added to the catalog will enter this tier.

Moonshot Kimi K2.5 (Kimi K2.5)

Premium

by Moonshot AI • Est. $0.0107 / turn • BYOK • ~$6.42/mo at 20 msg/day

Moonshot's advanced variant that balances powerful multi-turn reasoning and conversational nuance with cost efficiency.

Premium Reasoning

Google Gemini 2.5 Pro (Gemini 2.5 Pro)

Premium

by Google • Est. $0.0375 / turn • BYOK • ~$22.50/mo at 20 msg/day

Google's most capable model. Deeper reasoning and stronger coherence than Flash, especially for companions with complex backstories or nuanced dynamics.

Premium Deep Reasoning High Quality

OpenAI GPT-5 (GPT-5)

Premium

by OpenAI • Est. $0.0375 / turn • BYOK • ~$22.50/mo at 20 msg/day

OpenAI's latest flagship. Exceptional at complex roleplay and instruction following.

Premium Flagship

OpenAI GPT-5.1 (GPT-5.1)

Premium

by OpenAI • Est. $0.0375 / turn • BYOK • ~$22.50/mo at 20 msg/day

Fine-tuned version of GPT-5 with improved coherence and reduced refusals.

Premium Refined

OpenAI GPT-4o (GPT-4o)

Premium

by OpenAI • Est. $0.0550 / turn • BYOK • ~$33.00/mo at 20 msg/day

OpenAI's previous-generation flagship. Still very capable and well-liked by many. If you're familiar with ChatGPT, this model will feel familiar.

Premium Familiar Reliable

Anthropic Claude Sonnet 4.6 (Claude Sonnet 4.6)

Premium

by Anthropic • Est. $0.0720 / turn • BYOK • ~$43.20/mo at 20 msg/day

Anthropic's latest and most capable creative model. Excellent at nuanced companion voice and emotional range. An advanced option for maintaining complex relationship dynamics across long conversations.

Premium Highly Expressive

Anthropic Claude Sonnet 4.5 (Claude Sonnet 4.5)

Premium

by Anthropic • Est. $0.0720 / turn • BYOK • ~$43.20/mo at 20 msg/day

The previous Sonnet generation — still outstanding for creative writing and companion work. Slightly cheaper per token than 4.6 while delivering a very similar quality of experience.

Premium Excellent Value

Anthropic Claude Haiku 4.5 (Claude Haiku 4.5)

Premium

by Anthropic • Est. $0.0240 / turn • BYOK • ~$14.40/mo at 20 msg/day

A lightweight, fast model from Anthropic's Claude family. Great for quick, casual conversations. Less depth than the larger Claude models, but significantly faster response times.

Premium Fast Lightweight

OpenAI GPT-5.4 Mini (GPT-5.4 Mini)

Premium

by OpenAI • Est. $0.0195 / turn • BYOK • ~$11.70/mo at 20 msg/day

A cost-effective entry into the GPT-5.4 generation. Delivers solid performance for everyday companions at a lower price point than full GPT-5.4, with a 1M token context window and strong instruction following.

Premium Efficient

OpenAI GPT-4.1 (GPT-4.1)

Premium

by OpenAI • Est. $0.0440 / turn • BYOK • ~$26.40/mo at 20 msg/day

OpenAI's previous-gen flagship model. Excellent at complex instruction following, detailed persona adherence, and long-context conversations. A proven performer for companions that need reliable, high-quality responses at a lower price than the GPT-5 generation.

Premium Reliable Long Context

OpenAI GPT-5.4 (GPT-5.4)

Premium

by OpenAI • Est. $0.0650 / turn • BYOK • ~$39.00/mo at 20 msg/day

A high-quality model from OpenAI's GPT-5.4 generation with a very large 1M token context window. Strong at complex roleplay, nuanced companion dynamics, and conversations requiring depth. A great mid-tier option between GPT-5 and Claude Sonnet.

Premium High Quality

Anthropic Claude Opus 4.6 (Claude Opus 4.6)

Premium

by Anthropic • Est. $0.1200 / turn • BYOK • ~$72.00/mo at 20 msg/day

Anthropic's most powerful model. Exceptional reasoning and creative depth. ⚠️ This is one of the most expensive models available — it costs significantly more per message than Sonnet or any included model. Only recommended if you specifically want the absolute maximum quality and are comfortable with the higher API costs.

Premium Very Expensive Enthusiast

Anthropic Claude Opus 4.5 (Claude Opus 4.5)

Premium

by Anthropic • Est. $0.1200 / turn • BYOK • ~$72.00/mo at 20 msg/day

Previous-generation Opus. Deep reasoning and creative range. ⚠️ Also one of the most expensive models — similar pricing to Opus 4.6. Consider Sonnet 4.6 for a more balanced quality-to-cost ratio.

Premium Very Expensive Enthusiast

xAI Grok 4 (Grok 4)

Premium

by xAI • Est. $0.0720 / turn • BYOK • ~$43.20/mo at 20 msg/day

xAI's full-power model. Stronger reasoning and more creative range than the Fast variants, with a distinctive direct and engaging personality.

Premium Powerful

About cost estimates: Monthly estimates assume ~20 messages per day for 30 days; the platform's no-key allowance is 75 messages per day. For a detailed breakdown of rates and estimates, see our Model Pricing Guide. Actual costs vary based on conversation length and complexity. Prices are set by model providers via OpenRouter and may change at any time. Included models are fully covered by the platform. Premium model costs are charged to your OpenRouter balance.

The Full Model Lineup

This list mirrors the model selector in the app. See the pricing guide for current OpenRouter rates.

Model Provider Capabilities Tier Pricing
Claude Haiku 4.5 anthropic Multimodal, Tools Premium See pricing guide
Claude Sonnet 4.5 anthropic Multimodal, Tools Premium See pricing guide
Claude Sonnet 4.6 anthropic Multimodal, Tools Premium See pricing guide
Claude Opus 4.5 anthropic Multimodal, Tools Premium See pricing guide
Claude Opus 4.6 anthropic Multimodal, Tools Premium See pricing guide
Claude Opus 4.8 anthropic Multimodal, Tools Premium See pricing guide
Claude Sonnet 5 anthropic Multimodal, Tools Premium See pricing guide
Claude Fable 5 anthropic Multimodal, Tools Premium See pricing guide
Gemini 2.5 Flash google Multimodal, Tools Included See pricing guide
Gemini 2.5 Pro google Multimodal, Tools Premium See pricing guide
Gemma 4 26B (MoE) google Multimodal, Tools Included See pricing guide
Gemma 4 31B google Multimodal, Tools Included See pricing guide
Mistral Large mistralai Multimodal, Tools Included See pricing guide
Mistral Medium 3.1 mistralai Multimodal, Tools Included See pricing guide
Mistral Medium 3.5 mistralai Multimodal, Tools Premium See pricing guide
Mistral Small 4 mistralai Multimodal, Tools Included See pricing guide
GPT-4o openai Multimodal, Tools Premium See pricing guide
GPT-4.1 openai Multimodal, Tools Premium See pricing guide
GPT-4.1 Mini openai Multimodal, Tools Included See pricing guide
GPT-5 openai Multimodal, Tools Premium See pricing guide
GPT-5 Mini openai Multimodal, Tools Included See pricing guide
GPT-5.1 openai Multimodal, Tools Premium See pricing guide
GPT-5.4 openai Multimodal, Tools Premium See pricing guide
GPT-5.4 Mini openai Multimodal, Tools Premium See pricing guide
GPT-5.6 Luna openai Multimodal, Tools Premium See pricing guide
GPT-5.6 Luna Pro openai Multimodal, Tools Premium See pricing guide
GPT-5.6 Terra openai Multimodal, Tools Premium See pricing guide
GPT-5.6 Terra Pro openai Multimodal, Tools Premium See pricing guide
GPT-5.6 Sol openai Multimodal, Tools Premium See pricing guide
GPT-5.6 Sol Pro openai Multimodal, Tools Premium See pricing guide
Qwen 3 VL qwen Multimodal, Tools Included See pricing guide
Qwen 3.5 Flash qwen Multimodal, Tools Included See pricing guide
Qwen 3.5 35B qwen Multimodal, Tools Included See pricing guide
Grok 4.3 x-ai Multimodal, Tools Premium See pricing guide
Grok 4.5 x-ai Multimodal, Tools Premium See pricing guide
DeepSeek V3.2 deepseek Conversational, Tools Included See pricing guide
DeepSeek Chat V3.1 deepseek Conversational, Tools Included See pricing guide
DeepSeek Chat V3 0324 deepseek Conversational, Tools Included See pricing guide
Kimi K2 moonshotai Conversational, Tools Included See pricing guide
Kimi K2.5 moonshotai Conversational, Tools Premium See pricing guide
Gemini 3.1 Pro Preview google Multimodal, Tools Premium See pricing guide
Gemini 3.6 Flash google Multimodal, Tools Premium See pricing guide
Gemini 3.7 Flash google Multimodal, Tools Premium See pricing guide
DeepSeek V4 Flash deepseek Conversational, Tools Included See pricing guide
DeepSeek V4 Flash (0731 snapshot) deepseek Conversational, Tools Included See pricing guide
DeepSeek V4 Pro deepseek Conversational, Tools Included See pricing guide
Qwen 3.6 Flash qwen Multimodal, Tools Included See pricing guide
Qwen 3.6 35B qwen Multimodal, Tools Included See pricing guide
Qwen 3.8 27B qwen Multimodal, Tools Premium See pricing guide
Qwen 3.8 2.4T (MoE) qwen Conversational, Tools Premium See pricing guide
Gemini Pro Latest ~google Multimodal, Tools Premium See pricing guide
Gemini Flash Latest ~google Multimodal, Tools Premium See pricing guide
GLM 4.7 z-ai Conversational, Tools Premium See pricing guide
GLM 4.7 Flash z-ai Conversational, Tools Premium See pricing guide
GLM 5.1 z-ai Conversational, Tools Premium See pricing guide
GLM 5.2 z-ai Conversational, Tools Premium See pricing guide

A Note on Speed: "Flash" ≠ Always Fastest

Some models with "Flash" in the name (like Alibaba Qwen 3.5 Flash) are actually reasoning models that "think" internally before responding. This hidden reasoning step can make them slower than their non-flash counterparts, even though they cost less per token.

Similarly, premium reasoning models like Google Gemini 2.5 Pro will naturally have longer response times because they're doing deeper analysis — this is by design and produces higher quality responses, but it means they're not ideal if you want instant replies.

If speed is your priority: Qwen 3.6 Flash, DeepSeek V4 Flash, and Mistral models deliver the snappiest responses. If cost is your priority: the economy models are hard to beat — included models are fully covered for all users.

Tips

  • Experiment freely. You can switch models at any time without losing memories, goals, or conversation history. Everything is preserved.
  • Different models suit different personas. A companion with a poetic, emotionally complex personality may shine on Claude Sonnet, while a witty, factual companion might be better on Grok or Gemini.
  • Premium models fall back gracefully. If your BYOK credits run out, the system automatically falls back to an included model so you never lose access. See our BYOK guide for details.