Live coverage

Directory

Models vs apps & tools · live OpenRouter pricing · no fake benchmarks

FreeModelLive

Dots Studio: Dots3-Note Preview (free)

Dots Studio

Best for lightweight applications; weak at handling complex, resource-intensive tasks.

Free via OpenRouter (subject to Open512k
FreeModel

Gemini (Free tier)

Google

Best for quick multimodal Q&A at $0; weak when you need production rate limits or agent loops.

Free consumer / AI Studio tier
FreeModelLive

Google: Gemma 4 26B A4B (free)

Google

Best for high-quality instruction tasks; weak at handling extensive contextual information due to limited active parameters.

Free via OpenRouter (subject to Open262k
FreeModelLive

Google: Lyria 3 Clip Preview

Google

Best for generating short music clips; weak at creating longer compositions.

Free via OpenRouter (subject to Open1049k
FreeModelLive

Google: Lyria 3 Pro Preview

Google

Best for generating high-quality music tracks; weak at producing lyrics or vocal performances.

Free via OpenRouter (subject to Open1049k
FreeModel

Meta AI (Llama)

Meta

Best for casual chat inside Meta apps; weak as a developer API or offline runtime.

Free
FreeModelLive

MiniMax: MiniMax M3 (free)

MiniMax

Best for multimodal tasks like coding and creative projects; weak at real-time interactive applications.

Free via OpenRouter (subject to Open1049k
FreeModelLive

NVIDIA: Nemotron 3 Ultra (free)

NVIDIA

Best for complex reasoning tasks; weak at simple, repetitive queries.

Free via OpenRouter (subject to Open1000k
FreeModelLive

NVIDIA: Nemotron 3.5 Lightning (free)

NVIDIA

Best for high-throughput agentic workloads; weak at general-purpose conversational tasks.

Free via OpenRouter (subject to Open1000k
FreeModelLive

Nex AGI: Nex-N2.5-Mini (free)

Nex AGI

Best for coding tasks with visual feedback; weak at handling unstructured data or non-coding related inquiries.

Free via OpenRouter (subject to Open262k
FreeModelLive

Nex AGI: Nex-N2.5-Pro (free)

Nex AGI

Best for coding tasks with visual feedback; weak at handling unstructured text analysis.

Free via OpenRouter (subject to Open262k
FreeModelLive

Poolside: Laguna S 2.1 (free)

Poolside

Best for coding assistance; weak at complex problem-solving.

Free via OpenRouter (subject to Open262k
FreeModelLive

Poolside: Laguna XS 2.1 (free)

Poolside

Best for coding assistance; weak at handling complex multi-step reasoning tasks.

Free via OpenRouter (subject to Open262k
FreeModelLive

Thinking Machines: Inkling (free)

Thinking Machines

Best for general-purpose reasoning and coding; weak at highly specialized domain tasks requiring deep expertise.

Free via OpenRouter (subject to Open1049k
FreeModelLive

Thinking Machines: Inkling Small (free)

Thinking Machines

Best for efficient multimodal tasks; weak at handling large-scale data processing.

Free via OpenRouter (subject to Open1049k
FreeModelLive

inclusionAI: Ling 3.0 Flash Fin (free)

inclusionAI

Best for financial analysis and investment insights; weak at general conversational tasks.

Free via OpenRouter (subject to Open262k
FreeModelLive

inclusionAI: Ling 3.0 Flash Sante (free)

inclusionAI

Best for health-related inquiries; weak at general knowledge outside medical contexts.

Free via OpenRouter (subject to Open262k
PaidModel

Claude API

Anthropic

Claude is best used for generating human-like text responses in conversational applications.

Pricing is tiered based on usage; co
PaidModelLive

DeepSeek V4 Flash Latest

~deepseek

Redirecting users to the latest DeepSeek V4 Flash model.

$0.04 / 1M input tokens1311k
PaidModelLive

DeepSeek: DeepSeek V4 Flash 0731

DeepSeek

This model is best used for coding and reasoning tasks.

$0.05 / 1M input tokens1311k
PaidModel

GPT API

OpenAI

GPT-4o is best used for generating high-quality text across various applications.

Pricing is based on usage and may va
PaidModel

Gemini API

Google

Gemini is best used for advanced natural language processing tasks.

Pricing is tiered based on usage and
PaidModelLive

Google: Gemma 3 4B

Google

Best for multimodal tasks involving vision and language; weak at highly specialized technical queries.

$0.05 / 1M input tokens131k
PaidModelLive

IBM: Granite 4.0 Micro

IBM

Best for long-form content generation; weak at real-time interactive applications.

$0.02 / 1M input tokens131k
PaidModelLive

Meta: Llama 3.1 8B Instruct

Meta

Best for quick, efficient instruction-following tasks; weak at handling complex reasoning or deep contextual understanding.

$0.05 / 1M input tokens131k
PaidModelLive

Mistral: Mistral Nemo

Mistral

Best for multilingual text generation; weak at highly specialized technical topics.

$0.02 / 1M input tokens131k
PaidModelLive

Mistral: Mistral Small 3

Mistral

Best for low-latency AI tasks; weak at handling complex reasoning or nuanced conversations.

$0.05 / 1M input tokens33k
PaidModelLive

MythoMax 13B

Gryphe

Best for creative storytelling and roleplay; weak at technical or factual accuracy.

$0.06 / 1M input tokens8k
PaidModelLive

Nex AGI: Nex-N2-Mini

Nex AGI

Best for coding and tool use; weak at handling large-scale image processing tasks.

$0.02 / 1M input tokens262k
PaidModelLive

Qwen: Qwen3.7 Flash

Qwen

Best for multimodal tasks like visual coding; weak at handling purely text-based queries.

$0.03 / 1M input tokens1000k
PaidModelLive

Sao10K: Llama 3 8B Lunaris

Sao10K

Best for creative storytelling and roleplaying; weak at highly technical or specialized knowledge tasks.

$0.04 / 1M input tokens8k
PaidModelLive

Upstage: Solar Pro 4

Upstage

Best for long-horizon tasks and document management; weak at real-time conversational interactions.

$0.03 / 1M input tokens524k
PaidModelLive

inclusionAI: Ling 3.0 Flash

inclusionAI

Best for high-efficiency token processing; weak at handling complex multi-turn dialogues.

$0.02 / 1M input tokens262k
LocalModel

DeepSeek (via Ollama)

DeepSeek

Best for strong open-weight reasoning/coding locally; weak if you lack VRAM for mid-size tags.

Free8GB+
LocalModel

Llama (via Ollama)

Meta

Llama is best used for natural language processing tasks and conversational AI applications.

Free8GB+
LocalModel

Mistral (via Ollama)

Mistral AI

Mistral is best used for high-performance natural language processing tasks.

Pricing details may vary; please che6GB+
LocalModel

Stable Diffusion (via ComfyUI)

Stability AI / community

Best for unlimited local image pipelines; weak if you want one-click cloud simplicity.

Free6GB+