Every Model, One API Key
Call any model below through the OpenAI-compatible MixerLead API at https://api.mixerlead.com/v1. Credits used = (prompt + completion tokens) × the model's credit multiplier.
| Model | Model ID | Type | Context | Credit × |
|---|---|---|---|---|
Meta · Fast, low-cost workhorse for chat, summaries, extraction and agents. |
mixerlead-ai/meta/llama-3.1-8b-instruct-fast | chat | 131K | 1× |
Meta · Meta's mixture-of-experts Llama 4 with a 131K context window, tool calling and vision. |
mixerlead-ai/meta/llama-4-scout-17b-16e-instruct | chat | 131K | 0.75× |
Meta · Meta's flagship 70B instruct model. Strong general reasoning, coding and tool calling. |
mixerlead-ai/meta/llama-3.3-70b-instruct-fp8-fast | reasoning | 24K | 3× |
Mistral AI · Mistral's 24B instruct model with 128K context, vision and tool calling. |
mixerlead-ai/mistralai/mistral-small-3.1-24b-instruct | chat | 128K | 0.5× |
Google · Google's efficient open model with 256K context, reasoning, vision and tool calling. |
mixerlead-ai/google/gemma-4-26b-a4b-it | chat | 256K | 0.25× |
OpenAI · OpenAI's open-weight reasoning model for production agents and complex tasks. |
mixerlead-ai/openai/gpt-oss-120b | reasoning | 128K | 0.75× |
OpenAI · Lower-latency OpenAI open-weight reasoning model. |
mixerlead-ai/openai/gpt-oss-20b | reasoning | 128K | 0.25× |
DeepSeek · Reasoning model distilled from DeepSeek R1. Shows its thinking; best for math and logic. |
mixerlead-ai/deepseek-r1-distill-qwen-32b | reasoning | 80K | 4× |
Qwen · Efficient Qwen3 mixture-of-experts model with reasoning and tool calling. |
mixerlead-ai/qwen/qwen3-30b-a3b-fp8 | reasoning | 33K | 0.5× |
Qwen · Code-specialised Qwen model for generation, review and refactoring. |
mixerlead-ai/qwen/qwen2.5-coder-32b-instruct | coding | 33K | 1× |
Zhipu AI · Fast multilingual model (100+ languages) with multi-turn tool calling. |
mixerlead-ai/zai-org/glm-4.7-flash | chat | 131K | 0.5× |
Moonshot AI · Moonshot's frontier agentic and coding model with 262K context. |
mixerlead-ai/moonshotai/kimi-k2.6 | reasoning | 262K | 3.5× |
Meta · Small, quick Meta model for simple replies, tagging and sorting at low cost. |
mixerlead-ai/meta/llama-3.2-3b-instruct | fast | 80K | 0.5× |
IBM · IBM's tiny, very low-cost model for RAG, classification and tool calling. |
mixerlead-ai/ibm-granite/granite-4.0-h-micro | fast | 131K | 0.25× |
Meta · Meta's smallest model: very fast and cheap for yes/no answers, labels and routing messages. |
mixerlead-ai/meta/llama-3.2-1b-instruct | fast | 60K | 0.25× |
BAAI · English embeddings (1024 numbers per text) for search, FAQ matching and similarity. |
mixerlead-ai/baai/bge-large-en-v1.5 | embedding | 1K | 0.2× |
BAAI · Multilingual embeddings, 1024 dimensions, 60K input tokens. |
mixerlead-ai/baai/bge-m3 | embedding | 60K | 0.1× |
Qwen · Compact multilingual embeddings, 1024 dimensions, 8K input tokens. |
mixerlead-ai/qwen/qwen3-embedding-0.6b | embedding | 8K | 0.1× |
Mistral AI · Deprecated upstream. Use mixerlead-ai/mistralai/mistral-small-3.1-24b-instruct instead. |
mixerlead-ai/mistral/mistral-7b-instruct | Legacy | 33K | 1× |