MIXERLEAD

Every Model, One API Key

Call any model below through the OpenAI-compatible MixerLead API at https://api.mixerlead.com/v1. Credits used = (prompt + completion tokens) × the model's credit multiplier.

ModelModel IDTypeContextCredit ×
Meta logoLlama 3.1 8B Instruct (Fast)
Meta · Fast, low-cost workhorse for chat, summaries, extraction and agents.
mixerlead-ai/meta/llama-3.1-8b-instruct-fast chat 131K 1×
Meta logoLlama 4 Scout 17B 16E
Meta · Meta's mixture-of-experts Llama 4 with a 131K context window, tool calling and vision.
mixerlead-ai/meta/llama-4-scout-17b-16e-instruct chat 131K 0.75×
Meta logoLlama 3.3 70B Instruct (Fast)
Meta · Meta's flagship 70B instruct model. Strong general reasoning, coding and tool calling.
mixerlead-ai/meta/llama-3.3-70b-instruct-fp8-fast reasoning 24K 3×
Mistral AI logoMistral Small 3.1 24B
Mistral AI · Mistral's 24B instruct model with 128K context, vision and tool calling.
mixerlead-ai/mistralai/mistral-small-3.1-24b-instruct chat 128K 0.5×
Google logoGemma 4 26B A4B
Google · Google's efficient open model with 256K context, reasoning, vision and tool calling.
mixerlead-ai/google/gemma-4-26b-a4b-it chat 256K 0.25×
OpenAI logoGPT-OSS 120B
OpenAI · OpenAI's open-weight reasoning model for production agents and complex tasks.
mixerlead-ai/openai/gpt-oss-120b reasoning 128K 0.75×
OpenAI logoGPT-OSS 20B
OpenAI · Lower-latency OpenAI open-weight reasoning model.
mixerlead-ai/openai/gpt-oss-20b reasoning 128K 0.25×
DeepSeek logoDeepSeek R1 Distill Qwen 32B
DeepSeek · Reasoning model distilled from DeepSeek R1. Shows its thinking; best for math and logic.
mixerlead-ai/deepseek-r1-distill-qwen-32b reasoning 80K 4×
Qwen logoQwen3 30B A3B
Qwen · Efficient Qwen3 mixture-of-experts model with reasoning and tool calling.
mixerlead-ai/qwen/qwen3-30b-a3b-fp8 reasoning 33K 0.5×
Qwen logoQwen2.5 Coder 32B
Qwen · Code-specialised Qwen model for generation, review and refactoring.
mixerlead-ai/qwen/qwen2.5-coder-32b-instruct coding 33K 1×
Zhipu AI logoGLM-4.7 Flash
Zhipu AI · Fast multilingual model (100+ languages) with multi-turn tool calling.
mixerlead-ai/zai-org/glm-4.7-flash chat 131K 0.5×
Moonshot AI logoKimi K2.6
Moonshot AI · Moonshot's frontier agentic and coding model with 262K context.
mixerlead-ai/moonshotai/kimi-k2.6 reasoning 262K 3.5×
Meta logoLlama 3.2 3B Instruct
Meta · Small, quick Meta model for simple replies, tagging and sorting at low cost.
mixerlead-ai/meta/llama-3.2-3b-instruct fast 80K 0.5×
IBM logoGranite 4.0 Micro
IBM · IBM's tiny, very low-cost model for RAG, classification and tool calling.
mixerlead-ai/ibm-granite/granite-4.0-h-micro fast 131K 0.25×
Meta logoLlama 3.2 1B Instruct
Meta · Meta's smallest model: very fast and cheap for yes/no answers, labels and routing messages.
mixerlead-ai/meta/llama-3.2-1b-instruct fast 60K 0.25×
BAAI logoBGE Large EN v1.5 (Embeddings)
BAAI · English embeddings (1024 numbers per text) for search, FAQ matching and similarity.
mixerlead-ai/baai/bge-large-en-v1.5 embedding 1K 0.2×
BAAI logoBGE-M3 (Multilingual Embeddings)
BAAI · Multilingual embeddings, 1024 dimensions, 60K input tokens.
mixerlead-ai/baai/bge-m3 embedding 60K 0.1×
Qwen logoQwen3 Embedding 0.6B
Qwen · Compact multilingual embeddings, 1024 dimensions, 8K input tokens.
mixerlead-ai/qwen/qwen3-embedding-0.6b embedding 8K 0.1×
Mistral AI logoMistral 7B Instruct v0.1 (Legacy)
Mistral AI · Deprecated upstream. Use mixerlead-ai/mistralai/mistral-small-3.1-24b-instruct instead.
mixerlead-ai/mistral/mistral-7b-instruct Legacy 33K 1×

Get an API key →   Read the API docs →