No KYC
No identity-document upload or KYC step in the current sign-in flow. Access and usage remain linked to your account.
Choose the right balance of capability, context and cost.
No identity-document upload or KYC step in the current sign-in flow. Access and usage remain linked to your account.
WeightsAPI adds no content filter at the gateway. A model’s own training can still lead it to refuse a request.
Published model sources verified 2026-09-04. Check service status for current availability. API model list ↗
| Model | Input / 1M | Output / 1M | Context | Type |
|---|---|---|---|---|
| LLlama 3.1 8B | $0.03 | $0.05 | 128K* | Text |
| DDeepSeek V4 Flash | $0.09 | $0.18 | 128K* | Text |
| GGemma 3 27B | $0.10 | $0.15 | 128K* | Text + images |
| MMistral Small 3 | $0.10 | $0.20 | 32K* | Text |
| QQwen3-Coder-Next | $0.12 | $0.80 | 128K* | Text |
| HHermes 4 70B | $0.13 | $0.40 | 128K* | Text |
| MMistral Small 4 | $0.15 | $0.60 | 128K* | Text + image |
| QQwen3 32B | $0.15 | $0.20 | 128K* | Text |
| DDolphin Mistral 24B Venice Edition | $0.20 | $0.90 | 128K* | Text + image |
| QQwen3 235B A22B | $0.20 | $0.60 | 128K* | Text |
| DDeepSeek V3 | $0.25 | $0.85 | 128K* | Text |
| LLlama 3.3 70B | $0.35 | $0.40 | 128K* | Text |
| DDeepSeek R1 | $0.50 | $2.15 | 128K* | Text |
| MMixtral 8×22B | $0.60 | $0.60 | 64K* | Text |
| LLlama 3.1 405B | $1.20 | $1.20 | 128K* | Text |
Fast classification, extraction and lightweight chat.
A text model for long-context reasoning, coding and tool workflows, with a configured 1,048,576-token window.
Image understanding and multilingual conversation.
Efficient text tasks, function calling and European-language chat.
A text-only coding model for codebase exploration and tool workflows, with a published 262,144-token native context.
Creative conversation, roleplay and reasoning with configurable instructions, published by Nous Research.
An official Mistral model combining text and image input, tool calling and configurable reasoning, with a published 262,144-token window.
A balanced choice for code, multilingual chat and reasoning.
Conversations and characters with application-controlled instructions, from Dolphin and Venice.
Strong general-purpose instruction following and tool use.
Code generation, complex instructions and structured tasks.
General-purpose assistants, summarization and tool use.
Difficult reasoning, mathematics and code analysis.
Multilingual generation and function-calling workflows.
Demanding general-purpose generation and complex instructions.