No platform deposit fee
Pay the exact quote within 10 minutes and the full USD amount is credited within about a minute once the network confirmations are reached; only the network fee is extra.
One crypto top-up becomes a USD balance that pays for per-token calls to 15 open-weight models through an OpenAI-compatible API and for a dedicated GPU prepaid by the month. Monero is accepted next to Bitcoin and stablecoins, and signup is a username and a password.
Token aggregators sell tokens; GPU clouds bill by the hour. One WeightsAPI balance buys both, and a flat H100 SXM month costs $1,352, where RunPod’s Secure Cloud rate of $3.49 an hour adds up to about $2,548 over 730 hours.
Create your account now; check each model on /status before topping up. Dedicated GPUs: activation date confirmed before you pay.
Pay the exact quote within 10 minutes and the full USD amount is credited within about a minute once the network confirmations are reached; only the network fee is extra.
BTC, XMR, ETH, SOL, LTC, TRX, USDT and USDC.
Pay per million tokens, or prepay a GPU month from $416.
Every confirmed top-up becomes USD credit in one balance. The playground and the API draw on it per million tokens, input and output priced separately, from $0.03 / $0.05 on Llama 3.1 8B to $1.20 / $1.20 on Llama 3.1 405B. No subscription; a request rejected before dispatch costs nothing.
The same balance prepays a dedicated GPU for one month: L40S (48 GB) $416, A100 (80 GB) $792, H100 SXM (80 GB) $1,352, H200 (141 GB) $1,832, 2 × H100 $2,568 or 8 × H100 $9,736. No per-token charges apply on that machine; the month starts on activation and does not renew automatically.
What USD 100 means in tokens, computed from our published rates at 80% input and 20% output: about 2.9 billion tokens on Llama 3.1 8B, 625 million on Qwen3 32B or 270 million on DeepSeek V3.
Checkout offers nine asset and network options. A transfer is credited within about a minute once it reaches its confirmation count, and USDT and USDC are quoted at their market rate, not an assumed $1.
| Asset | Network | Confirmations |
|---|---|---|
| BTC | Bitcoin | 1 |
| XMR | Monero | 10 |
| USDT | TRON (TRC-20) | 19 |
| USDT | Ethereum (ERC-20) | 12 |
| USDC | Ethereum (ERC-20) | 12 |
| ETH | Ethereum | 12 |
| SOL | Solana | 32 |
| LTC | Litecoin | 6 |
| TRX | TRON | 19 |
Send only on the network shown on your funding order: USDT on TRON and USDT on Ethereum are not interchangeable. Full rules are in billing and crypto deposits.
An exact payment inside the quote window qualifies for the full USD amount on your order, credited within about a minute once the network confirmations are reached; your only extra cost is the network fee. OpenRouter adds a 5% fee to crypto credit purchases, about $5 per $100. With us that $5 stays in tokens or GPU time. Read the billing rules.
Eight coins fund the same balance, so there is no swap step. OpenRouter takes USDC for crypto payments, and Venice sells API credit in USD or USDC. Holding XMR or USDT on TRON? Pay us directly. Check your coin and its confirmation count.
Prototype on the shared API, then move a steady workload to a dedicated GPU without a second account, check or provider. An H100 SXM month is $1,352, against about $2,548 for 730 hours at RunPod’s $3.49 Secure Cloud rate, and RunPod’s billing docs ask for any required KYC before a first crypto payment. Compare the six GPU configurations.
Signup is a username and a password: no email, phone, ID document or card. No KYC is not anonymity: keys, usage and payment records (asset, network, address, amount, transaction details) stay linked to your account. See exactly what signup asks for.
Every rate is in prices.json. We beat OpenRouter on Mixtral 8×22B and Llama 3.1 8B, and on output for DeepSeek V3, Gemma 3 27B and Qwen3 32B; OpenRouter is cheaper on Llama 3.3 70B and DeepSeek V4 Flash. See the model-by-model table.
| Item | What applies |
|---|---|
| Token price and unit | USD per 1M tokens, input and output separate; $0.03 (Llama 3.1 8B input) to $2.15 (DeepSeek R1 output) |
| GPU price and unit | Prepaid per month: $416 (L40S) to $9,736 (8 × H100); no automatic renewal |
| Minimum | USD 20 per top-up, every time; smaller remainders stay usable |
| Crypto assets and networks | BTC, XMR, USDT (TRON or Ethereum), USDC (Ethereum), ETH, SOL, LTC, TRX; no card payments |
| Platform deposit fee | None; the sender pays the network fee |
| Quote | Fixed for 10 minutes; stablecoins at market rate |
| Credit timing | Credited within about a minute once the network confirmations are reached. A pending payment does not unlock access |
| KYC scope | No ID or KYC step in the current account flow; account required, payment records kept |
| Availability | Available — deposit from $100, credited in about one minute |
| Refunds and withdrawals | No wallet withdrawals; GPU orders cancelled before activation are refunded to the balance |
| Wrong network or asset | No recovery policy |
All rates are on the pricing page.
| Provider | Crypto accepted | Fee on crypto | Entry amount | Identity | Dedicated GPU billing |
|---|---|---|---|---|---|
| WeightsAPI | BTC, XMR, USDT, USDC, ETH, SOL, LTC, TRX | None | USD 20 per top-up | Username and password; no KYC step | Monthly, $416 to $9,736 |
| OpenRouter | USDC; never refundable | 5% | Not published | Email or other contact | None published |
| Venice.ai | USDC for API credit; BTC for yearly plans only | Same rates as USD | Not published | Account for most features; wallet access without one | None published |
| NanoGPT | BTC, Lightning, LTC, XMR, ZEC, USDC, USDT, ETH, SOL and more | None | From $0.10 | No account needed | None published |
| Chutes | TAO direct, other crypto via Stripe; TAO deposits non-refundable | Not published | Not published | Not addressed in its terms | Per-second private deployments |
| RunPod | Via processors; coins not published | Not published | From $10 | Any required KYC before first crypto payment | Per second; H100 SXM $3.49/h (Secure Cloud) |
Where another option fits better: NanoGPT to start with cents and no account; Venice for USDC wallet access without an account; RunPod for a GPU used under roughly 380 to 400 hours a month on Secure Cloud, or about 500 to 530 on Community Cloud, where per-second billing costs less.
| Model | WeightsAPI | OpenRouter | Lower price |
|---|---|---|---|
| Mixtral 8×22B | $0.60 / $0.60 | $2.00 / $6.00 | WeightsAPI |
| Llama 3.1 8B | $0.03 / $0.05 | $0.05 / $0.08 | WeightsAPI |
| DeepSeek V3 (0324) | $0.25 / $0.85 | $0.25 / $1.00 | WeightsAPI on output |
| Gemma 3 27B | $0.10 / $0.15 | $0.08 / $0.45 | WeightsAPI on output |
| Qwen3 32B | $0.15 / $0.20 | $0.08 / $0.28 | WeightsAPI on output |
| Llama 3.3 70B | $0.35 / $0.40 | $0.10 / $0.32 | OpenRouter |
| DeepSeek V4 Flash | $0.09 / $0.18 | $0.049 / $0.098 | OpenRouter |
OpenRouter rates exclude its 5% crypto fee. On Llama 3.3 70B, Venice charges $0.70 / $2.80 and Together AI $1.04 / $1.04. More rows: our OpenRouter alternative comparison.
| GPU | WeightsAPI per month | RunPod Secure Cloud per hour | RunPod over 730 h (computed) | Break-even (computed) |
|---|---|---|---|---|
| L40S | $416 | $1.09 | about $796 | about 382 h |
| H100 SXM | $1,352 | $3.49 | about $2,548 | about 387 h |
| H200 | $1,832 | $4.59 | about $3,351 | about 399 h |
RunPod’s Community Cloud is cheaper ($0.79, $2.69 and $3.59 an hour), which moves break-even to about 527, 503 and 510 hours (computed). Not like-for-like: our month is managed inference capacity, with no shell access, backups, failover or autoscaling by default, while RunPod bills per second.
We do not sell GPT or Claude. We sell open-weight models behind an OpenAI-compatible endpoint: set the base URL to https://weightsapi.com/v1 with a WeightsAPI key and an exact model ID. Chat Completions, Completions and the model list are covered; the Responses and Assistants APIs are not drop-in. Check the migration guide.
Open-weight picks by job, in USD per 1M input / output tokens. They are different models from GPT and Claude, so test them on your own prompts:
What will it really cost? Your top-up plus the network fee, then tokens at published rates (reserved for input plus maximum output, settled on confirmed usage) or a flat GPU month. A spending cap on each key stops a runaway script.
Can I trust it with my money? Rates, terms and the records we keep (on the trust page) are published, gaps included; no SLA is offered. Start with one USD 20 top-up and scale from there.
Is my model available? Shared inference is in launch preview: check your model on the status page (refreshed every 60 seconds) before topping up. For GPUs, the activation date is confirmed before you pay.
Will anyone ask for ID? Not in the current account flow, for tokens or GPUs. A mismatched transfer is reviewed, and a support report may include a transaction ID, which can link activity to a wallet.
What are the conditions? Credit cannot be withdrawn to a wallet; cancelling a GPU order before activation returns the payment to your balance (refund FAQ). A GPU month does not renew automatically.
Choose WeightsAPI if you hold BTC, XMR, USDT or another listed coin and want open-weight inference and GPU capacity without a platform fee, a card or an ID check, especially if your usage will grow into a monthly GPU or leans on the models where our rate is lower.
Look elsewhere if you want to test with less than USD 20, need GPT or Claude themselves, need an SLA, or need unused crypto returned to a wallet.
What to do now: create a username account, check your model on /status, then top up from USD 20 in the coin you hold. For a GPU, pick a configuration on dedicated GPUs and get the activation date confirmed before you pay.
Username, password, recovery code. Check your model on /status, then fund from USD 20 in the coin you hold, with no platform deposit fee.
Answers to check before your first top-up.
Yes. Choose BTC at checkout and send the quoted amount for a top-up of USD 20 or more; it is credited within about a minute once its 1 network confirmation is reached. The credit pays for tokens on 15 open-weight models or for a dedicated GPU month.
Yes, on TRON (TRC-20, 19 confirmations) or Ethereum (ERC-20, 12 confirmations); USDC is accepted on Ethereum. Stablecoins are quoted at their market rate, not an assumed one dollar. Use only the network on your order.
No, we do not resell OpenAI or Anthropic models. We sell open-weight models through an OpenAI-compatible API, so the OpenAI SDK pointed at https://weightsapi.com/v1 is compatible for Chat Completions and Completions.
Yes, for the open-weight DeepSeek models we host: DeepSeek V3 at $0.25 / $0.85, DeepSeek R1 at $0.50 / $2.15 and DeepSeek V4 Flash at $0.09 / $0.18 per million input / output tokens. This is our service, not DeepSeek’s own platform.
Yes. Dedicated GPUs cost $416 (L40S) to $9,736 (8 × H100) per month, paid from confirmed crypto credit, with no ID step in the current account flow. The activation date is confirmed before you pay, and the prepaid month starts on activation.
It is a floor per top-up, not a balance to keep; smaller remainders stay usable. With no platform deposit fee, fewer and larger transfers also cut your network fees and confirmation waits. To spend only a few dollars, a provider with a lower entry amount fits better.
About a minute once the network confirmations are reached; the count depends on your asset, from 1 for BTC to 32 for SOL. If a deposit stays pending, follow the support guide rather than paying twice.
There is no recovery or refund policy for wrong-asset or wrong-network transfers, so check asset, network, address and amount against your funding order before sending. If the 10-minute quote expires before you send, request a new one.
What signup asks for, what we record and what can still trigger a check.
Model-by-model prices, crypto fees and where each option wins.
Six monthly configurations, from L40S to 8 × H100.
Dolphin, Hermes and open models for creative and roleplay work.
Quotes, confirmations, ledger and withdrawal rules.
Per-model availability, refreshed every 60 seconds.