Spare Compute Serious Models
One OpenAI-compatible API for Qwen, Llama, DeepSeek and Mistral, served by idle computers around the world. Developers pay per token. Node owners get paid for every token their machine serves.
Change one base URL in the OpenAI SDK. No card, no subscription — just a wallet.
One API, Many Machines
A network of desktops and workstations running open-weight models, behind a router that picks a good one for every request.
Drop-in OpenAI API
Point the OpenAI SDK at /v1. Streaming, embeddings and JSON mode work as you'd expect.
Priced per model
Small models cost cents per million tokens. You pay for the tokens you receive — failed requests are free.
See live prices →Failover built in
If a machine drops mid-answer, another one picks up exactly where it left off. Your stream just keeps going.
How it works →Build On It, Or Power It
For developers
- Open-weight models from 8B to 70B, plus embeddings
- Prepaid credits in USDC on Base — top up from any wallet
- Per-key rate limits, usage history and a verified-nodes tier
- GPT, Claude and Gemini too, discounted, from sellers' spare capacity
For node owners
- A tray app runs models when your machine is idle, on your schedule
- Earn a share of every token you serve, paid in USDC to your wallet
- Apple Silicon and NVIDIA GPUs; the app picks models that fit
Live In Three Steps
-
Connect a wallet
Sign one message to prove it's yours. No password, no gas.
-
Add USDC credits
Send USDC on Base from the same wallet. Credits appear after confirmation.
-
Call the API
Create a key and swap your base URL. That's the whole migration.
import OpenAI from 'openai'; const client = new OpenAI({ apiKey: process.env.SPAREHASH_KEY, baseURL: 'https://your-coordinator.fly.dev/v1' }); const res = await client.chat.completions.create({ model: 'qwen3-8b', messages: [{ role: 'user', content: 'Hello!' }], });
Ready To Build?
Connect a wallet, add a few dollars of USDC and make your first call in minutes.