
Qwen3.8-Flash-NextMost runSolo load79.5
Qwen3.8-27B27B, SGLang72
Ling-3.0-FlashGLM 5.3 FlashEXL3 K2, one box18.6
Open weights. Uncensored models.
Choose your model and where it runs. Run locally to keep prompts on your device, or use independent hosts on the network.
Hosts keep 80% of each completed job. You choose your rate.
No chat logs. Prompts are erased when a host picks them up. Replies are erased 10 minutes after they finish. We keep only token counts for billing. How privacy works

Text inference on the network. Image, voice, and audio models to explore locally.
Checking availability…
Rates apply to input and output tokens. Host-reported speed; endpoint checks do not verify model weights.
TRENDING
STREET · DGX SPARK
Sample data · 2026-10-03
$8,050▼ 6.7% 7d
TAPE
@fillagrew “351 tokens/sec on Qwen3.5 2B at Q4_K_M”
@Youssofal_ “Peak Speed: 126.5 TPS.”
since Sep 20





Every liberated weight. A number is a quoted tok/s.
Quoted
No quote yet
US price. Each model at its suggested price, host keeps 80%. 20% duty. Left is faster.
| 1 | Gemma-4 26B-A4B-it · RTX 5090 32GB | 17 mo | 71% |
| 2 | Qwen3.8-27B · RTX 5060 Ti 16GB | 2.1 yr | 47% |
| 3 | Qwen3.8-27B · RTX 5060 Ti 16GB | 2.7 yr | 37% |
| 4 | Qwen3.5 2B · RTX 5090 32GB | 2.9 yr | 34% |
| 5 | Qwen3.8-27B · DGX Spark | 3.4 yr | 30% |
| 6 | Qwen3.8-27B · DGX Spark | 4.9 yr | 20% |
| 7 | Qwen3.6-35B-A3B · M4 Pro | 5.3 yr | 19% |
| 8 | Qwen3.5-4B · M4 Pro | 5.5 yr | 18% |