OpenWeightsTerminal

Open weights. Uncensored models.

Private inference.
Open to everyone.

Choose your model and where it runs. Run locally to keep prompts on your device, or use independent hosts on the network.

Start earning

Hosts keep 80% of each completed job. You choose your rate.

No chat logs. Prompts are erased when a host picks them up. Replies are erased 10 minutes after they finish. We keep only token counts for billing. How privacy works

Your model. Your call.

Text inference on the network. Image, voice, and audio models to explore locally.

Explore models
  • Text & codeOn the network
  • ImagesExplore locally
  • VoiceExplore locally
  • Speech-to-textExplore locally
  • AudioExplore locally
Uncensored, filtered, or unknown? Understand the labels.

Checking availability…

    Rates apply to input and output tokens. Host-reported speed; endpoint checks do not verify model weights.

    TRENDING

    1. 1DeepSeek v4.1 FlashHOT
    2. 2Empero Qwen3.8-35B-A3BHOT
    3. 3Gemma-4 26B UncensoredHOT
    4. 4Gemma-4 26B-A4B-itHOT
    5. 5GLM 5.3 FlashHOT

    STREET · DGX SPARK

    Sample data · 2026-10-03

    $8,050▼ 6.7% 7d

    TAPE

    @fillagrew “351 tokens/sec on Qwen3.5 2B at Q4_K_M”

    @Youssofal_ “Peak Speed: 126.5 TPS.”

    Street stacks

    since Sep 20

    View all →

    Payback ranked

    US price. Each model at its suggested price, host keeps 80%. 20% duty. Left is faster.

    Full table →
    1Gemma-4 26B-A4B-it · RTX 5090 32GB17 mo71%
    2Qwen3.8-27B · RTX 5060 Ti 16GB2.1 yr47%
    3Qwen3.8-27B · RTX 5060 Ti 16GB2.7 yr37%
    4Qwen3.5 2B · RTX 5090 32GB2.9 yr34%
    5Qwen3.8-27B · DGX Spark3.4 yr30%
    6Qwen3.8-27B · DGX Spark4.9 yr20%
    7Qwen3.6-35B-A3B · M4 Pro5.3 yr19%
    8Qwen3.5-4B · M4 Pro5.5 yr18%