earthruntimedeveloper preview by Provocative

Fast Qwen 3.6 inference without marketplace overhead.

We operate the hardware ourselves, giving you direct pricing, predictable performance, and full visibility into the model you are running.

Start with 100,000 free tokens. No card required.

Code sent to
Didn't get it? Send a new code β€” it expires in 15 minutes.
Generating your key…
This usually takes about 30 seconds. Keep this page open.
Check your inbox.
This address already joined the beta, so your key is on its way to by email.
Your key, shown once
performance that holds up under load
Median latency
2.2s vs 3.5s
P95 latency
3.6s vs 13.5s
Median TTFT at load
196ms

Aggregate throughput at 64 concurrent requests. Requests originated from the same independent cloud environment. Higher throughput and lower error rates are better.

earthruntime2,237 tok/s Β· 0% errors
Atlas Cloud1,736 tok/s Β· 77% errors
Open Router1,716 tok/s Β· 0% errors
Parasail1,327 tok/s Β· 74% errors
tested on real workloads
β€œThere were no latency issues using Qwen.”

Agentic coding

A developer tested Qwen on a Rust DSP mastering tool containing a deliberately misleading hardcoded ID. The model attempted the obvious fix, saw its tests fail, backtracked, identified the underlying cause, and produced a patch that two other models independently reviewed.

β€œMuch better than the OpenRouter options we tested.”

Production spam-call classification

After benchmarking earthruntime against OpenRouter-hosted alternatives in production, the team made earthruntime its primary provider and retained the marketplace as an outage fallback.

run your first request
curl https://api.earthruntime.com/v1/chat/completions \
  -H "Authorization: Bearer $EARTHRUNTIME_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen3.6-35b",
    "messages": [{"role": "user", "content": "Explain why tail latency matters for production inference in three sentences."}]
  }'
Trouble pasting? Use the single-line version.
curl https://api.earthruntime.com/v1/chat/completions -H "Authorization: Bearer $EARTHRUNTIME_KEY" -H "Content-Type: application/json" -d '{"model":"qwen3.6-35b","messages":[{"role":"user","content":"Explain why tail latency matters for production inference in three sentences."}]}'

Works with your existing OpenAI client
Change the base URL and API key. Your request and response formats stay the same.

Qwen 3.6 served in native BF16
262K context. No undisclosed quantization. Reasoning is disabled by default for lower latency and can be enabled when the workload benefits from it.

add credit when you need it
Checkout cancelled β€” nothing was charged. Pick a pack below whenever you're ready.

No subscription or account setup required.

$1
~1M tokens
$5
~5M tokens
$20
~20M tokens
Need a different amount or higher-volume pricing? contact@earthruntime.com

Direct infrastructure changes the economics

We operate inference hardware in cities, then look for local uses for its physical outputs. Our first system pairs compute with direct-air carbon capture and supplies recovered CO2 to nearby hospitality businesses.

Faster inference is the product today. A more useful relationship between cities and compute is what we are building toward.

We love to bottle up the CO2 for soda to sell to all our favorite local bars and restaurants ;)

earthruntime is a product of Provocative Science Holdings, Inc