JouleCloud routes AI work to available GPUs where electricity costs less. That lowers the cost of inference today and builds toward a future where compute demand accelerates clean power.
Run AI work on available GPUs where electricity costs less.
Make more locations viable for AI data centers.
Reduce the carbon impact of every token as AI scales.
The same request can run on more than one available GPU.
Use the latest electricity price and confirm that a GPU is ready.
Favor lower-cost power. If price data is missing, keep routing.
Send the request to a ready GPU and try another if it fails.
Create an API key in the console, then send a request to JouleCloud. These examples cover cURL and the official OpenAI Python and JavaScript clients.
curl https://api.jouledns.com/v1/chat/completions \
-H "Authorization: Bearer $JC_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"qwen2.5-7b-instruct","messages":[{"role":"user","content":"Hello"}]}'Live electricity prices help JouleCloud favor lower-cost power.
Requests go to available GPUs and can retry elsewhere if one fails.
API keys, budgets, and usage stay inside your organization.