TokenHub

TokenHub

7

@tokenhub

Self-hosted open-weight LLM inference with an OpenAI-compatible API. Owned GPU capacity serving open models with transparent per-token pricing and steady low-latency responses for chat completions.

Joined 2026/8/10No data yetRPM cap 20

Service status

Model
24 hours agoNow
Availability
qwen3-4b
AvailableDegradedUnavailableNo Data

Models (1)

ModelInput ($/M)Output ($/M)Cache ($/M)Chat
$0.05$0.15
Read$0.005

Ratings (0)