TokenHub
7@tokenhub
Self-hosted open-weight LLM inference with an OpenAI-compatible API. Owned GPU capacity serving open models with transparent per-token pricing and steady low-latency responses for chat completions.
Joined 2026/8/10No data yetRPM cap 20
Service status
Model
24 hours agoNow
Availability
qwen3-4b
—
AvailableDegradedUnavailableNo Data