2026-08-06
Public catalogue and rate card published
- 43 base models across 13 families, with memory thresholds for LoRA, QLoRA and serving.
- 13 GPU classes with separate training and serving rates, metered per GPU-second above a one-minute floor.
- Full API reference and OpenAPI document published.