Choose a model, call the API, and start iterating. There is no cluster to provision and no GPU to select.
Leading open-source models or your own LoRA weights, one API
Access a curated catalog of open-source models, or bring your own LoRA weights and serve them side by side. Switch between them without changing your integration.
Tracing, evals, and observability
Every request can be traced and evaluated through native observability features, so teams can debug and improve AI applications and agents without extra instrumentation.
Playground access with zero configuration
Explore and compare models directly in the playground before writing a line of code. No endpoints to configure and no keys to manage.




