Get Started

Build with Pearl Inference

OpenAI-compatible inference for open-weight models, a flat 10% below Together AI, on simple prepaid credits.

Pearl Inference runs production inference for a curated set of open-weight models on the Pearl network. You bring any OpenAI SDK (or plain HTTP), point it at our base URL, and pay per token from a prepaid balance, no GPU provisioning, no contracts, no switching costs.

Get started in minutes

Not sure where to start? Pick a model with the model selection guide, run the quickstart, and skim Concepts to understand keys, credits, and billing.

What you can build

The platform today

The platform currently serves chat-model inference through a deliberately curated catalog of 5 production models spanning reasoning, coding, general chat, and vision. Every model is served on the same OpenAI-compatible endpoint and billed per token against your organization's prepaid credits.

Resources

Pearl Inference Docs