Skip to main content
Use hosted models from the dashboard’s Endpoints catalog.

Discover Models

The catalog is public. Include a token to see models available to your workspace.

Call a Chat Model

Install the OpenAI Python client and set BEAM_MODEL to a chat model ID from the catalog:
Pass stream=True for streaming responses. Supported options vary by model.

Other Model Types

Use /v1/embeddings for embeddings or /v1/models/{model-id}/invoke for models with custom schemas. The model’s catalog page includes request examples, image routes, and pricing. Inference requires a workspace token and credits for prepaid workspaces. Keep the token on your server. To deploy your own model, use an endpoint or container.