Discover Models
Call a Chat Model
Install the OpenAI Python client and setBEAM_MODEL to a chat model ID from the catalog:
stream=True for streaming responses. Supported options vary by model.
Other Model Types
Use/v1/embeddings for embeddings or /v1/models/{model-id}/invoke for models with custom schemas. The model’s catalog page includes request examples, image routes, and pricing.
Inference requires a workspace token and credits for prepaid workspaces. Keep the token on your server.
To deploy your own model, use an endpoint or container.