Use one profile across model APIs
A gateway key is assigned to an access profile, not to a client API. An entitled application can call Forge with Chat Completions, Responses, or Anthropic Messages while Forge selects the configured provider operation and translates supported text and tool requests. Bedrock Converse and Gemini Generate Content are also supported client and provider operations. The same profile, identity, policy, and budget apply to translated requests. Unsupported fields or content are rejected rather than silently removed. Streaming responses use the client’s API format. For incomplete Chat tool histories, Forge records a bounded repair and supplies missing results before policy checks and the provider call. Orphan results are removed, and the last duplicate result within a tool exchange is kept. Results are never moved across conversation turns. For Claude models, Forge checks effort levels against the selected model before sending a request. Older models may supportlow through high but not xhigh or max; Haiku 4.5 does not support effort. Forge rejects per-message effort on known unsupported models and adds Anthropic’s required beta header when it is used. A translated continuation with signed Claude 4.5 thinking uses that generation’s manual thinking format and requires enough output tokens for its minimum budget.
The Models endpoint lists models available to that key. Your application chooses a model name; Forge chooses the eligible destination. For provider-specific features, check the selected provider and model’s capabilities before relying on them.