Anthropic (Claude)
Default model:claude-opus-4-6
Setup:
Pricing (per 1M tokens):
OpenAI (GPT-4o)
Default model:gpt-4o
Setup:
Pricing (per 1M tokens):
Smart routing
If you’ve configured multiple providers viakong setup, Kong uses your default provider. Override at runtime:
Rate limiting
Kong uses automatic token-bucket rate limiting to stay within API rate limits. No configuration needed — it proactively sleeps between requests when approaching the limit. This prevents429 Too Many Requests errors during large analyses.
Further reading
- Custom Endpoints — use local or third-party models
- LLM Models & Pricing — full pricing and limits reference
- Setup Wizard — configure providers interactively

