Key points at a glance
- BYOK: use your own Together account and quota.
- One endpoint reaches a large set of open-source models—no need to build your own GPU inference.
- A middle ground for "I want open models but don't want to operate them."
How to connect
In model settings choose an OpenAI-compatible provider, enter your Together API key, point the base URL at Together's endpoint, pick a specific model name, then set it as default or assign per character.
Who it's for
Users who want the flexibility of open models but don't want to buy and operate their own GPUs. With Enclave self-hosted, your data stays under your control while inference runs on Together's cloud.
Key facts
- Integration
- BYOK (OpenAI-compatible endpoint)
- Billing
- Billed at Together's official rates; no Enclave markup
- Best for
- Cloud-hosted open models, no GPU ops
- Data flow
- Direct from your instance when self-hosted
Ready to try it?
Open it in your browser — no credit card, no install.
Related
Connect other models
- Connect OpenAI (GPT-series) models in Enclave
- Connect Anthropic Claude models in Enclave
- Connect Google Gemini models in Enclave
- Connect DeepSeek models in Enclave
- Connect local Ollama models in Enclave (fully offline capable)
- Connect self-hosted vLLM inference in Enclave
- Use Mistral AI models in Enclave
- Use Groq in Enclave (ultra-low-latency inference)
- Use OpenRouter in Enclave (one key, many models)
- Use self-hosted Hugging Face TGI in Enclave