Key points at a glance
- BYOK: use your own Groq account and quota.
- Ultra-low-latency inference, ideal for real-time, high-frequency conversational characters.
- Mix with other providers: route speed-critical characters to Groq and others elsewhere.
How to connect
In model settings choose Groq (an OpenAI-compatible provider), enter your API key, point the base URL at Groq's endpoint, then set it as default or assign per character. Enclave lets multiple providers coexist, so you can move just the latency-sensitive characters to Groq.
Who it's for
Users who value response speed and want chat to feel smooth—especially latency-sensitive scenarios like voice companionship and real-time group chat. With Enclave's memory and proactivity, you get both speed and being remembered.
Key facts
- Integration
- BYOK (OpenAI-compatible endpoint)
- Billing
- Billed at Groq's official rates; no Enclave markup
- Highlight
- Ultra-low inference latency
- Data flow
- Direct from your instance when self-hosted
Ready to try it?
Open it in your browser — no credit card, no install.
Related
Connect other models
- Connect OpenAI (GPT-series) models in Enclave
- Connect Anthropic Claude models in Enclave
- Connect Google Gemini models in Enclave
- Connect DeepSeek models in Enclave
- Connect local Ollama models in Enclave (fully offline capable)
- Connect self-hosted vLLM inference in Enclave
- Use Mistral AI models in Enclave
- Use Together AI in Enclave (cloud-hosted open models)
- Use OpenRouter in Enclave (one key, many models)
- Use self-hosted Hugging Face TGI in Enclave