Overview
Beyond the built-in model providers covered in Model Config, Sulus supports Groq as a dedicated low-latency model provider, plus a broader umbrella of "any OpenAI-compatible endpoint" support that covers your own server or a range of third-party inference services. This page covers both. For a faster way to get started without picking individual providers, see Model Presets; for the general bring-your-own-endpoint pattern in more depth, see Custom LLM (Bring Your Own Model Endpoint).