Swfte Connect / Providers
Every model provider through one API, with your own keys if you prefer
Connect routes one OpenAI-compatible request to commercial APIs, fast-inference hosts, European clouds and open-weight models you run yourself. You choose the model; Connect handles the provider, the credential and the fallback.
How a request reaches a provider
Connect keeps a registry that maps a model name to a provider adapter. The application only sees one API.
A request names a model, either with a provider prefix such as provider:model or by a well-known model name that the registry recognizes. Connect resolves the provider, picks the credential, calls the provider in its native protocol and returns the answer in the OpenAI-compatible format. Chat completions, streaming, embeddings, image generation, speech, transcription and moderation go through the same gateway.
Because the interface does not change, switching a model, or a provider, is a parameter change in the request, not a code migration. That is the practical value of a gateway: your product code depends on one contract while the provider market keeps moving.
Provider families
The registry covers several kinds of provider. Examples below are drawn from the Connect registry; the full count is <provider count - founder to fill>.
| Family | Examples | Typical reason to use |
|---|---|---|
| Frontier model APIs | OpenAI, Anthropic, Google, xAI, Mistral, Cohere, Perplexity | Highest capability on hard tasks |
| Fast inference hosts | Groq, Cerebras, SambaNova, Fireworks, Together | Low latency and high throughput for open-weight models |
| Open-weight model hosts | DeepInfra, Nebius, NVIDIA, Novita, Hyperbolic, Hugging Face | Pay-per-use access to Llama-, Qwen- and Mistral-class models |
| European providers | Scaleway, OVHcloud, Mistral | Keeping the provider relationship, and often the processing, in Europe |
| Cloud and platform providers | Amazon Nova, Databricks, Cloudflare Workers AI, GitHub Models | Using an existing cloud relationship |
| Regional model makers | DeepSeek, Moonshot, Zhipu AI, Alibaba DashScope, SiliconFlow | Specific model families, subject to your own jurisdiction review |
| Aggregators | OpenRouter, Vercel AI Gateway | A long tail of models behind one upstream key |
| Self-hosted open-weight | Models you host on your own or Swfte GPUs | Full control of where the data is processed |
A provider appearing in the registry means Connect can route to it. Which providers a given workspace uses depends on the keys connected and the policy applied. The providers Swfte itself relies on are listed on the sub-processors page.
Open-weight models you run yourself
Connect is not only a router to other people's APIs. It also fronts models you host.
The self-hosted catalogue spans far more than chat. It covers text generation, embeddings, reranking, speech-to-text, text-to-speech, OCR, image and video generation, safety classifiers and several specialist categories. Examples in the registry include Whisper Large V3 for transcription, BGE-M3 for multilingual embeddings, Qwen3-Reranker for reranking, PaddleOCR-VL for document OCR and Granite Guardian as a safety model. Hosting them on GPUs you control, through the deploy models catalogue, puts them behind the same API as the commercial providers.
This is how sovereignty and convenience coexist. Sensitive workloads route to a model running in your own environment; low-risk workloads route to the best commercial option; both use one client, one set of keys for your applications and one audit trail. See self-deployment for running the gateway next to those models.
Bring your own keys
You can use Swfte's pooled access for simple billing, or connect your own provider accounts.
- Your key first. If your workspace has connected its own key for a provider, Connect uses it.
- Platform access second. If not, Connect uses Swfte's pooled provider access, where it is offered.
- Custom endpoints. A connected key can carry its own base URL, which is how you point a provider slot at a private or self-hosted endpoint.
Your keys are stored encrypted in a secrets manager, decrypted only at call time and cached briefly. They are never sent to the client. Using your own keys means you keep your commercial terms, quotas and data agreements with each provider, and the provider bills you directly. The security page covers credential handling in more detail.
Routing and fallbacks
Provider outages, rate limits and exhausted balances are normal events. Connect treats them as routing inputs.
- Fallback chains. When a primary model's provider is unavailable, rate-limited or out of funds, Connect walks an ordered chain of alternatives. Defaults are organized by pricing tier and cross providers, so one provider outage does not take a tier down.
- Your own chains. A workspace routing rule must define its own fallback chain, with priority and time-of-day conditions, so you decide the order and the models.
- Routing strategies. Strategies include cost-optimized, weighted, round-robin, least-used, priority and scoring-based selection, applied by name to a rule.
- Consensus. For high-stakes calls, a request can fan out to several providers and a decision step selects the best answer.
- Budget-aware downgrade. When a usage cap is reached, Connect can switch to a cheaper model rather than stop.
Fallbacks are a governance tool as well as a reliability tool. A chain restricted to approved models is how you make sure an outage never sends a request to a provider your policy has not cleared. That is the basis of EU-first routing.
Regional and EU routing
Connect is designed to let you route by where processing happens. In practice that means restricting a workspace to providers whose processing location you have confirmed, preferring European providers and European-hosted open-weight capacity, and running self-hosted models for the most sensitive traffic. Customer data in the managed service is stored in AWS eu-west-1 (Ireland) today. The details, and the limits of what is enforced versus designed for, are on the EU routing page.
Choosing a route
| If you need | Route to | Keys |
|---|---|---|
| Best quality on difficult reasoning | A frontier model API | Your own or pooled |
| Low latency at volume | A fast-inference host | Your own or pooled |
| Data processed in Europe | A European provider or your own EU capacity | Your own |
| Data never leaves your boundary | A self-hosted open-weight model | None needed for the model |
| Predictable spend | Cost-optimized strategy with usage caps | Either |
| Resilience | A fallback chain across two or more providers | Either |
Model availability, pricing and provider terms change. Check a provider's current terms before sending regulated data to it. Swfte provides the technical controls, governance mechanisms and evidence required to deploy AI within an organization's applicable regulatory, security and policy requirements. The exact posture depends on the customer's use case, jurisdiction, deployment and configuration.
Frequently asked questions
How many providers does Swfte Connect support?
Connect's registry maps model names to a broad set of commercial, regional and open-weight providers. The published count is <provider count - founder to fill>. Which providers a workspace can use depends on the keys it connects.
Can I bring my own API keys?
Yes. A workspace can connect its own provider keys, which Connect uses first. Without them, Connect uses Swfte's pooled access where it is offered. Keys are stored encrypted and never sent to the client.
Can I route to a model I host myself?
Yes. Open-weight models you host on your own or Swfte GPUs sit behind the same API, and a connected key can carry a custom base URL to point a provider slot at a private endpoint.
What happens when a provider is down?
Connect walks a fallback chain. Defaults are organized by pricing tier and cross providers, and you can define your own chain per workspace with priority and time-of-day conditions.
Is the API compatible with the OpenAI format?
Yes. Chat completions, streaming, embeddings, images, speech, transcription and moderation use an OpenAI-compatible request format, so moving an integration is mostly a base-URL and key change.
Does using a provider through Connect change that provider's data terms?
No. The provider processes the request under its own terms. Check each provider's terms, and the sub-processors page for the providers Swfte itself relies on.
Connect in the platform
Models and routing layer
Connect as the model gateway and routing layer of the platform.
Sovereign infrastructure
Where AI runs, from managed to dedicated.
Dedicated cloud
Private capacity operated for you.
GPU and inference
GPU capacity for the models Connect routes to.
Deploy open-source models
The catalogue of open-weight models you can host.
LLM API guides
Provider and model reference pages.
Trust centre
What Swfte can show today.
Sub-processors
Providers and services Swfte itself relies on.
Route every model through one governed gateway
Start managed in minutes, or plan a deployment inside your own boundary with the team.