Swfte Connect / Providers

Every model provider through one API, with your own keys if you prefer

Connect routes one OpenAI-compatible request to commercial APIs, fast-inference hosts, European clouds and open-weight models you run yourself. You choose the model; Connect handles the provider, the credential and the fallback.

How a request reaches a provider

Connect keeps a registry that maps a model name to a provider adapter. The application only sees one API.

A request names a model, either with a provider prefix such as provider:model or by a well-known model name that the registry recognizes. Connect resolves the provider, picks the credential, calls the provider in its native protocol and returns the answer in the OpenAI-compatible format. Chat completions, streaming, embeddings, image generation, speech, transcription and moderation go through the same gateway.

Because the interface does not change, switching a model, or a provider, is a parameter change in the request, not a code migration. That is the practical value of a gateway: your product code depends on one contract while the provider market keeps moving.

Provider families

The registry covers several kinds of provider. Examples below are drawn from the Connect registry; the full count is <provider count - founder to fill>.

Provider families Connect routes to
FamilyExamplesTypical reason to use
Frontier model APIsOpenAI, Anthropic, Google, xAI, Mistral, Cohere, PerplexityHighest capability on hard tasks
Fast inference hostsGroq, Cerebras, SambaNova, Fireworks, TogetherLow latency and high throughput for open-weight models
Open-weight model hostsDeepInfra, Nebius, NVIDIA, Novita, Hyperbolic, Hugging FacePay-per-use access to Llama-, Qwen- and Mistral-class models
European providersScaleway, OVHcloud, MistralKeeping the provider relationship, and often the processing, in Europe
Cloud and platform providersAmazon Nova, Databricks, Cloudflare Workers AI, GitHub ModelsUsing an existing cloud relationship
Regional model makersDeepSeek, Moonshot, Zhipu AI, Alibaba DashScope, SiliconFlowSpecific model families, subject to your own jurisdiction review
AggregatorsOpenRouter, Vercel AI GatewayA long tail of models behind one upstream key
Self-hosted open-weightModels you host on your own or Swfte GPUsFull control of where the data is processed

A provider appearing in the registry means Connect can route to it. Which providers a given workspace uses depends on the keys connected and the policy applied. The providers Swfte itself relies on are listed on the sub-processors page.

Open-weight models you run yourself

Connect is not only a router to other people's APIs. It also fronts models you host.

The self-hosted catalogue spans far more than chat. It covers text generation, embeddings, reranking, speech-to-text, text-to-speech, OCR, image and video generation, safety classifiers and several specialist categories. Examples in the registry include Whisper Large V3 for transcription, BGE-M3 for multilingual embeddings, Qwen3-Reranker for reranking, PaddleOCR-VL for document OCR and Granite Guardian as a safety model. Hosting them on GPUs you control, through the deploy models catalogue, puts them behind the same API as the commercial providers.

This is how sovereignty and convenience coexist. Sensitive workloads route to a model running in your own environment; low-risk workloads route to the best commercial option; both use one client, one set of keys for your applications and one audit trail. See self-deployment for running the gateway next to those models.

Bring your own keys

You can use Swfte's pooled access for simple billing, or connect your own provider accounts.

  1. Your key first. If your workspace has connected its own key for a provider, Connect uses it.
  2. Platform access second. If not, Connect uses Swfte's pooled provider access, where it is offered.
  3. Custom endpoints. A connected key can carry its own base URL, which is how you point a provider slot at a private or self-hosted endpoint.

Your keys are stored encrypted in a secrets manager, decrypted only at call time and cached briefly. They are never sent to the client. Using your own keys means you keep your commercial terms, quotas and data agreements with each provider, and the provider bills you directly. The security page covers credential handling in more detail.

Routing and fallbacks

Provider outages, rate limits and exhausted balances are normal events. Connect treats them as routing inputs.

  • Fallback chains. When a primary model's provider is unavailable, rate-limited or out of funds, Connect walks an ordered chain of alternatives. Defaults are organized by pricing tier and cross providers, so one provider outage does not take a tier down.
  • Your own chains. A workspace routing rule must define its own fallback chain, with priority and time-of-day conditions, so you decide the order and the models.
  • Routing strategies. Strategies include cost-optimized, weighted, round-robin, least-used, priority and scoring-based selection, applied by name to a rule.
  • Consensus. For high-stakes calls, a request can fan out to several providers and a decision step selects the best answer.
  • Budget-aware downgrade. When a usage cap is reached, Connect can switch to a cheaper model rather than stop.

Fallbacks are a governance tool as well as a reliability tool. A chain restricted to approved models is how you make sure an outage never sends a request to a provider your policy has not cleared. That is the basis of EU-first routing.

Regional and EU routing

Connect is designed to let you route by where processing happens. In practice that means restricting a workspace to providers whose processing location you have confirmed, preferring European providers and European-hosted open-weight capacity, and running self-hosted models for the most sensitive traffic. Customer data in the managed service is stored in AWS eu-west-1 (Ireland) today. The details, and the limits of what is enforced versus designed for, are on the EU routing page.

Choosing a route

Which route fits which need
If you needRoute toKeys
Best quality on difficult reasoningA frontier model APIYour own or pooled
Low latency at volumeA fast-inference hostYour own or pooled
Data processed in EuropeA European provider or your own EU capacityYour own
Data never leaves your boundaryA self-hosted open-weight modelNone needed for the model
Predictable spendCost-optimized strategy with usage capsEither
ResilienceA fallback chain across two or more providersEither

Model availability, pricing and provider terms change. Check a provider's current terms before sending regulated data to it. Swfte provides the technical controls, governance mechanisms and evidence required to deploy AI within an organization's applicable regulatory, security and policy requirements. The exact posture depends on the customer's use case, jurisdiction, deployment and configuration.

Frequently asked questions

How many providers does Swfte Connect support?

Connect's registry maps model names to a broad set of commercial, regional and open-weight providers. The published count is <provider count - founder to fill>. Which providers a workspace can use depends on the keys it connects.

Can I bring my own API keys?

Yes. A workspace can connect its own provider keys, which Connect uses first. Without them, Connect uses Swfte's pooled access where it is offered. Keys are stored encrypted and never sent to the client.

Can I route to a model I host myself?

Yes. Open-weight models you host on your own or Swfte GPUs sit behind the same API, and a connected key can carry a custom base URL to point a provider slot at a private endpoint.

What happens when a provider is down?

Connect walks a fallback chain. Defaults are organized by pricing tier and cross providers, and you can define your own chain per workspace with priority and time-of-day conditions.

Is the API compatible with the OpenAI format?

Yes. Chat completions, streaming, embeddings, images, speech, transcription and moderation use an OpenAI-compatible request format, so moving an integration is mostly a base-URL and key change.

Does using a provider through Connect change that provider's data terms?

No. The provider processes the request under its own terms. Check each provider's terms, and the sub-processors page for the providers Swfte itself relies on.

Connect in the platform

Route every model through one governed gateway

Start managed in minutes, or plan a deployment inside your own boundary with the team.

Ready to build with Swfte?

One platform for the agents, models and workflows your team ships. Free to start, no card required.