WebInfer

Supported Providers

WebInfer connects to 25+ AI providers. Use any combination of cloud APIs and local models.

Works with your AI Providers

Ollama
OpenAI
Anthropic
Google AI
Meta
Mistral AI
Hugging Face
Perplexity
Ollama
OpenAI
Anthropic
Google AI
Meta
Mistral AI
Hugging Face
Perplexity

Plus Chrome's built-in Gemini Nano, and many more.

Major Cloud Providers

10

Leading AI model providers with flagship models

Fast & Cost-Effective

7

Optimized for speed and budget-friendly pricing

Specialized Providers

4

Focused on specific use cases like search, deployment, or custom models

Gateway & Aggregators

4

Access multiple providers through unified APIs

Local & Self-Hosted

10

Run models on your own hardware with full privacy

Audio & Speech

4

Text-to-speech, speech-to-text, and audio generation

Vision & Image

3

Image generation and visual understanding

Open standard

One Protocol

All these providers speak the same language. Switch between them freely, combine local and cloud, or run your own infrastructure.

Switch Freely

Move from OpenAI to Anthropic to local models. Same code, different provider.

Federation

Nodes connect to each other. Your home server can fall back to cloud when needed.

Self-Host

Run your own gateway. Issue tokens to your team or app users. Keep data on your infrastructure or relay other AI inference nodes.

Future-Proof

New providers plug in automatically. Your apps gain new capabilities without code changes.

Add Your AI Service

Any AI service or gateway can advertise its inference capabilities by adding a webinfer.json manifest to their server. This enables automatic discovery and configuration.

Learn How
  • Automatic discovery via webinfer.json
  • Standardized capability declaration
  • Self-hosted gateway support
  • OpenAI-compatible API endpoints
  • Custom model configurations