Local hosting · Proxy · Analytics

OneAI

Your models.
One gateway.

Host AI models on your hardware and proxy requests to external providers through one unified gateway. Connect your applications once, choose where requests run, and see what happens.

Routing policy controls today. A dedicated policy page is planned. Explore policy controls →

OneAI’s folded coral and glass gateway

One connection. Local or external.

Your AI infrastructure,
working together.

OneAI sits between your applications and the models they use, bringing local hosting and external-provider access into a single operating view.

OpenAI- and Anthropic-compatible APIs let applications, IDEs, and agents connect through a consistent gateway. Published names, aliases, and routing determine which configured deployment handles each request.

Your applications, IDEs & agents
OneAIModel names · routing · policy controls
Local CPU / GPUExternal providers
Local hosting · External proxying · Analytics

01 / Gateway & local hosting

Make your hardware
an AI serving layer.

Load compatible GGUF models on your CPU or GPU, publish model names for clients, and manage routing alongside your external deployments.

OneAI / Gateway
OneAI Gateway overview showing provider status, routing policy summary, and locally hosted GGUF models with device selection
Actual OneAI interface with local development data. View full size ↗

02 / External-provider proxy

Connect providers.
Keep one front door.

Bring external model providers behind OneAI. Manage provider connections and health, then route client requests through the same gateway used for local models.

CONNECT

Your choice of provider

Configure connections to OpenAI, Anthropic, Azure OpenAI, Amazon Bedrock, Google Vertex AI, and OpenAI-compatible endpoints.

OBSERVE

Know what’s available

Review provider health and connection status alongside your local model infrastructure.

CONTROL

Keep spending visible

Track provider spend and configure monthly provider budgets to control external usage.

Requests routed to an external provider are processed by that provider. Local hosting and external routing are explicit configuration choices.

03 / Analytics & activity

See where requests go.
Understand how they run.

Follow request activity across local and external models. Inspect routing decisions and operational signals to understand usage and investigate failures.

OneAI / Activity & analytics
OneAI Activity page with operational analytics for model requests
Actual OneAI interface with local development data. View full size ↗

Policy-aware routing

Define where
requests may run.

Bring hosting, proxying, analytics, and routing policy controls together. Keep requests local, restrict eligible providers, and apply configured provider spending limits.

The models behind the gateway

Built for more than text.

Serve the model capabilities your applications need, from embedded inference to multimodal workflows.

01

Embedded local inference

Run compatible models directly on your own CPU or GPU, with a native engine at the heart of OneAI.

02

GGUF models

Load and serve compatible GGUF models, publish clear model names, and choose where they run.

03

Vision & images

Work with image inputs using vision models and generate images with supported image backends.

04

Voice

Turn speech into text and text into speech with compatible transcription and synthesis models.

05

Tool calling

Connect tool-capable models to application workflows through structured tool calls.

06

Embeddings & retrieval

Build search and retrieval workflows with embeddings, classification, and reranking models.

Capabilities depend on the model, backend, and available hardware. Configure compatible models for vision, image generation, voice, and tool calling.

Get OneAI

Your platform. Your next step.

Public binaries are coming soon. This is where you’ll find the installers, build details, and release updates.

Windows

Coming soon

The installer will be available here when the release is ready.

macOS

Coming soon

The installer will be available here when the release is ready.

Linux

Coming soon

The installer will be available here when the release is ready.

Want to explore OneAI now?

Let’s walk through local hosting, provider proxying, and your integration needs.

OneAI is proprietary software. Third-party components and model weights retain their own licenses. Asset notices.

A conversation starts here

Have something in mind?

An application to explore, a song that stayed with you, or an idea we could build together. I’d love to hear it.