Msty Nexus
Open web app

Msty Nexus - Inference gateway

Own the inference layer behind your AI tools.

Put one governed inference layer behind your AI tools for model routing, local runtimes, provider keys, usage visibility, and guardrails.

Open web app

Product info

A control plane for the AI stack you already use.

Msty Nexus sits between your applications and model endpoints. It standardizes the plumbing for local runtimes and online providers, protects credentials, applies routing and guardrails, and gives visibility into inference traffic.

Abstract access gateway approving business AI app and model connections.

Control AI access

Reduce shadow IT by giving your business one governed place to approve apps, providers, keys, and model access.

Abstract local AI server protected by a privacy boundary with optional cloud access.

Leverage local AI

Keep sensitive data private and secure on local infrastructure, while still using online models when the job needs them.

Abstract layered guardrails filtering requests before they reach model infrastructure.

Add guardrails

Apply extra protections like scoped tokens, PII redaction, key blocking, and routing rules before requests reach models.

How it works

One gateway between apps, models, and credentials.

Nexus gives teams and power users a single place to connect AI tools without scattering provider keys and runtime settings across every application.

01

Connect runtimes and providers

Bring local runtimes and hosted providers into one Nexus-managed catalog.

02

Set presets, aliases, and policies

Standardize model settings so applications call stable names instead of fragile provider details.

03

Issue app-scoped tokens

Give each approved tool its own local token and revoke access without rotating every key.

04

Route and monitor requests

Use compatible endpoints while Nexus tracks usage and keeps provider credentials protected.

Msty Nexus machine dashboard showing providers, runtimes, models, system metrics, and local inference status

Power features

Route, balance, and scale model work.

Nexus is more than a gateway. It can choose the right model path, distribute traffic, and coordinate inference capacity so every request does not land on the most expensive model.

Abstract gateway routing one request to multiple model destinations.

Dynamic model routing

Send each request to the right model for the job. Keep expensive models for work that needs them, and route routine tasks to cheaper or local models to reduce token costs.

Abstract gateway distributing request streams across multiple runtime nodes.

Load balancing

Distribute requests across available runtimes and machines so teams can keep throughput steady without pointing every app at one overloaded endpoint.

Abstract cluster of machines connected to a shared inference gateway.

Clustered inference

Coordinate multiple machines as shared inference capacity for larger local workloads, central servers, or teams that need more horsepower behind the gateway.

Features

Built for model access that needs control.

Use Nexus as the operating layer for local runtimes, hosted providers, governed app access, and practical usage visibility.

01

Model gateway

Give applications one governed catalog across local runtimes and online providers.

02

Credential control

Keep provider credentials in one controlled place instead of copying them across tools.

03

Compatible endpoints

Connect approved applications through familiar OpenAI- and Anthropic-compatible endpoints.

04

Runtime control

Manage supported Ollama, llama.cpp, and MLX runtimes alongside hosted model access.

05

App tokens

Issue scoped tokens per application instead of sharing one master key across every integration.

06

Usage visibility

See requests, latency, and usage by application and provider from one place.

07

Connection approvals

Review and approve new model or provider connections before applications can use them.

08

Network boundary

Keep the gateway off the LAN by default and open access only when you choose to.

Pricing

Free to start. Pro and Enterprise when you need shared control.

Nexus is available as a free download. Pro and Enterprise add managed scale, protection controls, logging, and team governance.

Local gateway

Free

$0Free download

Download Nexus and run the local gateway on your machine.

  • Free desktop download
  • Unlimited gateway requests
  • 1 managed machine
  • Local gateway for approved apps
  • Provider model catalog
  • Basic local usage visibility
Download Nexus
Enterprise edition

Enterprise

Starts at $499per month
Starts at $4,999/yearSave $989+ with annual billing

Enterprise controls for larger teams, stricter environments, and managed deployments.

  • Unlimited gateway requests
  • Starting at 125 managed machines
  • Teams/RBAC, rate limits, SSO, and priority support
  • Cluster, air-gapped support, and team/user logging
Contact sales
Feature / AttributeFreeProEnterprise
Usage & Billing
Gateway RequestsUnlimitedUnlimitedUnlimited
Managed machines13Starting at 125 included
Administration
Desktop Console
Cloud console-
Teams / RBAC--
Team / User rate limit--
Single Sign-On (SSO)--
Routing & Safety
Smart RoutesLimited
Load Balancer-
Cluster--
Fallbacks-
Logging & Support
Basic logging
Team/User logging--
Priority Support--

Download

Install the local Nexus gateway.

Download the desktop gateway for the machine that runs Nexus and routes approved apps, models, credentials, and local runtimes.