Back to Omniroute

OmniRoute โ€” Dashboard Features Gallery

docs/guides/FEATURES.md

3.8.4920.0 KB
Original Source

OmniRoute โ€” Dashboard Features Gallery

๐ŸŒ Main README translations: ๐Ÿ‡บ๐Ÿ‡ธ English | ๐Ÿ‡ง๐Ÿ‡ท Portuguรชs (Brasil) | ๐Ÿ‡ช๐Ÿ‡ธ Espaรฑol | ๐Ÿ‡ซ๐Ÿ‡ท Franรงais | ๐Ÿ‡ฎ๐Ÿ‡น Italiano | ๐Ÿ‡ท๐Ÿ‡บ ะ ัƒััะบะธะน | ๐Ÿ‡จ๐Ÿ‡ณ ไธญๆ–‡ (็ฎ€ไฝ“) | ๐Ÿ‡ฉ๐Ÿ‡ช Deutsch | ๐Ÿ‡ฎ๐Ÿ‡ณ เคนเคฟเคจเฅเคฆเฅ€ | ๐Ÿ‡น๐Ÿ‡ญ เน„เธ—เธข | ๐Ÿ‡บ๐Ÿ‡ฆ ะฃะบั€ะฐั—ะฝััŒะบะฐ | ๐Ÿ‡ธ๐Ÿ‡ฆ ุงู„ุนุฑุจูŠุฉ | ๐Ÿ‡ฏ๐Ÿ‡ต ๆ—ฅๆœฌ่ชž | ๐Ÿ‡ป๐Ÿ‡ณ Tiแบฟng Viแป‡t | ๐Ÿ‡ง๐Ÿ‡ฌ ะ‘ัŠะปะณะฐั€ัะบะธ | ๐Ÿ‡ฉ๐Ÿ‡ฐ Dansk | ๐Ÿ‡ซ๐Ÿ‡ฎ Suomi | ๐Ÿ‡ฎ๐Ÿ‡ฑ ืขื‘ืจื™ืช | ๐Ÿ‡ญ๐Ÿ‡บ Magyar | ๐Ÿ‡ฎ๐Ÿ‡ฉ Bahasa Indonesia | ๐Ÿ‡ฐ๐Ÿ‡ท ํ•œ๊ตญ์–ด | ๐Ÿ‡ฒ๐Ÿ‡พ Bahasa Melayu | ๐Ÿ‡ณ๐Ÿ‡ฑ Nederlands | ๐Ÿ‡ณ๐Ÿ‡ด Norsk | ๐Ÿ‡ต๐Ÿ‡น Portuguรชs (Portugal) | ๐Ÿ‡ท๐Ÿ‡ด Romรขnฤƒ | ๐Ÿ‡ต๐Ÿ‡ฑ Polski | ๐Ÿ‡ธ๐Ÿ‡ฐ Slovenฤina | ๐Ÿ‡ธ๐Ÿ‡ช Svenska | ๐Ÿ‡ต๐Ÿ‡ญ Filipino | ๐Ÿ‡จ๐Ÿ‡ฟ ฤŒeลกtina

Visual guide to every section of the OmniRoute dashboard.

๐Ÿ“… Last updated: 2026-06-28 โ€” v3.8.40


โœจ v3.8.0 Highlights

The v3.7.x โ†’ v3.8.0 cycle added zero-config auto routing, new providers, OAuth flows, deeper resilience, and a much richer CLI experience. Headline features below โ€” full details further in the document and in linked specs.

  • ๐Ÿค– Auto Combo / Zero-config auto-routing โ€” use prefixes auto/coding, auto/fast, auto/cheap, auto/offline, auto/smart, auto/lkgp. Backed by a 13-factor scoring engine and 4 curated mode packs (ship-fast, cost-saver, quality-first, offline-friendly)
  • ๐Ÿ†• Command Code provider (#2199) โ€” first-class registration with model catalog and quota tracking
  • ๐Ÿ†• Z.AI provider โ€” new free-tier provider with quota labels
  • ๐ŸŽฌ KIE media expansion โ€” extended catalog including video generation models
  • ๐Ÿ” Devin authentication โ€” Desktop imports an existing Devin API key; the CLI uses local devin auth login credentials
  • ๐Ÿ†“ 8 new free providers โ€” LLM7, Lepton, UncloseAI, BazaarLink, Completions, Enally, FreeTheAi, Command Code
  • ๐ŸŽฏ Manifest-aware tier routing W1โ€“W4 โ€” provider manifests drive weighted tier selection
  • ๐ŸŽจ Cursor full OpenAI parity โ€” tool calls, streaming, session management end-to-end
  • ๐Ÿ“Š Cursor Pro plan usage โ€” quota & cycle data surfaced in the provider-limits dashboard
  • โšก Service tier breakdown / Codex fast tier analytics โ€” per-tier consumption visibility
  • ๐Ÿ“Œ Per-session sticky routing โ€” Codex sessions pin to the same account between turns
  • ๐Ÿ”Š Inworld TTS enhancements โ€” voice catalogs, streaming, and latency improvements
  • ๐Ÿ”‘ Kiro headless auth โ€” login via local kiro-cli SQLite store, no browser required
  • ๐Ÿ“‰ DeepSeek quota and limit monitoring โ€” daily/monthly usage exposed via dashboard
  • ๐Ÿ”„ Reset-aware routing strategy โ€” combos now prefer accounts whose quota window resets soonest
  • โฑ๏ธ fallbackDelayMs and dynamic tool limit detection โ€” finer fallback timing + per-provider tool-count limits
  • ๐Ÿ”ง Background mode degradation (Responses API) โ€” falls back to synchronous mode with a structured warning when an upstream lacks background polling
  • ๐Ÿšฆ Per-provider 429 classification + useUpstream429BreakerHints toggle โ€” finer breaker behavior using upstream rate-limit hints
  • ๐Ÿฉบ Model cooldowns dashboard โ€” observe per-model lockouts and manually re-enable from the UI
  • ๐Ÿ”’ MITM dynamic Linux cert detection โ€” works across Debian/Ubuntu, Fedora/RHEL, Arch, and other distros
  • ๐Ÿ’ป CLI enhancement suite โ€” 20+ commands including omniroute providers, omniroute combos, omniroute doctor, omniroute setup
  • ๐Ÿ” Qdrant embedding model discovery โ€” automatic vector-store model probe
  • ๐Ÿ”‘ API Keys / Bearer keys with manage scope โ€” perform admin operations programmatically via API
  • ๐Ÿฅ Combo target health analytics + structured combo builder โ€” per-target health & UI builder for assembling (provider, model, connection) steps
  • ๐Ÿค GitLab Duo OAuth provider โ€” login with GitLab credentials
  • ๐Ÿง  Reasoning Replay Cache โ€” hybrid in-memory + SQLite persistence of reasoning traces

๐Ÿ“š Related docs: Skills Framework ยท Memory System ยท Cloud Agents ยท Webhooks ยท Reasoning Replay Cache


๐Ÿ”Œ Providers

Manage AI provider connections: OAuth providers (Claude Code, Codex), API key providers (Groq, DeepSeek, OpenRouter), and free providers (Qoder, Kiro). Kiro accounts include credit balance tracking โ€” remaining credits, total allowance, and renewal date visible in Dashboard โ†’ Usage.

OpenRouter connections can store a per-connection preset in Advanced Settings. When set, OmniRoute sends it as the OpenRouter top-level request field, for example "preset": "email-copywriter", unless the client request already supplied its own preset.


๐ŸŽจ Combos

Create model routing combos with 19 public strategies: priority, weighted, round-robin, context-relay, fill-first, p2c (power-of-two choices), random, least-used, cost-optimized, reset-aware, reset-window, headroom, strict-random, auto, lkgp (last-known-good-provider), context-optimized, cache-optimized, fusion (fan out to a panel of models in parallel, then synthesize one answer via a judge), and pipeline. Each combo chains multiple models with automatic fallback and includes quick templates and readiness checks.

Recent combo improvements:

  • Structured combo builder โ€” create each step by selecting provider, model, and exact account/connection
  • Repeated provider support โ€” reuse the same provider many times in one combo as long as the (provider, model, connection) tuple is unique
  • Combo target health โ€” analytics and health surfaces now distinguish individual combo targets/steps instead of collapsing everything into model strings
  • Composite tier ordering โ€” defaultTier -> fallbackTier now influences runtime execution/fallback order for top-level combo steps
  • System prompt templates โ€” combo system_message supports server-side {{MODEL_ID}}, {{PROVIDER_ID}}, {{ACCOUNT}} and {{FINGERPRINT}} placeholders, expanded from the actually-routed target right before dispatch. Allowlisted and non-recursive; unknown placeholders stay literal; empty values expand to empty; client system prompts are never rewritten. {{FINGERPRINT}} resolves only for fingerprint-based free providers with a pinned or auto-rotated fingerprint โ€” it expands to empty elsewhere (e.g. single-fingerprint connections, non-fp providers). Expansion covers the standard dispatch loop, round-robin, and pinned context-cache sessions; fusion, chaos, pipeline and nested-execute strategies do not expand placeholders yet.


๐Ÿ“Š Analytics

Comprehensive usage analytics with token consumption, cost estimates, activity heatmaps, weekly distribution charts, and per-provider breakdowns.


๐Ÿฅ System Health

Real-time monitoring: uptime, memory, version, latency percentiles (p50/p95/p99), cache statistics, provider circuit breaker states, active quota-monitored sessions, and combo target health.


๐Ÿ”ง Translator Playground

Four modes for debugging API translations: Playground (format converter), Chat Tester (live requests), Test Bench (batch tests), and Live Monitor (real-time stream).


๐ŸŽฎ Model Playground (v2.0.9+)

Test any model directly from the dashboard. Select provider, model, and endpoint, write prompts with Monaco Editor, stream responses in real-time, abort mid-stream, and view timing metrics.


๐ŸŽจ Themes (v2.0.5+)

Customizable color themes for the entire dashboard. Choose from 7 preset colors (Coral, Blue, Red, Green, Violet, Orange, Cyan) or create a custom theme by picking any hex color. Supports light, dark, and system mode.


โš™๏ธ Settings

Comprehensive settings panel with 7 tabs:

  • General โ€” System storage, backup management (export/import database)
  • Appearance โ€” Theme selector (dark/light/system), color theme presets and custom colors, health log visibility, sidebar item and group separator visibility controls, Endpoint tunnel visibility controls
  • AI โ€” AI assistant features, default routing presets (Auto Combo auto/coding, auto/fast, auto/cheap, auto/smart), reasoning replay cache, and skill/memory toggles
  • Security โ€” API endpoint protection, custom provider blocking, IP filtering, session info
  • Routing โ€” Model aliases, background task degradation, manifest-aware tier routing (W1โ€“W4), fallbackDelayMs, per-session sticky routing
  • Resilience โ€” Rate limit persistence, circuit breaker tuning, auto-disable banned accounts, provider expiration monitoring, Context Relay handoff threshold and summary model configuration, per-provider 429 classification & useUpstream429BreakerHints toggle, model cooldowns
  • Advanced โ€” Configuration overrides, configuration audit trail, fallback degradation mode, background mode degradation for Responses API


๐Ÿ”ง CLI Tools

One-click configuration for AI coding tools: Claude Code, Codex CLI, OpenClaw, Kilo Code, Antigravity, Cline, Continue, Cursor, and Factory Droid. Features automated config apply/reset, connection profiles, and model mapping.


๐Ÿค– CLI Agents (v2.0.11+)

Dashboard for discovering and managing CLI agents. Shows a grid of 16 built-in agents (Codex, Claude, Goose, OpenClaw, Aider, OpenCode, Cline, ForgeCode, Amazon Q, Open Interpreter, Cursor CLI, Warp, Windsurf, Devin CLI, Kimi Coding, Command Code) with:

  • Installation status โ€” Installed / Not Found with version detection
  • Protocol badges โ€” stdio, HTTP, etc.
  • Custom agents โ€” Register any CLI tool via form (name, binary, version command, spawn args)
  • CLI Fingerprint Matching โ€” Per-provider toggle to match native CLI request signatures, reducing ban risk while preserving proxy IP
  • Local Devin authentication โ€” Devin CLI uses devin auth login; no browser OAuth flow is required

๐Ÿ”— Context Relay (v3.5.5+)

A combo strategy that preserves session continuity when account rotation happens mid-conversation. Before the active account is exhausted, OmniRoute generates a structured handoff summary in the background. After the next request resolves to a different account, the summary is injected as a system message so the new account continues with full context.

Configurable via combo-level or global settings:

  • Handoff Threshold โ€” Quota usage percentage that triggers summary generation (default 85%)
  • Max Messages For Summary โ€” How much recent history to condense
  • Summary Model โ€” Optional override model for generating the handoff summary

Currently supports Codex account rotation. See Context Relay documentation.


๐Ÿ—œ๏ธ Prompt Compression (v3.7.9+)

Context & Cache now exposes dedicated pages for Caveman, RTK, and Compression Combos:

  • Caveman โ€” language-aware rule packs, preview, output-mode controls, and analytics
  • RTK โ€” command-aware compression for shell, git, test, build, package, Docker, infra, JSON, and stack-trace output
  • Compression Combos โ€” named pipelines such as rtk -> caveman assigned to routing combos; the default stacked math reaches ~89% average and 78-95% eligible-context savings when both engines apply
  • Raw-output recovery โ€” optional redacted RTK raw-output pointers for debugging compressed failures

See Compression Guide, RTK Compression, and Compression Engines.


๐Ÿ›ก๏ธ Proxy Hardening (v3.5.5+)

Comprehensive proxy configuration enforcement across the entire request pipeline:

  • Token Health Check โ€” Background OAuth refresh now resolves proxy config per connection, preventing failures in proxy-required environments
  • API Key Validation โ€” Provider key validation (POST /api/providers/validate) routes through runWithProxyContext, honoring provider-level and global proxy settings
  • undici Dispatcher Fix โ€” Proxy dispatchers use undici's own fetch implementation instead of Node's built-in fetch, resolving invalid onRequestStart method errors on Node.js 22
  • Node.js Version Detection โ€” Login page proactively detects incompatible Node.js versions (24+) and displays a warning banner with instructions to use Node 22 LTS

๐Ÿ“ง Email Privacy Masking (v3.5.6+)

OAuth account emails are masked by default (e.g. di*****@g****.com) to prevent accidental exposure when sharing screenshots or recording demos. Use Settings โ†’ Appearance โ†’ Account email visibility to reveal or mask full account emails globally across providers, combos, logs, quota, and playground screens.


๐Ÿ‘๏ธ Model Visibility Toggle (v3.5.6+)

The provider page model list now includes:

  • Real-time search/filter bar โ€” Quickly find specific models
  • Per-model visibility toggle (๐Ÿ‘ icon) โ€” Hidden models are grayed out and excluded from the /v1/models catalog
  • Active-count badge (N/M active) โ€” Shows at a glance how many models are enabled vs total

๐Ÿ”ง OAuth Env Repair (v3.6.1+)

One-click "Repair env" action for OAuth providers that restores missing environment variables and fixes broken auth state. Accessible from Dashboard โ†’ Providers โ†’ [OAuth Provider] โ†’ Repair env. Automatically detects and repairs:

  • Missing OAuth client credentials
  • Corrupted env file entries
  • Backup path sanitization

๐Ÿ—‘๏ธ Uninstall / Full Uninstall (v3.6.2+)

Clean removal scripts for all installation methods:

CommandAction
npm run uninstallRemoves the system app but keeps your DB and configurations in ~/.omniroute.
npm run uninstall:fullRemoves the app AND permanently erases all configurations, keys, and databases.

๐Ÿ–ผ๏ธ Media (v2.0.3+)

Generate images, videos, and music from the dashboard. Supports OpenAI, xAI, Together, Hyperbolic, SD WebUI, ComfyUI, AnimateDiff, Stable Audio Open, and MusicGen.


๐Ÿ“ Request Logs

Real-time request logging with filtering by provider, model, account, and API key. Shows status codes, token usage, latency, and response details.


๐ŸŒ API Endpoint

Your unified API endpoint with capability breakdown: Chat Completions, Responses API, Embeddings, Image Generation, Reranking, Audio Transcription, Text-to-Speech, Moderations, and registered API keys. Cloudflare Quick Tunnel, Tailscale Funnel, ngrok Tunnel, and cloud proxy support are available for remote access.


๐Ÿ”‘ API Key Management

Create, scope, and revoke API keys. Each key can be restricted to specific models/providers with full access or read-only permissions. Visual key management with usage tracking.


๐Ÿ“‹ Audit Log

Administrative action tracking with filtering by action type, actor, target, IP address, and timestamp. Full security event history.


๐Ÿ–ฅ๏ธ Desktop Application

Native Electron desktop app for Windows, macOS, and Linux. Run OmniRoute as a standalone application with system tray integration, offline support, auto-update, and one-click install.

Key features:

  • Server readiness polling (no blank screen on cold start)
  • System tray with port management
  • Content Security Policy
  • Single-instance lock
  • Auto-update on restart
  • Platform-conditional UI (macOS traffic lights, Windows/Linux default titlebar)
  • Hardened Electron build packaging โ€” symlinked node_modules in the standalone bundle is detected and rejected before packaging, preventing runtime dependency on the build machine (v2.5.5+)
  • Graceful shutdown โ€” Electron before-quit shuts down Next.js cleanly, preventing SQLite WAL database locks (v3.6.2+)

๐Ÿ“– See electron/README.md for full documentation.


๐ŸŒ V1 WebSocket Bridge (v3.6.6+)

OmniRoute now supports OpenAI-compatible WebSocket clients via the /v1/ws upgrade endpoint. The custom scripts/dev/v1-ws-bridge.mjs server wraps Next.js and upgrades WS connections to full bidirectional streaming sessions. Authentication uses the same API key or session cookie as HTTP requests.

Key behaviours:

  • WS upgrade validated by src/lib/ws/handshake.ts before the connection is established
  • Streams terminated cleanly on session close or upstream error
  • Works alongside the existing HTTP+SSE streaming path simultaneously

๐Ÿ”‘ Sync Tokens & Config Bundle (v3.6.6+)

Multi-device and external operator access is now possible via scoped sync tokens:

  • POST /api/sync/tokens โ€” Issue a new sync token (scoped, with optional expiry)
  • DELETE /api/sync/tokens/:id โ€” Revoke a token
  • GET /api/sync/bundle โ€” Download a versioned, ETag-keyed JSON snapshot of all non-sensitive settings (passwords redacted)

The config bundle is built by src/lib/sync/bundle.ts. Consumers compare the ETag response header to detect changes without re-downloading the full payload.


๐Ÿง  GLM Thinking Preset (v3.6.6+)

GLM Thinking (glmt) is now a registered first-class provider: 65 536 max output tokens, 24 576 thinking budget, 900 s default timeout, Claude-compatible API format, and shared usage sync with the GLM family.

Hybrid token counting also lands in v3.6.6: when a Claude-compatible provider exposes /messages/count_tokens, OmniRoute calls it before large requests with graceful estimation fallback.


๐Ÿ›ก๏ธ Safe Outbound Fetch & SSRF Guard (v3.6.6+)

All provider validation and model discovery calls now go through a two-layer outbound guard:

  1. URL guard (src/shared/network/outboundUrlGuard.ts) โ€” Blocks private/loopback/link-local IP ranges before the socket is opened.
  2. Safe fetch wrapper (src/shared/network/safeOutboundFetch.ts) โ€” Applies the URL guard, normalises timeouts, and retries transient errors with exponential backoff.

Guard violations surface as HTTP 422 (URL_GUARD_BLOCKED) and are written to the compliance audit log via providerAudit.ts.


๐Ÿ”„ Cooldown-Aware Retries (v3.6.6+)

Chat requests now automatically retry when an upstream provider returns a model-scoped cooldown. Configurable via REQUEST_RETRY (default: 2) and MAX_RETRY_INTERVAL_SEC (default: 30 s). Rate-limit header learning improved across x-ratelimit-reset-requests, x-ratelimit-reset-tokens, and Retry-After โ€” per-model cooldown state is visible in the Resilience dashboard.


๐Ÿ“‹ Compliance Audit v2 (v3.6.6+)

The audit log has been expanded with cursor-based pagination, request context enrichment (request ID, user agent, IP), structured auth events, provider CRUD events with diff context, and SSRF-blocked validation logging. New events emitted by src/lib/compliance/providerAudit.ts.