Open pricing
No sales call required.
Kairos is priced by token volume. Calculate your cost in seconds. Every feature is included — no feature walls, no per-module upsells. Start free — forever — without talking to anyone.
How you can start
Three clear options — free forever, full trial, or paid by volume. All features included at every tier.
- Sherlock monitoring on 1 network
- 100K governed tokens / month
- Every feature included
- Unlimited Sherlock networks & scanners
- 250K tokens, full enterprise access
- Drops to the free plan — never cut off
- Effective rate drops as you commit more
- Unlimited Sherlock, annual saves 20%
- No sales call — subscribe self-serve
Token pricing calculator
Slide to your monthly token volume. Price appears instantly — all features included.
Volume pricing — starts at $2.00/1M and your effective rate drops the more you commit.
We'll pre-fill checkout with 50M tokens/mo. No credit card to start. Switch to annual (−20%) anytime.
— All platform features included — routing, compression, compliance, DLP, agents, war room, RL feedback.
— Volume pricing: the more tokens you commit, the lower your effective per-1M rate.
— Kairos compresses context by >=25%, reducing tokens sent to your LLM provider on every call.
— Annual prepay applies an automatic discount across the full year.
— You're saving 24% vs the $2.00/1M entry rate at this volume.
How Kairos pays for itself — at 50M tokens/mo
25% context compression on input tokens saves real money across every model you use. Based on a typical workload split: 40% chat, 35% coding/analysis, 25% deep thinking.
* Assumes 70% input / 30% output token split. Model prices as of 2026-Q1. Compression rate ≥25% on input tokens only; output tokens are provider-generated and not compressed. Actual savings vary by workload.
Everything included — at every tier
Kairos is an AI governance platform. Compliance enforcement, security, and observability are the core value proposition — not raw LLM cost arbitrage. Context compression (≥25% input token reduction) and intelligent model routing reduce your provider bill as built-in bonuses on top.
22 compliance frameworks
EU AI Act, NIST AI RMF, CMMC L2/3, FedRAMP, GDPR, OWASP LLM Top 10 — enforced per request, not just logged
Intent Firewall
Semantic jailbreak detection, PII extraction guard, prompt injection blocking — evaluated before the LLM call
DLP — 152+ PII types
Pre-LLM redaction with tamper-evident strip proofs for auditors. Custom pattern rules supported
SHA-256 audit trail
Cryptographically signed hash chain per request — exportable as evidence packages for regulators
Intelligent routing
Route by intent, cost, compliance, and latency across 100+ models. Smart fallback chains included
Context compression ≥25%
Reduces input tokens before every LLM call, lowering your provider bill automatically on every request
Human-in-the-Loop review
Route high-stakes or flagged requests to human reviewers before they reach the model
War Room + RL Feedback
Simulate 10K scenarios pre-production. Export human corrections as JSONL training data for fine-tuning
Not sure how many tokens you need?
Answer 3 quick questions and we'll estimate your monthly usage and apply it to the calculator above.
1/3 — How many users will make AI requests?
Questions? [email protected] — no sales call required.
Full pricing details →Everything included. No exceptions.
Feature access is never carved into upsell tiers. Your subscription is based on usage — not on which features you need.
AI Gateway & Routing
OpenAI-Compatible API
Drop-in replacement — change your base URL, keep your SDK. Supports streaming, function calling, and tool use.
Intelligent Model Routing
Keyword, regex, semantic intent, and cost-optimized rules route every request to the right model automatically.
Semantic Intent Routing
Intent-based routing understands the purpose of a request and selects the best model — beyond simple keyword matching.
100+ AI Models Supported
GPT-4o, Claude 3.5, Gemini, Llama, Mistral, DeepSeek, and every major provider — routed intelligently.
BYOK — Bring Your Own Keys
Use your own provider API keys. Kairos routes through them, enforcing your cost and compliance policy.
Custom Model Endpoints
Register custom or fine-tuned models alongside commercial providers and route between them on the same rules.
Context Compression
≥25% Context Reduction
Kairos compresses conversation context before sending to any LLM — reducing token spend without degrading quality.
Auto & Manual Modes
Auto mode adapts compression level based on model, query complexity, and governance context. Manual gives full control.
Extractive Compression
Scores segments by query relevance, structural weight, and entity density — keeping the highest-signal content.
Compression Proof
Every compression decision is explained — which segments were kept, dropped, and why — in a human-readable proof.
Per-Agent Compression
Compression is configured per agent and applied to every message — trimming token spend on each call.
AI/LLM Compliance & Governance
22 Compliance Frameworks
NIST AI RMF, EU AI Act, OWASP LLM Top 10, CMMC L2/3, FedRAMP, ISO/IEC 42001, UNESCO AI Ethics, and 15 more.
Real-Time Enforcement
Controls are enforced on every request — not just logged. Blocking, masking, or logging based on your policy configuration.
Per-Control Configuration
Enable or disable individual controls per framework, per API key. Fine-grained governance without all-or-nothing tradeoffs.
Tamper-Evident Audit Trail
SHA-256 hash chain on every request — cryptographically verifiable, immutable, and downloadable for auditors.
Human-in-the-Loop (HITL)
Route high-stakes requests to human review before they reach the LLM. Configurable per mandate or request type.
Compliance Readiness Score
Aggregate score showing how well your deployment aligns with enabled frameworks — updated in real time.
Board-Ready Exports
One-click CSV/JSON compliance reports pre-mapped to all active frameworks — ready for third-party auditors.
Auto-Updating Controls
All 22 mandates reviewed against authoritative sources every 90 days. Controls update when regulations change.
DLP & Data Protection
152+ PII Types Detected
SSN, MRN, DOB, credit cards, IBANs, passport numbers, driver's licenses, biometric identifiers, and more.
Pre-LLM Redaction
PII is stripped before context leaves your perimeter — the model never sees sensitive data.
DLP Strip Proof
Per-call proof of exactly what was redacted — with the original and redacted versions available for audit.
Custom DLP Rules
Define your own pattern-matching or regex rules on top of the built-in PII library.
Secure Document Analysis
Upload documents for AI analysis — PII is stripped, mandates highlighted, and the response returned on clean data.
Agents
Default Chat Agent
Connect your AI provider API keys and start chatting in seconds — smart routing and compression on from the first message.
Custom Agent Builder
Pick allowed models, routing rules, intent firewall, DLP, compliance frameworks, MCP servers, and HITL — in a clean UI or conversationally with Build As You Go.
Projects with Persistent Context
Workspaces with personalization instructions, accumulated context, and documents. Swap agents in and out — every chat carries the project’s context.
Model Routing + Compression Per Agent
Every agent routes each ask to the cheapest capable model across your linked providers and compresses context on every message.
MCP Server Integration
Model Context Protocol Gateway
Connect any MCP server to your agents with health monitoring, auth, and scoped access controls.
Zero-Trust Tool Allow-Lists
Agents access only explicitly authorized MCP tools — no implicit trust, full auditability of every tool call.
Custom MCP Functions
Register HTTP-backed custom functions as MCP tools — tested, versioned, and policy-gated in one place.
Health Monitoring
Real-time health status for every connected MCP server with configurable alerts.
Agent Identity & Security
Reusable Agents Across Chat, Projects & Jobs
Save an agent once and reuse it everywhere — in Chat, in Projects, and in Automated Jobs — with spend and token limits enforced via its Kairos API key.
Semantic Intent Firewall
Evaluates the purpose of every interaction to detect prompt injection, jailbreaks, and privilege escalation.
Hard Budget Enforcement
Set dollar caps per API key, agent, or team. Requests are blocked — not just alerted — when limits are hit.
Per-Agent Token Limits
Assign token budgets per agent to prevent runaway autonomous loops from generating unexpected costs.
Situational Risk Guardrails
Context-sensitive enforcement based on requester role, intent, and timing. Anomalous requests get flagged in real time.
Pre-Production & Simulation
War Room Simulation
Run up to 10,000 synthetic scenarios against your agent configuration before deploying to production.
Pass-Rate Analytics
See what percentage of scenarios passed compliance, routing, and cost targets — with per-scenario breakdown.
Cost Projection
Simulate real LLM costs against your pricing config before committing to production workloads.
Regression Testing
Run simulations against historical "golden" datasets to detect regressions when agent config changes.
Governance & Feedback
Governance-to-RL Feedback
Human corrections on flagged responses are structured as JSONL training data — ready for fine-tuning.
HITL Review Queue
Review non-compliant interactions, add corrected outputs, and export as structured RL training datasets.
Policy-to-Proof Audit
Every policy check is packaged with the specific context, evidence used, and controls passed — legal-grade record.
Decision Audit Trails
Immutable record of every routing, compression, governance, and DLP decision — per request, per call.
Observability & Cost Intelligence
Real-Time Cost Dashboard
Per-model cost breakdown, compression savings overlay, latency trends, and projected monthly spend.
Request Log with Compression Proof
Every request logged with full compression proof, compliance results, DLP events, and routing trace.
Trust & Geopolitical Overview
CISO/CIO executive dashboard: compliance posture score, model distribution, geolocation of requests.
CSV Export
Export usage, cost, and compliance data to CSV for finance teams and external auditors.
Cost Savings Attribution
Every dollar saved through compression and routing is attributed and displayed — broken down per day.
Sherlock — Shadow AI Discovery
Shadow-AI Network Scanner
A lightweight scanner agent in the Kairos CLI detects traffic to known AI/LLM endpoints — OpenAI, Anthropic, Gemini, Mistral, Perplexity, DeepSeek, xAI, and more.
Unlimited on Paid Plans
Sherlock is included with every paid subscription — continuous monitoring across unlimited networks and scanners.
Unlimited During Trial
Sherlock is fully unlimited for your 14-day trial — monitor as many networks and scanners as you like while you evaluate.
Free Forever on 1 Network
After your trial, continuous Sherlock monitoring keeps running on 1 network with up to 2 scanners — free, forever.
1 Free Scan — No Signup
Enter your email and run one free scan without creating an account. See what's on your network in minutes.
Shadow-AI Cost Estimation
Byte volume is converted to estimated tokens and priced with a representative model per provider — unknown AI usage shows up in Usage & Cost.
Per-Network & Per-Device Attribution
Usage is attributed to the specific WiFi/LAN and device IP, with geolocation. Ignore or delete any network's costs.
Metadata-Only & Privacy-Respecting
Sherlock classifies by destination endpoint and byte volume. It never decrypts traffic or reads message contents.
Enterprise & Access Control
Command Center RBAC
Hierarchical workspaces with scoped privileges, sub-organization isolation, and per-unit usage limits.
SSO / SAML
Enterprise identity provider integration for seamless team onboarding and centralized access management.
API Key Management
Create scoped API keys with granular permissions, budget limits, and per-key audit log panels.
Team Seats & Roles
Owner, admin, member, and viewer roles with feature-level access control per command center.
Private VPC / On-Premises
Deploy Kairos inside your own private VPC or on-premise environment, isolated to your infrastructure.
Custom SLA & MSA
Enterprise agreements with configurable SLAs, BAAs for HIPAA environments, and custom procurement.
Common questions
Does Kairos replace my AI provider?
No — Kairos integrates with your existing provider relationships. You bring your own API keys. Kairos routes intelligently between them, compresses context to reduce your bill, and enforces compliance on every call. Your providers stay the same; Kairos makes them work smarter.
What is token-metered pricing?
You pay based on the volume of tokens that flow through Kairos each month. There are no feature gates — every capability is included regardless of token volume. The price per million tokens decreases at scale.
Is there a free trial?
Yes. Every new account starts with a full-access Enterprise trial — 14 days or 250,000 tokens, whichever comes first. No credit card required. Sherlock is unlimited for the whole trial.
What happens when my trial ends?
You drop to the free plan automatically — you're never cut off. The free plan is free forever and includes 100K governed tokens per month, continuous Sherlock shadow-AI monitoring on 1 network with up to 2 scanners, and every platform feature. Upgrade whenever you need more volume.
What's the minimum paid commitment?
Paid plans start at a 5M-token monthly commitment with a $19/mo platform minimum. Rates are graduated — $2.00 per 1M tokens for the first 10M, falling to $0.50 per 1M at 500M–1.5B — so your effective rate drops as you commit more.
Can I pay annually?
Yes. Annual prepayment receives an automatic 20% discount. If your usage exceeds your prepaid volume, you can purchase additional tokens at your per-token rate directly in the platform.
What if my usage is unpredictable month to month?
Kairos tracks your usage and can recommend a renewal contract once a stable pattern is established. Until then, monthly billing keeps you flexible.
Do I need to talk to a salesperson?
No. You can calculate your price, start a trial, and subscribe entirely self-serve. Enterprise teams that need security review, custom procurement, or a guided demo can email [email protected] — but it's never required.
Need enterprise procurement support?
Self-serve is always the default. If your team needs security review, legal/MSA, a custom BAA, or a guided walkthrough, reach out to [email protected]. No SDR call required.