Cruxy
Talk to sales
Try Cruxy
Try Cruxy
ProductsCruxyCruxy CodeCruxy CoworkCruxy BenchCruxy Guard

Cruxy. Thinks for itself.

© 2026 Cruxy. Built by the team behind mycrux.

[email protected]
PrivacyTermsUsageStatus
Model › Mira

Mira. Built for speed.

A sub-400ms latency design target, high-volume throughput, and pricing that works at scale. Mira is the model for the moments where latency matters more than depth.

Join waitlistAPI reference →
64K context·2K max output·Text·400ms first-token target·$0.20/1M input · $0.40/1M output

Capabilities

Mira is built for scale.

Real-time chat

Mira is designed to a sub-400ms time-to-first-token target so it feels instant. Build chatbots, support agents, and conversational interfaces where any perceptible lag breaks the experience.

High-volume classification

Tagging, routing, moderation, sentiment, intent - Mira handles the small decisions that happen millions of times a day. At $0.20 per million input tokens, the economics hold up at high volume.

Autocomplete and suggestion

Fast enough to live inside an editor or search box, Mira powers the inline AI experiences where response time is the product.

Pricing

Priced to run at scale.

Input

$0.20per 1M tokens

For prompts at real-time-chat volume.

Output

$0.40per 1M tokens

For generated responses, at any volume.

Per-token pricing that holds up at high volume, metered in USD.

Choose well

When to pick Mira.

Pick Mira when speed and cost are the product - real-time chat, classification, autocomplete, anything that happens thousands of times per minute.

Pick Vaani when you need more reasoning quality, and the workload isn't latency-critical. Most interactive apps fit Vaani, not Mira.

Pick Kavi when the task is rare, hard, and accuracy-critical.

Built for volume.

When latency matters, Mira answers first.

Join waitlistAPI reference →