Thursday, August 20, 2026

Ramp Launches Router.com to Help Businesses Reduce AI Inference Costs Through Automated Model Routing

Share

Corporate spend management platform Ramp has announced the launch of Router.com, a unified API gateway engineered to optimize artificial intelligence inference costs across enterprise operations. By acting as a single connection point to major commercial and open-weight artificial intelligence models, the system dynamically routes individual API prompts to the lowest-cost model capable of satisfying specific developer performance criteria.

The deployment integrates model routing directly into Ramp’s overarching financial visibility and spend control framework. According to early operational metrics, organizations utilizing the platform have reduced inference expenditures by an average of 40% while preserving output quality.

Curbing Unprecedented AI Costs with Algorithmic Routing

The launch comes amid a sharp rise in artificial intelligence investments across corporate sectors. Data from the Ramp AI Index indicates that business expenditure on AI models has expanded more than 20-fold since mid-2025. However, many software teams default to using high-tier foundational models for straightforward tasks, leading to unnecessary operational overhead.

Also Read: Mercury Expands Corporate Treasury Ecosystem with Custom Funds from Morgan Stanley and State Street

Router.com addresses these budget overruns by automatically matching workloads to appropriate models based on cost, latency, and reasoning requirements. The gateway includes more than 100 continuous optimizations covering prompt caching, context compression, latency timing, and automated provider fallbacks to maintain multi-region system reliability.

Key platform features include:

  • Unified Model Access: Connects via a single OpenAI- and Anthropic-compatible API to 27 leading models, including frontier architectures from OpenAI, Anthropic, and SpaceXAI, alongside open-weight options like DeepSeek, Nvidia, and Qwen.

  • Real-World Benchmarking: Evaluates model efficiency using Ramp SWE-Bench, an internal benchmarking framework trained on actual production engineering tasks rather than static public datasets.

  • Enterprise Spend Integration: Links token consumption directly to corporate finance controls, mapping costs to specific engineering teams, applications, and business units.

  • Sovereign Data Protection: Operates across U.S.-based cloud infrastructure, providing zero-data-retention parameters to secure proprietary corporate information.

AI is the fastest-growing line item at most companies, and the one they can least measure,” said Rahul Sengottuvelu, Chief Technology Officer at Ramp. “Router puts every token in one place and sends each request to the model that delivers the right performance at the right cost.”

Scalable Infrastructure for High-Volume Workloads

Originally built to manage Ramp’s internal engineering workflows, the underlying routing engine lowered the company’s internal model costs by approximately 30% over three years while processing trillions of monthly tokens.

Router.com is currently available globally, offering free request routing through 2026, allowing development teams to evaluate performance, automate model fallbacks, and optimize token utilization without upfront software fee overhead.

Read more

Local News