Nicheloom

Market intelligence for builders — see what's gaining traction before it's crowded.

Rayline routes Claude Code subagents to on-device and cheaper models

Details

External ID
48448372
Source
HN
Company
—
Product
Rayline routes Claude Code subagents to on-device and cheaper models
Website domain
rayline.ai
Launched
June 8, 2026
Cohort
—
Upvotes
11
Upvotes percentile
0.6284153005464481
Tags
—
Fetched at
Sept. 7, 2026, 9:26 p.m.
Updated at
Sept. 7, 2026, 9:26 p.m.

Description

Hi HN,I’m one of the builders of Rayline.Rayline is a Claude Code compatible LLM gateway. It intercepts and overrides claude code’s internal routing and lets you route subagent calls to different models instead. For example, you can run the main agent on Opus, some subagents on cloud-hosted open models, and other subagents on-device.We’ve seen others implement routing for claude code as tools the agent can invoke. In our experience, that doesn’t work well because it requires the main agent to use tokens to think about + call the tools, and LLMs are generally a very inefficient way to make routing decisions. By implementing Rayline as a gateway, we let users deterministically configure routing decisions, and you can optionally use our ML model to make routing decisions.We built it after noticing that Claude Code sessions contain a lot of subagent calls that don’t all need the same model. Other routers exist, but we built Rayline to let us continue using claude code (no separate harness), route tasks at a subagent level, and route across cloud and on-device. The main agent often benefits from Opus. But many delegated calls have narrow scope: search the repo, summarize context, inspect an error, poll for CI updates, etc.The thing we’re exploring is subagent-level routing. The main cost lever in coding agents is usually cached vs non-cached input. Subagent delegations are a natural point to make routing decisions because you avoid busting cache. We look at the message-thread context for a delegated call and choose a model for that call. At a task level, Sonnet and Haiku are almost always less capability-per-dollar than open models, so the main advantage is better + (much) cheaper subagents (60-90% in our private beta).The whole world seems to have started talking about model routing in the past two weeks, so apparently others agree it’s a relevant product area.We’d love to get feedback from the HN community!

Enrichment

Theme
Claude integrations and coding agents
Vertical
Horizontal
Function
Agent / copilot
Audience
Developer
AI stance
AI-native
Project type
Commercial product
Normalized one-liner
ai code agent router for cost optimization
Manually corrected
False

Could you build this?

Yes Rayline is an HTTP/proxy gateway that intercepts Anthropic API-compatible requests and proxies/routes them to alternative endpoints (local Ollama or cheaper cloud models) based on payload inspects.

Discussion

9 comments analyzed.

Competitors mentioned: RouteLLM, Not Diamond, OpenRouter, Google Gemma, Mistral

Concerns raised: Deciding when a task is actually safe to run on-device, Chinese-hosted/trained models may not be acceptable for some companies, How this differs from existing routing solutions like OpenRouter, Unclear capability-per-dollar comparison with Claude Sonnet/Haiku

Feature requests: Geographic restrictions for inference location, Support for non-Chinese model options

Competitors

Other products that read as similar to this one — 176 launches clear the similarity bar, closest 8 shown.

Attention rank: #69 of 177 (itself plus its competitors, highest first — normalized so YC and Product Hunt are compared fairly).

Launched 222 days after the earliest competitor.

Other launches for this product

Same idea, different domain

Nobody's really built a agent / copilot tool for Agriculture yet.