Mo

Model-literals, model-aliases, and preference-aligned routing for LLMs

Hacker News

Model-literals, model-aliases, and preference-aligned routing for LLMs

Today we’re shipping a major update to ArchGW (an edge and service proxy for agents [1]): a unified router that supports three strategies for directing traffic to LLMs — from explicit model names, to semantic aliases, to dynamic preference-aligned routing. Here’s how each works on its own, and how they come together. Preference-aligned routing decouples task detection (e.g., code generation, image editing, Q&A) from LLM assignment. This approach captures the preferences developers establish when testing and evaluating LLMs on their domain-specific workflows and tasks. So, rather than relying on an automatic router trained to beat abstract benchmarks like MMLU or MT-Bench, developers can dynamically route requests to the most suitable model based on internal evaluations — and easily swap out the underlying moodel for specific actions and workflows. This is powered by our 1.5B Arch-Router LLM [2]. We also published our research on this recently[3] Modal-aliases provide semantic, version-controlled names for models. Instead of using provider-specific model names like gpt-4o-mini or claude-3-5-sonnet-20241022 in your client you can create meaningful aliases like "fast-model" or "arch.summarize.v1". This allows you to test new models, swap out the config safely without having to do code-wide search/replace every time you want to use a new model for a very specific workflow or task. Model-literals (nothing new) lets you specify exact provider/model combinations (e.g., openai/gpt-4o, anthropic/claude-3-5-sonnet-20241022), giving you full control and transparency over which model handles each request. P.S. we routinely get asked why we didn't build semantic/embedding models for routing use cases or use some form of clustering technique. Clustering/embedding routers miss context, negation, and short elliptical queries, etc. An autoregressive approach conditions on the full context, letting the model reason about the task and generate an explicit label that can be used to match to an agent, task or LLM. In practice, this generalizes better to unseen or low-frequency intents and stays robust as conversations drift, without brittle thresholds or post-hoc cluster tuning. [1] https://github.com/katanemo/archgw [2] https://huggingface.co/katanemo/Arch-Router-1.5B [2] https://arxiv.org/abs/2506.16655

Share card

Actual performance

2points
Did not reach leaderboard

Launch Intel predictions

Analyze your own launch →
Product HuntOn track for Day 1 leaderboard · Strong signals: agents, agent, claude · Missing: mac, macos, cursor
97%97% predicted probability of success on Product Hunt, based on ML models trained on real launch data.
best fitHighest predicted score across all platforms for this description.
Indie HackersFits the IH revenue-focused audience · Strong signals: supports · Missing: reddit linkedin, podcasting, created
90%90% predicted probability of success on Indie Hackers, based on ML models trained on real launch data.
Hacker NewsMay not resonate with HN audience · Strong signals: lua, ide, io · Missing: https docs, excited, just released
42%42% predicted probability of success on Hacker News, based on ML models trained on real launch data.
nativeThis product was originally launched on this platform.
TrustMRRLess likely to generate early MRR · Missing: mobile apps, ios, personal
40%40% predicted probability of success on TrustMRR, based on ML models trained on real launch data.
AppSumoMay struggle as an AppSumo deal · Missing: plus, platform, intuitive
38%38% predicted probability of success on AppSumo, based on ML models trained on real launch data.
Acquire.comPre-revenue stage for this audience · Missing: arr, mrr, revenue
13%13% predicted probability of success on Acquire.com, based on ML models trained on real launch data.
BetaListMay not resonate with beta-testers · Missing: web3, chat, crypto
0%0% predicted probability of success on BetaList, based on ML models trained on real launch data.

Correct prediction on native model

Similar products

Ro
RouteGPT – model routing on ChatGPT aligned to user preferences32%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

RouteGPT – model routing on ChatGPT aligned to user preferences

Hacker News2
Pu
PureRouter – Multi-model AI routing across LLMs (beta and $10 credits)24%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

PureRouter – Multi-model AI routing across LLMs (beta and $10 credits)

Hacker News4
A
A System Model of Western Civilisation51%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

A System Model of Western Civilisation

Hacker News2
Cy
Cygnus-X1 a Thrust Vectoring Model Rocket51%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Cygnus-X1 a Thrust Vectoring Model Rocket

Hacker News1
Ho
How to Model Infectious Diseases51%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

How to Model Infectious Diseases

Hacker News1
Fi
Finetune a Gemma 2B model for codegen57%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Finetune a Gemma 2B model for codegen

Hacker News5
I
I made an ensemble model to find underpriced properties33%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

I made an ensemble model to find underpriced properties

Hacker News2
ja
javscript model of Ackermann steering51%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

javscript model of Ackermann steering

Hacker News3
MEJA MAKAN MEWAH MINIMALIS
MEJA MAKAN MEWAH MINIMALIS52%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Aneka model mеја makan mewah kayu јаtі уаng paling laris

Indie Hackerscommitment-full-time
Reliance MET Industrial Plots
Reliance MET Industrial Plots51%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Reliance Model Economic Township is developing an industrial

Indie Hackerscommitment-side-project