LLM & Text Generation3 min reading time

Everyone is building LLM routers, we deprecated ours

Hacker News
Read full post
Manifest launched an LLM router in March to reduce costs by routing requests to models based on complexity but deprecated it after mixed results and challenges in accurately assessing task complexity from prompts alone. They found caching more effective for cost reduction and emphasized the importance of engineers choosing models deliberately.

More in LLM & Text Generation

DeepSeek V4.1 Flash now available on AI Gateway

Covered by 2 sources

Cohere Debuts Open-Weight 218B Mixture-of-Experts Machine Translation Model

Unite.AI