LLM & Text Generation3 min reading time

Everyone is building LLM routers, we deprecated ours

Hacker News
Read full post
Manifest launched an LLM router in March to reduce costs by routing requests to models based on complexity but deprecated it after mixed results and challenges in accurately assessing task complexity from prompts alone. They found caching more effective for cost reduction and emphasized the importance of engineers choosing models deliberately.

More in LLM & Text Generation

Peter Thiel-Backed AI Startup Cognition Raises Funds at $48 Billion Valuation

Covered by 2 sources

Build more natural voice experiences with GPT‑Live‑1 in the API

Covered by 2 sources