Everyone is building LLM routers, we deprecated ours
manifest.build
Everyone is building LLM routers, we deprecated ours
1–10 of 94 posts
Re: Everyone is building LLM routers, we deprecated ours
#2> "A cache-aware model router will take that into account by adding stickiness to the initially chosen model and keeps querying it."
Re: Everyone is building LLM routers, we deprecated ours
#3Re: Everyone is building LLM routers, we deprecated ours
#4Ironically, my confidence that a human had at least an active part in writing/editing this article went up because of this train wreck of a sentence: > "A cache-aware model router will take that into account by adding stickiness to the initially chosen model and keeps querying it."
Re: Everyone is building LLM routers, we deprecated ours
#5Ironically, my confidence that a human had at least an active part in writing/editing this article went up because of this train wreck of a sentence: > "A cache-aware model router will take that into account by adding stickiness to the initially chosen model and keeps querying it."
what’s wrong with the sentence? reads fine to me
Re: Everyone is building LLM routers, we deprecated ours
#6Ironically, my confidence that a human had at least an active part in writing/editing this article went up because of this train wreck of a sentence: > "A cache-aware model router will take that into account by adding stickiness to the initially chosen model and keeps querying it."
what’s wrong with the sentence? reads fine to me
Re: Everyone is building LLM routers, we deprecated ours
#7Re: Everyone is building LLM routers, we deprecated ours
#8Re: Everyone is building LLM routers, we deprecated ours
#9If you really need more discrimination of the complexity of an input to get an efficient response, sft or rl tuning something for your harness would be more effective.