Balancing cache writes-vs-utilization, tool calling back-and-forth, reasoning effort, to improve subjective & perceived sense of performance of an LLM-backed agentic workflow, while having a classifier to route each inference call to a complex hierarchy & classes of LLMs, providers, and wrappers which also reroutes during outages, degraded outputs, and latencies, while ensuring a consistent personality to the end consuming entity is the funnest engineering challenge today.
Tag: agents
1 jot
Subscribe to this tag
RSS 2.0
https://jottings.vishalvshekkar.com/tags/agents/feed.xml
Atom 1.0
https://jottings.vishalvshekkar.com/tags/agents/atom.xml
JSON Feed
https://jottings.vishalvshekkar.com/tags/agents/feed.json