AI Engineer | Multi-agent LLM orchestration & RAG | Production AWS backend, multi-tenant authorization
- Taiwan
Pinned Loading
-
cost-aware-hybrid-router
cost-aware-hybrid-router PublicCost-aware hybrid router for multi-agent LLM systems: keyword → embedding → LLM cascade. Matches LLM-only accuracy (82.6% vs 82.9%, McNemar p>0.3) with 74% fewer LLM calls on CLINC150.
Python 2
-
tau-bench-decomposition
tau-bench-decomposition PublicCheap Decomposition, Expensive Execution: cost-aware task decomposition for LLM agents on τ-bench
Python 1
-
tinyrouter
tinyrouter PublicFine-tuning a small encoder for cost-aware LLM routing on CLINC150: sample efficiency, calibration, and LLM fallback.
Python
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.



