| ▲ | sangwook an hour ago | |
What online signal recalibrates simulated rankings against actual task success? Also do you have a plan to support semantic caching at the router level? | ||
| ▲ | kfallah15 an hour ago | parent [-] | |
For the online signal, we use a LLM judge with a rubric calibrated offline by the user via TUI. UX of the calibration is a major focus area. Semantic caching is interesting, open to supporting it but not currently planned. | ||