Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Theyre training the system to minimize compute,so most likely theyre dynamically downgrading quants in the first few turns hoping to find the cheapest model to run. The side effect may be excessive token gen


Sure. That’s possible. But that’s not what I was talking about.


Tomato, technical tomato




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: