I enable thinking, the issue goes back to a regression bug that I had reported last year, it appears I need to hardcode the number of tokens used for thinking and it won’t accept 0 or -1 and won’t accept nothing even though it says it’s optional.
It only seems to affect the Flash Lite model but not the Flash model (which works just fine even without any thinking tokens specified and thinking enabled).
The Pro model is completely broken, I can’t get it work at all after the upgrade, I get an internal server error.