Helping ecosystem to grow on vLLM, ex-Head of APAC ecosystem @huggingface. Ex-Googler on TFLite/micro. Ideas on my own. Interested in future tech. DM open
After using recent model, I'm having this feeling strongly, most likely someone has already done a research on this, that there is a thinking scaling law.
The longer model thinks, the more intelligence you can get, especially on hard problems.
This has been fueling reasoning models like R1 but would love to see a better explanation on what's happening inside with clear observability + interpretability trace that others can re-produce.