가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

xjdr
@_xjdr
building AI that wont embarrass me in front of my own standards
가입 December 2023
724 팔로잉 중    29.3K
spent yesterday on both grok 4.5 evals and using it in practice. it is a very good model for its price and intended use case. its smarter than glm5.2 (the model i would most immediately compare it to) in a lot of ways for ncode harness use and shows the beginning of real frontier RL / post training polish . it is very good at operating in ncode and makes excellent use of the tools at its disposal all while being extremely token efficient (i'd say its most distinguishing quality) . my 2 major complaints are no 1m ctx (which is basically standard now and i find myself missing often in real use) and its price point is just a _touch_ too high for where it fits (in my world at least) . ideally, i would use it as the subagent execution arm to replace glm5.2 or gpt5.5 med or to replace GLM 5.2 as the planner and delegate to dsv4-flash or gemma 4 31b . that said, i do plan to use it (unless gpt5.6 replaces it today) quite a bit from now on. if they can apply (or even improve) this post training polish to their next 2T base model (that is currently training as i understand it) , then they could have a very interesting next release
더 보기