Qwen updated the HF model card w/ some interesting details about 3.8:
- Native vision: images + hour-long videos (this one surprises me the most)
- support for popular agentic harnesses & coding tools
- thinking is toggleable per request
- supports reasoning_effort (xhigh default) - supports preserve_thinking
• 262K native ctx expandable to 1M ctx
Funny; they call it a casual model and a lot mentions about support for popular harnesses and development tools, making it easier to integrate into your existing stack.
This could be the best local model for Hermes and OpenClaw and such...
TBD....