๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Victor M
@victormustar
๐Ÿค— Head of Product @huggingface
๊ฐ€์ž… August 2012
2.3K ํŒ”๋กœ์ž‰ ์ค‘    30K ํŒฌ
here we go again: deployed a FREE public endpoint for Qwen3.8-Flash-Next ๐Ÿš€ (going at +100 tok/s) No token needed, OpenAI-compatible, vision + tool calls, 262K context, thinking from xhigh โ†’ off. Light rate limiting, be nice to your neighbors ๐Ÿค— 4ร— H200 ยท FP8 ยท SGLang cookbook ยท ~140 tok/s per stream ยท ~100 tok/s @ 16 concurrent ยท 0.8s TTFT Guide + chat UI ๐Ÿ‘‡
๋” ๋ณด๊ธฐ