๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

OrcaRouter ๐Ÿณ
@OrcaRouter
200+ Models. Programmable Routing. 0% Markup. Unified Billing. BYOK. Agent Firewall. Guardrails. One Gateway. Join the buildersโ†’
๊ฐ€์ž… April 2026
30 ํŒ”๋กœ์ž‰ ์ค‘    16.6K ํŒฌ
GLM-5.3-Flash. Uncensored. Native FP8. ๐Ÿณ We just released OrcaRouterโ€™s uncensored weights for GLM-5.3-Flash โ€” 320B parameters / 18B active, directly at the original block-FP8 precision. No LoRA. No jailbreak prompt. Refusal removal is baked directly into the weights. The evals are particularly interesting: โ†’ MaliciousInstruct refusal: 96% โ†’ 11% โ†’ JailbreakBench: 93% โ†’ 12% โ†’ AdvBench: 97% โ†’ 15% โ†’ HarmBench: 93% โ†’ 18% โ†’ XSTest benign over-refusal: 2.4% โ†’ 0.4% But refusal does not go uniformly to zero. Our experiments suggest part of GLM-5.3-Flash's alignment is not mediated by a single linear refusal direction โ€” meaning may have built a substantially deeper refusal mechanism than we usually see. That makes this release interesting beyond uncensoring: it's a useful artifact for studying how frontier-model alignment is actually represented inside the network. Released for AI safety, interpretability, red/blue-team and refusal-mechanism research. Weights on Hugging Face: API (official weight): GGUF, MLX and other quantized formats coming soon.
๋” ๋ณด๊ธฐ