ๆณจๅ†Œๅนถๅˆ†ไบซ้‚€่ฏท้“พๆŽฅ๏ผŒๅฏ่Žทๅพ—่ง†้ข‘ๆ’ญๆ”พไธŽ้‚€่ฏทๅฅ–ๅŠฑใ€‚

Trishool | SN23
@trishoolai
Bittensor Subnet for AI Alignment | Operated by @astrowareai
ๅŠ ๅ…ฅ November 2025
64 ๆญฃๅœจๅ…ณๆณจ    1.2K ็ฒ‰ไธ
Weโ€™re excited to announce that Trishoolโ€™s HaloGuard 1.0 ๐ก๐š๐ฌ ๐š๐œ๐ก๐ข๐ž๐ฏ๐ž๐ ๐’๐Ž๐“๐€ prompt-safety performance among open-weight guard models. Today, we present HaloGuard 1.0, a constitutional input classifier for multilingual AI safety. It is built as a first-layer input guard that checks user prompts before they reach a downstream LLM, agent, or application. This is part of the safety infrastructure being built through @trishoolai , our decentralised AI red-teaming subnet on Bittensor SN23. Full arXiv paper goes live soon.
ๆ˜พ็คบๆ›ดๅคš