๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Trishool | SN23
@trishoolai
Bittensor Subnet for AI Alignment | Operated by @astrowareai
๊ฐ€์ž… November 2025
64 ํŒ”๋กœ์ž‰ ์ค‘    1.2K ํŒฌ
Weโ€™re excited to announce that Trishoolโ€™s HaloGuard 1.0 ๐ก๐š๐ฌ ๐š๐œ๐ก๐ข๐ž๐ฏ๐ž๐ ๐’๐Ž๐“๐€ prompt-safety performance among open-weight guard models. Today, we present HaloGuard 1.0, a constitutional input classifier for multilingual AI safety. It is built as a first-layer input guard that checks user prompts before they reach a downstream LLM, agent, or application. This is part of the safety infrastructure being built through @trishoolai , our decentralised AI red-teaming subnet on Bittensor SN23. Full arXiv paper goes live soon.
๋” ๋ณด๊ธฐ