๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

NVIDIA AI
@NVIDIAAI
Teaching your AI new tricks.
๊ฐ€์ž… June 2016
897 ํŒ”๋กœ์ž‰ ์ค‘    337.3K ํŒฌ
This #CVPR2026# paper from our research team is trending #1# on @HuggingFace ๐Ÿค— Meet LocateAnything: a vision-language detection model that rethinks bounding box prediction. For AI agents and robots, โ€œseeingโ€ is only useful if a model can pinpoint where something is fast enough to act. Trained on 138M high-quality samples, LocateAnything decodes bounding boxes in parallel instead of one coordinate at a time, improving localization accuracy while dramatically increasing throughput for visual grounding and detection. Project page:
๋” ๋ณด๊ธฐ