注册并分享邀请链接,可获得视频播放与邀请奖励。

NVIDIA AI
@NVIDIAAI
Teaching your AI new tricks.
加入 June 2016
897 正在关注    337.3K 粉丝
This #CVPR2026# paper from our research team is trending #1# on @HuggingFace 🤗 Meet LocateAnything: a vision-language detection model that rethinks bounding box prediction. For AI agents and robots, “seeing” is only useful if a model can pinpoint where something is fast enough to act. Trained on 138M high-quality samples, LocateAnything decodes bounding boxes in parallel instead of one coordinate at a time, improving localization accuracy while dramatically increasing throughput for visual grounding and detection. Project page:
显示更多
0
58
2.2K
326
转发到社区