註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Gregor Zunic
@gregpr07
founder @browser_use
加入 August 2012
594 正在關注    31K 粉絲
One night autoresearch didn't generalize yet. It tried so much crazy stuff optimizing vLLM though. > optimized caching specifically for custom harness > implemented diffusion transformer on a 2048 block which made inference 17x faster
顯示更多
Jev really inspired me to build SUPER fast browser agents without sacrificing accuracy. I gave Codex access to vLLM on 2×B300 and let it change everything from the harness to inference. The constrained optimization: > min end-to-end task time > s.t. score ≥ baseline Caching, thinking, action batching, inference. Any part of the pipeline is fair game. First results below on 12 local form tasks. The goal: 10× faster on long tasks.
顯示更多