Register and share your invite link to earn from video plays and referrals.

Wanderer
@Wanderer_01_03
Waiting for DeepSeek v4.1.
21 Following    0 Followers
🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta! 🔷 We’ve massively upgraded its Agent capabilities—benchmark scores are now far surpassing the V4-Pro-Preview. Check out the massive performance leap below! 👇 🔷 The official V4-Flash now natively supports the Responses API format and is fully adapted for Codex! Check out the configuration details in our official API docs:
Show more
0
1.7K
30.1K
3.5K
Forward to community
You can try the following two prompts in the Thinking mode (via web/app) to get a better model experience in certain domains like counting (Note: keep a line break after the bracketed titles):: [Think with Grounding] ....... [Think with Pointing] ...... These two prompts encourage the model to adopt bounding boxes or points (which are classic fundamentals in computer vision) in its thought process. Personally, I love the pointing approach for solving abstract topological/reasoning tasks. Using points to represent continuous trajectories makes the MLLM's reasoning process feel much more human-like. Speaking purely from my personal exploration: Getting a multimodal model to accurately represent continuous trajectories with points is still a highly challenging frontier task for the entire industry. The current performance on real-world scenarios still has a long way to go.
Show more
Vision is now live on web and app. 👀 Come test the new eyes, but give its pure text capabilities a try while you're at it.
Just my opinion: The real reason PPO can handle long-horizon tasks? The Value Model for multi-step task. Honestly, it just shifts the training difficulty from GRPO’s GRM onto the Value Model. I mean, the problem isn’t solved — it’s just been relocated. The core question remains: How do you get stable process supervision in long-horizon tasks? That’s the real bottleneck.
Show more
🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length. 🔹 DeepSeek-V4-Pro: 1.6T total / 49B active params. Performance rivaling the world's top closed-source models. 🔹 DeepSeek-V4-Flash: 284B total / 13B active params. Your fast, efficient, and economical choice. Try it now at via Expert Mode / Instant Mode. API is updated & available today! 📄 Tech Report: 🤗 Open Weights: 1/n
Show more
0
1.7K
45.6K
7.5K
Forward to community