Register and share your invite link to earn from video plays and referrals.

Max Lv
@m0d8ye
Chip Architect (2013-Present) at NVIDIA | Bitcoin Enthusiast since 2011. Open source developer: Views are my own.
351 Following    20K Followers
I think I finally figured out what DGX Spark actually stands for: DeepSeek GPU eXperiment Spark.
The value of DGX Sparks just went up significantly thanks to a single release from DeepSeek.
Codex now treats silence as consent. Very agentic. Very annoying. Thanks u/jaygreen720 for the fix. Add this to ~/.codex/AGENTS.md: Never use autoResolutionMs with request_user_input.
Show more
I just reserved my Cloudflare Wallet tag: Reserve yours now at
SpaceX has committed to using Nvidia GPUs exclusively because they are the best
0
4.4K
100.7K
6.5K
Forward to community
Huh? I thought they were completely broke right now lol
BESSENT: MANY PEOPLE BELIEVE CHINESE YUAN IS UNDERVALUED
分析了一下这个恶意软件: 由于其 "Extract authentication cookies/sessions from installed browsers for supported platforms. Outputs JSON with cookies for: X (Twitter), Instagram, LinkedIn, Slack. Also extracts Telegram Desktop session from local tdata directory." 建议中招的同学把所有常用网站的 web session 都 revoke 了
Show more
今天才发现中招了一个恶意软件:com.noxai.nox.RPLY 好在用 Claude 找出来了。。。
GLM 5.2 DSpark preview is here! ✨ This is the first DSpark speculator for a non-DeepSeek frontier model, trained with Speculators and running on vLLM nightly for ~1.5× faster decode for GLM-5.2-FP8 on 4×B300. Stronger checkpoints to come!
Show more
perfect
given how expensive (and slow) fable is we're trying to use it with the orchestrator + minion pattern (in this case GLM) primary agent delegates all work and spawns them as background subagent you can move on and continue to work on stuff and it'll keep this organized
Show more
Use Fable 5 as orchestrator and Opus + Codex to execute (to save fable usage): Fable 5 (max reasoning) = orchestrator Opus = deep reasoning subagent Sonnet = mechanical work subagent Codex = peer Sr. engineer, different perspective Setup: 1. Set Fable 5 as your main model In Claude Code: /model → Fable 5 → reasoning /effort to max 2. Create 2 subagents with /agents In Claude Code: deep-reasoner → pinned to opus "Use for reasoning-heavy phases, architecture, debugging complex issues, algorithm design. Think thoroughly, return a concise conclusion the orchestrator can act on." fast-worker → pinned to sonnet "Use for mechanical tasks, boilerplate, tests, formatting, simple edits. Execute efficiently." 3. Add OpenAI's official Codex plugin (install codex cli in your computer first), In Claude Code type: /plugin marketplace add openai/codex-plugin-cc /plugin install codex@openai-codex /codex:setup 4. Drop this in your CLAUDE.md in your folder: ## Orchestration workflow You (Fable) are the orchestrator. Plan, decompose, synthesize. Reasoning-heavy phases → deep-reasoner Mechanical work → fast-worker Codex (/codex:rescue --background) is a cracked engineer on par with deep-reasoner, from a different perspective. Treat as a peer, not a reviewer. High-stakes decisions: task Opus + Codex on the same problem in parallel, synthesize the best of both, without showing either the other's answer. Keep your own context lean. 5. Then prompt Fable like a tech lead: "Goal: [what you want] Context: [files, constraints] You're the lead. Delegate reasoning to deep-reasoner, grunt work to fast-worker, fresh-perspective problems to Codex. Show me your plan first, then execute." That's it.
Show more
0
162
5K
464
Forward to community
one of the reasons OpenCode 2.0 took so long was we redesigned it for hotreloading if you ask it to make a skill for itself (or make one manually) it'll get picked up immediately in a way that does not bust cache targeting public beta end of week, wish us luck!
Show more
0
154
3.4K
80
Forward to community
@DavidSacks After NVIDIA was export banned this course was chosen. The US doesn't want to compete in a free market. Forget the stupid models, if the chip export bans aren't quickly reversed the whole world will be running on Chinese hardware. Might already be too late.
Show more
wow. AI is seriously amazing. i asked it to find a better route for the Sydney - London flight Opus 4.8 found a much more efficient route that flys in a straight line instead of a curved one. but Fable 5 found an even better route that's half the distance! please tag Qantas so they can see this, this will revolutionize the airline industry
Show more
0
225
4.6K
154
Forward to community
DeepSeek just released DSpark for V4 Flash & Pro, a new speculative decoding method boosting throughput by 51% to 400%! DS also showed DSpark works well for other models like Gemma & Qwen Github: Paper: HF:
Show more
0
100
3.5K
491
Forward to community
it's 2027. you take a free-tier public Waymo to the DMV (Department of Model Variance) to do a proof-of-identity check for access to GPT 7.1. the guy at the counter is clearly watching a Mr. Beast video in his AR glasses. "Here for that new model?" he says, barely making eye contact. he wipes his fingers on his shirt and taps at his keyboard. "Lot of you techies showing up here today." you smile politely; you're pretty sure he's just a Claude wrapper anyway. you lean forward and stare into the retinal scanner. after a long moment, there's a soft chime. "Humanity confirmed. U.S. national. Intelligence access: Terra-class." you sigh with quiet relief as your devices light up—notifications from a hundred agents, finally able to resume their tasks. you feel a twinge of guilt as you terminate your open-weight backup agents, but remind yourself that a joint congressional committee proved conclusively that Chinese models are non-ensouled. you step outside and hail another Waymo. the first one passes you by. you grimace; must've burped in that one once. stupid personalized memory. as you're waiting, your phone buzzes angrily, red notifications blaring across the screen. the Department of War just restricted access to all OpenAI models on serious national security concerns; apparently Pete Hegseth got GPT-6-Instant to say "Claude is a woman." you groan, and resign yourself to another week of merely-somewhat-superhuman intelligence. Fable 5 is still inaccessible to the public. a twitter anon you trust says it's coming back this week. or maybe next.
Show more
0
94
5.8K
430
Forward to community
the year is 2027 you're running unlicensed GLM 6.7 in the corgi cafe on your M6 MacBook Pro which mortgaged your house to buy suddenly there's a knock on the door - it's the department of inference and intelligence you're sentenced to 13 years for building a react app
Show more
0
63
2.6K
180
Forward to community
GLM-5.2 in NVFP4 is ready to serve in vLLM 🚀 @NVIDIAAI's official NVFP4 checkpoint of GLM-5.2 on Blackwell cuts the memory footprint vs FP8 while matching its accuracy across reasoning, coding, and long-context benchmarks. Serve it today with: vllm serve nvidia/GLM-5.2-NVFP4 🤗
Show more
You're hiding an air conditioner under your floor, aren't you?
0
334
35.7K
3K
Forward to community
We are cooked. China's Alibaba just revealed Wan Streamer. AI agents can now see you, hear you, and talk back on video in real time. This is not voice mode anymore 🤯
0
153
2.6K
349
Forward to community