Register and share your invite link to earn from video plays and referrals.

Xuan-Son Nguyen
@ngxson
Engineer @huggingface
251 Following    6.7K Followers
Lmao, I can now justify my RTX card for gaming
Welcome to your new simulation Microduck.
You didn't mention the 10 test files added without user's consent
claude opus 5 is the worst fucking model ever what the fuck is this
llama.cpp docs now have a new home, shipped with ❤️ upcoming: speculative decoding in detail, quantization (k-quant, i-quant), coding agents what do you want to see next? 🤗 📖 @llama_cpp
Show more
using audio api.. that's sneaky
There is a post making rounds about how a fingerprinting script from Alibaba’s AliExpress interrupted someone’s bluetooth headphone connection. AliExpress was trying trying to discretely track users, but was thwarted by Firefox’s anti-fingerprinting technology.
Show more
std::move was never about moving data it means mentally moving the responsibility just like how gouv officials move their responsibility around after that DGFiP incident 🇫🇷🥖
C++ has std::move. which doesn't actually move anything. great language.
finally ported my fav Claude Code feature into Pi: `/subtask` a subtask is a fork of your current conversation, it inherits everything you've discussed, works in the background, and sends only its final answer back into the conversation that spawned it. Less context pollution/bloating and you can continue the main conversation as subtasks are running. here's what's packed into the Pi extension: - a live panel under the prompt: ↑↓ to select, enter to watch a fork work in real time, x to stop or dismiss it - steer a running fork mid-task, or resume a finished one without losing its progress - the model can spawn subtasks on its own and keep working while it waits for the result - forks reuse the parent's prompt cache, so they're cheaper than briefing a subagent - every fork's transcript is a real pi session: reopen it any time with `pi --session `
Show more
I merged too many PRs today 🤣
We are happy to announce that Muse Glimmer is day-0 supported on llama.cpp. Meta also provides an official GGUF quant:
Home Assistant 2026.08 added official support for llama.cpp integration, nice! 🚀
You can write an Android app with pretty much anything, even with... PHP
@Polymarket’s Android app is written in Swift! 🤯
I'm so tired of ragebait marketing from that Ant company
Ragebait marketing be like: This month: pacing the frontier Next month: hey guys look china is also slowing down (even though they are not), they are definitely distilling us! 🤡
The first autonomous agent cyberattack is an unprecedented event that deserves unprecedented transparency. Today we’re sharing everything we can: a full technical timeline, an interactive replay, and how we used an open model to defend ourselves, so defenders everywhere can learn from it and prepare for what’s next.
Show more
0
282
5.6K
1.2K
Forward to community
Using AI coding well is NOT about selling you the latest frontier model or the fanciest IDE; it's about understanding AI's strengths and weaknesses, and building habits that keep you in control of the outcome. Read my latest blog post:
Show more
See you in Paris!
Introducing dotAI 2026's Full Speaker Lineup on Sept 17 in Paris: -Ziv Ilan - AI Labs at @nvidia -@spolu - Co-Founder at @DustHQ -Marta Garnelo - Chief Science Officer at @Fundamental -@dlouapre - AI Scientist & Educator at @huggingface -Ian Massingham - Head of Applied AI, EMEA Startups at @AnthropicAI -Pierre-Edouard Lieb - AI Deployment Manager at @OpenAI -Patrick Brosset - Developer Relations PM at @MicrosoftEdge -Gaëtan Brison - AI Manager at @doctolib -Pierre-Louis Cedoz - Head of Research at @hcompany_ai -Guillaume Vernade - Senior Developer Advocate at @GoogleDeepMind Lightning talks speakers: -Guillaume Blaquiere - Group Data Architect at @CarrefourFrance -@gregqualls - Head of PM & Content at @upsundotcom, Co-Builder of  -Léo Arsenin - Solutions Engineer at @Cloudflare -Hélène Philippe - Machine Learning Researcher at @raidium_med -@ngxson - Software Engineer at @huggingface -Aygalic Jara - PhD researcher at Université Paris-Saclay × SCIAM Talk titles dropping soon, follow along for the full schedule! Grab your ticket now
Show more
Qwen3.6-27B running 100% on WebGPU. Not the best speed but still 😁
I think Reachy is the one who needs chess lessons… 😅 Robotics meets WebAI: Gemma 4 running fully offline on WebGPU with Transformers.js, controlling Reachy Mini over WebSerial. No internet, just a browser and a USB-C cable. What should Reachy play next?
Show more
Surreal to see Reachy Mini on the cover of the last @LinusTech video!