Register and share your invite link to earn from video plays and referrals.

the tiny corp
@__tinygrad__
We make tinygrad; sell tinybox for the GPU middle class. Our mission is to commoditize the petaflop.
204 Following    82.6K Followers
Despite record usage of the models, ZAI stock is down over 2x from its peak. The Chinese markets have a real understanding of the economic impact of AI, the US ones do not. It's very important that the US government doesn't bail any of this out.
Show more
@chamath The cyber incidents weren't enough for them to get the regulatory capture they wanted. They need to make a bio incident next. I hope our government is on top of this, their ideology has no limits to what they will do for "good"
Show more
Overall, a very tasteful chip launch by OpenAI. It takes the best of both GPUs and TPUs. I wish they were for sale, but we'll all get the Chinese knockoffs in 3 years. I can't wait until China starts spamming fab capacity.
Show more
0
18
1.2K
33
Forward to community
We merged some high performance kernels for RDNA3 Qwen 3.8 27B. Written in the tinygrad kernel language, which are the same UOps with memory movement specified. 46 tok/s on 7900XTX, no spec decode.
Show more
I can't believe Qualcomm's market cap is only 168.79B. Does anyone know if a leveraged buyout is possible? With new leadership that actually wants to sell chips it could easily 5x, we'll get a Qualcomm chip in every robot. And save money by firing the whole 80s era sales org.
Show more
0
34
1.2K
34
Forward to community
@IanCutress Posting FLOPS, GB, GB/s, and W is not a meme dump, it's what every single respectable chip maker does. Compare the recent Cerebras press to the Etched one. Etched posted a video where they put an empty wooden box in a truck and said don't worry our lead investor thinks it's good.
Show more
Chestnut purchased, I'll shoot at the 3090
Installed the Chestnut eGPU from @comma_ai. Already noticing improvements, way sooner than I expected from these first chestnut-class models. Braking/stopping is noticeably smoother and more natural, and it starts slowing down much earlier. More videos coming soon.
Show more
We're making room in our CI racks and selling 6x4090 tinyboxes for $35k. This is the original tinybox green, and we have 3 for sale. They are great machines, we're just mainly focused on AMD these days and need the power for more AMD GPUs.
Show more
The team will be in Hong Kong in October, PM me on Discord if you are a known tinygrad contributor and want an invite.
tiny chestnut eGPU dock is getting rave reviews in the gpu-on-usb channel on our Discord. Shipping right away from the comma shop.
tinygrad finally has a SOTA speed result! This is MLPerf llama31_8b in 2h 6m, beating the 2h 7m time from AMD's docker on our computer. We aren't ahead in wall time yet because it's warm in San Diego right now and we are too cheap to buy AC. Can you spot the thermal throttle?
Show more
Oh, I didn't know Taalas got bought by AMD. They are an example of how you should market a new chip company, drop a benchmark with a demo to prove it! If there's no benchmark numbers, it's because they aren't good.
Show more
We are pleased to share that Taalas has agreed to join AMD. We built Taalas to rethink AI inference from the ground up: hardware designed around the model, rather than the other way around. The result is the world's fastest and most cost-effective inference silicon. Joining AMD gives us the scale, engineering resources, and global reach to bring that work to the world – and to accelerate what comes after it. We are also excited to build on AMD's long-standing presence in Canada and its continued commitment to the country's AI ecosystem. Many of us grew up at AMD Canada, and are looking forward to coming home. We are proud of what this team has built, and even more excited about what comes next.
Show more
@Etched There's been a rapidly changing product story, nonsensical technical claims, and zero third party validation of the Etched chips. We need to hold the individuals hyping this to a higher standard.
Show more
this smells very bad. none of the claims make sense. > low voltage hits 80% mfu high mfu is easy if your peak flops is low > low voltage solves power bottleneck bleeding edge wafers is more scarce than power and it doesn't make sense to tape out lower perf chips on these wafers to save power if you want perf/$. the true bottleneck to frontier perf is max flops density to minimize going off die. power is the most fungible elastic commodity on earth > half voltage = quarter power presumably: 1. only matmul gates (~60% of die power) at half voltage. half voltage -> ~3x slower -> need ~3x more transistors, and their power leaks scale linearly. ~20% net savings at chip level 2. half voltage -> engineering complexity to fix timing violations + exponentially scaled soft errors (~20x). matmuls will randomly corrupt undetectably from bit flips (and drop model intelligence). > btc miners run at 3x lower V btc workload is hashing, so 1) error checking is literally in the problem 2) arithmetic intensity is like infinity so they don't care about packing flops density in single die unlike AI workloads. doesn't make sense to copy > GPUs get low MFU from thermals no, they get low MFU from mem bw and non-matmul ops. > CSM 5x lower latency than Blackwell 4000ns Blackwell nvswitch is 300ns? unless they are comparing their bare hardware to nvidia hardware + software. the whole product design tradeoffs don't make any sense from first principles and only makes sense if their initial "transformer asic" had some horrible power issue that they solved by undervolting and they had to respin the whole story for investor marketing. would love to be proven wrong if anyone from Etched wants to educate me :)
Show more
Public service announcement about @Etched. It's possible they have a great chip and just very distasteful marketing. It's also possible they have no chip or a terrible chip and want you to fall for the smoke and mirrors. This type of marketing is bad for technology.
Show more
0
74
1.3K
44
Forward to community
Is this your chestnut being tested?
We really can't move to our self hosted Gitea soon enough. How is this an enterprise service people rely on? @Microsoft @github
0
81
1.7K
53
Forward to community
Here's Qwen 3.6 27B on an AMD 7900XTX over USB3 (literally any computer made in the last 10 years) at 34 tok/s. The eGPU dock that supports this launches on the 12th, with 100% open source firmware and an extra USB port for serial + unbrickability.
Show more
0
67
1.9K
102
Forward to community
Got code exec on AMD 7900XTX last night, Kimi's custom MEC firmware just ran its first kernel! After spec and HCQ2, the next phase of tinygrad will be operating system.
Kimi K3 running on AMD MI350X with SGLang. Amazing work to @AMD @sgl_project @Kimi_Moonshot this worked great out of the box!
0
28
1.5K
109
Forward to community