Register and share your invite link to earn from video plays and referrals.

Taalas Inc.
@taalas_inc
The Model is The Computer
3 Following    17.7K Followers
We are pleased to share that Taalas has agreed to join AMD. We built Taalas to rethink AI inference from the ground up: hardware designed around the model, rather than the other way around. The result is the world's fastest and most cost-effective inference silicon. Joining AMD gives us the scale, engineering resources, and global reach to bring that work to the world – and to accelerate what comes after it. We are also excited to build on AMD's long-standing presence in Canada and its continued commitment to the country's AI ecosystem. Many of us grew up at AMD Canada, and are looking forward to coming home. We are proud of what this team has built, and even more excited about what comes next.
Show more
RT @powerpig: Yes, it's fast. It "thinks" it's self-aware, believes it's Claude, and tries to reason that away after I confront it. No, I…
Some comments on Taalas HC1: - It’s real. Try it yourself. At ~16k tokens/sec, the output is instantaneous. - The current demo model is aggressively quantized (roughly 3–6 bits). The goal was to prove the system works end-to-end. Improving quantization quality, that's the easy part. - Their next iteration, a mid-size reasoning LLM, will be much more accurate. - The weights are frozen, but the chip supports LoRA adapters (high-rank), so you can still adapt it to your domain. In practice, you could also distill knowledge from newer/larger models into adapters to “refresh” what the chip can do without changing base weights. - Frontier open-weight models to land on the platform this year.
Show more
yesterday we chatted with @martin_casado and @sarahdingwang on the pod and he happened to do basic math™ on the logic of asics today @taalas_inc launched their HC1 asic that can inference 17k tok/s. Sure, it's a shitty 3.1 8B today which is a 1.5 year gap. But read the details to the HC2 this winter, and do the math — this timeline will converge to 0 in the next 2 years. Build accordingly.
Show more
I know I'm going to come back and re-read this quite a few times.
24 dedicated people. $30M spent on development. Extreme specialization, speed, and power efficiency. Today we launch Taalas’ first product. Check it out: Details:  Demo chatbot:  API: 
Show more
0
468
6.1K
583
Forward to community