Register and share your invite link to earn from video plays and referrals.

Tim Dettmers
@Tim_Dettmers
Creator of bitsandbytes. Professor @CarnegieMellon and Research Scientist @allen_ai . I blog about deep learning and PhD life at
Joined October 2012
917 Following    48.4K Followers
With our new efficiency methods, you will be able to run this on a single DGX Spark or AMD Strix Halo at 7 token/s decode and >250 tok/s prefill. Stay tuned!
Introducing GLM-5.3: Built to Code. Ready for Cyber Defense. - Top-tier coding and agentic capabilities, achieved through post-training on the 743B base model - A major leap in cybersecurity, setting a new standard among open models Tech Blog:
Show more