Register and share your invite link to earn from video plays and referrals.

Search results for holymoly
holymoly community
One keyword maps to one global community path.
Create community
People
Not Found
Tweets including holymoly
holy moly fable usage limits suck now
HOLY MOLY, THE L2D????
0
69
7.6K
371
Forward to community
Holy moly, Hoshino super serious series! 😳😳
just saw minh’s pnl holy moly
SpaceX Q2-26 Earnings call Bret Johnson, CFO Holy moly "in the first few weeks of the third quarter, we've already contracted an additional $6.7 billion of cloud services revenue over a six month period that begins ramping, starting in October of this year. We believe this puts us on a trajectory, including contribution from cursor to reach 100 billion of RR or annualized revenue run rate by the end of this year."
Show more
Okay, just finished my first livestream and holy moly, @nikitabier and team did such a great job with this. Everything worked perfectly the first try. I shared a completely new benchmark that nobody has done yet comparing Fable Low vs. GLM 5.2 High vs. Sonnet 5, and well, you'll just have to watch to see the results. This was not at all the outcome I had expected, and I'll tell you, it goes to show that price based on input/output tokens probably isn't as meaningful as you might think. And that you're also probably sleeping on Fable Low.
Show more
Was running my GLM 5.3 Flash benchmark on @VulcanBench and everything crashed. Investigated, and holy moly, Docker was taking up 900GB+ of space 😳 Fixing and resuming the sweep. Committed to having GLM 5.3 Flash benchmarks in this weekend 🖖
Show more
Okay, my Muse Spark 1.2 benchmark with @VulcanBench is done, and holy moly, this was a wild one. I benchmarked it on Eval Suite 3, with the same rules I give every model, and it just couldn't get a good chunk of tasks solved in the time limit. As a quick reminder, VulcanBench is focused on all real engineering tasks, things swe's would give to a model during regular daily work, no weird math puzzles or exotic architecture. And every model is given the same time constraints, scaled by task difficulty. This is how engineering leaders like me make decisions, we can't have models that solve something in an infinite amount of time, when another model can solve it 10x faster. Grok 4.5 High remains at the top of the leaderboard, you can see the full leaderboard here: For more details on the Muse Spark 1.2 benchmark, you can see the detailed report here:
Show more