Hermes Agent has had some MASSIVE updates the last month NOBODY is talking about
- Bot mode
- Local AI mode
- Improved subagents
- Hermes HUD
and much more
It's the most powerful open source agent ever
In this video I cover EVERY new change and how to get the most out of them
Show more
Everyone on planet Earth is talking about local AI right now
And for good reason
Governments are banning models. Hardware prices are 10xing
You NEED to be getting into local AI. The number 1 questions everyone has though is which computer to buy?
Here's your answer:
You basically have 3 options:
1. MAC STUDIO (high intelligence, lower speeds)-
Mac Studios are excellent devices for local AI. They can run MASSIVE models. I'm running GLM 5.2 right now on a single Mac Studio. The model is Opus 4.8 level
The issue is, Mac Studios are a bit slower at running intelligence (altho the M5 Ultra is promising)
Mac Studios are a good choice for you if you want frontier level intelligence, but are fine running the intelligence passively
Meaning you get top intelligence, but it runs more in the background rather than on demand
As an example, I have GLM 5.2 running security checks on my codebase every hour. It creates a report. I review this later in the day
2. POWERHOUSE NVIDIA CHIPS (RTX 5090, 6000 Pro)
Nvidia is the most valuable company in the world, and for good reason
They make the world's best GPUs.
They have decent VRAM (32gb on the 5090, 96gb on the 6000 Pro) and INSANE bandwidth. Meaning the local models run at unbelievable speeds
I'm running Qwen 3.8 locally on a 5090 and it's just as fast as cloud models
I'd go this route if you want to run an AI agent like Hermes off a local model, still get decent intelligence, but have it able to work lightning fast
3. AI WORKSTATIONS (DGX Spark type computers)
The DGX Spark is an excellent AI computer
It has high memory (128gb unified memory) and has decent speeds because of the Nvidia CUDA architecture
It is basically the sweet spot between a cutting edge Nvidia chip and a Mac Studio
You can run medium sized models, and get usable speeds out of them
You're not going to get the same performance as cloud models, but it will allow you to offload small secondary tasks to your local models for them to handle
They are also the absolute easiest to get up and running
You plug it in, then tell your agent on your main computer to go onto it and set it up. You don't even need it connected to a monitor
CONCLUSION
Here's what it comes down to: how high intelligence do you need, what speeds do you need, and how plug and play do you want?
Want the highest speeds, like you are used to with cloud compute? Build a computer around an RTX 5090
Want to run frontier level intelligence, and don't mind slower speeds, go with a Mac Studio
Either way, it's never been more important to get into local AI
Show more
Everyone on planet Earth is talking about local AI right now
And for good reason
Governments are banning models. Hardware prices are 10xing
You NEED to be getting into local AI. The number 1 questions everyone has though is which computer to buy?
Here's your answer:
You basically have 3 options:
1. MAC STUDIO (high intelligence, lower speeds)-
Mac Studios are excellent devices for local AI. They can run MASSIVE models. I'm running GLM 5.2 right now on a single Mac Studio. The model is Opus 4.8 level
The issue is, Mac Studios are a bit slower at running intelligence (altho the M5 Ultra is promising)
Mac Studios are a good choice for you if you want frontier level intelligence, but are fine running the intelligence passively
Meaning you get top intelligence, but it runs more in the background rather than on demand
As an example, I have GLM 5.2 running security checks on my codebase every hour. It creates a report. I review this later in the day
2. POWERHOUSE NVIDIA CHIPS (RTX 5090, 6000 Pro)
Nvidia is the most valuable company in the world, and for good reason
They make the world's best GPUs.
They have decent VRAM (32gb on the 5090, 96gb on the 6000 Pro) and INSANE bandwidth. Meaning the local models run at unbelievable speeds
I'm running Qwen 3.8 locally on a 5090 and it's just as fast as cloud models
I'd go this route if you want to run an AI agent like Hermes off a local model, still get decent intelligence, but have it able to work lightning fast
3. AI WORKSTATIONS (DGX Spark type computers)
The DGX Spark is an excellent AI computer
It has high memory (128gb unified memory) and has decent speeds because of the Nvidia CUDA architecture
It is basically the sweet spot between a cutting edge Nvidia chip and a Mac Studio
You can run medium sized models, and get usable speeds out of them
You're not going to get the same performance as cloud models, but it will allow you to offload small secondary tasks to your local models for them to handle
They are also the absolute easiest to get up and running
You plug it in, then tell your agent on your main computer to go onto it and set it up. You don't even need it connected to a monitor
CONCLUSION
Here's what it comes down to: how high intelligence do you need, what speeds do you need, and how plug and play do you want?
Want the highest speeds, like you are used to with cloud compute? Build a computer around an RTX 5090
Want to run frontier level intelligence, and don't mind slower speeds, go with a Mac Studio
Either way, it's never been more important to get into local AI
Show more
Meta Muse is an UNBELIEVABLE AI agent
You can literally have it go through all of your credit card bills, find your subscriptions, and autonomously negotiate them all down
Craziest part? It's free
In this video I cover how Muse works and how to master it:
Show more
Meta Muse is an UNBELIEVABLE AI agent
You can literally have it go through all of your credit card bills, find your subscriptions, and autonomously negotiate them all down
Craziest part? It's free
In this video I cover how Muse works and how to master it:
Show more
Stop scrolling for a second.
I just want you to pause for a moment. We live in the most incredible time to be alive in the history of this species
Every other day a new revolutionary product drops that gives you more freedom, power, and ability to do ANYTHING you want
LITERALLY every other day
A couple weeks ago Astra gave you super intelligence. Today Opus gives you super intelligence at lightning speeds for dirt cheap. Soon Meta glasses will come out that let you talk to a super intelligent personal agent everywhere you go
Every day technology makes your life better and better. Every day you're capable of accomplishing more. Every day you have the ability to help out your fellow man more than ever before
For thousands of years nothing happened. But today, everything is happening
Please just take a moment and realize how incredible, awe inspiring, violently beautiful this all is
If you are one of the pessimists that haven't realized the incredible nature of what is happening right now, I beg you to view this world through a different lens
Being alive right now is the greatest gift God has ever given us. I hope everyone realizes this soon enough
Show more
Claude Opus 5.5 is the greatest AI model ever released
It is the most intelligent model I've ever used, while be lightning fast and incredibly efficient
The perfect trifecta
In this video I cover why this should be your new daily driver and how to get the most out of it:
Show more
These next 2 weeks are going to be the most insane 2 weeks in technology history
All of these are rumored to drop:
1. ChatGPT 6.1 Astra
2. Claude Fable 5.5
3. Massive Grok Bot functionality upgrades
4. Muse hardware integrations
5. Tons of new products from OpenAI dev day
If you thought AI labs were going to slow down, you’re extremely mistaken
Opus 5.5 is the greatest model of all time and shocked the entire industry. Every AI lab is moving their release dates up
We will accelerate faster than ever
Moments like this present OUTRAGEOUS amounts of opportunity
If you get ahead and use these new pieces of tech immediately, you have an edge over your competition
You can build things faster, smarter, better than all the other people in your space
Cancel all your plans. All your appointments. All the people you were going to talk to
Stand by X. Don't move
The moment anything new drops, use it to its fullest.
I'll be dropping guides on each when they come out
The great lock in has begun
Show more
Claude Opus 5.5 is the best AI model I’ve ever used
I was lucky enough to have early access and I’ve been using it nonstop
It’s smarter than Fable and Astra yet it’s:
• Significantly faster
• A fraction of the price
• And most importantly: WAY better to talk to
My biggest complaint for ALL AI models the past few months is they’ve all been really annoying to talk to
Every frontier model from every company has all developed this weird AI language. They don’t feel ‘human’ anymore
You read paragraphs of text and it’s like you read nothing
Opus 5.5 changed that. It’s the first model in months to feel human again. It is just a total pleasure to talk to
Highly encourage you to try it out
Show more
Grok Bot just released for Tesla and I'm blown away
I was lucky enough to have early access. Having your car drive you around while you talk to an army of agents is incredible
In this video I take you for a ride in my Cybertruck and show you just how awesome this new release is
Show more
I believe Opus 5.5 is the first model that was made as a result of RSI
It's the first model I've used that improved and became near frontier on basically every metric, while also getting faster and cheaper
Basically no downsides
There feels like there's some sort of magic behind it that's hard to describe
I'd highly encourage you to use this model for more 'exploration'
Brain dumping ideas and thoughts, and asking it to explore what could come out of them. What you could build. How you could improve your internal operating systems
I've gotten a tremendous amount of novel ideas out of this model. Things that I've never thought of before
For instance I told it I bought a new Apple Watch Ultra 4. I said what should I do with it.
An hour later I had an app on my Watch that showed all my Herdr agents working and allowed me to dictate commands to them
Things I never even thought of doing just appeared in front of me
Do yourself a favor and just carve out an hour tonight to do this type of exploration with this model. I promise you'll get some amazing results
We truly live in the most amazing time
Show more
Claude Opus 5.5 is the best AI model I’ve ever used
I was lucky enough to have early access and I’ve been using it nonstop
It’s smarter than Fable and Astra yet it’s:
• Significantly faster
• A fraction of the price
• And most importantly: WAY better to talk to
My biggest complaint for ALL AI models the past few months is they’ve all been really annoying to talk to
Every frontier model from every company has all developed this weird AI language. They don’t feel ‘human’ anymore
You read paragraphs of text and it’s like you read nothing
Opus 5.5 changed that. It’s the first model in months to feel human again. It is just a total pleasure to talk to
Highly encourage you to try it out
Show more
Grok Bot just released for Tesla and I'm blown away
I was lucky enough to have early access. Having your car drive you around while you talk to an army of agents is incredible
In this video I take you for a ride in my Cybertruck and show you just how awesome this new release is
Show more
Grok 4.7 just released and it's an EXCELLENT model
It was trained FOR Grok Bot
Meaning this is a fully agentic model trained to do your knowledge work better than you can
In this video I show you how to use Grok 4.7 and a Grok Bot workflow that will 10x your productivity:
Show more
Grok 4.7 just released and it's an EXCELLENT model
It was trained FOR Grok Bot
Meaning this is a fully agentic model trained to do your knowledge work better than you can
In this video I show you how to use Grok 4.7 and a Grok Bot workflow that will 10x your productivity:
Show more
It happened. Grok 4.7 dropped
Better intelligence than Opus 5. Half the price
Fully baked into my favorite AI agent harness at the moment: Grok Bot
If you haven't tried using cloud cursor agents inside Grok Bot, now is by far the best time to do it
Choose a project you want to work on, connect your github, ask a grok bot to do work on it
It will spin up Cursor cloud agents and write code in the cloud. Lightning fast and incredibly smart
I recommend using a project management tools like Linear or Notion to make a bunch of tasks first, then have cloud agents just tear through them all 1 by 1.
You'll get a massive amount of work done without much oversight.
Big opportunity to lock in right now and get ahead of the curve with new tech
Take my steps up above and get to it
Show more
Grok 4.7 is here.
It's a notable improvement over Grok 4.6 at the same price and speed.
The AI agent race will be the most important tech race in human history
1000x bigger than the AI model race
Whoever wins gets ALL our personal data, gets a % of ALL our transactions, and basically determines the information we see and what we purchase
Somehow AI agents have made people completely disregard privacy. Nobody cares. They connect literally EVERYTHING to their agents
All their emails, personal texts, calendars, buying habits, everything. Never in history have humans cared so little about privacy
We are willing to give up quite literally every piece of data we have in order to save time on responding to emails
And this isn't criticism. I've done it. I've connected everything to these agents. They're wildly helpful. I don't regret it. I'm making thr sacrifice to increase efficiency
But we are about to enter an age where all of these tech companies will pour every penny they have into pushing their agent. The prize for the winner is too big
They'll have every piece of data from all users, plus you know they will start taking a transaction fee from every purchase the agent makes
Not to mention the advertising opportunities. They will 100% start charging for companies to be prioritized when the agent makes buying decisions. This will be a bigger ad revenue opportunity than anything we've seen before
Agents will be the most profitable area in tech ever and it wont be remotely close.
Now it's just up to you if you want to give every piece of your data to Meta, SpaceX, OpenAI, Anthropic, or Google
Show more
Linux is so obviously the future
Open source model controlling an open source agent harness working on an open source operating system
If you don't see this you're blind
Some AI labs are trying to get local models banned
It's critical you get into local AI before you no longer can
Luckily Hermes Agent just added a 1 click feature that allows ANYONE no matter which computer you have to load free, unlimited local AI
Here is how to get it set up:
Show more
This is what the future of AI is
One ‘god’ thread you talk to that manages tons of other agents and threads
It’s WAY simpler than managing hundreds of smaller threads. Also just more fun
Cursor just added it. ChatGPT Voice is basically this. Grok Bot w/ chief of staff is this
Highly recommend switching to this paradigm in all your agents. Just talk to 1 thread and have it spin up and manage other threads
Will make it way easier to quickly open your apps and get work done, because you always go to one thread
Show more
Projects now run from one conversation, starting in Claude Code. You describe what needs doing, and Claude directs parallel threads that keep working after you close your laptop.
In beta today for select Pro and Max users in cloud sessions; coming to all Claude users soon.
Show more