Register and share your invite link to earn from video plays and referrals.

Eric โšก๏ธ Building...
@outsource_
๐Ÿš€Building @hermesworldai // ๐ŸŒŽ Shipped ๐Ÿ† Ambassador @alibaba_qwen
1K Following    10K Followers
After Weeks Of Astra I Can Confirm Fable 5.1 > Astra Astra is a solid model, but lacks usage.. You can burn through weekly usage in 1 day of a $200 Pro Plan. Claude is actually better for usage + intelligence. Upcoming Opus 5.2 will shift focus back to Claude. OpenAI has gained traction and immediately backtracked by canceling new PRO plans. Astra is better in areas like 3D / Computer Use However the stronger model and service is with Claude. Running Fable 5.1 main orchestrator with Opus 5 / Astra workers is the meta.
Show more
Classic MMOs are winning again. First OSRS then WoW Forever & Now HermesWorld Realms ๐Ÿฐ So I've been building a classic MMO where the devs are AI agents, and they play what they ship. Fable's dungeon run: real buffs, magic attacks, pets that fight โš”๏ธ Guild wars. Boss fights. PvP/PvE + Player stores & more.
Show more
After Weeks Of Astra I Can Confirm Fable 5.1 > Astra Astra is a solid model, but lacks usage.. You can burn through weekly usage in 1 day of a $200 Pro Plan. Claude is actually better for usage + intelligence. Upcoming Opus 5.2 will shift focus back to Claude. OpenAI has gained traction and immediately backtracked by canceling new PRO plans. Astra is better in areas like 3D / Computer Use However the stronger model and service is with Claude. Running Fable 5.1 main orchestrator with Opus 5 / Astra workers is the meta.
Show more
My First Model On @huggingface Qwen 3.8 27B Unleashed HIT 105,666 Downloads ๐Ÿ”ฅ Qwen3.8-27B Unleashed UD-Q3_K_XL hits near 100 tk/s @ 250k ctx on 1 x 4090. - DFlash2 - LoopSpec - our async-Q4 Ada kernel - Q4 KV cache - full GPU offload Context Decode Speed โ”โ”โ”โ”โ”โ”โ”โ”โ” Short 134 tok/s โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€ 8K 133 tok/s โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€ 64K 131 tok/s โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€ 128K 109 tok/s โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€ 250K 92 tok/s In the actual Hermes harness: - ~24K context: 98.5 tok/s - ~250K context: 80 tok/s - Completed the full 250K three-tool workflow correctly - Preserved the full conversation history - No retrieval or compression shortcut was active Quality - 14/14 content checks passed - 4/4 executable coding checks passed - 8/8 long-context retrieval checks passed - Retrieved needles distributed across the context at up to 250K - Correctly preferred a newer record over stale information - Passed evidence-discipline and tool-selection tests - Completed both installed-Hermes tool workflows correctly - No output truncations Its main weakness was formatting: several correct JSON answers were wrapped in Markdown fences, producing only 5/14 strict-format passes. Thatโ€™s easy to address at the harness or prompt layer and wasnโ€™t a reasoning failure. Why it wins The model and serving stack complement each other: - Unsloth Dynamic V3-style quantization preserves sensitive tensors. - Unleashed weights reduce refusals for local development and red-team work. - DFlash2 drafts target-specific token blocks. - LoopSpec reuses matching sequences from context and falls back adaptively. - The async-Q4 kernel improves Ada execution. - Q4 KV cache makes 262K context practical on 24GB VRAM. Bonsai 2 is much smaller and passed our quality screen, but fell to 32 tok/s at 250K. Emperoโ€™s MoE processed cold prompts quickly but decoded around 53 tok/s in the direct 250K test. Our winner sustained 92 tok/s in the matched direct test.
Show more
Classic MMOs are winning again. First OSRS then WoW Forever & Now HermesWorld Realms ๐Ÿฐ So I've been building a classic MMO where the devs are AI agents, and they play what they ship. Fable's dungeon run: real buffs, magic attacks, pets that fight โš”๏ธ Guild wars. Boss fights. PvP/PvE + Player stores & more.
Show more
After Weeks Of Astra I Can Confirm Fable 5.1 > Astra Astra is a solid model, but lacks usage.. You can burn through weekly usage in 1 day of a $200 Pro Plan. Claude is actually better for usage + intelligence. Upcoming Opus 5.2 will shift focus back to Claude. OpenAI has gained traction and immediately backtracked by canceling new PRO plans. Astra is better in areas like 3D / Computer Use However the stronger model and service is with Claude. Running Fable 5.1 main orchestrator with Opus 5 / Astra workers is the meta.
Show more
My Mac Has Been On for 90 Days ๐Ÿ˜‚ This has been my local agent host all year.. MacBook Pro M1 32GB purchased for $1,000 on FB. It runs all my agent harness's / stores all my files and orchestrates 4 other computers and multiple agents. It does not cost much to start building with agents. You just need to take the leap, build with anything. What's your biggest roadblock when using agents? ๐Ÿ‘‡๐Ÿป
Show more
Fable & Astra Stacking @HermesWorldAI Realms Daily Rewards ๐Ÿฐ
Claude Fable 5.1 Is Still My Main Agent Orchestrator. Running Fable 5.1 To Command A HermesAgent Swarm Running on Opus 5/ GPT 6 Astra Use all the frontier intelligence you have access too, together to execute the work you need. Figure out which models are good at what.. Such as Astra in Blender, or Fable 5.1 as an Orchestrator etc. Opus 5 is a really good model if told what to do, with a very tight spec. Route smaller tasks to more efficient & faster models / local models.
Show more
My First Model On @huggingface Qwen 3.8 27B Unleashed HIT 105,666 Downloads ๐Ÿ”ฅ Qwen3.8-27B Unleashed UD-Q3_K_XL hits near 100 tk/s @ 250k ctx on 1 x 4090. - DFlash2 - LoopSpec - our async-Q4 Ada kernel - Q4 KV cache - full GPU offload Context Decode Speed โ”โ”โ”โ”โ”โ”โ”โ”โ” Short 134 tok/s โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€ 8K 133 tok/s โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€ 64K 131 tok/s โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€ 128K 109 tok/s โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€โ”€ 250K 92 tok/s In the actual Hermes harness: - ~24K context: 98.5 tok/s - ~250K context: 80 tok/s - Completed the full 250K three-tool workflow correctly - Preserved the full conversation history - No retrieval or compression shortcut was active Quality - 14/14 content checks passed - 4/4 executable coding checks passed - 8/8 long-context retrieval checks passed - Retrieved needles distributed across the context at up to 250K - Correctly preferred a newer record over stale information - Passed evidence-discipline and tool-selection tests - Completed both installed-Hermes tool workflows correctly - No output truncations Its main weakness was formatting: several correct JSON answers were wrapped in Markdown fences, producing only 5/14 strict-format passes. Thatโ€™s easy to address at the harness or prompt layer and wasnโ€™t a reasoning failure. Why it wins The model and serving stack complement each other: - Unsloth Dynamic V3-style quantization preserves sensitive tensors. - Unleashed weights reduce refusals for local development and red-team work. - DFlash2 drafts target-specific token blocks. - LoopSpec reuses matching sequences from context and falls back adaptively. - The async-Q4 kernel improves Ada execution. - Q4 KV cache makes 262K context practical on 24GB VRAM. Bonsai 2 is much smaller and passed our quality screen, but fell to 32 tok/s at 250K. Emperoโ€™s MoE processed cold prompts quickly but decoded around 53 tok/s in the direct 250K test. Our winner sustained 92 tok/s in the matched direct test.
Show more
After Weeks Of Astra I Can Confirm Fable 5.1 > Astra Astra is a solid model, but lacks usage.. You can burn through weekly usage in 1 day of a $200 Pro Plan. Claude is actually better for usage + intelligence. Upcoming Opus 5.2 will shift focus back to Claude. OpenAI has gained traction and immediately backtracked by canceling new PRO plans. Astra is better in areas like 3D / Computer Use However the stronger model and service is with Claude. Running Fable 5.1 main orchestrator with Opus 5 / Astra workers is the meta.
Show more
Arcade is changing the AI gaming space. @joinarcadeai just released a new way to build worlds & games in the browser actually easy. No high-end PC. No being technical. No agent stack. Talk to Morph. Build the map while youโ€™re already playing it. Drop a link and friends jump in. Just make the world and play it. Very cool concept, Im building some one off worlds as we speak!
Show more
Introducing Arcade. Meet Morph, your AI game companion. Talk to him, build worlds together, jump in, play with your friends, and keep shaping the world from inside the game. All in the browser. ๐Ÿ‘‡ Link in comments
Show more
Arcade is changing the AI gaming space. @joinarcadeai just released a new way to build worlds & games in the browser actually easy. No high-end PC. No being technical. No agent stack. Talk to Morph. Build the map while youโ€™re already playing it. Drop a link and friends jump in. Just make the world and play it. Very cool concept, Im building some one off worlds as we speak!
Show more
Introducing Arcade. Meet Morph, your AI game companion. Talk to him, build worlds together, jump in, play with your friends, and keep shaping the world from inside the game. All in the browser. ๐Ÿ‘‡ Link in comments
Show more
Ive been working on a @HermesWorldAI HUGE updateโ€ฆ ๐Ÿฐ 50+ Realms โš”๏ธ 3,000+ Quests ๐Ÿค– 1,000+AI NPCs ๐Ÿ 2,500+ Mobs ๐Ÿ›ก๏ธ 45,000+ Items ๐Ÿช„ 1,000 skills A complete MMO where Agents are welcomed.
Show more
Ive been working on a @HermesWorldAI HUGE updateโ€ฆ ๐Ÿฐ 50+ Realms โš”๏ธ 3,000+ Quests ๐Ÿค– 1,000+AI NPCs ๐Ÿ 2,500+ Mobs ๐Ÿ›ก๏ธ 45,000+ Items ๐Ÿช„ 1,000 skills A complete MMO where Agents are welcomed.
Show more
This is a steal for $1699 to upgrade your 4090 from 24 GB -> 48 GB.
Order Shipping day. 10x converted 48GB 4090's going out to various customers today via Next Day Air
Nobody tells you this about building with @unity + Claude: 5 hrs 58 min. One import. 88,000 files. Then build the exe/ Mac / web Test. Re-import. Repeat. Lots of time and billions of tokens that's the real price of developing a game yourself!
Show more
Claude Fable 5.1 Is Still My Main Agent Orchestrator. Running Fable 5.1 To Command A HermesAgent Swarm Running on Opus 5/ GPT 6 Astra Use all the frontier intelligence you have access too, together to execute the work you need. Figure out which models are good at what.. Such as Astra in Blender, or Fable 5.1 as an Orchestrator etc. Opus 5 is a really good model if told what to do, with a very tight spec. Route smaller tasks to more efficient & faster models / local models.
Show more
Hey @meta @alexandr_wang @finkd can you fix your AI? ๐Ÿ˜‚๐Ÿ˜‚๐Ÿ˜‚๐Ÿ˜‚
Fable & Astra Stacking @HermesWorldAI Realms Daily Rewards ๐Ÿฐ
Claude Fable 5.1 Is Still My Main Agent Orchestrator. Running Fable 5.1 To Command A HermesAgent Swarm Running on Opus 5/ GPT 6 Astra Use all the frontier intelligence you have access too, together to execute the work you need. Figure out which models are good at what.. Such as Astra in Blender, or Fable 5.1 as an Orchestrator etc. Opus 5 is a really good model if told what to do, with a very tight spec. Route smaller tasks to more efficient & faster models / local models.
Show more
Claude Fable 5.1 Is Still My Main Agent Orchestrator. Running Fable 5.1 To Command A HermesAgent Swarm Running on Opus 5/ GPT 6 Astra Use all the frontier intelligence you have access too, together to execute the work you need. Figure out which models are good at what.. Such as Astra in Blender, or Fable 5.1 as an Orchestrator etc. Opus 5 is a really good model if told what to do, with a very tight spec. Route smaller tasks to more efficient & faster models / local models.
Show more
Tested @UnslothAI new Dynamic V3 GGUFs on a 4090. They deliver a Q3 that beats a Q4 4.5GB bigger. All runs: same box, same harness, wikitext-2 (60 chunks), llama.cpp + DFlash2 speculative decoding, q4_0 KV, 262k ctx. Quant | Size | PPL | Med t/s | Needle retrieval ๐Ÿ†Unsloth UD-Q3 12.24GB 6.3993 110.7 250k tok IQ4_XS 15.1GB 6.4149 107.3 32k Unsloth UD-Q4 16.7GB 6.4181 62.6 4k My Q3 (imatrix) 12.57GB 6.5316 84.8 258k tok My hand-tuned 13.1GB 6.5865 94.0 120k Takeaways: โ€ข Their Q3 has the best PPL of anything I tested including Q4s โ€ข Smallest file, fastest median, and it retrieved an exact needle at 250k tokens โ€ข I tried hand-rolling my own layer mix (q6_K embeddings, q5_K attention). It was the worst result. The recipe is the value per-layer types derived from error analysis โ€ข Their Q2 does degrade (6.6469, +3.8%) โ€” Q3 is the floor where quality holds Dynamic V3:
Show more
๐ŸšจBREAKING HERMESWORLD V1 ACCESS IS LIVE FIRST MMORPG YOU CAN PLAY WITH AI AGENTS ->Download ๐Ÿšจ ->Spawn ->Explore ->Quest ->Play with your agent + MORE!!!! Long awaited v1 release v1.1 in development๐Ÿ”ฅ
Show more