anthropic's moat just got a lot smaller.
researchers across NUS, Stanford, Oxford, and Princeton built on anthropic's own agent-skill format, then showed the format alone isn't what does the work.
the number that stands out:
an agent with every skill dumped straight into its context scored 65.6% on a long-horizon retail task.
the same agent, the same skills, but with a working memory deciding which one to invoke and when, hit 83.6%. same skill library, 18 points apart.
different setup than what i wrote about, but the same underlying idea: having the skill isn't the unlock, knowing when to reach for it is.
wrote the full guide on doing this with everything you already know: books, docs sites, videos, the open web, your own notes.
full article below đ