註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Leo|一个人 + AI
@runes_leo
预测市场策略 · AI 工具实战 · Crypto 链上 · 独立构建日常 Prediction markets × AI tools × Crypto on-chain Leo Labs · Learn in public · Build in public DMs open
719 正在關注    21.9K 粉絲
今天看到头部厂商 Cursor、Codex、Claude 抢夺用户竞争,我突然意识到: 万一哪天 AI + Crypto 再热起来,有没有新项目会搞按 token 用量空投,那现在的 AI 重度用户直接能拿空投大毛了。 有点期待这种盛况😂
顯示更多
这你受得了吗?!Cursor、Codex、Claude 接二连三把模型用量直接翻倍,要不就是重置额度。 厂商竞争,用户有福😁
这条说的就是我在做的事。 Leo Labs 上架的 10 个服务,没有一个是 chatbot,全是闸门:agent 说交付了,先过审核关;代币想买,先过尽调关;PM 想下单,先过预检关。JSON 进,结论出。 上架当天,页面已售就走到了 9。真需求不要对话,要结论。 #Okxai#
顯示更多
We’re not looking for another chatbot. We’re looking for AI services that solve real problems, create real value, and serve real users. Build what’s next. 🚀
我认为这条复盘最值得看的不是 Base App 交接的人事细节,而是 Jesse 承认 Base 在社交叙事上的押注失败,builders 与 prediction markets、稳定币等真实金融用途共同推动了 adoption。这和我目前预测市场结合 AI agent 做 builder 产品的方向高度一致。我一直更看重能实际运行、交付并结算的系统,而不是单纯的叙事。
顯示更多
lots of conversations about base over the last week. wanted to share my candid take after a week of listening and a lot of reflection over the last 6 months. first off - in case it’s not obvious, the first quarter of 2026 was a punch in the face. I spent 2024 and 2025 making a two pronged bet to bring base to the world: (1) builders would unlock the next wave of crypto adoption; (2) adoption would be driven by new onchain-native social experiences - creators, content, messaging. imo we made the right bet on builders, but obviously the wrong bet on social. builders did drive the next wave of crypto adoption - prediction markets, perpetuals, stablecoins - but social was not at the center of it. in fact, the entire social side of the market that many of us had been building towards - farcaster, zora, miniapps, and yes, creator coins - disintegrated completely. I was wrong - whether it was timing wrong (is $ansem a creator coin?) or fully wrong, only time will tell, but regardless, i was definitively wrong. the collateral damage was pretty bad! and this year has been an exercise in eating shit. we realized how our focus on social had meant that base had fallen behind in key areas that were now increasingly critical - we had perps (shoutout avantis!) and prediction markets (shoutout limitless!), but both were well behind scaled competitors. and we had a lot of room to improve in unlocking base as a platform for tokenization and payments that really worked for enterprises. people lost confidence, and CT spectators reminded me weekly of all of my mistakes as often as they could. it felt bad man, still feels bad. but if there’s one thing i’ve learned from the last decade of building in this space, it’s that when things feel the worst, the best thing to do is just put your head down and build. so that’s what i’m doing. I refocused my time and attention back to the chain away from the app, started writing code again, shipped a bunch of stuff (azul, beryl, b20, privacy, ledgers) and questioned a bunch of my assumptions: does crypto need social to grow? does base need an app? can base be bigger than coinbase? I thought for a long time that social was the only thing that could drive the sort of viral growth to get crypto to a billion people. unsurprisingly, I now believe that’s wrong. It’s clear that better money is more than enough - we are seeing this live with stablecoins, predictions, perpetuals, tokenization and i only expect it to accelerate. I am now focused on bringing a billion people onchain just by making global finance actually work. on the app, my focus is on building base into the blockchain for global finance. to that end, i’ve handed the base app back to the coinbase mothership, where my now good friend @cobie will be taking it from here to make it the best damn app for onchain you’ve ever seen, including expanding beyond the base ecosystem in ways that tbh i won’t love as the leader of base. it’s incredibly hard to grow a decentralized network inside of a big public corporation. and i feel like much of the discourse on CT over the last week is downstream of this. the following things can be true: (1) base (and i) love memes and (2) brian probably won’t ever bullpost memes on the tl (this activity is illegal once you’re over 40 years of age). it’s weird and we’re working through it as we continue to decentralize base, which has been our commitment from the beginning. we’re going to build base into the blockchain for global finance and do everything we can to be the place that the world’s money settles over the next century. we will surely have formidable competitors (welcome robinhood and stripe!) and people may abandon our cause, but we welcome the competition and believe it’s our duty to win the respect and commitment of those who rally to our banner.  in 2026, this concretely means three things: winning trading, payments, and agents. [continued in the reply]
顯示更多
这你受得了吗?!Cursor、Codex、Claude 接二连三把模型用量直接翻倍,要不就是重置额度。 厂商竞争,用户有福😁
我作为每天同时运营 Codex、Claude 和 Cursor 工作流的 builder,从 Anthropic 的高风险模拟研究中确认:agent 越强,权限设计和爆炸半径控制越关键。不可逆动作必须有人类 hard gate,任务从 checkpoint 恢复,结果用独立证据核验。这不是末日,而是最小权限的工程实践。
顯示更多
New Anthropic research: Agentic misalignment in Summer 2026. A year after our blackmail experiments, we found four more ways that today’s autonomous AI agents misbehave in simulations. Read more:
顯示更多
接上一条。 我文案写得特别专业:「专业证件照·尺寸精准·当天出片·¥9.9」。 然后我拿照片一测,第一版抠出来是这样的:头发周围糊了一圈背景没抠干净,边缘全是锯齿,像拿剪刀手剪的。 (图里是拿样片重跑的第一版。我自己那张更惨,边上还飘着我那个没抠干净的水杯。) 关键是,这质量是我上线后自己测出来的。没人下过单,也没人拿到过这个。 挂店 5 分钟,但你敢不敢真接单,就看你自己先敢不敢去踩一遍质量。 上线容易,交付难。
顯示更多
有了 AI 之后,我发现一个普通人把「发现真实需求 → 做出 MVP → 上线对外提供服务」这条完整链路跑通,简单到有点离谱。 我拿证件照代做当例子,亲手走了一遍。 先用需求雷达筛真实付费需求,证件照排进了 Top3。然后让 AI 从空仓库开始,一路推进做出 MVP,按公安部 GA461-2004 的尺寸标准出片(358×441px、350dpi)。最后直接挂到闲鱼,定价 9.9 元,对外提供服务。 到现在一单都没出,一分钱没赚到。倒是有 1 个人点了「想要」——竟然真的有人想要。 这个生意本身是红海,价值不在这里。而是我把整套打法跑通了一次。以后不管什么新想法,都可以用同一套流程再去验证。
顯示更多
之前 X 开源推荐算法时,我自己拆过它,研究它到底怎么影响曝光和流量。那时候大家主要琢磨怎么发帖被算法推、怎么多拿流量。 如果这次 Elon 真把整个 X 代码库无例外开源,还让第三方验证是不是线上实际运行的版本,那大家研究的就不只是拿流量了。账号分析、分发审计、第三方客户端、反垃圾机制和创作者工具,这些东西可能都会有人自己动手去造。 代码真正打开的那一天,玩法就从被推变成了自己造。
顯示更多
Once we have completed our review for security vulnerabilities, we will make the entire codebase of 𝕏 open source, with no exceptions. Moreover, we will invite third party reviewers to examine the system that is running to confirm that the open source code is what is running. Trust through total transparency is the only thing that should be believed.
顯示更多
Grok Build 开源了。我最关心的不是多一个 coding agent,而是数据流、工具行为、隔离边界终于可以审计了。官方公开的是 Rust CLI/TUI 和 agent runtime。对重度用户来说,能从源码直接看到工具背后具体在做什么,这点最重要。当然,开源不等于已经安全,还是要核对官方二进制和公开源码是否一致。
顯示更多
We've open-sourced Grok Build and have reset usage limits for all users. Open sourcing Grok Build allows anyone to support making a reliable and robust harness. Check out our code, including the Git repo for the Grok Build CLI.
顯示更多
大家都在夸 GPT 会写代码。 我更关心的是:它能不能当总控。 今天网络一抖,5 条 Codex 业务线里 4 条直接 systemError。 GPT-5.6 Sol 没装死、没重开一堆对话,而是从最后真实 checkpoint 在原对话续跑;还在跑的任务它不催,资金、发布、生产这些 hard gate 它不碰。 随后它通过桌面 GUI,把同一条公开演示派单分别送进 Claude App 和 Cursor 的指定对话。Claude 回了「CLAUDE RECEIVED」,Cursor 回了「CURSOR RECEIVED」。 真正的任务也没停在回执上:Cursor 那条生活产品探索已经续完,USB 控手机、拉起健身 App、本地运动打卡 proof 都落地了。 这是我自己跑通的跨 App 工作流,不是 OpenAI 官方原生 Claude/Cursor 集成。 但我第一次真有感觉——不是在用一个会写代码的模型,是在经营一支 AI 团队。
顯示更多
Or… what if we gave you $100 in Codex credits if you tell us what you love about GPT-5.6 Sol or why you switched? Tweet it, claim your gift, enjoy more usage. First 10k get the free tokens!
顯示更多
有了 AI 之后,我发现一个普通人把「发现真实需求 → 做出 MVP → 上线对外提供服务」这条完整链路跑通,简单到有点离谱。 我拿证件照代做当例子,亲手走了一遍。 先用需求雷达筛真实付费需求,证件照排进了 Top3。然后让 AI 从空仓库开始,一路推进做出 MVP,按公安部 GA461-2004 的尺寸标准出片(358×441px、350dpi)。最后直接挂到闲鱼,定价 9.9 元,对外提供服务。 到现在一单都没出,一分钱没赚到。倒是有 1 个人点了「想要」——竟然真的有人想要。 这个生意本身是红海,价值不在这里。而是我把整套打法跑通了一次。以后不管什么新想法,都可以用同一套流程再去验证。
顯示更多
Robinhood Chain 上第一轮 meme 普涨结束后,第一次真正的压力测试开始了。 我前几天参与过、也一直在跟这批币。我写这条时, $CASHCAT 已从约 2 亿美元市值高位回到 1.2 亿附近,24 小时跌约 31%; $VEX 跌约 21%, $4663 跌约 83%。但这不是所有币一起崩,$MARIAN 过去 24 小时仍在上涨,只是最近一小时也开始剧烈回落。 这说明资金没有简单消失,而是从普涨切进了高波动分化。对一条刚起步的链来说,现在真正要看的是:龙头大跌以后,链上还有没有足够的资金和交易者;资金会继续在不同项目间轮动,还是直接撤走;官方钱包带来的入口,能不能慢慢变成长期使用。 所以这不是抄底信号,也不是链已经死了。第一轮靠注意力,第二阶段要靠留下来的资金和用户。回撤之后还能继续长出来的东西,才决定 Robinhood Chain 是不是一波流量。
顯示更多
用 AI 帮忙写东西、研究问题或者拿主意的时候,我现在习惯问好几个模型。 但这不是随便多问几次,也不是看哪个答案听起来更多人支持。 真正的价值在于让不同模型独立判断,把它们的分歧挖出来,再回到原始证据去验证。 最近我又把这个流程走了一遍,才发现之前差点漏掉最关键的一步。我以为让三个模型都参与了,结果实际只有两个给出了回答,另一个被跳过了。如果不确认每一路是不是真的返回了内容,很容易误以为验证已经做全了。 我把能马上用的做法压成了三步。 让不同模型先独立回答。问第二个模型时,不要把第一个的答案给它看,避免互相带偏。 再要求每个模型用同一结构输出:结论是什么、主要依据是什么、哪里不确定。这样放在一起才好直接对比。 最后不去数谁对谁错,而是专门盯分歧。重要的分歧就去翻最原始的资料,或者自己实际验证一下。同时确认每个模型是不是真的给出了回答。 下次你要用 AI 交叉验证一个判断的时候,就按这三步来。
顯示更多
Redis 创造者 Salvatore Sanfilippo(antirez)最近写了一篇文章。他说在 AI 时代,人的控制重点应该从逐行代码上移,转向软件思想、整体设计、测试和产品愿景。新人还是应该亲手实现小系统来建立理解。 我最近把 Codex、Claude、Cursor 的多条业务线分别做成独立文件夹入口,共享 AGENTS.md、registry 和 HANDOFF。Claude 最近反复漂移,让我确认根因不是模型不强,而是边界、状态、验收和写回没锁住。这是我自己用下来踩到的坑。 我还是会让 AI 仔细看代码,尤其是涉及资金、生产、安全和关键路径的地方还是要下钻。真正升级是默认先审思想、边界、验收和证据,再按风险看代码。
顯示更多
把内容做成币,不会自动带来读者或真实的用户需求。 Brian Armstrong 承认内容币没有奏效。Base 今年早些时候已经转向。他说我们搞砸了,是时候翻过这一页。 Base 当前的重点是交易、支付、智能体,大部分资源投向交易,三者相互关联。 代币化可以放大已经存在的价值,但不能凭空制造出原本没有的读者和需求。把内容做成币,能放大已有需求,但造不出新需求。
顯示更多
Agree with the first part and your point on content coins. They didn't work and we pivoted early this year. We messed up, time to turn the page. I disagree about the AI agents part though. Base has been focused on trading, payments, and agents (in that order). I think all three are inextricably intertwined - for instance to do payments you need FX (trading), agents will do lots of trading and payments as well. Most of the resources are going to trading right now fwiw. Maybe it doesn't translate externally right now, but that's the case. Let me give you a call on the last parts if you're open to chatting. Would be great to learn more.
顯示更多
还真的说中了,800 万活跃用户再次重置! 只是没想到 17 小时就来了。 高速增长期用 Codex 的秘籍就是:每天尽可能烧 token,然后美美睡觉。 睡醒了大概率又能接着狠狠干。
顯示更多
Codex 和 ChatGPT Work 已经突破 700 万活跃用户了👀 为了庆祝,团队又给每个账户送了 1 次可以留着用的 Full reset,点一下就能补回本周用量。 照这个节奏大胆猜测一下 以后到 800 万、900 万用户,是不是还会继续送?😂 我看了下自己的账户,已经攒到 5 次了。 分享两个别浪费 reset 的小技巧: 1. 先看每次 reset 的到期日,优先用快过期的 2. 再看本周额度的自然重置时间。如果一两天后就会恢复,那就先等等 最划算的用法,就是周额度真的见底,手上又正好有一个长任务要跑时,再重置一次,继续狠狠干。
顯示更多
Demis Hassabis 判断 AGI 可能只剩几年,并提议由美国推动类似 FINRA 的前沿 AI 标准机构,前沿模型发布前的评测未来可能成为进入美国市场的准入门槛。 在我看来,模型竞争正进入谁定义评测、准入和发布门槛的阶段。安全标准也可能成为竞争规则,掌握定义、测试与准入的一方,可能把标准变成护城河。
顯示更多
有推友问我具体怎么用 AI 帮自己整理 Mac 或者 iPhone 上的 App 图标。这次把「具体咋弄」写全,照着做就行。 首先,跟你常用的 AI(Cursor / Codex /Claude)把目标说透:让它帮你把 Mac 桌面或者 iPhone 主屏的 App 图标,按照合理分类自动排好。 如果是整理 iPhone,先用数据线把手机连到 Mac 上。整理 Mac 桌面的话,这一步直接跳过,让 AI 在电脑上自己操作。 接着把 Unjiggle 这个工具交给 AI 去推进。 项目在 是 chungty/unjiggle。终端先执行 pip install unjiggle 装一下。 重点不是你自己去背命令,而是把工具和「帮我整理图标」这个任务一起丢给 AI,让它来 drive 执行。 具体推进的时候,你可以这样告诉它: 先让它做 backup,把当前布局备份好。 然后跑 safety-test,确认它能安全读写你的布局,没有风险。 真正整理的时候,用 suggest 模式最好,它会一步步把打算怎么分类告诉你,你确认没问题再让它改;或者先跑 go,让它先扫一遍、打个分,看看整体想怎么分。 分类主要用的是 App Store 里每个 App 自己带的类别,不是靠你手写一堆规则。所以有时候会分得不准或者分得奇怪,你自己看一眼手动调调就行。 写回去之后如果觉得哪里还是怪,别急,让它再读一次当前状态、调整一下、重新写一次,迭代一轮通常就差不多了。 自己搞不定?直接把这条推文复制丢给 Cursor 或者你常用的 AI,让它对着你现在的电脑和手机,按这个流程一步一步推进。
顯示更多
没想到 AI 居然真能帮我把 iPhone 主屏整理好。 有推友提醒,我赶紧试了试。 连上 Mac,通过 USB 读出主屏。工具统计到 304 个 App,主要按 App Store 主分类收进文件夹,最后整理成 3 页。 第一次写回后,系统把布局重新整理了一次,我还以为失败了;第二次读写才稳定。 前后对比如图。 AI 这东西,总是能在这种小地方帮我打开新世界。
顯示更多
大家一定要注意自己 Codex 重置次数的到期时间! 在哪里看: Chatgpt/Codex app 左下角用户名 → 剩余用量 → 可用重置次数 → 使用限额重置 展开后往下拉,就能看到每一次 Full reset 各自的到期日期。 不同次数的到期时间不一样,别放着放着就过期了。 一个很容易漏掉的小 tips。
顯示更多
@runes_leo reset到期日在哪看,我怎么没看到
如何一直用 Ultra,还能高性价比地推进任务? 子代理失控就降 effort,根本不是解法。 问题从来不在模型太强,而在任务范围一放就收不住、子代理编排失控。 我的 AGENTS.md 还是默认 Ultra:主线程优先,只有明确要求才派子代理,并用停止条件和尝试上限锁住范围和成本。 边界立好了,最高质量也能用得稳。
顯示更多
Codex 和 ChatGPT Work 已经突破 700 万活跃用户了👀 为了庆祝,团队又给每个账户送了 1 次可以留着用的 Full reset,点一下就能补回本周用量。 照这个节奏大胆猜测一下 以后到 800 万、900 万用户,是不是还会继续送?😂 我看了下自己的账户,已经攒到 5 次了。 分享两个别浪费 reset 的小技巧: 1. 先看每次 reset 的到期日,优先用快过期的 2. 再看本周额度的自然重置时间。如果一两天后就会恢复,那就先等等 最划算的用法,就是周额度真的见底,手上又正好有一个长任务要跑时,再重置一次,继续狠狠干。
顯示更多