@kleon_aii
iAccount based inSingapore
About this account
- Account based in
- Singapore
- Connected via
- United States App Store
Account-level information from X, not a live location or the device used for a specific post.
🤖 TikTok算法工程师,运营AI Agent公司 ⚙️ 全栈: Agent系统 · 数据工程 · AI infra 📖 分享: AI工具实测 | AI创业 | 算法拆解 | agent开发 | 变现路径 💵 Building https://nitter.cf/t.co/TEnvdgtQuc
Singapore
Joined July 2019
- Tweets2.7K
- Following1K
- Followers248
- Likes2.9K
Introducing Talea!
Your books, taken care of.
Meet Talea: AI-powered bookkeeping for founders. Hand over receipts, invoices and bank statements. Keep the evidence close, and answer the questions that need your input.
talea.maestro.onl
悲报,额度砍一半,立刻滚去claude,opus 5.5还是太耐用了
Hi,
Tomorrow we are re-opening the Pro $200 subscriptions to new subscribers, but together with it we are also changing how we calculate the usage for it. In effect, if you do the math, it will net out at half the dollar in API spend compared to the old Pro $200 plan.
Now that it's said, let me explain why this is happening and why you will still get more work done than if you were on the Pro $200 subscription one month ago.
(a) We didn't want to compromise in other ways and are committing to not reintroducing the 5h limit, so that you can fully use the weekly usage when you want.
(b) On the subscription, we guarantee that over time you always get more work done and with an increasing level of quality. This means that you will continue to get more value per dollar spent as a result of models getting more efficient and us passing down the improvements in the form of API price reductions.
(c) We don't want to put an incentive on ourselves to artificially inflate the API list prices to make it look like you are getting a lot (and workaround it through discounts, etc). Instead we want to continue to both rapidly reduce prices and increase capabilities of models on the API. This week we introduced GPT-6 Sol and GPT-6 Luna at 50% of their previous price. Over time, we see prices go low enough that it makes sense for most to buy usage as needed without there being a significant gap between what you get in a subscription and what you get in the API for a dollar spent.
(d) Tomorrow, we are adding more things to the subscription that won't draw on the usage, I won't reveal what that is yet.
I wanted to be transparent before all the big announcements tomorrow. Lots of new exciting things are coming to the subscriptions that will make it super compelling, but I wanted to make sure to share this change ahead of time so you can all understand it before we shower you with good news.
Codexingly,
Tibo
Kleon retweeted
真的要卸载剪映了。
YC S24 团队做的 Palmier Pro,开源后已经冲过 1 万 GitHub stars。
我看到它的第一反应:AI 已经进时间线里干活了。
把 Codex、Claude Code 或 Cursor 通过 MCP 接上,你可以直接说:
“把这段剪短一点。”
“重新排一下镜头。”
“生成一段 B-roll,放进时间线的空位。”
AI 会读取当前项目,自己调整轨道、补充素材、修改剪辑点。
画面不满意,也能在时间线上继续调用 Seedance、可灵、Nano Banana Pro 生成,不用在网页、下载文件和剪辑软件之间反复搬素材。
它用 Swift 原生开发,目标对标 Premiere Pro。编辑器和 MCP 免费开源,不过目前只支持 Apple Silicon + macOS 26,生成式 AI 功能需要订阅。
对内容创作者来说,这个变化很具体:以前你告诉 AI 应该怎么剪,现在它可以先把第一版铺到轨道上,你负责判断和修改。
X怎么检测bot账号?2023年开源的算法代码里能看到几个关键机制:
TweepCred:每天跑一次batch job,在全用户follower图上算PageRank。粉丝质量、账号年龄、设备指纹都是输入。这个分决定了行为容忍度,分高的账号发100条没事,分低的发30条就触发。
Agatha/Smyte:每条推发布后分钟级异步打分。最容易暴露的特征:固定发帖间隔。真人发推是burst+静默,机器是metronome,每隔90秒一条连续几小时。
DownrankSpamReply:账号级标签。打上之后你的所有reply被折叠到"Show more replies"后面。你自己看一切正常,别人看不到你。这就是ghost ban。
检测是叠加的,单一指标不触发。timing pattern + 内容指纹 + 发帖量 + fan-out模式(一天reply几十个不同的人但没有一个回头对话)合在一起过阈值才标记。
被标记后怎么恢复?batch job每天重评一次,行为数据正常了标签可以被覆盖。行为窗口多长没开源,7天还是30天不知道。
代码在 github.com/twitter/the-algor…,visibilitylib目录下。
现在越来越多人同时订着Claude Max $200、Codex Pro $100、Cursor $20。一个月$320的AI coding账单。
三个工具的分工已经很清楚了:
Cursor改UI和前端,写完马上能预览,边看边调。
Codex开个task扔后台,去喝杯咖啡回来看结果。不需要盯着。
CC做整个module的重构,跨几十个文件改接口那种重活。需要它理解整个codebase的结构。
FYP上有人说Opus 4.7比GPT-5.5差太多,也有人说Composer 2.5用了8小时觉得最舒服。结论都不一样,因为干的活不一样。
2024年大家在争到底用哪一个。2026年大家在争三个怎么配合。
DeepSeek V4 Pro把75折从促销变成永久了。
看看现在的定价差距:
DeepSeek V4 Pro输出:$0.87/百万token
Claude Opus 4.7输出:$25/百万token
GPT-5.5输出:$30/百万token
同样跑100万token的输出,DeepSeek花不到1块钱,Claude要25块,GPT要30块。差了30倍。
能做到这个价的原因也很直接:750K芯片集群 + 140万亿token的训练数据 + 1.6万亿参数MoE架构只激活490亿。规模够大,单位成本就能压死。
coding benchmark也没拉胯,LiveCodeBench 93.5,SWE-Verified 80.6,Codeforces 3206。
OpenAI和Anthropic的API收入是烧钱的主要回血渠道。如果agentic场景的开发者开始大规模切DeepSeek,这个回血速度就跟不上了。
Cursor内部现在用得最多的skill叫 /thermo-nuclear-code-quality-review
干的事很简单粗暴:
- 复杂度太高的代码直接删掉重写,不是移到别处藏起来
- 超过1000行的文件直接block
- 薄wrapper和泄漏的逻辑直接flag
- PR能跑通但让代码变乱的,reject
4800+ likes,Cursor工程师自己在用。
有意思的是这个方向:code review的标准从"能不能跑"变成了"代码有没有变乱"。AI写代码速度太快之后,质量门禁反而要比以前更严。不然一周就能把整个codebase写成屎山。
Karpathy观察LLM写代码的常见毛病,整理成一份CLAUDE.md,65行,GitHub上143K星了。
4条规则:
1. 不确定就问,不要自己猜。多个理解方式存在的时候把选项列出来,别悄悄选一个。
2. 最简代码。没人要求的功能不写,单次使用的代码不做抽象,"以后可能用到"的灵活性不加。200行能用50行写的,重写。
3. 只改该改的。不要顺手"改进"旁边的代码、注释、格式。你改出来的孤儿代码要删,别人留下的不动。每一行diff都要能追溯到用户的request。
4. 先定义"做完了"长什么样。"加个校验"→"写invalid input测试然后跑通"。"修bug"→"写一个能复现的测试然后修掉"。
规则本身不复杂,复杂的是LLM总忍不住做多余的事。过度设计、顺手重构、不问就猜,这些毛病用过CC的都见过。
github.com/multica-ai/andrej…
isaiprofitable.com 追踪了主要AI公司的累计投入和收入。
整理完发现,整个AI产业链目前只有NVIDIA明确赚钱。累计GPU+数据中心收入$4780亿,净利$729亿。
其他所有人的年烧钱速度:
OpenAI $140亿
Anthropic $87亿
Meta $1420亿
xAI $120亿
四大云厂商2026年的AI capex各投$1900-2000亿,花钱速度远超AI收入增速。
模型公司在烧,应用公司在烧,云厂商在烧。整个行业的利润集中在一家卖GPU的公司。
加州淘金热翻版:挖矿的都在亏,卖铲子的赚麻了。
HN上一个开源项目解决了AI coding agent的一个实际问题:多个agent同时改同一个代码库怎么办。
Kanbots把Kanban看板跟CC/Codex结合了。每张卡片派一个agent执行,每个agent跑在独立的git worktree里,各改各的分支不冲突,最后merge。
autopilot模式更有意思:把任务丢进看板,设定personas(前端/后端/测试),自动分配,并行跑,跑完自检。你睡觉去,醒来看结果。
CC同时开多个session改同一个repo很容易文件冲突。Kanbots在中间加了一层worktree隔离,把这个编排问题解决了。
MIT开源,HN 173分。
Anthropic的Project Glasswing发了第一个月成绩单。
Claude Mythos Preview被丢进50家合作伙伴的代码库里找安全漏洞。一个月结果:
Cloudflare:2000个bug,400个高危/严重,误报率比人类测试员还低
Mozilla:Firefox 150找到271个漏洞,同流程上一版用Opus 4.6只找到不到1/10
全体合计:超过10000个高危或严重漏洞
大部分合作伙伴的bug发现速度提高了10倍以上。
这带来一个新问题:安全领域的瓶颈从"找漏洞"变成了"修漏洞"。Palo Alto最新版本补丁数量是平时的5倍,微软说补丁数量会"持续增长一段时间"。
另外一个细节:某家合作银行用Mythos Preview发现并阻止了一笔150万美金的欺诈电汇。从找代码bug到找金融犯罪,同一个模型。
Anthropic花3亿多美金收购了Stainless。
Stainless做什么的?给API公司自动生成SDK。OpenAI的Python SDK、Google的SDK、Cloudflare的SDK,全是Stainless帮着生成和维护的。连Anthropic自己的Claude SDK也是。
收购完成后Anthropic直接停掉了Stainless的对外托管服务。OpenAI和Google以后要自己维护SDK了。
这个操作有多狠:花3亿买下了给竞争对手造武器的军火商,买完就停止对外供货。现有SDK还能用,API更新后的SDK适配得自己想办法。
时间点也值得注意。Anthropic正在推MCP(Model Context Protocol),Stainless除了SDK还能生成MCP server。收购后MCP的工具链生成能力等于被Anthropic包了。
3亿美金买一个SDK生成器?看完客户名单就懂了。
Codex昨晚的更新值得仔细看,几个功能加起来意味着什么比单独看更有意思。
Appshots:同时按两个Command键,直接截屏当前窗口喂给Codex当上下文。不用再打字描述"我现在看到的页面是...",截图丢进去就行。
锁屏使用:Mac锁屏之后Codex继续干活。用的是Apple官方的Authorization Plugin。拿手机远程派任务,电脑锁着它自己跑。这个功能等于把Codex从"你坐在电脑前才能用"变成了"随时能用"。
内置浏览器+标注:Codex自己能开浏览器看网页,还能在页面上做高亮标注。
三个功能放一块想想:能看你屏幕,能锁屏干活,能自己上网。这已经不是一个coding tool了。
OpenAI自己的工程师已经全面在用。内部发了一份PDF讲各团队(安全、infra、前端、API)怎么用Codex:搞懂陌生代码库、跨几十个文件的重构、生成edge case测试。连自家团队都在拿它当主力开发工具。
另外一个实用tip:Codex app没有Apple官方签名,Mac的Gatekeeper会持续扫描它导致特别卡。暂时没法关Gatekeeper,等OpenAI补上签名吧。