@hellokaton

Building https://nitter.cf/t.co/Tiv6hh0Ur0 · AI product notes, vibe coding tips, and generative art experiments

Joined August 2014
Codex 貌似挂了?
4
6
3,467
刚才使用 codex app 一直在重试连接,看了下我网络没啥问题啊。问了下 hermes,不过官方状态也没问题啊。
1
929
Kimi K3 实测,一句话复刻动森风格游戏,美学和可玩性都很好,太惊艳了🤩 我是直接粘贴了 X 别人的视频推文,启用多模态分析自己完成,买的是 199/月 档位的订阅,消耗还是挺快的: - 周消耗 5%,4小时消耗 27% 这个任务耗时情况: - 游戏从零到可玩(含分析视频、写代码、浏览器实测):11:10–11:55 ≈ 45 分钟 - 二刷视频 + 细节打磨:≈ 5 分钟
我把「用 AI 一个晚上复刻一款爆火小游戏」打包成了开源 Skill 🐾 对你的 Agent 说一句"复刻 Water Sort",它会: → 调研游戏规则,和你确认范围 → 先写求解器再写生成器,每关可解性是证明出来的 → 三消的目标分由机器人构建期实跑标定 → 自己开浏览器玩通一局才算验收 产出是成品,不是 demo。仓库里有两款完整案例,全部开源: github.com/hellokaton/game-c…
2
6
49
9,984
我给它的提示词: """ nitter.cf/intheworldofai/status/… 复刻一下这个游戏,做一个中文版的,不要使用 skill """ - 源码:github.com/hellokaton/cozy-i… - 在线体验:cozy-isle.hellokaton8143.wor…
Kimi K3 is insane🤯 It built an Animal Crossing-style game, and it generated a fully playable experience with the cozy aesthetic, interactions, and gameplay loop in a single shot. Open-weight models are starting to rival the best closed models for game generation. The pace of progress is getting ridiculous. 🌱🎮
6
982
我把「用 AI 一个晚上复刻一款爆火小游戏」打包成了开源 Skill 🐾 对你的 Agent 说一句"复刻 Water Sort",它会: → 调研游戏规则,和你确认范围 → 先写求解器再写生成器,每关可解性是证明出来的 → 三消的目标分由机器人构建期实跑标定 → 自己开浏览器玩通一局才算验收 产出是成品,不是 demo。仓库里有两款完整案例,全部开源: github.com/hellokaton/game-c…
2
2
1
15
14,216
katon retweeted
我把「用 AI 一个晚上复刻一款爆火小游戏」打包成了开源 Skill 🐾 对你的 Agent 说一句"复刻 Water Sort",它会: → 调研游戏规则,和你确认范围 → 先写求解器再写生成器,每关可解性是证明出来的 → 三消的目标分由机器人构建期实跑标定 → 自己开浏览器玩通一局才算验收 产出是成品,不是 demo。仓库里有两款完整案例,全部开源: github.com/hellokaton/game-c…
2
2
1
15
14,216
katon retweeted
MXGA(Make X Great Again)上线 48h,已经识别出 6K+ Spam 账号,AI Agents 还在继续处理剩下的数据。 这是一个开源、免费、AI 驱动的 Chrome 插件,用来清理 X 评论区里的 spam bot。 v0.3.0 已经上架 Chrome 商店,新增「自动拉黑」设置,默认关闭,可自行开启。 实测随着数据量上来,识别效果和体验都好了不少,我自己的信息流已经越刷越干净了 😆 欢迎大家试用。 🔗 Chrome 插件:chromewebstore.google.com/de… 🌐 官网:x.zuoluo.tv
🤖 Made with AI
43
31
9
237
37,839
公众号这个广告植入也太丝滑了。。。 智商测不测另说,Hermes 先跑起来。
1
9
2,007
OpenAI 正在精简 ChatGPT 模型:GPT-4o、GPT-4.1 和 o4-mini 将于 2 月13 日从 ChatGPT 中下线。 官方解释是每天使用 GPT-4o 的用户只有 0.1%。不过这些模型会继续在 API 提供,供开发者使用。 openai.com/index/retiring-gp…
8
2,607
vibe coding 很酷,效率很爽,也成为了一种流行。 内行看门道,外行看热闹,我有几个标准区分 vibe 的基准吧: 1. 用户系统:能不能留下人和数据 2. 支付:敢不敢收钱、怎么交付、怎么负责 3. 营销:好的产品去看他的营销策略都是步步为营的 4. UI/UX:细节决定能不能长期用 据我观察这些做的好的要么是有 engineer 背景,要么就是在这块摸索的非常深入的科技行业从业者(比如立青和刘小排),而他们几乎不太过度吹嘘,更务实和实事求是,作品说话。 有人说在这个时代追求做好产品并不重要,应该抓住用户注意力,一会 skills 一会 clawdbot,从历史数据上看今天也没人再提及 gpt-3.5-turbo 了,90% 当下的流行都会变成垃圾,媒体只关心抓住你当下的注意力获得流量,仅此而已,连发文的人自己都不会去使用的。 看你要什么?做自媒体会去夸大任何一条新消息,让 AI 包装成病毒传播的帖文宣扬这是「价值」然后获得流量,继续下一个消息,如此重复。盈利产品更多的要去实践、去发现机会、去自我化,即便是个垃圾一旦完成了 PMF 就迅速投钱投资源占市场,在概率中坚持就可能中奖。而实打实做产品显然是一个更困难但成功收益率可能爆炸的事,比如玉伯团队做的事,橘子老师的团队我也看好,他们被称为创业者。一点感悟
看到一些网友介绍产品陷入了一个奇怪的误区,他们不介绍产品,却重点介绍了:用 AI vibe code 的、开发效率如何、用的什么技术栈。有的甚至产品本身完全不描述,也没截图,就丢个链接。向潜在受众介绍产品,为何完全不点题?
1
15
4,348
katon retweeted
腾讯刚刚发布了 HunyuanImage 3.0-Instruct,基于混元MOE结构的新一代视觉语言旗舰大模型。 我刚刚用官方的示例测了一下,中文有点拉啊😅
Today, we introduce HunyuanImage 3.0-Instruct, a native multimodal model focusing on image-editing by integrating visual understanding with precise image synthesis! 🚀 It understands input images and reasons before generating images. Built on an 80B-parameter MoE architecture (13B activated), it natively unifies deep multimodal comprehension and high-fidelity generation. 🧠 A "Thinking" Model with Native CoT & MixGRPO: The model doesn’t just execute commands, it processes them through a Native Chain-of-Thought (CoT) schema. Enhanced by our self-developed MixGRPO algorithm, it reasons through complex instructions to achieve flawless intent alignment and human-preference consistency. 🎨 Precise Editing & Multi-Image Fusion: The model enables accurate image editing by adding, removing, or modifying elements while keeping non-target areas perfectly intact. It also excels at seamless multi-image fusion, synthesizing complex scenes by extracting and blending elements from multiple sources into a unified, consistent output. 🏆 SOTA Performance: HunyuanImage 3.0-Instruct sets a new benchmark in visual quality and alignment, delivering performance that matches leading proprietary models. We aim to enable the community to explore new ideas with a state-of-the-art foundation model, fostering a dynamic and vibrant image generation ecosystem. 🛠️🎨 💻Try it at (PC only): hunyuan.tencent.com/chat/Hun…
1
11
3,133
腾讯刚刚发布了 HunyuanImage 3.0-Instruct,基于混元MOE结构的新一代视觉语言旗舰大模型。 我刚刚用官方的示例测了一下,中文有点拉啊😅
Today, we introduce HunyuanImage 3.0-Instruct, a native multimodal model focusing on image-editing by integrating visual understanding with precise image synthesis! 🚀 It understands input images and reasons before generating images. Built on an 80B-parameter MoE architecture (13B activated), it natively unifies deep multimodal comprehension and high-fidelity generation. 🧠 A "Thinking" Model with Native CoT & MixGRPO: The model doesn’t just execute commands, it processes them through a Native Chain-of-Thought (CoT) schema. Enhanced by our self-developed MixGRPO algorithm, it reasons through complex instructions to achieve flawless intent alignment and human-preference consistency. 🎨 Precise Editing & Multi-Image Fusion: The model enables accurate image editing by adding, removing, or modifying elements while keeping non-target areas perfectly intact. It also excels at seamless multi-image fusion, synthesizing complex scenes by extracting and blending elements from multiple sources into a unified, consistent output. 🏆 SOTA Performance: HunyuanImage 3.0-Instruct sets a new benchmark in visual quality and alignment, delivering performance that matches leading proprietary models. We aim to enable the community to explore new ideas with a state-of-the-art foundation model, fostering a dynamic and vibrant image generation ecosystem. 🛠️🎨 💻Try it at (PC only): hunyuan.tencent.com/chat/Hun…
1
11
3,133
katon retweeted
我开发的 Agent Coworker 应用正式 Release 啦。 他拥有以下功能: Amon 可以根据你发送的消息,进行思考,执行工具调用,完成你的任务。 Amon 拥有计划模式,可以先创建任务 TODO,然后按计划执行。 可以自定义添加多个 API 供应商。 可以 Claude Code 模式,完全当 CC 的可视化客户端。 Amon 支持 Agent Skills,你可以通过安装 Skills 为 Amon 添加专业能力。 大家感兴趣的可以在 GitHub 下载体验,点个 Star。 liruifengv.com/posts/amon-an…
1
6
1
29
9,643
如果你不擅长做计划、写一个 LLM 理解的好的 prd,玩 Ralph循环纯属给Anthropic 送钱做慈善。 直到你 7x24 小时跑任务烧掉数以亿计的 token 后,你就会调整了。 不知道为什么的可以看看下面这个访谈中提到的技巧👇
Ralph loop is extremely unhealthy. I now run 1-2 loops 24/7, tweaking, iterating. Before sleep I set off a loop, I wake up 3-5x a night thinking of it with excitement. My workflow is getting better, but it's a steep learning curve. I'll share it all when I'm a bit further, there are some things I still don't like. The best thing? I am doing projects I never had time to do and I can't stop. I am tired too 😅 But having something code while you sleep is pretty incredible.
1
4
1,850
vibe coding 开发自己使用的软件非常容易,开发提供给其他人使用的软件如果没有基础,大概率是 AI 垃圾。
83%认同
17%我有不同的意见
30 votes • Final results
2
2
6
1,925
不想在合拍里直出真人脸,但又不想整张图都变画风? 设计了一个提示词,把人物改成卡通插画风,环境保持真实不动。 特别适合:coffee chat、多人合拍、线下合照但想保护隐私的场景。 提示词: 把照片中所有人物转换成Q版萌系卡通插画:大头小身、圆润比例,厚实干净的黑色描边;五官极简(点状眼睛、小嘴、鼻子弱化),脸颊打淡淡的腮红;纯色块上色,几乎无渐变与阴影,整体可爱治愈。 除人物外一律不变:背景、环境、光线、透视、构图、颜色、清晰度全部保持原照片。 人物的姿势、表情、衣服颜色与细节、人数与相对位置完全一致,边缘清晰不糊。
6
59
1
271
21,400
直接在 Artifox 中一键使用 artifox.app/discover/nano-ba…
1
1
6
1,588