攻城狮/业余投机/右侧交易/游戏开发/Haskell/Rust/C++/Unity3D/C#/对java有偏见/RL/智商欠费/浅尝辄止故平庸/反乌托邦/竹林中/地狱变/手撕菠萝蜜/胸口碎榴莲/单机推特中/乐视一生黑(乐视已阵亡)/华为一生黑(迟早会阵亡 )/中华跪族/器材党/预防式B台支

上海, 中华人民共和国
Joined June 2012
日常焦虑帝
@gpuhell
18 Oct 2017
我觉得国内应该全面禁止AR。对于一群做着中国梦的人来说,“增强现实”是致命的 (手动滑稽
5
1
12
日常焦虑帝
@gpuhell
10h
CXMT (688825.SH) announced plans to invest RMB 24.1bn in a new technology R&D project, including RMB 13bn in excess IPO proceeds. It also plans to invest RMB 10.8bn in Phase II of its memory wafer back-end testing base, including RMB 5bn in excess IPO proceeds. The new facility will provide DRAM chip testing and module assembly. Total planned investment: RMB 34.9bn (~US$4.9bn). #CXMT #DRAM
2
124
日常焦虑帝
@gpuhell
16h
Muse is basically a data harvester that goes out of its way to collect your personal information, while offering very little practical value. Straight to the trash.
1
3
172
还有在中国用 Claude,一次都没被封过的人吗? 好奇你们都用了多久,平时怎么用的。有没有长期稳定用到现在的,出来给我一点信心
230
5
11
147
88,925
日常焦虑帝
@gpuhell
Sep 27
OAI 和 A\ 都没被封,都两网站刚上线时注册的老号。
1
5
2,653
用 Gemini spark 开云浏览器注册 muse 也失败了
19
1,510
日常焦虑帝
@gpuhell
Sep 27
今天打开 APP 直接注册就好了啊,哪来这么多戏
1
1
90
感觉最近 X 上的视频变卡了不少,经常播到一半卡住,只有我这样吗?ytb 都不卡。
27
27
17,799
图片加载也变慢了,经常要loading至少五秒以上才出来 – at Taipei City, Taiwan
3
1
1,740
日常焦虑帝
@gpuhell
Sep 25
你在台湾也这样吗
1
1
40
WquGuru
@wquguru
Sep 24
Opus 5.5 制作了一支音乐 MV —— 《对齐失败》 主演:达主席 配角:@wquguru 转发过 100 我就出一篇《从 0 到 1 精通 Opus 5.5 视频制作》😆
WquGuru
@wquguru
Sep 24
3天体验下来,关于 Opus 5.5 的一些感受: 1. 省,非常省!以往这个点,周用量基本到了40%,今天这才到 19% 2. 3D和2D理解能力都非常棒,和前一代相比有跨越式发展 3. 自主能力超强,更关键的是,往往都还挺一发即中的,比如我有一个 Skill,Opus 5.5 今天在做完某一个任务后直接用它去校验结果了,这种场景以前从来没遇到过 彩蛋:用 Opus 5.5 做了一个小视频《对齐失败》,敬请期待👇
12
1
3
32
10,997
日常焦虑帝
@gpuhell
Sep 24
还怪好听的 👍
1
1
108
日常焦虑帝
@gpuhell
Sep 24
my timeline after opus 5.5 release
1
2
134
日常焦虑帝 retweeted
🔥K3.1已确认泄露!! Coming Soon~
37
10
4
231
22,441
Anthropic 第一次把 Fable 系列的「前沿模型研发」限制放进 Opus。 Opus 5.5 加入了类似 Fable 的分类器。遇到少量前沿 LLM 研发任务,例如为特定 AI 加速器开发 kernel,系统会直接从 Opus 5.5 切到能力更弱的 Opus 5。Opus 5 没有这项限制。 Fable 5 此前已经会限制分布式训练基础设施、AI 加速器设计和部分 kernel 开发。Anthropic 当时解释,强模型已经可能加速下一代模型研发,它不希望 Claude 帮竞争者更快造出同级别的前沿模型。Fable 当前也会对一小部分这类任务进行拦截或切换模型。 Anthropic 自家的评测也开着这套护栏。前沿 LLM 研发题目一旦触发限制,就改由 Opus 5 作答,Opus 5.5 的相关 benchmark 成绩也可能因此被拉低。 Anthropic 强调,这类分类器只覆盖很小一部分前沿 LLM 研发任务,绝大多数普通 AI、机器学习研究和日常编程不会受到影响。触发降级后,Claude 会明确显示已经切到 Opus 5,后续对话也会继续使用 Opus 5。
Claude Opus 5.5 will be the first Opus meant to fall back to a less capable model for "a small set of capabilities related to the development of frontier LLMs, such as kernel development.."
5
13
5,442
日常焦虑帝
@gpuhell
Sep 23
用意很明显了。。。🤔
1
1
137
日常焦虑帝
@gpuhell
Sep 23
Only Google has truly responded to the call of "Pace the Frontier." 😆
只有谷歌真正响应了ai发展减速的号召。
1
1
4
282
日常焦虑帝
@gpuhell
Sep 22
Replying to @Kimi_Moonshot
工具多解决不了用量抠门的问题
73
我告诉Astra,我的Foxmail是Windows程序,2006年停用,登录密码我怎么也记不得了,你研究下,而且做一个Linux版本的邮件阅读器,因为我不用Windows了,你把邮件解密后都导过去。然后半个小时后,结果如图。我太幸福了吧?我看到了之前我以为再也找不回来的情书。
53
15
2
302
94,047
日常焦虑帝
@gpuhell
Sep 22
很容易因此收到 cybersecurity abuse 的警告邮件。
1
1,769
山东财经大学は中級の大学だけど、中国で一番有名な数学の先生がいらっしゃるようですね
現地の学生さんに大学を見てみないかとお誘いいただき、滅多にない機会なので行ってみることにしました! 学食の15.9元 380円の米线がうまい 後で講義を受けに行きます
20
3
220
30,974
日常焦虑帝
@gpuhell
Sep 22
宋浩的讲的段子比知识多...
1
606
日常焦虑帝
@gpuhell
Sep 22
Qwen 4 is coming soon. Qwen 4.5 and Qwen 5 are targeting a parameter size of 4–10T.
14
823
日常焦虑帝
@gpuhell
Sep 21
MIMO’s RL training has stopped at step 30. We can observe the following: The Pro model achieved a DeepSWE score of 72.57, but this score was reported at step 28; data for steps 29 and 30 are missing. The number of active environments for Pro began increasing at step 23, then dropped rapidly after step 26. The duration of each training step also rose sharply.
2
1
5
989
Do you have any idea, why did they stop at 30? I think both the models still have the capacity to climb up. What do you think?
1
29
日常焦虑帝
@gpuhell
Sep 22
they decided to release the checkpoint 😀 nitter.cf/XiaomiMiMo/status/2102…
Introducing Xiaomi MiMo-V2.6 — Pro & Flash. Frontier intelligence, all the modalities, built in public. 🔹 Two omnimodal models, advancing through scaled reinforcement learning 🔹 Pro performs on par with Claude Opus 5 and GPT-5.6 Sol across most agent benchmarks 🔹 Pro scores 46 on the Artificial Analysis Intelligence Index — the highest among open-source models 🔹 Stronger coding, computer use, 3D reasoning and creative capabilities 🔹 Open model weights, technical report, RL environments and training code Blog:mimo.xiaomi.com/mimo-v2-6
20
日常焦虑帝
@gpuhell
Sep 21
The previously missing data from MiMo’s RL training has now been added.
日常焦虑帝
@gpuhell
Sep 21
MIMO’s RL training has stopped at step 30. We can observe the following: The Pro model achieved a DeepSWE score of 72.57, but this score was reported at step 28; data for steps 29 and 30 are missing. The number of active environments for Pro began increasing at step 23, then dropped rapidly after step 26. The duration of each training step also rose sharply.
5
624
日常焦虑帝
@gpuhell
Sep 21
76
Looks like the MiMo RL runs have completed. The pro model went from 58.41 to 72.57 on DeepSWE. For context, the highest score on DeepSWE is 74 by Astra, Gemini 3.8 Flash and Opus 5
27
21
12
641
113,734
日常焦虑帝
@gpuhell
Sep 21
mimo-v2.6-pro stopped at step 30. but the DeepSWE score 72.57 is from step 28.
4
488
日常焦虑帝 retweeted
Looks like the MiMo RL runs have completed. The pro model went from 58.41 to 72.57 on DeepSWE. For context, the highest score on DeepSWE is 74 by Astra, Gemini 3.8 Flash and Opus 5
27
21
12
641
113,734