@gpuhell

攻城狮/业余投机/右侧交易/游戏开发/Haskell/Rust/C++/Unity3D/C#/对java有偏见/RL/智商欠费/浅尝辄止故平庸/反乌托邦/竹林中/地狱变/手撕菠萝蜜/胸口碎榴莲/单机推特中/乐视一生黑(乐视已阵亡)/华为一生黑(迟早会阵亡 )/中华跪族/器材党/预防式B台支

上海, 中华人民共和国
Joined June 2012
我觉得国内应该全面禁止AR。对于一群做着中国梦的人来说,“增强现实”是致命的 (手动滑稽
5
1
12
Muse is basically a data harvester that goes out of its way to collect your personal information, while offering very little practical value. Straight to the trash.
3
80
my timeline after opus 5.5 release
1
2
131
日常焦虑帝 retweeted
🔥K3.1已确认泄露!! Coming Soon~
37
10
4
231
22,399
Only Google has truly responded to the call of "Pace the Frontier." 😆
只有谷歌真正响应了ai发展减速的号召。
1
1
4
279
Qwen 4 is coming soon. Qwen 4.5 and Qwen 5 are targeting a parameter size of 4–10T.
14
821
The previously missing data from MiMo’s RL training has now been added.
MIMO’s RL training has stopped at step 30. We can observe the following: The Pro model achieved a DeepSWE score of 72.57, but this score was reported at step 28; data for steps 29 and 30 are missing. The number of active environments for Pro began increasing at step 23, then dropped rapidly after step 26. The duration of each training step also rose sharply.
5
624
MIMO’s RL training has stopped at step 30. We can observe the following: The Pro model achieved a DeepSWE score of 72.57, but this score was reported at step 28; data for steps 29 and 30 are missing. The number of active environments for Pro began increasing at step 23, then dropped rapidly after step 26. The duration of each training step also rose sharply.
2
1
5
989
日常焦虑帝 retweeted
Looks like the MiMo RL runs have completed. The pro model went from 58.41 to 72.57 on DeepSWE. For context, the highest score on DeepSWE is 74 by Astra, Gemini 3.8 Flash and Opus 5
27
21
12
641
113,681
kimi 3.1 👀
🚨 K3.1 要来了 今天早上 @Kimi_Moonshot 官方在 @ZhihuFrontier 突然发了一串完全没解释的数字: 415926535897932384626433832795... 一开始看很莫名其妙,但答案其实就在圆周率里。 π 是: 3.1415926535897932384626433832795... 把最前面的 3.1 删掉: 415926535897932384626433832795... 刚好就是 Kimi 发出来的这一串。 也就是说,这个谜语很可能只有四个字: Kimi K3.1。 目前 Kimi 官方还没正式公布 K3.1,官网和文档里的旗舰模型仍然是 K3。 但大半夜专门发一个「少了 3.1 的 π」,这暗示已经有点明显了😂
5
375
日常焦虑帝 retweeted
So now we know that before Kirin9050, there only has one N+3 chip: Kirin 9030. (Kirin9030/ Kirin9030 Pro logic die is the same) It seems that SMIC is still struggling on N+3 process even under small die size. Kirin9030 chip is roughly ~5M shipping so far (Mate 80 Pro/ ProMax/ RS and Pura X Max). For Kirin9030 die size of 136mm2, one wafer can divide into ~400 chip. So even we assume N+3 yield is as poor as 50%… 5M chip only need 25K wafer in total… Even we estimate that there should be some stockpile, like extra 3M chip so 8M chip in total, still translate to 40K wafer only. It seems that Huawei preserve most of its N+3 volume to Mate 90 series. Mate90 basic/ Pro will use Kirin9030 chip, only ProMax and RS use Kirin9050. We can expect that Kirin9050 volume constrained will be more severe than Kirin9030 on Mate80 (Mate80 Pro use Kirin9030 but Kirin9050 is only for ProMax/RS). And if we give N+3 process 10K wpm (actually very low) and 1 year manufacturing time, there should be 120K wafer. So ~80K wafer available for Mate90. 10M Kirin9030 chip for Mate90 only need 50K wafer, and 30K wafer can divide into~3.75M Kirin9050 chip under 25% yield (considering the combination of package yield, and Kirin9050 die size is ~114mm2). Actually, I do not think Huawei now have 10M Kirin9030 chip/ 3.75M Kirin9050 chip for Mate80, so my calculations assumption condition should be lower- either the capacity (10k wpm) or the yield (50% for Kirin9030 and 25% for Kirin9050). No matter which one, it means SMIC still need to take a long road to increase the yield or output of the N+3 process even on small die. I do not expect that we can see the large ASIC/HPC die size chip on N+3 for soon. Ascend series chip will use logicfolding by 2030, perhaps we can also see the large AI chip with N+3 process by that time? Remember, steady and large amount of N+2 Ascend chip shipping is occurring on Ascend 950 with 0.5x mask size, three years after first SoC chip Kirin9000s was shipped. As for more challenging N+3, it may take longer time.
Shocking, Kirin 9030S is N+2, not N+3…. ————————— Mobile: Pura 90 Pro Max Die size: 122 mm2 Marker: Hi36D0G-GPCV100 CPP:63nm Cell height: 252nm SRAM cell area: 0.0315 —————————— Btw, Kirin 9050 Pro decap is on-going, and it is a chip using N+3 bond with N+2. Die size is even smaller than Kirin 9030S, hence far smaller than Kirin 9030.
3
3
26
4,426
日常焦虑帝 retweeted
Shocking, Kirin 9030S is N+2, not N+3…. ————————— Mobile: Pura 90 Pro Max Die size: 122 mm2 Marker: Hi36D0G-GPCV100 CPP:63nm Cell height: 252nm SRAM cell area: 0.0315 —————————— Btw, Kirin 9050 Pro decap is on-going, and it is a chip using N+3 bond with N+2. Die size is even smaller than Kirin 9030S, hence far smaller than Kirin 9030.
1
2
2
43
19,642
“Chinese researchers need to increase their development speed so that, like their American counterparts, they can see the dangers of AI development,” Huawei rotating chairman Xu Zhijun said at a press conference on Thursday, in a brilliant response to Dario Amodei’s “We Must Pace the Frontier.”
3
257
It also surprised me. The 950DT’s production hasn’t ramped up this year, yet the 960DT—originally slated for Q3 next year—has been moved up to Q1. The only reasonable explanation I can think of is that LogicFolding is working very well, allowing the memory chips originally allocated for the 950DT to be shifted over to the 960DT.
The more surprising fact is that they are offering 288GB with Ascend 960DT in Q1 2027 I wonder where the supply is coming from 🤔🤔
5
2
66
16,096
I’m trying to check this model’s fingerprint, but it isn’t outputting any tokens right now. It's even slower than moonshot.ai
🥷 New stealth model: Union Alpha (@unionalphaai) A multimodal model for research, coding, and agentic workflows. - Free to use - 256K context - Tool calling - Frontier-level general-purpose performance Try it now and share your feedback: openrouter.ai/stealth/union-…
1
217
日常焦虑帝 retweeted
Pace the frontier
238
692
147
15,872
1,400,555
the whale bros are saving the world.
Replying to @teortaxesTex
At least DeepSeek researchers will never be able to work with the safety advocates in the Anthropic camp
3
513
日常焦虑帝 retweeted
Open source must win; humanity has no other option.
51
138
11
1,416
20,599
日常焦虑帝 retweeted
Hi Dario and Sam 👋 If you want to slow down your AI development, feel free to do it. Anyone preventing you? Slow down, stop, or even shut down your AI. Why do you need the government to intervene?* *Answer: To stop your emerging competitors.
190
729
45
5,279
150,047
Kimi K2.8 Preview offers a 1M context window, and its performance is officially claimed to be close to K3. kimi.com/code/docs/en/kimi-c…
11
1,058