@PaXError

十年测试开发工程师 | 有说实话的坏习惯

香港
Joined July 2024
MUSE AI智能体手把手教程来了,这里还有两个邀请码P6P1EH,YT04Y2,每个邀请码仅有30次使用次数
手把手教你注册 Muse AI智能体:全程免费注册,共耗时5分钟,获得永久有效的10亿token和2核CPU、8G 内存、100G磁盘的容器。
1
震惊瘫坐,我在mimo desktop用了两周preview觉得一般般呀,正式版这么强吗
Introducing Xiaomi MiMo-V2.6 — Pro & Flash. Frontier intelligence, all the modalities, built in public. 🔹 Two omnimodal models, advancing through scaled reinforcement learning 🔹 Pro performs on par with Claude Opus 5 and GPT-5.6 Sol across most agent benchmarks 🔹 Pro scores 46 on the Artificial Analysis Intelligence Index — the highest among open-source models 🔹 Stronger coding, computer use, 3D reasoning and creative capabilities 🔹 Open model weights, technical report, RL environments and training code Blog:mimo.xiaomi.com/mimo-v2-6
42
Step 5 Preview breaks the Pareto frontier. Strong agentic performance, great at coding & professional tasks. Sign up with my invite link and grab at least 30 days Token Plan to test it: platform.stepfun.ai/?invite_… #StepFun #Step5Preview
Introducing Step 5 Preview: Advancing the Pareto Frontier. Step 5 Preview is our new flagship model for agentic work, delivering frontier-level performance across software engineering and professional knowledge work, with particular strength in finance. - 600B total / 27B active MoE, with 1M context + Vision - Substantially lower task cost at comparable intelligence - Broad software engineering capabilities with sustained execution over long horizons Try Step 5 Preview: platform.stepfun.ai Model page: stepfun.com/step-5-preview Open weights on Oct 15.
27
最近爆火的jev,其实以前就有类似的小模型做路由、护栏、级联,帮大模型判断“这题难不难、危不危险、该不该拦”,典型的例子就是A/会判断你的需求是不是太蠢不配用他的fable。 Jev就是个极致的产物,不说话,只输出概率、选项、分数。 类似RAG 外挂知识,而Jev 外挂决策。
1
33
提了点速度就涨价,不知道的还以为提到了1000 tokens/s呢,结果只是最高触碰到了Deepseek-v4.1-flash的基础速度。
智谱 Infra 是真的强啊,刚刚又推出了 GLM-5.3-FlashX,是此前 GLM-5.3-Flash 的高速版,官宣最高速度 200 tokens/s。 💰价格约为 Flash 的 2.5 倍:输入 $0.37 / 百万 Tokens,输出 $1.25 / 百万 Tokens,缓存命中 $0.075 /百万 Tokens。
32
PaX retweeted
智谱 Infra 是真的强啊,刚刚又推出了 GLM-5.3-FlashX,是此前 GLM-5.3-Flash 的高速版,官宣最高速度 200 tokens/s。 💰价格约为 Flash 的 2.5 倍:输入 $0.37 / 百万 Tokens,输出 $1.25 / 百万 Tokens,缓存命中 $0.075 /百万 Tokens。
41
10
8
121
32,717
据我观察那些说flash模型不好用的人,基本都是codex或claude的重度订阅者。他们习惯了许愿式编程,不懂软件工程,不懂兜底,不懂测试,只是盼着强大的模型给他完成任何事情😄
2
52
Anthropic‌下一代模型将会继续利用deepseek开源的技术降本,但未必会降价😄
1
44
我发现codebuddy和qoder上的glm5.3-flash的极低积分消耗是个幌子,这个模型实际上非常磨叽,不仅推理时间过长,积分也消耗得也更多,用起来没有dsv4f干脆利落。 现在正确的使用姿势应该是是梁文峰时期用glm,梁文谷时期才用回deepseek
54