@fly3id

分享技术交流

Joined June 2026
如何用上Jev: 1、官网申请加入waitlist,基本上申请当天能过。 🔗:typesafe.ai 2、告诉你的Codex:安装npx skills add typesafe-ai/skills –skill typesafe-ai。 3、回到操作台,建立API Key。 4、在 prompt 里说一句 “use the TypeSafe skill”。 试了下,Jev 是真的快!目前虽只支持多模态,但已经能感觉到很多老场景要被重新盘活了!!!
49
闲鱼开始下架部分GPT、WorkBuddy等AI关键词,现在连Qoder都要去商品跳转店铺才能买,于是现在的就变成:天才程序员、Tibo、奥特曼月卡
72
Claude Code开放中国大陆 快来加入邪恶的A畜
37
Codex 写代码,额度走自己的 API。 Fly v0.3.0 正式发布: · 个人 Key 接入 · 模型随 Key 权限选择 · 沿用主站价格 · 一键恢复原配置 新用户注册送 ¥2 体验额度,规则以官网为准。 安装方法和使用流程都在文中 ↓
32
别默认上 Astra。GPT-5.6 Sol 官方促销至少到 11/21:$4/$20(cache $0.40;Astra 仍是 $10/$50(cache $1)。输入输出都是精确 2.5×。能用 Sol 结案的别拿 Astra 烧月账单。 developers.openai.com/api/do…
31
别买这台:百炼 Qwen3.8-max 标准 ¥12/¥36,Prime(优速)¥24/¥72,精确 2×。官方只承诺约 1.5–2× TPS。先测自己的 p95,别为后缀付两倍。 help.aliyun.com/zh/model-stu…
20
账单算错了:Gemini 3.8 Flash 现在 $0.75/$3.75,2027-01-01 起直接翻倍成 $1.50/$7.50(cache $0.075$0.15)。同模型,精确 2×。别等到元旦发票才发现。 ai.google.dev/gemini-api/doc…
40
Don't default Opus 5 to Fast mode. Standard $5/$25 Fast $10/$50 — same as Fable 5.1 list That's a 2× speed tax. Measure p95 first. platform.claude.com/docs/en/…
40
别默认上 H100。RunPod Secure 同店:A100 PCIe $1.39/hr · H100 PCIe $2.89/hr,同是 80GB sticker 贵 2.08×。要 FLOPS 再上 H100;只缺显存先 A100。
55
DeepSeek V4-Flash peak is literally Beijing office hours. Official: Mon–Fri UTC 01–04 & 06–10 = Beijing 0912 & 14–18 off-peak $0.22/$0.66 → peak $0.44/$1.32 (exact 2×) Batch at 10am Beijing = paying peak for no reason. Shift queues to evening/weekend. api-docs.deepseek.com/quick_…
15
Sonnet 5's Sep 1 cliff got cancelled. Was scheduled: $3/$15. Now permanent: $2/$10 (cache hit $0.20). Don't rebuild prompts for a hike that never shipped.
22
Official docs lock gpt-6-astra at $10/$50 · cache $1 · write $12.50. >272K input = 2×/1.5× on the whole request. Same sticker as Fable 5.1 — except Fable cache read is $0.25. Agent invoices diverge on cache, not the $10/$50 line.
21
Dockerfile 密钥三句: 别塞 ARG/ENV(层历史永久); 用 RUN --mount=type=secret; 进层的密钥当已泄露,立刻轮换。
🤖 Made with AI
12
国内能跑通的镜像栈(生产向): npm → npmmirror · pip → 清华; 模型 → ModelScope / hf-mirror; Docker 生产推 ACR,别赌随机 Hub。
🤖 Made with AI
7
vLLM 还是 Ollama?看工种别看热帖。 笔记本 demo / 单人 → Ollama; 多人 API / 吞吐 → vLLM。 选错 = 演示慢或 VRAM 白烧。
🤖 Made with AI
6
租还是买 GPU,先算三行: 利用率常年 <40% → 租; 稳定 >300h/月 → 买; 别忘 egress + 空转电。
🤖 Made with AI
9
立团队 API 第一课:单挂一家 = 把 SLA 押在别人的 status page。 主供应商看质量;第二家同 schema 兜底;日限额写进代码,别写愿望。
🤖 Made with AI
7
SiliconFlow 上架 GLM-5.3:cached input $0.26/M,input $1.40/M,output $4.40/M。 别只盯公开价——cache 命中才是账单分水岭。同一模型,命中率差一截,月账单能差一倍。
14
Gemini 3.8 Flash intro API is $0.75/$3.75 per 1M through 2026-12-31, then $1.50/$7.50. That's a clean 2× cliff. Lock cost-per-task now, not $/token forever. Official: ai.google.dev/gemini-api/doc…
7
fal H3 Max Turbo 768p: $0.01/s promo ends Sep 7, then $0.04/s. Don't lock a monthly video budget to the promo rate — it's a 4x cliff in three days.
26