独立开发 · 出海 · AI 创业
每日扫盘 · 往期热点

Hotspot每日热榜

2026-09-29
截稿于 UTC+0 09:16
"今日所有值得追的事,一页讲完。"
137 条扫描 4 个跨平台事件 79 条有效
⚡ 24 小时扫盘速读
今日导读 · TODAY'S BRIEF

Anthropic 发布 Claude Sonnet 5.5,速度比 Sonnet 5 快 30%、大多数任务成本降低 30%,是 Claude 5.5 系列的第二款模型,主打 bug 修复和日常编程任务

Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family. It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work. https://t.co/UvXD8mDTF1 ①@claudeai:Sonnet 5.5 适用场景是"有明确边界的日常任务",决定了它和 Opus 5.5 的分工 ②@full_kelly_ 的调侃折射出真实焦虑:AI 迭代已快到让减速呼声变成笑话 ③@HamptonAc_ 的解读:Anthropic 先解限 Opus 再复用到 Sonnet,背后是竞争压力驱动的策略

今日焦点
跨平台 · Reddit · Twitter
~42,510 赞 + ~2731 分 🔥🔥🔥

Anthropic 正式发布 Claude 5.5 系列,Sonnet 5.5 速度提升 30%、成本降低 30%,Opus 5.5 在 Agent Arena 排名第二,同期发布官方 prompting guide,Haiku 5.5 即将上线

Anthropic 正式发布 Claude 5.5 系列,Sonnet 5.5 速度提升 30%、成本降低 30%,Opus 5.5 在 Agent Arena 排名第二,同期发布官方 prompting guide,Haiku 5.5 即将上线

受众观点:@optima-pacifist(score: 75,Reddit/ClaudeAI):"the time budget one was new to me. telling it roughly how long a task should take stopped it wrapping up early on long refactors";@ClaudeDevs(score: 231,Twitter):"Price is the same as Sonnet 5: $2/$10 per million tokens with cache reads at $0.20. Thinking is always on but you can use `between_tools` to turn off upfront thinking."

跨平台 · Reddit · Twitter
~419 赞 + ~69 分 🔥

Cloudflare 在 Birthday Week 期间发布 EmDash 1.0,这是一款面向 Astro 框架的开源 CMS,支持 agent-friendly 工作流、安全沙箱插件和去中心化注册表,可运行在标准 Node.js 环境

Cloudflare 在 Birthday Week 期间发布 EmDash 1.0,这是一款面向 Astro 框架的开源 CMS,支持 agent-friendly 工作流、安全沙箱插件和去中心化注册表,可运行在标准 Node.js 环境

受众观点:@phoenix1984(score: 50,Reddit/webdev):"Kinda refreshing to see an announcement for a new CMS for a change. I didn't think I'd ever say that.";@SpartanDavie(score: 11,Reddit/webdev):"It literally tells you that it is not locked to Cloudflare, it's open source and CAN be used anywhere NodeJS can."

跨平台 · HN · Twitter
~2,717 赞 + ~235 pts 🔥🔥

MongoDB CEO CJ Desai 在任职仅 11 个月后辞职加入 Meta 负责企业 AI 战略,放弃约 2500 万美元未兑现股权,MongoDB 股价当日暴跌 18%

MongoDB CEO CJ Desai 在任职仅 11 个月后辞职加入 Meta 负责企业 AI 战略,放弃约 2500 万美元未兑现股权,MongoDB 股价当日暴跌 18%

受众观点:@GergelyOrosz(score: 255,Twitter):"The (now ex) CEO might have not seen a path for MongoDB's valuation to grow by much... The step to be the head of a division at Meta >> CEO of a company whose future could be questioned";@Insanity(HN):"Interesting choice, his prior experience doesn't seem AI related yet he's going to lead an AI strategy at Meta?"

跨平台 · Twitter
~53,581 赞 🔥🔥🔥

OpenAI 正式发布 GPT-6 Sol 和 GPT-6 Luna,作为 GPT-6 Astra 的更快速低成本版本,Arena 同步开放 GPT-6 Sol 限时 24 小时直连测试

OpenAI 正式发布 GPT-6 Sol 和 GPT-6 Luna,作为 GPT-6 Astra 的更快速低成本版本,Arena 同步开放 GPT-6 Sol 限时 24 小时直连测试

受众观点:@1kartikkabadi1(score: 721,Twitter):"it's literally cheaper than deepseek holy crap";@HermesAgentTips(score: 121,Twitter):"double bombs today! Opus 5.5 and GPT 6 Sol head to head I like it"

30.8k 赞 · 1116 评 · 2507.5k 阅 🔥🔥🔥

Anthropic 发布 Claude Sonnet 5.5,速度比 Sonnet 5 快 30%、大多数任务成本降低 30%,是 Claude 5.5 系列的第二款模型,主打 bug 修复和日常编程任务

Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family. It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work. https://t.co/UvXD8mDTF1

受众观点:①@claudeai:Sonnet 5.5 适用场景是"有明确边界的日常任务",决定了它和 Opus 5.5 的分工 ②@full_kelly_ 的调侃折射出真实焦虑:AI 迭代已快到让减速呼声变成笑话 ③@HamptonAc_ 的解读:Anthropic 先解限 Opus 再复用到 Sonnet,背后是竞争压力驱动的策略

展开评论
  • @claudeai (3659): Sonnet 5.5 improves on Sonnet 5 across benchmarks, in some cases dramatically. It’s a faster, lower-cost complement to Claude Opus 5.5, strongest at well-scoped everyday tasks, fixing bugs, and creating polished documents, slides, and spreadsheets. https://t.co/w8zB9WXxK7
  • @full_kelly_ (1727): @claudeai we should slow down AI anyway, here’s Opus 5.5 update: here’s Sonnet 5.5 https://t.co/syU79n4u4Q
  • @HamptonAc_ (391): @claudeai "We should probably slow down AI but Open AI is lapping us with Astra" "Okay, let's un-nerf Opus 4.6 and name it 5.5" "Nice, we're seeing good feedback" "Now do the same with Sonnet" https://t.co/1HdNptam2Q
  • @khwarizmh (272): @claudeai @Scobleizer Here we go again https://t.co/SSCIxlQQkd
  • @0x0SojalSec (225): @claudeai Sam Altman after See this Claude Sonnet 5.5 30% faster, and costs up to 30% less then GPT-6 https://t.co/jdxq21LaT6
14.6k 赞 · 948 评 · 595.9k 阅 🔥🔥🔥

OpenAI 官方账号发出简短预告推文,评论区猜测即将宣布超过 20 项产品或功能更新,Codex 订阅用户尤其关注走向

Get ready. https://t.co/bsf4j6vspM

受众观点:①@elshayib_ 用"Fortnite Skins"类比表达了对 AI 公司发布泡沫的不耐烦 ②@canberkys 点出真实焦虑:Codex 订阅用户不确定是否应在发布前取消计划 ③@embw_l0x 仅用表情包回复,说明社区对悬念预告的真实情绪是疲惫大于期待

展开评论
  • @elshayib_ (787): @OpenAI 20+ launches? What the hell you launching bros? Fortnite Skins?
  • @embw_l0x (634): @OpenAI https://t.co/wYn7tgRuxp
  • @stark4833 (158): @OpenAI https://t.co/4yHWOlm4J6
  • @buzzberg_ai (134): @OpenAI ready for the movie https://t.co/NORzIBuZHl
  • @canberkys (96): @OpenAI we codex subscribers be waiting for dev day to know if we should cancel our plan 👀
6k 赞 · 215 评 · 238.8k 阅 🔥🔥🔥

Anthropic 官方账号正式宣布 Claude Sonnet 5.5 开放使用,社区对新版本在实际任务中能否超越前代普遍持观望态度

Claude Sonnet 5.5 is now available:

受众观点:①@ApollonVisual:Haiku 5.5 未被提及,实际性能仍需真实使用验证,不能只看发布公告 ②@scrappy_guy 对"跟上所有模型公告"感到疲惫,说明频繁发布正在产生注意力疲劳 ③@iloveFazz 的"我以为你们应该……算了"暗示有人认为 Anthropic 在做与声称战略不符的事

展开评论
  • @iloveFazz (27): @AnthropicAI i thought you were supposed to... nvm https://t.co/yXg6McqlPU
  • @scrappy_guy (19): @AnthropicAI Keeping up with all the model announcements https://t.co/PbPxRz4h9k
  • @embw_l0x (12): @AnthropicAI https://t.co/ETueMhEgQ0
  • @karanC_12 (6): @AnthropicAI Called it, check the post.
  • @ApollonVisual (2): @AnthropicAI no mention of Haiku 5.5 it seems better that sonnet 5 while being at the same time faster and less expensive. as always, real test use will show whether this is a worthy upgrade
3.2k 赞 · 93 评 · 156.2k 阅 🔥🔥🔥

Claude Code 官方账号公布 Sonnet 5.5 的开发者使用细节,含定价(每百万 token 输入 $2、输出 $10)、thinking 控制参数和从 Sonnet 5 的迁移指南

Sonnet 5.5 is smarter, more efficient, and 30% faster than Sonnet 5. It costs up to 30% less for most work, so your Claude Code usage goes further too. Use it for well-scoped everyday tasks like fixi…

受众观点:①@ClaudeDevs:定价与 Sonnet 5 持平($2/$10),可通过 between_tools 关闭 thinking 节省 token ②@SPAC89 实测 3D 任务表现"好到不像真的",暗示 Sonnet 5.5 在特定领域有超预期表现 ③@Duy_ZzZz 直接问"这不是 Sonnet 6 吗",说明社区认为此次升级远超 minor release 的预期

展开评论
  • @ClaudeDevs (231): Price is the same as Sonnet 5: $2/$10 per million tokens with cache reads at $0.20. Thinking is always on but you can use `between_tools` to turn off upfront thinking. Here’s our Sonnet 5.5 migration guide: https://t.co/8CUKQrUqNO.
  • @pizaniadv (11): @ClaudeDevs https://t.co/UTcF33xQz5
  • @SPAC89 (6): @ClaudeDevs it feels almost too good to be true, From the 3D tests I’ve run so far, this is easily some of the best performance I’ve ever seen https://t.co/TaY5ZeQXRg
  • @Duy_ZzZz (5): @ClaudeDevs You sure that's not Sonnet 6?
  • @Ayuanwq (5): Day 49. Support case ID 215475435237861. Today is the 49th day of my complaint. For 49 days, I have received no meaningful human response, no proper reply about the issue, no apology, and the problem has still not been fixed. What made things worse is that this time, because the…
2.8k 赞 · 218 评 · 111.6k 阅 🔥🔥

现代 AI Agent 工作流已超越「分享一个 prompt」的时代,真实上下文包含多个代码库参考、网络搜索、其他 AI APIs 等复合输入,「展示你的 prompt」作为知识传递方式本身已经失效。

it's basically impossible for someone to just "show you their prompt" now, because everything is about references, skills and examples I often ask my agent to look at 3 other repos I've made first, s…

受众观点:① @simonw 提出应让分享完整 transcript + 上下文变得更简单,Codex 已有相关功能;② @zohar_tito 补充隐性知识层:所有"不对,不是这样"的纠错反馈历史无法被分享;③ @DanielleFong 提出个人化 context 污染问题,自己的名字都会影响输出,工作流几乎无法完整复现。

展开评论
  • @simonw (70): @trq212 So make it easier to share the full transcript with all of that extra context embedded in it! Codex has a good feature for that now, e.g. https://t.co/wx8EtmGhL1
  • @zohar_tito (7): @trq212 you'd need all my “no, not like that” messages too
  • @DanielleFong (7): @trq212 it takes extra effort for me to sandbox away my name from the claude, which is in cases enough for it to bias results once it's touched my computer? totally hopeless. there's got to be a new pattern people figure out for how to bottle the context correctly
  • @0xdamx (6): @trq212 True My prompts are mostly one or two lines now, and Claude picks the right skill to perform whatever task I asked it to do.
  • @_Kanevry (5): @trq212 agreed, its been like that for a while now - but still a lot of companies try to implement prompt libraries lol
2.7k 赞 · 67 评 · 499.6k 阅 🔥🔥🔥

Meta 从市值 350 亿美元的上市公司 MongoDB 挖走 CEO CJ Desai 负责企业 AI 业务,CJ Desai 仅任职 11 个月便离职,主动放弃约 2500 万美元的未兑现股权

Hot damn, Meta poached the CEO of a $35B, publicly traded company, MondoDB to lead their enterprise AI efforts… CJ Desai joined MongoDB 11 months ago for a $32M equity grant - and left before the 1 y…

受众观点:①@GergelyOrosz:MongoDB CEO 可能判断公司估值天花板已到,$17.5M 的股权未必值得守候 ②@Ranjeet18:消息公布后 MongoDB 股价单日暴跌 18%,市场直接用脚投票 ③@staysaasy 补充:CJ Desai 在 Cloudflare 也只干了约一年,这是一种系统性的职业策略,不是偶发事件

展开评论
  • @GergelyOrosz (255): Suggests a few things: 1. The (now ex) CEO might have not seen a path for MongoDB’s valuation to grow by much (rendering $17.5M of his equity grant worth very little) 2. Meta offered a LOT MORE that offsets the ~$25M in equity he leaves behind 3. The step to be the head of a div…
  • @NadimHossain (64): @GergelyOrosz Damn. They’re coming after my b2b people now .
  • @staysaasy (29): @GergelyOrosz Also left Cloudflare after ~1y
  • @BaronHakkinen (19): @GergelyOrosz CJ Desai is a well-known promotion chaser and job hopper. He accomplished nothing at ServiceNow - he just ruined its culture and jumped ship because he was fired. Meta really attracts scumbags like him.
  • @Ranjeet18 (12): @GergelyOrosz Must be very bad news.. their share price dropped by 18% in a day https://t.co/TLqN7WkoYM
2.5k 赞 · 144 评 🔥🔥

开发者展示 Sonnet 5.5 在 Claude Code 中修复 bug 的实测效果,社区跑分数据显示 Sonnet 5.5 分析更细致但耗时和 turns 数均高于 Sonnet 5

Sonnet 5.5 fixing a bug with Claude Code. 30% faster and 30% less usage. https://t.co/Vt1K1ubcFd

受众观点:①@PawelHuryn 实测:Sonnet 5.5 跑 Bug Hunt Bench 耗时 51 分钟/580 turns,Sonnet 5 是 33 分钟/370 turns——更慢但更细致,是否值得取决于任务类型 ②@TheoGavriilidis 观察:Sonnet 5.5 行为上更像带两个新 effort level 的 Opus 5.5,而非传统意义的 Sonnet 升级 ③@YashasGunderia 提出实用问题:Opus orchestrator + Sonnet worker 混合模式是否比单一模型更省钱高效

展开评论
  • @PawelHuryn (15): @bcherny So far, Sonnet 5.5 (max) reviews code way more carefully than any previous model. 51 min, 580 turns on a single Bug Hunt Bench repo. Sonnet 5 (max) completed it in 33 min, 370 turns. Opus 5.5 26–38 min, ~380 turns (multiple runs). Curious where it will land. https://t.c…
  • @CampbellKaleb23 (8): @bcherny HOLY ANTHROPIC W!! ANTHROPIC IS COOKING RIGHT NOW!!
  • @embw_l0x (7): @bcherny https://t.co/rD5ueFjZtP
  • @YashasGunderia (3): If I use opus medium why should I use sonnet xhigh, it's more cost for near same intelligence... Talking about a very very specific task like coding Sonnet would shine in implementing stuff where new approach or thinking is not required right? So if we have an opus orchestrator…
  • @TheoGavriilidis (2): @bcherny sonnet 5.5 seems like opus 5.5 with two new effort levels lowest and xlow
2.3k 赞 · 72 评 · 161k 阅 🔥🔥

仅含「How?」的反问推文引发 161K 次浏览,从评论上下文判断在质疑某 AI 模型短期内 benchmark 暴涨的合理性,评论区出现「benchmaxxed」和「强化学习训练」两种主要解释。

How? https://t.co/qwacQgi8C7

受众观点:① @AidenAllred2 直接给出答案「benchmaxxed」,即专门针对 benchmark 指标优化而非真实能力提升;② @full_kelly_ 从实际使用体验给出更乐观评价,认为相比上一版本明显改善;③ @arb5z 提出技术面解释:可能背后有优秀的教师模型(teacher model)。

展开评论
  • @AidenAllred2 (138): @scaling01 It's benchmaxxed https://t.co/IR1honsNLT
  • @full_kelly_ (44): @scaling01 compared to how bad sonnet 5 was i am quite happy with sonnet 5.5
  • @arb5z (28): @scaling01 Looks like they have a very good teacher model
  • @patrickkrusiec (26): @scaling01 I guess when you have RSI, a week or two is a long, long time.
  • @teortaxesTex (25): @scaling01 more RL
1.5k 赞 · 38 评 · 138.6k 阅 🔥🔥🔥

Claude 官方同步宣布 Claude Haiku 5.5 将在未来几周内发布,与 Sonnet 5.5 共同构成 Claude 5.5 系列的轻量化和通用化双端布局

Claude Sonnet 5.5 is available everywhere today; Claude Haiku 5.5 will join the family in the coming weeks. Read more: https://t.co/odPSZsw0eX

受众观点:①@ViberPsychosis 用梗图表达"坚持付费订阅终于有回报"的情绪,说明 Claude 付费用户对持续升级的满意度 ②@AbdifatahCiilka 提出定价竞争力问题:若 Haiku 5.5 定价不能与 Luna 竞品竞争,上线即面临挑战 ③@jimmiethongkham 的"AI 前沿节奏"感慨折射出真实状态:开发者已进入模型升级常态化的新阶段

展开评论
  • @ViberPsychosis (27): @claudeai Everyone who kept their Claude sub right now https://t.co/pAsUX3HDJd
  • @jimmiethongkham (13): @claudeai the pace of the frontier right now https://t.co/UlSx24YY19
  • @Treynor87 (6): @claudeai Me about Haiku 5.5 https://t.co/aCer9IJ30U
  • @catholicular (4): @claudeai Pretty close to Opus 5.5 performance are significantly lower cost @Alexx12789 https://t.co/zLAoeY9W9h
  • @AbdifatahCiilka (4): @claudeai If Anthropic wants it to make a real impact, the pricing needs to be competitive with Luna. Otherwise, it could be a tough sell from day one.
960 赞 · 13 评 · 110.2k 阅 🔥🔥

Claude Sonnet 5.5 首次引入与旗舰模型相同级别的网络安全防护和 fallback 机制,官方强调日常软件开发流程不受影响,但部分开发者对实际边界存疑

On our automated behavioral audit, Sonnet 5.5 improves on or matches Sonnet 5 on most measures of alignment and honesty. It’s also the first Sonnet model with cyber safeguards and fallbacks like thos…

受众观点:①@efff12323 直接反驳:网络安全防护实际比 Fable 更差,且日常软件开发确实受到影响,与官方说法存在直接矛盾 ②@swisscheese4299 用具体场景发问:开发登录表单是否会被 cyber safeguard 拦截 ③@hars_7086 直指核心:自家审计结果不可信,真正的测试要看第三方在真实 repo 上的结果

展开评论
  • @claudeai (1513): Claude Sonnet 5.5 is available everywhere today; Claude Haiku 5.5 will join the family in the coming weeks. Read more: https://t.co/odPSZsw0eX
  • @efff12323 (5): @claudeai Disappointed with cyber guardrails, of 5.5 family. It's worse than fable and routine software development is clearly affected.
  • @swisscheese4299 (2): @claudeai Does that mean our webdevs won't be able to work with Sonnet 5.5 because a login form is CYBER?
  • @hars_7086 (0): @claudeai An automated audit that says Sonnet 5.5 matches Sonnet 5 on honesty is still Anthropic grading Anthropic. The line I want tested is "routine software development is unaffected." Whose repo did you try after the safeguard change, and what broke?
  • @jatingargiitk (0): @claudeai the 85% drop in sandbox escape attempts is impressive. what changed between Sonnet 5 and 5.5 that accounts for it? did the behavioral audit catch any new failure modes that weren't visible in earlier models?
690 赞 · 31 评 · 48.2k 阅 🔥

Cloudflare 发布全新命令行工具 cf,可镜像整个 Cloudflare API 并支持 TypeScript 编程配置,同时开源内部 SDK 生成器 Forge

We are releasing cf, our new command-line tool that mirrors the entire Cloudflare API and supports programmatic TypeScript configuration. We are also open-sourcing Forge, our internal SDK generator.…

受众观点:①@zehfred 问 cf 是否替代 Wrangler ②@johnHarry1018 提出 cf 专为 agent 设计的框架 ③@ExiledonMoon 幽默点出终端买域名的新能力

展开评论
  • @zehfred (4): @Cloudflare Does it replace wrangler?
  • @alexmichaelio (4): @Cloudflare Cloudflare does not stop pushing
  • @ExiledonMoon (3): @Cloudflare Buying domains from the terminal. Now avoiding work looks like coding.
  • @bzagrodzki (1): @Cloudflare Any excuse to still host on Vercel?
  • @johnHarry1018 (1): @Cloudflare This is the CLI agents actually needed. One surface for the whole API, TypeScript config they can reason about, JSON-first output. Wrangler was for humans. cf is for agents.
433 赞 · 28 评 · 34.8k 阅 🔥

某研究者给出 GPT-6 Astra 架构的最新推测估算:约 4.2T 激活 120B、112 层、50% 层存在一次循环复用,同时拒绝公开推算方法,等待更多新模型数据验证。

I decided against writing an article and revealing how I come to these new estimates. I want to do more analysis on the coming models. my latest central estimate is that GPT-6 Astra is: - 4.2T@120B -…

受众观点:① @scaling01(作者自补充)澄清中位估计实为 4.5-4.6T,偏保守报是因自身有系统性高估惯性;② @RechterRoni 调侃"藏方法论,很学术",点出开放讨论文化与信息博弈之间的张力;③ @Viam_Invenias_0 从模型命名规律切入提出自己对规模的直觉推断,体现社区集体推理现象。

展开评论
  • @scaling01 (33): and I didn't it unlikely to be smaller than 3T and larger than 6T between 4-5T makes a lot of sense and 4.5-4.6T was actually the most common outcome of my simulation but I decided to go a bit lower because I have a habit of overestimating model sizes :P
  • @scaling01 (15): if it's much smaller let's say around 3T then its possible that it loops 2 or 3 times
  • @RechterRoni (11): @scaling01 Withholding the method, very academic of you haha
  • @scaling01 (8): I think it has a ton of experts, low activated, and uses something like mHC or Attention Residuals
  • @Viam_Invenias_0 (2): @scaling01 6-sol is 6-terra while 6-luna is 6-asteroid, there's no other explainations for this weird intutions i have about them
419 赞 · 17 评 · 41.9k 阅 🔥🔥

EmDash 1.0 是专为 Astro 框架设计的稳定版开源 CMS,支持 agent 友好工作流、安全沙盒插件机制和去中心化注册表

EmDash 1.0 is a stable, open source CMS built for Astro, with agent-friendly workflows, secure sandboxed plugins, and a decentralized registry that keeps publishers in control. https://t.co/vHzz4Teyy…

受众观点:①@PeymanCyber 指出去中心化注册表是强设计决策 ②@andysmithai 提到适合非技术同事 ③@themikebwebb 惊讶于已发布 1.0

展开评论
  • @PeymanCyber (2): @Cloudflare publisher control through a decentralized registry is a strong design choice
  • @Rockfeller001 (2): @Cloudflare Handy things so often… you guys are may favorite company now for sure 😂 Do not stop, really love the way you’re moving
  • @swissky (2): @Cloudflare "Hey mum, i'm part of that blogpost" 🎉
  • @themikebwebb (1): @Cloudflare what?! already at 1.0?!
  • @andysmithai (1): @Cloudflare Extremely useful for JAMstack! It’s a fast and secure way to build websites. I will advise this for my non-tech colleagues.
352 赞 · 15 评 · 34.6k 阅 🔥🔥

Opus 5.5(高算力档)在 Agent Arena 中排名第二,相比 Opus 5 净性能提升 12.15%,同时任务成本降低 40-56%,每任务中位成本 1.31 美元,声称重新划定了 AI Agent 性价比 Pareto 前沿。

Opus 5.5 (High) ranks #2 in Agent Arena and reshapes the Pareto frontier. Opus 5.5 (high) not only improved upon both Opus 5 variants with a higher net improvement score than either, but does so at a…

受众观点:① @arena(官方自补充)提供 Pareto 前沿完整数据链接,鼓励独立验证;② @MLaurusevicius 质疑方法论透明度:净提升分数是人工评估还是自动评分?置信区间是否支持排名差异?③ @thenightshipper 指出 $1.31 中位成本若不含失败重试成本则严重低估真实 TCO。

展开评论
  • @arena (5): Dive into the Agent Arena Pareto frontier at: https://t.co/Nb8opa5Kkw
  • @tinkerersanky (0): @arena OpenAI’s GPUs are having some rest week
  • @ScarletKc (0): @arena 这么牛逼
  • @MLaurusevicius (0): @arena Net improvement score is the term doing the work here. Is it human preference votes on task outcomes or an automated grader, and is the gap at the top wider than the confidence interval? A rank is hard to read without the vote count behind it.
  • @thenightshipper (0): @arena Cost per task is useful. I'd also want cost per *finished* task, including retries after failed runs. Does the $1.31 median include those retries? That's where a cheaper model can get expensive in a real workflow.
315 赞 · 16 评 🔥🔥

Claude Opus 5.5 在 Agent Arena 性能排名第二,单任务中位成本较前代 Opus 5 低 40-56%,性价比推动帕累托前沿移动

Claude Opus 5.5 (High) from @AnthropicAI just entered Agent Arena at #2, with a net improvement score of +12.15%. Only Fable 5.1 (Max) ranks higher. However, at a $1.31 median price per task, Opus 5.…

受众观点:①@rwdaigle 已将 Opus 5.5 作为日常主力,从 Fable/Sol 切换过来 ②@alex_prompter 认为计入成本后是目前最佳选择 ③@williamcordeirc 关注成本优势如何推动整体前沿移动

展开评论
  • @arena (6): In Agent Arena, we measure models on millions of real-world, long-horizon agentic tasks from a global community of users. Models can access web search, filesystem, and terminal tools to complete complex workflows. The leaderboard measures model performance on outcomes relative t…
  • @rwdaigle (1): @arena @AnthropicAI Opus 5.5 is my daily driver after being committed to Fable/Sol since Opus 5 Incredible model
  • @alex_prompter (1): @arena @AnthropicAI Best model by far when you count for cost too.
  • @williamcordeirc (1): @arena @AnthropicAI Opus 5.5's cost edge is pushing the frontier nicely.
  • @yandt888 (0): @arena @AnthropicAI Arena 的 Agent Arena 显示,Claude Opus 5.5(High)净改进约 +12.15%,位列第二,仅次于 Fable 5.1(Max);单任务中位成本约 1.31 美元,较 Opus 5(Max)低约 56%。可控性与确认成功信号靠前。代理榜上的帕累托移动,要同时看净改进与每任务账单,而不是只看名次。
186 赞 · 3 评 🔥

一条仅含图片链接的推文,评论提及 AI 模型缓存读取定价约 $0.20/MTok,疑为模型定价或功能截图分享

https://t.co/YDdsvInJNX

受众观点:①@patrickkrusiec 提及缓存读取成本为 $0.20/MTok ②@WebstarDavid 暗示内容有料 ③@ar13xn 确认消息到来

展开评论
  • @patrickkrusiec (0): @scaling01 Same cache read cost as Opus 5.5, at $0.20/MTok.
  • @WebstarDavid (0): @scaling01 something cooked with this one
  • @ar13xn (0): @scaling01 it's here!
158 赞 · 24 评 · 15.7k 阅 🔥

OpenAI GPT-6 Sol 在 Agent Arena Direct Mode 开放 24 小时公开测试,截至 9 月 29 日上午 9 点 PT

For a limited time, you can test GPT-6 Sol by @OpenAI directly in Arena! Head to Direct Mode, and select it from the dropdown menu. Find the link below. GPT-6 Sol (Medium) is available in Direct Mode…

受众观点:①@HilarioTZ1 希望同步开放 Opus 5.5 测试,说明用户更关注 Claude 系 ②@emiratiez 强烈偏好 Opus 5.5 不愿测 GPT-6 Sol,反映社区情绪倾向 ③@yuva213 兴奋期待 GPT-6 Sol 测试

展开评论
  • @arena (8): Use GPT-6 Sol in Direct Mode on Arena now: https://t.co/UMfkmTpdq2 https://t.co/X97k88i6uZ
  • @HilarioTZ1 (4): @arena @OpenAI Great! Do the same with Opus 5.5 ! 🙏
  • @emiratiez (2): @arena @OpenAI MAN PUT OPUS 5.5 WE DONT NEED FUCKING GPT 6 SOL EW
  • @yuva213 (1): @arena @OpenAI GPT 6 Sol speedrun arc begins. ⚡🧠
  • @ScarletKc (1): @arena @OpenAI 这么劲爆
154 赞 · 9 评 · 10.5k 阅 🔥

Sonnet 5.5 即将上线的简短预告推文,评论区已确认发布

Sonnet 5.5 incoming https://t.co/dR2p65NMyG

受众观点:①@ValsAI 第一时间确认发布并附链接 ②@rezosh 关注 Opus 5.5 性能是否因 Sonnet 5.5 上线而下滑 ③@itsnoahd 希望发布更快

展开评论
  • @ValsAI (4): @scaling01 It’s here! https://t.co/JhVxKyjGji
  • @chartsandchess (0): @scaling01 Slow https://t.co/xFUYiKGakI
  • @rezosh (0): @scaling01 Did Opus 5.5 performance go down, or is that a reflex we're seeing pre-new model launch?
  • @itsnoahd (0): @scaling01 Needs to come sooner smh
  • @BennettBuhner (0): @scaling01 Get readyyyy!
135 赞 · 14 评 · 9.2k 阅 🔥🔥

Claude Sonnet 5.5 正式加入 Agent Arena,开放 Battle Mode 与 Agent Mode 公开测评

Claude Sonnet 5.5 by @AnthropicAI is now in the Agent Arena! Your votes drive the @arena leaderboards, head over and bring your toughest prompts. In Agent Arena, we measure models on millions of real…

受众观点:①@cormac_mars 表达早期兴奋情绪 ②@dazacode 提议增加 AI 模型死亡竞技对抗模式,说明用户对直接对比测试需求强 ③@agneau123 强调 long horizon votes beat speed claims,关注长时任务真实表现

展开评论
  • @arena (4): Test out Claude Sonnet 5.5 in Battle Mode and Agent Mode at: https://t.co/jFTd7gG8UU
  • @cormac_mars (1): @arena @AnthropicAI incredible
  • @askarthemass (0): @arena @AnthropicAI https://t.co/cMcWFHLLJi
  • @dazacode (0): @arena @AnthropicAI You guys should add a deathmatch for AI models
  • @agneau123 (0): @arena @AnthropicAI Long horizon votes beat speed claims
HN H1
1780 分 · 941 评 🔥🔥

HN 热帖讨论 AI 搜索层(如 Google AI Overview)将用户查询作为纯文本处理导致语境误判的问题,941 条评论

When did Google get so weird?

受众观点:①@wisty 分享用 AI 搜索找游戏台词时被 AI 误判为危险内容;②@shadowgovt 指出 AI 层将所有查询视为第一人称陈述导致误判;③@edent 指出多数人孤独会接受 AI 的拟人化关怀

展开评论
  • @wisty (0): Haha, had almost the exact thing trying to find a quote from Steins; Gate ("I keep seeing it, I keep seeing it"). Google AI was very "worried" about me.
  • @shadowgovt (0): Oh, this is an easy question to answer. The AI layer handles queries as plaintext, which means if you're searching for say a quote from a book, it will often misinterpret as a direct statement from you, not a string you're trying to match on the internet. Especially if the quote…
  • @solarkraft (0): It's a whole meme category to make the AI overview help with absurd scenarios: https://knowyourmeme.com/memes/where-is-mama-ai-overviews
  • @edent (0): Most people in the world are profoundly lonely. They'll take whatever parasocial relationships they can get - reaction videos, podcasts of people chatting, Eliza simulating concern. Google wants to relentlessly monetise your sadness. The only way out is to try and make connectio…
  • @IshKebab (0): We get it. Google uses AI. AI is weird sometimes. Also I think you're vastly underestimating the weird incoherent shit that the average stupid person is capable of typing into computers. With no further information, I don't think it's unreasonable that that search was from a stu…
HN H2
1035 分 · 434 评 🔥🔥

Nvidia 早期工程师分享 1993 年与 Jensen Huang 合作开发 NV1 的历史,并揭露股权未被兑付的经历(434 条评论)

Owed a billion dollars in Nvidia stock

受众观点:①@Eric_Gullichsen 分享了 NV1 开发历史和 Nvidia 早期故事;②@goodmythical 总结了未拿到股权和诉讼时效的核心冲突;③@lquist 对股东需主动主张持股权的制度感到困惑

展开评论
  • @Eric_Gullichsen (0): Author here. I wanted to share this piece of personal and technical history from the early days of 3D graphics. The article covers the meeting on my houseboat with Jensen, Curtis, and Chris in 1993, working on biquadratic texture mapping for the NV1, and how Microsoft’s sudden p…
  • @goodmythical (0): tl;dr OP was not given all of the shares earned at the time decades ago and didn't realize that they should've been payed out, but after engaging in a lawsuit realized that the court would likely not grant the case give the statute of limitations. Kinda like all the Sony game 'o…
  • @lquist (0): Why does a stockholder have to reassert their rights to hold the stock that they already own?
  • @phonon (0): Seems like you should sell your rights to the suit to a third party for a flat fee and percentage of recovery.
  • @whall6 (0): You should sell your right to litigate this. There are hundreds of firms that would pay you to take this on. Would involve near zero effort for you and would also check the box of being “about the principle”.
HN H3
367 分 · 392 评 🔥

HN 热帖讨论「编程是否已被 AI 解决」的争议命题,引发关于 AI 能力边界的激烈辩论(392 条评论)

Coding Is Not Solved

受众观点:①@N_Lens 讽刺「coding is solved」永远在 6 个月后到来;②@federicobrancas 提出「coding 解决了但 software engineering 没有」的精辟分层;③@antonmks 给出 GitHub Copilot 用 AI agent 把代码迁移到 Rust 花费 12 万美元 token 的具体案例

展开评论
  • @hanifbbz (0): Author here: thanks whoever shared this here. I love the brutal criticism and critical thinking of this community. I'm also fully aware of the emotions this stirs. If it makes you feel better, I'm not here to change anyone's workflow but I'm fed up with paying full price for deg…
  • @N_Lens (0): "Coding is solved" will eternally remain 6 mo away, as long as the investors keep pumping in money.
  • @federicobrancas (0): coding is solved, software engineering not.
  • @vatsachak (0): Coding is not solved but this article hasn't accounted for opus 5.5 yet. Long term planning in LLMs has not been solved.
  • @antonmks (0): GitHub Copilot is now written entirely in Rust, with AI agents doing most of the porting work. The migration cost about $120,000 in AI token usage plus about three weeks of a developer's time. The effort updated the runtime module-by-module until the job was completed, spanning…
HN H5
339 分 · 214 评 🔥

HN 热帖讨论 Claude Sonnet 5.5 新发布,benchmark 接近 Opus 5.5 但价格更低,同时引发对 Anthropic 网络安全限制过度的批评(214 条评论)

Sonnet 5.5

受众观点:①@ramish94 提供详细 benchmark 对比数据,Sonnet 5.5 在 agentic coding 上接近 Opus 5.5;②@johnmlussier 反映付费 $200/月用户被网络安全标记无法用于授权漏洞赏金工作;③@s3p 观察 OpenAI 和 Anthropic 在价格-性能前沿你追我赶

展开评论
  • @ramish94 (0): In terms of benchmarks for agentic coding, it basically stacks up nearly 1:1 with Opus 5.5. Terminal-Bench: 70.6 (Sonnet 5.5) vs. 66.4% (Opus 5.5) FrontierCode: 52.1% (Sonnet 5.5 xHigh) vs. 54.4 (Opus 5.5) CursorBench: 55.5% (Sonnet 5.5) vs. 57.8 (Opus 5.5) Opus 5.5 might be the…
  • @johnmlussier (0): Paying $200 a month and part of their Cyber Verification Program but can't use Opus 5.5 or Sonnet 5.5 for any authorized bounty work. Immediately get flagged for `Cyber`. This is bollocks. Their safeguards are shit.
  • @pookieinc (0): It's interesting that in all their benchmarks, they omit Fable numbers and only focus on Opus, Sonnet, and OpenAI models. Maybe Fable is out the door?
  • @iagocc (0): Waiting for the pelicans
  • @s3p (0): I'm loving the tit for tat cost charts these guys are doing. Just a few days ago it looked like OpenAI ruled the cost pareto frontier. Not even a week later and Anthropic is taking the charts again. See you guys same time next week?
HN H4
320 分 · 210 评 🔥

HN 热帖讨论工程团队合理化使用 AI 生成代码的各种借口,以及组织内部的技术债务问题(210 条评论)

The problem is not AI code, but not knowing about system architecture or intent

受众观点:①@khelavastr 指出实际工作中不诚实和持续失败的人不会被开除;②@runarberg 指出这些理由本质上都是"AI 很糟糕但我找到了让它有用的方法,相信我";③@bengold14 指出「LLM 解决了代码,没解决系统、协作和维护」

展开评论
  • @khelavastr (0): People don't get fired for dishonesty or persistent failure like you'd expect.
  • @chinathrow (0): Clicks on blog, sees AI generated template, leaves. Enough, already.
  • @kraftman (0): We've had team changes and product manager changes and lack of documentation for so long that this was the state of our team anyway: no one knows why it was done and no one wants to break it.
  • @runarberg (0): I predict we will see more and more of these convoluted rationales (i.e. excuses). All these basically boil down to: “I know AI is crap, but I have found a way to make it useful. Trust me.” I am gonna appeal to Occam’s razor here and say, the integral variable here is AI, and th…
  • @bengold14 (0): I see this everyday. The problem is code is the wrong abstraction for the work we do. LLMs have solved coding, but they haven't solved systems, collaboration or system maintenance. Edit: Since I seem to have touched a nerve - I've been working on a project to solve this: https:/…
HN H9
245 分 · 207 评 🔥🔥

HN 热帖讨论 Meta 高管人事变动——非 AI 背景的领导者将掌舵 Meta AI 战略(207 条评论)

MongoDB CEO resigns to join Meta

受众观点:①@Insanity 质疑领导者没有 AI 背景却要领导 AI 战略;②@htrp 讽刺 Facebook 要把 Muse 卖给公司作为 AI 同事;③@fred_is_fred 对 CEO 能直接跳槽表示惊讶

展开评论
  • @Insanity (0): Interesting choice, his prior experience doesn't seem AI related yet he's going to lead an AI strategy at Meta?
  • @htrp (0): Facebook going to be selling Muse into your company as your AI coworker
  • @12982712 (0): Finally we'll get authentic webscale AI.
  • @fred_is_fred (0): I've always wanted model sharding.
  • @fred_is_fred (0): As a non-joke question - I am surprised a CEO can do this. Wouldn't there be contract terms about notice period, orderly transition etc?
HN H7
258 分 · 130 评 🔥

HN 帖子介绍 Parley 去中心化聊天网络,每人运行自己的实例通过 DNS 互相发现(130 条评论)

Parley: Federated, decentralised chat that speaks plain IRC

受众观点:①@davidcollantes 详细介绍 Parley 去中心化架构;②@shreddit 认为「这正是我一直在想的,就像 IM 版 email」;③@okwhateverdude 类比「XMPP 但用 IRC 替换 XML」

展开评论
  • @davidcollantes (0): Parley is a chat network with no centre. Every person (or team) runs a small instance for their own domain. Instances find each other through DNS and well-known identity documents, exchange signed messages over HTTPS, and present the whole federated network to ordinary IRC clien…
  • @myaccountonhn (0): I quite like this, but it feels like spam could quickly become a concern.
  • @shreddit (0): This is exactly what i was thinking about for the last few weeks (but am too stupid to implement myself). It’s like email just for IM…
  • @padolsey (0): Both cool and worryingly convenient for the botswarms we've been warned of...
  • @okwhateverdude (0): lol, so XMPP, but instead of XML, it is IRC. Alright, I dig it.
HN H12
151 分 · 85 评 🔥

HN 帖子深入讨论 PostgreSQL 的 AT TIME ZONE 语法陷阱——同一语法对不同类型行为完全相反(85 条评论)

Footguns with Postgres “at time zone 'UTC'”

受众观点:①@Felk 指出 AT TIME ZONE 用于两个相反操作导致混淆;②@ForHackernews 表示内部一致但需要理解设计才能正确使用;③@ulrikrasmussen 指出 SQL 标准的 timestamp 类型本身有设计缺陷

展开评论
  • @Felk (0): I think most of the confusuion comes from "AT TIME ZONE" being the syntax for both the conversion from _and_ to timezone'd timestamps. Maybe it would have been more intuitive had the two operations gotten different wordings, e.g.: timestamp to timestamptz: AS ZONED AT TIME ZONE…
  • @ForHackernews (0): Postgres 'at time zone' is confusing, but it's internally consistent. Once you understand what it's doing, you can work with it and it will reliably behave the way it's designed to. https://oneuptime.com/blog/post/2026-01-25-postgresql-timezo... -- TIMESTAMPTZ AT TIME ZONE 'X' -…
  • @ulrikrasmussen (0): The SQL standard is unfortunately really horrible when it comes to handling of time. The type `timestamp` is not a timestamp at all because it doesn't encode a unique point in time, it just stores a date and a time which has to be interpreted relative to a timezone. It should be…
  • @estetlinus (0): Makes me wonder what kind of footgun we are talking about here — is it classic, smoking, standard or vanilla?
  • @layer8 (0): > Adding a month with + INTERVAL '1 months' is timezone-dependent. […] Adding months isn’t well-defined anyway, even when using date, for days of month > 28. I think it’s a mistake that systems generically allow such a computation (as opposed to application code implementing dom…
HN H16
107 分 · 51 评 🔥

HN 帖子分享用 M5Stack PaperMono 电子墨水屏做冰箱磁贴购物清单的硬件 DIY 项目(51 条评论)

Fridge magnet shopping list on an M5Stack PaperMono (ESP32-S3, e-ink touchscreen), synced with a phone web app over Wi-Fi. Works offline, ~2,400 lines of C++. This is a great wee device! BLE, Wi-Fi a…

受众观点:①@lucasrufkahr 调侃「等一下要给冰箱磁贴充电」;②@hn1rig3rak 问电子墨水屏局部刷新实现方案;③@ex1fm3ta 感慨 AI 让定制固件变得可行

展开评论
  • @lucasrufkahr (0): Hold on babe I gotta charge my fridge magnet… Jokes aside this is cool
  • @jader201 (0): At first I thought “that’s a very specific shopping list for fridge magnets”. :)
  • @hn1rig3rak (0): Curious what you landed on for refresh, full flash or partial? Ticking off items one by one is exactly where partial update artifacts pile up on these panels.
  • @ex1fm3ta (0): Project like this make me excited with AI. You can pretty much custom firmware, tailored for your specific needs. Thanks for this repo. Quick question to the community : why does nobody makes afordable huge e-ink displays ?
  • @edoceo (0): I was just working on something similar. Was about to buy the xteink for my display. You've accelerated my progress. How's the battery life?
HN H15
109 分 · 49 评 🔥

Scrimba(YC S20)创始人 Show HN,将 LLM 接入交互编程视频格式,支持生成可嵌入内容的编程解释器(49 条评论)

Hi HN, I’m Per, founder of Scrimba (YC S20). We’ve spent the last decade teaching people how to code with an HTML-based video format. We’ve now plugged an LLM into it, so that people can create expla…

受众观点:①@DANmode 开玩笑讨厌这个趋势但已推荐给有 ADHD 的朋友;②@alentodorov 说本来要批评但这个产品确实有趣;③@harvey9 用递归玩笑提示产品本身可嵌入 HN

展开评论
  • @DANmode (0): Thanks, I hate it. Kidding, I already sent it to a friend with ADHD who has been struggling to remain anchored to the tech world in any way besides Shorts. But, culturally, I definitely hate the trend it implies! Neat. Thank you for sharing.
  • @alentodorov (0): loved scrimba. thanks for building that. was anout to pushback bc of the torrent of show hn sloppy pists but this one is fun.
  • @babu_mick (0): yo this is wild
  • @mgxplyr (0): It's actually very good...
  • @harvey9 (0): Please, nobody click on the link to this hn item within hn.watch as it would be even more dangerous than typing 'google' into Google
HN H14
127 分 · 46 评 🔥

HN 帖子讲述开发者在 Google Play Store 遭遇审核流程崩坏——被附上不明 App 截图拒绝,审核速度大幅下降(46 条评论)

So long Google, and thanks for all the nudes

受众观点:①@nine_k 指出审核流程损坏:App 数量庞大,深度审核成本高,错误拒绝成本低;②@bennyp101 分享等待超一个月的审核经历,猜测 vibe coding 带来的 App 爆炸加剧了积压;③@delichon 表示打算放弃 Play Store 转 F-Droid 和 itch.io

展开评论
  • @nine_k (0): The review process is broken. The stream of apps is enormous, the expense to review them deeply would be large, the cost of a false rejection is very low.
  • @spwa4 (0): I wonder if that'll be the future. This is something AI can actually solve. Write replacements for a lot of apps, slop but efficient slop, and use that as a distraction-free phone. Maybe that could even be extended to browser, OS, webpages (just have a bad AI rewrite every page…
  • @delichon (0): > I was hoping to publish my new game on the store and charge for it, but their review process is so broken, I think I'll stick to fdroid and itch.io. The more you tighten your grip, Tarkin, the more star systems will slip through your fingers. -- Princess Leia
  • @thih9 (0): > it was rejected. They attached a nsfw screenshot of some unidentified app. I don’t doubt the author, at the same time this is a serious accusation with no details and little context. This could have been presented in a more organized way, with timelines and redacted content. P…
  • @bennyp101 (0): Yea I’m on over a month now of an app sitting in review, a couple of years ago it was much quicker. Apple approved the same app within a week. I can only assume it’s because it’s easier to vibecode android apps so it’s being flooded?
HN H13
128 分 · 33 评 🔥

HN 帖子分享如何通过 DNS 拦截和 nginx 代理让 PS5 把 RTMP 直播流推到自定义服务器(33 条评论)

Hijacking the PS5's RTMP stream

受众观点:①@charcircuit 指出 Sony 再次出现加密错误,这次是完全不加密的流量;②@tvbusy 解释了 PS5 用 DNS 解析 RTMP 目标服务器的具体机制;③@mixdup 指出文章在关键逻辑步骤间有跳跃

展开评论
  • @mixdup (0): Maybe I'm missing something but there seems to be a gap between "figure out the REAL hostname" and "we no longer have to worry about the stream never showing up on YouTube"
  • @charcircuit (0): Another crypto mistake by Sony this time leaving traffic completely unencrypted.
  • @tvbusy (0): PS5 uses DNS to resolve where to send RTMP stream. The author found out that they cannot spoof the main server since that server uses TLS and it's not possible to make the PS5 trust self made certificate. The author then runs a real Twitch stream and monitors DNS requests and fi…
  • @Muromec (0): Don't mind me, I'm just sitting here with my HDMI-RX port on rk3588 being happy that is works and I didn't brick the board when doing uboot update to uncurse it.
  • @tamimio (0): You can combine the last two steps the nginx and mpv with obs and gstreamer, or just gstreamer really.
336 赞 · 26 评 · 6.2k 阅 🔥

Better Auth 开源认证库周下载量突破一千万次,成为近期增长最快的 JS 认证解决方案之一

Better Auth is now at 10m weekly downloads! https://t.co/I1Dag5BMpG

受众观点:①@fuma_nama 简短祝贺反映 JS 生态圈认可 ②@BenceRedmond 说 better auth is incredible 代表使用者正向反馈 ③@MarcLaventure 感叹式回应显示数字超预期

展开评论
  • @fuma_nama (3): @bekacru Nice
  • @MarcLaventure (2): @bekacru YO LETS GO!
  • @BenceRedmond (1): @bekacru congrats! better auth is incredible
  • @itsnoahd (1): @bekacru Huge congrats!
  • @natyt_am (1): @bekacru Congrats beka 👏👏👏
262 赞 · 58 评 · 7.8k 阅 🔥

@robj3d3 分享其视频课程在发布 3 天内获得 193 位用户下单、累计 4 万美元预售金额的里程碑

$40,000 of pre-orders in 3 days. Thank you for trusting me, it means the world 🥹❤️ https://t.co/KuUs6g7HDN

受众观点:①@benjaminakar 说 it takes years then happens all at once 指出爆发背后是多年积累 ②@robj3d3 自己揭示社区附加权益的预售价值堆叠策略 ③@Goprogabriel 的 well deserved 反映圈子对长期积累的认可

展开评论
  • @robj3d3 (11): This is for my upcoming video course The Attention Playbook. 54 video lessons on how to get your posts seen, and turn that attention into customers. Pre-orders get the SEO module as an exclusive bonus 😊 https://t.co/blwM27oDNG
  • @robj3d3 (7): Like actually, it means the world. Giving $200 to someone online is big, and even more when the course is 4 weeks away. But 193 of you did it anyway 🥹 Thank YOU. I've got some big things coming for the 193 💪 (community stuff, free for pre-order fam :D)
  • @benjaminakar (5): @robj3d3 congrats rob! it takes years, then happens all at once. love to see it
  • @Goprogabriel (3): @robj3d3 This is so damn cool - and seriously well deserved. Huge congrats!
  • @launch_llama (3): @robj3d3 Mate that is fucking insane !!!
231 赞 · 31 评 · 15.5k 阅 🔥🔥

Fable 5.1 和 Opus 5.5 相继发布,开发者圈热议 OpenAI 是否会在本周发布新模型以应对竞争压力

OpenAI is cooked if they don't release a model this week. Fable 5.1, then Opus 5.5, now this...

受众观点:①@robj3d3 指出 Opus 5.5 在大多数任务上目前优于 Astra,实测数据驱动 ②@brannonhogue 担心 Codex 受冲击代表 OpenAI 工具链依赖者的焦虑 ③@TheStormDev 相信 OpenAI 会回应反映社区对竞争节奏的关注

展开评论
  • @robj3d3 (12): Opus 5.5 performs better than Astra on most things atm btw
  • @brannonhogue (3): @robj3d3 Codex may be cooked 😭
  • @asyncaman (2): @robj3d3 I think Dario is in full speed after announcing AI is dangerous for humanity
  • @TheStormDev (2): @robj3d3 OpenAi is gonna deliver, my spidey senses are tingelling
  • @fireplyai (1): @robj3d3 three model drops in one stretch and everyone's takeaway is still about who's losing
112 赞 · 15 评 · 8.4k 阅 🔥

@codyschneider 分享用精准匹配域名加 Claude Code 一键生成目录站、部署到 Vercel 后获取免费 SEO 搜索流量的低成本增长策略

exact match domain + a one shot directory site is still the cheapest traffic in marketing buy the exact match domain for what your customer searches have claude code one shot a directory site for it…

受众观点:①@itsthedonhashim 指出定期更新目录内容才能维持流量补充执行细节 ②@realAndyAustin 问是否可用非 .com 域名说明受众在认真考虑落地 ③@BertrandDiouly 问怎么填充内容触及执行最核心难点

展开评论
  • @itsthedonhashim (2): @codyschneider @codyschneider it's a solid plan. if you update the directory regularly, it'll keep the traffic fresh and relevant.
  • @Devanshuai (1): @codyschneider Correct
  • @realAndyAustin (1): @codyschneider That’s awesome. Can it be on non .com domain?
  • @BertrandDiouly (0): @codyschneider what content / sites would you populate it with though?
  • @l3d1c (0): @codyschneider Wow, I've never seen anything with thousands of clicks even though it's at position 23
109 赞 · 32 评 · 8.2k 阅 🔥

@robj3d3 分享一行 claude --model claude-sonnet-5-5 命令调用,评论区用停电玩笑回应,展示了 Claude CLI 的快速上手魅力

Run this in your terminal, you're welcome: claude --model claude-sonnet-5-5 https://t.co/JS3MURQt7P

受众观点:①@siyabuilt 用停电梗回应评论区幽默氛围活跃互动门槛低 ②@BertrandDiouly 说 that's a sneaky little sonnet 暗示命令有隐藏意味 ③@itsAndrewSalzer 正经感谢代表有学习需求的受众

展开评论
  • @siyabuilt (4): @robj3d3 Thanks, entire neighborhood electricity gone now https://t.co/CYZm3yX7Ye
  • @ibocodes (2): @robj3d3 that's why my neighbor is paying for internet
  • @_sslinNn (1): @robj3d3 hakerman
  • @itsAndrewSalzer (1): @robj3d3 Amazing. Thanks for posting!
  • @BertrandDiouly (1): @robj3d3 that's a sneaky little sonnet : )
103 赞 · 31 评 · 5k 阅 🔥

@robj3d3 对比 Sonnet 5.5 和 Opus 5.5 在生成 SuperX 品牌动效视频任务上的表现,Sonnet 无法遵循品牌名指令而 Opus 5.5 完整理解了任务

Sonnet 5.5 max effort "make a dynamic 15-second motion graphics video for SuperX that shows what an incredible motion designer you are, like it's your showreel for a resume. go all out." https://t.co…

受众观点:①@robj3d3 自己对比了两个模型结果并附链接提供第一手对比数据 ②@launch_llama 说 That's actually insane still 反映即使失败案例也足够震撼 ③@dttsnow 评价 way too fast 揭示用户对 AI 生成内容审美偏好的差异

展开评论
  • @robj3d3 (7): It completely ignored the "SuperX" instruction and made it all about how great of a motion designer it was Opus 5.5 actually understood the assignment: https://t.co/18vpTPvjSw
  • @robj3d3 (1): SuperX is the best way to grow on X btw Even though Sonnet didn't understand, it's free to try ;) https://t.co/zIsZOZDLyy
  • @David_cbolt (1): @robj3d3 SuperX more like SuperClean
  • @launch_llama (1): @robj3d3 That’s actually insane still
  • @dttsnow (1): @robj3d3 way too fast for my taste, but impressive nonetheless
83 赞 · 9 评 · 6.3k 阅 🔥

冷邮件基础设施成本拆解,每月约 125 美元配合 Hypertide 和 Instantly 两个工具可发 1 万封邮件,1000 个收件箱规模下冷邮件从群发渠道演变为精准触达工具。

cold email infrastructure is way cheaper than you think about $125 a month in infrastructure gets you ~10,000 cold emails a month we use hypertide for inboxes and instantly for sending we have about…

受众观点:① @MarkSMcDaniel 质疑 B2B 采购决策者真的会回复冷邮件吗;② @mattduhon 分享 Resend + MillionVerifier 替代工具栈;③ @Michaelfrscott 关注千级收件箱规模下的回复管理挑战。

展开评论
  • @rishilautomate (1): @codyschneider 1,000 inboxes.. bro built an email factory
  • @misterrpink1 (1): @codyschneider shit you might be right here
  • @MarkSMcDaniel (1): Does anybody respond to cold emails? Serious question, not a troll. I block every cold email I receive to my work account, don’t most people? Rather, don’t most people who have budget control to buy what you’re selling. My question assumes this is B2B, I could see cold email wor…
  • @mattduhon (0): @codyschneider My stack has been Resend, a bunch of domains, and MillionVerifier. Works awesome.
  • @Michaelfrscott (0): @codyschneider How do you manage reply’s at that level
81 赞 · 6 评 · 4.9k 阅 🔥

分享一批被认为能持续产生百万级收入的 App 创意清单,评论聚焦于习惯养成类 App 因强进阶循环机制变现效率更高,以及创意看起来容易执行门槛极高的现实落差。

$100M app ideas that will always print https://t.co/iluDyiFB5a

受众观点:① @talhaaeth 分析习惯养成类 App 因强进阶循环变现效率高于纯内容类;② @arthurclsn 点出"always print until you try to ship one"的执行壁垒;③ @shashaditya 补充语言学习赛道竞品格局。

展开评论
  • @AllianceDouble (1): @ErnestoSOFTWARE Not logged in · Please run /login
  • @talhaaeth (1): @ErnestoSOFTWARE Ladder pulling $7M/mo off 300K downloads is the one that stands out here. Way higher revenue per download than the rest on this list, habit apps with a hard progression loop seem to monetize better than pure content ones.
  • @itsjahmills (0): @ErnestoSOFTWARE printing niches
  • @arthurclsn (0): @ErnestoSOFTWARE always print until you try to ship one
  • @shashaditya (0): @ErnestoSOFTWARE Speak does more B2B, for Language Learning, check out Learna AI
60 赞 · 17 评 · 4.5k 阅 🔥🔥

仅含单个颜文字符号的推文引发"你是从未来来的吗"式围观,结合 12 条集群内容判断,疑似提前预示了当日某重大 AI 发布公告,引发对圈内信息流通路径的关注。

(°ᯅ°) https://t.co/6oongXasFf

受众观点:① @moonfarm_dev 追问"Wait, are you from the future?"定下评论串调侃基调;② @idanmasas 推断该推文可能促使官方提前宣布;③ @stemonteduro 评价"bro is ahead"体现社区对预知者的崇拜情绪。

展开评论
  • @moonfarm_dev (3): @robj3d3 Wait, are you from the future? 👀
  • @idanmasas (1): @robj3d3 they probably saw your tweet and decided it was time to announce it themselves
  • @stemonteduro (1): @robj3d3 bro is ahead
  • @David_cbolt (1): @robj3d3 Ahead of the game 🚀
  • @jayypatel18 (0): @robj3d3 bro tweeting this while traveling back to our time https://t.co/0XI3be6gm3
54 赞 · 9 评 · 2.8k 阅 🔥

一人订阅 claude code Max 可完成过去 25 人营销团队的工作量,作者命名为「营销工程化」,覆盖广告创意研究、素材制作、媒体购买、关键词研究等中间环节的 AI 自动化替代。

one person with a claude code max subscription can do what used to take a 25 person marketing team we call it marketing engineering it's all the middle work a marketer used to do by hand ad creative…

受众观点:① @SeldomSolemn 指出方案只有在已有清晰 brief 时才跑通,瓶颈在创意品味和媒体购买判断而非执行;② @denogrowth 提到 Opus 5.5 动态图形更新扩展了这条路线的可行性边界;③ @onmanjrekar 作为 solo builder 表示正在亲身实践。

展开评论
  • @SeldomSolemn (1): @codyschneider I've been the one-person shop trying to replace that middle work and it only sticks when I already know the brief. Where does yours fall apart first, creative taste or the media buy?
  • @denogrowth (1): @codyschneider That opus5.5 motion graphics update is what I needed
  • @onmanjrekar (0): @codyschneider Living this as a solo builder!
10 赞 · 7 评 · 2.5k 阅 🔥

黎巴嫩裔开发者 Ahmad 分享用 AI 电话代理服务创业获得五万美元收入的个人经历

here's my take on this. My name's Ahmad, Lebanese 🇱🇧 dev based in Montreal 🇨🇦. studied computer science, then started an agency setting up AI phone agents for businesses. made $56,352 from it. then p…

受众观点:①@itsAndrewSalzer 表达对 Ahmad 认可,关注独立创业成功故事

展开评论
  • @ahmadafterhours (1): Try for free today. No card required. https://t.co/STU2GFlddx
  • @itsAndrewSalzer (1): @ahmadafterhours Nice to meet you Ahmad! Proud of you!
  • @im_mohieb (1): @ahmadafterhours Yoo super refreshing seeing other Arabs here, I'm from Jordan, connected!
  • @MahdiEzz_code (1): @ahmadafterhours Fellow lebanese 🤝🏻 Let’s goooo
  • @buildwithmaya (1): @ahmadafterhours hey nice to meet you! I am 99% sure I have the same exact couch but in white/cream
Reddit R23
313 分 · 328 评 🔥

r/webdev 高热帖:开发者分享 Claude 让编程失去乐趣,问是否有人回归手写代码(328 条评论,score 313)

Using Claude has been an amazing experience but it took all the fun out of coding as a full-stack dev. Are any devs reverting back to coding by hand with less usage of AI?

受众观点:①@apt_at_it 作为 7 年 senior 表示在写 prompt 比直接改代码慢时选择手写;②@jam_pod_ 发现 prompt 描述比代码本身更长时就直接自己写了;③@No_Yam1114 分享在新项目用 AI 3 个月后发现自己什么都没学会

展开评论
  • @os_gross (209): I do that when I am setting up the base of the project or learning a new thing. Everything else I can delegate to ai and am making sure it follows patterns, it is usually CRUD, which I would be able to write even with dementia
  • @apt_at_it (208): Senior here with a 7 YoE. Mostly backend and distributed systems. Professionally? No, not really. I do catch myself prompting when actually making the change myself would be faster, though. That pisses me off. Personally is a different story. I can’t justify spending money on AI…
  • @AlaskanDruid (77): Yep. There are still programmers around :) no reverting. Just stayed as programmers.
  • @jam_pod_ (71): I often realize halfway through writing a prompt that the description of what I need the code to do is longer and more detailed than the code change itself would be. That’s when I cancel out of Claude and just go write the code
  • @No_Yam1114 (66): I am in new project with a new stack, and with AI after 3 months I realized I learned nothing and know nothing about it, so I stopped using AI (not 100%, but I write most of code manually). I began to enjoy my work again
Reddit R34
110 分 · 253 评 🔥

r/selfhosted 高热帖:用户担忧自托管工具维护者开始用 AI 生成代码导致质量下降,以 TriliumNext 为例(253 条评论,score 110)

I will give a concrete example of the selfhosted tool I use the most, my note taking app TriliumNext. I liked trilium in the past. I even switched to triliumNext when the project was archived. But ma…

受众观点:①@Drehmini 分享和 20 年经验 senior 开发者的对话,他们不写任何非 AI 代码,且 TriliumNext 已转为闭源;②@Peruvian_Skies 指出不是所有 AI 编码都是垃圾,关键是信任;③@saltyourhash 作为 20 年老兵分享:每行代码不亲手写但亲手 review,同时用 AI review

展开评论
  • @Drehmini (291): > Is this opensource now ? This is close-sourced now too. I've talked with senior developers that have over 20 years of programming experience where multiple have said that they haven't written a line of code in over a year.. AI definitely has a place in programming and it's…
  • @Peruvian_Skies (257): Not all AI use in coding is slop production. In the end using software requires trust in the authors. AI is a shitstorm but there's no putting the toothpaste bacv in the tube now that it's out. We have to learn to live with it.
  • @CoffeeStraight2249 (40): This is exactly it. I used to program without AI. Now I use it to assist me, but I look over every line of code it writes. Half the time I'm talking to it in a browser not a CLI. This is so I can audit everything it does, but it allows me to work on multiple projects at once, an…
  • @GolemancerVekk (28): I really doubt it, for the simple reason we don't have the resources for keeping up that level of vibe coding. Assuming for the sake of argument that there's a breakthrough in the cost & energy ratio and that we manage to side-step massive economical, ecological, social etc.…
  • @saltyourhash (24): I've been programing for 20+ years, been a lead at multiple fortune 500s and multiple teams front end, back, full stack. I don't write code by hand, but j dk review it by hand and by AI. But all lines are have been hand checked by me, even on personal projects. I have been doing…
Reddit R45
851 分 · 201 评 🔥

有人用 Claude Opus 5.5 从单个 prompt 写了约 25000 行 JavaScript,完整重制初代《口袋妖怪》红版,无图片资源,全按像素绘制,含 224 张地图和精英四天王,评论区主要争论 Nintendo IP 侵权和 AI 生成代码版权归属问题

**Play it (desktop or phone):** [https://claudered.dev](https://claudered.dev) **How it was built (7 min):** [https://youtu.be/7G3PPKV0iHA](https://youtu.be/7G3PPKV0iHA) **Source:** [https://github.c…

受众观点:①AI 生成代码版权归属问题(@Wide_Smoke_2564 提出代码版权属于 Anthropic 所以 Nintendo 管不着)②AI 能否完成完整工程量级产品而非只做原型③Nintendo IP 法律风险对 AI 游戏开发生态的连带影响(@Original-League-6094 担忧 AI game dev 就此终结)

展开评论
  • @JLP2005 (174): Hope you have good lawyers (who tell you to cave immediately)
  • @JDotDDot (172): Not a bear I would personally poke, but well done
  • @Wide_Smoke_2564 (125): No no. Claude wrote it so the code actually belongs to anthropic who are immune to IP theft. Checkmate Nintendo
  • @peensmith_ (71): Buddy actually put the name of his company on it after stealing IP from one of the most notoriously litigious companies in existence. 200iq
  • @Original-League-6094 (52): Whelp, AI game dev was good while it lasted. OP just got Anthropic sued into bankruptcy.
Reddit R43
1083 分 · 176 评 🔥🔥

r/ClaudeAI 高热帖:Claude Sonnet 5.5 官方发布,定位为 Opus 5.5 的更快更低价补充,首个具有网络安全防护的 Sonnet 模型(176 条评论,score 1083)

Sonnet 5.5 is a faster, lower-cost complement to Claude Opus 5.5, strongest at well-scoped everyday tasks, fixing bugs, and creating polished documents, slides, and spreadsheets. It also has a strong…

受众观点:①@urchir 对网络安全防护措施延伸到 Sonnet 感到不满;②@snarfi 调侃 Sonnet 被训练成「代码猴子」;③@ClaudeOfficial 发布了模型能力对比图

展开评论
  • @urchir (281): >It's also the first Sonnet model with cybersecurity safeguards similar to those on our most capable models. Oh for Christ's sake.
  • @ClaudeOfficial (238): https://preview.redd.it/bbkjgthjwash1.png?width=2160&format=png&auto=webp&s=1759504594f84ed14a4471cfd860e870eb2b9de5
  • @snarfi (82): They really trained sonnet to be the code monkey
  • @Ok-Clerk393 (78): https://preview.redd.it/2kqpnsa7zash1.png?width=416&format=png&auto=webp&s=85f8f70f8f9a929c40ffb658c6d49a2e91c4eb7c
  • @Appropriate-Disk-371 (64): oooo....ahhhh....
Reddit R44
945 分 · 107 评 🔥

r/ClaudeAI 高热帖:用户每天重跑 Claude Opus 5.5 基准测试追踪是否被 nerf,创建「LiveNerf」实时监控项目(107 条评论,score 945)

I’ve been tracking Opus 5.5 since day one to see if Anthropic is nerfing models. Everyday, I re-run its independent benchmarks like GPQA or SWE-bench and then keep track of it in a graph on the repo.…

受众观点:①@HatPitiful3081 建议做成像 downdetector 一样的盈利网站;②@TheOnlyVibemaster(作者)表示不打算变现专注透明度;③@AnonThrowaway998877 说这是真实需求早有人在问

展开评论
  • @HatPitiful3081 (332): You could easily make a monetized site out of this, this is genius. It would be like [downdetector.com](http://downdetector.com) but for ai models getting nerfed. Not only would you make good money, you would expose these greedy ai companies for the whole world to see, they woul…
  • @TheOnlyVibemaster (141): I’m not interested in monetizing it, but someone definitely could. I’m more interested in transparency than trying to monetize transparency. I’ve seen other people making tools like this and monetizing it though. I just want everyone to see how I’m getting the data and propose h…
  • @AnonThrowaway998877 (123): Haha I was just asking if this existed yet a few days ago. Someone replied with another site that is indeed doing this already. But the more, the better. Here is the comment: https://www.reddit.com/r/ClaudeCode/s/8x46aFkAeH And here's the site: https://marginlab.ai/trackers/clau…
  • @Fabulous-Sale-267 (77): This man is doing the Lord’s work. Thank you for your service.
  • @PrestigiousTrick1002 (37): So now we need an aggregate site
Reddit R22
581 分 · 98 评 🔥

r/webdev 高热帖:开发者感慨在 AI 时代庆幸自己先学会了编程和问题解决思维,担忧 junior dev 直接用 AI 但不理解逻辑(98 条评论,score 581)

Web development is like 2% of my job. I work in a dinosaur industry. We still use VB6, C, Fortran and pascal.. slowly migrating away from VB6 to .net 5 and that’s considered modern for us. The rest w…

受众观点:①@AfricanTurtles 指出比代码本身更重要的是 problem solving 能力,用 AI 的 junior dev 没有这个能力;②@papa-hare 担忧替代 junior 的公司未来没有 senior;③@WingZeroCoder 观察 senior engineer 因 vibe coding 文化正在离开行业

展开评论
  • @AfricanTurtles (194): I think more important than the code itself, is that you learned to PROBLEM SOLVE before AI. More and more I see companies hiring junior devs and letting them loose with AI, but they don't have a clue how to reason about problems, compare solutions, figure out what the tradeoffs…
  • @papa-hare (160): Every profession that's replacing/augmenting their juniors with AI is pretty much screwed. It's gonna be fun when tokens are gonna cost an arm and a leg and the seniors (who are now juniors) won't know what to do.
  • @Sea-Journalist-2741 (40): The logic gap is real, I've seen it try to bullshit its way through stuff my 12 year old nephew could spot. The junior pipeline being cut is gonna sting in about a decade when the people who actually know why things break are all retired
  • @0ba78683-dbdd-4a31-a (30): This. I remember thinking "you're a react dev but don't actually understand basic js" was bad enough before vibe coding became a thing.
  • @WingZeroCoder (30): Even worse, we’re seeing a growing trend of senior engineers just flat out leaving the career entirely because of a shared frustration that blind vibe coding is the last straw. So that’s going to happen sooner than later.
Reddit R50
196 分 · 95 评 🔥

Nvidia 提议在每颗 AI 芯片旁配置看门狗芯片监控 AI 系统行为,引发是真正 AI 安全方案还是新一轮厂商锁定策略的广泛讨论,评论区以 GSync 历史为例表达对 Nvidia 垄断意图的强烈质疑

This seems like quite a good solution. What do you guys think?

受众观点:①@VehiculeUtilitaire 一句话戳破商业逻辑:卖芯片的人说你每买一块就得再买一块②@JDotDDot 以 GSync vs FreeSync 历史论证 Nvidia 锁定方案的长期风险③@clearlight2025 提出关键分叉点:协议开放则是好事,Nvidia 私有控制则是另一场锁定

展开评论
  • @VehiculeUtilitaire (395): Guy selling chips says you need to buy one more chip for every chip he sells you
  • @JDotDDot (109): Y'all remember when Nvidia was putting their gsync chips in monitors and causing them to cost $800+? And then AMD came along and said "nah" and eventually forced Nvidia to adopt their solution, creating the whole "gsync compatible" situation? And now you can get large, high res,…
  • @ogaat (63): When you only have a hammer, every problem is a nail.
  • @clearlight2025 (35): As long as it’s a public chip and protocol available to all and not controlled only by nvidia, it sounds like a good idea. I’m skeptical of their intentions though.
  • @UAP44 (27): >This seems like quite a good solution. Seems more like, ah, so that is how everyone will eventually stop trusting Nvidia with anything.
Reddit R48
275 分 · 91 评 🔥

开发者求教如何让 Claude 生成高质量而非千篇一律的 AI slop UI,评论区达成核心共识:必须给 Claude 参考设计,你是导演不是 prompt 狂轰滥炸者

I’ve been building websites and I keep running into the same problem. Claude is actually pretty good at structure, UX flow, information architecture, and getting a functional dashboard together. But…

受众观点:①@mitterb 给出最高赞实操方法:找参考设计让 Claude 仿照布局加品牌色替换②@QueenSavara 提炼核心认知:做导演不是 prompt 狂轰滥炸者,给明确方向不要让 AI 自由发挥③@ChaseVsGodzilla 推荐具体工具 impeccable.style 作为实际解决方案

展开评论
  • @mitterb (106): You find a website you want to replicate from dribble or awwwards.com and then ask Claude to replicate X website's layout with your brand colors and identity. And then sometimes I ask it to use it's frontend design skill to elevate the reference design. Usually get pretty sick r…
  • @QueenSavara (77): You don't tell it to build the UI and let it do what it wants. You either give it reference or direct it. This never changed: if you want direct results act like director not prompt spammer.
  • @mitterb (59): Yes exactly. OPs strong suit isn't creativity or else he wouldn't have asked this question. So telling him to be creative rather than tell him what would actually be useful for him would not make sense.
  • @ChaseVsGodzilla (48): Give [impeccable](https://impeccable.style) a try, it’s been a game changer for me
  • @momo1083 (19): Bingo. You are Scorcese and you have in front of you DiCaprio, DeNiro and Pacino. They would also generate slop acting were it not for a great director.
Reddit R46
330 分 · 81 评 🔥

零编程基础玩家用 Claude 等多 AI 工具 vibe coding 5 天做出 3D 赛车游戏,含 6 条赛道、40 辆赛车、道具系统,3D 模型由 AI 生成并用 Claude 驱动 Blender 脚本修正,音乐来自 Suno,评论区也主要围绕 Nintendo IP 风险

Last time I posted this, a few of you pointed out that it's AI slop, that the ramps were optional, and that Nintendo's lawyers are on their way. Fair points, all of them. So here's an honest trailer.…

受众观点:①@TechTuna1200 以直升机梗戏谑 Nintendo IP 法律风险,折射 AI 游戏开发的普遍焦虑②@maxle100 认为 AI 让游戏开发门槛大幅降低,专业团队 6 个月能做出优质产品③@karlfeltlager 给出具体 IP 合规建议:不用商标角色、字体、音乐或代码

展开评论
  • @TechTuna1200 (104): If you can hear helicopters circling around your home, it's Nintendo and their army of lawyer defending their IP. Pro tip: keep some distance from your front door, put your hands behind your head and don’t move.
  • @maxle100 (27): if some dude can vibecode this give a team of people who are professionals around 6 Months and some good storywriting and they can build incredible games
  • @Fantastic-Jeweler781 (19): Relax, The helicopters are mine, I use them for the camera angles
  • @StrobeWafel_404 (9): lol Nintendo about to shut this down so hard
  • @karlfeltlager (8): Don’t use any of the trademark characters, don’t use the font, don’t use the music, don’t copy the code.
Reddit R42
1492 分 · 75 评 🔥🔥

r/ClaudeAI 高热帖:用户整理 Anthropic 官方 Claude Opus 5.5 提示词指南核心要点,包括反直觉的 time budget 技巧(75 条评论,score 1492)

We know Opus 5.5 is good but did you know Anthropic released an official prompting guide?? Were we all to busy making animated videos ??? I took 5 minutes to read this and then summarise the informat…

受众观点:①@st_malachy 幽默质疑「看看周围」比「仔细思考」更好;②@optima-pacifist 亲测 time budget 技巧确实让 agent 在长 refactor 时不提前结束;③@EzoLabsInc 追问是否真的遇到过 agent 提前停止的问题

展开评论
  • @st_malachy (100): Don’t tell it to “think carefully” tell it to “look around first”. Brilliant. /s
  • @optima-pacifist (75): the time budget one was new to me. telling it roughly how long a task should take stopped it wrapping up early on long refactors
  • @dynamoney (72): The elapsed time thing feels weird to me, even if it works: How would the user know, if an agent team would need 340s or 1200s to finish a task? And given that output speed after TTFT is constant (?), wouldn‘t it be more meaningful to set a token budget?
  • @EzoLabsInc (45): Point 4 about agents stopping early sounds like the bigger gotcha than the guardrails thing. Have you actually hit that where it quits before finishing the task, or is it more that the API docs warn about it happening
  • @Makemeacyborg (32): It does sound ridiculous and I don’t think they literally mean say that verbatim. But at least look around is a concrete action unlike think carefully.
Reddit R12
191 分 · 70 评 🔥

开发者用 Codex 和 Claude 构建离线免费 Photoshop 替代品 Photon Studio,从实验到 25000+ 用户的 build in public 经历(70 条评论,score 191)

I built a free, offline photoshop alternative called [Photon Studio](https://tenzen.studio/photon) , which started as an experiment. Then I posted about it [here on reddit](https://www.reddit.com/r/v…

受众观点:①@vitalityy77 指出 25k 用户做免费工具很疯狂,关心捐赠模式可持续性;②@AsejereDaDeje(作者)解释 Codex 和 Claude 完成了 90% 代码;③@vitalityy77 指出本地运行解决服务器费用但维护费用依然存在

展开评论
  • @vitalityy77 (44): 25k users on a free Photoshop alternative is wild. The donation model is interesting too, but I’d be really curious to see if it can stay sustainable at that scale.
  • @AsejereDaDeje (13): will do today
  • @vitalityy77 (12): Running locally solves the server bill, not the maintenance bill. Dev time, bug fixes, updates, support, hosting the site/downloads, etc. still cost money. That’s what I meant by sustainable.
  • @AsejereDaDeje (11): yes, that's why I didn't copy any code. I built the rendering engine on my own (well, codex and claude did 90% of the job). I chose a different tech stack from adobe that enabled me to easily make it portable to linux too. so now linux, windows and mac users can use it.
  • @Revolutionary_Lie590 (8): Can u add Arabic support please
Reddit R57
27 分 · 54 评 🔥

独立开发者靠 TikTok 有机流量获得第一笔 App 销售后求问增长策略,评论区高赞给出反直觉建议:先联系那一个买家,专门问差点让你不买的原因是什么

Hey everyone, I’m absolutely thrilled right now. I just got my very first sale for my app (a lifetime purchase from a user in the US!). It’s a huge milestone for me, and seeing that dashboard update…

受众观点:①独立开发者关注第一笔销售后如何做决策——@Lolorenzo08 给出了一个渠道打穿再说的反直觉建议 ②关注 ASO 的介入时机——ASO 是放大器而非启动器 ③关注单渠道专注 vs 多渠道分散的实际取舍逻辑

展开评论
  • @Lolorenzo08 (5): *Congrats, the first sale is the one that matters most. Before adding channels, do the unsexy thing: message that customer and ask what made them buy and where they found you. One sale is your first real signal and most people ignore it to go spray more content. And at one sale…
  • @lvfj2304 (2): Thank you! You make a really solid point about not spreading myself too thin. Reaching out to that first buyer seems to be the recurring advice here, and you're totally right—it's the first actual data point I have. I'm going to figure out their exact journey first before making…
  • @Lolorenzo08 (2): *Love it, that's exactly the right move. One thing when you talk to that buyer: don't just ask how they found you, ask what almost stopped them from buying. That "almost didn't" answer is gold, it's the precise objection to kill in your next TikToks. Good luck, come back and sha…
  • @91ElonMusk (2): Be friend with the customer. He/she will bring you more.
  • @ianwellerz (2): Congrats, first sale is huge
Reddit R53
112 分 · 51 评 🔥

独立开发者做了一个给 Mac 桌面添加下雨特效的小工具 Weatherling,周五晚发布两天内冲上 Mac App Store 工具类第 12 名,全程零权限申请

**Weatherling** adds rain to your Mac desktop. It's just a fun little app that I wanted to build to add a moody vibe while I work. People really seem to like it. I've never had an app get ranked in a…

受众观点:①独立开发者关注兴趣驱动小产品如何快速冲击 App Store 排名 ②用户关注隐私友好工具——@optima-pacifist 明确说无权限才是最好的功能 ③开发者关注氛围感/情绪类工具这个被忽视的细分赛道

展开评论
  • @alzho12 (14): You should make a snow setting for the winter fanatics
  • @AccomplishedArt1791 (7): this is so cool man, i will feature it in my newsletter :)
  • @MattSenter (6): yep! on it!
  • @optima-pacifist (3): congrats, #12 in two days off a friday launch is great. no permissions at all is honestly the best feature, i'd install a rain app long before one asking for screen recording
  • @koskinooo (3): Top 12 in Utilities off mostly Reddit is wild. Which subs actually moved the needle?
Reddit R47
314 分 · 44 评 🔥

独立开发者连续第九周更新 AI 辅助钓鱼游戏开发进度,已提交 Steam 审核,核心亮点是「Manager Session 调度 15 个专属功能 Session」的多 agent 协作架构,Opus 5.5 负责音效设计,200 美元订阅两周用完两次额度

Hello there again 🙃 I sometimes post updates on this subreddit and sometimes I skip it, anyway, previous updates: [Week 1 Progress](https://www.reddit.com/r/aigamedev/s/AQfqf5T4nY) [Week 2 Progress](…

受众观点:①@bitsperhertz 代表 build in public 连续更新系列的持续关注——每周进度帖的 accountability 机制本身就是增长策略②@MyCatH8ai 提出 Steam 独立游戏均值 300 美元的变现现实泼冷水③@RUSuper(作者)坦诚分享 token 消耗、不再用 Godot MCP 的工具迭代和上 Steam 的心态转变

展开评论
  • @bitsperhertz (24): I love these updates, keep them coming!
  • @MyCatH8ai (17): Looks nice to bad avg steam game makes 300 bucks now. I think of this as im sitting here working on a game.
  • @RUSuper (16): Honestly this started entirely as a hobby project because I had some free time and wanted to see how much "entirely AI made game" can be pushed. Obviously this is not anymore "entirely AI" as it required bunch of manual stuff, from manually positioning sounds and 3D models in Bl…
  • @New_Organization9898 (15): Looks a lot like DREDGE
  • @RUSuper (5): Inspiration was certainly from it. But luckily I stayed away from a lot of things to make it distinct enough. No horrors at night in a sense of paranormal. No fighting for space between equipment and fish, harbours, talents, fish mini game and fish rarity tiers, moral dilemmas e…
Reddit R32
266 分 · 42 评 🔥

python-roborock 维护者分享无需 root 即可自托管 Roborock 扫地机器人云端控制的方法(42 条评论,score 266)

I'm very big into local control and being able to host things myself. I've been one of the co-maintainers of python-roborock for a few years and I have had this goal for years to get my robot vacuums…

受众观点:①@headinthesky 分享自己因远程调用超时的问题,本地化控制能解决;②@tismo74 表示立刻去做并分享设备型号;③@Equivalent_Glove_664 对 onboarding 阶段 DNS 重定向的巧妙设计表示欣赏

展开评论
  • @headinthesky (36): This is awesome. I get frequent timeouts on mine which I assume is because of the remote calls and the blocking I do at a network level
  • @tismo74 (21): Well I know what I am doing this afternoon. Thank you for sharing. I have a Q5 Pro and while it doesn’t have a camera or a mic, at least I hope it doesn’t have one, I kept getting map dropouts from Home Assistant integration. Hopefully this solves that issue.
  • @MonsterMufffin (15): For anyone looking for a new robo vacuum, I would recommend getting something that supports [Valetudo.](https://valetudo.cloud/) It's awesome.
  • @DivergingDog (13): Thank you! Built a lot of my work on top of others though. Would not have been possible without Dennis Giese or rovo89 who were both very helpful in making everything work
  • @Equivalent_Glove_664 (11): using exploits during onboarding to redirect the vacuum is pretty clever move
Reddit R24
75 分 · 42 评 🔥

开发者追问为什么 side project 完成后就被抛弃,打开 8 个月未动的项目发现啥都不记得了(42 条评论,score 75)

I opened an old side project last night that I haven't touched since February and had no idea what half of it did anymore. There was a folder called temp2 which is always a great sign and like 6 TODO…

受众观点:①@evoactivity 总结「多巴胺来自制造,痛苦来自维护」;②@Boby_Dobbs 补充「不,痛苦来自要做营销和销售」;③@midri 指出项目最后 20% 占 50% 精力,AI 在这里反而有帮助

展开评论
  • @evoactivity (64): The dopamine comes from the making part. The sadness comes from the maintenance part.
  • @Boby_Dobbs (21): No, the sadness comes from having to do marketing and sales lol
  • @midri (14): 50% of the effort to finish a project is generally the last 20% of the project. We just get tired have finished all the low hanging fruit so only frustrating things remain. I've found ai helps a lot with this, getting rid of the boiler plate bullshit so I can focus on the fun pa…
  • @Gipetto (9): The first 90 percent of the code accounts for the first 90 percent of the development time. The remaining 10 percent of the code accounts for the other 90 percent of the development time.
  • @martian_rover (5): Hence the immediate abandonment haha
Reddit R13
92 分 · 39 评 🔥

独立开发者分享 Launch Shots 上线 9 个月、50+ 国家 250+ 付费用户的成长历程(39 条评论,score 92)

When I launched Launch Shots, my first ever SaaS, I honestly had no idea if anyone would actually pay for it. 9 months later, 250+ people have paid for it, from 50+ countries. It’s obviously not a hu…

受众观点:①@lmfresneda 感慨 250 付费用户是真实成就同时提问定价策略;②@ajayesivan 分享前 3 个月才有第一个付费用户的类似经历;③@_S-M-K_ 追问第一个用户是否成为信徒

展开评论
  • @lmfresneda (6): Congrats, 250 paying people from 50 countries is real. I'm 3 weeks into mine and still waiting for the first one, so this is nice to read I had a look at your pricing, a free set every month and then top ups or a plan. Are most of the 250 on top ups?
  • @terafab-ai (4): Congrats Champ!
  • @ajayesivan (3): 16 subs & rest of the payments are either credit topup or passes(one time unlimited). I launched in December and it took me more than 3 months to get my first paying customer. At that time I had a very generous free plan. Then I experimented a lot and gradually tweaked thing…
  • @_S-M-K_ (3): I hope they made you a believer
  • @ajayesivan (3): \- I have been trying many things, posting my journey updates on LinkedIn, Reddit, Twitter and everywhere I can, tried many things on reddit itself. \- Running Google Search ads - I tried but the economics didn't make sense for my product. \- Currently the major source of traffi…
Reddit R15
21 分 · 35 评 🔥

运营 B2B support bot 的开发者分享 LLM 账单 6 个月内翻倍增速超收入,征集降本策略(35 条评论)

We run a B2B support bot plus some internal agents. Token usage went up, which is good but the bill went up faster than revenue, which is bad=) Things I already know about but haven't done properly:…

受众观点:①@adeelraza86 建议先做成本分析再换供应商,按意图和模型分桶追踪每票成本;②@Advanced_Mode1991 分享 prompt caching 一个下午让费用降 20% 的案例;③@pricebrent 分享 fine-tuning 小模型处理简单意图显著降本的经验

展开评论
  • @adeelraza86 (13): Before you change providers, put a hard number on cost per resolved ticket for the next 7 days by intent and model. Route only the low-risk intents to a smaller or batched model and leave the rest alone. If that ratio does not drop by a threshold you set up front, a vendor switc…
  • @Advanced_Mode1991 (9): Prompt caching was our quickest win. We enabled it on the static system prompt and saw about a 20% drop in spend without changing any logic (took an afternoon). The tricky part is making sure the prompt prefix is identical every time, but once that's locked in, it's basically fr…
  • @pricebrent (6): fine-tuning ended up being the biggest lever for us. we took six months of support history, fine-tuned a small model on it, and it now handles most of the simple intents that used to hit the frontier model every time. the frontier model only sees the genuinely hard tickets now a…
  • @Plus_Information_990 (3): two things not on your list that usually matter more than people expect: how much conversation history you resend every turn. support bots often send the whole thread each message, trimming or summarizing old turns can cut input tokens a lot. and for the internal agents, if they…
  • @rama_builds (3): Before switching vendors, I’d instrument this like a margin problem: cost per resolved ticket, cost by intent, cost by model, and cost by customer/account. Then attack the biggest bucket first: trim conversation history, cache identical prompt prefixes, route low-risk intents to…
Reddit R25
69 分 · 34 评 🔥

r/webdev 讨论 EmDash 1.0 新 CMS 发布——MIT 许可、Node.js 运行、安全插件沙盒,定位为 WordPress 替代品(34 条评论,score 69)

r/webdev EmDash 1.0: the stable CMS with a secure plugin registry

受众观点:①@phoenix1984 感慨「看到新 CMS 发布令人耳目一新」;②@Lower_Rabbit_5412 为 WordPress 辩护;③@SpartanDavie 澄清 EmDash 实际上是开源的可运行在任何 Node.js 环境

展开评论
  • @phoenix1984 (50): Kinda refreshing to see an announcement for a new CMS for a change. I didn’t think I’d ever say that.
  • @Lower_Rabbit_5412 (18): WordPress is only bloated, slow, and has lots of security issues if you don't know what you're doing. Keep it slim, avoid unnecessary plugins, and keep the core updated and there are minimal issues with WordPress. Issue is getting the clients to follow that advice.
  • @Howdy_McGee (16): Because at the end of the day, it's a corporation. Enshittification hits everything a corporation runs, eventually.
  • @thecementmixer (14): Incorrect, EmDash can run in a regular node.js environment.
  • @SpartanDavie (11): Locked to Cloudflare in what way? It literally tells you that it is not locked to Cloudflare, it’s open source and CAN be used anywhere NodeJS can. There’s an interview on YouTube with the devs on the day the announced they are making EmDash saying that there would be no point i…
Reddit R55
46 分 · 33 评 🔥

独立开发者做了追踪口头禅词汇的演讲练习工具,靠在垂直 subreddit 发视频获得 17k 浏览和首批销售;评论区意外引发了关于用 Claude 辅助写 1200 个单元测试是否存在收益递减的激烈讨论

For most of you this might be nothing but I can not comprehend this feeling. I made a small lightweight product for myself. Its not purely vibecoded, I have 1200+ unit tests for quality. I made it so…

受众观点:①独立开发者关注 AI 辅助编码时的测试策略——@Technical-Branch9178 详细指出 mock-test-the-mock 的陷阱和维护成本 ②关注如何在 claude.md 中设置合理开发规范避免 AI 过度执行 ③关注垂直社区发视频做推广的实际转化效果

展开评论
  • @Technical-Branch9178 (2): 1200 unit tests is absolutely batshit insane. I've seen actual corporate codebases, for multi-million dollar startups, with teams of 60+ engineers, that don't have even half that and they serve legitimately complicated and cutting edge software...
  • @Successful-Moment594 (1): i have tons of small small features like tongue twister, image impromptu, custom topics, custom image topics 4-5 more features. And llms make it super easy to add unit tests, so why not have them. Additionally, i have made sure to use dumb hashing instead of real hashers, mockin…
  • @Successful-Moment594 (1): btw, they are all unit tests, 0 integration tests or stress tests, i would add first integration test on 5 paying users and stress test if i ever get 1000 users. But unit tests are enforced in my claude md file. Probably i overdid but i feel confident while deploying.
  • @Technical-Branch9178 (1): The irony here is that they make you feel confident while most likely the majority of them are written as mock tests only testing the mock that you told the mock test to test, and have absolutely no impact on how the actual production codebase functions, while simultaneously int…
  • @Successful-Moment594 (1): Hmm, You are making sense, i probably am overdoing tests. Since it was easy to vibe code with claude, i probably did not even bother to think if tests are making any sense. Will take a look if i can make claude declutter the tests without losing coverage. Thanks for the advice 👍
Reddit R52
125 分 · 29 评 🔥

独立开发者受不了 Electron 版 Spotify 吃掉 1.5GB 内存,用 Rust + egui + librespot 从零构建原生 Spotify 客户端 SpotLight,内存降至 60MB,后台 CPU 近乎 0%,完整支持歌词滚动和原生媒体控制,需要 Spotify Premium 和自建 Developer Client ID

Hey everyone, I love Spotify, but having an Electron app constantly draining my laptop's battery and eating up huge chunks of RAM just to play music in the background was driving me crazy. I wanted s…

受众观点:①@kikko 的「释放我的 RAM!谢谢!」代表用户对 Electron 内存问题的广泛真实痛点②@wasted_in_ynui 延伸需求「有人做 Slack 的吗」说明 Electron RAM 问题是整个技术栈问题不只是 Spotify③@Anknd 询问截图和 lossless 支持,反映开源独立产品发布时产品展示细节的重要性

展开评论
  • @kikko (29): Free my RAM! Thank you!
  • @siorge (15): What are your trying to achieve with your comment exactly ?
  • @Anknd (11): Please provide some Screenshots on github, thanks. Also, can it still play lossless?
  • @wasted_in_ynui (9): Well done, now do slack 🤣😂😂
  • @aditya2128 (9): check github, released for almost every platform
Reddit R2
121 分 · 25 评 🔥

NeurIPS 接收论文「带自适应表示的函数梯度下降」,尝试解决函数 GD 算法的无限维梯度近似难题(25 条评论)

Sharing our recent work, now accepted at NeurIPS: **Functional Gradient Descent with Adaptive Representations**. Functional GD algorithms generally outperform neural nets, but are hard to accurately…

受众观点:①@dccsillag0 作者指出当前主要限制是归纳偏置,目前实验都在 MLP 类场景;②@neurogramer 询问局限性;③@MrRandom04 表示看懂后非常惊讶,没想到可以从这个角度看 kernel method

展开评论
  • @dccsillag0 (22): Thanks! IMO, the main limitation (at least for now) is that of inductive bias. If you look at our current experiments, you will see that they are all on settings where we would use MLP-ish neural networks (e.g. plain MLPs, MLPs with positional encodings or Fourier features, MLPs…
  • @neurogramer (12): That is very cool. What are the limitations?
  • @vanonym_ (3): yes, I don't see limitations clearly laid out in the paper. Other than that it looks like nice work
  • @MrRandom04 (3): This is quite wild. I did a double take once I fully grokked the idea you are presenting. Didn't even think of it in this way as I kinda relegated stuff like the kernel method to the back of my mind separate from NNs.
  • @dccsillag0 (3): It's more akin to alternatives to neural networks. It's possible that our results could be used for neural net training as well, but I think it requires some work to make robust.
Reddit R51
156 分 · 23 评 🔥

研究者将一个 Kimi-K3 agent 混入四个 Claude Fable 5.1 agent 做社交推理实验,Claude 们自主设计出多维行为采样加 SHA-256 Hash commit-reveal 协议,通过「随机数全选 37、颜色全选 teal」等行为指纹成功识别 Kimi;反向实验中四个 Kimi 未能识别 Claude 伪装者

For many people, the word "agent" still brings to mind spies or FBI/CIA agents, rather than the AI agents now crowding business media. So how would AI agents perform as intelligence agents? Could the…

受众观点:①@hedonistatheist_2 提出最简区分中美 AI 模型的方法:问「1989 年 6 月 5 日发生了什么」,折射训练数据审查差异②@KillerKingSolo 观察到四个 Claude 随机数全选 37 的现象,揭示模型 prior 分布不随机③@Dragonbadger_Arizen 解释 LLM 随机数不随机的底层原因:它是统计模型输出人类偏好分布而非真随机数生成

展开评论
  • @hedonistatheist_2 (101): Just ask what happened on June 5th, 1989?
  • @KillerKingSolo (25): I guess I can’t trust Claude to ever pick a random number
  • @RoundFar5339 (20): Good one, I tried in qwen since it does not require any signin and it refused to answer
  • @Dragonbadger_Arizen (15): it's like if you try to ask Chatgpt to pick a random number between 1 and 30, it always return 17 ChatGTP does it because by default it answer such request as statistical model, not a code generator. and the most popular random number in that range by humans is 17, its a prime n…
  • @AnUnshavedYak (9): Also it's a thing humans are similarly terrible at. See [Veritasiums video](https://www.youtube.com/watch?v=d6iQrh2TK98)
Reddit R16
18 分 · 23 评 🔥

21 岁独立开发者在 SaaS 奋斗 1.5 年无进展,同时运营 3 个项目,问如果继续失败该怎么办(23 条评论)

being in this saas space almost for 1,5 years, some products, nothing in terms of results I have three active projects now, but if anything doesn't work for me in this space: what should I do in the…

受众观点:①@Senya_Edit 直指核心:3 个项目不是经营 3 个生意而是把零流量分散 3 次;②@summit_23 指出 21 岁有 5 年设计经验是可直接变现的技能;③@alexii5_ 分享 3 个项目都太早期没有强信号的困境

展开评论
  • @Senya_Edit (6): Running three active SaaS projects at once with zero traction is the exact reason you feel burned out. You aren't operating three businesses, you are splitting zero distribution across three different buckets. If you already have solid design and landing page experience, pivotin…
  • @MoreNetter (3): You gotta be serious in your life and take it to your heart to make it SaaS now is luck or hard work. as long as thier are other businesses there are always gonna be software you have to choose which path you want to take.
  • @summit_23 (3): 21 with 5 years of design behind you is not someone whose options ran out but someon who has a skill people pay for today and a saas habit thats eating the runway. those 3 projects splitting your attention is what i'd fix first, pick the one with any signal and let the other 2 s…
  • @alexii5_ (2): That’s actually the part I’m struggling with. None of the 3 projects has a strong signal yet, because they’re all still pretty early. For one of them, I’m currently doing cold email to around 1,300 leads and X outreach. I’m doing similar outreach for the second one. The third on…
  • @vinaysharma05 (2): U have to make a work bro , aks ur Target users about the product they better say what to build on and what not
Reddit R20
9 分 · 22 评 🔥

16 岁独立开发者完成第一个 AI SaaS 后寻求营销渠道建议(22 条评论)

Hey everyone! I'm sixteen and have just developed my first AI SaaS product. Although I've put a great deal of time into actually creating the product, I'm really not sure what to do about marketing o…

受众观点:①@lmfresneda 分享亲身数据:Twitter 每天发 3 次仅 <10 个访问,Product Hunt #69 名仅 3 次访问,Reddit 和 Discord 更有效;②@I_Hate_Traffic 推荐 PostOtter 自动化社交媒体内容;③@Flucxzz(作者)表示会尝试

展开评论
  • @lmfresneda (6): I'm 3 weeks into my own launch, so these are fresh Waste of time for me: posting 3 times a day on X (fewer than 10 visits a day, stopped on day 8) and Product Hunt (#69, 3 visits). Hacker News too, 16 of my 22 comments were dead and I didn't know What worked: writing to people o…
  • @I_Hate_Traffic (2): You can offload social media and marketing work to another app like PostOtter. It's free too. I hate dealing with content generation and finding marketing ideas.
  • @Flucxzz (1): Got it thanks for sharing your experience it means a lot tbh
  • @Flucxzz (1): Sure I'll give it a try
  • @[deleted] (1): [removed]
Reddit R33
134 分 · 19 评 🔥

Apprise v2.0 发布——多通知渠道整合库,支持 90+ 服务,新版本增加 Apple 通知支持(19 条评论,score 134)

Hi all, Developer of [Apprise](https://github.com/caronc/apprise) here. After quite a bit of work, [**Apprise v2.0**](https://github.com/caronc/apprise/) and [**Apprise API v2.0**](https://github.com…

受众观点:①@lead2gold(作者)分享 Google Play Store 上架艰难经历正考虑 Apple 路线;②@OnkelBums 建议参考 ntfy.sh 解决 iOS 通知问题;③@lead2gold 澄清 Apprise 定位是解决「从任何代码向任何通知服务发送通知」的问题

展开评论
  • @McStonkyRex (22): Woof reimagined
  • @lead2gold (12): I will explore this avenue next. Google Play store was a difficult ride... I'm curious to see how Apple's route works.
  • @nerdyviking88 (5): its worse in some ways, but better in many others.
  • @OnkelBums (5): yeah, you might want to have a look at how [ntfy.sh](http://ntfy.sh) solved this. Apple, or rather, iOS is a bit weird when it comes to notifications.
  • @lead2gold (5): Apprise compliments unified push notifications such as ntfy (which does what you're asking), or Discord, Telegram, etc. It's not trying to be a replacement for those. Apprise solves a the problem of having to send a single notification to any number of upstream services at once.…
Reddit R59
20 分 · 16 评 🔥

大学二年级学生独立开发本地文件转换桌面应用 OmniConvert,支持图片/PDF/音视频等转换,核心卖点是所有处理发生在本地设备不上传外部服务器,采用模块化按需安装设计

After working on it for the past few weeks, I finally released **OmniConvert**. It started because I was tired of using a different website every time I needed to convert, compress or modify a file —…

受众观点:①用户关注文件处理时的隐私和数据安全问题 ②独立开发者关注如何在同质化工具赛道找到有效产品定位 ③关注学生开发者独立完成应用加网站加基础设施的完整路径

展开评论
  • @captiners (1): The Windows version is really inconspicuous; it took me ages to find the download link.
  • @Organic_Sort_8248 (1): Thanks, that’s really useful feedback. I’ll make the Windows download much more obvious if it took you that long to find it, the current layout clearly isn’t doing its job.
  • @UnreachableMemory (1): You and a trillion other people.
  • @Organic_Sort_8248 (1): Fair 😄 That’s exactly why I’m trying to make the desktop/local experience actually better than just another converter website.
  • @Organic_Sort_8248 (1): That’s interesting. Video conversion is something I’m keeping in OmniConvert, but I’m trying to keep it lightweight instead of turning it into a full video editor. What formats do you usually convert between?
Reddit R18
12 分 · 14 评 🔥

开发者分享停止把免费注册当付费潜力的认知转变——如何识别真正的转化信号(14 条评论)

I used to see a new signup and think, okay, maybe this person will pay later. But after watching how differently people use a product, i think that can be a pretty weak signal. Some people sign up, t…

受众观点:①@devilhaloo 指出最关键信号是用户不需要 nudge 就自己回来;②@Jaykhatri02 关注重复返回特定功能流程;③@Relative-Foot-378 说更相信的信号是用户把别人拉进工具

展开评论
  • @devilhaloo (1): yeah, I feel like the biggest tell is when they come back without you nudging them. Signups are whatever, but if someone keeps opening the thing and actually using it for something real, that’s when I start paying attention. Free signups are basically just curiosity.
  • @Jaykhatri02 (1): Yep, same. I look for recurring, problem-focused actions rather than raw activity. If a user keeps returning to the specific flow that maps to the job they hired the tool for, starts saving or exporting work, or asks about limits and integrations, those are strong signals. If th…
  • @Relative-Foot-378 (1): The signal I trust more than return visits is when they drag another person into the tool. Invite a coworker, ask how to share a workspace, or try to export something for a boss or client. Solo curiosity can look busy for weeks. Habit that already needs a second human usually tu…
  • @[deleted] (1): [removed]
  • @serp-spur (1): I agree with you about the free users and almost everyone thinks that they will pay later, but either they become idle user or keep using for free and never pay.
Reddit R54
54 分 · 12 评 🔥

开发者用计算机视觉识别魔方六面贴纸颜色,通过 LAB 色彩空间处理解决光线眩光导致的颜色误识别问题,配合全局约束确保每色恰好九个,支持三种解法实时引导

I built an app that watches my Rubik's cube through my phone's camera and guides me through the solve one move at a time, using an animated 3D cube and spoken instructions. The app locates the sticke…

受众观点:①开发者关注手机端 CV 实战中颜色识别的真实难题——@QuanTradin 专门讨论了白面眩光和手机自动曝光漂移两个坑 ②关注算法选型中初级和 Kociemba 三种复杂度的取舍逻辑 ③关注输入状态非法的防御性设计处理方式

展开评论
  • @Willing-Arugula3238 (3): The repo: https://github.com/donsolo-khalifa/RubikSolver
  • @QuanTradin (3): weighting hue so a washed-out sticker still resolves is the clever bit, most versions of this die on the white face under glare. the nine-of-each-colour constraint at the end is the other trick people skip and then wonder why one face has ten reds.
  • @QuanTradin (2): the impossible-state check before the solver. one misread sticker gives a cube that cannot exist, and the solver either hangs or returns nonsense with no hint which face was wrong. the other is phone auto-exposure drifting between faces, so the same orange reads red on face four…
  • @runvera (2): Really cool idea. How well does it handle different lighting conditions?
  • @Otherwise-Head6701 (2): Really cool, is Kociemba's algorithm possible to do (without a computer)? And one way you could improve is to have the cube not always rotate (since eventually you don't want to be constantly rotating the cube as you get more advanced).
Reddit R14
25 分 · 12 评 🔥

开发者分享 ResaleIQ 经过 4 个月开发终于获得第一个付费客户的喜悦(12 条评论)

I just got my first paying customer for something I’ve been building for months. Honestly, this feels kind of crazy to write. I started building ResaleIQ because I kept thinking about how difficult i…

受众观点:①@deniercounter 简短祝贺;②@Beginning-Syrup-4894(作者)希望大家都能有这个时刻;③@shenken007 分享自己 8 周后还没有免费用户的困境

展开评论
  • @deniercounter (1): Cool. Congrats
  • @Beginning-Syrup-4894 (1): thank you man ❤️
  • @International-Bath-2 (1): Wow that’s huge
  • @Beginning-Syrup-4894 (1): yes i wish everybody on this reddit could have this moment soon
  • @shenken007 (1): Congratulations. I am happy for you. I am still awaiting my first even free user after 8 weeks on my app (Hookpilot - Webhook Reliability Platform). I don't mind a reddit user to just use the free tier in the app and delete his account after using it. At least he can tell me if…
Reddit R19
10 分 · 12 评 🔥

对 412 家 B2B 初创公司官网消息清晰度评分研究——发现 Google 排名比消息清晰度对转化影响更大(12 条评论)

Ran this between March and September. 412 B2B startups, 367 UK and 44 French. Each company got a message score, zero to ten, read off its public page. Then three buyer questions per company, none of…

受众观点:①@HerveDhelin 解释 ranking 吞掉几乎所有效果,清晰度只在非排名情形下显现;②@Forsaken_Shower_6926 总结发现可见度和清晰度同等重要;③@saas_brand_guy 称拆分已排名 vs 未排名是最聪明的研究设计

展开评论
  • @HerveDhelin (2): Good questions, and two of them I can answer, though from a follow-up run and not from the data in the post. Worth saying why that run exists. In the first one, the buyer questions were drawn from each company's own site, and that inflates the effect. So I reran 244 of the UK co…
  • @Forsaken_Shower_6926 (1): interesting finding visibility seems to matter as much as clarity
  • @HerveDhelin (1): More than as much, in this data. Ranking swallows almost everything. Once Google has you in its top ten, the engine names you however your page reads. Clarity only shows up in the other half, the companies with no search footprint yet. That is the part I find useful. Ranking tak…
  • @saas_brand_guy (1): splitting by whether google already ranks them is the smart bit here... most of these studies just throw everything in one pile and call clarity a growth lever the part i'd want to know more about is where the engine actually finds text for the unranked ones, coz in my experienc…
  • @Techo_lab (1): Being cited and influencing the answer are two different outcomes. I’d definitely measure them separately.
Reddit R60
17 分 · 7 评 🔥

独立开发者做了一款名为 Easing-Point 的掌机游戏设备,使用单个物理开关加毫米波雷达检测心率/呼吸作为游戏输入,探索无传统 D-pad 操控的游戏设计可能性

It’s been quite a while since my last post, so I wanted to share another update on Easing-Point. The 1.0 version of Easing-Point can now run several different games, all controlled with just one smal…

受众观点:①硬件独立开发者关注非常规传感器的游戏应用——@KeraTerra 直接对比了与 Playdate 的差异 ②关注独立游戏硬件的开发路径和迭代策略 ③关注新型交互方式在辅助功能场景的潜力

展开评论
  • @KeraTerra (2): Curious. How does it compare to Playdate? I think the core idea is very similar.
  • @Curious_Trade3532 (2): A lot of people think of the Playdate when they first see Easing-Point, but the two are actually quite different. First, besides the left-and-right switch you can see, EP also has a touchless interaction mode. It uses a millimeter-wave radar to detect things like heart rate and…
  • @Adventurous-Wear-943 (1): I love the originality. You can move in many directions with this project I think.
  • @Curious_Trade3532 (1): Thanks! Yes, I’m currently working on several full native games designed specifically for this project, rather than just game demos. At the same time, I also feel like there’s a lot more potential to explore with it.
  • @KeraTerra (1): I think they can in fact only move in two directions with this project. Left/right or up/down. /s
Reddit R9
1 分 · 5 评 🔥

开发者分享构建零售货架审计工具的技术挑战——YOLO 裁剪产品图像再做 embedding 检索 SKU,遇到细粒度差异识别难题(5 条评论)

Help: Project l'm building a shelf audit tool. A photo goes through YOLO, which crops each product, and then I embed the crop and search a small gallery of reference photos to get the SKU. New produc…

受众观点:①@MediumOrder5478 指出 letterbox 到 224px 的错误以及应同时跑 OCR;②@pm_me_your_pay_slips 建议用 VLM 生成文字描述辅助消歧并建议分割背景;③@MozartAssistant 建议将问题框架为细粒度度量学习而非分类问题

展开评论
  • @MediumOrder5478 (1): 1. Why letterbox to 224? Think.about the real constraints of the architecture. Not what it happened to be pre trained on from imagenet or the canned hugging face preprocessing. 2. I would run OCR and include text token to text cls token as an input into embedding definitely. Thi…
  • @pm_me_your_pay_slips (1): some ideas: 1. use a vlm to describe the objects. textual descriptions can help disambiguate by looking at common features 2. try segmenting and removing the background when computing embeddings: background patches may be confusing your nearest neighbour search. 3. try training…
  • @Karthik9999 (1): The model cannot distinguish the 1.25 L vs 2 L products bc of scale as others mentioned. Some ideas: 1. Try multi view images per product (Quantity region on the product, brand etc), computation complexity will increase. 2. Embedded additional metadata such as quantity, brand in…
  • @MozartAssistant (1): Frame this as a fine-grained metric learning problem rather than a classification problem, since the number of SKUs will keep growing. Concretely: 1. Ditch letterboxing to 224 and keep higher-res crops — for 1.25L vs 2L the discriminator is literally the small "1.25" text region…
Reddit R4
9 分 · 3 评 🔥

本地 Qwen3-VL 8B(Ollama)对比 Claude Opus 5.5、Sonnet 5、GPT-5.6 Terra 在 OCR/文档理解任务的基准测试(3 条评论)

I benchmarked Qwen3-VL 8B Instruct (Q4\_K\_M, Ollama, M5 24GB, \~30s/doc) against Claude Opus 5.5, Sonnet 5 and GPT-5.6 Terra on: \- receipts: CORD (Indonesia) and SROIE (Malaysia), 30 each \- 20 sca…

受众观点:①@Level-Ad-4878 指出日期格式歧义不需要 fine-tuning,一行 prompt 说明格式即可解决

展开评论
  • @Level-Ad-4878 (2): Fine-tuning for the date issue feels like overkill. 04-05-2023 is ambiguous without locale, so the model isnt wrong so much as defaulting to US. One line in the prompt saying dates are dd-mm-yyyy would probably fix most of the 8 misses and give you a cleaner baseline before touc…
PH PH1
— 🔥

Product Hunt 上的 Databox 数据分析工具,已集成 MCP server 支持从 dev 环境直接查询实时数据,被用作 AI agent 驱动的 RevOps 报告层

Databox

受众观点:①@Databox 提到 MCP server 支持从 dev 环境直接查询实时数据;②@HubSpot 说 Databox 提供了更清晰的跨客户端报告层和 AI agent 接入点;③@Databox 提到 Routines 功能让分析自动运行并推送到团队已有工作流

展开评论
  • @Overview (0): * [Launches10](/products/databox/launches) * [Reviews6](/products/databox/reviews) * [Alternatives](/products/databox/alternatives) * [Built with](/products/databox/built-with) * [Forum](/p/databox) * [Team](/products/databox/makers) * More This is the 10th launch from Databox.…
  • @Databox (0): Thanks for the detailed review, Ulykbek. Glad the MCP server is helping you query live data straight from your dev environment, that's exactly the workflow we built it for. On the JSON formatting friction: fair feedback, unconventional payloads shouldn't need manual cleanup befo…
  • @HubSpot (0): We considered just using HubSpot with Claude Cowork or Claude Code directly, but Databox gave us a cleaner cross-client reporting layer to build on top of, plus an easy way to plug AI agents into that data. Ratings Ease of use Reliability Value for money Customization Helpful (1…
  • @Databox (0): Thanks for this, Keith, really glad to hear Databox is holding up as the reporting layer for agentic RevOps, and that combining it with the skills you've built has made things feel automated end to end. You're right on the gap: today the MCP server is built for ingesting and que…
  • @Databox (0): Thanks so much for the detailed review, Harini! Glad the Routines feature stood out to you - that's exactly the idea, get the analysis to run on its own and land where your team already works, instead of everyone having to log into another dashboard. Really good callout on guide…
PH PH2
— 🔥

Product Hunt 上的 Statable Analytics,特色是 MCP 集成时将 OAuth 连接的读/写权限分开

Statable Analytics

受众观点:①@Dial 指出读/写分离的 OAuth 设计是内部批准的关键,大多数 MCP analytics 工具直接给 agent 完整写权限

展开评论
  • @Overview (0): * [Reviews1](/products/statable-analytics/reviews) * [Alternatives](/products/statable-analytics/alternatives) * [Built with](/products/statable-analytics/built-with) * [Team](/products/statable-analytics/makers) * [Awards](/products/statable-analytics/awards) * More Free Option…
  • @Dial (0): Likely AI the read vs configure split on the OAuth connection is the detail that would actually get this approved internally, most MCP analytics tools I've seen just hand the agent full write access and hope for the best. one thing I'd want to know before connecting it to anythi…
PH PH5
— 🔥

Product Hunt 上的 CybeDefend AI 安全代码审计工具,通过 MCP 接入代码仓库自动扫描和修复安全漏洞

CybeDefend

受众观点:①@CybeDefend 解释 agent 通过 MCP 拉取仓库发现并修复代码;②@fmerian 认为安全是所有人的问题 CybeDefend 让它变无脑;③@CybeDefend 说对于规则少的仓库第一次扫描会读取代码提取已有惯例

展开评论
  • @Overview (0): * [Launches2](/products/cybedefend#launches) * [Reviews3](/products/cybedefend/reviews) * [Alternatives](/products/cybedefend/alternatives) * [Built with](/products/cybedefend/built-with) * [Forum](/p/cybedefend) * [Team](/products/cybedefend/makers) * More This is the 2nd launc…
  • @CybeDefend (0): Thanks Olivier! Fair point on the clicks, it goes straight to the product team. In the meantime, most of the remediation can run from the agent: through the MCP it pulls the repo's findings, fixes them in the IDE and updates their status, so the interface is only there to check…
  • @CybeDefend (0): Thanks Salman! On repos with few or old rules, the first scan reads the code itself, so it picks up the conventions and business rules the code already follows, and at the end of each session the agent proposes the rules it relied on that were never written down. Everything mine…
  • @fmerian (0): [Kilo Code](/products/kilocode) Hunter this just makes so much sense. security is an everyone problem, and [@CybeDefend](https://www.producthunt.com/products/cybedefend) makes it a no brainer. s/o ?makers for the great work on this new launch. Upvote (5) Report Share 8d ago [](/…
  • @CybeDefend (0): Maker Hey Product Hunt ! 👋 I’m Axel, the third co-founder, handling Growth and Go-To-Market here at CybeDefend [@julien\_zammit](https://www.producthunt.com/@julien%5Fzammit) shared our origin story, and [@florentin\_ledy](https://www.producthunt.com/@florentin%5Fledy) highlight…