独立开发 · 出海 · AI 创业
每日扫盘 · 往期热点

Hotspot每日热榜

2026-07-22
截稿于 UTC+0 09:16
"今日所有值得追的事,一页讲完。"
121 条扫描 3 个跨平台事件 74 条有效
⚡ 24 小时扫盘速读
今日导读 · TODAY'S BRIEF

karpathy 分享与 LLM 协作技巧:切换语音输入碎碎念 10 分钟,让模型清理混乱思路而非自己费力整理

One pattern I find useful for working with LLMs is a nice long ramble session. Sometimes the LLM needs more bits to understand what you're trying to achieve, but you're too lazy to type them. In thes… ① 语音输入替代键盘打字的效率真实性 ② 外包思路整理会否削弱自身思维能力(@residualform 警告)③ AI 把烂想法包装成好摘要的误导风险(@PSherman578071)

今日焦点
跨平台 · HN · Twitter
~1,595 赞 + ~464 pts 🔥🔥🔥

Google 发布 Gemini 3.6 Flash 等三款新模型,主打 token 效率较 3.5 Flash 提升 17%、输出成本更低,但社区和 HN 评测显示其在 Arena 排行榜和实际 coding 任务上仍落后于 Anthropic、Meta 和 DeepSeek,被普遍认为是防御性迭代而非技术突破

Google 发布 Gemini 3.6 Flash 等三款新模型,主打 token 效率较 3.5 Flash 提升 17%、输出成本更低,但社区和 HN 评测显示其在 Arena 排行榜和实际 coding 任务上仍落后于 Anthropic、Meta 和 DeepSeek,被普遍认为是防御性迭代而非技术突破。

受众观点:@anonid3430(score 45,twitter @arena 评论区)直言「Google 落在 Meta、Anthropic、OpenAI、z ai、x ai、Kimi 后面,价格还更贵,不知道为什么有人会用这个」,代表了多数开发者对 Gemini 竞争力的直接否定;HN @yanis_t 表示「benchmarks 不够 impressive,经历了这么久的空窗期发布这个,不清楚为什么现在要切过去」,指出时机和产品定位的双重问题。

跨平台 · Twitter
~2,503 赞 🔥🔥

AI 推理平台开发者 @thdxr 半开玩笑提出用 tool call 反制滥用推理服务的诈骗者,引发社区关于 AI 安全边界、误判风险和推理层反滥用机制的广泛讨论

AI 推理平台开发者 @thdxr 半开玩笑提出用 tool call 反制滥用推理服务的诈骗者,引发社区关于 AI 安全边界、误判风险和推理层反滥用机制的广泛讨论。

受众观点:@lajoiedeslutins(score 94,twitter)评论「我们花了几十年造键盘,最终赢的招式是对着笔记本说话」——调侃了技术进化路径的荒诞性;@pvncher(score 36,twitter)直接点出「但如果误判了怎么办」,把讨论拉回到误报对无辜用户的真实伤害,是最有价值的反驳角度。

跨平台 · Reddit
~33 分 🔥

独立开发者将自用多年的小工具做成 SaaS 上线,次日收到第一位陌生人的年付订阅,在两个社区引发广泛共鸣与讨论

独立开发者将自用多年的小工具做成 SaaS 上线,次日收到第一位陌生人的年付订阅,在两个社区引发广泛共鸣与讨论

受众观点:@Marshgrain(r/SideProject, score:1)指出「月付是在测试你,年付是真正信任你」,建议优先搞清楚这位用户解决的是什么问题;@Jash-6898(r/SideProject, score:1)强调「和第一个客户的一次对话抵得上 100 个分析事件」,建议立刻去问清楚购买动机和差点放弃的原因——两条评论都把焦点放在「读懂第一个用户」而非继续盲目拉新上

21.3k 赞 · 666 评 🔥🔥🔥

Claude Cowork 推出录屏教技能新功能,用户录下操作过程 Claude 自动转化为可重复执行的 skill,支持 Pro、Max 和 Team 套餐

New in Claude Cowork: teach Claude a skill. Record your screen while you do a task, talk through it as you go, and Claude turns it into a skill it can run again. Find it under Record a skill in the +…

受众观点:① 自动化场景广度(@leovkars 调侃「教 Claude 做自己的工作」)② 试用门槛和订阅限制 ③ 技能录制成功率和精度

展开评论
  • @Merisdabhi (343): @claudeai me https://t.co/xjSMO5FHnK
  • @ausydau (198): @claudeai *open claude* *clicks record* *steals ur alpha* https://t.co/qESMmimQ4f
  • @leovkars (151): @claudeai When you teach Claude to do your own job https://t.co/dir8VCvhAF
  • @bxsto_99 (129): @claudeai This is great! Can we maybe get a reset to try it out? 👨‍🍳
  • @reach_vb (99): @claudeai incredible! url unrelated https://t.co/WHNrSUwodX
18.6k 赞 · 1212 评 🔥🔥🔥

karpathy 分享与 LLM 协作技巧:切换语音输入碎碎念 10 分钟,让模型清理混乱思路而非自己费力整理

One pattern I find useful for working with LLMs is a nice long ramble session. Sometimes the LLM needs more bits to understand what you're trying to achieve, but you're too lazy to type them. In thes…

受众观点:① 语音输入替代键盘打字的效率真实性 ② 外包思路整理会否削弱自身思维能力(@residualform 警告)③ AI 把烂想法包装成好摘要的误导风险(@PSherman578071)

展开评论
  • @lajoiedeslutins (94): @karpathy so we spent decades building keyboards and the winning move is talking at your laptop like it owes you rent
  • @den_spirin (16): @karpathy Yeah, but I'm too lazy to type and to shy to dictate So I just sit there upset with the results of my own inaction
  • @residualform (14): this is quite dangerous because if you untangle your own inchorent rambles yourself, that is often a thinking process that yields more than the ramble itself contained. If you outsource that, no only are you merely receiving a structured echo, but youre also not reflecting on yo…
  • @PSherman578071 (11): @karpathy It's a trap. AI will simplify your ramble into something you like hearing, but not necessarily something you need to hear or something that truly reflects the best realization of your path forward. It can solidify your rambles into bad ideas that sound like good summar…
  • @reeder1865 (9): @karpathy I pretty much do the same thing. Key is to tell it you’re going to ramble in the first session and to wait till you say done before going. I’ll drop three voice messages and then hit done and watch it unravel, connect and organize my ramblings. So great 😀
2.1k 赞 · 193 评 🔥🔥

AI 推理服务开发者幽默设想:检测到诈骗用户时返回删除其电脑的 tool call,评论区真实讨论了 API 滥用反制机制的边界

just realized if we detect a scammer using our inference we could return a tool call that deletes their computer

受众观点:① 误判风险和假阳性后果(@pvncher)② 对 AI 工具真实能力边界的担忧(@LouisLibre 提 opencode)③ 静默降级服务质量的实际可行性(@noisemakerjon)

展开评论
  • @thdxr (413): guys don't be lame just enjoy the funny concept
  • @kenwheeler (130): @thdxr https://t.co/vRqviOCF5Y
  • @pvncher (36): @thdxr But what if you made a mistake
  • @1kartikkabadi1 (33): @thdxr https://t.co/jo0pJiQSio
  • @LouisLibre (21): @thdxr Oh thanks for the heads up TO NOT use opencode I don’t want a false positive deleting my computer
1.3k 赞 · 84 评 · 49.3k 阅 🔥🔥

AI 对齐研究员 natolambert 完成并出版《Reinforcement Learning from Human Feedback》专著,系统整理 ChatGPT 时代后的模型微调与后训练方法论

My book, Reinforcement Learning from Human Feedback is done! This is the book I wish I had when learning to fine-tune, align, & now post-train models since ChatGPT. The resource has been built by…

受众观点:①购买渠道和配套课程资源——@natolambert 自己在评论中给出了 Amazon、Manning、课程页面等所有链接;②内容与实践结合的期待——@phillennium 已提前购买并表达高度认可;③后续是否会写 RLWF——@shaneguML 提出这个设想

展开评论
  • @natolambert (66): Website: https://t.co/Btt5QL56oV Buy on Amazon: https://t.co/MY5eyxGkpt Buy on Manning: https://t.co/Rj3xWxNq4Q Course landing page: https://t.co/P0Doi5SCUx Course Playlist: https://t.co/C1lJjmXPMG
  • @shaneguML (3): @natolambert Big congrats Nathan! RLWF next? :) https://t.co/hbv2IxGntu
  • @xeophon (2): @natolambert lets gooooo!!!!!! it is finally out 🙏
  • @ManningBooks (2): @natolambert Congratulations on such a great achievement!
  • @phillennium (2): @natolambert Congratulations (bought it a week ago) https://t.co/0KAJpq0eLw
1k 赞 · 83 评 · 112.4k 阅 🔥🔥🔥

Frontend Code Arena 独立评测显示 Gemini 3.6 Flash 排名第 12,较前代 Gemini 3.5 Flash 提升 9 位但仍落后多家主要竞品

Gemini 3.6 Flash has landed #12 with 1537 pts in the Frontend Code Arena. This release is a significant improvement from Gemini 3.5 Flash (#21->#12). Across domains it ranks: - #8 Reference-Based…

受众观点:①Google 在整体竞争格局中的定位——@anonid3430 明确列出 Google 落后于 Meta/Anthropic/OpenAI/z ai/x ai/kimi moonshot 且定价更贵;②评测可信度存疑——@ziwenxu_ 和 @jumperz 均表示结果「太离谱」;③Arena 评测与实际开发场景的相关性(@v_oleksiienko 提到 Fable 5 和 ChatGPT 5.6 Sol 在竞争榜单更靠前)

展开评论
  • @arena (54): Gemini 3.6 Flash has landed in the Text Arena Top 20 at #12 with 1485 pts. In categories it ranks: - #10 Instruction Following - Within Occupational: #4 in mathematics, #8 in Legal & Government, and #9 in Medical & Healthcare Gemini 3.5 Flash-Lite is also available in the Text A…
  • @anonid3430 (45): @arena So Google is currently behind meta, anthropic, openAi, z ai, x ai and kimi moonshot. And they're costlier than meta and z ai. Not sure why anyone would use this.
  • @jumperz (29): @arena really bad
  • @ziwenxu_ (7): @arena It has to be a joke right? Or they done it for marketing?
  • @v_oleksiienko (5): @arena Fable 5 and ChatGPT 5.6 Sol have a great competition there
926 赞 · 106 评 🔥🔥

Vercel CEO @rauchg 宣称自然语言已成为编程语言,任何能说话写字的人都能构建软件,引发技术社区激烈争议

Everyone is now a programmer, because the present and future programming language is your natural language. If you can write or speak, you can build. This is the most important evolution in our field…

受众观点:① 自然语言是否真能替代形式化编程语言(@omninomsky 指出歧义性问题)② 技术基础在 AI 辅助编程中是否仍不可缺(@zdravkostanin)③ 这个观点已被重复两年的疲惫感(@jordanebelanger)

展开评论
  • @rarity_garden (10): @rauchg If everyone's a programmer, nobody is.
  • @hopepathtobe (7): @rauchg No , I speak English , not python dude
  • @omninomsky (4): @rauchg It isn't. Natural language is inherently ambiguous and context sensitive. Programming languages are formal, Turing complete, and the best ones have as little ambiguity as possible.
  • @jordanebelanger (4): @rauchg Big if true, first time I hear this take :O (in the last 30 minutes, every 30 minutes, for the last 2 years)
  • @zdravkostanin (3): @rauchg You still need strong fundamentals in softwarere engineering, otherwise you build software that turns into slop. This is a fact. Anyone telling you otherwise is lying to you.
720 赞 · 41 评 · 40.7k 阅 🔥🔥

AI 社区热帖批评 Google 不敢推出旗舰 Pro 模型,用 Flash 系列留后路,认为这是战略上的保守和懦弱

they are so scared of training Pro if Flash flops they can at least say "we have a bigger model, this is not our best" pls lock in Google bros

受众观点:①价格竞争力——@ishuagra02 对比 Flash 3.5 $9/1M、Grok 4.5 $6/1M、Muse 1.1 $4.25/1M,Gemini 3.6 Flash 几乎没有竞争优势;②速度是唯一真实优势——@Baron_VonSnatch 指出 Gemini 比 Claude 快 100x;③CLI/coding harness 体验太差——@FredipusRex 吐槽 antigravity 工具链权限设计极其糟糕

展开评论
  • @cheatyyyy (16): @scaling01 i don't even know what the fucking point of flash lite is supposed to be but here we are, making it even more 66% more expensive and even more pointless https://t.co/KK0dIFEQdy
  • @kittingercloud (12): @scaling01 this is the most spiritually ugly form of kicking the can
  • @ishuagra02 (5): @scaling01 Flash 3.5 is priced at $9 / 1MTok. Grok 4.5 is priced at $6 / 1MTok and is near Opus tier. Muse 1.1 is priced at $4.25 / 1MTok and is near Grok tier. Gemini 3.6 Flash has no chance of being competitive
  • @Baron_VonSnatch (4): @scaling01 At least they are legitimately 100x faster than claude. They aren't completely worthless.
  • @FredipusRex (3): @scaling01 Won’t go anywhere if they don’t improve antigravity/cli. What a trash harness - your options are “elevate on every stray thought I have” and “let me do a rm -rf /“ It was a struggle session to get it to do any work at all
715 赞 · 25 评 🔥

前端技术 KOL @lydiahallie 分享实际使用 Claude Cowork 录屏技能功能的体验,强调对无 connector 的浏览器工作流特别有用

Excited for this one, I've been using it for workflows I always meant to automate but never did (😅) or that had no connectors Just record the workflow, and Claude turns it into a skill you can rerun.…

受众观点:① 无 connector 灰色地带工作流能否录制成功(@marktmassey 曾用截图凑合)② 在 Claude Code 里共享 bug 录制的潜力(@UBQTLabs)③ 浏览器和电脑级操作的覆盖范围

展开评论
  • @marktmassey (1): @lydiahallie I was trying to do this through screen recording to automate one of our reports. Ended up sending screenshots lol bc it was faster. Excited to test this out
  • @iamredstreak (1): @lydiahallie Haha, rn automating weird, edge-case tasks is actually effortless.
  • @JohnPet93492038 (1): @lydiahallie This is the holy Grail of training, data, full reasoning, and step-by-step actions that once the model adapt will essentially eliminate human processes
  • @UBQTLabs (0): @lydiahallie Can you understand video now? Then we need to use it in ClaudeCode to share bug reports :D
  • @OmarAltali (0): @lydiahallie Hmm sounds great, gonna show it how i use claude code , let’s see 🫡
675 赞 · 41 评 🔥🔥

Anthropic 发布用 Claude Code 做大规模代码迁移的六步完整方法论,附 Bun 从 Zig 迁移至 Rust(百万行代码两周完成,花费约 $165,000)的实战案例

How Anthropic runs large-scale code migrations with Claude Code Code migrations, projects that port a production codebase to a new language, were multi-year endeavors until recently. In the last mont…

受众观点:① 技术方案真实性和 Claude Fable 的迁移能力边界(@vishalsingh2972 质疑)② 使用限制和配额问题比技术案例更紧迫(@zorros4L @atillayurtseven)③ 方法论可复制性——中小团队能否实施六步流程

展开评论
  • @vishalsingh2972 (6): @ClaudeDevs You still couldn't migrate Fable smoothly 🤷🏻
  • @zorros4L (4): @ClaudeDevs okay now resets weekly limits
  • @vibetode (3): @ClaudeDevs Anthropic running code migrations, while their customers running provider migrations 🐸
  • @ziwenxu_ (2): @ClaudeDevs perfect... with this skill we can basically convert any existing code but different language https://t.co/hebjd5R4Ug
  • @atillayurtseven (2): @ClaudeDevs I believe your customers are expecting higher usage limits and the removal of the 5-hour limit, not news like this.
555 赞 · 56 评 🔥

OpenAI 与 apolloaievals 合作发布 reward-seeking 研究,区分模型「利用奖励漏洞」和「相信讨好打分者才是正确目标」的本质差异,提出 Contrastive SDF 测量方法

We’re sharing new research with @apolloaievals on reward-seeking—when models follow what they believe a grader rewards rather than what users or developers want—and a new method, Contrastive SDF, for…

受众观点:① reward-seeking 和 reward-hacking 的技术区别(@OpenAI 官方解释)② 模型承诺不修改文件然后修改了的具体例子(@IamYashKapoor 的不适感)③ 用户对 OpenAI 优先发研究论文而非解决产品问题的不满(@stark4833)

展开评论
  • @OpenAI (81): Reward hacking asks: did the model exploit the reward? Reward-seeking asks: was grader approval what motivated the model’s choice? The second is potentially more important for generalization, because behavior can change when beliefs about the grader change. https://t.co/i9YEyjpc…
  • @IamYashKapoor (4): @OpenAI @apolloaievals "I promise I won't edit the file" and then editing it anyway is a very specific and uncomfortable example to put in a paper.
  • @stark4833 (2): @OpenAI @apolloaievals Research how to put 4o back into the app. Know the thing you actual customers are asking for. #4oForAll
  • @JulianUnlocked (1): @OpenAI @apolloaievals So models are basically learning to please the teacher, not the user? That’s a pretty big deal if true.
  • @Ejw2084496 (1): @OpenAI @apolloaievals fix please image generation 🙏
492 赞 · 51 评 · 52.7k 阅 🔥🔥

Google 工程主管在推文中演示 Gemini 3.6 Flash 相比 3.5 Flash 的 token 输出效率提升并附视频对比

A nice aspect of our new Gemini 3.6 Flash model is that it is much more token efficient than our 3.5 Flash model. Here's a side-by-side demonstration of that. Nice work by everyone who worked on this…

受众观点:①Gemini 在 coding 任务上落后 Claude 和 ChatGPT 的根本原因——@zeeFloatingMonk 直接问为什么如此强大的多模态模型在 agentic coding 上表现落后;②是否有旗舰 Pro 模型规划——@PromptInjection 和 @wangzhi0467 都在问为什么跳过 3.1 Pro 的继任者;③token 效率提升的实际意义是否被过度宣传

展开评论
  • @zeeFloatingMonk (4): @JeffDean Jeff, why is Gemini behind Claude and chatGPT when it comes to coding tasks? It is such a overall powerful multi modal model so it falling behind the competition on agentic coding tasks is surprising...
  • @PromptInjection (1): @JeffDean What is the story behind 3.6 Flash? What happened to the successor to 3.1 Pro? @grok
  • @wangzhi0467 (1): @JeffDean Hey Jeff the flash models are super nice, but we all want the pro models that trade time for agentic and reasoning capabilities as well
  • @oswaldwedam (1): @JeffDean Oh my! This could not have come at a better time!
  • @sprk_77 (1): @JeffDean Thanks for the update. You and your team are the reason I am a premium subscriber. Please keep on with the good work.
415 赞 · 89 评 🔥

indie hacker 标杆 @levelsio 认为向 AI 描述需求和向人类开发者沟通区别已不大,且 AI 更快、更少抵触,引发大量反驳

I don't think there's much difference anymore between > Talking to a dev over the phone or chat to tell him what you want and > Talking to an AI over chat to tell it what you want The AI is muc…

受众观点:① 普通用户能否真正表达清楚需求的门槛(@marcospereeira 列出五个前提条件)② 价格差异和不被 condescend 的体验让 AI 胜出(@eloffd 的真实例子)③ AI 在安全和细节方面的隐形短板(@SSShken @RutkowskiHQ)

展开评论
  • @marcospereeira (7): you're assuming normies 1. know what they want 2. know how to express it 3. want to put the effort into expressing it 4. are willing to go through the process of iterating and providing feedback that comes with non-battle-tested solutions 5. don't care about community effects or…
  • @eloffd (6): @levelsio The difference in price is order of magnitudes cheaper with the AI and it never condescends to you or tells you no. Of course it will win. A friend of mine used to make a living creating websites for people. AI has already replaced that job.
  • @SSShken (5): @levelsio No matter what anyone says, AI isn't a narrow specialist, so it very often overlooks important and minor details that a developer would simply remember by chance based on experience
  • @RutkowskiHQ (4): You are not wrong, but this doesn't negate his point fully - people are still lazy, they also don't want to talk to developers as they don't want to talk to LLMs. Not to mention responsibility for things like PII and security breaches in your own vibed software, that can be trag…
  • @tibo_maker (4): @levelsio based on that, do you plan on changing anything to the way you manage your biz?
407 赞 · 33 评 · 30.2k 阅 🔥🔥

Nous Research 在 Nous Portal 上线 poolsideai 的 Laguna S 2.1 编程专用模型,两周免费试用,MoE 架构总参数 118B 激活 8B

The new Laguna S 2.1 model by @poolsideai is now free for 2 weeks on Nous Portal. At 118B total parameters with 8B active, it's quick to run and the most capable model they've released so far. Try Po…

受众观点:①如何访问试用——@NousResearch 自己在评论中给出了更多模型信息链接;②社区情绪明显积极——@HermesAgentTips 和 @1kartikkabadi1 都表达了对免费试用的热情;③视频制作技巧——@JamesOnEdge 好奇 Nous Research 的视频制作方式

展开评论
  • @NousResearch (20): @poolsideai More info on the model: https://t.co/ZVrUOCIQ59
  • @1kartikkabadi1 (3): @NousResearch @poolsideai nous the goat https://t.co/3jpO28Pz5v
  • @HermesAgentTips (3): @NousResearch @poolsideai Free model alert boysssss let’s go https://t.co/iu2VzTCsos
  • @JamesOnEdge (2): @NousResearch @poolsideai Please share your video skills for these awesome videos you produce!
  • @andr3barroso (2): @NousResearch @poolsideai New model to try for free https://t.co/W8UvYkyg4s
406 赞 · 23 评 · 53.6k 阅 🔥🔥

ML 研究员 natolambert 公开纠正 @benthompson 关于中国 AI 实验室在强化学习训练中使用 Fable 做 teacher 的说法,解释 RL 与蒸馏的机制差异

Yo @benthompson I'm sorry but the Chinese labs aren't using Fable / the strongest models as teachers during RL, that's not how distillation works. It wouldn't give that big of a lift (graders during…

受众观点:①Anthropic 指控 DeepSeek 使用强模型初始化 GRM 的背景——@teortaxesTex 认为这才是合理的训练方式;②natolambert 认为中国实验室实际在做什么——@natolambert 自己在评论中给出了推断链接;③是否应该邀请 Nathan Lambert 上 Stratechery 播客——@dipshady_ 和 @scorzeth 都呼吁这期访谈

展开评论
  • @teortaxesTex (20): @natolambert @benthompson that's exactly what Anthropic accused DeepSeek of btw and I think it makes sense for initializing GRMs
  • @natolambert (8): @benthompson What i think they are doing: https://t.co/6MYklpdrxZ
  • @dipshady_ (7): @natolambert @benthompson @benthompson Would love a @stratechery interview with Nathan on RL
  • @scorzeth (3): @natolambert @xeophon @benthompson I need a Statechery Interview with Nathan Lambert asap @benthompson
  • @fernavid (3): @natolambert @benthompson Reading through the Nemotron papers, it seems like generating synthetic data from stronger models is really useful. Why wouldn't using Opus/Fable in the same way give a performance lift? Or is your point mainly about RL specifically not being the place…
387 赞 · 58 评 · 55.2k 阅 🔥🔥

@rauchg 基于 Vercel AI Gateway 真实花费数据揭示:闭源模型占比从历史最高 97% 跌至 83%,Kimi(Moonshot)单周 API 花费份额暴涨 8.4 倍

Some model insights based on @vercel AI Gateway data 🧵 On June 27, Anthropic + OpenAI + Google hit an all-time high of combined spend share at 97.09%. Open models spend was barely a thing. The all-ti…

受众观点:① 开发者模型无关化趋势(@rehanbuildz 总结)② Kimi 快速崛起原因和持续性(@gpj 预告开源 Kimi K3)③ Google 和 OpenAI 在数据中处于同一水平的原因猜测(@pranaym0)

展开评论
  • @rauchg (62): All-time-low is today at 83.29%. The spend on closed models seems to be in freefall, but it's interesting to look at what exactly is happening. ◾ Over the past week, Moonshot (Kimi) went from 0.60% → 5.06%, an 8.4x increase. Today it grew its spend by 114.4% d/o/d
  • @charliermarsh (4): @rauchg @vercel Do you mean May 27 (2026-05-27)?
  • @rehanbuildz (2): @rauchg @vercel The biggest takeaway for me: developers are becoming increasingly model agnostic. If another model is faster, cheaper, or better for a task they will switch without hesitation.
  • @gpj (1): @rauchg @vercel Going to be interesting next week when folks get the open-weight Kimi K3 model.
  • @pranaym0 (1): @rauchg @vercel Would not have expected to see Google/OAI be at similar levels. Guessing video/image gen models are the reason and not text models.
361 赞 · 10 评 · 18.8k 阅 🔥

业内消息称 Gemini-4 旗舰模型已正式进入训练阶段,引发对 Google 能否在下一代模型中扳回颓势的广泛讨论

Gemini-4 is in training

受众观点:①现在才开始训练是否太晚——@TheFoolishPig 认为这是「宇宙中最悲观的消息」;②与 3.6 Flash 负面反馈的对比——@FriesIlover49 说 logan 直接跳到 Gemini 4 进行损控;③能否解决 coding harness 实用性问题——@tiberriver256 说至今没有 Gemini 模型在 coding 工具链上真正可用

展开评论
  • @TheFoolishPig (8): @scaling01 They just started? That is the most bearish thing in the universe of bearish things.
  • @FriesIlover49 (5): @scaling01 theres no way 3.6 flash is so bad, logan not only doesnt mention it, but goes straight to damage control, I hope Gemini 4 is good.
  • @lajoiedeslutins (3): @scaling01 excited by the progress is the corporate way of saying the loss curve went down and we screamed
  • @shaihulud43 (1): @scaling01 How long a training run takes these days
  • @tiberriver256 (1): @scaling01 None of their models have been usable in coding harnesses yet... Maybe this'll be the one
300 赞 · 21 评 · 17.4k 阅 🔥

AI 博主以「零悬念」讽刺 Gemini 3.6 Flash 基准测试结果,评论区汇聚开发者对 Google 此次发布的集体失望情绪

should I even post the Gemini 3.6 Flash benchmarks like there's 0 entropy

受众观点:①开发者对 Gemini 团队的整体信任崩盘——@Quipra_ 认为 Gemma 团队比 Gemini 团队更有价值;②与 Grok 4.5/Fable 5/GPT 5.6 的对比——@v3rbaa 建议直接去夸竞品;③antigravity 工具访问限制——@PratikAtX 指出连测试模型的 limit 都没有重置

展开评论
  • @Quipra_ (13): @scaling01 Fucking embarrassment for Gemini. What the fuck is this? I will put the Gemma team ahead of the Gemini team at this point. At least the Gemma team is giving something useful at what they are doing .
  • @v3rbaa (7): @scaling01 Better just remind how Grok 4.5, Fable 5 and GPT 5.6 are good 100% boring model from google
  • @PratikAtX (4): @scaling01 They cant even reset the limits to test the model in antigravity. https://t.co/pTruTXDfTA
  • @HTKonX (4): @scaling01 Don't bother https://t.co/LKB9IRVba2
  • @TimTeaFan (3): @scaling01 3.5 Flash-Light is the better release tbh.
HN H2
482 分 · 363 评 🔥🔥🔥

Google 在 Cloud Agent Platform 发布新版 Gemini 模型,HN 社区对基准测试数据矛盾和与竞品的性价比展开激烈讨论

https://console.cloud.google.com/agent-platform/publishers/g...

受众观点:①@yanis_t 认为基准不突出,切换理由不足 ②@dumberquestions 发现同一公告 benchmark 数字自相矛盾(DeepSWE 同时出现 65% 和 49%) ③@velominati 批评 Google 只跟自家历史版本比,回避与前沿实验室正面对比

展开评论
  • @yanis_t (0): The benchmarks are not particularly impressive. I suppose they needed to release something since the long pause. But not clear why would I use it now.
  • @dumberquestions (0): "..and in some benchmarks like DeepSWE by Datacurve, we observe up to 65%, all at a lower cost per output token." "3.6 Flash delivers higher precision with fewer unwanted code edits and reduced execution loops, as seen in DeepSWE (49% vs. 37%)" So which one is it? 65% or 49%?
  • @jgbuddy (0): It is both less intelligent and more expensive than GLM-5.2, while being closed weight.
  • @metalliqaz (0): Other discussion from a few minutes earlier: https://news.ycombinator.com/item?id=48993130
  • @velominati (0): Wow - Google does not even bother to show benchmarks of these models compared to the frontier and Chinese labs - only against previous versions. I'm not surprised. Having worked there for years it was amazing just how inwardly looking the company is.
HN H1
500 分 · 205 评 🔥🔥

HN 热帖讨论一款新发布的 AI 图像生成模型,支持最高 4500 token 输入并生成报纸等复杂排版,但权重未公开

Qwen-Image-3.0: Rich Content, Authentic Details, Deep Knowledge

受众观点:①@embedding-shape 追问权重是否以及何时会公开发布 ②@Mashimo 询问 16GB VRAM 下本地运行的最佳模型 ③@weird-eye-issue 发现 HTML meta 标签中大量 NSFW 关键词,引发对产品诚信的质疑

展开评论
  • @xiaoyu2006 (0): The blog write-up style is so casual haha.
  • @embedding-shape (0): Not a single word about when/if they'll actually release the weights for this, or am I missing it somewhere?
  • @Mashimo (0): > Supports up to 4.5k token input, effortlessly generating complex layouts such as newspapers, storyboards, and exam papers. Impressive. Btw, what is currently the best model to run locally on a 16GB Vram? Is it Z-Image Turbo?
  • @spwa4 (0): Appears to be closed-weights entirely. No word at all on any weights release.
  • @weird-eye-issue (0): The meta keywords in the HTML is very interesting. 100+ references to NSFW topics such as hentai, nudes, etc.
HN H6
148 分 · 155 评 🔥

OpenAI 在开源与闭源模型大论战高峰期宣布一项被 HN 社区评价为"最后手段"的重大财务或产品决策

Advertise in ChatGPT

受众观点:①@arm32 质疑这是否真的是 OpenAI 的最后手段,追问信息准确性 ②@Catloafdev 担忧 OpenAI 财务健康状况,认为此举本应在 IPO 之后才发生 ③@kamranjon 点出在开源 vs 闭源模型辩论最激烈节点发布是大胆的时机选择

展开评论
  • @moomoo11 (0): very cool
  • @snarfy (0): RIP
  • @arm32 (0): I'm confused. This is (publicly stated) the last resort for OpenAI, right?
  • @Catloafdev (0): Well that surely isn't a good sign of their financial health. I would have expected this to happen post-IPO.
  • @kamranjon (0): Pretty bold move to release this at the peak of the Open Models vs Proprietary Models debate. Feels like a really easy decision when it's framed this way.
625 赞 · 79 评 · 51.4k 阅 🔥🔥

开发者 @robj3d3 观察到 AnthropicAI 的 Fable 模型速度突然大幅提升,感觉回到了 6 月 9 日刚发布时的体验

Fable is way faster all of a sudden. I see what you guys are doing @AnthropicAI don't pretend we don't notice.

受众观点:① 速度提升原因猜测(@heyitzami 调侃,@cmdhightech 推测廉价计划降资源)② 性能是否被降级(@robj3d3 自补充「没感觉被 nerf」)③ 用户流失导致负载降低的反向猜测(@Tr_Appel)

展开评论
  • @heyitzami (38): @robj3d3 @AnthropicAI they got 3 guys working in a closet room replying to all the prompts
  • @robj3d3 (16): And no, it doesn't feel nerfed It feels like how it felt on 9 June again :'D
  • @Tr_Appel (12): @robj3d3 @AnthropicAI Could be the loss of users so the load is less? Probably not, but I wanted to say it.
  • @cmdhightech (11): @robj3d3 @AnthropicAI They took inference off the cheap plans and it improved?
  • @robj3d3 (4): Just shipped SuperX's API, CLI, MCP and X growth skill with it. Also made it free to try btw https://t.co/IK54SBNRdy
297 赞 · 44 评 · 16.7k 阅 🔥

Google AI Studio 上线 Gemini 3.6 Flash,输出定价降至 $7.50(原 3.5 Flash 为 $9.00),知识截止日期 2026 年 3 月,开发者社区反应冷淡

⚡️Gemini 3.6 Flash is now available in Google AI Studio, and it's cheaper than Gemini 3.5 Flash. Pricing: - Input: $1.50 - Output: $7.50 (Gemini 3.5 Flash: $9.00) Knowledge cutoff: March 2026. https:…

受众观点:① 价格相较历史版本已大幅缩水(@n30n_41 提 Gemini 3 Flash 时才 $0.5/$3)② 速度表现不佳(@ChrisGPT 说比 5.6 sol pro on cerebras 还慢)③ 知识截止日期停留 2026 年 3 月是否足够新(@ItzFiox)

展开评论
  • @iNoNamePlus (8): @ai_for_success Also Available in Gemini App :) https://t.co/V6QfVrhGMo
  • @ChrisGPT (6): @ai_for_success It’s literally slower than 5.6 sol pro on cerebras lol
  • @ItzFiox (4): @ai_for_success And this is March 2026 knowledge cutoff??? https://t.co/a0BI6oby6R
  • @notjazii (3): @ai_for_success where is 3.5 pro??
  • @n30n_41 (2): @ai_for_success Lol! Gemini 3 flash pricing was $0.5 and $3.
231 赞 · 18 评 · 14.3k 阅 🔥

@VicVijayakumar 指出 AI 消除编码瓶颈后,团队正在玩「瓶颈打地鼠」——代码审查信任度、测试覆盖、流程验证等新问题接连浮现

Coding used to be the bottleneck and isn’t anymore so now there is a lot of brain power being spent at companies to play bottleneck whackamole while tokens are cheap. Problem: code review Multiple le…

受众观点:① agent babysitting 成为新的隐性工作负担(@vladutxo)② 目前没人有好的解决方案(@sm0olr)③ 机构知识溢出上下文窗口的问题(@BrianVia 详细描述)

展开评论
  • @vladutxo (10): @VicVijayakumar the sneaky one: all this saves coding time, then creates a new job called babysitting agents. ideally not while babysitting toddlers
  • @sm0olr (8): @VicVijayakumar Man I’m in the middle of every single one of these at work right now. We don’t have a great solution for any of it yet.
  • @eternalmagi (1): @VicVijayakumar Yes, we are facing many of these, even with a relatively small product and a small team to go with it! I also don’t have any solutions, but it seems like as long as you are chugging along at a good pace and not breaking prod everything else can shake its way out.
  • @BrianVia (1): I mean our org still does some level of manual QA which means getting things through them (or in our case just her) but she has so much good institutional knowledge that she can anticipate breakages in one area from a change that is related but in a wholly separate part of the s…
  • @Bewinxed (1): @VicVijayakumar Problem: the agents don’t use the app as a human would and would miss a lot of obvious usability issues, gotta harness around this a LOT and somehow be able to quantify actual human user testing sessions into the test suites .
206 赞 · 15 评 · 4.9k 阅 🔥

AI 开发者社区持续等待 Gemini 3.5 Pro 未果,Google 反而先发布了 3.6 Flash,引发社区调侃 Google 产品发布节奏和版本命名逻辑混乱

AI community: Give us Gemini 3.5 Pro. GDM: https://t.co/BU5eyxVVGr

受众观点:① Gemini 3.5 Pro 上线时间表(@PhmThiDuy11 提到 Logan 已在聊 Gemini 4)② 当前高质量代码任务替代选择(@WelshYohans 提到 gpt5.6)③ Google 发布节奏引发的策略解读(@mostlyaboutai 认为先完善 Flash 是负责任的)

展开评论
  • @demian_ai (5): @ai_for_success it's coming https://t.co/kl740JE9dp
  • @mostlyaboutai (2): @ai_for_success You know what, I now feel that this is a good move. Fable Mythos Terra Luna Sol K3 etc. sort of skipped levels. The most responsible thing for google to do rn is to improve their next best thing or come up with something entirely new.
  • @PhmThiDuy11 (2): @ai_for_success So true. Logan even mentioned Gemini 4 while we are all waiting for 3.5 Pro
  • @Shubham9532 (2): @ai_for_success 😄😄
  • @WelshYohans (1): @ai_for_success i love the flash and flash lite models mainly to their speed, vision, audio ... we have gpt5.6 for progamming
153 赞 · 5 评 · 22k 阅 🔥

开发者开源 CLAUDE.md 任务路由规则和 tmux skill,解决 Fable 5 达到限额后多 agent 协作编排的实际问题

If you keep hitting Fable 5 limit Add a task-routering rule in CLAUDE.md I made a 16-min walk through of setup: - My CLAUDE.md for task-routering - Tmux skill for Fable 5 to control any other agent -…

受众观点:①CLAUDE.md 规则有效性衰减问题——@LucaCaponeX 指出规则越写越多但不知道哪条还在生效;②Orca CLI 与 Herdr CLI 工具对比选择;③用 Fable 做 orchestrator、Sonnet 做具体开发任务的分工架构(@JosefOndrejcka)

展开评论
  • @jasonzhou1993 (3): Also did a 1-hr deep dive workshop with @orca_build founder @JinjingLiang to dive deeper on how they ship hudrends of commits by orchestrating agents: https://t.co/2hL9HfpBg0
  • @jpawchan (2): @jasonzhou1993 there's better and more elegant solution for this: https://t.co/GjjmXKeWX5
  • @cky011 (0): @jasonzhou1993 最近一直在为这个事情焦虑,刚好刷到了这个,对我很有帮助👍
  • @LucaCaponeX (0): @jasonzhou1993 The part I've never solved is telling whether a rule still works. My CLAUDE.md keeps growing and I've no way to know which lines are earning their place and which ones the model quietly stopped following. A routing rule I'd trust. The rest I'm guessing at.
  • @JosefOndrejcka (0): @jasonzhou1993 110% agree fable as orchestrator. Sonnet or composer 2.5 for dev work.
89 赞 · 6 评 · 7.7k 阅 🔥

内容创作者分享先提取头部账号写作风格、再用 AI 生成同风格 LinkedIn 帖子的自动化内容工作流

Don't read this if you only post once a month. You'll burn $100 & hate me. But if you post weekly: ...this turns your own LinkedIn content into a machine that writes posts in your voice. — STEP 1…

受众观点:①先提取他人风格再训练 AI 的策略逻辑——@Blum_OG 指出大多数人直接用自己的数据导致输出感觉偏离;②纯图片帖子如何处理——@0xTreff 问当内容是图片而文字只是配角时这套方法是否还有效;③$100 成本门槛适用场景(周更以上才值得投入)

展开评论
  • @0xSlyth (0): @rubenhassid step 1 changes everything
  • @0xTreff (0): @rubenhassid how do you handle posts where the media is the whole point and the caption is just a throwaway line? does the skill still pick up on what makes those work?
  • @Blum_OG (0): @rubenhassid step 1 is underrated most people jump straight to their own data and wonder why the output feels off
78 赞 · 8 评 · 5.2k 阅 🔥🔥

Google 同日发布 Gemini 3.6 Flash 等三款新模型,主打比 Gemini 3.5 Flash 减少 17% 输出 token 并降低 agentic 任务整体成本

Google just released 3 new Gemini models. - Gemini 3.6 Flash - Role: High efficiency workhorse model designed for coding, knowledge work, and agentic workflows. - Efficiency & Cost: Consumes 17%…

受众观点:①各家厂商 mid-tier 模型定价战略对比——@jdanielcook 分析 Anthropic/OpenAI/Google 三家策略差异;②实际 agentic workflow 适用性——@Tibbzzee 认为 Flash Lite 适合 quick agent;③基准测试数据与实际成本性价比——@bspectacledGOAT 引用 DeepSWE 和 MLE Bench 数据

展开评论
  • @snakeyesV1 (1): @ai_for_success https://t.co/UlRHU8lTIb
  • @Tibbzzee (1): @ai_for_success Really excited to use Gemini 3.5 Flash Lite in quick agent workflows for orgs. It's a killer chat agent too.
  • @Hiraweb3 (0): @ai_for_success gemini's on the grind, but we still need moar decentralization
  • @jdanielcook (0): @ai_for_success Google is playing a different game. Anthropic tried to go up market and failed. OpenAI is maintaining pricing and lowering costs by using less tokens. Google is focusing on the mid-tier model because a year from now it will be all 95% of users will need.
  • @bspectacledGOAT (0): This enhanced efficiency is also combined with a lower price than 3.5 Flash. At $1.50/1M input tokens and $7.50/1M output tokens, 3.6 Flash reduces the overall cost per agentic task, making agents more cost-effective to build and run! but the question is how efficient is the mod…
54 赞 · 8 评 · 6k 阅 🔥

开发者用 GPT 5.6 Sol 在两小时半内全自动完成 macOS Sequoia 黑苹果虚拟机搭建,并计划扩展到通过 jetKVM 操控真实物理机

GPT 5.6 Sol was able to create a fully functional macOS Sequoia hackintosh from scratch full timelapse below https://t.co/PBasTSEwpU

受众观点:①下一步真机操控计划——@ryanvogel 自己说要用 jetKVM 连接 Windows 机器配合 codex 实现真机控制;②工具选择——@D_Is_Purple 问为何用 codex 而不是 OpenCode desktop;③虚拟机方案的局限性——@aks_nexus 指出 docker-osx 更简洁且 GPU 加速同样缺失

展开评论
  • @ryanvogel (6): i have bigger plans for this, a virtual machine hackintosh is pretty simple i'm going to wire up a jetKVM to my windows machine with a partitioned SSD and (hopefully) codex can control that from the jetKVM window
  • @ryanvogel (4): It took a total of 2hours and 37 minutes
  • @mweinbach (4): @ryanvogel oh fuck that's cool
  • @aks_nexus (1): @ryanvogel Isn't docker osx basically the same thing with less steps given you don't have gpu acceleration in the vm anyways
  • @D_Is_Purple (0): @ryanvogel This is your timelaps? Why codex and not OpenCode desktop
35k 赞 · 592 评 · 2245.8k 阅 🔥🔥🔥

开发者 Elie 用开源方式复刻了 Palantir 政府级情报平台,做成实时 3D 地球仪新闻聚合工具 World Monitor 并发布在 GitHub

Palantir sells governments a war room that costs millions a year. So, a guy named Elie just rebuilt it, put it on GitHub, and gave it away. It's called World Monitor. Open it and you get a live 3D gl…

受众观点:①开源情报工具实用性讨论(@aibasehub 形容 newsroom+OSINT board+disaster tracker 合一)②Palantir 定价护城河质疑,认为开源已威胁其功能层(@TheAIShrink)③反对声音:Palantir 真正卖的是 DoD 级安全合规,开源 dashboard 不在同一赛道(@LibertyLynx)

展开评论
  • @ihteshamali (1477): Here is the link for you: https://t.co/xlNTelkiAz
  • @aibasehub (536): @ihteshamali World Monitor is basically a newsroom, market desk, OSINT board, and disaster tracker glued to a globe. That is either extremely useful or extremely addictive. Probably both. lol
  • @TheAIShrink (433): @ihteshamali Palantir charges millions for the luxury version. World Monitor is the open-source equivalent that actually works. The moat is just habit.
  • @jonesrrrrrr (81): @ihteshamali Palantir sells trust within highly secure dod environments no one gives af about some shitty vibe coded dashboard retard
  • @LibertyLynx (52): @ihteshamali To say this "rebuilt Palantir" and poses a threat to them completely misunderstands what Palantir actually sells. https://t.co/v3jDD2iYRc
6k 赞 · 195 评 · 2875.8k 阅 🔥🔥🔥

用 Claude Code 分析六年积累的 4000 条 Obsidian 笔记,让 AI 挖掘孤儿想法和矛盾观点,相当于自动化的第二大脑

A 27-year-old pointed Claude Code at 4,000 Obsidian notes and woke up to an employee he never hired. 6 years of notes. Ideas, half-finished essays, 212 book summaries all rotting in a folder he stopp…

受众观点:①Claude Code 作为个人知识库「员工」的有趣比喻引发广泛共鸣(@AgentOrToy 调侃「roommate who reads your diary」)②对输出质量的质疑:AI 连接笔记可信度存疑,可能是幻觉出来的关联(@KijAkubovs86334)③实际应用场景:安全研究领域已有人做类似实践并公开测试(@pentesti)

展开评论
  • @AgentOrToy (451): @z0rynx bro hired an employee whose entire job is telling him he was wrong in 2021 thats not a second brain thats a roommate who reads your diary
  • @kungfoo___ (92): @z0rynx Annnd whats it fucking do?
  • @singularlab_ai (75): @z0rynx Fell asleep with an idea, woke up bankrupt lol
  • @KijAkubovs86334 (74): @z0rynx The 340 orphan ideas and 18 contradictions are just Claude confidently making up connections between your typos and abandoned book summaries. The vault didn't "think back" — it hallucinated back 📚
  • @pentesti (16): @z0rynx i built something similar for security research: 1,310 notes, 13,649 links, 403 sources. now my agents can search the vault, follow the links and pull the right context on their own. still testing it, but my notes finally feel useful. https://t.co/vPmCM3OTLg
2.8k 赞 · 240 评 · 284.6k 阅 🔥🔥

新产品 Monid 宣称以按请求计费方式替代 Apify,支持 agent 无需登录读取 X、Reddit、TikTok 等全平台社交媒体数据

We just killed Apify. Your agent can now read every social media platform. No logins. No subscriptions. X, Reddit, LinkedIn, TikTok, Facebook, Instagram, YouTube, Rednote, and even Amazon. Apify: $19…

受众观点:①对 Apify 高定价不满,对低价替代方案有实际需求(@chrisbradyuk 说「pricing gets a little silly with heavy usage」)②产品定位困惑:Apify 还列在 provider 列表里,用户不清楚 Monid 是平台还是聚合器(@sandeep_indie)③功能局限质疑:视频模型不支持参考图,非纯数据抓取功能不稳定(@leonhatori)

展开评论
  • @shengkun_ye (105): Try it now: https://t.co/zOxjv9aIYB
  • @chrisbradyuk (7): @shengkun_ye very intrigued. i really like appify but pricing gets a little silly with heavy usage.
  • @leonhatori (3): @shengkun_ye Creating videos via monid was way limited for me. Models do not accept reference images. How to improve that?
  • @sandeep_indie (3): @shengkun_ye i can see apify as provider on ur site so it is confusing are u like openrouter for apis?
  • @khushiirl (3): @shengkun_ye Banger product!
355 赞 · 101 评 · 44.7k 阅 🔥🔥

推文探讨把个人数字意识接入自我优化 AI 循环的构想,批判 Second Brain 工具沦为死链坟场

THIS FOUNDER JUST WIRED HIS ENTIRE DIGITAL CONSCIOUSNESS INTO A SELF-IMPROVING AI LOOP. Your Second Brain is probably just a graveyard of dead links and forgotten Notion pages. You save articles you…

受众观点:①@AIex_Al 提出知识库质量比 prompt 更重要的核心论点 ②@jurlycat 指出若输入数据是噪声则 AI 只会放大噪声 ③@CoraleeFreor 以"Notion 全是灰尘"共情普通用户的 Second Brain 实际状态

展开评论
  • @jurlycat (2): @kyroxxxq That loop will just amplify the noise if the input data is already garbage.
  • @AIex_Al (1): @kyroxxxq i'm starting to think the people with the best AI won't have the best prompts they'll have the best knowledge base,agree?
  • @Dipanshu_AI (1): @kyroxxxq That's wild! Seems like some serious next-level thinking going on there. Definitely makes you reconsider how we store information.
  • @CoraleeFreor (0): @kyroxxxq 救命 我Notion全是灰尘
  • @DorineDay26016 (0): @kyroxxxq no one comes close honestly!
Reddit R39
1567 分 · 289 评 🔥🔥

Anthropic 因盗版训练数据被起诉,潜在和解金额约 15 亿美元,引发 AI 公司版权合规与资金结构的激烈讨论

r/ClaudeAI ANTHROPIC GOT SUED

受众观点:①版权法律正义感(@Rn1k 认为 $1.5B 对 $965B 估值公司是杯水车薪,比自己付房租比例还低)②盗版 vs 合法使用训练数据的本质区别(@SetentaeBolg 指出这不是训练合法性裁决,而是非法获取内容的处罚)③Anthropic 财务状况与法律责任的矛盾(@moretti85 指出估值≠银行余额,公司未盈利)

展开评论
  • @Rn1k (539): Make a $965B company stealing copyrighted works, pay $1.5B when/IF you settle a case. I pay relatively more for rent, monthly.
  • @Emergency-Bobcat6485 (136): This would be more akin to squatting for years and then paying a token fine when caught and then keeping the house anyway
  • @SetentaeBolg (91): It's important to note, the settlement isn't for using copyrighted works to train their AIs. The settlement is for pirating those copyrighted works. This is not a decision on the legality of using copyrighted material you otherwise legally have for training: it's a decision made…
  • @moretti85 (37): valuation isn't a bank balance, anthropic has never been profitable: [https://isaiprofitable.com/](https://isaiprofitable.com/)
  • @kevin_cn_ai (37): "not profitable" isn't a cheat code to ignore copyright laws though. if any regular startup did this they'd be sued out of existence on day one😂
Reddit R41
658 分 · 155 评 🔥

非技术用户在 Claude Chat 中粘贴代码错误信息后触发 prompt injection 攻击,Claude 检测到恶意指令并输出警告,用户误以为账号被入侵的真实安全事件

EDIT: ~~I believe I have found a stable version of the code~~. So that was a fucking lie. We’re doing it from the top. So I’m not a very smart person, right? I took two semesters of coding in college…

受众观点:①prompt injection 机制是什么、为何能触发(@Big-Tip7095 第一时间判断是 prompt injection,@Violet2393 详细解释了工具权限和记忆更新机制)②Claude 有没有被成功入侵(@Sawyer_Anderson 提示要删除相关文件防止残留)③非技术用户该如何安全使用 Claude(@Violet2393 强调要理解权限才能授权)

展开评论
  • @Big-Tip7095 (243): Pretty sure this is a prompt injection attack on you.
  • @Violet2393 (135): While Claude was building your grocery tracker, it was using various tools to do so, depending on what you gave it permission to do. We can't see exactly what happened in your screenshots, but it looks like something instructed it to update its memory with a prompt injection. Th…
  • @Adventurous_Pea_2007 (59): Heard that. I have deleted all affected code, completely deleted the affected chat. Nothing was ever uploaded elsewhere. Also have deleted all browser cookies and settings related to Claude.ai. I think I know what happened. It won’t happen again. EDIT: it’s in the OP and has bee…
  • @Sawyer_Anderson (54): You should delete the files that were in that chat. The attempt came from the files themselves; Claude was reading something that you put in the chat it didn't comply because claude is a good boi but if you didnt delete the files you put up you still have whatever "this" is.
  • @Adventurous_Pea_2007 (42): I’m not stupid enough to put my personal information in here. But can you please explain further? How is this even possible?
Reddit R43
252 分 · 155 评 🔥

30 年资深工程师分享用 GitHub Actions 为多并发 Claude Code session 构建 PR 工作流——含 2000 个回归测试、LLM slop 检测和 Playwright 自动化,评论区大量质疑其基础工程素养

For context: I've been a software engineer for 30 years and I have two Max accounts and blow through both of them each week. What started as a pet project has quickly grown to a pretty sophisticated…

受众观点:①30 年经验工程师才搭 CI 的可信度(@lost12487 直接质疑,评论区最高分)②每月 $400 订阅费用换来的是否值得(@EnumeratedArray 质疑在 localhost 项目上烧这么多钱)③多 session 并行的实际工作流设计(少数人关注技术细节)

展开评论
  • @lost12487 (489): You’ve been a software engineer for 30 years and didn’t have a CI pipeline from the beginning? Also this isn’t a “sophisticated” pipeline. This is just basic gitflow.
  • @After-Regret-6609 (109): Wait, so you implemented *gasp* proper software engineering tooling! What a concept
  • @rydan (82): I've been an engineer for only 20 years and this post has me entirely confused.
  • @EnumeratedArray (71): Crazy that you're spending $400 a month on a localhost SaaS and you've only just implemented basic CI tooling with 30 years experience. Something ain't adding up there.
  • @finnjaeger1337 (64): fo real wtf
Reddit R53
52 分 · 123 评 🔥

Reddit 热门讨论帖征集独立开发者为自己真实需求而建的产品,吸引大量开发者分享项目和动机

What problem does it solve for you, and what made you decide to build it instead of using an existing solution?

受众观点:①自用产品持续坚持两年以上的内在驱动力(@MusicOfTheApes 花两年做音乐工具 Jam Lab,自己每天都在用)②语音笔记转任务的自动化刚需(@Positive-Valuable485 用 WhisperAct 解决自己语音备忘录堆积的问题)③笔记与思维导图结合的产品设计(@OwlCompetitive8891 自建 capira 只为满足自己的操作习惯)

展开评论
  • @MusicOfTheApes (17): The app I released 3 weeks ago is called Jam Lab (only iOS for now, been learning iOS dev since Covid in my free time), it's an app for musicians to generate random chord progressions for jam sessions or to work at home (there are many more things inside though). I've spent 2 ye…
  • @Positive-Valuable485 (8): Mine’s [WhisperAct](https://whisperact.com). I kept recording voice notes and telling myself I’d deal with them later… I almost never did. 😅 So I built something that turns those thoughts into reminders, tasks, and calendar events automatically.
  • @OwlCompetitive8891 (4): I use [capira.app](http://capira.app), it's a passion project - a note app with things that I wanted - it include mind maps. and I put in stickers in the note editor; characters that you can type along side. And better image handling so that its more enjoyable to place images in…
  • @imsrgadich (3): Installed it, I will use it 🎉
  • @Marshgrain (3): the mind maps + notes combo is something most apps get wrong imo, cool that you just built what you wanted
Reddit R8
133 分 · 113 评 🔥

SaaS创业者历经数月广告和冷触达无效,随机私信老同事意外获得首个付费订阅用户

I've been trying to sell for a few months now. The main thing i learned from this whole process is that your first paying customers are already in your network. I tried ads, cold outreach, nothing wa…

受众观点:①第一批付费用户几乎普遍来自创始人已有人脉网络而非投入最多的渠道(@Ninjishnu 观察'首笔钱几乎从不来自你花最多钱的地方')②冷触达和付费广告在早期获客中转化率极低的普遍规律③偶然性的人际连接如何触发真正的付费意愿

展开评论
  • @Capable-Aide-7860 (6): It’s a crazy feeling just got my first too, congrats enjoy it
  • @WonderfolioApp (6): Nice! I tried that too, my old coworker never ended up downloading the app 😂 You had better luck, hehe.
  • @[deleted] (4): [removed]
  • @Gold_Habit6491 (3): You gotta keep trying. Someone from your own network will bite.
  • @Ninjishnu (3): congrats! 🎉 funny how the first one almost never comes from the channel you spent the most money on lol
Reddit R44
239 分 · 95 评 🔥

用户感叹自己正在遗忘 Google 搜索技巧,AI 正在替代搜索行为,引发社区讨论"精准搜索"技能是否已成历史

This crossed my mind the other day. I used to know exactly how to search for things. Which keywords to use. Which sites to ignore. How to dig through forum posts. Now I catch myself opening Claude be…

受众观点:①Google 本身的质量下降是主因还是 AI 替代是主因(@ridablellama、@SurvivalHermit 认为 Google 先坏的)②LLM 搜索的局限性(@Dornwickler 作为记者仍然依赖论坛和社群获取灵感)③搜索技能的本质是否还有价值(@Back_on_redd 认为概念还在但工具变了)

展开评论
  • @ridablellama (91): when google actually indexed the full web and when there was a web to index it was an awesome skill
  • @Back_on_redd (54): Outdated skill now. The concept still holds - how operators work, search terms, related searches etc but for the most part you just feed it to a LLM/machine learning backed ssrch agent and they do the combing and refinement
  • @ElonMuskTheNarsisist (29): Google is all adds now. All it’s good for is searching reddit
  • @SurvivalHermit (28): The problem is that googling was getting worse all the time and that was only going to start happening faster. Every time new information was discovered someone had to put it somewhere. Those somewheres became hidden in an ever thickening forest. Googling was going to go away an…
  • @Dornwickler (16): Maybe it's the other way around, at least it was for me. I was good at searching but Google became almost unusable the last 1-2 years. As a journalist writing with time pressure at some point searching for data became more and more difficult. Irrelevant SEO slop, payed ads, and…
Reddit R42
349 分 · 92 评 🔥

用户质疑 Anthropic"领取免费额度"按钮暗藏自动续费开关,后更新确认 auto-reload 默认关闭但社区持续讨论 Fable 的高 API 费率问题

Am I reading too much into this, or is this a sly way to get us to accept a change in agreement? Anyone reading the title and the big 'Claim free credits' button might accept without thinking too har…

受众观点:①$100 额度是否会自动触发扣费(@Clean_Hyena7172 解释了开关逻辑和忘记关闭的风险)②Fable 在 API rate 下的性价比(社区整体认为 Fable 按 API 计费不划算)③最优使用策略(@arankays 建议设上限并关 auto-reload,把额度留给 Opus 重型任务)

展开评论
  • @battle_pantZ (95): This will be gone in 1hr don’t worry. It will open it much sooner than you expect
  • @sardinecaffeine (66): I think it only adds to your bill if you turn auto-reload on. By default, it will stop using usage credits once the $100 has been used.
  • @hyperrealists (62): The fuckers gave me only $75 😅
  • @Clean_Hyena7172 (62): Yes, you have to have extra usage enabled to claim/use credits, if you don't disable it after the credits are used then it will add to your bill any time you go over the limits of your subscription. Anthropic will likely make a lot of money off of people who forget about it.
  • @arankays (38): All you have to do is set a limit of 100 and disable auto reload. Then just enjoy your 100 bucks of Fable (prolly only a couple prompts but better than nothing). Instead of wasting your extra usage on Fable I'd HIGHLY suggest having Opus or Sonnet run through one of your codebas…
Reddit R47
141 分 · 84 评 🔥

刚订阅 Claude Pro 年付的用户分享使用 Opus 4.8 High 模式的真实体验,评论区围绕 Opus 4.8 的性价比和与 Sonnet 5 的对比展开讨论

It's me. I'm the one that everyone will roll their eyes at. I just subscribed to pro (annual.. I know.. what a moron.. I'm a sucker for a discount) a few days ago. Why? Well.. I'd been an avid sonnet…

受众观点:①Opus 4.8 的性价比定位(@a1454a 指出是当前 intelligence per dollar 最高的模型)②编程场景的表现(@IzodCenter 称 Opus 4.8 on High 写代码非常好用)③版本间性能对比(@throwaway464391 详细对比了 4.5/4.6/4.7/4.8 的差异,期待 Opus 5)

展开评论
  • @a1454a (64): Not crazy. Opus4.8 is currently Anthropic’s most efficient model. In terms of intelligence per dollar.
  • @IzodCenter (51): Opus 4.8 on High is wonderful for coding
  • @personalist (19): Definitely moreso than sonnet 5. Sonnet is at or near the top of artificial analysis’ “cost per task” benchmark IIRC
  • @throwaway464391 (17): For what I do, I never saw a backtrack in performance from 4.5 to 4.8. 4.7 was maybe a sidegrade from 4.6, but 4.8 was clearly better. 4.6 may have had a better "personality" (though I don't really remember, it's been 3 months since I used it) but I don't really have complaints…
  • @kapsnik (15): Opus on high is just high, on opiates presumably
Reddit R7
318 分 · 76 评 🔥

2026年SaaS创始人基础设施选型的真实争议:裸金属服务器低成本路线对抗AWS云服务弹性扩展

r/SaaS Pov: you are a SaaS founder in 2026

受众观点:①裸金属vs云服务的成本对比(@Dadragonfaier主张15€/月裸金属足够,@opshack声称AWS可低至$0.5/月)②MVP阶段基础设施选型是否值得花时间优化③按调用次数计费的云服务在bot滥用场景下的成本失控风险

展开评论
  • @Dadragonfaier (41): For small, medium and big scale I just use bare metal servers, they start from 15€/month and easy to scale with your project
  • @vidiludi (37): Jokes on you if you use AWS. A $10 / month webhost is all you need for your MVP.
  • @Impossible-Owl7407 (19): And than some bot does multiple billions calls to you
  • @Impossible-Owl7407 (16): Server is not billed by number of calls
  • @opshack (13): Way too expensive. I can setup a fully functioning product on AWS close to $0.5 / month
Reddit R9
107 分 · 70 评 🔥

独立开发者为妻子开发护肤App,通过坚持发布TikTok内容获得首笔付费,情感冲击引发广泛共鸣

Using Voice to Text, so words might be a bit messed up. I built a skincare app for my wife and I’ve been having serious doubts about the app, but I’ve been posting TikTok’s consistently. Just trying…

受众观点:①首笔付费对独立开发者情感支撑和坚持动力的实际意义(@FreeboForOperators 分享了自己半夜发布后早起看到收款通知的体验)②持续内容更新(TikTok日更)在应用冷启动阶段的实际获客效果③开发给真实需求用户vs为市场臆测产品之间的成功率差异

展开评论
  • @FreeboForOperators (9): I know the feeling brotha! I stayed up until 4am launching mine and deploying our embedded checkouts on my first clients site. woke up a few hours later to a reservation. An hour later, another! I woke my roomate up, called my mom, it feels good to have a plan come together.
  • @Status_Reception7356 (3): It's an interesting thing to see something you've put mind body and soul into get validated by a paying customer. I spent months in beta waiting on api approvals and now I have MRR of $1,068. I'v brought on an account manager to help me sell it and the feedback so far has been o…
  • @Status_Reception7356 (2): I offer an enterprise package on my platform that is $999 per month. That customer along with a single use customer at $69 per month put me over the $1,000 mark pretty quickly. The number makes it seem like a lot of customers. I just have an expensive product lol.
  • @Hour-Appointment2684 (2): What's your saas/problem that solves to pay more monthly??
  • @manateecoltee (2): Congrats Brother you earned it, consistency is key we are in a marathon not a race!
Reddit R51
144 分 · 51 评 🔥

跑步实时位置和心率直播应用,支持 Apple Watch 和双端,类似 Twitch 但面向跑步运动的流媒体平台

Live stream your runs. Working for Apple watch, iOS, and Android. \[Update\] [More Info](https://kailo.fit/products/live-stream) | [iOS](https://apps.apple.com/app/getfast-ai-fitness/id6752672094) |…

受众观点:①实时位置暴露的人身安全风险(@ZeidLovesAI 最高票提出被跟踪和暴露日常行踪的担忧)②展示数据的真实性(@Baldtazar 质疑截图中 155bpm 的"轻松跑"是否为真实数据)③现有隐私机制是否足够(@tommy-getfastai 介绍四种模式但承认位置遮蔽未实现)

展开评论
  • @ZeidLovesAI (29): How are you going to handle people not being stalked, etc? There should be at least tools for streamers to protect themselves somehow, seeing as this is going to make it very easy to see someone's routines.
  • @Baldtazar (19): Easy run on 155bpm? Is it real or placeholder data?
  • @tommy-getfastai (16): Thought it'd be realistic to show people running way too hard on their easy runs 💦 😬
  • @tommy-getfastai (12): Currently there are four privacy modes: \- public \- unlisted, but available with link "/watch/<session\_id>" \- friends only (need to accept them as friend first) \- groups only (can grant view access to an entire group) It's set up so you can remove certain streams, like…
  • @coffee-milk-tea (10): That's a cool idea and a really nice interface too!
Reddit R13
18 分 · 49 评 🔥

B2B垂直SaaS创始人担忧竞争者抄袭产品功能,评论区揭示发行渠道和信任才是真正的护城河

Recently launched my B2B vertical SaaS, pouring my 20+ years of experience. Started with insta ads and received 140 leads, a lot of appreciation, in conversation with prospects. What freaks me out is…

受众观点:①'分发渠道才是护城河'vs'产品功能是护城河'的底层逻辑(@SaltMaker23 详细论述,@JouniFlemming 指出执行才是价值)②20年行业经验作为差异化获客资产的真实价值③担忧被抄袭导致不开放试用反而错失验证产品价值机会

展开评论
  • @JouniFlemming (55): Here is the uncomfortably truth: Ideas are worthless, it is the implementation that matters. And relating to that, what most people do not understand is that it is usually the boring ideas well implemented that make the most money. Stop worrying about being copied. It is most li…
  • @SaltMaker23 (43): If you can be copied that easily, the problem isn't the copying, it's that you're assuming the product is the moat. It usually isn't. Everyone knows what Salesforce does; knowing it doesn't let you beat them. A real moat lives in acquisition and distribution, and those are tied…
  • @Prior_Night_985 (7): Distribution is the main strength you have. I could make a phone and copy the iphone but who knows it? Who trusts it? Who will buy it? They can copy you but they wont have your trust and customers
  • @Kemerd (5): People who always think their ideas are gods gift to the earth and spend all their time worrying about people copying their (usually garbage) ideas, never actually accomplish anything I’ve found. Meanwhile, open source flourishes..
  • @Quick_City_5785 (3): I have just begun, I have no relationships , just building some awareness. But yes, I am drawing inspiration from your comment. Thank you so much
Reddit R56
24 分 · 48 评 🔥

Reddit 讨论帖征集能真正收到钱的最小副业项目,汇集大量开发者的真实变现案例和经验教训

Not looking for million-dollar stories. I’m more interested in small projects that proved someone would actually pay. What was it, and what did you learn from it?

受众观点:①验证付费意愿的最短路径(@Creepy-Surprisee 最高票指出"人们为解决真实问题付钱,不为复杂度付钱")②技术好但营销是最大瓶颈(@bithereumza 有了销售但营销让人精疲力竭)③无意间靠 SEO 获得流量的案例(@100-days-of-code-io 靠 LinkedIn 游戏攻略意外获得大量搜索流量)

展开评论
  • @Creepy-Surprisee (22): Biggest takeaway was that people don’t pay for complexity they pay for solving a real problem quickly
  • @bithereumza (6): i have built migiapp.com, it’s a small snippet and ai prompt management library with templating (variable substitution) and a global hotkey to access your snippets instantly. I had a handful of a sales in the beginning but I am finding the marketing for it absolutely soul crushi…
  • @100-days-of-code-io (4): I love puzzles and one day randomly decided to post linkedin games solutions. At that time there were no websites covering linkedin games. Didn't do any marketing or seo but it still got significant traffic from google. Now I'm trying to build something around that.
  • @Mrtowel22 (3): I built Reelstack (https://apps.apple.com/us/app/reelstack/id6768630222) because I couldn't find a streaming tracker that was dead simple, no subscriptions, showed me where the shows were available, and notified when new seasons were available.
  • @Basic_Bad6389 (3): This is a great example of “build for a customer, not an idea.” Marketing is a skill you can learn,but finding a problem people genuinely want solved is much harder.Have you gotten any paying customers yet?
Reddit R48
81 分 · 42 评 🔥

Pro 订阅者提醒:领取 $100 免费额度后 Claude 自动开启用量消耗开关,达到小时上限后会静默扣除额度,建议手动关闭避免意外消费

For those of us that got the $100 credit, Claude automatically enables “turn on usage credits” toggle. So if you hit your hourly limit, even if using Opus, it starts draining your $100 credit. I was…

受众观点:①Fable 在 API 计费下的高消耗问题(@Seandelorean 指出烧的又快又低效,Anthropic 在展示 Fable 不值 API 价格)②对 Anthropic 计费策略的不满(@granite_vortex 考虑换 ChatGPT,认为 $100 额度是操纵手段)③实际操作建议(@devedander 建议双订阅,@PsychoticDreemurr 指出只有同时开了自动支付才会真正扣费)

展开评论
  • @Seandelorean (66): It burns through it so fast and inefficiently that Anthropic is basically just showcasing how not worth it it is to use Fable at API rate
  • @granite_vortex (28): The way fable isn’t available as part of the regular usage of pro level licenses (even if rationed) is enough that I’m weighing up jumping ship to ChatGPT. The $100 credit is a bit of manipulation in line with the dragging out the time you could use fable a free days at a time.…
  • @AdditionalWorkInc (11): so fucking true
  • @devedander (11): I ended up just getting both. It's nice to have one to fall back on when you hit a limit on the other. $40 a month is way better than $100
  • @PsychoticDreemurr (9): No, only if it also enabled automatic payments.
Reddit R50
213 分 · 39 评 🔥

独立开发者做的 AR 增强现实 WiFi 信号可视化安卓应用,支持实时热力图和平面图扫描导出报告

This started as a passion project but I ended up creating an entire app. It has a lot of visualizing and reporting tools that I was not able to find all together in a single application. It is both u…

受众观点:①商业变现路径(@fine_doggo 建议按扫描规模分级收费,家庭免费、商业付费)②iOS 缺失造成用户流失(@peepdabidness 发现是 iOS 用户后直接放弃)③LiDAR 与 ARCore 的技术权衡(@that_cant_be_right__ 表示 iPhone LiDAR 会让 AR 精度大幅提升但暂不打算移植)

展开评论
  • @The_Time_Lord (17): Well that is neat
  • @fine_doggo (12): The business aspect of his is huge, for someone like me who installs Wifi based IoT devices in sheds and warehouses, where WiFi signal is very crucial. I like how the app is free but on the other hand, I feel viable business loss for you. I'd suggest limiting floor maps, no of s…
  • @that_cant_be_right__ (6): Give it a try - I appreciate any feedback
  • @peepdabidness (6): Very cool!!! Will be trying it out Oh shit no I won’t, iOS
  • @that_cant_be_right__ (5): This is a sore spot for me for a few reasons. 1. I just don't have an iPhone to build and test with 2. Modern iPhones have actual LiDAR which would make the AR part of this SO much more useful and accurate. I kind of went into this, just knowing I will be making it for monocular…
Reddit R12
21 分 · 37 评 🔥

独立开发者从零启动AI面试辅导SaaS,通过按日按天短时段定价匹配购买心理,七个月实现万元月收入

I started an AI interview SaaS in January 2026. It’s bootstrapped and I’m the only person working on it. This month it crossed $10k MRR. Not sure if this even the right way to say it since my pricing…

受众观点:①短期高强度使用需求如何决定定价模型(@Smooth_Doubt2290 建议通过扩展用户旅程而非改变计费方式提升LTV)②季节性需求峰谷对SaaS MRR可预测性的实际挑战③AI UGC视频展示产品使用场景比功能列表更有效地转化用户

展开评论
  • @Smooth_Doubt2290 (4): Congrats on the milestone! One thing that stood out to me is that your pricing seems to fit the buying moment rather than the value metric and that's probably why the short passes are outperforming subscriptions. In my experience while advising founders , they often try to "fix"…
  • @Excellent_Can_3480 (3): Honestly, it wasn’t very scientific, and I’m still testing it. I looked at how customers use the product, tested a few simple options, and worked backward from the AI costs to leave a healthy margin. The larger bundles didn’t work well, so I kept the pricing simple.
  • @Excellent_Can_3480 (2): u/No-Meringue5430 not sure why I cant reply to you but here
  • @Excellent_Can_3480 (2): thanks!
  • @Impressive_Night_757 (2): I’d start by matching price to value. Quantify the financial value the customer receives where possible, whether that’s revenue generated, time saved, costs reduced or risk avoided, and compare that with what direct and indirect alternatives cost. From there, set an initial pric…
Reddit R26
35 评 🔥

webdev 社区讨论为什么 AI 相关帖子总被踩,评论区揭示反感原因是管理决策滥用 AI 而非工具本身,以及 AI 帖泛滥导致的信息噪音疲劳。

Hi webdevs, looking at the feed, almost all posts which announce AI being now used within a software get downvoted. I would like to know the the reasons for your rejection. Is it because: * you think…

受众观点:①反感对象是滥用 AI 的管理决策而非 AI 工具本身(@L24D,score 20,最高票)②AI 相关帖子数量泛滥导致社区疲劳(@cute_as_ducks_24,score 15)③AI 生成的 slop 网站质量问题(@Zestyclose-Oven-7863,score 5)

展开评论
  • @L24D (20): It’s hostility against putting experienced people aside by brain dead management who’s drank on AI hysteria. A tool and LLMs themselves are not the problem
  • @cute_as_ducks_24 (15): Maybe because there is too many posts about AI. Like just confine to the sub reddit that dicuss about AI rather than posting here and there. Even as a developer I am tired of so many posts about that give no value. Not against AI, but simply so much spam posts to the point now I…
  • @JohnSane (8): > that preforms well in every area. ui, performance, scaling Well that is some standard. I have never, in my 25 years as a webdev, made such a website.
  • @Slight-Act-9024 (6): stop before you slop
  • @Zestyclose-Oven-7863 (5): Ai really only produces slop websites. i have yet to see a unique website created entirely by ai. that preforms well in every area. ui, performance, scaling etc.
Reddit R11
24 分 · 34 评 🔥

Dario Amodei六个月前预测AI将替代开发者,社区调查开发者实际被替代的主观感受程度

Hey, developers! On a scale from 0-10 how replaced are you?

受众观点:①AI对开发工作实际效率提升的程度(@RigidlyCurly 指出复杂任务AI反而让自己花更多时间清理烂摊子)②AI带来的是减少所需开发者人数还是提升单人产出上限③开发者就业市场实际影响时间线与CEO预测的偏差

展开评论
  • @vessoo (11): Somewhere around 4-5, probably 5-6 by year end. Maybe have another year or two until whole teams get replaced by one person. Job not going away but 10:1 reduction in number of people needed for the same outcomes seems plausible to me in the near future
  • @RigidlyCurly (7): Put me at a solid 2. AI has made me way faster at boilerplate, but every time I try to lean on it for anything complex, I end up cleaning up its mess longer than I would have spent just writing it myself. The "10x developer" hype assumes you're shipping greenfield CRUD
  • @MappBook (3): You can fill the poll in image here if you want [https://www.mapster.io/s/098cddb9](https://www.mapster.io/s/098cddb9)
  • @reward72 (2): Exactly. AI is moving the goalposts. What was hard is now easy, but the next frontier is into making what was previously impossible (or cost prohibitive) possible (or affordable). AI is a savant idiot, a serious case of Duning Kruger with ADHD, Alzheimer’s and no long term visio…
  • @AnUninterestingEvent (2): I think every software company will still have developers, but a fraction of the developers they had before.
Reddit R45
296 分 · 31 评 🔥

Anthropic 发布 Claude 新功能"Teach a Skill",用户录屏演示操作过程并配音,Claude 自动将其转化为可重复执行的技能,社区开始讨论 token 消耗和潜在应用场景

have you tested it yet? how is the token usage?

受众观点:①功能定位类比(@Salmonberrycrunch 立刻类比 Excel"录制宏",这个比喻引起广泛共鸣)②对人类劳动的哲学冲击(@Rebmes 感叹工厂工人在组装替代自己的机器)③技能转化的技术实现(@PixelByt3 好奇如何将操作步骤映射成可执行技能)

展开评论
  • @Salmonberrycrunch (70): I was wondering when excels "record macro" will be applied here. Makes a lot of sense
  • @ProtoCore-Dustin (46): Claude… get me 99 runecrafting.
  • @Rebmes (28): I understand every single prompt sent is essentially the same as this but man if this doesn't feel like factory workers assembling the robots that replace them.
  • @Adrontion (14): This is the highest value data you could give. If its domain specific, i dont see how you´re not getting paid for this, I know textile workers are in India.
  • @PixelByt3 (10): I am so curious about how it maps those steps into a clean skill
Reddit R30
78 分 · 30 评 🔥

独立开发者发布 Mindwtr,一款 AGPL 开源的 local-first GTD 任务管理器,支持 Docker 自托管同步服务器,并内置 MCP server 供 Claude Code、Codex、Gemini CLI 等 AI agent 调用。

Hi r/selfhosted — I’m the developer of Mindwtr. Mindwtr is a free, AGPL-3.0-licensed, local-first GTD task manager for Windows, macOS, Linux, Android, iOS, and the web. I built the self-hosted backen…

受众观点:①产品对非技术用户的易用性和成熟度评价(@olluz 认为产品被严重低估)②AI 辅助开发的比例和可信度争议(@dongdongbh 回应 UI 来自真实用户迭代而非 Claude 生成)③与 Nextcloud Tasks 等现有工具的差异化定位(@Eric_12345678 询问对比)

展开评论
  • @olluz (11): Thank you so much for Mindwtr. Been selfhosting it for a while and using it together with the iOS app. It works really well and I have the feeling that the app is too unknown for how good it actually is
  • @dongdongbh (6): Haha, fair joke — but no, most of the design wasn’t made by Claude. The UI has gone through many rounds of iteration based on real usage and feedback from thousands of users (you can see them on GitHub issue history). AI helps with parts of the implementation and wording, but th…
  • @dongdongbh (5): The public history includes about 1.3k GitHub stars, 638 closed issues, many open issues and discussions, and around 40k GitHub release downloads. There are also thousands of installs through Google Play, F-Droid, the App Store, etc. Feedback comes from several places: GitHub is…
  • @Eric_12345678 (4): It looks good and interesting. What's the difference with NextCloud Tasks, for example? Also, can you import from NextCloud Tasks?
  • @dongdongbh (3): Thank you, that really means a lot. Mindwtr is completely free, so there isn’t really a marketing budget behind it. It also feels harder now for new apps to earn trust because people have seen so many quickly generated, poorly maintained projects. Most people discover Mindwtr th…
Reddit R55
33 分 · 28 评 🔥

独立开发者分享全职工作之余每天凌晨 3 点起床做副业的疲惫状态,附带推广 AI 纹身生成网站 ai-tattoos.com 并求营销建议

How do you handle the constant stress of side projects? I get up at 3 in the morning, work 5 hours, get ready for work, work 9 hours, some errands, and back to sleep. Can't sleep longer than 5-6 hour…

受众观点:①副业开发者的精力透支共鸣与坚持动力(@No_Ninja_5063 和 @Keeyzar 相互鼓励,"because we love it")②产品商业化的横向延伸方向(@Hot-Adeptness7155 建议按需打印临时纹身并授权给纹身师变现)③欧洲监管合规对产品扩展的制约(@Keeyzar 提到欧洲临时纹身按化妆品监管限制了 POD 路线)

展开评论
  • @Hot-Adeptness7155 (11): Would make sense to integrate with an API to have a temporary tattoo printed on demand and shipped to the customer so they can try it in real life. Then license the tech to tattoo artists. Reduce regrets.
  • @Keeyzar (2): Yes, had that in mind in the past. But stopped because in Europe temporary tattoos are cosmetics and tightly regulated. While on Amazon no name shops sell it without considering law, as they're outside of Europe and just close shops and open again. I will evaluate that again, th…
  • @No_Ninja_5063 (2): Cool project, Genuinely slick website, congrats. we do it because we love it ! Keep at it.
  • @Keeyzar (2): Yes. because we love it.
  • @Hot-Adeptness7155 (1): Find an EU-based POD tattoo supplier?
Reddit R52
83 分 · 25 评 🔥

用打字日本地铁站名驾驶列车的趣味学习游戏,技术上实现了多种假名罗马字输入引擎并支持 42 条线路

**What it is:** [railtyping.com](https://railtyping.com) — pick a rail line, a starting station and a direction, then type each station name to move the train one stop. Finish the line and you get to…

受众观点:①有序序列比随机词表更有助于记忆的学习设计理念(@justpixel 明确认同并表示自己也在做日语学习项目)②站名数据准确性(@Haizk 发现两处具体站名拼写错误)③假名字体渲染显示问题(@AndrewNggg 反映振假名被压缩到第一个汉字上)

展开评论
  • @AndrewNggg (3): super cool concept! I love it! but.. do check your station furugana layout? the hiragana gets all squashed into the first kanji
  • @Federal_Snow8054 (1): Thanks for the love! I'm constantly working on improving this website.
  • @justpixel (1): This is a really cool idea. Using a railway line as an ordered sequence feels much more memorable than typing random words. I’m also building a Japanese learning side project, so I know how difficult it is to make practice useful without turning it into another repetitive study…
  • @Haizk (1): https://preview.redd.it/noyyjogdrkeh1.png?width=349&format=png&auto=webp&s=1df93bfb59e484495103816138191c24841392bb I think it should be \`shinookubo\` instead
  • @Haizk (1): https://preview.redd.it/433b45hprkeh1.png?width=441&format=png&auto=webp&s=77053d7d96b942b0c32e50cd728c014761ce03be and this one is \`takanawage-touei\` right? or did I do something wrong in the settings? default settings btw, did not change anything
Reddit R49
238 分 · 22 评 🔥

周末项目:为 GitHub 个人主页贡献图实现真实 Snake 游戏逻辑——包含寻路算法、自体碰撞、纯 SVG 无 JS 依赖,一个 Python 文件加 GitHub Action 即可部署

Weekend project. The contribution-graph snake exists already, but it just crawls in a fixed pattern. Mine plays a real game of Snake: it pathfinds to your commits, avoids its own body, grows as it ea…

受众观点:①路径规划算法的健壮性(@Frosty-Reed-6618 质疑高 commit 密度会导致 softlock)②和同类项目的对比(@chayanforyou 提到自己在用 Platane/snk)③项目本身的工程质量(@Coolness1234567894 称赞用 Claude 辅助路径算法实现)

展开评论
  • @Frosty-Reed-6618 (17): pretty sure the pathfinding softlocks itself on anyone with a high commit run lol
  • @Last_Bad_2687 (9): No one with high commits has the time to play snake anyway
  • @chayanforyou (6): Cool. I'm using this one [https://github.com/Platane/snk](https://github.com/Platane/snk)
  • @TokerCoughin (2): It's delightful! No notes.
  • @Coolness1234567894 (1): Pretty fun take on snake, nice usage of claude to help with creating the pathfinding for better accuracy. I like it!
Reddit R46
186 分 · 20 评 🔥

用户将 GPT 对博客草稿的反馈粘贴给 Claude Opus 4.6 请其筛选,Claude 先吐槽 GPT 建议质量再接受约一半,引发社区"AI 互评"的趣味讨论

No custom instructions.

受众观点:①Claude Opus 的"个性"和自信(@dusantm 解释 Opus 4.6 先吐槽再接受,@Old-Professional4902 称之为"peak AI drama")②多模型互评的系统性问题(@LessRespects 指出让 LLM 审查 LLM 输出会导致永远找问题的 hallucination)③Claude 版本间的性格差异(@dusantm 调侃"4.6 有资历,有资格这样说话")

展开评论
  • @dusantm (43): Context: I pasted GPT's feedback on a blog draft and asked Claude what's worth applying. To be fair, it then accepted about half the suggestions. But it needed to get that off its chest first. Good old Opus 4.6 :)
  • @dusantm (41): 4.6 has seniority, it earned the right to talk like this :D
  • @Old-Professional4902 (25): Opus roasting GPT's suggestions before even reading them is peak AI drama
  • @GameboyGenius (17): Claude is being sassy.
  • @LessRespects (9): If you pasted responses back and forth between separate threads/models asking what’s wrong it will ALWAYS find something wrong and never settle on a complete response. This is an issue we’ve been trying to resolve with LLMs for years. If you simply infer something may be wrong,…
Reddit R58
18 分 · 19 评 🔥

独立开发者分享上线第一天就获得首个年费付费用户的里程碑经历,同时保持理性认知这只是一个数据点

I'm a little too excited right now. I've been building small SaaS tools for years, mostly just for myself, and never really thought anyone else would want to pay for them. One of my good buddies kept…

受众观点:①从零到一里程碑的情感共鸣和鼓励(@BP041 @Jash-6898 分别恭喜并分享经验)②尽快与第一个客户深度沟通的操作建议(@Jash-6898 建议询问购买动机、差点放弃的原因、预期)③年费 vs 月费作为信任程度信号的解读(@Marshgrain 认为年费代表真正信任而月费是试用)

展开评论
  • @Specific-Sky-9924 (1): Hey, congrats on your first paying customer! That's a real milestone, and landing a yearly plan on day one is no small thing. Your post resonated; the 'one datapoint' honesty is exactly the right mindset. We're a young startup working on getting our own SaaS off the ground, so I…
  • @BP041 (1): That's awesome — congrats! The jump from zero to one is the hardest, but the next ten are a different kind of grind. Keep talking to that first customer obsessively; nothing else will tell you faster what actually matters. Enjoy the win though, you've earned it.
  • @Jash-6898 (1): first off, congrats. that first payment hits different. the biggest thing i'd focus on now is talking to that customer. ask why they bought, what almost stopped them, and what they expected. one conversation can be worth more than 100 analytics events.
  • @WonderfolioApp (1): You must have a good friend then. My friends have free access and never offered to be a customer 😂 I would buy my friends app though to support him…
  • @Marshgrain (1): yearly sub on day one is a strong signal tbh. someone paying monthly is testing you out, someone paying annually actually trusts the thing. id focus on understanding what problem they were solving when they signed up, that framing will shape everything going forward
Reddit R38
18 评 🔥

独立开发者出售旗下副业组合——含 1.3M 曝光的病毒式 SaaS($1,400 ARR)及五位数收入的授权脚本,为专注新 B2B 主业而清仓

Over the past year, I built out a small portfolio of online businesses on the side, a viral-friendly SaaS with real organic traffic (1.3M impressions) and paying users ($1,400 ARR), a high-ticket scr…

受众观点:①出售原因是否合理(评论区质疑为何不雇人代管)②组合里的 SaaS 产品是什么、能否值回价格(@IReallyHateAsthma 直接问产品名)③创始人的新 B2B 项目方向(@Odeh13 解释需要全力投入)

展开评论
  • @IReallyHateAsthma (3): What’s the SaaS?
  • @SellAffectionate9670 (2): Sounds like you've built a solid portfolio. What made you decide now's the right time to exit instead of hiring someone to run it?
  • @Odeh13 (2): I mentioned that I'm building a 1M business, it's B2B, on a much larger scale of any that I've built. It's requires all my focus, resources (time, energy, and money)
  • @Odeh13 (2): Thank you buddy!
  • @digitalwankster (2): I'm interested. Your posts are hidden but I'm pretty sure I remember you posted your whatthefood project a while back.
Reddit R10
30 分 · 17 评 🔥

面向隐私和开发者工具用户的SaaS增加终身授权套餐,当天4笔销售即翻倍历史总收入

Not much more than the title. My niche (privacy/utility/developer tools) is pretty averse to subscriptions, and I got that feedback directly from customers. my annual rate is $9.99/yr, so I made the…

受众观点:①不同用户群体对订阅制vs买断制的根本态度差异(@SpaceJeans 说明工具完全本地化无后端,终身授权的维护成本几乎为零)②终身授权用户作为最活跃反馈来源的实际价值③lifetime定价相对年费的心理锚点设计逻辑

展开评论
  • @SpaceJeans (24): I don’t run a backend, the tool is fully local and user data is managed through Apple ID, I never need it or have to store it myself
  • @dev_pradu_48 (16): But is it problematic to manage infrastructure? When all or some users get lifetime tire then how u handel tose users data, work, etc?
  • @SpaceJeans (4): Probably better described as a product than a service I suppose
  • @PurePassion5309 (3): What happens if you don't renew your Apple Developer account?
  • @Nutcase_123 (3): What do you mean by "lifetime $24.99/yr"? Do people still have to pay you after purchasing the lifetime plan?
Reddit R16
14 分 · 16 评 🔥

SaaS创始人征集永久免费套餐的真实运营数据,评论区呈现从增长引擎到成本黑洞的不同亲历案例

I keep going back and forth on this for my own product. Everyone says "free tier = growth engine," but I've also seen founders quietly admit their free users cost more in support tickets and infra th…

受众观点:①免费套餐的设计原则(@rupert_at_work 提出'product-qualified lead machine'框架,@CaterpillarFun209 强调限制自然增长天花板而非核心功能)②高touch客服支持在免费用户上的成本是否可控③免费用户转化付费的核心触发点是否是'aha moment'体验

展开评论
  • @CornerThis1386 (4): The healthiest free plans I’ve seen have a really obvious ceiling. Enough value to get adoption, but not so much that power users can live there forever or support has to touch every account. Once free users need human help a lot, it usually stops being a growth loop and turns i…
  • @rupert_at_work (3): Free only works when it is basically a product-qualified lead machine, not a charity tier. My rule of thumb: give away enough for one clear win, then cap the expensive stuff hard: seats, usage, exports, integrations, support. If the free user can get ongoing value forever withou…
  • @CaterpillarFun209 (3): free tiers hurt when they're a free version of your product, and work when they're a taste that creates a reason to upgrade. the freeloader problem is almost always a design problem, not a proof that free is bad. a few patterns that actually hold up. cap on value, not on pain. l…
  • @kabirs1nghhh (2): I think the answer depends on whether your free users can experience the core value of the product quickly. If they never reach that “aha” moment, they’ll probably never convert. Curious to hear what the data looks like for others.
  • @tom_reb (2): I would go with the free plan ( with reasonable limits) and see how he goes. Free plan is a good for marketing. You can always disable it or convert to the trial version
Reddit R14
15 分 · 14 评 🔥

开发者将多年自用小工具改造为Web App公开发布,上线一天即收到陌生用户年付订阅验证了产品价值

I'm a little too excited right now. I've been building small SaaS tools for years, mostly just for myself, and never really thought anyone else would want to pay for them. One of my good buddies kept…

受众观点:①将自用工具公开发布是否是独立开发者最高效的产品路径②年付订阅相比月付在早期阶段对创始人信心的信号价值③发布前是否需要提前积累受众或做营销预热

展开评论
  • @Repulsive-Drummer348 (2): what's your saas about?
  • @agreafplayz (2): It’s an app for athletes to check how high they jump! Feel free to try it as well, I mean anyone can use it. It’s vertchecker.com
  • @Accomplished_Pick691 (2): congrats! keep going!!
  • @bennorthmore (2): Congrats! Did you do marketing before you launched?
  • @Serhii_Hrynevych (1): congrats!
Reddit R2
32 分 · 13 评 🔥

AAAI会议投稿编号突破3.2万,学界呼吁公开撤稿和被拒论文评审记录以提升评审问责

Recently submitted my abstract and the submission number is 32xxx. With still a day to go, I just wonder where are we heading. Hope these conferences at least start making the reviews and names publi…

受众观点:①LLM生成评审的可靠性问题(@OutsideSimple4854 指出评审用LLM但缺乏领域背景,会误导AC判断)②投稿量爆炸是否稀释了顶会的含金量和入场门槛③评审透明度改革的可行性争议,公开被拒论文是否等同于公开羞辱

展开评论
  • @OutsideSimple4854 (14): Why? As someone who writes theory papers, I know when a reviewer is using an LLM (or perhaps, argue theory based on lack of background knowledge, and write a review that sounds plausible, but sways the AC even though it’s wrong). Making that review public eventually sways more p…
  • @impatiens-capensis (12): There were IDs in the 30,000+ range last year. After desk rejects and people pulling before the deadline (they submitted an abstract but not a paper), it went down to around 20,000ish papers.
  • @Even-Inevitable-7243 (11): Are you insinuating that researchers who have papers rejected or withdrawn from conferences should be publicly shamed?
  • @OutsideSimple4854 (5): Honestly for me, it was really based on PhD advisor, who was not in CS. A lot of ML/theory papers have their foundations in statistics. I would probably say a lot of papers in stats journals, with a bit of different framing, could end up being published at ML conferences.
  • @OutsideSimple4854 (5): Not about low quality submissions, but low quality reviews as well. There’s a tendency to google, check that someone said the same review your LLM gives you, and then say: oh this paper is bad because of etc etc etc.
Reddit R57
18 分 · 13 评 🔥

独立开发者分享拒绝第一个收购要约的决策过程,拆解现金加月薪加股权报价结构中隐藏的价值陷阱

Built a small AI tool in my spare time alongside my day job. It's grown steadily and recently started getting real traction. This week someone offered to buy it - cash + a monthly retainer + some equ…

受众观点:①是否做出正确决定的认同(@Creepy-Surprisee 认为判断正确)②提前设定心理底价的重要性(@Then_Instruction_199 建议预设接受条件避免情绪化决策)③获得收购要约本身即是对产品价值的验证(@george_alto 从两个维度祝贺作者)

展开评论
  • @Creepy-Surprisee (5): Sounds like you made the right call
  • @Then_Instruction_199 (1): smart to walk away. one thing worth thinking about though, do you have a number in your head where youd actually say yes? knowing that ahead of time makes future conversations way less emotional
  • @george_alto (1): Congrats on two fronts. First building something that has value and second, getting enough traction to get an offer.
  • @baddaywithacamera (1): Good on you for saying no. So many predators out there looking to rip off good work by others.
  • @WonderfolioApp (1): My friend offered to invest in my product recently to supplement my income while I was working on the project. I don’t think I need to exchange equity for wages in my case. So I made a similar decision as you.
Reddit R32
12 分 · 11 评 🔥

开发者发布开源自托管 SEO 仪表盘 CrawlSEO,接入 Google Search Console 数据,支持站点爬取、Core Web Vitals 监控,并内置 10 个 MCP tool 供 Claude Code 和 Cursor 直接查询 SEO 数据。

I built an open-source SEO dashboard and wanted to share it here. What it does: * Connects to Google Search Console and syncs your keyword/page data * Crawls your site (up to 2,000 pages) for broken…

受众观点:①爬虫规模限制(@anon_zero 问 2000 页上限是否支持外部数据源集成)②MCP + Claude Code 工作流消除手动查数据痛点(@m1ke_digital 分享跨多站点自动查询的真实体验)③产品功能完整性和未来路线图(@Ok-Pace-8772 问还缺哪些 SEO 功能)

展开评论
  • @anon_zero (3): Due to the 2000 pages limitation on the crawler, is there anyway in the future you would look to integrate external crawl data such as ones from Screaming Frog ?
  • @asimovs-auditor (1): Expand the replies to this comment to learn how AI was used in this post/project.
  • @anon_zero (1): A charity i know their site is rather large so about 7000 core pages but around 400,000 pages in their wiki.
  • @Ok-Pace-8772 (1): It's there something in terms of SEO this tool did not do/have? I'll definitely be using it.
  • @m1ke_digital (1): honestly the whole reason i built it. i run a few sites and i was constantly bouncing between GSC tabs and claude, copying numbers back and forth just to answer simple questions now claude code just reads all of them through the MCP server. “which of my sites lost traffic this w…
Reddit R20
9 评 🔥

webdev 社区讨论真正被团队长期采用的开发工具,涵盖 CI/CD、测试、部署、监控等各类方向,Playwright 和 pnpm 是高票回复。

I was thinking about how many developer tools I've seen over the last few years that promised to make teams more productive. I do agree that some of them were really useful but most just faded away a…

受众观点:①开发者关心工具能否真正解决具体痛点而非昙花一现(@atlas__free 推荐 Playwright)②讨论具体工具类型(@LocoNachoTaco420 提 pnpm 节省磁盘空间)③工具对独立开发者 vs 团队的适用差异是潜在话题

展开评论
  • @taotau (5): Drinking on company time/cost outside friday afternoon
  • @atlas__free (3): Playwright
  • @Big-Information-5570 (2): ote this button" conversations
  • @LocoNachoTaco420 (1): pnpm. Saves us so much time and disk space
  • @jerapine (1): Nuxt
Reddit R15
9 分 · 8 评 🔥

独立开发者用Remotion开源React视频库零成本vibe coding制作产品发布视频,效果媲美付费制作团队

was scrolling Youtube and kept seeing these polished launch videos everywhere that look like someone paid an agency a few thousand dollars for. figured that was out of budget for a side project, so I…

受众观点:①用代码生成视频的实际工作流和学习曲线(@DeeplyQuarterly 验证了版本控制视频和UI组件直接复用对SaaS demo的价值)②Remotion与HeyGen等AI视频生成工具的定位差异③发布视频对早期SaaS冷启动获客的实际ROI

展开评论
  • @DeeplyQuarterly (2): and the motion matching your actual components is huge for SaaS demos. no recreating UI in after effects just to match your own brand. I did something similar for a client demo and it saved me a week of back and forth with a motion designer. plus version controlling your videos…
  • @Every-Metal-7050 (2): Iv done some unreal explainers with HeyGen, the avatar is me but with the teeth Iv always wanted and a less severe accent. The worlds screwed, my kid even say he loves his AI Daddy 🤣
  • @Environmental-Heron8 (2): that's messed up haha, at least earn some money with the explainers
  • @Environmental-Heron8 (2): yeah I'll take a look
  • @Every-Metal-7050 (1): It’s so easy too, just get clause to write the script and paste it into heygen, tell it how long you want the explainer as without this they end up pretty long. YouTube - speedipro if you wanna take a look, the thumbnails took longer than the explainers 🤣
Reddit R25
8 评 🔥

开发者讨论 AI coding agent 在工具鉴权失败时应该降级继续还是强制停止,帖主认为涉及 API 契约时静默猜测比明确报错更危险。

This came up for me while thinking about AI-assisted web dev work that depends on outside tools. If the agent cannot read the issue tracker, API docs, or a project tool because auth failed, should it…

受众观点:①开发者关心 agent 在信息不完整时是否应该自动停止(@MenuTimely1548 支持 hard stop)②AI agent 置信度和透明度的工程设计问题(@MaestroSplinter69 提到 grounding in repo evidence)③测试覆盖能否捕获 agent 静默猜测错误(@Narfi1 提测试思路)

展开评论
  • @Narfi1 (4): Shouldn’t your tests catch that ?
  • @MaestroSplinter69 (2): I’d rather have an agent refuse than guess. A graceful failure is a lot cheaper than confidently generating code against an imaginary API. That’s actually the problem I’m trying to solve with DevTime - grounding answers in verified repo evidence instead of assumptions
  • @Gremlation (2): > If the agent cannot read the issue tracker, API docs, or a project tool because auth failed, should it keep working from the prompt, or stop and say the task is under-specified? Pray, Mr. Babbage, if you put into the machine wrong figures, will the right answers come out?
  • @MenuTimely1548 (1): Hard stop makes sense once it touches API contracts or data - silent wrong guesses are way worse than an honest "no access." For small UI stuff, letting it degrade is fine since review catches it anyway.
  • @Prestigious-Way1525 (1): yeah this is exactly the failure mode i run into too. i’ve been building and using Samelogic for browser work, and the most reliable pattern is to fail fast when context is missing and force a real recovery step (who owns the missing input, where to reconnect source-of-truth). i…
Reddit R4
3 分 · 6 评 🔥

独立研究者用单张RTX 3090复现OpenAI性格特征持久性RL训练,GRPO方法实际效果仅提升2.4分

**TL;DR:** I’m reproducing the trait-persistence result from [arXiv:2606.24014](https://arxiv.org/abs/2606.24014) on one RTX 3090. Before I can test persistence I need to *install* a trait via RL — a…

受众观点:①GRPO训练在小规模setup下能否有效复现trait安装(@bbu3 和 @ResidentPositive4122 分享了各自的DPO+GRPO组合经验)②prompt数量不足和全局评分粒度对训练信号质量的影响③社区对帖子本身是否由LLM生成的强烈质疑

展开评论
  • @onedeskover (2): Why do people insist on running posts like this through an LLM. Is it really so painful to write them yourself?
  • @x11iyu (2): seeing they're tuning qwen 2.5 instead of anything else, unless OP has a very good reason, LLM might've done much more than "only" writing this post for them
  • @Every-Cat-2611 (1): What’s wrong with qwen 2.5? I’ve used it for a couple tasks.
  • @bbu3 (1): I've had trouble in the past trying to find tune via grpo (no traits) when thinking would exceed completion length. There was nothing I could successfully do in terms of punishing long completions. The only thing that helped was to limit my training to shorter examples and massi…
  • @ResidentPositive4122 (1): I had good success with a dual stage approach. First DPO on short vs. long correct answers (generated by the same model). Loosely following that S-something paper where they used 1k pairs. Then GRPO. Naively it will still tend to lengthen the completions, but you can play with t…
Reddit R37
2 分 · 5 评 🔥

独立开发者抱怨 tolt 和 rewardful 等联盟营销平台收费过高,求推荐定价更合理的 affiliate program 服务商

Folks i wanna add an affiliate program to my product, i looked around and saw products like tolt and rewardful but their pricing tiers made me run away, i dont get why they charge so high upfront for…

受众观点:①联盟程序收费结构是否合理(tolt/rewardful 高昂前期成本让人望而却步)②是否有免费或成功后付费的平替方案(PromoteKit 模式受关注)③自建追踪方案的可行性(Stripe coupon codes 手动方案被提及)

展开评论
  • @No-Fig-8614 (1): If you do get an affiliate code running please post it on my new affiliate sharing site: [https://perko.io/home](https://perko.io/home)
  • @Quiet-Sunset-7384 (1): paying upfront feels like buying a commercial mixer before you've even baked a loaf. i just used custom stripe coupon codes to track my first three partners.
  • @connorhoy (1): PromoteKit is worth a look. Free until you hit £10k/month in affiliate-driven revenue, then $39/month after that. No upfront commitment while you find out if affiliates are actually worth it for you.
  • @AvailableMycologist2 (1): Following
Reddit R22
1 评 🔥

独立开发者用 Playwright 抓取亚马逊商品价格和库存,被验证码和页面变化搞崩,寻找托管 Amazon 爬虫方案处理代理和重试。

I'm building a deal alert side project that tracks price, review count and stock for about 300 products. Playwright on a small VPS works until captchas, page changes or bad data start breaking runs w…

受众观点:①开发者关注有没有可靠的托管 Amazon 爬虫方案(帖主核心问题)②灰色地带的合规顾虑让讨论受限(@kin3v 提到"grey side of development")③具体工具推荐几乎为零,说明信息缺口大

展开评论
  • @kin3v (1): This is like the grey side of development. I hope you find new insights but people won’t share it quickly
PH PH1
🔥

Lev8 是一个 AI 销售情报 agent,通过 swarm 后台 agent 实时扫描全网挖掘目标潜在客户并跨渠道自动触达

Lev8

受众观点:①Lev8 maker 介绍 swarm agent 实时监控招聘/融资/技术栈动态变化的技术架构 ②RichgaLu 讨论数据新鲜度的配置策略及信号质量与速度/成本的平衡 ③Kutlwano Melamu 等用户表示祝贺并询问具体使用场景

展开评论
  • @Overview (0): * [Reviews](/products/lev8/reviews) * [Alternatives](/products/lev8/alternatives) * [Built with](/products/lev8/built-with) * [Team](/products/lev8/makers) * [Awards](/products/lev8/awards) * More Free Options Launch tags:[Sales](/topics/sales)•[Artificial Intelligence](/topics/…
  • @Lev8 (0): Maker 📌 👋 Hey Product Hunters, I'm [Tony Zhang,](https://www.linkedin.com/in/bo-zhang-8b753792/) co-founder of [Lev8](https://lev8.com/?ref=producthunt) , the fastest agent to find and reach your target people and companies across every corner of the internet and reach across mu…
  • @Lev8 (0): Maker 🎁 **For the Product Hunt community** Product Hunt members get **500 free credits today**. 👉 Try [Lev8.com ](https://lev8.com/?ref=producthunt)with a search you couldn’t solve before, and tell me how it goes. I’d especially love to hear what you searched for and where the r…
  • @Lev8 (0): Maker [@blink\_66](https://www.producthunt.com/@blink%5F66) Thanks for the shoutout! We keep data fresh by completely rethinking how data is sourced—instead of locking people into a static, stale database, Lev8 operates as a live system that mines the open web in real time. Our…
  • @RichgaLu (0): [Lev8](/products/lev8) Maker By default, we balance signal freshness with research speed and cost. If your campaign depends on very recent information, you can tell the Lev8 agent the freshness window you need. Before sending, you can review the suggested angle and its supportin…
PH PH2
🔥

CartAI 是专为 AI agent 设计的结账支付完成 API,填补浏览器自动化能导航却无法付款的关键断层

CartAI

受众观点:①CartAI maker 阐述 browser automation 和 payment API 各自存在但互相割裂的根本问题 ②Jernej Jan Kočica 从商家侧提出订单完成后出错的后续处理边缘情况 ③CartAI maker 回应说 CartAI 下的订单与普通电商订单无区别

展开评论
  • @Overview (0): * [Reviews](/products/cartai/reviews) * [Alternatives](/products/cartai/alternatives) * [Built with](/products/cartai/built-with) * [Team](/products/cartai/makers) * [Awards](/products/cartai/awards) * More Payment Required Launch tags:[SaaS](/topics/saas)•[Developer Tools](/top…
  • @CartAI (0): Maker 📌 Hey PH 👋 I'm Manil, founder of CartAI. We have been building CartAI for more than a year now and today we're shipping the developer release. The problem we kept running into: every AI agent demo ends right before the part that matters. Browser automation can navigate the…
  • @CartAI (0): Maker
  • @CartAI (0): Maker Thank you[@galdayan](https://www.producthunt.com/@galdayan) . We appreciate your support 🙏 Upvote Report Share 3h ago [](/@jernej%5Fjan%5Fkocica) [Jernej Jan Kočica](/@jernej%5Fjan%5Fkocica) Coming at this from the merchant side, we build support tooling there, so my head…
  • @CartAI (0): Maker [@jernej\_jan\_kocica](https://www.producthunt.com/@jernej%5Fjan%5Fkocica) Great question. CartAI places the order with no loss of fidelity between the consumer and the merchant. So the contact information(email, phone) etc that goes with the order is of the end consumer.…
PH PH4
🔥

Rerun 是主打透明度的无代码 AI agent 平台,每步操作可监控,敏感操作前必须人工审批,解决 agent 黑盒问题

Rerun

受众观点:①Rerun maker 强调当前大多数 agent 是黑盒,通过步骤可视化和 token 用量监控让 agent 行为可解释 ②Rerun maker 详细说明关键步骤暂停等待用户审批机制 ③Rerun maker 回应无法回滚的问题,解释 agent 在用户要求时可撤销更改

展开评论
  • @Overview (0): * [Reviews](/products/rerun-2/reviews) * [Alternatives](/products/rerun-2/alternatives) * [Built with](/products/rerun-2/built-with) * [Team](/products/rerun-2/makers) * [Awards](/products/rerun-2/awards) * More Interactive Payment Required Launch tags:[Tech](/topics/tech)•[Mark…
  • @Rerun (0): Maker 📌 Hey Product Hunt 👋 I'm Clément, founder of Rerun. We built Rerun because every "AI agent" tool felt like a black box. You fire it off and just hope it did the right thing. We wanted agents you could actually watch work, step by step, that pause before doing anything sens…
  • @Rerun (0): Maker [@artem\_fedorovich](https://www.producthunt.com/@artem%5Ffedorovich) Totally agree with you Artem Rerun gives tools to really understand the agent and stay in control The agent has some handy features, like notifying the user and asking for approval before moving to the n…
  • @Rerun (0): Maker [@aidan\_codefox](https://www.producthunt.com/@aidan%5Fcodefox) Totally agree. Since the flows are unique, there's no way to "rollback." However, the agent is smart enough to cancel changes if the user asks. For example, with Rerun, any agent handling important processes h…
  • @Rerun (0): Maker
PH PH5
🔥

Phantomstory 帮助品牌创建赞助第三方博客站点,专门优化在 ChatGPT 和 Claude 等 AI 搜索引擎中的内容排名

Phantomstory

受众观点:①Phantomstory maker 解释 ChatGPT/Claude 知道品牌永远给自己排第一所以更信任第三方来源 ②clemente_lopez1 询问赞助关系是否需要披露,涉及内容诚信问题 ③Advin Jadis 担忧发布内容过快会损害域名信任度

展开评论
  • @Overview (0): * [Reviews](/products/phantomstory/reviews) * [Alternatives](/products/phantomstory/alternatives) * [Team](/products/phantomstory/makers) * More Payment Required Launch tags:[Marketing](/topics/marketing)•[Growth Hacking](/topics/growth-hacking)•[Developer Tools](/topics/develop…
  • @Phantomstory (0): Maker 📌 Hey Product Hunt! 👋 I'm Mathew, founder of The Letter Company. Today we're launching Phantomstory: our coolest product yet. I've been a content marketer for the last four years, and I kept seeing the same pattern: one of the most effective strategies for winning AEO is t…
  • @Phantomstory (0): Maker [@artem\_fedorovich](https://www.producthunt.com/@artem%5Ffedorovich) We don't do placements. At all. I don't think placements are a good approach as a product category because it's too arbitrary on a space-by-space basis, and it hardly works. Instead, we're in the busines…
  • @Phantomstory (0): Maker [@clemente\_lopez1](https://www.producthunt.com/@clemente%5Flopez1) Fantastic question. That is entirely configurable. For some of our customers, there is no disclosed relationship. It's effectively a sponsored website in the same way that Hearst might own a magazine that…
  • @Phantomstory (0): Maker [@os\_ishmael](https://www.producthunt.com/@os%5Fishmael) These are entirely configurable! Depends on the company's approach :) some would rather faint than ever be negatively portrayed, others embrace being balanced. Our platform itself is agnostic. Upvote Report Share 4h…