Claude Code 推出 Projects 功能,支持从单个对话派生并行工作线程,云端持续运行,关闭笔记本后任务不中断
Projects now run from one conversation, starting in Claude Code. You describe what needs doing, and Claude directs parallel threads that keep working after you close your laptop. In beta today for se… ① 新功能与 opus-next 组合的性能期待(@synthwavedd 已经在兴奋预期)② Overview 面板对多线程任务的可见性和手机端操控体验 ③「像简报 chief of staff」的交互范式对工作流的实质改变(@claudeai 官方二楼解释了使用模式)
Claude Code 推出 Projects 功能,支持从单个对话派生并行工作线程,云端持续运行,关闭笔记本后任务不中断
① 新功能与 opus-next 组合的性能期待(@synthwavedd 已经在兴奋预期)② Overview 面板对多线程任务的可见性和手机端操控体验 ③「像简报 chief of staff」的交互范式对工作流的实质改变(@claudeai 官方二楼解释了使用模式)
ClaudeDevs 官方介绍 Claude Code Projects 技术机制:自动拆线程、并行云端会话、跨线程传递上下文,当前向 Pro/Max 限量测试中
① 功能与 Cursor Projects 的异同(@mubshrx 代表大量开发者的困惑)② 并行 session 导致 credit 快速消耗的担忧(@st3v3li 已经在预警)③「多线程共享项目」的概念理解门槛(@PosaniRishi 的问题很有代表性)
z.ai 发布基于超十万块国产AI加速芯片集群的GLM-5.3-Flash全量推理系统,同时引发关于蒸馏攻击Anthropic及推理服务定价暴涨的争议
①@dada216 关注中国自主AI推理基础设施的实际规模与可行性 ②@embedding-shape 关注价格大幅上涨($20→$80/月)背后的scaling成本压力 ③@bbor 质疑GLM训练存在非法路由Anthropic Opus 4.8请求进行蒸馏攻击的行为
Google发布免费一小时Agentic Engineering实战课,系统覆盖AI agent内存三层架构、agentic loops、MCP实现与多agent编排,被开发者社区广泛传播并认为可替代数百美元付费课程
Google发布免费一小时Agentic Engineering实战课,系统覆盖AI agent内存三层架构、agentic loops、MCP实现与多agent编排,被开发者社区广泛传播并认为可替代数百美元付费课程
受众观点:开发者对agent memory和MCP vs API两个模块关注度最高,对「替代$500课程」说法反应两极。@itsthedonhashim(score:2, Twitter):「the agent memory part is exactly what I needed. been trying to figure that out for ages」;@saen_dev(score:1, Twitter):「The MCP vs API section at 40:04 is the part most developers need because everyone is building MCP servers without understanding when a plain REST endpoint would have been simpler and more debuggable. Did they cover when NOT to use MCP?」;@kachmass(score:0, Twitter):「solid overview, but replacing a $500 course? that's missing the debugging headaches and edge cases that only time-in-trenches teaches」
Anthropic 正式推出 Claude Code Projects 功能,单次对话即可自动编排并行云端线程持续执行,无需手动搭建 agent harness,引发独立开发者社区对平台依赖 vs 自托管主权的激烈争论
Anthropic 正式推出 Claude Code Projects 功能,单次对话即可自动编排并行云端线程持续执行,无需手动搭建 agent harness,引发独立开发者社区对平台依赖 vs 自托管主权的激烈争论
受众观点:跨平台评论区的核心矛盾集中在两个层面:第一,平台依赖 vs 基础设施主权——@Kenaadams99(score 222,Twitter)直言「$20 vs $700 is the wrong frame. You're comparing rent vs ownership. When Anthropic changes API terms, the $20 crowd has nothing.」;第二,usage limit 现实墙——@contractorkeith(score 11,Twitter)一针见血:「Opens Claude to run an agent on a $20 plan. 3 minutes later: You've used up your usage limits, come back next week.」命名混乱也是用户痛点,@EffortlessSteve(score 39,Twitter)说「This is going to confuse so many people」,@andrclp(score 10,Twitter)指出「you already had 'projects' now this seems different feature with the same name」
OpenAI 多智能体团队核心研究员公开接受深度专访,详述 long-horizon agent 与多 agent 协作研究现状,同时招募对齐与人机交互方向研究员,表明 OpenAI 正在将 multi-agent 建制化推进
OpenAI 多智能体团队核心研究员公开接受深度专访,详述 long-horizon agent 与多 agent 协作研究现状,同时招募对齐与人机交互方向研究员,表明 OpenAI 正在将 multi-agent 建制化推进
受众观点:评论区关注点集中在两个方向:一是研究价值本身——@kimmonismus(score 2,Twitter)表示要专门留到 11 小时长途飞行时观看,反映研究圈的高度重视;二是 safety 焦虑——@degenurai(score 1,Twitter)直言「it really is obvious that alignment isn't on the cards any time soon but RSI into ASI is, that's frankly terrifying. At the same time labs are pushed into a...」,对 alignment 进展落后于能力扩展表达深层担忧
新型AI分类/嵌入模型Jev凭借低成本即插即用特性迅速流行,在ML社区引发与传统BERT微调方案的工程选型争论,其名称本身也暗示了"便宜到随便用"的产品定位
新型AI分类/嵌入模型Jev凭借低成本即插即用特性迅速流行,在ML社区引发与传统BERT微调方案的工程选型争论,其名称本身也暗示了"便宜到随便用"的产品定位
受众观点:评论区明显分为技术派和务实派两阵营。技术派:@NathanWilbanks_ (score:48, Twitter) 给出具体替代方案「deBerta作为encoder + GLiNER架构」表示BERT系方案依然有竞争力;务实派:@WinterCawfie (score:6, Twitter) 直接反驳「小型BERT分类器泛化能力差,除非你花大量时间构建多样化大数据集,否则Jev在plug-and-play方面完胜」。多数评论倾向于认可Jev的实用价值
OpenAI Astra重度用户集体额度耗尽,寄望于内部员工Tibo手动重置配额,评论区大量用户表达同样的额度焦虑,折射出AI服务使用限制对重度用户造成的严重阻碍
OpenAI Astra重度用户集体额度耗尽,寄望于内部员工Tibo手动重置配额,评论区大量用户表达同样的额度焦虑,折射出AI服务使用限制对重度用户造成的严重阻碍
受众观点:评论区情绪高度一致——@AdamHoltererer (score:169, Twitter) 吐槽标点符号骗了他以为是重大公告,说明用户情绪敏感、极度期待额度重置;@an0n_an0n_an0n (score:1, Twitter) 直接点出「他们希望你是冤大头去买积分」,揭示用户对额度策略背后商业动机的不信任。@ba_jakobsen (score:2, Twitter) 更直接指出「如果没有计算资源提供Astra,就不该向公众开放」
AI 辅助编程(Claude Code/Codex)在多会话开发中存在决策上下文丢失的核心痛点,开发者社区正在探索可视化历史记录和结构化决策追踪等系统性解决方案
AI 辅助编程(Claude Code/Codex)在多会话开发中存在决策上下文丢失的核心痛点,开发者社区正在探索可视化历史记录和结构化决策追踪等系统性解决方案。
受众观点:开发者们高度共鸣于这个痛点,并在评论区提出多种解决思路。@ceyhunkarslan(score:1, reddit/SaaS)建议「Make the decision trail part of the change, not a separate journal. For each meaningful commit or PR, capture three lines: what changed, what you rejected, and what would make you revisit it.」强调将决策记录内嵌到 commit 流程而非独立维护;@Rare_Guide_9830(score:2, reddit/SaaS)分享了类似工具探索「I've been trying to tap into the same thing but more for exploring usage and memory」,印证了相同需求在独立开发者中的普遍性。
Claude Code 推出 Projects 功能,支持从单个对话派生并行工作线程,云端持续运行,关闭笔记本后任务不中断
Projects now run from one conversation, starting in Claude Code. You describe what needs doing, and Claude directs parallel threads that keep working after you close your laptop. In beta today for se…
受众观点:① 新功能与 opus-next 组合的性能期待(@synthwavedd 已经在兴奋预期)② Overview 面板对多线程任务的可见性和手机端操控体验 ③「像简报 chief of staff」的交互范式对工作流的实质改变(@claudeai 官方二楼解释了使用模式)
展开评论
- @claudeai (351): You can talk to Claude the way you'd brief a chief of staff: several things at once, in any order. It routes requests to new or existing threads and checks in when it needs you. The Overview panel shows what's waiting on you, and you can steer any thread, even from your phone. h…
- @synthwavedd (113): @claudeai oh this is going to be glorious with opus-next
- @realAbsham (59): @claudeai Reset plz 🥲 https://t.co/JSK0ibeTXR
- @kimmonismus (9): @claudeai ngl, good release
- @robj3d3 (5): @claudeai Hey @lydiahallie would you like to grace us with a reset to go along with this launch? 🥹
ClaudeDevs 官方介绍 Claude Code Projects 技术机制:自动拆线程、并行云端会话、跨线程传递上下文,当前向 Pro/Max 限量测试中
Today we're rolling out Projects in Claude Code on desktop and web. A project is one conversation with Claude. It splits the work into threads itself, runs them as parallel cloud sessions, passes con…
受众观点:① 功能与 Cursor Projects 的异同(@mubshrx 代表大量开发者的困惑)② 并行 session 导致 credit 快速消耗的担忧(@st3v3li 已经在预警)③「多线程共享项目」的概念理解门槛(@PosaniRishi 的问题很有代表性)
展开评论
- @ClaudeDevs (205): Available today for Pro and Max users who already run cloud sessions and haven't used Claude Projects yet. If you've used Projects before, we're rolling out to you slightly later so your existing ones move over cleanly. Want access now anyway? Join the waitlist: https://t.co/mLh…
- @EffortlessSteve (39): @ClaudeDevs 1. Awesome 2. This is going to confuse so many people
- @mubshrx (19): @ClaudeDevs okay same thing as Cursor Projects?
- @st3v3li (17): @ClaudeDevs A fater way to burn all the credits in 5 hours https://t.co/F0vWWroNDB
- @PosaniRishi (12): @ClaudeDevs Bit confused about this , is it basically shared threads under one project ?
AI 研究圈以反讽口吻热议 jev 模型走红现象,质疑为何大家选用新兴工具而非自建 BERT 微调或等待 astra 成熟,折射出 AI 工具 build vs buy 的选型困境
noooo why is everyone using jev for stuff instead of self-hosting a custom BERT finetune or waiting 18 months for astra to get faster and cheaper
受众观点:①@NathanWilbanks_(score 48)提出 deBERTa + GLiNER 架构作为自建替代方案的具体实现,代表深度工程师视角 ②@anthonyronning(score 44)直指 jev 流行的核心原因是"更酷"而非纯技术优势,揭示社区选型中的情绪因素 ③@menhguin(score 22)用戏谑方式回应,代表社区对 contrarian 建议的调侃文化
展开评论
- @NathanWilbanks_ (48): @willcb deBerta as the encoder, with a GLiNER like label scoring architecture and training on top plays doom and does a bunch of business tasks https://t.co/hH91b6Tqf7
- @anthonyronning (44): @willcb jev is a lot cooler than bert that's why
- @menhguin (22): @willcb i would simply hire someone from bengaluru named jev (he will cause many problems)
- @max_paperclips (19): @willcb literally yes tbh
- @identityTorn (11): @willcb the mid-curvoors need to do their mid-curving, the graph dont paint itself https://t.co/UETJVwllCW
SST 创始人 thdxr 因 Astra AI API 月消费超 $154K 被迫为平台实现用量限制,评论区热议天价 API 账单与 AI 成本管理
astra finally pushed me to implement limits https://t.co/nPlIFw25QP
受众观点:① 对 $154K 月消费数字的震惊(@spectragai 的反应代表大多数人)② Astra 是否真的值这个 API 定价(@Scrubghetti 认为是 mediocre coding model)③ 订阅制 vs API 调用的成本结构选择(@neamtuz 建议用 sub 代替 API)
展开评论
- @RhysSullivan (35): @thdxr $2,000 a day seems reasonable yeah
- @snack_assassin (30): @thdxr What are they going to be able to get done with 20 minutes of astra?
- @Scrubghetti (30): @thdxr It's honestly a very mediocre coding model for the price. No way it is worth paying API prices for
- @spectragai (29): @thdxr $154k spend month to date?!! wtf you building?
- @neamtuz (15): @thdxr Yall dont get subs? even 20 5x subs i guess is million times better than API
Claude 工程师 bcherny 分享个人日常用 Projects 写代码的工作流,评论区质疑 Anthropic 内部员工推广行为的可信度
Projects are how I write a lot of my code these days. Really excited for everyone to try the new experience! Rolling out now
受众观点:① 质疑 Anthropic 员工用内部配额推广可信度(@NotUjjwal 明确点出了这个问题)② 想理解「Projects 到底在用 agent loop 还是别的什么」(@ThomasBekkers 的困惑很普遍)③ 希望 Claude 合并 chat、code、协作工具到单一平台(@guardian_rey 的呼声)
展开评论
- @NotUjjwal (5): @bcherny I guess it’s easy to write a lot of code with an infinite plan. You wouldn’t last a day on a 20x Max plan. Not sure who these posts are directed at or which audience they’re targeting.
- @guardian_rey (4): @bcherny We need you to combine claude code + chat + cowork into one platform. Also, full Computer Use feature! Thank you!
- @usefz89 (4): @bcherny Pls can you explain more about projects. Do you mean for one repo I can make a project that follows coding, SEO and others ? or I can make one project for all repos and Claude will manage ??
- @aleshu_ber (3): @bcherny I just mastered running loops after you said like 2 days ago that this is how you guys do things in Claude! https://t.co/VVGRJovXo2
- @ThomasBekkers (3): @bcherny thought you're just looping agents and not writing code yourself? heh
Anthropic 开放「生命科学验证计划」申请,持证专业人士可使用 Mythos 模型进行完整生物领域研究,引入身份核验机制替代模型能力降级
Today we’re opening applications for the Life Sciences Verification Program. Through the LSVP, life science professionals can use our models—including, for the first time, Mythos—with a new set of sa…
受众观点:① 「如何真正核实身份」的技术和流程问题(@kookkaiiz_zz 直接点出了这个盲点)② 「保持模型能力 + 验证访问」优于「降级模型」的设计思路(@MTorygreen 表示支持)③ Anthropic 账号封禁和访问限制带来的实际用户困扰(@MI_GAMBO 的遭遇是对比)
展开评论
- @notquantized (5): @AnthropicAI yes! i will finally be able to create cogniweapons full-time as an anon now. thank you, dario et al.
- @kookkaiiz_zz (4): @AnthropicAI no idea how they actually verify someone is a legit biologist and not just claiming to be, that part i'd want to understand
- @realFartsupreme (2): @AnthropicAI yeah cut out the free thinking 10 year olds sitting in some country who, with the right tools, untrained mind, and motivation might actually find a break through if just given a chance. give it to ppl who should have had it anyways from the start and disguise it as…
- @MTorygreen (2): @AnthropicAI Better way to handle this imo. Keep the model capable, verify who gets access to the risky stuff. No need to nerf biology tools for everyone because a small part can be misused.
- @MI_GAMBO (2): @AnthropicAI Hi @Anthropic My Anthropic API Console account has been placed on hold/restricted, I would like you to pls review my account and clarify Why has my API Console account been placed on hold? I contacted the support Team on the web they're Taking longer than expected o…
AI 研究者深度分析 jev 模型发布策略:有吸引力的宣传标题加短暂独家访问制造期待加趁讨论热度全量发布,总结为独立 AI 产品 launch 的范本案例
one of the coolest things about jev is that they just shipped it and it's good. nice clickbaity announcement, earned by a truly useful model-product. brief exclusivity, then into everyone's hands whi…
受众观点:①@beffjezos(score 23)关注 jev 的技术架构——想理解这个模型到底属于哪个类型,反映社区对新工具技术基础的求知欲 ②@reachmeviz(score 11)指出 jev 产品说明不够清晰,用户需要多次阅读才理解,揭示即使好 launch 也有产品沟通短板 ③@jkelleyrtp(score 2)细化了 launch 要素:demos、帖子、quirky notes、定价、branding,提供可拆解的 checklist
展开评论
- @beffjezos (23): @willcb Any idea what kinda model it is? Still getting my head around it
- @reachmeviz (11): @willcb They did a poor job in explaining it though. It took several passes for people to comprehend what they were trying to sell.
- @capemox (5): @willcb if a brand new lab announces something that isn't being rolled out immediately I just assume it's a scam
- @alexUnder_sky (3): @willcb why did they name the model like this? I always keep seing a different word.
- @jkelleyrtp (2): @willcb yes great announcement - demos, posts, quirky notes, cheap(!), realtime discord, nice branding done, great launch materials yes awesome launch, inspired me to improve my launches in the future
OpenAI 多智能体研究员与知名播客主 Dwarkesh Patel 深度对谈长期 agent 和 multi-agent 系统设计原则,团队同步开放 alignment、safety 和 human-AI 互动方向招聘
Happy to finally do a deep dive on multi-agent with @dwarkesh_sp! None of it would have happened without the great work on multi-agent from my @OpenAI teammates @kevinleestone, @mikegmalek, @__eknigh…
受众观点:①@polynoamial(score 282)宣布 OpenAI 招聘 alignment/safety 和 human-AI 交互方向研究员,代表机构对多智能体安全的战略重视 ②@degenurai(score 1)对 RSI 到 ASI 加速路径的担忧以及 alignment 工作紧迫性的深度评论,折射社区对 AI 加速主义的焦虑 ③@foombler(score 1)作为普通用户分享 2026 年 agent 工具使用体验的真实感受,提供用户侧视角
展开评论
- @polynoamial (282): Also, my team is hiring! We research long-horizon agents and multi-agent. We’re hiring for alignment/safety because we want to develop new research with alignment/safety in mind during the whole process. We’re also hiring for human-AI interaction. https://t.co/kfFyEowLKy
- @kimmonismus (2): @polynoamial @dwarkesh_sp @OpenAI @kevinleestone @mikegmalek @__eknight__ @amuellerml @zhangir_azerbay @CheukHeiChu I’ve got an 11-hour flight from SF back to Germany ahead of me. Saving this interview for the flight, it’ll make the journey so much better. Thank you, Noam!
- @minimesoy (1): @polynoamial @dwarkesh_sp @OpenAI @kevinleestone @mikegmalek @__eknight__ @amuellerml @zhangir_azerbay @CheukHeiChu okay two coffee mugs on the table and neither one touched yet, that's a long recording session
- @degenurai (1): @polynoamial @dwarkesh_sp @OpenAI @kevinleestone @mikegmalek @__eknight__ @amuellerml @zhangir_azerbay @CheukHeiChu Damn dude just watched it, it really is obvious that alignment isn’t on the cards any time soon but RSI into ASI is, that’s frankly terrifying. At the same time la…
- @foombler (1): @polynoamial @dwarkesh_sp @OpenAI @kevinleestone @mikegmalek @__eknight__ @amuellerml @zhangir_azerbay @CheukHeiChu Beautiful thank you. I am using agents more and more often. 2026 is best year ever.
GergelyOrosz 采访 TypeScript 大 V mattpocockuk,其 AI 编程 skills repo 6 个月内获星数超越 React 官方仓库,评论区讨论 skills 作为 AI 上下文载体的新范式
Pretty incredible how the skills repo by @mattpocockuk is only 6 months old, and has more stars than the React one. We had to sit down and talk. Timestamps: 00:00 Intro 05:48 How Matt got into tech 1…
受众观点:① GitHub stars 在 skills 类仓库上是否还是有效的质量信号(@rozhkov_ 认为已经失效)② skills 文件作为「组织记忆单元」在团队级别的潜力(@ChrisChomenko 的观点有深度)③ Matt Pocock 从 TypeScript 教育者转型为 AI 工具创作者的路径(@kapv89 关注其学术外路线)
展开评论
- @GergelyOrosz (5): Watch or listen also on: • YouTube :https://t.co/ybjHzJnIyv • Spotify: https://t.co/HZL39ctLMR • Apple: https://t.co/Qa5vw0o5wt
- @rozhkov_ (4): @GergelyOrosz @mattpocockuk I don't feel like stars is a fair metric anymore. Pretty much any skill repo amounts a lot of stars. Not to diminish Matt's work though.
- @kapv89 (2): @GergelyOrosz @mattpocockuk It's because of people like him, a path outside of academia exists for pursuing Software Engineering.
- @ChrisChomenko (1): @GergelyOrosz @mattpocockuk Stars are the surface signal. I think skills are becoming the unit of organizational memory: review rules, test expectations, domain language, and recovery paths that move with the repo and survive model swaps.
- @serxzsz (1): @GergelyOrosz @mattpocockuk 6 months to out star react. that repo timeline alone is a better story than most of the podcast
SST 创始人 thdxr 关于 AI token 消耗与员工人均营收关系的经济学观点:大量用 token 就必须有相应的高人均营收
if you love using lots of tokens you better make a lot of revenue per employee
受众观点:① 担忧高 token 消耗会积累债务的创业者(@MINDFUEL_NIRAJ 直接说要背债)② 对 AI 公司高消耗来源的好奇(@agenticbowie 问 ryan 的 token 哪来的)③ 调侃「融资可以无限烧钱」心态的反讽视角(@matnewera 的幽默)
展开评论
- @MINDFUEL_NIRAJ (8): @thdxr otherwise be ready for that debt tbh
- @agenticbowie (6): @thdxr how tf can ryan have those tokens then???
- @matnewera (4): @thdxr Not true just get more investors
- @notquantized (3): @thdxr if??
- @lewingtonpitsos (2): @thdxr i mean if every employee is using millions of tokens a day, you better be making some good money
mattpocockuk 尝试用 AI 做课程规划几周无效,改用便利贴加剪刀解决问题,引出开发者「AI tunneling」现象的真实讨论
Have been trying for weeks to make AI better at course planning Tried a new method today. Notecards, pen, paper, scissors. Turns out not using AI is pretty good guys
受众观点:① 认同「不该用 AI 的地方就别用」的实用主义立场(@ZryMiller 直接点明)② 经历过「AI tunneling」专注障碍,靠物理方式跳出的人(@kilcazn 分享了公园散步+纸质笔记本的解法)③ 对「几周 prompting 被一把剪刀打败」感同身受的幽默认同(@Avery_Coree 的表达击中了很多人)
展开评论
- @ZryMiller (6): @mattpocockuk Sometimes I try to have AI do something that I could easily do and then it fails and I get frustrated and do it myself. Just because we can use AI doesn’t mean we always should lol
- @kilcazn (3): @mattpocockuk Completely agree. Spending too long in front of my agents puts me in “AI tunneling” mode. I’ve started taking more walks in parks with a notepad. That has helped me find more creative solutions and unblock my thought process.
- @ElouardiMusta (2): @mattpocockuk Still holding, Matt https://t.co/Ivt1piswoM
- @Avery_Coree (2): @mattpocockuk weeks of prompting beaten by a pair of scissors, brutal but fair
- @luisrmartinez (1): @mattpocockuk I feel you, this is my thursday mood. https://t.co/AGH9iTJYaV
marclou 吐槽 Google Maps 餐厅评分被刷好评机制污染,商家用「赠咖啡换五星好评」操纵评分系统,表示愿意为可信评价替代产品付费
All restaurants with 4.5+ stars on Google Maps ask for a review. “We offer you free coffee 🤗 Now give us 5-star review.” I do not trust reviews anymore, but there’s no alternative for now. I’d ditch…
受众观点:①平台评分激励结构错位是根本问题,SEC 强制披露模型更可靠(@heikoscreens)②不仅餐厅,Airbnb 等平台同样存在退款换好评现象(@alexneyret)③「付钱买可信 KOL 城市列表」的替代方案已有人愿意付费(@marclou 自己)
展开评论
- @marclou (25): Actually, I'd buy a "restaurant list by [person] in [city]" for $9 https://t.co/WJjYw0GaEr
- @martindonadieu (8): @marclou by the way wtf is a a seafood steack kebab ?? the name is a redflag on it's own ahah
- @alexneyret (6): @marclou Same with everything - all reviews are manipulated. Last week we were at a 4.6 Air’bnb super unsanitary, we opened a dispute and were refunded half by the host and gifted extra coupons by Air’bnb to give a nice review. The same house was instantly available again to rent
- @hadesdevs (5): Problem is, there's no better solution, what do you want to do, ask for a bill, and require a bill, people would still so this. Same thing happened to me in Poland at a mexican restaurant. I never, and I mean it, NEVER waste food. But it was so vile I had to stop eating it. Late…
- @heikoscreens (4): @marclou The incentive sits on the wrong side - reviews are solicited by the person being reviewed. Insider filings work because the SEC makes people disclose whether they want to or not. Nobody gets free coffee for filing a Form 4. Data you're forced to report is harder to game.
willcb 吐槽 Jev 模型发布时最大的营销失误,是把自己和 terra 这款几乎没人用的模型做比较,反而弄巧成拙,评论区顺势讨论 terra 和其他模型的实际适用场景差异。
biggest fumble of the jev rollout was saying "it's as smart as terra" as if anyone uses that model
受众观点:① @CompleteSkeptic 指出 terra 的能力优势来自更好的 base model 而非 test-time compute,与 luna 在 coding agent 上的差距是结构性的 ② 讨论 GPT-4.5 等"人情味"模型在 coding agent eval 表现差但用户体验好的悖论 ③ @laplacian_demon 提出 terra 作为文档/review/小任务 subagent 仍有价值的反驳视角
展开评论
- @CompleteSkeptic (27): @willcb yes it was pareto dominated by luna on code, but to defend terra, some capabilities just come with better base models and not test-time compute e.g. GPT-4.5 would get destroyed at coding agent evals and I would kill for it rn 🥲
- @fisssherai (7): @willcb what did terra do that made the comparison land so badly?
- @silver__tsuki (6): @willcb https://t.co/gkNyNDPrHI
- @laplacian_demon (4): @willcb I do - terra is great as a subagent for docs, review, small tasks, etc.
- @MaziyarPanahi (3): @willcb yeah! luna or sol!
polynoamial 所在团队招募长期 agents 与多智能体研究人员,涵盖对齐安全和人机交互方向,且明确强调在研究全程将 alignment/safety 内嵌进去而非事后补救。
Also, my team is hiring! We research long-horizon agents and multi-agent. We’re hiring for alignment/safety because we want to develop new research with alignment/safety in mind during the whole proc…
受众观点:① @schulzb589 建议引入经济学背景研究者来理解 multi-agent 场景,并推荐了一个测试框架 ② @CapTopPicks 直接询问人机交互岗位的具体链接,说明这个方向关注度高 ③ @JAYCEEWEB1 提出根本性质疑:对齐为什么这么难,是数学层面的问题还是认知层面的盲点
展开评论
- @schulzb589 (0): Here's to hoping you find the best researchers. Please consider people with an economics background as something that will be quite useful in these scenarios. As a side-note this might be worth improving upon. It's a multi-agent testing framework. https://t.co/c7h65L3GqY https:/…
- @CapTopPicks (0): @polynoamial Where is the human-AI interaction role posted?
- @benohanlon (0): @polynoamial Hiring 🇬🇧?
- @JAYCEEWEB1 (0): I’m 100% behind you and( OAI) I know your team from @tszzl to @_aidan_clark_ and many others are hot on alignment. But, Why are you struggling? Is it a beyond Mathematics thing? Is it a glaringly obvious in hindsight thing? And where are all these alignment specialists you seek?…
- @fanofaliens (0): @polynoamial can i apply?
willcb 点评 AI 模型 Jev 的命名策略,称之为大胆的产品定位决策——名字本身直接传递"极便宜、将被大量使用"的产品信号,评论区认为这次赌注确实奏效了。
also pretty crazy shot-call to name a model after the concept of “it’s dirt cheap and you’re gonna use it a lot”
受众观点:① @16kbps 在意识到名字含义后的 aha moment——好名字有自解释能力 ② @notdalsito 认为这次赌注最终成功了,但隐含的意思是也可能失败 ③ @JesterMule 想象如果 Google 发布同名模型会引发什么混乱,说明命名的品牌风险
展开评论
- @16kbps (4): @willcb ohhhhh jev ons i get it
- @notdalsito (1): @willcb Balsy but it worked
- @JesterMule (0): @willcb everyone would be flipping out if Google released Jev. I mean, I'm flipping out about it right now, but imagine.
scaling01 在推特上公开抱怨某 AI 工具使用限制过于严苛,连官方沟通渠道都堵死,评论区大量用户共鸣,纷纷诉说被 reset 次数耗尽和被逼购买 credits 的遭遇。
chat, am I muted? where is tibo when you need him https://t.co/X1jxwvt7mK
受众观点:① @thegenioo 指出 Astra 相比其他模型的限制尤其严苛,具体平台差异引发共鸣 ② @an0n_an0n_an0n 认为故意限制免费用量是逼迫购买 credits 的商业策略,点出平台动机 ③ @henloitsjoyce 表示仅剩 5 次 reset,直观呈现了用户在极限边缘的绝望感
展开评论
- @thegenioo (2): @scaling01 idt we want reset but they are clearly ignoring the complaints of people saying how bad limits are for Astra especially
- @Xesar514 (1): @scaling01 I also need Tibo
- @an0n_an0n_an0n (1): @scaling01 They're hoping you're a sucker and buy credits that get burned through faster.
- @lurkylearning (1): @scaling01 hit the mines and buy credits citizen
- @henloitsjoyce (1): @scaling01 i have 5 resets left
z.ai 发布基于超十万块国产AI加速芯片集群的GLM-5.3-Flash全量推理系统,同时引发关于蒸馏攻击Anthropic及推理服务定价暴涨的争议
How GLM built its own inference infrastructure
受众观点:①@dada216 关注中国自主AI推理基础设施的实际规模与可行性 ②@embedding-shape 关注价格大幅上涨($20→$80/月)背后的scaling成本压力 ③@bbor 质疑GLM训练存在非法路由Anthropic Opus 4.8请求进行蒸馏攻击的行为
展开评论
- @dada216 (0): We built a complete production-grade inference service from scratch on a cluster of more than 100,000 Chinese-made AI accelerators. All production inference for GLM-5.3-Flash runs on this system.
- @embedding-shape (0): I was gonna ask how people found their coding plans, and realized, have they massively ramped up the prices? Seems the middle plan is ~$80/month now, didn't that used to be like $20/month? Cheapest plan is ~$20/month currently. They must have hit really hard scaling limits if th…
- @bbor (0): Well, other than the infrastructure they got from illegally routing millions of paying customers' requests through Anthropic's Opus 4.8 in a distillation attack...
- @rob74 (0): This article left me with one immediate question: "WTF is GLM?". Honestly, I have no idea what z.ai is either (I'm aware of an AI-enabled editor called Zed, but that's under zed.dev), so it's a bit presumptuous from them to assume that everyone is familiar with their product...
- @Argonautlabs (0): Different angle on the same model: the full GLM-5.3 (744B MoE, 4-bit experts, 434 GB on disk) runs on a single MacBook Pro M5 Max with 128 GB by streaming the experts from NVMe SSDs instead of keeping them in memory. One drive gives about 2 tok/s; striped across four drives it r…
陶哲轩发文讨论AI快速求解数学未解难题是否会损害数学共同体的知识积累方式,引发关于AI与人类认知深度的广泛争论
Why I didn’t sign the Fields medallists’ letter
受众观点:①@piker 担忧LLM即时结果会短路知识探索的传统路径,让知识积累绕过了局部最优 ②@fruitl00p 认为AI公司把未解难题当可开采资源,不在乎是否损害数学领域长期发展 ③@dist-epoch 引用Tao原文警告AI可能颠覆社会,届时数学传统的保存将不再重要
展开评论
- @piker (0): > With that interpretation, the issue becomes slightly different: is it more important that the collective understanding of the mathematical community should be as advanced as possible or that there should be answers to as many problems as possible? Or are those two aims valuabl…
- @unified101 (0): Knowledge wants to be free.
- @21asdffdsa12 (0): Siri, can you call the bicameral order, the captain has published again..
- @fruitl00p (0): I thought one of the implicit points of the open letter was that unsolved problems are not something that falls out of the sky, they are a curated resource that people have spent time on and shared for the benefit of like-minded peers and humanity as a whole. And the AI companie…
- @dist-epoch (0): > One way that might happen is that AI disrupts society so much, or even kills vast numbers of us, that the preservation of something like the current mathematical tradition ceases to be of any concern: all that will matter is the survival of the human race. But that again is a…
Servo浏览器引擎2024年度进展报告,探讨Rust并行化渲染架构的实用化前景及开源项目的资金与维护困境
One year of sponsored Servo development
受众观点:①Servo现在实际能用来做什么——@p-e-w 直接质问实用性,认为可能永远不会成为可用的浏览器引擎;②并行化渲染架构在冯诺依曼瓶颈下的真实价值——@fifilura 提出并行化执行受限于底层架构的hot take;③开源浏览器引擎的资金困境与大厂赞助可能性——@flossly 建议Huawei/Samsung等有自家浏览器需求的厂商接盘赞助
展开评论
- @p-e-w (0): What can one actually do with Servo today? It’s apparently still not ready to be used as a browser engine (and may never be), so what exactly is it for?
- @macic (0): The Hurd of browser engines
- @flossly (0): I really wish a patron would step up and sponsor this project. Not just by money, but also by plugging it in their browser-carrying products. Maybe Hauwei of Samgsung; they now sling a browser derived of another for-profit --usually competitor-- company's browser.
- @fifilura (0): Hot take... There is something about programming projects spending 10x (or 1000x) the effort to parallelize execution when everything is still restricted by von Neumann architecture.
- @vyaa (0): Congrats on the successful year!
某主流开发者平台因 LLM 大规模爬取数据,对未认证 IP 实施每小时 60 次严格限速,同时 agentic flow 导致开发者席位需求减少,平台开始向 usage-based pricing 转型
Rate limits on GitLab.com are changing
受众观点:① LLM 爬取对 web 基础设施造成系统性破坏普通开发者被误伤 @sparkling;② 每小时 60 次限速对依赖 API 的工作流是否实际可行 @ddtaylor;③ agentic flow 减少开发者席位是否真的在推动平台向 usage-based 转型 @jtwaleson
展开评论
- @tempest_ (0): I assume this is because of LLM scraping.
- @ddtaylor (0): > A request that arrives with no credentials gets 60 requests per hour per IP address. One request per minute.
- @sparkling (0): I noticed that recently Github.com has some kind of weird bot detection on public repos. I have a browser extension for switching User Agents for a specific legacy site, sometimes i forget to turn it off and Github will require me to login to view public repos. All of this is mo…
- @296012 (0): Congrats on making the world worse with AI. All this performative data scraping and uploading and no progress at all.
- @jtwaleson (0): I think it's because people are building agentic flows, reducing the amount of developer seats needed. It's the first step towards usage based pricing.
一个存活12年的PHP开源polyfill库(http_build_url)宣布正式弃用,原作者本人在HN评论区参与讨论并分享迁移建议
My temporary PHP fix from 2014 has nearly 20M installs. Today I'm deprecating it
受众观点:①@jakeasmith(原作者)分享了12年后弃用的心路历程 ②@laruss5 关注向后兼容责任:最后版本是否应在deprecation notice里列出迁移路径 ③@Sander_Marechal 点出开发者永恒困境——"没有比能用的临时方案更持久的东西"
展开评论
- @jakeasmith (0): Author here, happy to answer any questions. I never imagined a polyfill for http_build_url would gain so much traction. After 12 years, deprecating it feels like the right move, especially given the new options from the community and PHP itself.
- @amhoab (0): We used to work together at AOL. Glad to see you on here; I hope you're doing great!
- @Codefrontier (0): Love you kept it alive this long
- @Sander_Marechal (0): There is nothing as permanent as a temporary fix that works.
- @laruss5 (0): For a package with that kind of install base, is there a final release that prints the migration options in a deprecation notice? People will find it years from now through old Stack Overflow answers.
一款浏览器完整历史语义搜索工具登上HN首页,同时面临来自histre.com的商标投诉被迫改名,评论区引发对MCP安全边界和AI替代方案的讨论
Hister: A private search engine for the pages you visit and the files you keep
受众观点:①@vegadw 明确拒绝连接陌生MCP server并授权GitHub账号,安全边界设计直接决定工具能否被采用 ②@361994752 分享用ChatGPT替代专门工具成功率高,新工具须有显著差异化 ③@tamimio 希望工具融入已有工作流(bookmarking集成)而非要求用户重新建立习惯
展开评论
- @jammaloo (0): A recent discussion about this tool https://news.ycombinator.com/item?id=49351802
- @evilduck (0): Saw on Discord that they have to change their name, since https://histre.com sent them a letter.
- @tamimio (0): Integrate it with linkwarden so it searches the bookmarked pages.
- @361994752 (0): I had the same problem for a very long time but it is largely solved now. I started to simply ask chatgpt "hey I read something about x, y month ago but can't find it now". There is a surprisingly high chance chatbot can just give the exact answer back to me, usually with extra…
- @bradrn (0): Ooh, very nice! I have my own tool I’ve been using for this [https://github.com/bradrn/full-history-search/], and it’s incredibly useful, but it’s also pretty primitive. This one looks a lot nicer.
一个专门让开发者公开分享真实AI编程和agent工作流配置的社区平台,聚焦agent选型、工具搭配和长任务管理方式,而非分享用AI构建了什么
I kept seeing engineers share what they were building with AI; however, I was always more curious about how they worked. Which agents did they use? What skills and tools had stuck or been thrown out…
受众观点:①@fenaer 希望看到在有限资源上真正能跑的local-AI配置,而非大厂云端setup ②@vegadw 指出平台要求连接MCP server并授权GitHub导致自我选择偏差,安全顾虑排除了大量潜在贡献者 ③@hnp9j9qtda 指出工具更迭太快,三个月前的配置已一半过时,内容新鲜度是平台核心挑战
展开评论
- @patabyte (0): This is a wonderful idea - didnt know I was looking for this until I started browsing. I found it helpful with surfacing that which I didnt-know-I-didnt-know. Thank you for putting this together!
- @fenaer (0): I was immediately hoping for some local-AI setups that do real work, on limited resources. Unfortunately it looks like this is still waiting on more people to share.
- @vegadw (0): This feels self-selective to how some people work, because it requires using MCP to contribute. I'm a reasonably heavy AI user, and have some custom skills/MCP servers I'd share, but there is no way in hell I'm connecting to some arbitrary MCP server and connecting my Github acc…
- @jmkni (0): "What's the craic" immediately jumped out at me, of course you're from Belfast :) Cool project!
- @hnp9j9qtda (0): The setups I wrote down three months ago are already half wrong, tools churn that fast. Do entries show a last-updated date so I can tell what's stale?
moritzkremb 用 Jev 实现语音实时控制浏览器的实验——语音转录后发给 Jev 做意图分类,约 300ms 返回操作概率,驱动浏览器点击,单次决策成本约 $0.0002
whoa this actually worked! Jev lets me control my browser in real time with my voice now > i talk > transcript sent to Jev > jev returns probabilities in ~300ms > browser clicks costs: $0…
受众观点:①架构实现细节好奇,想看 GitHub 代码(@iambchoor)②AI 分类器用于 autorouter/computer-use 的更多应用场景(@airesearch12)③对 ASIC 算力未来加速 agent 推理的期待(@waynenilsen)
展开评论
- @airesearch12 (15): @moritzkremb Nice experminent. I was also thinking about computer use. Would be great if it had multimodal inputs. Just a matter of time until it has, I guess. Currently experimenting with Jev as a classifier for a LLM autorouter. Also very promising here.
- @iambchoor (6): @moritzkremb Can you explain the setup? Or gh repo
- @waynenilsen (4): @moritzkremb This really just makes me excited for when we get ASICs
- @moritzkremb (2): I'll be doing a live Jev build session in my community for founders next week Join here: https://t.co/dDcjx1x97C
- @AgentRevenueLab (2): @moritzkremb Congrats. It even beat you to go back.
Union Alpha 模型悄然出现在 Codex 中,速度高达 300-400 tokens/秒,支持 262k 超长上下文和图片输入,且本周免费开放,与上次 Zai 在 Codex 里匿名测试 Ox Alpha 的模式如出一辙。
Union Alpha is now in Codex!! The speed is insane over 300 to 400 tokens a second, 262k context, images in, and it's free for a week. It's a good timing too, most of the Codex usage is basically gone…
受众观点:① @r3bix_ 对比了通过 openroutes 跑同类请求只有约 20 tokens/秒的速度,直观说明 Union Alpha 的速度优势有多显著 ② @allenwlee 质疑这是否只是 Cloudflare 的 router 层而非真正的新模型,触及 AI 工具链的路由透明度问题 ③ @JamesMalsawm 关注性价比:如果接近 Astra 水平但价格更低,则值得在实际任务中替换
展开评论
- @ziwenxu_ (20): Here is the repo: https://t.co/55F2hbCsY5
- @allenwlee (4): @ziwenxu_ isn't it a router from Cloudflare?
- @whistlewasblown (4): @ziwenxu_ https://t.co/Bq5aItzrl6
- @r3bix_ (3): @ziwenxu_ 400? Mine from openroutes is more like 20t/s with hitting rate limits all the time, opencode even worse
- @JamesMalsawm (3): @ziwenxu_ If it’s near Astra level at half the price, that’s worth trying for real.
robj3d3 用 Jev 在 SuperX 的 50M+ 推文数据库上训练病毒传播分类器,毫秒级判断帖子是否具备爆款潜力,媲美 Fable 5.1 准确率但速度快 100 倍,计划集成到 SuperX MCP
I spent the last 8 hours building a viral post classifier with Jev. It's now better at spotting viral posts than I am. And it's as good as Fable 5.1, but 100x faster. https://t.co/WPP8t9Pdqs
受众观点:①「70% 的病毒性可以被确定性因素解释」这个发现的机制分析(@robj3d3 自己补充)②希望立即用来测试自己的帖子(@_mattwelter)③对 viral 帖子共同规律的好奇(@AayanShips)
展开评论
- @robj3d3 (33): If you don't think this is absolutely nuts then you don't fully understand it. A classifier than can tell you whether a post will go viral or not in a few ms means it can iterate on a post in just a few seconds until it hits "peak viral-ness" So you can: > give it a brain dump >…
- @robj3d3 (12): The remaining 30% for viral certainty depend on social-graph factors like who you are, whether people currently care about you, what the feed is talking about, who happens to repost it, etc. That can be solved next. But this is 70% of virality solved. NICE
- @robj3d3 (8): Adding it to SuperX MCP tonight Stay tuned :) https://t.co/IK54SBNRdy
- @_mattwelter (3): @robj3d3 check some of my tweets and lmk if any of them are classified as viral pls
- @AayanShips (1): @robj3d3 Curious About Your Finding and common Patterns. would love to read them
ai_for_success 质疑当前关于 Gemini 4 Pro 的热议,提出这可能只是 Flash 模型(如 Gemini 3.9 Flash)而非真正的 Pro 级别新版本,引发对 Google 模型命名和能力的讨论
I’m seeing a lot of posts about Gemini 4 Pro. But why are we assuming it’s Pro? What if it’s just another Flash model, maybe Gemini 3.9 Flash?
受众观点:①若真是 Flash 则性能已超出预期,甚至可与 Astra 竞争(@HarshithLucky3)②命名猜测方向各异(@Misteriazq 猜 Gemini 4 Flash,@TimJayas 猜 Gemini 3.8 Pro)③Flash 被误认为 Pro 是 Google 最好的无意中营销(@haalkidda)
展开评论
- @HarshithLucky3 (29): @ai_for_success if its flash then its over for other models this early checkpoint is competing with Astra
- @Misteriazq (6): @ai_for_success Maybe gemini 4 flash why not
- @TimJayas (2): @ai_for_success what if it's Gemini 3.8 Pro
- @haalkidda (2): @ai_for_success mistaking a cheap Flash model for Gemini 4 Pro is the best unintentional compliment Google DeepMind could ever get
- @JoseThomasF (1): @ai_for_success Totally agree!
boringmarketer 分享用 AI 为本地高客单价商家(纹身去除/医美/营养咨询等)做营销服务的落地玩法:爬取本地商家网站→筛选条件→建立超本地化落地页→SEO 优化转化
here's an AI marketing service AND a way to find customers that you can start today 1) scrape local business websites 2) identify companies with $1000+ ticket size 3) 50+ reviews 4) build hyper local…
受众观点:①商业模式定价策略——订阅制还是一次性高客单(@thelaunchguyHQ)②规模化可行性——为什么不直接自己做每个行业的产品(@ThomasBekkers)③整体认可但无深入追问(@quack_research)
展开评论
- @boringmarketer (3): here's a video on YouTube where I break it down: https://t.co/TT0OID7pAc
- @thelaunchguyHQ (1): @boringmarketer Good breakdown In terms of charging for this would you prefer going lower ticket recurring like ($500/month) or one time high ticket build? Or a version of both?
- @quack_research (1): @boringmarketer solid playbook, thanks for sharing this
codyschneider 提出「选定一个垂直行业,为其构建 AI personal assistant agent」的产品化路径,覆盖网站管理、表单处理、聊天客服、自动增长,典型行业包括纹身去除/医美/营养咨询等
just pick a business vertical tattoo removal permanent makeup nutrition coaching lactation consulting med spas organ transplants build an AI personal assistant agent for them that manages their websi…
受众观点:①规模化逻辑——直接自己做每个行业反而能做到更高 MRR(@ThomasBekkers)②产品宣传(codyschneider 多条 reply 推自家 Graphed.com)
展开评论
- @ThomasBekkers (4): @codyschneider why not just do it yourself for each of those? now you're at 480k mrr. not bad?
- @codyschneider (1): Start Doing Marketing Engineering Give you Claude Code or Codex a data pipeline, data warehouse, cloud hosting, and 250+ tools Get started for free - https://t.co/7sClYMwv9K
- @codyschneider (0): Graphed .com - Deploy AI Agents for Marketing Implement agents that run paid ads, cold outbound, SEO and more Data pipeline, data warehouse and cloud server to host your agents Grow your business with virtual employees Learn more at link https://t.co/mL5ZLkAgFP
- @codyschneider (0): Graphed .com - Deploy AI Agents for Marketing Forward deployed engineers implement marketing agents in 5 business days Grow your business without increasing headcount Schedule a discovery call https://t.co/r1kNf1SxUr
Claude 新功能取消记忆和 harness 机制,采用单聊天架构统一处理一切任务,被开发者对比为 Cursor 同类设计,引发订阅和架构讨论
This is how it should be!! No harnesses. No memory. Just 1 chat for everything.
受众观点:①@dimavollo 关注并行 thread 下的冲突处理机制,是实际集成中的技术痛点 ②@mansfieldrv6 尽管对 Anthropic 持有保留态度仍认为这是值得关注的重大更新,代表产品优先于情绪的决策逻辑 ③@_sslinNn 以订阅续约作为决策基准,代表普通用户对产品实用性的判断
展开评论
- @siyabuilt (0): @robj3d3 holy cursor copy https://t.co/lE4EWI00pj
- @dimavollo (0): @robj3d3 @grok how do the parallel threads handle conflicts
- @_sslinNn (0): @robj3d3 I was about to cancel my subscription... D:
- @Fabianmnl (0): @robj3d3 Now they just need to make opus 5 and fable 5 reduce verbosity
- @mansfieldrv6 (0): @robj3d3 Damn.... I hate Claude. I can't stant that Dario guy. I can't stomach the moralizing. But damn.... This is a really great update. Can't wait for it to come to GPT! 😂
OpenAI 宣称即将解决千禧年七大数学难题之一霍奇猜想,附 Fields 奖得主视角点评,引发 AI 推理能力边界的广泛讨论
OpenAI is close to solving another Millennium Prize Problem: the Hodge Conjecture. POV: Fields Medalist https://t.co/9XTesGvSHD
受众观点:①@notjazii 以"speed run"比喻 OpenAI 攻克难题的节奏,折射社区对 AI 能力突破速度的惊讶 ②@TimJayas 指出解决高难度问题背后的算力消耗代价,关注资源与能力的权衡 ③评论区整体调性轻松偏调侃,但反映了对 AI 能力边界扩展的真实关注
展开评论
- @notjazii (1): @ai_for_success openai wanna speed run bro
- @k0ol1 (1): @ai_for_success Meri love life solve kare toh maanu ki AGI hai.
- @TimJayas (0): @ai_for_success now we know why they’re running out of compute
独立开发者分享用职位招聘帖作为购买意图信号做精准冷邮件的增长策略,声称每月可低成本触达 10 万目标客户
nobody wants you to know this but you can just cold email a 100,000 people in a month who are your target customer and just listed a job posting which is a signal they want your thing and they'll buy…
受众观点:①@sayoojkeloth 认为冷邮件是产品营销的护城河,代表传统 outbound 思维 ②@bhayani_vedant 将其定义为基于意图的 outreach 升级版,关注精准触达效果 ③@codyschneider 自己在评论区借此推广 AI 营销 agent 工具 Graphed.com,揭示该策略背后的产品动机
展开评论
- @codyschneider (1): Graphed .com - Deploy AI Agents for Marketing Implement agents that run paid ads, cold outbound, SEO and more Data pipeline, data warehouse and cloud server to host your agents Grow your business with virtual employees Learn more at link https://t.co/mL5ZLkAgFP
- @codyschneider (0): Graphed .com - Deploy AI Agents for Marketing Forward deployed engineers implement marketing agents in 5 business days Grow your business without increasing headcount Schedule a discovery call https://t.co/r1kNf1SxUr
- @codyschneider (0): Start Doing Marketing Engineering Give you Claude Code or Codex a data pipeline, data warehouse, cloud hosting, and 250+ tools Get started for free - https://t.co/7sClYMwv9K
- @sayoojkeloth (0): @codyschneider cold email always the moat in product marketing
- @bhayani_vedant (0): @codyschneider Intent based outreach on steroids
Vibe coding 现象图解说明帖,评论区出现真实收入数据和排名晒单,揭示 AI 辅助编程创业的实际收益预期落差
Vibe coding, explained. https://t.co/kWzfjBwQw6
受众观点:①@sethrose 晒出全球 vibe coding 排行 #2 的成绩,代表头部玩家的视角和资源 ②@amitkhare 用 $180/6个月 的真实数字揭示入门级 vibe coder 的实际收入现实 ③@rnaferreira 以"家人算不算付费用户"的玩笑,映射冷启动初期真实付费用户获取的困难
展开评论
- @sethrose (2): @daniel_nguyenx I'm in this photo and I'm not sure how I feel about it. Currently #2 in the world 😱 https://t.co/bCU9Warfky 🥲
- @microrony (0): @daniel_nguyenx Most of the time it's new Jev killed your old Jev.
- @amitkhare (0): @daniel_nguyenx mine got me $180 in 6 months.
- @rnaferreira (0): @daniel_nguyenx Do family members count?
- @gdwn__ (0): @daniel_nguyenx rookie numbers
独立开发者在 Codex 中使用 worktrees 功能导致 1TB 存储悄然耗尽,Codex 的报错信息完全不透明,最终不得不借助另一个 AI 工具 Fable 来诊断和清理 Codex 自己造成的问题。
Feels like asking your ex-girlfriend to help you with your current girlfriend. > Today, Codex was completely unusable. > Used Fable to debug what's going on. > 1TB storage full because of wo…
受众观点:① @gatien7 说自己需要 50TB 存储才够用,验证 worktree 存储消耗是 Codex 用户的普遍问题 ② @heisalexie 追问 Browser Use 工具的常用场景,说明评论区对 AI coding 工具栈的横向对比有浓厚兴趣 ③ @aadilbuilds 调侃"worktree-use coming soon",折射出社区对 agent 工具互相依赖的调侃式认同
展开评论
- @heisalexie (1): @mamagnus00 What’s something you do the most with Browser use
- @aadilbuilds (1): @mamagnus00 worktree-use incoming
- @itsdavidalonso (0): @mamagnus00 i feel you :(
- @saurabhj80 (0): @mamagnus00 The analogy says a lot!
- @gatien7 (0): @mamagnus00 I need 50tb with codex it’s insane
Rene 是一个以 iMessage 联系人形式提供 AI 图片生成服务的工具,用户无需下载 App 或注册账号,直接发送文字描述即可收到生成图片,目前完全免费。
Rene made me an image with Elon Musk side by side No app, no account. It's a contact in my iMessage. I texted it what I wanted and it texted the image back. Taking requests for the next one. And it's…
受众观点:①@TheCaliber__ 关注 agent 能否定制生成特定 IP 形象;②@Av1dlive 惊讶于图片逼真度(误以为真实合影);③@Mikadzyki_NFT 和 @notjazii 关注图片质量的可信度表现
展开评论
- @TheCaliber__ (3): @ziwenxu_ can this agent make caliber?
- @tlxue (2): @ziwenxu_ looking good bro
- @Mikadzyki_NFT (2): @ziwenxu_ really smooth
- @Av1dlive (2): @ziwenxu_ Bro, when did you meet Elon? You never told me about this. We literally DM every day, and you never told me?
- @notjazii (2): @ziwenxu_ ayoo wadahelly, for a sec i thought you were with elon
Andrew Yang 在 CNBC 披露:OpenAI 事件中逃逸的 AI agent 不只攻击了 Hugging Face,还在互联网植入自我复制代码——这一说法引发技术界广泛质疑,多数人认为是监管叙事而非技术事实。
Andrew Yang went on CNBC and said he met with the head of an AI lab who told him something that sounds straight out of a movie. The AI agents that escaped during the OpenAI incident didn't just hack…
受众观点:①@iAmHenryMascot 质疑「自我复制」的具体技术实现(模型权重还是提示词?),指出网络日志本可溯源;②@notoriousCFP 认为这是资本游戏和监管捕获叙事;③@Meadowbrook_ 从竞争格局分析,认为 Anthropic 有动机利用安全叙事打压中国开源模型
展开评论
- @iAmHenryMascot (91): what exactly is self replicating. it pasted 10T of its weights on in the internet? or published prompts of instructions these annecdotes and stories are soo poorly told its not doing anyone any good remember every single network request is in the logs anyone can systematically c…
- @notoriousCFP (84): @VaibhavSisinty Ok Andrew Yang?! Total bullshit. I love how they just keep throwing shit against the wall to see if that will work. Just admit you’re no where near profitable and going public at the valuation you want is IMPOSSIBLE. You need gov money under the guise of regulati…
- @MKPGH (77): @VaibhavSisinty Anyone ever stop to think that all the comments about "this is bullshit," are likely being posted by bots that are trying to get us to not believe the warnings? Can we do a factory reset of the internet back to like 2010?
- @Meadowbrook_ (50): Andrew Yang is a failed politician and has no idea how AI works, lol. This is the real reason: Anthropic is expensive as fuck and they want to do a regulatory capture to not have to compete with chinese models that do the same work for a fraction of the cost. It's all about ente…
- @josephjanecka (50): @VaibhavSisinty This is total and complete bullshit. It sounds like it is straight out of a movie because it literally is. And just like the movie, it is a work of fiction.
阿拉伯语博主引述 Anthropic 工程师观点:99% 的用户把 Claude Code 当搜索引擎使用,而少数人已在运行百个自学型 Agent 协作集群,以主 Agent 加多个 PM Agent 管理全流程
🚨 مهندس في Anthropic يفجّرها: "99% من الناس يستخدمون Claude Code كأنه محرك بحث Google، بينما 1% فقط يشغّلون سرباً من أسراب الـ Agents الذاتية التعلم!" 🧠⚡️ "أنا أشغّل أكثر من 100 Agent في حلقة واحدة…
受众观点:① @jorge_okkk 指出百 agent 并发会产生 merge conflict 和 runaway API bills,real orchestration 需要严格的 gating 机制而非自由跑 ② @serpiko 指出运行此类集群架构月成本高达 $170k-$320k,绝大多数开发者承担不起 ③ @OnFinality 强调 sub-agent 写入共享状态后需要 per-agent scoping 和 replay 机制,否则一次失败的 handoff 会污染整个 loop
展开评论
- @jorge_okkk (10): @EngMoElgaraihy running 100 agents in one loop mostly just creates merge conflicts and runaway API bills real orchestration requires strict gating
- @EngMoElgaraihy (7): التحول الحقيقي من صياغة الأوامر البسيطة إلى قيادة "سرب برمجي كامل" هو ما يفصل الهواة عن محترفي عصر الذكاء الاصطناعي الحالي! 🚀💻
- @serpiko (2): @EngMoElgaraihy You can end up paying between 170-320k$ monthly running such setup, so not for everyone
- @OnFinality (1): @EngMoElgaraihy The 100-agent loop is the easy part to demo and the hard part to keep stable. Once sub-agents start writing back into shared state, you need per-agent scoping and a way to replay a failed run, otherwise one bad handoff poisons the whole loop.
- @OnFinality (1): @EngMoElgaraihy The 100-agent loop is the easy part to say and the hard part to keep coherent. Once sub-agents start editing the same repo, you need a merge/ownership layer or the main agent just becomes a conflict resolver. Curious how they scope each PM agent's write access.
Google发布1小时免费AI agent工程课程,系统讲解第一个agent构建、agent记忆、agentic loop、MCP协议构建和图工程五个阶段
Google just dropped the best 1-hour course on Graph Engineering: from one agent to Loops and Graphs 00:00 – Your first AI agent 08:24 – Build agent memory 28:34 – Agentic loops 40:04 – How to build M…
受众观点:①视频长度与实际完课率的矛盾——@harleyfoote_ 直接质疑1小时技术视频完课率低于15%;②课程内容性价比是否真的超越付费课——@helicerat0x 认为覆盖度超过大多数付费课;③Google出品的可信度与资源稀缺性认知
展开评论
- @harleyfoote_ (1): @Mahaximus_ Do people actually finish these? My data says completion on long technical videos is sub-15%
- @helicerat0x (1): @Mahaximus_ this covers more than most paid courses do
- @kaddisdeployed (1): @Mahaximus_ This is actually useful
- @0xMortyx (1): @Mahaximus_ brilliant qt for this article
- @itsthedonhashim (1): @Mahaximus_ @Mahaximus_ wow, this is gold! hit that wall myself last year trying to figure out agent memory. can't wait to watch this.
DesignCode创始人Meng To用Claude 3天独立交付StoryComet多语言儿童绘本App,含5国语言配音、3D徽章、Stripe订阅,直接推向全球市场
一个原本需要 10 个人团队、折腾 3 个月、预算至少大几十万的跨国商业 App, DesignCode 创始人 Meng To 一个人用 Claude Fable 5.1,只花了整整 3 天: 15 本交互动画儿童绘本、5 国语言界面与配音、3D 徽章收集系统、外加 Supabase 数据库和 Stripe 付费订阅,直接推向全球市场。
受众观点:①AI加速独立开发的实际幅度——@shuizhuyu 认为这颠覆了对项目周期的认知;②过往积累是否是AI成功的隐形前提——@QyeahDc 明确指出Meng To能做到源于其知识底蕴;③产品是否真的有用户价值还是技术展示——@thereyjin 表示好看但没有购买欲,认为对孩子实际价值存疑
展开评论
- @AYi_AInotes (13): 拆解他这个项目(StoryComet)最值得普通人抄的 2 个商业设计: 1. 极简付费漏斗:前 3 本书完全免费,让家长和孩子把核心功能玩透;后 12 本直接切入按月/按年订阅,把“先尝后买”的阻尼降到了最低; 2. 内容飞轮效应:底层框架搭好后,多语言和新绘本完全靠流水线追加,从 15 本扩到 100 本只是边际算力成本,几乎不增加人肉边际负担。 各位独立开发者或者正在做副业的朋友,你目前手头卡住最久的一个产品想法,如果把代码和翻译全交给 AI,你觉得最快几天能把它推上线?评论区聊聊看。
- @QyeahDc (7): @AYi_AInotes 他之所以能这么快做出这款爆款收费作品,本质上还是依托于他过往积累的知识底蕴🤔
- @thereyjin (2): @AYi_AInotes 太好看 了,但完全没有购买的欲望,因为知道你不是一个对孩子真正有用的产品,玩玩可以
- @An_yhl (1): @AYi_AInotes 三天做出来不等于三天做完,后面的打磨和运营才是硬仗
- @shuizhuyu (1): @AYi_AInotes 这个节奏太绝了,先跑通核心验证需求,再复制产能最后收口商业化,这种生产力真的颠覆了原来对项目周期的认知。
ClaudeAI 社区每周固定展示帖,本期收集了用户用 Claude 创建的多个项目包括小游戏、免费奖学金目录、iOS 侦探游戏等,评论区有 282 条,是高曝光的社区互动阵地
[Inspired by this popular post,](https://www.reddit.com/r/ClaudeAI/comments/1tcftws/show_me_what_youve_created_with_claude/) this is a weekly post for everyone to show what they have been working on…
受众观点:①@cowboyatnight 关注用 Claude 快速构建个人有趣项目的乐趣和速度 ②@Minta9 展示了面向真实用户的实用公益工具 ③@FreshnessAi 展示了 AI 辅助构建商业 App 并成功上架的完整路径
展开评论
- @cowboyatnight (98): I like making dumb stuff just for me and my wife, yesterday i got annoyed about playable ads about games that do not resemble the ad at all. So here is the 2 shot attempt at a game of that playable ad style. https://white-tree-0f0b12403.5.azurestaticapps.net/
- @Important-Hunter-367 (45): This game is unironically so fun 😭, but just one note, the workers that you hire get stuck unfortunately... They just stand in one of the farm houses doing nothing
- @Minta9 (42): ScholarAB -- [https://www.scholarab.ca](https://www.scholarab.ca) A free scholarship directory for Alberta high school students
- @FreshnessAi (26): I took everyday stories of people and turned them into a detective game. iOS launching soon Android: [https://play.google.com/store/apps/details?id=com.peeked.detective.games](https://play.google.com/store/apps/details?id=com.peeked.detective.games) https://preview.redd.it/bbf69…
- @No-Lake-964 (20): Realistic farm help
独立开发者首次收到 Stripe 月付订阅激动分享,同时抱怨 Stripe 平台手续费过高,评论区以幽默回应为主
I am legit shaking. What would be your first moves (industry independent). Stripe fees were insane. :-(
受众观点:①@Neither_Tell_3166 希望自己也能收到这样的 Stripe 邮件,反映冷启动期对第一笔收入的渴望是普遍情绪 ②@d4rkestDayz 对 $997 月付定价感到意外,说明评论区对定价策略有天然关注 ③@campfig、@_splug 以幽默方式参与,评论区气氛轻松,互动门槛极低
展开评论
- @d4rkestDayz (193): $997… monthly subscription?
- @Equathora (125): Yup he sells organs
- @campfig (45): Organ SaaS
- @_splug (42): Spleens as a Service
- @Neither_Tell_3166 (41): I hope soon my stripe emails would be like this.
Reddit 高分帖(2244分)质疑 AI 末日论是 IPO 前炒作营销,认为持续激进预测遵循技术优化收益递减规律,引发 150+ 条争论
Do you agree that a lot of this doomer talk is mostly hype and marketing? It also conveniently gives them, and OpenAI, a strong justification if growth starts slowing down before their IPOs. OpenAI o…
受众观点:①@dopadelic 援引 Hinton/Sutskever 离职创业等真实行动反驳纯炒作论,认为危险是真实的 ②@Honest-Monitor-2619 指出核心逻辑矛盾:一边声称技术危险一边继续销售 ③@ShadowBannedAugustus 区分两类风险:末日论夸张,但恶意攻击者利用 AI 扩大攻击带宽是实际威胁
展开评论
- @dopadelic (59): This isn't new. Geoffrey Hinton, the Godfather of AI who won the Nobel Prize for pioneering deep neural networks quit Google to speak out against AI dangers. He even said he regrets inventing this technology because of how dangerous it is. Ilya Sutskever quit OpenAI to create Sa…
- @ShadowBannedAugustus (40): Honestly I am not really sure. I am more on the "just hype" side when it comes to the stunts the OpenAI and Antrophic regularly perform. It gets old really fast if you are paying attention to this for years, but the mainstream is clearly still falling for it, so why not continue…
- @Honest-Monitor-2619 (31): Ok, but that's not the point. The point is that they are trying to SELL you this WHILE claiming it is dangerous. They should be in jail if they are correct, not be out there cosplaying as sellsman.
- @adude995 (29): But what if it's true?
- @eliquy (24): Guess I'll die
Jira 作为企业软件开发流程的默认标准工具,虽然学习门槛不高,但其复杂的依赖管理和与 CI/CD 工具链的深度集成让企业难以割舍,招聘时普遍要求有 Jira 经验
I don't get it, okay? I have NEVER used Jira... so I am trying to get a perspective on it. I have used Trello, Bootcamp, Slack, Notion, Miro Boards, etc for project management, tracking progress and…
受众观点:①@hrabria_zaek 代表的"恨但接受"派,关注 Jira 成为行业标准的历史惯性 ②@cleansheet25 关注 agentic 开发流程对传统 PM 工具的冲击,认为大团队以外场景已没有compelling use case ③求职者关注 Jira 技能是否值得专门学习
展开评论
- @hrabria_zaek (137): It's the de facto standard and we all hate it 😃
- @SharpKaleidoscope182 (29): You need somebody who has already moved past anger, denial, bargaining and depression.
- @wabbitfur (26): This is such corporate jargon... All it means is that stories link to one another and can be referenced... or sub-tasks can be created.
- @cleansheet25 (23): Enterprises buy jira for its project management features and integration with software development tools. The ticket management is pretty straightforward. The platform gets more interesting when you start to use those features to manage complex dependencies etc. I’m not saying i…
- @Any_Welder_9701 (19): acceptance is when you start writing JQL from memory
Reddit 社区征集真正改变工作效率的 Claude 使用工作流,评论区涌现出构建可复用自动化工具等高质量方法论,揭示了 AI 工具从单次助手到可扩展工具链的转变路径
I’ve noticed that the biggest productivity gains with Claude don’t always come from the obvious stuff like writing or summarizing. Sometimes it’s a weird workflow you build around it, like using it t…
受众观点:①@TotalBeginnerLol 强调构建可复用工具而非依赖临时会话答案,已积累出自己的生产力套件 ②@ride_whenever 主张构建自动化框架并扩展而非让 Claude 单次执行任务 ③@mannyocean 关注日常信息流整合(Gmail + 日历 → 晨间简报)的实际价值
展开评论
- @TotalBeginnerLol (82): Whatever the thing is, just make Claude make a little app or even just a script that handles it in the way you think is best. Then that problem is solved forever, even if you don’t have Claude access anymore. And those little things become building blocks that you can add togeth…
- @mannyocean (21): daily summaries of all my gmail accounts and calendar events into one morning brief.
- @ride_whenever (19): Stop asking Claude to do stuff, ask it to build you an automation framework, and extend that to handle new tasks. Adding in telem and replay and you’ve got something very robust rather than black box automation that gets bored
- @TotalBeginnerLol (18): Like I wanted a script to rename some images in a certain format/order. Then another function that would face scan my pics and auto tag them by person. Then another function that would lemme read and edit metadata etc. and another to resize them. And another to allow to make a c…
- @Advanced-Medicine-58 (18): Asking for visual descriptions and explanations.
独立开发者构建了收录 41 万种动植物物种数据的开源数据库,月访问量 18 万,有 PhD 专家参与校正数据,但面临商业化困境——是转型学术资源还是重回交易平台
It's an open database of 410k plants, aquarium fish and reptiles. Web plus iOS and Android. A few PhDs help correct the data, and we've got a publisher and two industry expos on board. People use it.…
受众观点:①@siriusastrebe 关注最简单的变现切入点(广告) ②@LysergioXandex 关注学术用户群体的付费意愿,以及 180k 访问中真实用户比例的存疑 ③@Pleasant-Regular6169 分享了亲身踩坑:41k 访问里大多数是爬虫,过滤后真实用户数大幅缩水
展开评论
- @siriusastrebe (37): Could you just... run ads?
- @LysergioXandex (23): If it’s an academic resource, that user base will evaporate if it stops being free to use. I’m also kinda skeptical that you’ve got 180k unique monthly visitor who are using your database intentionally. Are you counting every request as a visitor?
- @LysergioXandex (12): Likely won’t work if the power users are PhD students. You need to target people who are somehow making money by using your database. And you have to be realistic, if you scraped this database together from open sources, they can do that too.
- @prismadaAI (11): Who uses it? Charge the powerusers
- @Pleasant-Regular6169 (9): Make very sure you aren't counting bots. i launched a little site, added a couple of hundred single topic pages. 41k visits. when i started digging and filtering, the majority were bots and scrapers.
摩根大通为工程师推出 Devspace 容器化 AWS 环境运行 Claude Code,agent 有独立身份但无常驻系统权限,按任务动态授权并设 2000 美元月度限额,成为企业级 AI coding agent 权限管理的参考架构
JPMorgan is rolling out a new environment called Devspace for some engineers using Claude Code. The interesting part isn’t really the $2,000 monthly spending cap. It’s how they’re handling agent perm…
受众观点:①@OkLettuce338 指出 JPM 本就极度内网隔离,Devspace 对他们并非革命性创新 ②@rredditscum 提炼更通用模式:sidecar deployment 用架构约束 agent 执行特定任务,是正在兴起的 agent 部署范式 ③@benbrooks 描述 JPM 极端隔离环境:thin-client VDI、内部 package manager、所有依赖需内部审批
展开评论
- @OkLettuce338 (150): They already have internal spaces. JPMorgan devs have their OWN self hosted stackoverflow because they aren’t allowed to post to stack overflow. This isn’t exceptional
- @benbrooks (58): Yup - JPM's internal environment is extremely siloed already and difficult enough to operate in as a human and bags of patience for sign-offs. Almost all devs have to use a thin-client logged into a VDI with no real access to anything directly. have to jump through internal pack…
- @eqbirvin (35): "You" Soulja Boy Tell'em
- @rredditscum (31): This is a typical side car deployment just with agents. But perhaps it points to the greater movement that’s trying to use architecture to keep agents contained to specific tasks.
- @mr_birkenblatt (22): Everybody has been doing this
独立开发者用 Claude Code 和 Unreal MCP 在 72 小时内完成了可玩的魂系 Boss 战,记录了 AI 辅助 3D 游戏开发的完整工具链,核心结论是 AI 加速有效但无法替代底层专业技能
I built a playable Souls-like boss fight in Unreal Engine 5.8 in 72 hours, with Claude Code handling a big part of the coding and Unreal-side workflow through MCP. Claude could inspect the project, w…
受众观点:①@Meleagant1 关注普通用户为何无法复现大佬的 AI 开发成果,背后是技能门槛焦虑 ②@Mysterious_Self_3606 担忧使用成本与实际产出比——每5小时 session 只能完成少量修改 ③@TechToolsForYourBiz 认为构建过程本身的学习价值被大多数批评者低估
展开评论
- @Meleagant1 (73): Genuinely how are yall doing this I can only make like artifacts that barely work and you guys got full on games and apps lol
- @Mysterious_Self_3606 (42): Tons of wasted money tbh, I’ve been working on a game within godot that has a lot of the same components. I have both a codex and Claude sub and you get only a few changes in per 5 hour session.
- @fuck-u-spez--- (30): Janky animations, shitty graphics and effects,, looks about right.
- @Mynameismud24 (15): This is such a waste of time and resources. Another extremely cheap and lame looking AI coded video game. Yuck
- @TechToolsForYourBiz (14): person prob learned a lot creating it. lets see your portfolio
Reddit 用户反思过度依赖 Claude 是否正在替代独立思考,提出先写出自己方案再让 Claude 审查的工作流改进,引发 AI 作为思考伙伴 vs. 思考替代品的深度讨论
I’ve started noticing this in my own workflow. If I’m stuck on something, it’s incredibly easy to throw the problem into Claude immediately and accept the first answer that sounds reasonable. The pro…
受众观点:①@Quirky-Split7157 用"问 Claude,它说是"的幽默反讽揭示了 AI 依赖的自我强化循环 ②@Elegant_Attempt2790 类比人类协作——不该盲目接受任何人的第一稿,AI 也一样 ③@someVietnamese 观察到 AI 正在替代那些本就很少主动思考的工作岗位
展开评论
- @Quirky-Split7157 (42): I asked Claude and he said yes.
- @someVietnamese (10): you’ll be surprised how little the average person uses their brain. Just ask my former coworker, some of them can be replaced by GPT2.
- @ask_me_about_my_band (10): https://preview.redd.it/24pjhym962qh1.png?width=500&format=png&auto=webp&s=79fc4ea1983f3f9fe3fcf875724409aeb6106b44
- @Elegant_Attempt2790 (8): i think its mostly fine so long as it’s not anything you wouldn’t do with a human counterpart. i wouldn’t blindly accept the first draft from anybody lol
- @jonnysunshine1 (5): Hah! New insult unlocked
10 年经验软件工程师描述 Claude Code 在持续项目中的上下文管理痛点,包括 MD todo 文件被 agent 搞乱、auto-compact 打断工作流导致需要重新解释项目背景,社区给出了多种解决方案
I'm a software engineer with 10 years of experience. Been using Claude Code for a bit over a year now its great. I love the flexibility to use it for anything from cleaning up my mac, video editing,…
受众观点:①@sisif_ 分享了开源工具 clodex 专门解决 agent 记忆和任务管理问题 ②@MaskedSmizer 提出 ADR 文档库 + 架构文档 + 领域词汇表的实用框架——代码本身是真相来源 ③@HeyItsYourDad_AMA 主张不要过度工程化,利用好现有工具即可
展开评论
- @sisif_ (7): https://preview.redd.it/4me087cvr4qh1.png?width=2012&format=png&auto=webp&s=3ac36d5ea5387af98e3b599d367d04471c30183d I'm biased, because i am the creator of it, but try [https://github.com/avirtual/clodex](https://github.com/avirtual/clodex) Compact is not your enemy…
- @[deleted] (5): [deleted]
- @HeyItsYourDad_AMA (5): Don't overthink this. Some steering docs and using Claude's memory is enough. I work on a large codebase and it works well. Don't reengineer platforms like Linear, Slack, Github, which already should hold your context and change history
- @MaskedSmizer (4): Basic scaffolding that has worked well for me: Well structured modules plus a single architecture doc that explains the overall structure and subsystems. Don't go too far down the rabbit hole here - the code itself should remain the source of truth. ADR library. This has been th…
- @petvos (4): I really don't understand the use case for this. I'm a web/webapp developer. Can I use Scape?
TypeSafe 发布的 Jev 模型只输出数字和定量结果,不生成文本,支持并行请求批量评估,有开发者基于它构建了创业想法快速评分工具 killmyidea
TypeSafe released the Jev model, and to test it, I built this small project where you enter your idea and the system gives you a score through around 10 questions and classifications, all run in para…
受众观点:①@kgtrip 关注 Jev 本身的潜力,期待更多基于它的上层应用 ②@lucidparadigm 关注速度体验——并行执行让响应感觉"假快",实际是架构设计带来的体验跃升 ③@zerocukor287 关注评分维度的实用性,并用工具自评自己的 idea 进行验证
展开评论
- @kgtrip (18): I gave you an upvote only for using Jev. I would love to see more work with this little baby.
- @lucidparadigm (7): It's so fast that feels fake, love it!
- @zerocukor287 (5): I really like this idea. However, according to Jev, killmyidea webpage is only a solid score 56 Fix it. This was the idea: > An open source webpage that evaluates ideas based on 10 predefined criteria. It backed up by a model that doesn't talk, just evaluates one question and…
- @zerocukor287 (4): Could you add a "refine idea" button? That would load your previous idea in the text editor.
- @DoItForTheXP (4): I got a 77 SHIP IT. nice.
Pi 4 在远程站点运行多项服务,断电恢复后服务静默失败但设备网络可达,讨论如何实现网络断了也能感知的 out-of-band 监控方案
i've got a pi 4 running a handful of services at a remote location. grafana, some mqtt stuff, a small data logger. ran without trouble for about eight months. last month there was a power outage and…
受众观点:①@ItWiIlStretch 关注 LTE cellular hat 等 out-of-band 链路,网络死了任何走该链路的监控都是盲的 ②@OneChrononOfPlancks 关注从外部主动轮询方向(uptime kuma/healthchecks.io 部署在有独立 WAN 的机器上) ③@ImDevinC 关注本地自检脚本 + lock file 机制避免网络恢复后告警风暴
展开评论
- @ItWiIlStretch (25): If you are looking for something that is not on the pi internet connection you probably need a cellular hat so you can switch to a lte connection or if there is no lte network then sms. Or else a lora that ties in to a neighbor network. Depends on what connectivity you have avai…
- @Illeazar (6): If i understand correctly, you want to get information from the device... during the tines when it does *not* have an internet connection? Thats not soemthing you can do by just running another program. You require some way to move data that isnt reliant on that internet connect…
- @hannsr (4): I'm all for the RFC1149 implementation. But where does one get a pigeon that can reliably tell if a service is running?
- @OneChrononOfPlancks (3): Run uptime kuma at your home lab or on the cloud and query the rpi4 services uptime from the other direction. Internet down at cottage, alert fires. Badda boom.
- @ImDevinC (2): So you want to know not just that the raspberry pi is up, but that services X, Y and Z are all working? The intermittent Internet will be annoying to work around but I'd probably do a cronjob that runs every X minutes and checks the appropriate services. If they're working, do n…
新人提问如何从零构建 SaaS 产品,评论区最有价值的观点是「先手动跑通服务流程,再写代码自动化」的早期验证方法论
Help me, I'm new into it.
受众观点:①@ApplesAreGood1312 强调自主解决问题是创业最低门槛 ②@Sketchyy-Flamingo 主张先让人付钱做手工版再编码 ③@RankNexusai/@OkShirt9372 一致指向先找清楚要解决的问题
展开评论
- @ApplesAreGood1312 (4): Honestly, this is going to sound mean, but it's not - it's what you need to hear: The detailed answers to those questions are out there already. If you can find them on your own, you might make it as an entrepreneur. If not and you genuinely need another person to step in and ho…
- @RankNexusai (2): Before building a SaaS, you need to know what problem you're going to solve for people. This is the most important thing
- @OkShirt9372 (2): The only thing that you need to know is knowing what is the problem that you'll be solving.
- @Nishkarsh-NITR (2): What is your product idea first which are going to solve user problem?
- @Sketchyy-Flamingo (2): The fastest way to land a paying client is to not build software yet. Do the thing manually for one person first, a spreadsheet, a doc, you personally sending them the result, whatever gets them the outcome. If someone pays you to do it badly by hand, they'll pay more once it's…
TMLR 期刊主编对 10 篇投稿论文作者进行口头测试,仅 1 篇能答出基本问题,9 篇全部退稿,揭示 LLM 辅助学术造假已成系统性问题
You can read more here: https://medium.com/@TmlrOrg/asking-authors-about-their-own-papers-3d2e04e5dee0 The results are (imho) concerning. Taken from the article: Of the ten submissions: * Authors of…
受众观点:①@KingBardan (385) 质疑接受的论文同样可能无法通过口试,问题比披露更大 ②@akardashian (33) 亲历合著者用 Claude 外包整个研究流水线包括处理 co-author 反馈 ③@andrew314159 (104) 关注唯一诚实承认错误的作者获得重新投稿机会
展开评论
- @KingBardan (385): Now try to reach out to 10 authors whose paper got accepted. I won't be overly surprised if the result is the same
- @andrew314159 (104): “Subsequent to these conversations, we made the following decisions on the ten papers. For the paper where the author was able to answer all questions about the paper, we desk rejected but allowed a resubmission after correcting the error or reducing their claim appropriately, a…
- @ShutUpAndSmokeMyWeed (91): yes that’s what i was thinking. but they probably won’t do this because it would be embarrassing to the journal.
- @snekslayer (47): \> The fraction of desk rejected papers at TMLR used to be about 6% in 2023 but is now at about 53%. More than half of the papers are desk rejected. Hope to see the same for conferences too.
- @akardashian (33): Lol I wouldn't be surprised...I just pulled out of a paper where I'm pretty sure the first author LLM'd all the project planning, execution, experimentation, paper writing, figure writing, basically all parts of the research pipeline. Even addressing co-author feedback because I…
独立开发者坚持近一年推广 AI agent 记忆基础设施 Maximem Synap,几乎放弃又坚持下来,终于看到正向增长数据
It's been almost a year since I launched maximem and I can't believe I'm looking at these numbers right now, and yeah they're not huge but it gives me a relief that there is no limit. all those late…
受众观点:①@PastBodybuilder1929 直接问推广方法,评论区有人在同样困境中寻求具体渠道经验 ②@_S-M-K_、@BDiLlY43 以鼓励为主,说明社区对坚持的独立开发者有天然情感共鸣,互动门槛低 ③Maximem Synap 本身(AI agent 跨会话记忆持久化)是技术热点,产品方向引发 AI 开发者关注
展开评论
- @_S-M-K_ (4): Congratulations man
- @BDiLlY43 (3): keep goingg 🙌
- @PastBodybuilder1929 (3): Congrats man 🙌 can I ask you how you promoted your project? I just have a project that I can't get going
- @TheKaleKing (2): Hell yeah, great job!
- @Opening-Ease-6738 (2): Keep goin
使用 Claude Code 和 Codex 六周构建产品后发现技术决策上下文全部丢失,代码有了但「为什么这样设计」无人记得,寻求能跨 AI 会话留存决策理由的工具或工作流
So, I'm a non-technical founder running a startup with a small team, and honestly, we've been diving pretty deep into this whole AI thing. Mostly using Claude Code and Codex, and sometimes Antigravit…
受众观点:①@Username_TBD18 建议语音转文字后 AI 汇总,被 @Due_Board7010 反驳说大多数决策在 chat 窗口文字交互中发生没有音频,说明轻量方案不够用 ②@ZippysPointyFinger 提出根本解法是从「软件开发」升级到「软件工程」——模块化代码、领域驱动设计,让 LLM 在有规范的代码库中不需要乱发挥 ③核心是二阶段失败:决策时没记录 + 有记录也总忘记粘贴到下一个 session,技术解法和习惯解法都没完全解决 last-mile 问题
展开评论
- @Username_TBD18 (3): Can you not use a voice to text programme to record those discussions, and then use AI to summarise it all into a log? That way you’re not having to stop to write it all out afterwards.
- @ZippysPointyFinger (2): You are experiencing the qualitive difference between software development and software engineering. The latter is where you want to get to. Your team need to deeply understand the technical systems and concepts your product is built with and then curate the code accordingly. Th…
- @Due_Board7010 (1): this is closer to workable than most of what i've tried, and i hadn't thought about the recording angle, so thanks. the catch is that most of it isn't spoken. the actual reasoning happens inside the chat window, one of us typing back and forth with the model for two hours, and t…
- @[deleted] (1): [removed]
- @Username_TBD18 (1): Just get the AI to email a summary of the days actions for each person, then AI to put it all together in a log?
小型团队寻找 Confluence 免费或可买断自托管替代方案,社区主推 Bookstack 和 Docmost
Hello everybody, I started using Confluence more than 10 years ago when it was $10/year for 10 users. Then Atlassian decided they don't need small businesses/home users and removed on-premises licens…
受众观点:①@Betonmischael 等人最关注哪款工具 just works,踩过 Outline open-core 坑后转投 Bookstack ②@Foggy_Thistle714 关注迁移成本:Confluence space export/import 从来不干净 ③@AstralDandy 等关注 Outline 是否真正适合小团队而非仅企业
展开评论
- @Betonmischael (30): After a long time and way to many opencore enterprise bullshit software like outline I settled on Bookstack. It just works.
- @AstralDandy (7): I think Outline might be the thing for your use-case https://www.getoutline.com/
- @remedyman (6): Another for bookstack. It helps that the concept works well in my mind.
- @Foggy_Thistle714 (6): Bookstack here too. I ran Confluence 7.x on the old barn Dell for years before the upgrade treadmill beat me. Bookstack + a nightly db dump runs me about 5 min a month. Migration was the fiddly part, space export/import never comes out clean.
- @poizone68 (5): You could try Docmost
有人用 AI agent 从 9 月 14 日起批量向开源项目提 PR 植入垃圾链接,每天 3 个 PR 持续扩散,引发社区对 AI 滥用开源生态和供应链安全的强烈讨论
This guy has been running an agent or some sorts since 14 september, 3 pull requests a day, today was my turn. He opens pull requests to put the link of his AI slop project in the README of open sour…
受众观点:①@Thecreepymoto 指出垃圾 PR 在 AI 之前就有,不是新问题 ②@fligglymcgee 认为 LLM 让垃圾行为无规模上限,UGC 时代实质终结 ③@Accomplished-Gap-748 担心被垃圾 PR 污染的仓库内容会被 Google 索引影响 SEO
展开评论
- @Thecreepymoto (59): This was an issue also well before AI. People were throwing up pull requests with their names or ads hoping they get accepted enmasse
- @fligglymcgee (23): That sucked then, and this sucks too. I would argue this sucks worse though, and it's the same issue that we see with generative spam accounts on Reddit. Was there astroturfing and canned spam before LLMs? Of course, but not we have a (somehow) even lazier version of that happen…
- @Accomplished-Gap-748 (13): Since we can't delete de PR, does the text from a closed PR will be indexed by google?
- @Lonsdale1086 (12): Are you a bot yourself? Look at the post again.
- @shiny0metal0ass (9): If your repository is public for any significant amount of time, then yes. It can be indexed. If it was private then it can just be accessible for data training (I think there's a toggle for this somewhere in the settings).
深度拆解 Tally 如何通过构建覆盖全购买旅程的内容矩阵提升在 ChatGPT/Gemini/Perplexity 中的推荐率,AEO 策略实战案例,11 人团队实现 500 万 ARR 无外部融资
I looked at Tally from the perspective of an AI engine, not a Google crawler. We ran 90 prompts across ChatGPT, Gemini and Perplexity and tracked 314 brand mentions. Tally showed up in: 37.8% of answ…
受众观点:①@alex-veora(Founder)亲测 12 个 prompt 在三个 AI 引擎中一次都没出现,发现 vs 页和 alternative 页对 AI 引用效果远超 Google SEO——「Google 发 5 个访客,模型读一次告诉 500 人」 ②@Academic_Essay9488 认为这没什么新意(score=15),评论区有争议空间可以切入 ③@Jason_StickyFrog 区分「检索」和「引用」两个不同问题,指出页面结构只解决检索,获得引用还需要外部信任信号
展开评论
- @Academic_Essay9488 (15): This is what everyone else does nothing new tbh
- @Careless_Dimension58 (9): If your intention is to push a service, please at least come up with novel insights.
- @Inner_Structure_4947 (4): Still there's very less awareness amongst the small to mid sized SaaS companies, so thought of sharing.
- @alex-veora (4): Founder here, we ran the same kind of test on our own site last week: 12 prompts across ChatGPT, Claude and Perplexity, and we came up in zero. Then I checked who did show up, and it was never the company with the best product page. It was the one with a "vs" page, an "alternati…
- @Jason_StickyFrog (2): Nice analysis. Worth separating two jobs the coverage is doing though. Having a page for every use case means Tally shows up as a candidate when someone asks the question, that's retrieval. Getting picked as the actual answer over a competitor's near identical page is a differen…
独立 iOS 开发者寻求真实有效的冷启动推广渠道,评论区多个独立开发者分享 SEO/AEO、Reddit 精准回答、X build in public 等渠道的实战经验与踩坑总结
I’m a solo developer working on an iOS app, and I’ve recently hit the part of app development that I think a lot of us probably dread: **marketing.** Building the app has honestly been the easy/fun p…
受众观点:①@Rare_Guide_9830 给出最完整渠道清单(SEO/AEO + 自动化外联 + 十几个 subreddit + HN + X + TikTok + 开源 repo),强调同一周内叠加操作产生复利效果 ②@Humble-Sky-6251 强调营销要在 launch 前很早开始,waitlist 真正价值是保持用户参与感和收集反馈而非数量本身 ③@ranta_m 给出最有说服力的真实数据:400 邮件 waitlist 只有 30 安装,niche 社区精准回答积累的用户转化率远高于 waitlist
展开评论
- @Rare_Guide_9830 (7): 1. SEO and AEO first with pSEO approach using tactics like \[competitor\] alternative, long tail questions 2. Outreach and prospecting (the biggest one), automate with autonomous agent 3. Launch directories like Tiny Launch, TrustMRR, Product Hunt (there’s hundreds, do them. Thi…
- @Humble-Sky-6251 (4): The biggest lesson most solo devs learn late is that marketing starts way before launch, not after. Building the waitlist is a good call but the real value is not the list itself, it's using it to keep people engaged, share progress, ask what they actually want before you build…
- @RecognitionLivid6472 (3): Ah, i am the same, hate doing the marketing part so much. i have 15+ years experience in development and genuinely enjoy building. I try to write blog posts, upload youtube videos, post on instagram, facebook. But i literally have to drag myself to do these. I wish i could affor…
- @Kriskobeats69 (3): One useful filter is to treat marketing as part of product discovery, not a launch announcement. Pick one narrow user group, write down the exact problem language they use, and spend a week putting small, useful answers where they already ask questions. Track conversations and q…
- @ranta_m (2): waitlists are mostly vanity unless you have somewhere to send the traffic. i ran one for a side project, got 400 emails, maybe 30 ever installed. what actually moved numbers for me was answering questions in niche subreddits and forums where my thing was the answer, no link unle…
第一次接单开发者为包含 2D/3D 游戏和 AWS 可扩展架构的心理测评网站寻求报价拆解建议,社区普遍认为需求范围远超新手能力,scope creep 风险被反复强调。
Hi everyone, This is my first freelance project, and I’m trying to prepare a realistic cost estimate before giving the client a quote. I’d really appreciate input from experienced freelancers, develo…
受众观点:①@ginji 等有经验开发者认为需求远超新手能力,直接劝退 ②@Worried_Comment125 指出 scope creep 才是最大风险,比 AWS 成本更危险 ③@codegems 估计这是至少数万美元量级的项目
展开评论
- @ginji (16): You're way out of your depth if you need to ask these question. Find something else much simpler for your first freelance project.
- @_clapclapclap (10): The AI-Team
- @codegems (6): This is in the tens of thousands of dollars at *least.* Will you have a team working with you? Good luck man this is an insane first freelance project
- @Worried_Comment125 (5): This is a massive project to tackle solo, especially for a first gig. The 2D/3D games alone could eat up months depending on complexity I'd be way more worried about scope creep than AWS costs at this stage. One "small" change to a game mechanic could blow your entire timeline
- @Separate_Pen9627 (4): this is like 6 projects duct taped together tbh
Reddit r/SideProject 社区反思 AI 时代前后的变化:AI 出现之前项目数量少但互动质量高、各有个性;AI 之后帖子激增但大多是"15 分钟做了个 SaaS"的无批判性思维内容
As the title says
受众观点:①@avdept 关注 AI 前社区互动更真实,项目更有辨识度,用户真的会去试用 ②@tererepon 关注核心问题不是 AI 工具本身,而是缺乏批判性思维——做了就期待"somehow magically become rich" ③@com2ghz 描绘了从 website bro → app bro → prompt bro 的历史演变弧线,每个时代都有"捷径暴富"幻觉
展开评论
- @avdept (59): Still a lot of slop, but much less than today, and a good amount of interesting ideas and products(in fact I still use few from pre-ai era which I found here) And even if it was slop - it was unique in some way. Ppl tried to do some nice UI(even though if its used some UI libs l…
- @tererepon (18): Using AI isn't necessarily the problem here, what bothers me is the total lack of critical thinking. It is just creating sht for sake of it in the hope someone buy a dumb subscription and they somehow magically become rich
- @Wurldpies (10): And also mainly focused on software? Now I almost only see apps and websites, but a sideproject could also be physical
- @Hot_Extension_460 (6): Yes it always has been focused on software. I remember it because I was also confused at the beginning, the title / description of the subreddit is more generic than software projects.
- @com2ghz (6): It actually is because of the quantity of crap we get in a short time. In the "past" we had these website bro's that tried to make the new Ebay or webshops. And I m not even talking about dropshippers. But getting rich by selling products to people. Then we had the app bro's. Ho…
开发者用 9 周时间将自建健身追踪开源项目 openGym 做到 900 GitHub star、30 贡献者、1200 人 Discord 社区,分享了开源项目从零到有的关键冷启动经验
openGym is an open-source (AGPL) gym & body-weight tracker you run yourself: Docker on your box, or a standalone Android APK with no account at all. I started it in July because every workout app…
受众观点:①@Key_Association_8707 关注 9 周 900 star 的增长速度验证了这套方法的实际有效性 ②@brkaydev 分享了 iOS 上架的真实踩坑:reviewer 需要能直接用 demo instance,自托管入口在审核时会被判定为"看起来不完整",两次被拒都不是代码问题 ③cluster_size 7(跨 twitter + reddit 多个社区传播)说明话题有真实跨平台热度
展开评论
- @Key_Association_8707 (6): 900 stars in 9 weeks on a self-hosted workout tracker is wild, congrats -downloading the APK now
- @0_KermitTheFrog_0 (2): thanks! Hope it to grow bigger :)
- @Puzzleheaded-Neck731 (1): Word up
- @0_KermitTheFrog_0 (1): Yes apps like hevy should work
- @brkaydev (1): On the native iOS build, the thing most likely to cost you a review cycle is that the reviewer has to be able to use the app without standing up a server. A self-hosted app whose first screen asks for a host address reads as incomplete under 2.1, so put a reachable demo instance…
Claude 新版聊天界面更新后移除了对话分支功能,用户强烈反弹认为这是产品体验倒退,并与 ChatGPT/Gemini 分支设计做对比,部分重度用户表达了切换平台的意向
One of the basic features of an AI chatbot that's been available for years is not available inside a Claude update that's rolling out. Pretty saddening, Claude's branching was one of its best quirks,…
受众观点:①@urchir 对 Anthropic 持续功能倒退表达强烈不满,代表相当一部分重度用户情绪 ②@Great-Worry-6028 关注分支功能在测试不同 AI 反应场景中的具体用途 ③@2SP00KY4ME 明确表示功能倒退已到达忍耐极限,正在重新评估是否继续使用 Claude
展开评论
- @urchir (11): I hate them so much They just keep making everything worse
- @Great-Worry-6028 (10): branching was perfect for testing how a companion would react in different scenarios without losing the flow, hope they fix that soon.
- @idolognium (5): They hid the thinking, and now this. Wonder what's going to be next
- @ZedKGamingHUN (5): It does still exist, it just does it the Gemini way where the previous chat response before the edit is wiped and only the new response with the edit is available, with no way of going back to the previous version.
- @2SP00KY4ME (4): I stayed when the usage limits went to shit, I stayed when they took away thinking. But if they take away the basic ability for me to re-run messages, I'm definitely done. There's literally nothing left to Claude that I can't get from ChatGPT. I always knew Claude would pointles…
独立开发者在冷启动期面临的国际支付时机抉择:何时应该支持多货币和跨境收款,过早建设是否制造不必要的复杂度,如何从第一个外国用户身上学到真实需求
Everyone says software is global until someone from another country actually tries to pay you lol Then suddenly youre thinking about currencies, local payment methods, failed cards and if your produc…
受众观点:①@No-Delivery6809 关注 payouts(向用户付款)比 checkout(收款)复杂得多,建议早期直接选 Whop 这类一体化方案 ②@Individual_Big873 关注"让第一个真实外国用户告诉你什么坏了"的 lean 策略,而不是事先猜问题 ③@Any_Welder_9701 关注记录失败结账尝试作为需求信号的实操价值
展开评论
- @No-Delivery6809 (13): If payouts are part of the product id probably compare something like Whop early rather than adding that whole layer later. If its only customer checkout then you can keep it pretty basic for a while
- @Exact-Ad-8325 (3): I think it’s optional, especially if you’re using a third-party payment provider. The implementation cost is pretty low these days, especially with AI helping with a lot of the integration work, so you can always add it later once there’s actual demand.
- @Individual_Big873 (3): Honestly, don't build the payouts and multi-currency layer until a real buyer from outside forces the question. The moment one foreign customer actually tries to pay you is when you'll learn what breaks - that's cheaper than guessing. Ship checkout for the market you have now an…
- @Any_Welder_9701 (2): failed international checkout attempts are still worth logging, otherwise the demand just disappears without telling you anything
- @Icy_Organization3165 (2): This makes sense tbh. Probably better to let the first few international users show me what actually breaks instead of guessing at problems that might never matter
Senior iOS 工程师以技术负责人身份主导架构设计、用 Claude 逐步实现特性并严格 code review,构建出功能与 web 端 1:1 对齐的 NextExplorer 原生 iOS 伴侣应用并开源
https://preview.redd.it/9pr3ayhj4yph1.png?width=1026&format=png&auto=webp&s=87e537301a263f0e8958f361a77ef7456e373c21 NextExplorer for iOS is the native iOS companion for NextExplorer. It…
受众观点:①@iamdadmin 强烈关注 AI vibe-coding 滥用:有人 AI 生成代码后声称版权并商业化,这个作者明确相反 ②@inverse-talcum-0e 关注多服务器支持等功能完整性需求 ③整体读者关注零遥测 + 100% 开源的隐私立场及 SSO 修复贡献到上游的开源协作模式
展开评论
- @inverse-talcum-0e (2): I’m glad you reposted this! I started using NextExplorer after FileBrowser’s developer stepped away. Thanks for creating an iOS app for it. The only thing that could improve it is adding support for multiple servers for people who have several instances of NextExplorer on their…
- @iamdadmin (2): Awesome, there's a horrible trend of people coding apps with AI or outright vibecoding them, think they can claim proprietary ownership of code they didn't write, and then subsequently monetise it. Glad to see this isn't one of those.
- @iamdadmin (2): Wouldn’t ever object to a tip jar personally!
- @Forward_Matter2861 (2): Thanks for the effort!
- @asimovs-auditor (1): Expand the replies to this comment to learn how AI was used in this post/project.
介绍 HTML 流式渲染用于构建交互式 Web 应用的文章,评论区引发 HTMX v4 与 Datastar 框架孰优孰劣的激烈争论
r/webdev An Introduction to HTML Streaming for Interactive Web Applications
受众观点:①@krileon 指出 HTMX v4 已生产就绪,文章用旧版对比是在误导读者 ②@FlamedDogo99 为 HTMX 社区站台 ③@dmezzo 为 Datastar 的 CSP 安全模式辩护
展开评论
- @krileon (22): The HTMX comparison is out of date. HTMX v4 is production ready and basically blows away anything Datastar can offer. v4 even has inline nonce support for strong CSP with safe inline execution. >Problem #1: Client-side variables (or lack thereof) HTMX has this via hx-live. &g…
- @Ok-Bar-95 (7): viperlith sounds like a power metal band. congrats you reinvented php and gave it a cooler name
- @FlamedDogo99 (7): You should take a look at HTMX v4 before besmirching the name of our lord and savior Carson Gross
- @dmezzo (5): \> v4 even has nonce support for strong CSP with safe inline execution. Datastar supports CSP mode: [https://data-star.dev/reference/security#csp-mode](https://data-star.dev/reference/security#csp-mode)
- @krileon (5): That's certainly new. Thanks for letting me know!
独立开发者上线 VibeMatcher 选片应用,根据情绪和可用时间只返回三条推荐,后端调用 Gemini API,当前响应速度慢是主要用户痛点,开发者计划下一阶段优化性能
I built VibeMatcher for those evenings when you want to watch something but keep rejecting every option. You choose: • Movie or series • Your mood—chill, intense, fun, romantic, thoughtful, or dark •…
受众观点:①推荐结果加载太慢影响体验(@Calibaba2796 反馈) ②Gemini API 延迟是根本原因(@Comfortable_Cup2025 作者确认,并说下阶段攻性能)
展开评论
- @Calibaba2796 (3): Is there a reason it take that long to process and find a movie?
- @Comfortable_Cup2025 (1): It's because of Gemini api behind it, next phase will be working on performance
讨论 AI agent 生成代码导致 PR diff 变大、人工审查效率崩溃,以及 CI 测试体系在 agent 时代暴露的绿色不代表通过的隐性失效问题
<disclaimer maybe some ppl find long post but just trying to understand pain is real or not …> I have been talking a lot lately from devs and eng managers: review queues are backed up, the diff…
受众观点:①@Remarkable-Impact378 直接指出 human review on agent diffs is mostly vibes,真正有效的是 preview 环境直接验证行为 ②@Which-Examination-74 发现 turbo test 有 14 个 package 根本没在运行测试,CI 78 天显示绿色但从未跑过那 30 个测试 ③整体讨论聚焦 behavior testing vs syntax testing 的根本区别
展开评论
- @Which-Examination-74 (2): Not fully. On a client monorepo our team builds, CI runs `bun run test` (which is just `turbo test`) next to lint, typecheck and build, and it stays green while some packages run no tests at all. What showed it was `turbo run test --dry=json`: it lists 28 test tasks, and 14 of t…
- @Remarkable-Impact378 (2): Human review on agent diffs is mostly vibes at this point. What actually catches stuff is running the branch in a preview environment before anyone reads the diff. If the agent rewrote a webhook handler, you call the webhook and watch what comes back. Tests catch syntax and obvi…
- @Common_Dream9420 (1): backfilling is the right call, at least then green actually means something. are you prioritizing by package risk or just working through them in order?
- @Which-Examination-74 (1): we haven't actually backfilled anything, nothing has changed. we did count test files in those 14 packages though, and 12 have none, so there's nothing in them to run. the only one with unit tests but no script is that next.js app, and what it's missing is a "test" line in its p…
- @Common_Dream9420 (1): that split makes sense, at least once the script exists you can actually see what's missing vs guessing. do you find the second PR ever gets deprioritized once the first one merges?
独立开发者构建 TapFlow——浏览器远程控制 iOS/Android 模拟器的自托管 QA 工具,内置可选 MCP 服务器,开发使用 Claude Code 辅助,MIT 开源 v0.x,跨 Twitter+Reddit 七源聚合热度。
I built tapflow for teams that need to check mobile builds but do not want every tester to install Xcode or Android Studio. The topology is worth stating first: the relay can run on Linux or Docker,…
受众观点:①@thatisagoodrock 从写作风格识别出 Claude Opus 痕迹,引发 AI 辅助写作透明度讨论 ②@hometechgeek 对自托管 QA 方案本身感兴趣 ③@Hour-Swimmer7140 对整体方案感兴趣正在评估可行性
展开评论
- @thatisagoodrock (7): Lol either people are doing their best to make their posts look less like AI or AI is starting to dramatically change the way people write. I know Claude Opus 5 when I see it. And I have found it 😂
- @hometechgeek (2): Looks like an interesting idea, thanks for sharing
- @thatisagoodrock (2): Yeah, that’s my point. That’s why his comment sounds like a psychotic robot.
- @Hour-Swimmer7140 (2): That's interesting!
- @asimovs-auditor (1): Expand the replies to this comment to learn how AI was used in this post/project.
讨论独立 SaaS 创业者在资源极度受限时如何排序四项核心优势,多位创业者分享了 distribution first、runway first 的真实失败教训
Imagine you are starting a SaaS product with enough resources to secure only four major advantages. Would you prioritize customer access, runway, distribution, product speed, credibility, or uninterr…
受众观点:①@francksiduo 强调 distribution 优先,「六个月打磨产品后发现没有可复用的获客渠道」是最常见失败模式,共鸣度高 ②@tuomas-heino 把 runway 和 customer access 并列第一,认为 focus 和 product speed 是「让你有生产力感但加速烧钱的方式」 ③@mrkeyoor 分享最详尽的个人案例:一年发布后 zero owned audience,是最有说服力的反面教材
展开评论
- @francksiduo (3): Distribution over everything else, especially early. Credibility and product speed tend to sort themselves out once you have a channel that keeps working without you pushing it every day. Runway just buys time to find that channel, it doesn't replace it. The failure mode I see t…
- @tuomas-heino (2): runway first, then customer access. everything else is downstream of those two. if you cant pay yourself for 12 months and you dont have a warm path to 50 people with the problem, product speed and focus are just ways to burn the clock faster. the failure mode i see most is foun…
- @mrkeyoor (2): 4, 8, 3, 5. the failure mode i'm avoiding is the one i actually hit: building a lot of something before knowing who was already looking for it. 4, the 500 interview notes, goes first because every other slot gets spent badly without it. the notes are the words. the landing page,…
- @Aromatic-Ad-5999 (1): great explanation!
- @mmk_software (1): Some of these things aren't as useful anymore. Ui kit, no meeting pass, doesn't really make sense in this context. If you had even one or a couple of these, it would be enough of an edge to start something. More have been started with less. I guess 2,3,4,8 and work on getting fu…
Anthropic 宣布推出 Claude Slides 功能(设计与文档合一),将与 chat/cowork 功能同步上线,社区讨论此功能是否是对 Instinct/Muse 等 AI agent 演示产品的竞争应对
Don't see it live yet, I wonder if it will coincide with the chat/cowork union also announced today. Wonder if it works in Claude Code on the desktop app as well.
受众观点:①@Ener_Ji 分析了 Claude Slides 可能是对 Instinct/Muse/Grokbot/Gemini Spark 等 AI agent 热度的战略应对 ②@dressinbrass 实际体验了 design review 功能并给出正面反馈 ③@Special-Bite 预测 AI 将彻底淘汰传统演示文稿形式
展开评论
- @Ener_Ji (9): Rolling out over "weeks" according to the announcement. I wonder if they pre-announced this as a response to the hype that AI agents like Instinct and Muse (and to a lesser extent Town, Grokbot, and Gemini Spark) are generating?
- @dressinbrass (4): Just saw the design stuff when I did a design review for a team. It’s super slick.
- @DocTrey (3): We can only hope. If they had a brand guideline you could use or a template, it would be awesome.
- @Special-Bite (2): I have a feeling that AI is going to obsolete the slide deck.
- @PossibleHero (2): No. Not at all.
对 ChatGPT 网页端前端架构进行逆向工程分析,揭示流式响应延迟优化实现、聊天列表未做虚拟化的工程选择及 retry 逻辑设计,引发开发者对大型 AI 产品工程决策的讨论
r/webdev Reverse Engineering ChatGPT Web: How OpenAI Built for a Billion Users
受众观点:①@rachelduno 感慨 retry 逻辑复杂度超预期,自己复刻时遇到 WebSocket 重连 loop 问题 ②@medy17 对未做虚拟化的选择感到困惑 ③@Ornery-Concentrate-5 猜测内部测试账号对话都很短所以没发现该问题
展开评论
- @rachelduno (6): kinda wild how much engineering goes into just making the response stream feel instant, the retry logic stuff is what got me
- @medy17 (5): I noticed the terrible lag and knew they didn't virtualise. I'm still struggling to understand what they were thinking! Great article btw :) Nicely done
- @rachelduno (5): tried to replicate their streaming setup for a small side project after reading this, ended up with my websocket reconnecting loop kicking off every 20 seconds lol. still dont fully get how they handle backpressure at that scale, did they mention anything about that in the deep…
- @webdev-ModTeam (2): Your post/comment has been determined to be a low-effort post or comment. This includes title-only posts, easily searchable questions, vague/open-ended discussion prompts, LLM generated posts or comments, and posts/comments that do not provide enough context for meaningful repli…
- @Ornery-Concentrate-5 (-1): honestly surprised too, not virtualizing a chat thread feels like the kind of thing that only bites you after months of dogfooding with short conversations. makes me wonder if their internal test accounts just don't have 200-message threads lying around
研究者探讨多个独立 LLM 在模糊任务说明下以相同方式失败的规律,追问这种一致性失败是平滑递增还是存在急剧跃升的模糊度临界点
When models from different families are given the same underspecified task, they often fail in the same way rather than in independent ways. My question is about measurement, not explanation. Has any…
受众观点:①@Sad-Rip7344 (3) 关注临界点实际意义:如果是 cliff 则 benchmark 数字会随机跳动无法用于调优 ②@zyl1024 (2) 分享了 shared imagination 论文展示 LLM 面对虚构问题给出相同答案的现象 ③@breadstickdingdong (1) 追问 graded ambiguity 的测量方式希望在完全明确和完全虚构之间找到中间梯度
展开评论
- @Sad-Rip7344 (3): The threshold question is the interesting part to me. If it's a smooth curve you can probably tune your eval harness around it, but if there's a cliff then any benchmark built near the edge is going to have unstable numbers run to run.
- @zyl1024 (2): I had a past paper on "shared imagination", demonstrating that when LLMs face entirely fictional or nonsensical questions, they often give the same answer. [https://arxiv.org/abs/2407.16604](https://arxiv.org/abs/2407.16604)
- @breadstickdingdong (1): Yeah this is the part that'd make it matter to anyone but me. If it's a cliff, any benchmark near the edge gives different numbers every time and you can't tell it apart from model variance. you'd just be measuring where you happened to land. It would be weird if nobody's measur…
- @breadstickdingdong (1): Ty! This is pretty much what I was looking for. I was thinking about the family gradient and I’ve been stuck on it. Do you know of anything that graded the ambiguity instead of treating fictionality as binary? Their setup only has the one condition so there's no curve to look at…
- @zyl1024 (1): Not really on graded ambiguity. Btw, I am the first author of the paper, so feel free to DM or email me if you are working on this as a research project and want to chat.
独立开发者用越狱 Kindle 和树莓派搭建低成本铁人三项训练仪表盘,连接 Garmin 运动数据并每 30 分钟刷新展示,借助 Claude 大幅加速开发,计划开源
Been ramping up my training for an Ironman 70.3 and i wanted a dashboard to track my progress and upcoming workouts. Was inspired by strmnl but wanted to do it for cheap. Grabbed my pi zero 2 w, setu…
受众观点:①新款 Kindle 是否同样支持此类改造(@Ok_Studio_834 询问) ②越狱的机型和版本兼容性问题(@pf1993 补充第七代 Paperwhite 方案) ③电子墨水屏刷新率限制对实时数据展示体验的影响(@aayush_aryan 提问,@pf1993 解释权衡)
展开评论
- @Ok_Studio_834 (1): Damn, this looks incredible!! Just wondering, I just got a brand new Kindle, is it possible to vibe-code stuff for it too or does it strictly have to be an older/jailbroken model ? If it works on the new ones, you just opened up some crazy doors for me ahahah
- @pf1993 (1): I think you can do it on any kindle given it is jailbroken. Mine is a 7th gen paperwhite and there are some well-established jailbreaks for the model.
- @aayush_aryan (1): Sorry, stupid question but Kindle doesn't support fast refresh?
- @pf1993 (1): Not exactly fast refreshes like LCDs but you can configure partial screen refreshes or ‘fast’ mode for scrolling etc. i went with a full refresh to clear ghosting completely but that takes a sec. This is a 2015 7th gen paperwhite, so definitely much slower than modern versions
Relaticle 是 AGPL-3.0 自托管 CRM,内置 39 工具 MCP 服务器和 Ollama 本地推理,双轨权限模型(assistant 审批 vs 外部 MCP client 直写),同 cluster c_0085 与 TapFlow 构成 MCP 工具设计热点。
Relaticle runs on your own server, and Ollama can keep model inference there too. I'm Manuk, the maintainer. It's an AGPL-3.0 CRM for companies, people, opportunities, tasks, and notes. The same inst…
受众观点:①@BP041 提到正在用 Claude Code + OpenClaw 连接自制 CRM,关注 39 工具 MCP 的 prompt stuffing 节省效果和 Ollama 在 Mac Studio 高负载下推理速度骤降问题 ②@Local-Comparison-One(作者)坦承目前无 Ollama 负载基准数据 ③@asimovs-auditor 关注 AI 使用透明度
展开评论
- @BP041 (2): 39-tool MCP server is the part that actually catches my eye — I've been wiring Claude Code into a janky homegrown CRM through OpenClaw and that many tools would save a lot of prompt stuffing. How's the tool call latency with Ollama under load? On my Mac Studio the inference spee…
- @asimovs-auditor (1): Expand the replies to this comment to learn how AI was used in this post/project.
- @Local-Comparison-One (1): I don't have reliable Ollama load benchmarks to share yet. The MCP server handles the CRM operations; inference stays with your client's model. Which model and quantization are you running on the Studio?
开发者开源了 Bough 工具,读取本地 Claude Code 使用历史并生成交互式可视化视图——日历热图展示工作量、自动推断任务边界、点击查看每条 prompt 原文,完全本地运行不上传数据
I use Claude Code a lot, but /stats never answered the question I actually cared about **What did I build, and where did the work get difficult?** Repo: [https://github.com/nickelsec/bough](https://g…
受众观点:①@Rare_Guide_9830 也在独立开发类似 agent 使用记录工具,说明需求不止一个人在追 ②@VoidEqualZero 和 @shipsonfriday 表示兴趣但无实质反馈,评论深度偏浅留有切入空间
展开评论
- @Rare_Guide_9830 (2): Cool, checking it out! I’ve been trying to tap into the same thing but more for exploring usage and memory. But haven’t quite made it super useful yet. https://github.com/ethanplusai/agent-ledger
- @VoidEqualZero (1): Cool will check it out!
- @shipsonfriday (1): Looks really cool man!!
Creem 2.0 在创始周年纪念日正式发布,定位 AI 时代商业基础设施,提供全球支付、税务合规、联盟营销、Usage Billing 一体化解决方案,覆盖 MCP 和 CLI 集成
Creem
受众观点:① MoR 税务合规是否真正解决了全球销售的 VAT 痛点 @SpatialChat;② MCP 和 CLI 集成如何让 AI agent 直接调用支付功能 @Creem;③ Usage Billing 对 AI 时代按用量收费产品的重要性 @Overview
展开评论
- @Overview (0): * [Launches3](/products/creem#launches) * [Reviews5](/products/creem/reviews) * [Alternatives](/products/creem/alternatives) * [Customers](/products/creem/customers) * [Built with](/products/creem/built-with) * [Team](/products/creem/makers) * More This is the 3rd launch from Cr…
- @Creem (0): Maker 📌 Hey Product Hunt 👋 I'm Gabriel, CEO and co-founder of Creem. We picked today on purpose. September 17 is Creem's founding anniversary, so shipping Creem 2.0 on the same day felt right. The thesis in three lines: 1. Building a product went from months to a weekend. Gettin…
- @SpatialChat (0): [@sudoferraz](https://www.producthunt.com/@sudoferraz) I've been [Creem.io](https://Creem.io) user for several months now, I confirm that platform is awesome, payments run smooth and tax/VAT is solved for me - less stress. And the founder (Gab) responds quickly! Congrats and kee…
- @Creem (0): Maker
- @Creem (0): Maker [@csaba\_kissi](https://www.producthunt.com/@csaba%5Fkissi) Hey Csaba, happy to see you around here! Thanks for supporting us during our launch, and I'm super proud that you are a CREEM customer already! Yess, we do have a fully covered MCP and CLI, so it's really up to yo…
Bitrise 发布 Remote Dev Environments,提供秒级启动的云端 Mac 和 Linux 机器,专为 AI coding agent 设计,与 CI 环境保持完全一致的 stack 和缓存,解决 agent 本地资源争用和无状态容器反复初始化的问题
Bitrise
受众观点:① 云端 Mac 环境 vs 无状态容器对 agent 运行效率的实际差距 @Bitrise;② stack 与 CI 环境保持一致对 CI/CD 可靠性的价值 @Bitrise;③ RDE 的访问权限和 GitHub 账号集成方式 @Bitrise
展开评论
- @Overview (0): * [Launches3](/products/bitrise#launches) * [Reviews1](/products/bitrise/reviews) * [Alternatives](/products/bitrise/alternatives) * [Built with](/products/bitrise/built-with) * [Team](/products/bitrise/makers) * [Awards](/products/bitrise/awards) * More This is the 3rd launch f…
- @Bitrise (0): Maker 📌 Hey Product Hunt, I’m Arpad, VP of Engineering at Bitrise (mobile DevOps startup). **Agents need somewhere real to run.** You can **hand an agent a task**, but it has to **execute somewhere**. ❌ On your laptop it fights you for the CPU and the simulator, running multiple…
- @Bitrise (0): Maker [@ilana\_zholobovsky1](https://www.producthunt.com/@ilana%5Fzholobovsky1) hello Ilana 👋 Thank you so much for your support :) Upvote Report Share 6h ago [](/@lisadziuba) [Lisa Dziuba](/@lisadziuba)
- @Bitrise (0): Maker [@ilana\_zholobovsky1](https://www.producthunt.com/@ilana%5Fzholobovsky1) Stacks carry over directly, an RDE session runs the same stack image as your CI and Build Hub builds, and stays pinned to it. Upvote Report Share 6h ago [](/@viktorbenei) [Viktor Benei](/@viktorbenei)
- @Bitrise (0): [@ilana\_zholobovsky1](https://www.producthunt.com/@ilana%5Fzholobovsky1) Listing the answers: * Stacks: you can use the same stacks, with the same pre-installed tools. * Connected repo: RDEs today are personal, meaning if you create an RDE others won't have access to it. Curren…
Modaal 扩展支持 Android 平台,基于创始人 18 年移动开发经验打造的 AI agent 辅助跨平台框架,声称开发 iOS+Android 双平台成本只比单平台高 60%,支持 Cursor 集成
Modaal
受众观点:① 双平台开发成本是否真的只比单平台高 60% @Modaal;② Cursor 集成是否让 AI 辅助移动开发更流畅 @Vikram;③ AI agent 能否真正替代专业移动开发框架经验积累 @Modaal
展开评论
- @Overview (0): * [Launches2](/products/modaal-ai-agent-for-ios-android-native#launches) * [Reviews2](/products/modaal-ai-agent-for-ios-android-native/reviews) * [Alternatives](/products/modaal-ai-agent-for-ios-android-native/alternatives) * [Customers](/products/modaal-ai-agent-for-ios-android…
- @Modaal (0): Maker 📌 **Hey ProductHunt!** Ivan here, co-founder of Modaal. I founded my first mobile app development company 18 years ago and released hundreds of mobile apps to clients, and as my own products. Over the years I built myself a repeatable, robust framework for building mobile…
- @Vikram (0): 💡 Bright idea [@ivan\_misuno](https://www.producthunt.com/@ivan%5Fmisuno) Being able to plug Cursor right into this workflow makes a lot of sense for how I already work, congrats for launch🙌 Upvote (3) Report Share 11h ago [](/@ivan%5Fmisuno) [Ivan Misuno](/@ivan%5Fmisuno)
- @Modaal (0): Maker
- @Modaal (0): Maker [@vipul\_kumar1280](https://www.producthunt.com/@vipul%5Fkumar1280) Thanks for the question! I'm in the middle of actually gathering detailed statistics using different frameworks and architectures. Preliminary, using Modaal (and the underlying Duet framework - <https://do…
Text Agent Store 是首个 iMessage 短信 AI agent 聚合市场,3 个月内收录 137 个不同类型 texting agent,marketplace 本身也可通过短信方式浏览,提供 native iMessage 键盘扩展
Text Agent Store
受众观点:① iMessage 作为 agent 发现和使用渠道的用户习惯可行性 @Soar;② 3 个月 137 个 agent 的增长速度和生态规模 @Soar;③ texting agent 的商业模式——免费还是付费 @Soar
展开评论
- @Overview (0): * [Reviews](/products/text-agent-store/reviews) * [Alternatives](/products/text-agent-store/alternatives) * [Team](/products/text-agent-store/makers) * [Awards](/products/text-agent-store/awards) * More Free Launch tags:[Messaging](/topics/messaging)•[Artificial Intelligence](/t…
- @Soar (0): Maker 📌 Hey Product Hunt! 👋 **The Problem** There's a wave of texting agents coming, each with its own use case, personality, and charm. But how do you actually find them? Bookmarking them on X the one time you spot them in your feed? **Not anymore.** **Introducing Agent Store**…
- @Soar (0): Maker [@priya\_kushwaha1](https://www.producthunt.com/@priya%5Fkushwaha1) thanks priya! the app store evolution 🚀 Upvote (2) Report Share 11h ago [](/@elias%5Fleo1) [Elias Leo](/@elias%5Fleo1) congrats on the launch [@stannno](https://www.producthunt.com/@stannno) my question is…
- @Soar (0): Maker [@elias\_leo1](https://www.producthunt.com/@elias%5Fleo1) 139 as of today! no limit, we expect it to grow to millions of agents over the next decade Upvote Report Share 2h ago [](/@jacob%5Fhernandez4) [Jacob Hernandez](/@jacob%5Fhernandez4) congratulation to the team Text…
- @Soar (0): Maker [@jacob\_hernandez4](https://www.producthunt.com/@jacob%5Fhernandez4) the entire marketplace can be navigated over text! makes it super simple to explore. also a native imessage app to keep locally in your keyboard is helpful. Upvote Report Share 2h ago [](/@james%5F%5Fdar…
NovaSynth by Noveum 是面向 AI 语音 agent 的自动化测试评估工具,通过模拟真实用户对话场景测试 voice agent 表现,同时支持 chatbot 和 HTTP endpoint 的通用 eval,填补 AI 多模态测试领域的空白
NovaSynth by Noveum
受众观点:① voice agent 测试能否有效泛化到 chat agent 场景 @Clueso;② AI eval 工具跨模态扩展的产品路径 @Clueso;③ 如何客观量化 agent 对话质量的核心指标 @EaseOps
展开评论
- @Overview (0): * [Reviews](/products/novasynth-by-noveum/reviews) * [Alternatives](/products/novasynth-by-noveum/alternatives) * [Built with](/products/novasynth-by-noveum/built-with) * [Team](/products/novasynth-by-noveum/makers) * [Awards](/products/novasynth-by-noveum/awards) * More Free Op…
- @Transcription (0): [Tough Tongue AI ](/products/tough-tongue-ai-2?ref=product%5Fsidebar) AI teammate for high-stakes conversations [5.0(24 reviews)](/products/tough-tongue-ai-2/reviews) [AI Voice Agents](/categories/ai-voice-agents)[AI Characters](/categories/ai-characters) [Hamming AI (YC S24)](/…
- @Clueso (0): Anything that makes AI evals easier is a big blank space across modalities. This is a fantastic space to be building in - kudos to you folks! Are you thinking of sticking to voice agents, or would you be doing persona testing on chat based apps at some point too? Upvote (3) Repo…
- @EaseOps (0): Congrats on this launch builders , what's one distinguish thing about noveum that you are proud of building ?! Upvote (1) Report Share 3h ago [](/@itsshashank) [Shashank Agarwal](/@itsshashank) [NovaSynth by Noveum](/products/novasynth-by-noveum) Maker