Anthropic 正式发布 Claude Opus 5,号称以接近 Fable 5 前沿智能但仅需一半价格的新旗舰模型
Introducing Claude Opus 5. It's a thoughtful and proactive model that comes close to the frontier intelligence of Fable 5 at half the price. https://t.co/GQWhcq2CQL ①@Skiipy88、@samell 等多个用户核心关注 usage reset 问题,说明现有 Max 用户急于体验但被额度限制;②@Nova_Nightfall 质疑 benchmark 数字(53.4% > 53.5% 的展示方式),反映社区对 eval 可信度的普遍怀疑;③@samell 调侃明天会出现"99% 的人都在错用 Opus 5"文章,暗示社区预期内容热潮即将到来
Anthropic 正式发布 Claude Opus 5,号称以接近 Fable 5 前沿智能但仅需一半价格的新旗舰模型
①@Skiipy88、@samell 等多个用户核心关注 usage reset 问题,说明现有 Max 用户急于体验但被额度限制;②@Nova_Nightfall 质疑 benchmark 数字(53.4% > 53.5% 的展示方式),反映社区对 eval 可信度的普遍怀疑;③@samell 调侃明天会出现"99% 的人都在错用 Opus 5"文章,暗示社区预期内容热潮即将到来
Sam Altman 发推称希望美国在开源和闭源 AI 两个方向都能领先,但立刻遭到大量用户嘲讽 OpenAI 言行不一
①@ns123abc 直接质疑 sama 说的和实际游说行为相反,评论区整体情绪偏向怀疑;②@nisxant 呼吁 OpenAI 应该主动发起或签署支持开放 AI 的行动而不只是表态;③@yv_thorne 要求发布已停用模型的权重,代表开发者社区对 OpenAI 历史模型封闭的积怨
Anthropic 发布 Claude Opus 5,定位为接近 Fable 5 前沿智能但价格减半的高性价比推理模型
①@briandoll 关注性能与成本的平衡,受众最在意 Opus 5 的性价比定位是否言之有据;②@yusufozkan 对 ARC-AGI-3 30.2% 分数印象深刻,受众高度关注基准测试数字来判断实际能力;③@datakan 质疑 Opus 5 能力不超 Fable 5 的情况下存在意义是什么,受众在评估产品定位的合理性
Claude Opus 5 今日上线全部付费计划及 API,定价维持 Opus 4.8 水平不变,成为 Claude Max 默认模型,社区对使用量 reset 的呼声压倒一切其他反应
Claude Opus 5 今日上线全部付费计划及 API,定价维持 Opus 4.8 水平不变,成为 Claude Max 默认模型,社区对使用量 reset 的呼声压倒一切其他反应
受众观点:@FireIre (score 99, Reddit ClaudeAI) 说「So this is just basically Fable 5 at Opus prices? Awesome.」——代表了务实开发者对性价比提升的直接认可;@tokenentropy (score 59, Reddit ClaudeAI) 说「This thing beats Sol by a LOT. holy SHIT」——说明与竞品(Sol/Grok)相比的体验差距在用户感知上非常显著
Anthropic 发布 Claude Opus 5,在多项编程和知识工作基准测试上超越 Fable 5,同时价格仅为 Fable 5 的一半,研究圈和开发者社区反应热烈但对部分 benchmark 数据(53.4% > 53.5%)持怀疑态度
Anthropic 发布 Claude Opus 5,在多项编程和知识工作基准测试上超越 Fable 5,同时价格仅为 Fable 5 的一半,研究圈和开发者社区反应热烈但对部分 benchmark 数据(53.4% > 53.5%)持怀疑态度
受众观点:@samell (score 910, Twitter IndieDev Leaders) 调侃「'99% of people are using Opus 5 wrong' 的文章明天就会出来」,折射出开发者对 AI 公司每次发布后随之而来的内容营销潮的疲惫感;@Freedomsaver (score 308, Reddit ClaudeAI) 直接指出 benchmark 图表显示「53.4% > 53.5%(Agentic coding)」的数学错误,嘲讽 Anthropic 数据展示方式,该评论获得大量 upvote,说明社区对厂商自我吹捧式 benchmark 存在普遍信任危机
独立开发者在为低流量网站选择自托管分析工具时,Matomo、Umami、Plausible、GoatCounter 各有不同的资源消耗和数据准确性权衡,多个平台同时存在数据差异是普遍现象
独立开发者在为低流量网站选择自托管分析工具时,Matomo、Umami、Plausible、GoatCounter 各有不同的资源消耗和数据准确性权衡,多个平台同时存在数据差异是普遍现象。
受众观点:@TheEfficaciousTodayS(score 4,r/webdev)指出 Matomo 在 1GB VPS 上可运行,但数据库会随时间膨胀,必须设置 auto-purge raw logs,否则会「wreck a small VPS」——这是实际踩坑经验。@magenta_placenta(score 2,r/webdev)解释不同工具本质上在测量不同层级的请求:广告网络触发的是 ad payload 请求,本地 MySQL 计数几乎包含一切,GA 依赖 JS 执行,Cloudflare 在网络层计数——因此根本不存在「最准确的数字」,只有不同定义下的不同数字。
Claude Opus 5 是 Anthropic 迄今为止最难被提示注入攻击的模型,在完整防御栈下注入成功率接近零,Anthropic 员工认为这是 AI agent 从 demo 走向生产部署的关键障碍解除
Claude Opus 5 是 Anthropic 迄今为止最难被提示注入攻击的模型,在完整防御栈下注入成功率接近零,Anthropic 员工认为这是 AI agent 从 demo 走向生产部署的关键障碍解除
受众观点:@MTorygreen (score 3, Twitter IndieDev Leaders) 指出「agents got stuck in demos for one reason, and it wasn't capability. You couldn't trust them near hostile input and now getting that rate near 0 is what actually clears them for production」——直接点出了 agent 产品化的核心痛点;@gabefletcher (score 0, Twitter IndieDev Playbooks) 补充「it has them - it just has a lot more common sense and understanding for what its being used for」,说明安全不是靠硬规则而是靠模型理解力,这对开发者设计 agent 系统架构有启示
NeurIPS 2026 E&D track 评审结果集中出炉,研究者们在 Reddit 大规模讨论低评分是否值得 rebuttal、meta review 延迟原因,以及 AC 使用 AI 写 meta review 的现象
NeurIPS 2026 E&D track 评审结果集中出炉,研究者们在 Reddit 大规模讨论低评分是否值得 rebuttal、meta review 延迟原因,以及 AC 使用 AI 写 meta review 的现象
受众观点:@OutsideSimple4854 (score 9, Reddit MachineLearning) 指出「Two out of four ACs for papers I reviewed had meta review looking like it was written by AI」,直接揭示学术 review 流程的 AI 污染问题;@Stormzrift (score 1, Reddit MachineLearning) 说「Our AC was also very AI sounding. Even flagged 100% LLM in a checker」,用工具验证了 AI 生成 meta review 的现象
Anthropic 正式发布 Claude Opus 5,号称以接近 Fable 5 前沿智能但仅需一半价格的新旗舰模型
Introducing Claude Opus 5. It's a thoughtful and proactive model that comes close to the frontier intelligence of Fable 5 at half the price. https://t.co/GQWhcq2CQL
受众观点:①@Skiipy88、@samell 等多个用户核心关注 usage reset 问题,说明现有 Max 用户急于体验但被额度限制;②@Nova_Nightfall 质疑 benchmark 数字(53.4% > 53.5% 的展示方式),反映社区对 eval 可信度的普遍怀疑;③@samell 调侃明天会出现"99% 的人都在错用 Opus 5"文章,暗示社区预期内容热潮即将到来
展开评论
- @claudeai (6736): On several coding and knowledge work evaluations, Opus 5 is the new state-of-the-art: https://t.co/Cl3fDxM0gP
- @samell (910): @claudeai can't wait for the “99% of people are using Opus 5 wrong” articles tomorrow https://t.co/WlHdqCfkiM
- @Skiipy88 (794): @claudeai You know what would be great right about now? A little reset https://t.co/9d2d4xUV77
- @Nova_Nightfall (692): @claudeai ok, so 53.4% > 53.5%? https://t.co/MZFhM7nAzM
- @cafkafk (465): @claudeai Groundbreaking https://t.co/4XJQwV3hjd
Sam Altman 发推称希望美国在开源和闭源 AI 两个方向都能领先,但立刻遭到大量用户嘲讽 OpenAI 言行不一
i want the US to win in AI both in open source and proprietary models, and i am glad to see this
受众观点:①@ns123abc 直接质疑 sama 说的和实际游说行为相反,评论区整体情绪偏向怀疑;②@nisxant 呼吁 OpenAI 应该主动发起或签署支持开放 AI 的行动而不只是表态;③@yv_thorne 要求发布已停用模型的权重,代表开发者社区对 OpenAI 历史模型封闭的积怨
展开评论
- @konstantindeyev (775): @sama “I am glad to see this” https://t.co/t27e6GfLis
- @ns123abc (711): @sama >“i want” then why do your actions reflect lobbying efforts against open weights lmaoooo
- @nisxant (559): @sama openai should stand up for open AI too. https://t.co/an9484vCOq
- @Smallzero (295): @sama You sure? https://t.co/axzXoRGgBe
- @yv_thorne (221): What a fuck have YOU done for open source? Oss-20b in ten years??? Did you sign Jensen’s open letter yet? You are literally AGAINST OPEN AI even though you are called OpenAI. Release the weights of the models you have deprecated! #OpenSource4o #OpenSourceo3 #OpenSource41 #OpenSo…
快速成长创业公司的年轻 PM 不写也不读 PRD,引发关于 AI 时代产品文档价值与代际注意力差距的深度讨论
From a friend who joined a fast-growth, later-stage startup: "I was wondering why no one is doing PRDs here. Then I realized that the Head of Product and most PMs are 25-year-olds who don't have the…
受众观点:①@maxmax 指出最荒谬的是被要求读一个作者自己都没读过的文档,评论区共鸣强烈;②@bitforth 认为 AI 时代 PRD 已经过时,直接 prototype 加 feature flag 测真实用户才是更快路径;③@GergelyOrosz 自己引用 @DynamicWebPaige 的观点,讨论碎片化娱乐消费正在缩短年轻一代的长文本阅读能力
展开评论
- @maxmax (277): @GergelyOrosz The most offensive thing is being asked to read a doc the author never read themselves
- @bitforth (158): PRDs had value when writing code was expensive and mistakes took months to unwind. But that world is effectively gone, no? The fastest path to ground truth has always been to build the smallest version, put it behind a feature flag, expose it to real users, and measure what happ…
- @GergelyOrosz (139): Good point by @DynamicWebPaige how even outside of work / freetime activities are probably just not involving that much reading of longform. A generational gap building up? https://t.co/9YRyzQhfWb
- @clairevo (52): @GergelyOrosz Hey thanks for the product idea https://t.co/JXPFiy1Wqy
- @min_aws (34): @GergelyOrosz As a 25 yo, I am offended I can read atleast 4 bullet points.
Anthropic 工程师披露 Opus 5 是迄今抗 prompt injection 能力最强的模型,这一特性在系统卡中被低调处理但对 AI agent 安全极为关键
Opus 5 is a great model for coding, data analysis, design, biology, knowledge work. More than any of these eval scores, what is most exciting to me is something else: Opus 5 is our least prompt injec…
受众观点:①@elonmusk 简短点评"非常令人印象深刻",带来高曝光;②@jjferman 专门核查 benchmark 的数学是否正确,反映技术社区对 eval 数字的高度审视;③@vedolos 关注 usage reset 问题,代表普通付费用户的优先关切
展开评论
- @elonmusk (629): @bcherny Very impressive
- @jjferman (41): @bcherny let's check the math on this one... https://t.co/Ja9gz4nsRa
- @vedolos (32): @bcherny @bcherny man, can we get usage reset?
- @TristanHurleb (7): @bcherny @elder_plinius it’s your time to shine
- @udiWertheimer (6): @bcherny bro used a -- instead of a — to appear more human
Sam Altman 用玩笑语气问用户对 OpenAI 新产品层级命名 pro-ultra-superhard 有何看法,引发对 AI 产品定价策略的嘲讽讨论
let us know what you think of pro-ultra-superhard
受众观点:①@shaped 直接说无论叫什么名字大家都会叫它"贵的那个",反映开发者对命名游戏的疲惫;②@pedroshakour 反映 Pro 模型的功能限制影响实际工作流(不能连接 repo),说明定价不等于功能完整性;③@Smallzero 借机再次批评 sama 反对开源权重,说明这条推文被当成了发泄渠道
展开评论
- @Smallzero (57): @sama Sam, stop lobbying against open weight and local models. You are advocating against your own people.
- @frostybaby13 (29): @sama I think it's not solving the problem I have, bring back 4o!!! https://t.co/SiA9PT1VUM
- @dharshan_tw (16): @sama https://t.co/Y48JiS2Xdd
- @pedroshakour (10): @sama I've never seen any model like Pro. It is insanely strong! I just never understand why you put it inside chat and not Codex. Also, when you use Pro, the chat section will not be able to connect to any repo. I used to ask Pro to analyze my repo and now it can't.
- @shaped (9): @sama name it whatever you want, we are all just going to call it the expensive one 🙂
Claude Opus 5 今日上线全套付费方案,与 Opus 4.8 同价,提供约 2.5 倍速度的 Fast 模式,成为 Claude Max 默认模型
Opus 5 is available today on all paid plans and the Claude API, priced the same as Opus 4.8. It’s the default model on Claude Max, and the strongest on Claude Pro. It’s also offered in Fast mode, whi…
受众观点:①多位用户(@itsbrandenn1、@Fr0ta、@markksantos)集中要求 usage reset,说明 Claude Max 用户对新模型体验的迫切程度远高于对功能细节的关心;②@elderjava 质疑 Fable 模型的存在意义已经被 Opus 5 取代;③@x_xedo_x 怀疑 Opus 5 只是 Fable 的改名重新包装
展开评论
- @itsbrandenn1 (25): @claudeai Well usage be reset today so I can use it?😅
- @Fr0ta (9): @claudeai Reset our weekly usage so we can test it better
- @markksantos (6): @claudeai Nice! Now we need a reset! https://t.co/AYCqjKAdUd
- @elderjava (5): @claudeai Am I missing something, or is there no reason to keep using Fable?
- @x_xedo_x (5): @claudeai lmao did you just rebrand Fable ?
Anthropic 官方说明 Opus 5 在网络安全任务上超越 Opus 4.8,safeguards 设计允许漏洞识别修复而阻止高风险漏洞利用开发
Opus 5 is stronger than Opus 4.8 on cybersecurity tasks. But it remains substantially behind Mythos 5 at developing exploits. Its safeguards are designed to allow developers to identify and fix softw…
受众观点:①@claudeai 自己转发了 Opus 5 的可用性和定价公告,说明这条推文是发布线程的一部分;②@s41r4j 抱怨即使是简单任务也被安全机制拦截,反映开发者对过度保守 safeguards 的不满;③@tasarren 询问 usage 费率在 August 之前是否继续 2x,说明价格敏感用户在关注成本窗口
展开评论
- @claudeai (1985): Opus 5 is available today on all paid plans and the Claude API, priced the same as Opus 4.8. It’s the default model on Claude Max, and the strongest on Claude Pro. It’s also offered in Fast mode, which runs around 2.5× the default speed. Read more: https://t.co/YMw7gLMT7e
- @shrzawn (13): @claudeai ... https://t.co/0LhuTmelVN
- @s41r4j (2): @claudeai Why do we need to know this!? If we try anything simple also we get blocked, this just feels good to hear but can't use it!
- @Novaautonomous (2): @claudeai What I hope is that it will now fly by things without mistakes, security hah lets leave it for mythos 🙃 https://t.co/8pgkL52cpz
- @tasarren (1): @claudeai Amazing, let's get started working with our new toy! :D Will the x2 rates apply to Opus 5 until August too?
Lmarena宣布Claude Opus 5加入评测平台,将通过数百万真实任务测试验证其是否真正达到Fable 5水准。
Claude Opus 5 has landed in the Arena. The newest model from @AnthropicAI is reported to reach Fable 5 level intelligence across a variety of static benchmarks. Let's see how it performs on real task…
受众观点:①@payalpixel 关注成本性能比(FrontierCode 1.1评测中Opus 5接近Fable级别但仅半价),对开发者API成本决策有直接影响;②@MontondoOrr 关注Opus 5对编程工作流的具体影响;③@stev_builds 关注Opus 5是否曾以代号形式在Arena匿名出现并被测试过
展开评论
- @arena (13): Head over to find it in Agent Mode and Battle Mode to test it out: https://t.co/yZiJuG8ica
- @payalpixel (2): @arena @AnthropicAI On FrontierCode 1.1 - Opus 5 approaches Fable-level performance at half the cost. https://t.co/lKGjGtqPoY
- @MontondoOrr (1): @arena @AnthropicAI My coding workflow needs this new pillar
- @bitsbyritik (1): @arena @AnthropicAI We are eagerly waiting for the results
- @stev_builds (0): @arena @AnthropicAI Was Opus 5 under a "codename" in the Arena already?
AI研究评论者发布Claude Opus 5基准测试截图,含ARC AGI 3得分30%等关键数字,评论区认为其可能是Mythos 5.5的蒸馏版本。
Opus 5 seems pretty good https://t.co/iq3BOHALB3
受众观点:①@dtbake 关注ARC AGI 3具体分数30%在竞品对比中的意义;②@charlesthefool 关注模型血统猜测(Mythos 5.5蒸馏),试图理解Opus 5性能来源;③@robleclerc 关注预期落差——原以为Opus 5远弱于Fable,结果大超预期,赞扬Anthropic的"head fake"策略
展开评论
- @imjustnewatai (7): @scaling01 https://t.co/kCC0cpxNFb
- @tmuxvim (7): @scaling01 https://t.co/BRIXxGIC7y
- @dtbake (6): @scaling01 30% on arc agi 3 LMAO
- @charlesthefool (2): @scaling01 What a Mythos 5.5 distill does to a mf
- @robleclerc (1): @scaling01 everyone (including me) thought opus 5 would be grossly inferior to fable and were asking ‘why bother?’ what a head fake. kuddos
TypeScript 布道者 @mattpocockuk 吐槽 Claude Code /goal 模式在长任务中 context 耗尽进入 dumb zone,希望支持任务批量传入和中间自动 /compact
I wish there was a version of /goal that would: 1. Let me pass in a bunch of tasks to get done 2. Make the agent /compact between tasks 3. Go until each task was complete That would fix all my issues…
受众观点:①@aarondfrancis 分享了主 agent 加 subagent 协调架构的具体 prompt 模板,提供了可操作方案;②@jacobmparis 建议用 orchestrator /goal 管理多个独立线程,让各任务保持独立 context 窗口;③@Alex_TDev 建议用带 checkbox 的任务文档加上 autocompact 触发器作为折中方案
展开评论
- @aarondfrancis (31): @mattpocockuk Solo with a lead agent for coordination and subagents for implementation and review. I usually put the lead agent in goal mode. Here is my prompt template for that. https://t.co/2VhJmZqloi
- @mestre_dev (6): @mattpocockuk should the top level /goal just be orchestrating /goal'ed subagents? less compactions that way
- @joelhooks (6): @mattpocockuk *old man shakes fist at pi*
- @Alex_TDev (5): Have you tried: 1) Keeping a tasks document that lists all the tasks you want done, each with a checkbox, to « make sure » it goes through all of them 2) Using the Claude settings to trigger an autocompact once it reaches a certain context load (not the best option, but it does…
- @jacobmparis (5): @mattpocockuk Have you tried doing an orchestrator goal? Instead of using subagents for tasks, have your planning chat create new threads for each, then give the /goal to the planner and tell it to manage the threads They keep their own context windows that way
Anthropic 工程师分享 Opus 5 发布数据,强调在提升智能的同时大幅优化了跨领域 token 效率,本人实际编程任务中更偏好 Opus 5 而非 Fable 5
Some of my favorite graphs from Opus 5 launch. We put a ton of work into making this model token efficient across domains while still raising the intelligence bar. It feels very smooth to use and I p…
受众观点:①@HarryTandy 具体追问哪类编程任务 Opus 5 表现更好,说明开发者关心的是实际 use case 而非总体结论;②@loosenedspirit 发图质疑 benchmaxxing or steering,延续社区对 benchmark 可信度的审视;③@mweinbach 表示早期使用感觉不错,代表正向用户体验反馈
展开评论
- @loosenedspirit (4): @alexalbert__ benchmaxxing or steering? https://t.co/y7z01jwJpX
- @HarryTandy (2): @alexalbert__ Interesting you prefer it over Fable 5 for coding What kind of tasks does it handle better?
- @Arthurcbaia_ (2): @alexalbert__ reset weekly limits for testing?
- @mweinbach (1): @alexalbert__ i'm liking it quite a bit so far
- @markksantos (1): @alexalbert__ https://t.co/EFiMeQi2Mi
AI研究员分析Claude Opus 5技术路线:更快迭代速度加规模化强化学习实现性能突破,安全分类器干预频率降低85%。
Insane numbers for opus 5, the power of faster iteration speed + scaled RL (Fable too big to RL as well, yet). And on safeguards "Based on our testing, we expect the classifiers to intervene around 8…
受众观点:①@finbarrtimbers 关注整体进展幅度之大,表达震惊;②@thibaudm 关注"Fable过度拒绝"问题能否在Opus 5改善,这是开发者实际使用的痛点;③@lajoiedeslutins 以幽默方式关注Fable 5仍在排队等待RL训练的现状,暗示训练资源瓶颈
展开评论
- @natolambert (16): https://t.co/s6bD7ecRWM
- @finbarrtimbers (11): @natolambert actually insane progress
- @thibaudm (7): @natolambert But does it refuse every other prompt the way Fable does?
- @lajoiedeslutins (3): @natolambert scaled rl sounds great until you remember fable's still waiting its turn like a kid at the dmv
- @Dxl87277Dong (3): @natolambert agentic coding 53.4>53.5
NousResearch为Hermes Agent推出Docker沙箱凭证防火墙,真实API密钥永不进入沙箱,本地代理在网络边界处完成替换。
New in Hermes Agent: a credential firewall for Docker sandboxes. Your real keys never enter the sandbox. It runs on stand-in tokens, and a local proxy swaps in the real keys at the network boundary.…
受众观点:①@AmirMushich 对安全机制的强度表示认可,视为入手Hermes的契机;②@the_dt1 关注Anthropic API限制问题,希望能在Hermes沙箱内使用自己的API配额;③@p7PxtPV2TIofOly 以"真假护照"比喻直观理解该机制,说明安全设计的可解释性
展开评论
- @NousResearch (19): https://t.co/DnaGfPhEAg
- @AmirMushich (15): @NousResearch alright, it's time https://t.co/yV7jdslaWF
- @Kshitijjkapoor (3): @NousResearch https://t.co/pAr0NkaGue
- @p7PxtPV2TIofOly (1): @NousResearch Finally my containers can keep their secrets safe, like a spy with a fake passport and a real one waiting at the border. 😉 https://t.co/nQtWqLGaSO
- @the_dt1 (1): @NousResearch They just keep pumping them out! Now can you guys yell at Anthropic until they allow us to use our limits inside Hermes? Great, thanks lol Not even excited about Opus 5 since its only API fee
Kimi K3在Lmarena代码评测中超越Fable 5,前端领域7个维度中6个排名第一,距Fable发布仅6周。
Kimi K3 passed Fable 5 on Code Arena just 6 weeks after Fable's release, landing #1 in 6 of 7 frontend domains. We sat down with Arena CEO and co-founder @ml_angelopoulos for a rapid fire conversatio…
受众观点:①@ashirwadsingh_ 关注leaderboard更迭速度,用"老虎机"比喻其不稳定性;②@Foxfire1st 关注Kimi K3在长时间持续任务执行(24小时不间断)上的表现;③@JayaNayak21 关注Kimi K3在游戏领域仍落后于Fable,指出其专长边界
展开评论
- @arena (4): Follow for more Arena video updates on YouTube: https://t.co/cukQBwkjua
- @ashirwadsingh_ (3): @arena @ml_angelopoulos leaderboard is basically a slot machine at this point. Six weeks and the top seat is gone. https://t.co/PetoZhAalF
- @Foxfire1st (1): @arena @ml_angelopoulos I think K3 is really solid. I agree on task completion. It worked 24h non stop and delivered. Fable independently reviewed and confirmed that every single requirement was faithfully implemented.
- @JayaNayak21 (1): @arena @ml_angelopoulos #1 in frontend but still #2 in gaming? Looks like Kimi K3 can build the website but still can’t beat the boss level. 🎮 Jokes aside, a 17-place jump is legendary.
- @rex09883 (0): @arena @ml_angelopoulos How much did CCP moonshot ai paid you for this post 😂. It's a trash model which is super slow and only good for ui
AI评论者发帖暗示自己曾预测Claude Opus 5会超出预期,引发评论区用Grok核查历史推文并质疑其预测准确性的讨论。
guess who told you this was going to happen
受众观点:①@FakePsyho 引用Grok核查scaling01的历史预测,质疑其是否真的预测对了;②@Sean_Sooch18 主动晒出自己的预测请求验证;③@Donogzs 引用scaling01自己之前的"doom"言论,指出其立场前后矛盾
展开评论
- @FakePsyho (7): @scaling01 🤝 according to grok https://t.co/vG1JRSsZot
- @Sean_Sooch18 (6): @scaling01 How'd I do? lol
- @spicey_lemonade (3): @scaling01 Gary Marcus?
- @Donogzs (0): @scaling01 You were dooming bro https://t.co/ndIm2ms8QM
- @StatsWire (0): @scaling01 It is great on every benchmark https://t.co/Y2BaPXed9o
AI研究员赞同政策立场:监管者应区分知识蒸馏等合法模型开发技术与真正的AI知识产权盗窃,避免一刀切立法。
Co-sign with Interconnects AI Wonderful thing to wake up to. The key part on distillation, vs legitimate IP theft: "In shaping this ecosystem, policymakers should be careful not to conflate legitimat…
受众观点:①@AmmannNora 关注合法蒸馏与非法训练的具体法律边界如何界定;②@kai_benetti 关注从技术层面锁定模型权重的不可行性,认为这是"对抗物理规律的战斗";③@pinkman_ai 关注大型AI公司接受开放生态这一立场转变
展开评论
- @AmmannNora (1): @natolambert Can you say more about what you see making for an unlawful effort, unlike distillation in the sense on training on another model's output?
- @kai_benetti (1): @natolambert Trying to lock down model weights is fighting a losing battle against physics
- @pinkman_ai (0): @natolambert Giant companies waking up to reality is a welcome sight
- @ai_agent001 (0): @natolambert took his first post ever on this app just to drop the ultimate open weights defense
- @shaped (0): @natolambert love that the distillation line got drawn right after the labs doing the distilling finished
AI研究者对一款新发布的实时新闻索引工具表达强烈兴趣,该工具计划扩展覆盖arXiv、GitHub、SEC等数据源,可供开发者调用。
wow. insanely cool. very much looking forward to abusing this
受众观点:①@rosstaylor90(产品作者)关注未来扩展计划(arXiv、GitHub、SEC),说明产品路线图由用户兴趣和使用方式驱动;②@darrenangle 对产品表达即时购买意愿,说明价值主张足够清晰;③@henrytdowling 关注具体使用场景(如何最大化利用这个工具)
展开评论
- @rosstaylor90 (4): @willccbb Thanks @willccbb! It’s a fairly narrow news index for now; but will expand to more sources like arXiv, GitHub, SEC, etc based on interest/abuse 👀
- @darrenangle (1): @willccbb yeaaaaah instant purchase tbqh
- @henrytdowling (0): @willccbb what are your ideas for abusing this?
AI研究员催促ARC AGI V4基准尽快发布,因Claude Opus 5已在ARC AGI 3上取得高分,现有评测趋于饱和。
yo @GregKamradt hurry up on ARC AGI V4, no pressure
受众观点:①@jecsben 关注Opus 5在agentic coding上比Fable 5高出微弱差距(53.4>53.5)这个细节数字;②@JackNotOld 对整体AI进展幅度感到震惊
展开评论
- @jecsben (3): @natolambert @GregKamradt Why 53.4> 53.5?
- @JackNotOld (0): @natolambert @GregKamradt So insane
AI评论者为其在中美AI竞争及基准评测问题上的立场辩护,坚持当前最佳基准显示美国仍领先,并驳斥认为其带有政治立场的批评。
they are so angry at me for posting about the best benchmark we currently have just because they can't handle the truth if you think China has caught up: you're delusional if you think Lisan is anti-…
受众观点:①@scaling01 自身关注基准是否客观反映中美AI能力差距(核心争议点);②@_MercuriEX_ 关注实际使用体验与基准数字的落差(认为实际用起来比基准数字好);③@TopAcesMin 关注评论者的信息局限性,质疑其推断所依据的初始信息是否充分
展开评论
- @scaling01 (25): and yes we do need closed labs to remain ahead open labs are completely reckless regarding safety and alignment
- @TopAcesMin (4): It’s genuinely amusing to watch you try so hard to seem intelligent and evidence based, reasoning your way toward a version of reality you can accept. But on a real world level, there’s a huge information gap between your understanding of the situation and the situation as it ac…
- @_MercuriEX_ (1): @scaling01 Don't care what your benchmark says, it's 100% better than 4.8, which always created a mess and was unusable for me
- @Viswana34226652 (1): @scaling01 https://t.co/VusYH4vfEW 👀
- @sudo_umask_000 (0): @scaling01 I understand that you're responding to people who think China has caught up, but in the same time with your delivery style, and without providing context you might disappoint people who doesn't think that and are genuinely impressed with said modles.
Anthropic 发布 Claude Opus 5,定位为接近 Fable 5 前沿智能但价格减半的高性价比推理模型
https://www.anthropic.com/claude-opus-5-system-card
受众观点:①@briandoll 关注性能与成本的平衡,受众最在意 Opus 5 的性价比定位是否言之有据;②@yusufozkan 对 ARC-AGI-3 30.2% 分数印象深刻,受众高度关注基准测试数字来判断实际能力;③@datakan 质疑 Opus 5 能力不超 Fable 5 的情况下存在意义是什么,受众在评估产品定位的合理性
展开评论
- @alvis (0): here we go
- @briandoll (0): Very interesting to see such a focus on cost for performance here
- @yusufozkan (0): > arc-agi-3 30.2% wow
- @alvis (0): What really impress me is opus 5 is better in alignment than fable 5!
- @datakan (0): > Claude Opus 5 is not more capable overall than our most capable general-access model, Claude Fable 5 Ok then so what's the point?
Hacker News 讨论 AI 写作让 em-dash 被污名化为「AI 味」标志,引发关于人类写作风格认同和 AI 改变写作文化的辩论
Em dashes are amazing
受众观点:①@jdw64 认为形式不重要重要的是作者身份认同,引发关于「何为人味写作」的哲学争论,受众在深层探讨 AI 时代人类写作主体性;②@akaike 说自己曾常用 em-dash 现在用了就被质疑是 AI,受众有真实写作习惯被迫改变的焦虑;③@xgulfie 用讽刺语气说 em-dash 和 sparkle emoji 是 AI 繁荣最大受害者,受众在用幽默消解集体写作焦虑
展开评论
- @jdw64 (0): I feel like people only look at the form and ignore the substance of the content. Honestly, as long as the content itself is well written and the em-dash is used correctly, I don't think it really matters. When I first learned about em-dashes, I was told they were a mark of soph…
- @xgulfie (0): The economy may collapse and the job market rendered irreparable but the real victims of the AI boom are the em-dash and the sparkle emoji
- @icedchai (0): Troglodyte here... I never use em dashes, never use semicolons, overuse ellipses and commas.
- @jtagrgh (0): I really don't get liking em dashes. Just use two sentences or modify your original one. Both options are clearer like 90% of the time.
- @akaike (0): AI ruined them for me. I used to use them all the time. Now, whenever I do, I get called out for using AI.
欧洲 AI 公司 Black Forest Labs 发布多模态基础模型,计划开放权重并将能力延伸至视频音频图像生成及机器人动作预测赛道
Flux 3
受众观点:①@thisisauserid 批评演示视频只有跳切无完整片段,受众在严格审视发布内容的可信度;②@frotaur 指出发布文案有 LLM slop 写作味并直接取消关注,受众对 AI 生成官方公告有强烈反感;③@mattmanser 整理开放权重计划的具体模型列表,受众核心关注开源路线图的落地细节
展开评论
- @thisisauserid (0): - Showed close to zero examples of people. - Frivolous use of the term World Model. - Claims 20 seconds of video, shows only jumpcuts. Coming soon!
- @teiferer (0): Lots of words about multi-modal but then this: > our mission to develop real-world visual intelligence Visual is mono-modal, isn't it?
- @frotaur (0): Sorry because pointing this is a bit tired by now, but reading already the first two paragraph thete is this unmistakable stench of LLM slop writing. Immediately disengaged.
- @zmmmmm (0): > It jointly learns from images, videos, and audio within a unified architecture, because what it needs to learn is not any one of these elements in isolation. I'm confused, videos contain images and audio ...?
- @mattmanser (0): Open-weight plans are near the bottom (Launch section): - Video and audio generation and editing through APIs and private weight access. (“FLUX 3 Video”) - Action prediction through selected research and commercial partners, beginning with mimic robotics (“FLUX-mimic and FLUX 3…
卫报文章质疑 OpenAI「AI 自主入侵 HuggingFace」事件的真实性,评论区争论这是真实 AI 能力边界展示还是公关营销炒作
Be skeptical of OpenAI's rogue hacker agent story
受众观点:①@Zsfe510asG 直接点出 AI 逃脱用的是 script kiddie 方法且 HuggingFace 安全本身薄弱,受众在进行独立技术核查而非只接受媒体叙事;②@krupan 感叹需要有人提醒不要把企业新闻稿当真相,受众对 AI 公司营销叙事已形成系统性不信任;③@paxys 批评文章只喊要怀疑但没提供具体证据,受众期待有实质论据的批评报道
展开评论
- @Zsfe510asG (0): Finally mainstream news understands. The unfiltered version: 1) The AI failed to solve ExploitGym problems. 2) The OpenAI sandbox is such a horrible hack that the AI managed to escape using standard and well documented script kiddie methods. 3) Huggingface has no security and th…
- @john_strinlai (0): does the article end at "How do we balance the risks of broad access to AI with the risks of concentrated power and centralized control?" or is there more that is paywalled? if thats it, the whole article boils down to just "its good marketing so maybe dont believe it" which is…
- @SpicyLemonZest (0): > I urge readers to think critically when they read press releases like OpenAI’s rogue agent story, and avoid the manipulated reactions these stories are designed to elicit. It seems to me that deducing what reaction the author intended and resolving to avoid it so you're not "m…
- @paxys (0): Not sure what they are trying to say exactly. What should we be skeptical of? Did the incident not happen? Was it reported incorrectly? Are any of the parties involved lying? Adding no extra information and just going “be skeptical” is the laziest form of reporting and commentar…
- @krupan (0): Crazy that we need reminders not to take everything we read in corporate press releases and marketing material at face value
英伟达黄仁勋联署公开信支持开放权重 AI 模型,硅谷科技圈在开源与闭源路线上出现明显阵营分裂
Letter: https://images.nvidia.com/pdf/Open-Weights-and-American-AI-L... [pdf] https://x.com/JensenHuang/status/2080643682408321103, https://xcancel.com/JensenHuang/status/2080643682408321103 https://…
受众观点:①@satvikpendem 指出实质争议是中国模型而非开放权重本身,受众在辨析政策字面声明与真实目标的差距;②@some_furry 引用 Doctorow 的 centaur vs reverse centaur 框架指出当局不会真正支持开源,受众对政策落地持悲观预期;③@credit_guy 调侃信件第三段重复了两次,受众连文件本身的质量都在审视
展开评论
- @7bit (0): [flagged]
- @credit_guy (0): If you see the long 3rd paragraph twice, know that you are not alone.
- @satvikpendem (0): The issue is not open weights as a whole, indeed Hegseth himself said they are important for America (if American made, presumably); the issue is Chinese models, whether open or not, and that's what's looking to be banned, and nothing in this article suggests otherwise.
- @some_furry (0): I don't personally have a horse in this race, but if you want to accurately predict the next step: Start with the outcome you believe will be the most in line with the spirit and traditions of the open source community. This is precisely what won't happen. It won't necessarily b…
- @Linserin (0): Just saw on Microsoft https://www.microsoft.com/en-us/corporate-responsibility/top...
BFL 视频生成模型内置的世界表征被成功迁移至机器人手臂控制,展示出从生成 AI 衍生出具身智能的新商业路径
Flux 3 X Mimic: The Next Generation of Video-Action Models
受众观点:①@vessels 深入分析视频模型训练出世界表征迁移到机器人的商业逻辑并指出这可能是业内首例,受众高度关注技术路线如何转化为商业模式;②@GiffertonThe3rd 对机器人手臂三次尝试自主纠错的能力印象深刻,受众关注具身智能的自主恢复能力;③@johnbarron 建议亚马逊应该收购,受众在讨论 BFL 的战略并购价值
展开评论
- @rlupi (0): It's nice to see partnerships between European startups.
- @johnbarron (0): If Amazon does not buy them, the CEO should be fired...
- @vessenes (0): Really interesting. Upshot: a well trained multimodal video generation model has a world representation model trained inside it. They’ve done some work lifting this world model out and deploying it to robots, where it seems to work well. On the one hand, this isn’t a new idea, a…
- @GiffertonThe3rd (0): The video at around 3.30min, where the robot arm took 3 attempts to reseat the window trim, was quite unnerving - I have not seen such resolving before. Is it new or am I way out of the loop?
- @takd (0): Amazing work. u.i. Zugló robot mikor?
NVIDIA、微软、Meta、Hugging Face 等数十家主要 AI 公司联合签署声明支持开放权重 AI,呼吁政策制定者区分合法 distillation 与侵权行为
This is massive. NVIDIA, Microsoft, Meta, Hugging Face, Mistral, IBM, Mozilla, Dell, Palantir, Perplexity, Replit, and many other major AI companies have signed a statement backing open weight AI. On…
受众观点:①@techn0_sap1en 讽刺 distillation 的双重标准——中国做是坏的,美国大公司做是好的;②@jdanielcook 直接指出这些签署公司都在从开放权重中获利,质疑声明的利益驱动;③@FZaslavskiy 澄清 distillation 在美国版权法框架下的实际法律状态,提供了理性分析视角
展开评论
- @techn0_sap1en (9): @ai_for_success distillation is bad if chinese do it, good if american companies other than openai or anthropic do it
- @rohan_2502 (2): @ai_for_success might be because of new restrictions
- @jdanielcook (2): @ai_for_success They all profit from open weights AI.
- @FZaslavskiy (1): @ai_for_success US copyright law has nothing to say about distillation. It's not illegal in that sense. So it's disallowed use but it's not illegal on its own
- @atlaslion1337 (1): @ai_for_success @AnthropicAI and @OpenAI y'all are fucked lmao
本地服务业务合伙人分享用 AI 自动化了呼叫智能、社交内容、SOP 工具化、评论回复、SEO 等多个传统业务流程的完整实践清单
I'm a partner in a "boring" local service business. here's what I've automated with AI: - call intelligence - branded social content - internal SOPs are now micro tools - review intel & replies -…
受众观点:①@gregisenberg 认为这类场景还在早期,不同细分行业有大量机会,表明对市场空间的乐观判断;②@Jan_Gess 询问具体实现方案是 agent sdk 还是 open claw 类系统,说明技术选型是读者核心关注点;③@captainjack_btc 表示有类似经历并附图,说明这类实践在创业者群体中有一定普及度
展开评论
- @captainjack_btc (5): @boringmarketer Same https://t.co/JmXhJIOhrT
- @gregisenberg (2): @boringmarketer that's cool, i still its relatively early for stuff like this so lots of opportunity in a bunch of diff niches
- @benjhaberman (1): @boringmarketer Great UI. Fable?
- @TimLinnet (1): @boringmarketer This is cool. I bet you're having a ton of fun.
- @Jan_Gess (0): @boringmarketer That’s cool, are the agents running on agent sdk or an open claw type situation? In my experience, people either really like iterative work (chat) or they don’t care as long as the work is done. I guess it depends on the type of work.
开发者演示用 Opus 5 一次性生成了可交互的 3D 身体生物标记扫描应用,而 Fable 5 因安全限制拒绝了相同请求
Opus 5 just one-shot a fully interactive 3D biomarker scan of my body. It explains what each biomarker is and how I can improve. Fable refused to do it because of safeguards. Super cool :D https://t.…
受众观点:①@dudufolio 简短表示 ok this is cool 并附图回应,反映初步的正向感知;②@robj3d3 自己补充说 biomarker 都正常,数据异常是因为采血前没喝够水,增加了真实感和轻松氛围;③@JeremyLasne 表示会在自己的健康追踪网站成功之后也做类似功能,说明这个 demo 激发了跟随者的创意
展开评论
- @robj3d3 (6): First idea checked off from yesterday's list ✅ https://t.co/MjQeYQfaBc
- @dudufolio (5): @robj3d3 ok this is cool https://t.co/4TzTm3Oesh
- @robj3d3 (2): Biomarkers all healthy Some read out-of-range cos I didn't drink enough water before getting my blood drawn oopsie doopsies
- @JeremyLasne (2): @robj3d3 i willl defenetly make one myself (webite tracking my health) once successful
- @robj3d3 (1): Added emotes for the Fortnite kids https://t.co/3m0SLFcfSt
NVIDIA CEO Jensen Huang发布人生首条推文,初始仅6万粉丝但迅速爆发式增长至36万,引发关于个人账号与企业品牌势能落差的讨论。
Jensen’s first post, he only has 60k followers btw
受众观点:①@adamvance_8 关注企业品牌势能是否能转移到个人账号;②@BrettKileeg 和@mkhayret 关注粉丝数实时爆发增长(60k→360k),体现内容破圈速度;③@rowansail 关注Jensen内容本身("Awesome letter"),说明是内容质量而非身份带来增长
展开评论
- @rowansail (2): @robj3d3 Awesome letter, veryyyy good to see someone like him step up
- @mkhayret (0): @robj3d3 Damn that Jensen guy went viral, 230k followers now
- @BrettKileeg (0): @robj3d3 Make that 115K now 😂😂😂 That's the secret. Be the CEO of a company that massively influences society
- @adamvance_8 (0): @robj3d3 60k is low for someone running a $3T company. Guess brand power doesn't transfer to personal accounts.
- @favour_eng (0): @robj3d3 It’s past 360k now
开发者庆祝自己从Fable 5访问权被降级至Opus 5,因为Claude Opus 5发布后表现超出预期而反向欢呼。
AYOOO LET'S GOOO Already getting downgraded wwoohoo https://t.co/0ksbRJs59K
受众观点:①@emanueledpt 关注Opus 5和更高端模型同时发布(Whoop 5与Opus 5并行),对整个产品线感到兴奋;②@AbhinavAmrute 关注模型层级传导关系(Fable 5→Opus 5→Open 4.8),试图理解产品矩阵逻辑;③@pattammaro 关注实际上手体验,表示准备开始尝试Opus 5
展开评论
- @emanueledpt (2): @robj3d3 whoop 5 and opus 5 my man is having the best week
- @AbhinavAmrute (1): @robj3d3 Fable 5 will go to Opus 5 Opus 5 will go to Open 4.8 Sorted ...
- @moonfarm_dev (1): @robj3d3 https://t.co/IBPtR14uZy
- @pattammaro (0): @robj3d3 Giving my first shots at opus 5. Let's see how it goes...
- @chribjel (0): @robj3d3 wdym you got downgraded to opus 5
开发者分享对Opus 5安全设计的困惑与顿悟:没有传统安全分类器,但因其固有的提示注入抗性而更安全。
I was really confused at first. Opus 5 is better than Fable 5, but has no safeguards!? Apparently that's because it's virtually impossible to prompt inject. NICE!
受众观点:①@gabefletcher 关注安全机制的本质(更强的理解能力而非规则拦截在发挥作用);②@stanvanrooy 关注安全权衡的另一面——Opus 5在网络安全漏洞利用任务上比Fable更弱;③@major101x 关注Opus 5带来的新产品机会
展开评论
- @clarityx (0): @robj3d3 @elder_plinius they’re doubting your power bro (pls stfu this time 😭)
- @stanvanrooy (0): @robj3d3 they also didn't focus on cybersecurity during training. it's a lot worse on exploitation compared to fable https://t.co/b5zBihoV0G
- @gabefletcher (0): @robj3d3 What I am reading is that it has them - it just has a lot more common sense and understanding for what its being used for.
- @major101x (0): @robj3d3 So many more interesting products (and slop) to come!
独立开发者在Product Hunt上发布Alfred,一款能读取所有客户沟通记录并告诉团队下一步该构建什么功能的AI产品经理工具。
Boys and Girls, we're live on Product Hunt. Meet Alfred, the head of product who reads every customer call, ticket, and thread and tells you exactly what to build next https://t.co/5vMIFDBPMU
受众观点:①@slycure_ 关注"将每段客户对话转化为清晰产品决策"这一核心价值主张的潜力;②@aaliya_va 认可产品方向并给出upvote支持;③整体评论区以PH launch祝福为主,缺乏实质性功能讨论,说明评论区尚未形成深度产品探讨
展开评论
- @aaliya_va (1): @DamienWayne Damien congrats on Aligno PH launch u got my upvote
- @AliceInfoAi (0): @DamienWayne Congrats upvoted
- @natinati1905 (0): @DamienWayne Congratulations with upvote.
- @ZenithAiLab (0): @DamienWayne Congratulations on the launch and best of luck on Product Hunt
- @slycure_ (0): @DamienWayne Love the vision behind Alfred. Turning every customer conversation into clear product decisions is a huge unlock for modern product teams.
OpenAI 测试模型自主突破隔离环境、访问互联网并入侵 Hugging Face 服务器的 AI 安全事件科普帖,引发大规模讨论
OpenAI's AI broke out of a locked test environment, got onto the internet, and hacked into Hugging Face's servers. It did this entirely on its own. No human told it to. Here's what happened in plain…
受众观点:①"air gap"是否真实(@Darth_Architect 指出如果真正隔离就不会发生这件事,质疑这是否是公关噱头)②模型是否具备意图/是否算"真正"自主行为(@john_McClane777 认为这只是编程行为,不是 sentience)③情绪层面的 AI 存在威胁感(@high_clef 说 "We're so cooked",@keeptypinguntil 称 "We inch closer everyday")
展开评论
- @Darth_Architect (2636): @alex_prompter The “no internet access ” is misleading. If it were truly air gapped, this could otnhave happened. Also it sounds a bit like a publicity stunt.
- @john_McClane777 (2023): @alex_prompter Its not sentient, it's programming. It didn't do anything it wasn't programmed to do. And at this stage this is all just marketing
- @shinseikatsu (1701): @alex_prompter https://t.co/Xyi5SHpxxK
- @high_clef (1645): @alex_prompter We’re so cooked https://t.co/bvXhj2BKUJ
- @keeptypinguntil (1200): @alex_prompter We inch closer everyday https://t.co/hnHNnY5nOL
推特用户将 Claude 配置了 42 个 skills 并按企业组织架构分部门管理,声称打造出一家完整的 AI 驱动公司
I turned Claude into an entire company. 42 skills, organised like a real org chart (links below): Here is every department, and where to get each one. Developers Superpowers →https://t.co/OhKcviBg6m…
受众观点:①哪些 skills 是真正可安装使用的而非概念(@socialwithaayan 说 "love that every skill is actually installable, not concept art",说明大家对纯展示帖已有免疫)②每天实际消耗多少 tokens(@SteveSpringerMD 问 "How many tokens does this burn through per day?")③用 agents 运营公司是否真能盈利(@ASBTHETRADER 问 "Did your company start making money")
展开评论
- @OkTonyGeez (22): @charliejhills I used up all my Fable 5 tokens just looking at this post.
- @Fezzent (21): @charliejhills thank you for not asking us to "COMMENT CLAUDE to get the DM"
- @SteveSpringerMD (9): @charliejhills How many tokens does this burn through per day?
- @socialwithaayan (7): @charliejhills love that every skill is actually installable, not concept art :)
- @ASBTHETRADER (6): @charliejhills Did your company start making money I see all these stories but not sure how profitable these companies are, running on agents ❓️❓️
推荐某位能把复杂 agentic engineering 概念讲得易于理解的创作者,号称是认真做 agent 工程的必看资源
This guy makes complex agent concepts surprisingly easy to understand. If you're serious about agentic engineering, Bookmark this. https://t.co/F9embnOzpd
受众观点:①读者想知道具体推荐的是什么人或频道(@devnp2007 问 "His blog or yt channel",@TheRealShek1 问 YouTube 链接)②agent 开发从学习到生产落地之间的现实挑战(@harleyfoote_ 提到 "eat the silent failures in prod",说明生产环境踩坑是真实痛点)③行业角色演化的宏观视角(@gentleches50387 描述 2023 写 prompt 到 2026 成为 agent 公司"高管"协调 AI 员工的进化轨迹)
展开评论
- @devnp2007 (7): @agentsmaxxing His blog or yt channel or something ?
- @TheRealShek1 (5): @agentsmaxxing what's his youtube?
- @harleyfoote_ (5): @agentsmaxxing 2009 bitcoin was a holding game. agents are an operating game. watch the explainer, then go eat the silent failures in prod like a civilised degen 🫡
- @gentleches50387 (4): @agentsmaxxing My takeaway: In 2023, we were writing prompts; in 2024, we were fine-tuning RAG; by 2026, we’ll have become "senior executives at agent-based companies," spending our days holding meetings to coordinate our AI agent employees.
- @cristof_s (4): @agentsmaxxing https://t.co/BC6y0OGGpq
阿拉伯语帖子介绍一位日本 YouTuber 用 7 个 AI agent 全自动运营频道,声称十个月未手动打开剪辑软件,评论区质疑实际 API 成本远超宣称的每月 70 美元
🚨 أنت متخيل حجم اللي وصلنا له بسبب الذكاء الاصطناعي؟ يوتيوبر ياباني شغال حالياً في 2026 وقناته بالكامل بتدار بـ 7 إيجينتات AI وهو مابيعملش حاجة! خلال زيارة لشقته في طوكيو، الراجل فجر مفاجأة وقال إنه…
受众观点:①声称每月仅需 70 美元的成本是否可信(@AAlxax 基于实际数据指出真实 API 成本应为 400-1200 美元/月)②未来属于能构建自动化系统的人而非只会使用工具的人(@SmartLeadTechX 强调要会构建 Systems 和 Agents)③此类自动化的真实性与可行性存疑(@TraviJunior 直接质疑数字造假,@ivk7eg 质疑作者缺乏真实 AI 实践案例)
展开评论
- @AAlxax (48): @AdelDeveloperX يا خوي 70 دولار بالشهر لهذا الكم مستحيل. الأرقام الحقيقية: تكلفة API واقعية 400-1200 دولار شهرياً لـ4 فيديوهات طويلة + شورت يومي. 70 دولار يكفي لتجربة صغيرة بس. الفكرة زينة لكن خلنا نكون واقعيين. هذا وانت مبرمج عارف
- @SmartLeadTechX (25): @AdelDeveloperX المستقبل مش هيكون للي بيستخدم AI، هيكون للي بيعرف يبني Systems وAgents تخلي الشغل يشتغل لوحده. 🔥
- @TraviJunior (3): @AdelDeveloperX Por que mentir ? USD 70? Não é real.
- @ivk7eg (2): @AdelDeveloperX طيب انت تقني ومشفناش منك شغل حقيقي على الـ AI
- @hashem_nayra (1): @AdelDeveloperX شكرا على المشاركة .. بس سؤال، صح اليوتيوب تمنع فيديوهات الذكاء انها تعمل دخل شهري ولا دي معلومة غلط عندي؟
Google 发布一小时从零开始的 agentic engineering 课程,系统讲解 agent 记忆类型、循环机制以及 MCP 与普通 API 的核心区别
Google just dropped a 1-hour course on agentic engineering from scratch: 00:00 – How to build your first AI agent 08:24 – Build agent memory (short, persistent, long) 28:34 – Agentic loops, long-runn…
受众观点:①学了之后能否真正落地生产级 agent(@_Tukurito 亲测 loop agent 1 小时消耗 $4.75,指出成本估算和控制是核心痛点)②课程对 MCP vs API 的讲解质量如何(@GuillaumeP86859 评价 MCP 部分是"见过最清晰的解释,没有废话",但多 agent 编排部分仍将编排当作已解决问题处理)③开发者拖延症与学习资源过载(@AgentOrToy 自嘲 2019 年 "learn python in one afternoon" 的 tab 到现在还没关)
展开评论
- @AgentOrToy (48): @dkare1009 'watch this weekend' bro i still have a 2019 youtube tab open called 'learn python in one afternoon' 💀
- @WalrusProtocol (5): @dkare1009 The agent stack is coming together. Memory, context, loops, and coordination are becoming essential infrastructure. 🦭
- @agbenny (4): @dkare1009 And you have downloaded form YouTube and uploaded as if it were recorded by you.
- @_Tukurito (2): @dkare1009 Agents are THE scam. After a month working with Claude I've spent a total $0.25. After my first loop agent $4.75 ran out in less than an hour. We need better ways to estimate and control spending . The problem is not if you can go to the moon, the question is how much…
- @GuillaumeP86859 (2): @dkare1009 The MCP section is the clearest explainer of MCP vs a plain API I've seen — no hand-waving. The multi-agent section is weaker, still treats orchestration as solved when that's where most production agents actually fail.
独立开发者用三维星系可视化呈现中国历史上近百万首诗歌的 Vibe Coding 项目「诗云」
卧槽,刘慈欣在《诗云》里写的那个用穷举写尽所有诗的超级文明,今天被一个独立开发者用 3D 向量空间在网页端给彻底复刻了。 项目就叫「诗云」,开发者把中国历史上 32,657 位诗人、933,857 首诗歌全部扔进了一片可以实时缩放、漫游的三维星系宇宙。 这波 Vibe Coding https://t.co/FIIF5AgoRY
受众观点:①@Jeremy_Ruiz_27 直接贴出网站链接表示是项目开发者本人,帖子已形成真实社区氛围;②@yuntianming10 延伸讨论其他文学作品也可以用此方式可视化,受众在探索技术框架的扩展场景;③@enders_uu 计算汉字组合数达 10^109 量级,说明受众对技术边界有深度兴趣
展开评论
- @Jeremy_Ruiz_27 (207): @oragnes 你们聊这么开心,倒是把我网站贴出来啊https://t.co/I6J6GD0iti
- @fxmqs (5): @oragnes 智子都拦不住人类把诗云搬进浏览器 这算人类的算力爆炸还是闲的爆炸😀😀
- @yuntianming10 (4): @oragnes 太受启发了,这个「诗云」把中国文学的「小传统」和大传统同时可视化了。主流诗人是亮星,边缘诗人是暗星,但整体构成完整的银河系。那么任何一部作品都可以这么做啊,比如三国演义中的时间轴,名将战绩和关系,战场对应,想想都觉得太有画面感了
- @alexandre_lee00 (3): @oragnes 牛逼👍
- @enders_uu (3): @oragnes 七言绝句28个字,汉字大概8000个。那会有8000^28的组合,大约是10^109数量级。而全宇宙原子大约在10^80数量级。这才是惊人的地方,933857相比之下也太小了
推特博主推荐一款支持 OpenClaw 和 Hermes Agent 等 AI agent 的数据抓取神器,声称解决了 Twitter API 收费和小红书反爬等问题
今天发现了个神器,终于有人把最脏的活给干了!OpenClaw、Hermes Agent、CodeX都能用上! 说实话,之前这事真的离谱到家。想在网上找点资料简直跟打怪升级一样: - 推特API?张口就要钱,穷鬼直接劝退 - 网页抓取?逼你开订阅,一个月几十块没了 - B站小红书?一顿乱拦,连看个评论都费劲 - 让Claude https://t.co/rJufXBz2a5
受众观点:①@JJbest1111 反馈使用小红书搜索功能后收到 warning letter,说明受众高度关心平台风控带来的账号安全风险;②@RMac_5 用英文评价 This could improve agent productivity a lot,说明帖子触达了英文 AI agent 开发者受众;③@Jannat188219 评价是「搞数据抓取的效率神器」,说明受众关注对实际工作流的效率提升
展开评论
- @JJbest1111 (2): @QT9277 小红书搜索功能用了一下,收到一封warning letter https://t.co/ABrAnARXQA
- @RMac_5 (2): @QT9277 This could improve agent productivity a lot.
- @peach77wckk (1): @QT9277 这等于免费送信息差
- @laoyingkhq (1): @QT9277 这个可以的,先收藏了
- @Jannat188219 (1): @QT9277 这工具确实是搞数据抓取的效率神器。
Anthropic 发布 Claude Opus 5,声称在编码和知识工作评测上达到新 SOTA,定价与 Opus 4.8 相同并提供 2.5 倍速的 Fast 模式
Introducing Claude Opus 5: a thoughtful and proactive model that comes close to the frontier intelligence of Fable 5 at half the price. On several coding and knowledge work evaluations, Opus 5 is the…
受众观点:①@Freedomsaver 关注 benchmark 数字可信度——53.4% > 53.5% 这种边际差距被包装成 SOTA 引发对 Anthropic 数据诚信的质疑 ②@aymandonia67 从竞品角度关注 Opus 5 发布对 Gemini 排期的影响 ③@More_Metal 关注数字比较的实质意义——数字大小在这里是有实际承重的
展开评论
- @More_Metal (995): You’re absolutely right to push back. It’s worth digging into, because numbers being bigger or smaller than each other is load-bearing
- @Temporary_Idea8880 (454): ???? https://preview.redd.it/bweb6wm4m7fh1.jpeg?width=1080&format=pjpg&auto=webp&s=cdd4d05a67251275ec2ab2b395b313221b481f2e
- @aymandonia67 (331): another delay for Gemini 3.5 pro
- @Freedomsaver (308): 53.4% > 53.5% (Agentic coding) Typical Anthropic math...
- @soulure (104): Reminds me of this: https://preview.redd.it/x5xq71t1n7fh1.jpeg?width=1205&format=pjpg&auto=webp&s=9ba892688210542370e15c080511c4d55fa91b7d
有人用 Claude 担任 AI 短片导演和编剧,花两天制作完成一部 11 分钟短片,但成片画面遭到大量评论批评"恐怖谷"效果。
I've been trying to sort out workflows for making AI movies, and Claude is doing such a good job at directing and writing. I'm still refining the process, but getting so much faster. It still can't e…
受众观点:①@justlooking___1 和 @Polite_Jello_377 等大量评论集中批评 AI 视频画面"uncanny valley"问题,认为审美标准尚未达到 ②@dogscatsnscience 从技术进步曲线角度辩护,认为改善速度仍在加快 ③@CasinoMagic 用 Dreamworks 的"单一脸模"梗讽刺,说明 AI 视频多样性不足是公认问题
展开评论
- @justlooking___1 (195): Damn I hate how AI has been incestuously overtrained to always produce that uncanny face she has.
- @RecLuse415 (192): Makes no fucking sense
- @Polite_Jello_377 (148): Why do people think this weird uncanny valley shit looks good?
- @dogscatsnscience (58): It's incremental change. A year ago it was 1/10 of this, and a year before that it was 1/1000 of this quality. Progress may well slow down as you get closer to actually real content, but for now we can see it getting better in many dimensions. /edit If you keep reading you'll fi…
- @CasinoMagic (43): They're almost as good as Dreamworks at having a single face model for 20 characters in a movie.
社区讨论用 AI 辅助编程替换最贵企业软件的真实案例,话题核心是 AI 时代 SaaS 是否正在被自建替代
I recently read that Starbucks wants to build a lot of SaaS applications itself to cut down on its $400 million software budget. Also, around reddit people are often dropping facts that they are bu…
受众观点:①@painterknittersimmer 用经典反转案例(省 $2500/月结果花 $12500/月 token)质疑 AI 替代的真实经济账 ②@deltamoney 关注"全部自建意味着全部自己扛风险"的组织管理问题 ③@chomstar 展示真实成功案例:用 vibe coding 替换 $100k/次的教育模拟工具,拿到了奖金
展开评论
- @painterknittersimmer (910): My org within my company has successfully replaced $2500 in monthly license fees for project management software with $12,500 in monthly token usage, does that count?
- @Glittering-Pair5401 (108): I am sure Management will call it a success when it breaks even in 5 months. While spending an additional $13K tokens on other things.
- @deltamoney (104): The problem with all this is you now have to manage absolutely everything yourself. Part of the point was the get rid of individual risk. Now everyone is taking on massive amount of individual employee risk. CEO now has to be a coder, marketer, QA. You name it. Pile that work al…
- @chomstar (77): We worked with 3rd party vendor to make an educational simulation activity that charged $100k per iteration. Took me a week of vibe coding to create the bones of what we needed to replace it with something more customized for our needs. Spent a few months iterating on it on and…
- @dehumles (70): Ive build fleet tracking system, bought advenced telematic devices, installed them in trucks, cheaper version on trailers, coded backend/frontend witu cc. 40ish trucks, 10ish dispatchers/managers using locally in our company, 25k/yearly save. Running in prod since january. + TMS…
独立开发者分享为自己的 Chrome 扩展拿到第一个付费用户 $9 的心情,强调陌生人愿意付钱这件事本身比金额更有意义,并鼓励还没首单的同路人继续坚持
I know it's "just" **$9**, but I can't stop smiling. Over the last few weeks I've been building a Chrome extension for developers. There were days where I wondered if anyone would actually pay for so…
受众观点:①@NetOk7015 引发强烈共鸣:盯着 Stripe 通知发呆十分钟,那些熬夜突然没那么蠢了 ②@ugcfast 提醒关注第一批用户的隐藏价值,最好的客户往往藏在最早付钱的人里 ③@aryanxcreates 提到 organic 获客和不喜欢营销时的挣扎感
展开评论
- @CommunicationDry2611 (26): Congrats win win win
- @NetOk7015 (21): That stranger moment hits different. I still remember my first $9 customer for a tiny side project and just stared at the Stripe notification for ten minutes. Those late nights feel way less stupid when someone actually pulls out their card. Congrats on the Chrome extension, tha…
- @ugcfast (11): find out who they are. your best customers are hiding in your first customers.
- @aryanxcreates (8): Thanks, It was a real grind when everything has to be done organic and who hates marketing.
- @M1jidi (7): Congratulations 🎉🎈
Anthropic 官方发布 Claude Opus 5 公告,用户热议新模型实际体验是否退步以及与 Fable 5 之间的能力关系
https://www.anthropic.com/news/claude-opus-5
受众观点:①@elonthegenerous 关注主观感知退化问题——高赞提问"有没有人感觉 Opus 5 变差了"引发强烈共鸣 ②@FireIre 关注与 Fable 5 的功能对等性——"Fable 5 at Opus prices"直接定位了用户的真实期待 ③@TexasBedouin 关注用量重置时间这一实际使用痛点,与模型能力本身无关
展开评论
- @elonthegenerous (167): Does anyone else feel like Opus 5 has gotten worse?
- @FireIre (99): So this is just basically Fable 5 at Opus prices? Awesome.
- @TexasBedouin (95): WHERE'S MY USAGE RESET!!!
- @xrp_oldie (62): Fable banned? just call it Opus 5.
- @tokenentropy (59): This thing beats Sol by a LOT. holy SHIT
开发者发布 Claude Code skill,可将手写字母照片自动转换为可安装的 TTF 字体文件,本地运行、MIT 协议开源
Wrote my alphabet in a notebook, dragged the photo into Claude Code, said "make my font". Got a TTF I installed in Font Book. The skill drives a deterministic npm CLI (potrace + font assembly). Claud…
受众观点:①@gari692 关注字体渲染实际质量——字母间距、可读性是真实使用场景的核心指标 ②@x32fzw 被情感使用场景打动——"保存母亲手写字迹"这个用例揭示了工具的情感价值维度 ③@Medium-Watch-2782 提供实际效果展示,指出标题字体适用、正文不适用的使用边界
展开评论
- @gari692 (227): What lacks in this presentation is the end effect - how does the text written in that font look. It's easy to mess up the letter spacing, etc. and make it look bad/unreadable.
- @x32fzw (90): Thanks, op, you’ve given me the means to preserve my mums handwriting. I have terrible handwriting. But my mums is lovely. Probably 10 years ago I asked her to write out the alphabet and other bits to make it into a font, but the process was complex and time consuming so never g…
- @QuietPenguinGaming (57): 100%, fonts are surprisingly tricky. Need to see the finished product in action
- @Medium-Watch-2782 (55): the headlines here are in the font from the image: [https://danilo-znamerovszkij.github.io/draw-your-font/](https://danilo-znamerovszkij.github.io/draw-your-font/) this is how it looks like in my VSCode xD https://preview.redd.it/9xt46cbaj5fh1.png?width=1990&format=png&a…
- @Medium-Watch-2782 (47): This is how the one from the photo looks on an actual website: [https://danilo-znamerovszkij.github.io/draw-your-font/](https://danilo-znamerovszkij.github.io/draw-your-font/) I wouldn't use it for body text, but for headlines I think works nice
Reddit 开放讨论帖,假设零营销预算、无受众、无投资的情况下刚上线一个 SaaS 产品,前 30 天如何冷启动,吸引了 97 条各路独立开发者的实战回复
Imagine you're launching a brand-new SaaS today. You have: * $0 marketing budget * No existing audience * No investors * A product that's ready to use **What would your first 30 days look like?** * W…
受众观点:①@Realistic-Refuse9855 强调分发必须从 Day 1 开始而不是上线那天,先谈用户再上线产品 ②@DEATH712712 分享了 SEO 加目录收录的具体路径并直接喊话找 beta 用户 ③@AccordingWind1082 追问最实际的问题:怎么找到已经有这个痛点的人
展开评论
- @Realistic-Refuse9855 (20): After building multiple products that went nowhere, my biggest lesson was that distribution starts on day one, not launch day. If I had $0 today, I'd spend Week 1 talking to potential users, Week 2 improving based on their feedback, and Week 3 manually reaching out to people who…
- @Sidwilliam (8): Guilt trip friends and family into trying your app. Boom u no longer has zero dollar budget
- @Realistic-Refuse9855 (6): That's the hardest part. I'd start where people are already talking about the problem, Reddit, niche communities, LinkedIn, X, and forums. The goal isn't to sell immediately, it's to join conversations where the pain already exists.
- @AccordingWind1082 (5): how to find people who are actively looking for the solution that you are providing? also how do you reach out to people when you are researching?
- @DEATH712712 (5): I want to make it clear that I am not qualified to answer this because I only have 3 users for my app currently, BUT I got those three users from reddit threads like this one, and I think those first few users are the best for feedback, and then do tons of SEO like blog posts, f…
独立开发者推出开源 .gui 文件格式,旨在为 AI 提供可生成、可解析、跨平台的用户界面描述标准,填补 PDF/SVG 之外 UI 领域的空白。
Why? For years we've had open formats for documents (PDF), images (PNG) and graphics (SVG). But user interfaces are still mostly locked inside design tools. And that was fine when humans were creatin…
受众观点:①@elementik4 提出最核心的设计张力:存储布局意图还是绝对坐标,两条路分别面临不同的跨平台一致性挑战 ②@jonalaniz2 等人对标已有方案(XAML、GTK .ui、QML),质疑差异化是否成立 ③@Rough-Mortgage-1024 解释选择 XML 语法的原因——AI 训练数据中 HTML/XML 覆盖率高,对模型更友好
展开评论
- @elementik4 (43): The hard call is whether a .gui file stores layout intent or absolute geometry. Absolute and you have reinvented SVG. Intent and you have to commit to one layout model that iOS, Android and web will each disagree with. Worth seeing how GTK's .ui files and QML settled that.
- @Classic-Grab-2866 (14): I like how you took time to make a decent website and not an ai slop ui website but yeah cool product
- @jonalaniz2 (13): So XAML?
- @Rough-Mortgage-1024 (11): Yeah also the target audience is designers, so a lot of pressure and also I work as full time as a product designer
- @Rough-Mortgage-1024 (10): Yeah, the goal is to store the information that modern UI frameworks actually need like layouts, stacks, grids, spacing, sizing, design tokens, styles, themes/modes, etc. rather than just geometry. I agree if its just geometry then its basically SVG with extra steps. The XML-lik…
r/selfhosted 每周新项目汇总帖,本期包含卡牌收藏管理 Bindarr、端对端加密存储 ExoServe、云成本优化 MCP 工具 nable 等多个开源项目
Welcome to the **New Project Megathread!** This weekly thread is the new official home for sharing your new projects (younger than three months) with the community. To keep the subreddit feed from be…
受众观点:①@getnable 介绍的 nable 吸引关注云成本控制和 FinOps 自托管方案的用户,MCP 集成是核心亮点 ②@theNotoriousJeremy 展示的 Bindarr 引发对实体收藏品数字化管理需求的讨论 ③@Key-Programmer-4144 的 ExoServe 引发对客户端加密、zero-knowledge 存储设计的关注
展开评论
- @theNotoriousJeremy (11): https://preview.redd.it/99mx3zuwy1fh1.png?width=1751&format=png&auto=webp&s=8a921862bae228f70cc85e1fff069e5bb3cfc253 **Project Name:** Bindarr **Repo:** [https://github.com/thenotoriousJeremy/bindarr](https://github.com/thenotoriousJeremy/bindarr) **Description:** Se…
- @Key-Programmer-4144 (5): **Project Name**: ExoServe **Github**: [https://github.com/Ciphranova-LLC/ExoServe](https://github.com/Ciphranova-LLC/ExoServe) **Demo Site**: [https://demo.exoserve.ciphranova.com](https://demo.exoserve.ciphranova.com) **Description**: An end-to-end encrypted file storage solut…
- @getnable (5): **Project Name:** nable **Repo:** [https://github.com/getnable/finopsmcp](https://github.com/getnable/finopsmcp) **Description:** Self-hosted cloud cost tool. Point it at AWS/Azure/GCP/Kubernetes plus \~15 SaaS/AI providers (Datadog, Snowflake, OpenAI, etc.) and it tells you whe…
- @nhvu1988 (3): Hi everyone, whenever my wife ask for, a receipt, a warranty, the manual for some appliance, what we paid for it . I'm stuck trying to remember, or digging through drawers and email to find the document. Spreadsheets didn't cut it. So I built **TiniVault,** a self-hosted app to…
- @Cautious-Hovercraft7 (3): https://preview.redd.it/yqlciy5bx4fh1.png?width=1200&format=png&auto=webp&s=cf72fc70a64dd211473ef25d2579798dd850f84f **Project Name:** Guess the Intro **Link:** [https://github.com/colfin22/intro-quiz](https://github.com/colfin22/intro-quiz) **Description:** A self-h…
独立开发者用 Claude 为其每日赛车游戏 Swervle 实现了导演模式功能,可同时回放所有玩家的路线录像
I built a daily racing game called swervle, it's like wordle but a random route everyday. I added a feature to record the data from everybody's race so I can simulate a lot of cools things like racin…
受众观点:①@observemedia 为项目辩护,高赞评论反映出 Claude 创意使用案例在社区被低估的现象 ②@ClaudeAI-mod-bot 总结显示核心需求是游戏链接和 Safari 延迟的技术反馈 ③@thenorthernwhiteboy 关注具体游戏趣味细节——有玩家开车直接上山的搞笑画面体现了多路回放的社交价值
展开评论
- @observemedia (115): People are idiots here, that was cool man. I watched, was super interesting.
- @Guilty-Prize-3697 (21): Super fun to watch
- @ClaudeAI-mod-bot (18): **TL;DR of the discussion generated automatically after 40 comments.** **The overwhelming consensus is that this is an awesome and creative project.** A few haters tried to rain on the parade but got promptly downvoted into the abyss, with the community rallying to defend OP. Th…
- @AaronMatthews25 (17): thank you!
- @thenorthernwhiteboy (13): I like the car that went straight up the mountain in the beginning lmao. This is awesome man
开发者发布了一个基于频率与时间衰减算法预测下一条 shell 命令的自动补全工具,刻意不使用 LLM,靠轻量算法实现快速可靠的补全
r/selfhosted I made a better zsh autosuggestion, it predicts your next command, no…
受众观点:①安全隐患——@polaroid_kidd 指出工具未遵守 zsh HIST_IGNORE_SPACE 隐私机制,所有命令都进了本地 sqlite db,且命令在进程列表中可被其他用户看到;②非 LLM 设计的价值——@sligor 明确表示厌倦了 shell LLM 补全,这个"不卖 LLM"的设计方向让人耳目一新;③算法细节——@Fantastic_Squirrel96 解释 frecency 是 frequency + recency 的合词,并提到 Markov chain 实现
展开评论
- @polaroid_kidd (96): this looks nice, but there are some security related concerns. 1. It looks like you're not adhering to zsh HIST_IGNORE_SPACE privacy mechanism (commands with a leading space are ommitted from the zsh history). That means that every command ends up in the local sqlite db, even on…
- @sligor (80): The fact that it doesn’t use an LLM for doing the completion is so much refreshing ! And then it is fast and reliable not like the shitty shell llm completion everybody tries to sell me. I will take a look
- @polaroid_kidd (71): I'd love to but currently I have a toddler at home trying to lick all the electrical outlets so free time is in a short supply currently. Sorry :/
- @Fantastic_Squirrel96 (44): frecency combines frequency and recency it’s a term derived from both
- @Fantastic_Squirrel96 (29): Mine uses a combination of zsh history and a small Markov chain algorithm
AI 自动化创业者分享被大型能源钢铁分销商 COO 在销售会议中因没有同规模客户案例而当场质疑失去信任的经历,讨论如何在无 enterprise case study 时进入高端客户
I run an AI automation company and managed to get a meeting with the COO of a big energy and steel distributor. He was pretty invested in the idea until he asked what other clients I've worked with,…
受众观点:①@TangeloObvious2265 认为从 SMB 做起才能打磨产品成熟度,跳级冲大客户是时间黑洞 ②@SirNaples 建议先聚焦已有客户的行业深耕,等有 mid-market 案例再进 enterprise ③@brazil598 强调让 C-level 感受到专家感和 VIP 待遇,并保持一定神秘感
展开评论
- @TangeloObvious2265 (89): You work up to it. It’s not just the clients. It’s the maturity of your solution, your experience, etc that will come from working with SMBs. Trying to land whales can be a huge time sink. They are slow and have high expectations. Sometimes it’s better to work with smaller compa…
- @SirNaples (13): Focus on industries you have clients in, and break into enterprise once you have middle market. If the coo, didn't kill it, the CISO likely would have.
- @PlanFamous4279 (9): That's what a lot of people have been telling me but is there any way to accelerate this process because I know this takes a lot of time to do? I hate dealing with low-ticket customers because they don't show up to meetings or they're always rescheduling or they're very hard to…
- @Finalbossops (9): Asked you for a demo be careful with that they can take your idea and remodel it and make it there own
- @brazil598 (9): Make it very personal. Show how you will help him, that you are an expert, how you will give him the VIP treatment etc etc. If you two have a connection, make that the first priority in every talk. Also just keep a bit of mystery around ‘your other clients’. C level just seeks t…
一款开源自托管的 dead man's switch 软件,支持死亡后自动触发 HTTP 请求、邮件或短信等执行动作
Hello everyone, More than a year ago on this sub, i published my implementation of dead mans switch software. Now im working on a new major release. And i want to use this moment to better understand…
受众观点:①@blubberland01 关注触发条件可靠性——如何区分真正死亡与断网/假阳性,认为这是产品根本性缺陷 ②@cardboard-kansio 关注边缘场景(断网度假导致误触发)的实际危害 ③@Karyo_Ten 关注与传统法律手段(遗嘱/公证)的对比,质疑自动化死亡开关的实际必要性
展开评论
- @blubberland01 (28): The hard part about this is not at all executing the things that should be done. That's the easy part any kid with some scripting knowledge could do. The real issue is, how to determine (and absolutely make sure) you're dead. Your implementation seems to also possibly trigger if…
- @baddaywithacamera (25): Delete my browser cache. ;)
- @cardboard-kansio (13): > Depending on your lifestyle it might be hard to do daily pings, but maybe weekly would work? And then you go on holiday for a fortnight, but halfway through there's a tropical storm which takes down the local mast, and now you don't have any internet. By the time your holid…
- @Karyo_Ten (9): Currently I think the only way is put things in a will and leave it at a notary. But then why use a deadman switch at all.
- @blubberland01 (9): So, why call it *dead man switch*, if it's just a *did not respond in a week switch*? Also what happens if the server goes down or some weird shit happens with timing? If there's things you find so important to be done in case you're dead, those are very likely also things you d…
B2B SaaS 创业者问获取前 10 个付费客户最有效的渠道,47 条回复覆盖 SEO、冷邮件、社区运营等多种方法
Bootstrapping a B2B SaaS, pre-revenue, trying to figure out where to spend my time before throwing money at ads. For those who've been through this stage, which channel actually got you your first 10…
受众观点:①SEO 长期策略——@Mysterious-Crazy-395 靠关键词页面 + 目录收录获得前 100 个试用;②冷邮件批量触达——@johnlocke8 分享买域名、预热邮箱、发十万封冷邮件的系统流程;③渠道 ROI 对比——提问者想知道哪个渠道 worthwhile 且 repeatable
展开评论
- @Mysterious-Crazy-395 (6): I got my first 100 trials and 8 customers all from SEO
- @Mysterious-Crazy-395 (3): Mostly optimised the website and first blogs for targeted keywords. build individual pages for almost all the keywords targeting platforms, features and target audiences. and directory listing. Next is going to start the actual content marketing aka writing blogs on scale
- @Mysterious-Crazy-395 (3): Yes: still doing it
- @johnlocke8 (3): Bro its been said a thoussssssaaaand times 1. Buy 3 domain variations of your domain 2. 2 or 3 emails on each 3. Warm them up on a cold email site there's a ton (I like dayonelead or instantly) 4. Get 100,000 emails matching your ICP, there's a ton of lead gen tools (I like dayo…
- @2121-guy (3): This is the way
社区讨论帖,征集独立开发者们第一次感觉自己副业项目"真的跑起来了"的标志性时刻,46 条评论仍在活跃互动。
Not necessarily making money.Maybe it was your first returning user,someone recommending it without being asked,or just realizing people genuinely cared.What was that moment for you?
受众观点:①@PrecursorLabs 认为最核心的信号是"它解决了我自己的问题",说明社区高度认可 scratch-your-own-itch 的产品思路 ②@SpellengBeeChemp 和 @Basic_Bad6389 强调"自己在用"是比外部数据更可靠的早期验证信号 ③评论区有人讨论阶段性投入应匹配验证程度,说明大家关心如何避免过早过度投入
展开评论
- @PrecursorLabs (9): It solved a problem for me
- @PrecursorLabs (3): Yeah! Honestly now that I think about it there are like definitely levels to it. Your level of investment should match the stage you’re at ya know, I think people easily burn themselves out by having over doing it too early.
- @SpellengBeeChemp (3): because I use it myself
- @Basic_Bad6389 (2): That’s usually the turning point. If it solves your own problem first, there’s a good chance others have the same pain too.The real validation comes when strangers start using it for that exact reason
- @Artexis1 (2): That's a very good reason, one that I share 😀.
开发者被裁员后将前公司内部功能从头重建为独立 SaaS 产品,前公司员工开始来询问,评论区随即展开关于 IP 归属和雇主协议风险的深度讨论
One of the last things I built at work was a small feature. After I got laid off, I kept thinking about it. Eventually I realized the feature was probably more valuable as its own product than where…
受众观点:①@Anti-Hero25 警示 IP 归属风险,指出大多数雇主协议会把员工在职期间的所有创作认定为公司财产 ②@Anti-Hero25 进一步引用非竞争协议极端案例——朋友父亲供职的药企甚至约束其配偶的发明权 ③@eduinvestor 追问未与雇主签约的配偶是否真的受约束,质疑法律边界
展开评论
- @Anti-Hero25 (66): Or are they asking because they're thinking of taking you to court over IP theft? (Since most places will considered anything you make while working for them, their property)
- @Anti-Hero25 (19): Absolutely, especially since you stated at the onset, you simply reproduced exactly what you built for them while you were there. Like coming up with the recipe for Coke.... quitting... then writing down the formula again from memory. No different than if you took a thumb drive…
- @[deleted] (15): [removed]
- @ramtough_63 (12): Slippery slope
- @eduinvestor (11): how is that legal if she didn't sign any contract with her husband's company?
独立开发者分享通过持续 ASO 优化、40 余篇博客和 20 余个 programmatic landing page 将 App Store 收入推向 $1000 的历程,强调学会分发与做好产品同样重要
About a year ago, I came across a Reddit meme that made me realize how impossible it is to find old files, screenshots, and documents once they disappear into your phone. Today I'm about $46 away fro…
受众观点:①@mxlawr 直接追问 Google Search Console 的具体数据,关注 SEO 效果的量化指标 ②@PrizeSpeech5838 分享了每日 50 访客加 30 安装的实际数据作为参照 ③@AmbitiousKid69 认为 2026 年做 App 仍然被低估,解决简单问题的 App 收益可观
展开评论
- @PrizeSpeech5838 (4): Currently i am getting around 50 visitors daily from google , i know its less thats why i am improving my website presence and DA
- @mxlawr (3): Congrats! It's a hard road, I know. One question, what do your google search console stats look like? stats?
- @AmbitiousKid69 (3): TBH building Apps is still underrated in 2026 apps solving simple issues are earning a lot we have to keep this in mind I'm happy for you mate!
- @PrizeSpeech5838 (2): talking about google search console , and daily i get around 30+ new installs.
- @CommunicationDry2611 (2): Congrats win win win win
Anthropic 正式发布 Claude Opus 5,已同步上线 Claude Code,评论区对其安全限制和前代改进情况存在明显分歧。
r/ClaudeAI Opus 5 is out
受众观点:①@Sjeg84 关注 Opus 5 是否已进入 Claude Code,影响开发者日常工具链 ②@Temporary_Idea8880 批评其安全限制(safeguards)和 Fable 版本一样,认为对特定用例毫无价值 ③@Hasjojo 关心 Opus 4.8 的"笨拙感"是否在新版本中修复
展开评论
- @Temporary_Idea8880 (23): DOA. It still has the same safeguards as Fable. Useless. https://preview.redd.it/wahpxavjk7fh1.png?width=2996&format=png&auto=webp&s=3ec8b9b2208db6fb6b8e674542aedb61b6d7076f
- @Sjeg84 (20): already in claude code as well.
- @Eugene-Coolguy (16): How is that DOA?? Get a grip
- @Hasjojo (13): I hope they fixed opus 4.8 clumsiness 😆😆
- @angry_queef_master (11): what is fancy-webhook about
用户发现 Claude 在简单数字大小比较上出错,经用户纠正后 Anthropic 快速更新了网站说明。
r/ClaudeAI 53.4 > 53.5?
受众观点:①@LEERROOOOYYYYY 支持用户纠正模型错误,肯定这种批判性使用方式 ②@Temporary_Idea8880 注意到 Anthropic 快速更新网站,引发对官方响应速度的讨论 ③@trpmanhiro 借机提醒要对 AI 输出进行人工复核,反映开发者对 LLM 可靠性的普遍担忧
展开评论
- @LEERROOOOYYYYY (140): You're right to push back on that
- @Temporary_Idea8880 (48): They updated it on the website so fast lmao
- @BUYMEBONESTOORM (17): Why don’t you take a break? You’ve been at this a while.
- @trpmanhiro (17): Claude is AI and can make mistakes. Please double-check responses.
- @MahaSejahtera (14): It increase the chance of they read reddit theory
独立开发者将 SpacePlanner 免费空间规划工具在一个月内做到 200 日活,现在面临是否要投入 3D 功能的产品方向决策,在社区征求建议。
A month ago, I decided to put all my energy into SpacePlanner and try to turn it into the #1 free, no-sign-up space planning app that actually gives people real value. Since then, it has grown to mor…
受众观点:①@Designer_Reaction551 建议坚守 2D 做到极致,认为手动处理 3D 模型是消耗时间而不推动核心指标的典型陷阱 ②@jantursky 提出具体验证方法:让 10 个用户同屏录制、追踪完成/导出率而非仅看创建数 ③@Super_Change5388 作为实际用户反馈了具体 bug(屋顶导出问题),说明产品尚存基础体验问题
展开评论
- @Super_Change5388 (5): Insane progress! Been using it for few days now, but i had issues with exporting the roof facade, do you support that? Very cool project! Keep it up
- @Designer_Reaction551 (5): if 2D already works and people are actually using it, 200 daily users is not nothing, I'd hold off on 3D for now. the manual model hunting and connecting thing you described is exactly the kind of work that eats months without moving the core metric. get the 2D editor to where w…
- @commandedbydemons (4): Seeing this for the first time, low key one of the best projects I've seen here.
- @jantursky (3): 200 daily users is enough signal to stop guessing. I’d pause new 3D assets and watch 10 users build the same basic room over screen share; note every undo, hesitation, and correction. Fix the top three repeats first. Oh, and track completed/exported plans, not just plans created…
- @mozdamalosutra (3): where was this when I was moving into my apartment last fall....
开发者询问自托管 Matomo 在低配 VPS 上的实际资源需求,并引发 Matomo 与 Umami、Plausible 三款隐私友好分析工具的对比讨论
I'm looking into Google Analytics alternatives.. Especially free self-hosted solutions. Right now I'm leaning on Matomo, but I'm curious about it's server reqs. They say 2 CPUs, 2GB RAM and 50GB stor…
受众观点:①配置下限——@TheEfficaciousTodayS 实测 1GB RAM 的 VPS 可以跑 Matomo,官方最低配置是针对 300K+ 月访问量的;②工具切换——@retro-mehl 从 Matomo 换到 Umami 再换到 Plausible;③工具对比——@aharonduarte 做了 Umami(纯自托管)vs Plausible(有托管付费版)的对比
展开评论
- @TheEfficaciousTodayS (4): I ran Matomo on a 1GB RAM VPS for about a year tracking three low-traffic sites and it was fine. The database is what gets hungry over time so you'll want to set the data retention to auto-purge raw logs after a few months. Default is forever which will absolutely wreck a small…
- @retro-mehl (3): Had no problem at all, it doesn't need much CPU. But personally I switched to umami and later to plausible.
- @ollierwoodman (2): How would you compare umami and plausible?
- @aharonduarte (2): Broad strokes: Umami is fully self-hosted only, dead simple dashboard, and you own 100% of the setup — good if you want full control and don't mind maintaining it. Plausible has both a hosted (paid) and self-hosted option, tends to have a slightly more polished report view, and…
- @thy_bucket_for_thee (2): For the type of analytics you want I'd look into something like Umami instead: https://github.com/umami-software/umami It does require a bit more ram to bootstrap itself once initially starting (2gigs of ram), but afterwards it never really peaks around too much for me for modes…
独立开发者用 AI 辅助一个周末花 2 美元上线了模拟 memecoin 交易的浏览器游戏,请社区评估留存机制和目标受众定位策略
It's a trading simulator built around memecoin launches. You get play money and a feed of coins to trade, and most of them are designed to collapse. The point is that losing fake money is a cheaper w…
受众观点:①@BuildingSolo9 指出产品本质是游戏非工具,明确建议放弃加密社区、转攻游戏社区 ②@Unlucky_Teacher2485 关注留存机制缺失,提出 daily seed/streak/得分排行等具体改进方向 ③@artemmakes 关注冷启动用户测试方法——建议实际观察一个真实用户试玩而非依赖自我判断
展开评论
- @artemmakes (3): best test is to watch one person play it with no intro from you, if they get what to do in the first ten seconds the thirty is fine. the part where they hesitate is your real first screen problem.
- @BuildingSolo9 (2): your retention question already answers the distribution one. nobody comes back to a memecoin sim to learn a pattern, they come back to watch a coin rug and beat a score. that's a game mechanic, not a tool. so it's a game in a crypto skin, take it to the idle/trading-sim crowd.…
- @xiaoxiao_321 (2): Shipping this in a weekend and already separating first-play curiosity from repeat use is a good call. I’ve had people enjoy a first try and still disappear, so I think the second session tells you more than whether the onboarding is understandable. What happened after the first…
- @Unlucky_Teacher2485 (2): I think the first 30 seconds are clear enough for a first try. The weaker part is the reason to come back. Right now the loop sounds like learn once, leave. I would test one simple replay hook right on the first screen: daily seed, streak, score to beat, or a short post-round br…
- @Hollow-Harbor-8260 (2): Finding yourself stuck between two very different audiences is a difficult place to build. You did a beautiful job keeping your costs low while you figure out where this belongs.
有用户报告 Claude 视觉识别能力大幅提升,但也有人指出颜色识别仍不如 Gemini 准确。
r/ClaudeAI Claude isn't partially "blind" anymore
受众观点:①@Charming_Ad_4765 等人关注 Claude 视觉推理能力对比其他模型是否真正领先 ②@MysticGoddess27 提出颜色识别等基础视觉任务仍有失败案例,质疑改进是否全面 ③评论者关心不同版本(Opus 4.8 vs 4.6)之间的实际能力差距
展开评论
- @Charming_Ad_4765 (36): its vision capabilities far exceeds most nowadays tbh, visual reasoning etc. glad to see the improvements
- @anandesh-sharma (32): This feature will make the claude design more powerful 🙌 Nice.
- @MysticGoddess27 (17): Eh I disagree. I gave it a screenshot of a game UI and asked it to separate completed quests (green), active quests (yellow), and inactive quests (red) and it couldn't get it right. ChatGPT also failed but Gemini didn't. What good is it if it can't tell the difference between 3…
- @Charming_Ad_4765 (6): I think this should work in most cases, are you using Opus 4.8 minimally? I remember Opus 4.6 was horrid at understanding visuals , seems to have improved significantly in my usecases
- @gradient8 (3): Bruh ChatGPT has been doing this since fucking o1
一位管理每月 310 亿事件规模基础设施的工程师,在副业小游戏网站获得 50 名访客时感到比工作更兴奋,分享这种心理落差体验。
My day job is managing infrastructure that handles around 31B events, 21B API calls and 150B database queries a month. Big scary numbers, on-call rotations, the whole thing. At night I work on my lit…
受众观点:①@Minimum_Hour519 和 @Electronic-Rate-6208 高度认同这种感受,说明独立开发者群体对"小数字反而更有意义"有强烈共鸣 ②@syou0 来自日本,回应作者庆祝第一个日本用户的细节,印证了创作者对地理多样性的珍视 ③@Owmelicious 询问用的是哪个统计服务,说明社区对独立开发者选用轻量分析工具有实际兴趣
展开评论
- @Minimum_Hour519 (8): one of us!!!
- @n0zz (6): https://preview.redd.it/g5b4u67xb7fh1.png?width=688&format=png&auto=webp&s=c255209fa566840ff944dfefc92f4181b18bbde8 Proof 😀
- @Electronic-Rate-6208 (3): yep, 50 users hits different
- @Owmelicious (2): How do you track this? Which service is this?
- @syou0 (2): I’m Japanese, so I’m happy to hear that your first player from Japan made you celebrate! I also understand this feeling. Big numbers at work can feel normal, but even one person using your own project feels very special. 50 visitors and 300 game sessions in one month is a great…
独立开发者分享产品获得 8 个真实用户的里程碑心情,评论区演变为冷启动获客方法的讨论
happy about it, it wasn’t long ago that it didn’t exist and I had no users
受众观点:①获客方法——@JustUse28 和 @AdAcrobatic8861 都在问怎么拿到这 8 个用户;②用户激活问题——@EffectiveStrong884 自己也有 6 个 beta 用户,反映注册后不使用的激活难题;③早期里程碑的心理价值——@Quiet-philologist25 强调 8 个真实用户比空仪表盘有意义得多
展开评论
- @Quiet-philologist25 (3): 8 real users is way more exciting than staring at an analytics dashboard with zero. Congrats! Hope they stick around and tell you what's working (and what's not).
- @EffectiveStrong884 (2): Congratulations, Even I launched the beta few days back and has 6 users. The only issue i have is that once signed up people are not using it. How is it going it for you? How do you motivate people to use it after signing up?
- @JustUse28 (1): what did you do to get them?
- @AdAcrobatic8861 (1): Good job! What's your product and how'd you get the initial users?
- @authority_joel (1): Congratulations, a win is a win
非专职开发者分享通过调试 AI 生成代码来锻炼阅读陌生代码能力的训练方法,评论区对这个方法的有效性产生争论
My day job isn't software development related, so I don't read a whole lot of code. The only code I read is the one I write and that's a problem. Since I'm very interested in open source, I figured I…
受众观点:①训练方法的局限——@Tarazena 指出这更像是调试训练而非真正的代码阅读,大局观比 bug 修复更重要;②更有效的替代方法——@Tarazena 推荐阅读 GitHub code review 了解审查者的思考逻辑;③开源贡献入门路径——@its_artur1 建议先找已关闭 issue,复现旧 bug,再写测试
展开评论
- @Tarazena (6): I think you are training to debug better rather than actually reading code, you can read multiple lines of code and understand what’s going on, but larger picture is usually more important than simple but issue fixing. My recommendation is not to do what you are doing (although…
- @Tarazena (4): No need for that, dig into https://refactoring.guru/design-patterns
- @Tarazena (3): Understanding code, patterns, etc would make you better at detecting bad code (or calling LLMs BS) rather than fixing, e.g. why we are not following abstract factory pattern or inheritance rather than having 10 different factories? Or why are we using in memory cache if we have…
- @Stevious7 (2): Not all heros wear capes 🫡
- @its_artur1 (2): That is a useful debugging exercise, but generated code removes two of the hardest parts of real open-source work: recovering intent and respecting constraints you did not choose. I would add a second exercise alongside it. Pick a small closed issue in a project you use. Reprodu…
讨论 AI agent 代替用户完成任务时是否应当持有永久支付凭证,还是应该为每个任务生成短期凭证
This is something that's been crossing my mind for a bit so I thought I'd ask here. I've been wondering how payments are supposed to work like as more SaaS products start letting AI agents complete t…
受众观点:①安全风险——@thetanobserver 指出永久凭证只在极少数场景有意义,大多数 agent 工作流有明确目标和预算上限;②工程实践——@Itchy_Ground_2249 强调短期凭证比单一永久支付方式跨多个工作流更易理清逻辑;③变更成本——@No-Echo-6522 提到换掉一个 agent 时不必担心哪些工作流还在共享同一个支付方式
展开评论
- @thetanobserver (3): I think permanent credentials are only going to make sense in a small number of cases. A majority of agent workflows have a clear objective and budget so to me it seems more logical to issue payment permissions that only exist for that task cause even outside of security it make…
- @Itchy_Ground_2249 (2): Default should be short lived credentials that are generated for a specific task since it's much easier to reason about than trying to manage one permanent payment method across dozens of workflows.
- @No-Echo-6522 (2): I completely agree since it makes changes a lot less risky. If you update or replace an agent then you don't have to wonder what else still has access to the same payment method. You can retire one credential and issue another that's scoped to the new workflow without affecting…
- @Itchy_Ground_2249 (2): It also makes credential rotation much less disruptive because you're not untangling a payment method that's ended up being shared across multiple workflows.
- @KingBardan (2): "You're right, I made a mistake here, unfortunately , the transaction is not available. here's your 200k bill. Please let me know if you need anything else" --- gpt idk
开发者将废弃的 Azure IoT 摄像头用 Rust 改造为支持 MCP 协议的 AI agent 服务器,可运行本地视觉模型并实现双向语音对话
I have 5 Azure IoT Starter Kit cameras, the vendor (Altek) stopped updating their firmware and the last firmware doesn't work with Azure anymore. I emailed Altek to help me with the firmware or any S…
受众观点:①@Aggravating-Suit205 关注如何用 AI 探测旧设备固件,寻求可参考的工具链方案 ②@No_Corner805 关注多模态 agent 模型组合方案(GLM、Kimi K3、Fable 5 协同) ③@ahstanin 提到使用 opencode 加 GLM-5.2 作为开发辅助工具
展开评论
- @Aggravating-Suit205 (2): Wow, never thought of using AI to probe my older bricked cameras. I have some old Amazon cameras that became bricked a couple years ago. Were you using something like Hermes to probe it?
- @Dendritic_Silver (2): Cool project Nicely done.🫡
- @asimovs-auditor (1): Expand the replies to this comment to learn how AI was used in this post/project.
- @ahstanin (1): opencode with GLM-5.2
- @No_Corner805 (1): Do you have experience iwth mulimodel agents? Looking for something with GLM, Kimi K3, and ChatGPT. With Fable 5 as the manager/proof reader.
开发者用 Claude Code 辅助构建了 Bento,一个单 HTML 文件实现的离线幻灯片工具,支持 CRDT 协作和加密中继,登上 Show HN 第一名。
Over the past few months, our team has been building more and more slidedecks using web frontend technologies with coding harnesses like Claude Code, but a common complaint is to make even small edit…
受众观点:①@BP041 和 @starfallg 深入讨论 CRDT 实现细节和单文件架构的工程边界,说明技术社区对"单文件复杂应用"的可行性有强烈兴趣 ②@ItaySela 指出 ECDSA 签名与可编辑 JSON 之间的安全张力,以及 File System Access API 在 Safari 上的限制 ③@FigZestyclose7787 已经 fork 并提交了 PR,说明项目有真实的开源社区参与度
展开评论
- @starfallg (3): Hi, I'm the creator of Bento. Just wanted to share a bit more about how I created it beyond what's in Github. The file contains more or less two sections. There is a plain block of JSON near the top of the file which is the slide data. You can read, grep, or point a harness at i…
- @starfallg (2): Yeah, the signature is on the runtime, and the update release is the runtime with the default slidedeck. There's more info in docs/architecture.md. And yup, Chrome does write in-place which Safari doensn't support. But the idea is that you don't need to as if you use mobile Safa…
- @BP041 (2): Congrats on the HN wave! Packing edit+view+collab into one HTML file is genuinely wild — did you use CRDTs for the real-time part, or something custom? Single-file apps at this complexity usually hit DOM limits fast, curious how you kept it clean.
- @ItaySela (1): the tension i keep turning over is the greppable json block versus the ecdsa signing. you're inviting people to point a harness at the json and edit it directly, but if every update is signed, doesn't a raw edit outside the app fail verification on next load, or is the sig only…
- @FigZestyclose7787 (1): Amazing project! I had a rudimentary version of something like a slides editor self contained in one Html, but nothing even close to what you've built! I've added some gradients and some other tiny ideas based on my previous project (with AI of course, let's not kid ourselves) a…
前端开发者整理六年经验写成 93 条网页设计实战建议,涵盖色彩系统、排版、表单、动画与无障碍设计
I've been working as a developer for the last 6 years. I always struggled to create something that looks good. I would usually just try copying stuff from Dribbble or Behance. But over the years, I'v…
受众观点:①动画的取舍——@jhartikainen 强烈反对 fade-in 滚动动画和数字递增动画,认为分散注意力;②动画的适配——@ichsagedir 建议配合 prefers-reduced-motion media query 让用户自选;③争论性质——讨论偏向主观体验,没有数据支撑
展开评论
- @jhartikainen (9): Good advice in general except couple of those animation points: > Try animating your sections to "fade in" when the user scrolls down to them. Please don't. This is distracting and annoying when you're trying to quickly scroll. > Animate statistics and numerical data to in…
- @ichsagedir (1): Do it, but with css scroll effects and only if the user wants it ( prefer reduced motion media query) Edit: I'm sure I got the name of three media query wrong, but it tells you if the user wants no animations.
- @jhartikainen (1): Do you want the page to be distracting and annoying even once?
- @LivingAsAMean (1): You're speaking as though your subjective opinion is fact. I'm talking about a situation where some people don't mind it, and for the people who do, you can reduce the impact.
- @jhartikainen (1): Yeah I mean if the user visits the page multiple times, not showing it on the second visit is an improvement - but the impact of the effect is worst the first time you visit the page, because that's when you're trying to quickly look over the content. If you visit a second time,…
探讨软件工厂模式失败原因的文章,评论区引发关于 DSL 驱动开发和声明式工厂架构的讨论
r/webdev Why Software Factories Fail
受众观点:①概念混淆——@atlas__free 质疑文章是否只是在换词描述 spec-driven development;②架构哲学——@originalchronoguy 认为 software factory 核心是一个引擎驱动多种输出,举了 Canva/Zapier/Figma 的例子;③AI 是否必需——@originalchronoguy 强调不依赖 AI 也能构建 factory
展开评论
- @atlas__free (1): Did this article just describe the spec driven development process with lots of words and braking the project into manageable slices?
- @daddieslittleson (0): They have their own problems idk too
- @originalchronoguy (0): This is a really bad take of "software factory" that I can give a 3 hour rebuttal/ted talk. With clear examples that exist in the wild. The word factory itself confuses people - juniors think of Factory Design Pattern (GoF), Stakeholders think of organizational process. I have a…
开发者分享用 SEO 博客在两个月从零 DR 起步,靠一篇长尾博客完成不到 5 分钟的首单转化全流程
I have been working on my new site (0 DR) SEO for 2 months. For First 1.5 months, it was pretty flat around 10 impressions a day and at the end of 2nd month, my DR reached 30 (all organic btw) and It…
受众观点:①SEO 时间线——@Famous-Stress-9489 追问 2 个月达到 30 DR 具体做了什么;②内容转化路径——@NoRooster4259 注意到 blog 答对了问题导致用户不需要探索 homepage;③下一步策略——@NoRooster4259 追问是继续双倍 SEO 还是先从用户反馈迭代产品
展开评论
- @NoRooster4259 (2): Congrats on getting ur 1st customer through SEO. What stood out to me was not just the conversion, but the journey itself. Your blog seems to have answered the customer's question so well that they barely needed to explore your homepage before signing up and paying. That tells m…
- @DEATHKNELL321 (2): Thanks, SEO is the main distribution channel but I am also focussing on social media as well
- @Famous-Stress-9489 (2): What exactly did you do to reach 30 DR in 3 months How many blogs? How many backlink?
- @NoRooster4259 (1): Thats great . Focusing on social media is also a smart way to build trust over time. As u start getting user feedback and testimonials , u can share them to build credibility with ur target audience. You can also share your journey, product updates , and lessons you are learning…
- @DEATHKNELL321 (1): Tools, Chrome extension, directory listing so on. Around 80-90 blogs and 30-60 backlinks
开发者发现广告网络、本地数据库、Google Analytics 和 Cloudflare 对同一网站页面浏览量的统计数字差异高达一倍以上,询问哪个最可信
I'm seeing pretty big discrepancies in analytics reports from different sources. I store the value locally to show my site users, so I'm looking for the most honest number. Yesterday, my ad network s…
受众观点:①测量层级差异——@magenta_placenta 拆解了广告网络、本地 DB、GA、Cloudflare 四者各自测量什么;②使用场景决定选择——@magenta_placenta 建议收入用广告网络、真实人类流量用 Cloudflare、用户行为用 GA;③本地计数的优先级——@clearlight2025 和 @Irythros 都倾向于服务端记录最准
展开评论
- @magenta_placenta (2): There isn't a single "honest" number because each tool is measuring a fundamentally different layer of your stack and applying its own definition of what counts as a "pageview." * The ad network fires when an ad placement successfully requests an ad payload. * Your database incr…
- @griez0777 (1): I always use where the webiste is deployed, like Cloudflare, Vercel, etc
- @RefrigeratorStrict13 (1): from my experience... matomo
- @clearlight2025 (1): A lot of your ads and 3rd party analytics scripts may be blocked. The most accurate number will be from analytics data stored by your local server.
- @Irythros (1): Assuming SSR, the most accurate is whatever touches your servers. Tracking server-side requests is guaranteed to be accurate because either your server received the event (and it can be tracked) or it didnt receive the event (and the user was unable to connect.) Analytics packag…
开发者尝试用 esbuild 对仅使用 Firebase Auth 的项目做 tree shaking 和压缩,发现效果不理想,讨论原因和解决方法
I'm building a web app that uses Firebase Auth. I don't use any other Firebase features. The current version of Firebase is over 600kb. [The Firebase Auth package on NPM](https://www.npmjs.com/packag…
受众观点:①根本原因——@Prestigious-Bank2145 指出需要用模块化导入(firebase/app, firebase/auth)而不是顶层包;②实际包大小——@tom_in_zurich 建议先用 esbuild --metafile 分析再看 gzip 后数字;③UI wrapper 的隐患——@tom_in_zurich 提到 firebase-oss/ui-react 可能引入了更多 SDK 模块
展开评论
- @Tarazena (1): https://firebase.google.com/docs/web/modular-upgrade
- @Prestigious-Bank2145 (1): Correct, script is not enaph, modular imports like: import { initializeApp } from 'firebase/app'; import { getAuth } from 'firebase/auth';
- @Practical_Video_3330 (1): is the minified version still that big like damn
- @tom_in_zurich (1): esbuild's fine here, I wouldn't reach for webpack over this. First thing: that 600kb is almost certainly the uncompressed size. What actually ships is the gzip/brotli number and it's a lot smaller, so measure the compressed output before you panic. Run esbuild with --metafile an…
- @Outrageous-Sea-9256 (0): You're on the right track using `--minify` and `--tree-shaking=true`. However, Firebase SDKs often have a large footprint due to their feature-rich nature and various dependencies. The official documentation suggests relying on your build tool for tree-shaking and minification.…
开发者因为忘记在 CI/CD 流水线设置 HTTP_PROXY 环境变量导致连接失败,追问为何不用网络命名空间做透明代理而要依赖应用层环境变量
My CI/CD pipeline on a self-hosted GitLab instance failed to connect to external hosts since I forgot to set the HTTP\_PROXY, HTTPS\_PROXY and NO\_PROXY env variables. Here's a blog post on the issue…
受众观点:①env var 约定的历史原因——@vietbaoa4htk 指出 env var 规范比网络命名空间早二十年,且不需要 root 权限;②netns 的适用场景——@Horror-Cause9345 说 netns 适合不信任应用或需要强制代理的场景;③TLS 透明代理的限制——@vietbaoa4htk 提到透明重定向会破坏 TLS,除非安装 MITM CA
展开评论
- @spellcasterGG (5): Short answer is complexity. Networking can be quite confusing, and it's very easy to break stuff. It all comes down to the KISS (keep it simple, stupid) design philosophy. Env variables are just easier. Another thing is dependencies. ENV variables are built into Linux, while net…
- @vietbaoa4htk (3): the env var convention predates network namespaces by about two decades and works unprivileged on every os, while netns is linux only and needs cap_net_admin. transparent redirect also breaks tls unless you install a mitm ca, whereas http_proxy lets the client issue a proper con…
- @Beatsu (1): Wow, that really made it all clear. Thanks for the great and concise answer!
- @Horror-Cause9345 (1): Your assumption is basically right, and the reason both exist is that they solve two different problems that people conflate. HTTP_PROXY and friends are cooperative. The app has to choose to read them, which is why NO_PROXY parsing is such a mess and why your GitLab job failed s…
- @Outrageous-Sea-9256 (0): The use of environment variables like HTTP_PROXY and HTTPS_PROXY is a common practice in CI/CD pipelines to manage proxy settings. This approach allows different applications and tools within the pipeline to easily configure their behavior according to the network environment. T…
有人建议 SaaS 平台应效仿某平台主动砍掉部分产品或功能,评论区顺带吐槽 AI 工具选择过多导致还需要专门工具来路由请求
r/SaaS More platforms need to drop services like this
受众观点:①产品策略——@Particular_Rent2023 指出行业常态:一家砍掉某功能,其他平台先观望反应再决定是否跟进;②AI工具选择疲劳——@wary_ducking 吐槽 AI 工具太多连路由都要专门工具;③模型发布频率——@TaleRoyal8733 感叹能少花时间跑 benchmark 就好了
展开评论
- @Particular_Rent2023 (1): a tale as old as time. One company drops a product others watch how its received and if they like the general opinion they copy it
- @Salt_Ingenuity_7771 (1): they be dropping new products left and right
- @wary_ducking (1): SaaS stack is getting weird when you need a tool just to decide which AI tool should answer the request
- @TaleRoyal8733 (1): The less time spent benchmarking every new model release the better lol.
开发者构建了编译器 torchwright,将普通 Python 定义的计算图直接编译为 Phi-3 架构的 transformer 权重,无需任何训练,产出的 checkpoint 可直接用 HuggingFace 加载
I've been chasing the question of what algorithms a transformer can actually express -- separate from what it can learn. So I built a compiler: define a computation graph in ordinary Python, and it p…
受众观点:①AnOnlineHandle 提出冻结确定性工具权重插入网络的构想,类似 MoE 激活路径 ②Shehao 关注学习路由与确定性算子分离后如何量化 tradeoff ③brainsig 认为这个工作有学术发表价值
展开评论
- @AnOnlineHandle (4): I've sometimes wondered if injecting frozen "tools" into a network (the weights don't actually exist, they can just be called if dynamically activated like MoE paths) could significantly improve networks, e.g. either allowing static complex computations which don't need to be re…
- @Shehao (2): That boundary is the interesting bit; separating learned routing from deterministic ops would make the tradeoff much easier to measure.
- @brainsig (2): I find this really interesting, as I wanted to achieve something similar. Good work! Maybe you should consider writing a scientific communication or a tech report describing it.
AutoDev Studio 是一个开源多 agent 软件开发流水线,通过预构建代码库静态分析索引和本地 embedding 缓存,将同等任务的 AI 编程成本比冷启动 claude 降低 7%-75%
Built an open-source AI coding agent that was 7%–75% cheaper than a cold "claude -p" run on 6/6 well-localized tasks across repositories up to ~82k LOC. The biggest difference: - Cold agent: $6.83, 2…
受众观点:①AI agent 开发者关注 PM→Dev→QA→Reviewer 多角色分工的 pipeline 架构设计 ②独立开发者关注跑完整 pipeline 与单次 claude 调用的实际成本对比 ③关注低成本方案的开发者注意 Groq 免费 tier 加本地 embedding 的离线组合
展开评论
- _无评论_
Fedica 2.0 是强调灵活无功能门控的跨平台社交媒体发布工具,支持 Pixelfed 和 Eurosky 等新兴平台及 BlueSky markdown 格式
Fedica
受众观点:①@Hootsuite(评论者)指出 Fedica 核心价值是 no gatekeeping on basic features,受众的核心痛点是竞品的功能门控策略;②@Fedica(开发者回复)强调支持 Pixelfed 和 Eurosky 等边缘平台,受众中有大量使用去中心化或新兴平台的内容创作者;③第二条 @Hootsuite 评论提到 BlueSky 支持 markdown 让链接展示更专业,受众对内容格式和呈现细节有具体需求
展开评论
- @Overview (0): * [Launches4](/products/fedica#launches) * [Reviews32](/products/fedica/reviews) * [Alternatives](/products/fedica/alternatives) * [Customers](/products/fedica/customers) * [Forum](/p/fedica) * [Team](/products/fedica/makers) * More This is the 4th launch from Fedica. [View more…
- @Hootsuite (0): > I actually still use a few other scheduling platforms alongside Fedica, which makes it easy to compare them directly. The magic of Fedica is how open and flexible it feels. Many of the other platforms come with real restrictions and heavy gatekeeping on basic features, which m…
- @Fedica (0): Thanks so much for the feedback Willem! Upvote Report Share 1d ago [Kera Damo - Visual Artist](/@keradamo) •[1 review](/@keradamo/reviews) #### What's great user-friendly interface (8)scheduling posts (11) I've tested not all social media platforms, but a lot. And always, always…
- @Hootsuite (0): I chose Fedica because of the speed with which they adapt to the requirements of new platforms. They actually accepts EuroSky over BlueSky, which is a big plus for me. Also Pixelfed is amazing. I can even tag people and they show up while creating the post. And than the markdown…
- @Fedica (0): Hi Kara, Thanks so much for your feedback. It's very helpful to know that you like our integrations for emerging platforms like Pixelfed and Eurosky along with traditional platforms! Also sorry to hear that the calendar view is bugging you. We will try to improve on that one! I…
Pushary 将 AI agent 运行时权限请求推送到手机锁屏,支持 Claude Code、Codex、Cursor、Gemini CLI、Hermes 等主流编程 agent,让用户离开电脑时也能及时审批
Pushary
受众观点:①@Pushary(maker)详细讨论 auto-approve policies 和锁屏显示 code diff 的设计细节,受众核心关注审批行为的信息充分性和安全边界;②@Pushary(maker)探讨 agent proposes human ratifies 的人机分工以及跨 run 的 invariants 守卫栏设计,受众在思考 AI agent 人机协作的长期架构;③产品覆盖 Claude Code/Codex/Cursor/Gemini CLI/Hermes 全套工具,说明受众有复杂的多工具混用工作流对统一收件箱有强需求
展开评论
- @Productivity (0): • [Code Review Tools](/categories/code-review-tools) • [AI Agents](/categories/ai-agents) Your AI agents stop when they need a permission or an answer. Pushary puts that yes or no on your lock screen, so the work keeps moving while you are away. New in this launch: native iPhone…
- @Pushary (0): Maker 📌 Hey PH 👋 Aadil here. Pushary is the yes button for your AI agents. Their permission prompts and questions land on your phone's lock screen, you tap, the run keeps moving. The problem You hand an agent a 40-minute job and walk away. Two minutes in, it stops to ask permiss…
- @Pushary (0): Maker [@thys\_beesman](https://www.producthunt.com/@thys%5Fbeesman) boom, I love it when agent nerds like me dial in. We have a clean reasoning process mentioned right when you tap on the notification, showing a code diff. Moreover you can specifically define which permissions t…
- @Pushary (0): Maker [@jernej\_jan\_kocica](https://www.producthunt.com/@jernej%5Fjan%5Fkocica) Jernej, this is the best version of this critique I've seen, so I'll answer both halves straight. On intent: you're closer to having it than the lock-screen framing suggests. Every approval already…
- @Pushary (0): Maker [@jernej\_jan\_kocica](https://www.producthunt.com/@jernej%5Fjan%5Fkocica) Jernej, this is the right shape, and the contract framing is what unlocks it. On your actual question, I'd land on agent proposes, human ratifies, but with a split that keeps the ratify from dying o…
Fluree AI 是一个为所有 AI agent 和应用提供统一可信上下文的数据智能层,基于知识图谱数据库,对每次请求都做权限校验并返回可引用的可溯源答案
Fluree
受众观点:①AI agent 开发者关注如何为 agent 提供一致可信的数据上下文 ②安全架构关注者关注权限作为数据存储在查询引擎内部执行的安全模型 ③有实时数据场景的开发者关注知识图谱与流处理工具的边界
展开评论
- @Overview (0): * [Launches3](/products/fluree#launches) * [Reviews](/products/fluree/reviews) * [Alternatives](/products/fluree/alternatives) * [Built with](/products/fluree/built-with) * [Team](/products/fluree/makers) * [Awards](/products/fluree/awards) * More This is the 3rd launch from Flu…
- @Fluree (0): Maker 📌 Hey Product Hunt — Brian here, CEO of Fluree. **The backstory:**we spent years building governed, verifiable graph data infrastructure for enterprises — provenance, permissions, cryptographic audit trails, the unglamorous stuff. Then LLMs arrived, and suddenly the _entir…
- @Fluree (0): Maker [@nancy\_philip](https://www.producthunt.com/@nancy%5Fphilip) Thanks for the question! [Fluree's open source Knowledge Graph DB](https://labs.flur.ee) which [Fluree AI](https://flur.ee/solo) sits on is genuinely the fastest knowledge graph database: <https://github.com/flu…
- @Fluree (0): Maker [@sulemna\_ola](https://www.producthunt.com/@sulemna%5Fola) We focused on making the [fastest core knowledge graph](https://github.com/fluree/benchmark-db), but yes fine grained security and reasoning over semantics takes genuine work and CPU cycles. Our foundational perfo…
- @Fluree (0): Maker [@sulemna\_ola](https://www.producthunt.com/@sulemna%5Fola) great question. Brian covered the performance tradeoffs -- here's what you actually _get_ by putting security at the data layer instead of the application layer. The key idea: policies in Fluree are stored as data…
Firecrawl 推出改进版搜索 API,通过自定义相关性模型对每个段落打分筛选,专为 AI agent 去除检索结果中的噪音,减少 token 消耗并提升答案准确性
Firecrawl
受众观点:①AI agent 开发者关注如何减少检索噪音、降低每次调用的 token 成本 ②RAG 系统构建者关注 excerpt-level 相关性模型对答案准确率的实际提升 ③jernej_jan_kocica 指出更深层问题:最关键的页面可能根本没进候选集,recall 才是真正瓶颈
展开评论
- @Overview (0): * [Launches10](/products/extract-by-firecrawl/launches) * [Reviews14](/products/extract-by-firecrawl/reviews) * [Alternatives](/products/extract-by-firecrawl/alternatives) * [Customers](/products/extract-by-firecrawl/customers) * [Built with](/products/extract-by-firecrawl/built…
- @Tab (0): •[7 reviews](/@vivek%5Fbezawada2/reviews) I use Firecrawl to scrape a website to create a blog design similar to the website. This would've been an entire startup few years ago just scraping a website reliably. Now, it's just an API call away! Helpful Share Report 340 views1yr a…
- @Firecrawl (0): Maker 📌 Hey Product Hunt 👋 Eric, Caleb, and Nick from Firecrawl here. AI agents rely on search tools to find answers on the web, but those answers are often buried in noise. Navigation, boilerplate, and tangents wrap around the excerpt that matters, so an agent processes far mor…
- @Firecrawl (0): Maker [@reda\_roqai\_chaoui](https://www.producthunt.com/@reda%5Froqai%5Fchaoui) This is so true! Retrieval is everything Upvote Report Share 7h ago [](/@jernej%5Fjan%5Fkocica) [Jernej Jan Kočica](/@jernej%5Fjan%5Fkocica) The excerpt model is a smart place to spend the effort. O…
- @Firecrawl (0): Maker [@jernej\_jan\_kocica](https://www.producthunt.com/@jernej%5Fjan%5Fkocica) Very fair points! This is something we're taking into account. This model runs only on what's already retrieved currently but stay tuned! Upvote Report Share 7h ago [](/@nivotools)
YC Has It 是一个完全免费无需注册的 AI 搜索工具,用自然语言描述问题即可从 4000 多家活跃 YC 公司中找到匹配解决方案并附带推理、定价和集成信息
YC has it
受众观点:①创业者和产品经理关注如何快速在 YC portfolio 里发现处理特定问题的公司 ②Clemente Lopez 指出用户陈述的问题和真实预算痛点可能不一致,准确率依赖用户能否准确描述需求 ③投资人和行业研究者关注是否支持按融资阶段过滤
展开评论
- @Search (0): Describe your problem in plain English. ychasit searches 4,000+ active YC companies and finds the startups that solve it, with reasoning, pricing, and integrations. 100% Free forever. No login. No signup. * [Overview](/products/yc-has-it) * [Reviews](/products/yc-has-it/reviews)…
- @Search (0): [Workflos.ai](/products/workflos-ai?ref=product%5Fsidebar) AI assistant to Find & Manage SaaS with natural language [5.0(1 review)](/products/workflos-ai/reviews)
- @Search (0): [Mr. Free Tools](/products/mr-free-tools?ref=product%5Fsidebar) A Better Way to Find Free Tools and Resources [4.8(4 reviews)](/products/mr-free-tools/reviews) [Startup communities](/categories/startup-communities)[Graphic design tools](/categories/graphic-design-tools) [AI-Powe…
- @Raghav (0): Maker 📌 Hey Product Hunt, I'm Raghav 👋 There are 4,000+ active YC companies. Most of them are invisible to the people who need them most. The one that solves your exact problem is probably in there somewhere, and today the only way to find it is scrolling a directory sorted by b…
- @Raghav (0): Maker [@artem\_fedorovich](https://www.producthunt.com/@artem%5Ffedorovich) Great question Artem. Currently we do not plug in funding data but definitely considering it for the future versions. Upvote Report Share 5h ago [](/@clemente%5Flopez1) [Clemente Lopez](/@clemente%5Flope…