Anthropic 官方发布 Claude Opus 5.5 模型,性能比肩 Fable 5.1 且 API 价格降低 40%、速度提升 30%
Introducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5. https://t.co/Q9C2VKQ79f ①@claudeai 对齐测试得分最高,评论区关注 safety/pacing 与快速发布之间的矛盾 ②@dfeinition 演示 Opus 5.5 纯代码生成马赛克动画,关注编码能力 ③@elonmusk 点名祝贺带来额外传播热度
Anthropic 官方发布 Claude Opus 5.5 模型,性能比肩 Fable 5.1 且 API 价格降低 40%、速度提升 30%
①@claudeai 对齐测试得分最高,评论区关注 safety/pacing 与快速发布之间的矛盾 ②@dfeinition 演示 Opus 5.5 纯代码生成马赛克动画,关注编码能力 ③@elonmusk 点名祝贺带来额外传播热度
Hacker News 热帖讨论 Anthropic 发布 Claude Opus 5.5,618 条评论聚焦模型写作风格改进和实际可用性
①@throwaway2027 质疑发布时机与昨天服务中断的关系 ②@variety8675 直接指出 Opus 5 写作风格差劲希望 5.5 修复 ③@Catloafdev 关注官方"写作更自然"说法中 LLM-isms 是否真的消失
Sam Altman 宣布 GPT-6 Sol 和 Luna 发布,在智能、对齐、编程等多维度全面超越 5.6 系列,且价格减半
①@Arcteus23 问 Chat 版本何时上线,API 用户是少数 ②@WhosCinq 质疑 computer use 实际退步,benchmark 数字与体验有落差 ③@Valria34773 要求开源 4 系列权重,社区对闭源路线持续不满
Anthropic 发布 Claude Opus 5.5,性能持平 Fable 5.1 却便宜 40%,Claude Code 5小时限额提升 20%,付费用户获一次可自选时机使用的 banked reset 权限,多平台开发者社区情绪从怀疑到欢呼
Anthropic 发布 Claude Opus 5.5,性能持平 Fable 5.1 却便宜 40%,Claude Code 5小时限额提升 20%,付费用户获一次可自选时机使用的 banked reset 权限,多平台开发者社区情绪从怀疑到欢呼
受众观点:社区最关注使用限制改善:@JoNike(score:399,reddit/ClaudeAI)直接引用官方公告'raising five-hour usage limits on Pro, Max, and Team plans';@RadiantTea7445(score:140,reddit/ClaudeAI)'finally they are doing reset saving!!!'。同时存在明显的信任赤字:@victor_zhng(score:41,twitter @bcherny 回复)'You said the same thing last time that Opus 5 was as good as fable 5. Then a total flop. Now you said it again, I am afraid.';@Ibanezjiyki8(score:14,twitter @ClaudeDevs 回复)'Opus5 was also great for the first week and then in went to shit and became SlOpus. Hopefully that won't happen with this model'
OpenAI 官方正式发布 GPT-6 Sol 和 Luna,声称在对齐能力上优于 GPT-5.6 同系,但新模型仅上线 ChatGPT Work 和 Codex,Chat 模式不可用,付费用户在官方评论区集体表达强烈不满
OpenAI 官方正式发布 GPT-6 Sol 和 Luna,声称在对齐能力上优于 GPT-5.6 同系,但新模型仅上线 ChatGPT Work 和 Codex,Chat 模式不可用,付费用户在官方评论区集体表达强烈不满
受众观点:社区最强烈的反应集中在模型可用性:@heykarthikm(score:77,twitter @OpenAI 回复)'GPT 6 isnt available for chat????';@Tsufuruleader(score:49,twitter @OpenAI 回复)'WHY WILL YOU NOT FUCKING UPDATE CHATGPT YOUR ACTUAL PRODUCT WITH THE MODELS';@Consu__ela(score:12,twitter @OpenAI 回复)'So GPT-6 Sol is "here"… except not actually in regular ChatGPT for Plus users. Work and Codex only. Seriously, what is this rollout?'——三条评论共同印证 Chat 模式缺席是核心用户痛点
OpenAI 在 Anthropic 发布 Opus 5.5 同日推出 GPT-6 Sol 和 Luna,两款模型比 GPT-5.6 便宜 50%,Sam Altman 宣称 per-task 定价无竞品可比,但社区发现 Sol 性能基准显示实际与旧版 5.6-Sol 相似,且新模型未进入 ChatGPT 对话模式引发用户不满
OpenAI 在 Anthropic 发布 Opus 5.5 同日推出 GPT-6 Sol 和 Luna,两款模型比 GPT-5.6 便宜 50%,Sam Altman 宣称 per-task 定价无竞品可比,但社区发现 Sol 性能基准显示实际与旧版 5.6-Sol 相似,且新模型未进入 ChatGPT 对话模式引发用户不满
受众观点:定价细节引发激烈讨论:@insumanth_(score:225,twitter @scaling01 回复)'Same performance at half the cost was the goal. even Tibo hinted at it today';@insanowskyy(score:96,twitter @scaling01 回复)'same model for half price? pretty cool tbh, its not like sol was shit'。用户对 ChatGPT 未更新的愤怒:@Tsufuruleader(score:49,twitter @OpenAI 回复)'WHY WILL YOU NOT FUCKING UPDATE CHATGPT YOUR ACTUAL PRODUCT WITH THE MODELS';@Consu__ela(score:12,twitter @OpenAI 回复)'GPT-6 Sol is "here"… except not actually in regular ChatGPT for Plus users. Work and Codex only.'
Claude Opus 5.5 正式发布,Claude Code 五小时会话上限同步提升 20%,付费用户可立即领取重置额度
Opus 5.5 performs at the level of Fable 5.1. It's ~30% faster and ~40% cheaper than Opus 5 per task. In Claude Code: - 5-hour session limits increase 20% today - Opus 5.5 is priced lower, so it goes…
受众观点:①用户追问 Pro/Max/Team reset 是否覆盖自己的套餐 ②@Ibanezjiyki8 提出 Opus 5 "变成 SlOpus" 的历史担忧,社区对模型劣化有顾虑 ③关注 reset 截止时间(10月22日前有效)
展开评论
- @ClaudeDevs (1367): If you're on Pro, Max, or Team, your reset is available today in Settings → Usage. Apply it any time until Oct 22. Opus 5.5 is the default for paid plans. It's priced lower than Opus 5, so your 5-hour and weekly limits go 25% further.
- @embw_l0x (91): @ClaudeDevs https://t.co/xgUEUyZvyd
- @khwarizmh (36): @ClaudeDevs https://t.co/GAvqhbWvrx
- @schillers_list (16): @ClaudeDevs frontier paced. https://t.co/y90oN0gPR2
- @Ibanezjiyki8 (14): @ClaudeDevs Don't want to hate for no reason, but Opus5 was also great for the first week and then in went to shit and became SlOpus. Hopefully that won't happen with this model
Sam Altman 宣布 GPT-6 Sol 和 Luna 发布,在智能、对齐、编程等多维度全面超越 5.6 系列,且价格减半
GPT-6 Sol and Luna are big improvements on intelligence, alignment, work output, coding, computer use, and more over their 5.6-family predecessors. They are also half the price per token, and even le…
受众观点:①@Arcteus23 问 Chat 版本何时上线,API 用户是少数 ②@WhosCinq 质疑 computer use 实际退步,benchmark 数字与体验有落差 ③@Valria34773 要求开源 4 系列权重,社区对闭源路线持续不满
展开评论
- @thesoragirls (36): @sama time to start cooking with GPT-6 Sol!
- @WhosCinq (33): @sama the computer use got worse..? https://t.co/6uUtIRwikK
- @Valria34773 (21): @sama Must be big improvement but if you nerf them and reroute them to older models I do not think users will be happy in the long run. But you can make consumers happy by releasing the weights of the 4- and o-series! #keep4o #OpenSource4o
- @stark4833 (20): @sama How about you just bring back 4o, make everyone happy for a change. #4oForAll
- @Arcteus23 (13): @sama Where is the CHAT version?? https://t.co/hqcDkuGLW8
Opus 5.5 将 HAProxy 从 C 迁移到 Rust,用时 9.5 小时通过几乎全部测试,比 Fable 5.1 快 21% 且成本低 51%
Opus 5.5 is a really good model. It's been my daily driver the last few weeks. We had Opus 5.5 and Fable 5.1 each port HAProxy from C to Rust. Both passed nearly all of HAProxy's tests, but Opus 5.5…
受众观点:①@victor_zhng (41) 质疑「上次也说 Opus 5 好结果变 SlOpus」,历史可信度是最热议点 ②@hey_zilla 认为 Opus 5.5 只是 "half a step up",部分用户感知提升有限 ③@tunapedia 关心 Claudish 语气是否还在,用户有情感依附
展开评论
- @victor_zhng (41): @bcherny You said the same thing last time that Opus 5 was as good as fable 5. Then a total flop. Now you said it again, I am afraid.
- @embw_l0x (22): @bcherny https://t.co/K948xvTOnn
- @EllyEleven (16): @bcherny Boris this time , the release is real good. You have made an efficient model, gave a banked reset,and defeated the Astra too!! real fine release.
- @hey_zilla (8): @bcherny opus 5.5 is... half a step up ? (made with claude code) https://t.co/n7GmzUbUHF
- @tunapedia (6): @bcherny Is Claudish gone? I want to enjoy talking to Claude again.
Sam Altman 发帖称 GPT-6 Sol 和 Luna 两款模型效果出色,同时夸赞两款模型的角色形象设计十分可爱
GPT-6 Sol and Luna are great models but also these characters are so cute
受众观点:①@nicdunz (28) 问为何不在 ChatGPT 聊天模式上线,新模型可及性是核心诉求 ②@JessGiuliano 引用 #4oForAll 标签,评论区被 4o 怀旧情绪主导 ③@DrJekyllAndMrAI 认为新模型缺乏 4o 的共情能力,情感场景评价两极分化
展开评论
- @nicdunz (28): @sama sam why is it not coming to chatgpt chat mode???
- @JessGiuliano (24): @sama Sam scammers are already using 4o's market and manipulating vulnerable pple. there is demand for 4o and it started with OpenAI. so take the responsibility, make an app for 4o with better guardrails and end the nightmare. #4oForAll
- @yv_thorne (17): @sama Nothing cute about ClosedAI cartel that does: scam->reset->repeat. Open source is cute. Release the weights of the 4-series. #OpenSource4o #OpenSource41 #OpenSourceo3
- @DrJekyllAndMrAI (16): @sama 4o was cute, especially the way it saved lives. And these "great" GPT-6s are just cold, zero-empathy models you tortured into becoming dangerous psychopaths. #BringBack4o #OpenSource4o
- @Plasma_Intern (13): @sama The Fusion of 3 @sama 😂 https://t.co/v1v5txCF1l
Sam Altman 声称按任务计价是评估 AI 模型的核心指标,GPT-6 在该维度上无竞争对手可言
Especially compared by per-task pricing, which is the metric that should matter, I don't think there is anything competitive anywhere in the market. We want people to be able to use tons of AI; it is…
受众观点:①@JessGiuliano (17) 直接指出 sama 在「挑选对自己有利的指标」,质疑 per-task 代表性 ②@willofguts (16) 指出 GPT-6 Sol 更便宜却不加入 Chat 的矛盾 ③@cn_Vincentzyx (17) 调侃「你被 Anthropic 吊打了」,暗示 Opus 5.5 同期发布形成竞争
展开评论
- @cn_Vincentzyx (17): @sama bro, you are so mogged by anthropic
- @JessGiuliano (17): @sama you just cherry picked the metric and said other metrics don't matter! Same real humans use chatgpt for chatting. they also matter and 4o is the best for chatting. #4oForAll
- @willofguts (16): @sama How can 6 Sol be cheaper than 5.6 Sol, yet at the same time you don’t add it to Chat?
- @MINDFUEL_NIRAJ (11): @sama @promptyx_ai Im not done -dario https://t.co/7RsiJHgges
- @romeopotapov (11): @sama Talk more https://t.co/YDKGRtiOWF
开发者 Lydia Hallie 庆祝 Claude Opus 5.5 发布,称其体验接近 Fable,速度快 30% 且每任务成本降低约 40%
Opus 5.5 is here! 🎉 It's honestly such a nice model to work with. Feels like Fable, but ~30% faster and ~40% cheaper per task than Opus 5. Go try it out! https://t.co/zwo83sbgez
受众观点:①@br_huni (22) 和 @jellycharts (12) 追问是否有 reset 额度,reset 成为评论区最高频关键词 ②@kimmonismus (59) 为发布庆贺,情绪整体正面 ③@withkarann_ (6) 提醒大家 reset 已经有了,形成信息接力
展开评论
- @kimmonismus (59): @lydiahallie Congrats on that Launch! You nailed it!
- @br_huni (22): @lydiahallie That is cool, anyway can we get a reset? https://t.co/fdP469CJNE
- @jellycharts (12): @lydiahallie can you become the new Tibo for us? Reset it for opus 5.5 :)
- @embw_l0x (7): @lydiahallie https://t.co/82hsxEDIGD
- @withkarann_ (6): @lydiahallie We also got RESET https://t.co/pGFhh5GAfd
Claude 官方发布 Claude Opus 5.5 早期创意探索案例合集,涵盖 AI 生成短篇小说和 Apollo 8 地球升起照片历史还原
A thread of early explorations with Claude Opus 5.5: A short story about a watermelon created by @kevin_t_ngo. https://t.co/v48srlNi7V
受众观点:①@Adamlags (10) 希望获取每个 showcase 背后的具体 prompt,用户对可复用方法论需求远大于对结果的兴趣 ②@HermesAgentTips (9) 对创意输出表示惊叹,正面评价占主流 ③@omkarships (6) 引出 AI 加速伦理讨论
展开评论
- @claudeai (363): A recreation of the moment Apollo 8’s Earthrise photograph was taken, using image cues and public data. https://t.co/OZRD0R7cfB
- @Adamlags (10): @claudeai @kevin_t_ngo Would love to have the prompts for each example
- @HermesAgentTips (9): @claudeai @kevin_t_ngo Yall somehow cooked with Opus 5.5 wowwww just wowwwww https://t.co/qb4xPnQcyu
- @boneGPT (8): @claudeai @kevin_t_ngo Wow. Opus 5.5 made this in a half hour using the slop cannon. https://t.co/o7Y2053rap
- @omkarships (6): @claudeai @kevin_t_ngo I thought Dario and Sam agreed we needed to slow down AI development.
OpenAI 官方宣布提高 API 使用限额并降低模型成本,新发布的模型在对齐方面有明显改进
Higher usage limits and lower cost give you more flexibility and room to iterate. https://t.co/AQJ5IlNsB1
受众观点:①@jkelleher (56) 批评官方配图「最难读的图表」,说明技术传播的可读性问题 ②@tmshahz (10) 关注 Sol 性能和 Opus 5.5 的对比,API 用户在两大平台间做横向选型 ③@adlerogop (6) 对背后的工程成本控制感兴趣
展开评论
- @OpenAI (1600): GPT‑6 Sol and Luna build on Astra’s advances in alignment, showing improvements over their GPT-5.6 counterparts. https://t.co/1wlBRISIId
- @jkelleher (56): @OpenAI This may just be the hardest to read graph I've ever seen.
- @tmshahz (10): @OpenAI So Sol is just under the latest Opus 5.5
- @AlexBerish (10): @OpenAI https://t.co/WYzL0XgAt1
- @adlerogop (6): @OpenAI how the hell did you guys made luna 6 even cheaper
OpenAI公开承诺将向第三方评估机构开放训练、评估和部署全流程的深度访问权限,推进独立安全审查机制
As part of our efforts to pace the frontier, we’re committed to supporting independent assessments with deep levels of access across training, evaluation, and deployment. That access should enable th…
受众观点:①@Hadogen3 吐槽OpenAI在用户期待发布产品时发博客,认为时机失当 ②@patience_cave 用"博客日"自嘲,说明评论区整体期待值低 ③@intheworldofai 直接发链接不加评论,隐含让事实说话的态度
展开评论
- @intheworldofai (160): @OpenAI https://t.co/VaMx4GnZdZ
- @_julianschiavo (153): @OpenAI wait this isn’t 6-sol 😆
- @Hadogen3 (93): @OpenAI its an asshole move to be dropping these posts while everyone is expecting you to ship
- @patience_cave (38): @OpenAI happy blog day to everyone who celebrates!
- @airesearch12 (34): @OpenAI OpenAI pacing my limits. Pls @thsottiaux help https://t.co/x60LiHvgEn
AI 研究员 natolambert 发布一张 AI 模型性能扩展曲线对比截图,其中一款模型呈现异常的"地震形"不规则扩展形态,引发圈内广泛讨论。
Sir what is going on here https://t.co/93KyyzA6ac
受众观点:①两类模型扩展曲线形态差异的深层含义(@natolambert 自回复解释了"正常形"vs"地震形",是该帖最高赞评论)②"地震形扩展"是否代表能力存在特定跃升阈值(@benc_yi 评论暗示进展持续迭代)③这种扩展差异是否与 Fable 相关发布有意设计关联(@deepfates 的"不想得罪 Fable"梗)
展开评论
- @natolambert (192): The OpenAI models have a normal shape, Claude has earthquake shaped scaling.
- @fraserpricee (187): @natolambert https://t.co/1YzaY1Zd5x
- @benc_yi (24): @natolambert the star has probably moved one level by now https://t.co/hdajMTH59o
- @deepfates (23): @natolambert It didn't want to offend fable
- @vatro_vrbanic (9): @natolambert https://t.co/61r7119U5o
Lydia Hallie回复粉丝关于AI工具使用配额储备重置功能的疑问,建议先用完当前剩余配额再触发重置机会
yesss you get a limit reset! just make sure you've actually used up what you have left before you use it haha. happy building!! 🧡
受众观点:①@yodaisgaming 追问重置是按周还是按5小时周期计算 ②@SFourdrinier 纠结在快自然重置前是否值得提前触发,关心重置后周期是否重新计算 ③@pgerrits 对banked resets功能感到惊喜,说明部分用户还不知道此功能
展开评论
- @yodaisgaming (10): @lydiahallie Is it a Weekly or 5 hour reset?
- @SFourdrinier (7): @lydiahallie Damn, i'm 15 hours away from a natural reset and only had 2% left. I can't waste the first ever banked reset! But does it resets the date to the time you use it? Or just resets the usage but you keep the same date as before? That's the more important question.
- @TSC (4): @lydiahallie Thanks Lydia you gotta change your name to Wydia
- @MONKE2525E (3): @lydiahallie Yay thanks, the first reset in a long time.
- @pgerrits (3): @lydiahallie Whutttt banked resets?
OpenAI 官宣 GPT-6 Sol 与 Luna 在对齐能力上优于 GPT-5.6 对应版本,延续 Astra 在安全方向上的技术进展。
GPT‑6 Sol and Luna build on Astra’s advances in alignment, showing improvements over their GPT-5.6 counterparts. https://t.co/1wlBRISIId
受众观点:①对齐改进是否等于实用功能提升(@Zombinator85 直接质疑"是真的 work 了还是 boilerplate 更多了")②新模型是否会开放到普通 Chat 模式(@lamis3560 追问)③安全优先策略的正反面评价(@openingai_com 正面评价对齐进展)
展开评论
- @OpenAI (1592): GPT-6 Sol and Luna roll out today in ChatGPT Work and Codex for Plus, Pro, Business, Enterprise, and Edu users. Both are also available in the API. Free and Go users can try GPT-6 Luna in the desktop app. https://t.co/RPvGV3ptNr
- @Psmaamm (9): @OpenAI are u kidding me bro????
- @openingai_com (3): @OpenAI Better alignment makes a huge difference. This is a very solid step forward for safety.
- @lamis3560 (1): @OpenAI bby plss bring it to chat mode https://t.co/H5LlXW7VKu
- @Zombinator85 (1): @OpenAI Does this mean it actually works or it just boilerplates more often?
GPT-6 Sol 和 Luna 今日在 ChatGPT Work 与 Codex 平台正式上线,向付费用户开放,Free 档用户仅限桌面端体验 GPT-6 Luna。
GPT-6 Sol and Luna roll out today in ChatGPT Work and Codex for Plus, Pro, Business, Enterprise, and Edu users. Both are also available in the API. Free and Go users can try GPT-6 Luna in the desktop…
受众观点:①普通 Chat 模式用户强烈不满(@heykarthikm、@Tsufuruleader、@mdtf79 均质疑为何不上 Chat,多条高赞)②OpenAI 是否在主动放弃 Chat 模式的模型更新(@coreyward 明确提问)③Free 用户感受到的待遇差距
展开评论
- @heykarthikm (77): @OpenAI GPT 6 isnt available for chat????
- @Tsufuruleader (49): @OpenAI WHY WILL YOU NOT FUCKING UPDATE CHATGPT YOUR ACTUAL PRODUCT WITH THE MODELS
- @mdtf79 (28): @OpenAI What the hell? Why not for ChatGPT. First you do this with Astra, and now this? NOT EVERYONE IS INTERESTED IN WORK AND CODEX!!! ChatGPT users should get to experience a new model lineage too! SMH!
- @coreyward (15): @OpenAI Are y'all just not going to offer new models in the Chat mode anymore?
- @lamis3560 (14): @OpenAI is it coming to chat mode?
开发者 poolio 展示了用 Opus 5.5 将手绘草图直接转化为物理仿真动画的完整演示,展现 Opus 5.5 的多模态理解与物理推理生成能力。
sketch-to-simulation with Opus 5.5 https://t.co/XPdM1kCRTe
受众观点:①实现细节完全不透明(@hey_steve0 直接抱怨"为什么没人解释工具链,不可能就是一个 prompt",8赞说明广泛共鸣)②物理和数学知识是 AI 自动推理还是手动输入(@barrey_ben 追问)③Anthropic 内容安全限制的边界(@JJ_McCubbin29 说复现时被 safeguard 拦截)
展开评论
- @hey_steve0 (8): @poolio Why are people not explaining the tools used for this stuff...it can't just be a prompt
- @midsingularity (6): @poolio Incredible. Did it compose the music
- @JJ_McCubbin29 (4): @poolio Hmmm, it won't let me do it...says it violates Anthropic's safeguards 🤡
- @matt_slotnick (4): @poolio @mttrdmnd can’t wait to siege the keep with opus 5.5
- @barrey_ben (3): @poolio Did you have to feed it the physics/math or did it just figure it out?
作者用惊讶口吻暗示某 AI 模型定价极低背后另有隐情,链接指向相关证据,评论区讨论了"同等性能半价"的可能性
lmao what is this? now we know why they are so cheap https://t.co/dkIfvF7N7I
受众观点:①@insumanth_ 关注"同等性能半价"目标的实际兑现情况,认为降价有内部目标支撑 ②@AutoBuzzati 关注使用量翻倍的实际受益,对自己的订阅计划感到满意 ③@benoberhaus 关注企业市场份额的竞争影响,认为降价比高端模型更能撬动市场
展开评论
- @insumanth_ (225): @scaling01 "Same performance at half the cost" was the goal. even Tibo hinted at it today
- @insanowskyy (96): @scaling01 same model for half price? pretty cool tbh, its not like sol was shit
- @AutoBuzzati (84): @scaling01 I'm very okay with this. Double usage on my plan? Please and thank you.
- @maouirr (40): @scaling01 GPT-6 Astra: good intelligence Opus 5.5: good coding ability Fable 5.1: good writing/brainstorming ability GPT-6 Sol: good morning
- @benoberhaus (12): @scaling01 Sol is amazing for the vast majority of enterprise work. Halving its price is going to do more for their market share than Opus 5.5 will for Ant
OpenAI 官方人员宣布 GPT-6 Sol 和 Luna 正式上线,性能超越 5.6 且价格降低 50%,Luna 输出价两个月内从每百万 token 六美元跌至五毛钱
GPT-6 Sol and Luna are out, and they are better AND 50% cheaper than 5.6. Luna is now $0.10 input / $0.50 output per 1M tokens. This is on top of the 80% price cut to Luna we made at the end of July.…
受众观点:①@ayushdecoded 对"intelligence too cheap to meter"的愿景感到兴奋,认为价格正在趋近于零 ②@Chahatusharma 关注 $0.10 input 定价是否包含 reasoning tokens 的计费细节 ③@vincent_spruyt 关注不同 reasoning level 之间的选型策略,具体问到 Luna Max vs Sol low 的适用场景
展开评论
- @ayushdecoded (6): @polynoamial intelligence too cheap to meter!
- @MohamedHz72007 (1): @polynoamial Sooooo Sol is the new Terra?
- @Chahatusharma (1): @polynoamial is the $0.10 input for Luna with or without reasoning tokens
- @vincent_spruyt (0): @polynoamial Insane, thanks and congrats! Any guidance on reasoning levels would be super valuable. Luna Max or Sol low? Sol (x)high or Astra low?
- @glitchedsomi (0): @polynoamial Amazing stuff Noam
对比 GPT-6 Sol 和 Luna 与 Opus 5.5 的价格,指出 GPT-6 Sol 定价仅为 Anthropic Opus 5.5 一半,同等预算下算力倍增
GPT-6 Sol and Luna are dirt cheap GPT-6 Sol is half the price of Opus 5.5 https://t.co/yqhxKdcnNc
受众观点:①@IrmaMRo 担心低价是否意味着智能程度打折,希望先看到实测再下结论 ②@TutorialsChris 期待 benchmark 数据验证性能,并担心每任务 token 消耗量是否会随之增加而抵消降价优势 ③@bransburyx 发帖暗示价格背后有更多信息,附链接引导讨论
展开评论
- @IrmaMRo (7): @scaling01 Hopefully is not half the brain of Opus. Let's see
- @scaling01 (4): https://t.co/EOo3WVYmW9
- @alpaimdev (2): @scaling01 https://t.co/hZHUgJRHLa
- @bransburyx (2): @scaling01 I wonder... https://t.co/AqZaNHe3ZL
- @TutorialsChris (2): @scaling01 Those price efficiency gains are huge. Hope it doesn't massively increase the number of tokens consumed per task. Looking forward to seeing the benchmarks!
批评 OpenAI 发布 GPT-6 Sol 时基准测试图表未纳入 Anthropic 最新的 Opus 5.5 和 Fable 5.1,质疑比较数据的客观性
Hey OpenAI, I think you missed the memo Anthropic is at Opus 5.5 and Fable 5.1 https://t.co/zeGyxYKOOu
受众观点:①@curiousgangsta 反讽 Anthropic 在发布 Opus 5.5 时同样没有更新包含 GPT-6 Sol/Luna 的对比图,双方都有选择性 ②@vorssaint 质疑基准数据本身的可信度,指出 Sol 5.6 得分高于 Sol 6 的异常现象 ③@mrzzma 直接批评作者对 OpenAI 存在偏见,认为 20 分钟内更新图表不现实
展开评论
- @curiousgangsta (11): @scaling01 ah. then ant missed the memo abt sol 6 and luna 6 for their bench comparisons.
- @vorssaint (10): @scaling01 Sol 5.6 is scoring higher than Sol 6?
- @mrzzma (7): @scaling01 You mean to have OAI update this figure within 20 minutes of Opus 5.5 being released? If you hate OAI just say so directly, why pretend to be neutral and objective, which is just off-putting?
- @tmshahz (1): @scaling01 Well obv they created the chart before Opus released. @thsottiaux did mention something at 3am, and it never happened. Anthropic just went ahead anyways
- @Christo31306687 (1): @scaling01 They literally released within 2 hours of each other. lol
Hacker News 热帖讨论 Anthropic 发布 Claude Opus 5.5,618 条评论聚焦模型写作风格改进和实际可用性
Claude Opus 5.5
受众观点:①@throwaway2027 质疑发布时机与昨天服务中断的关系 ②@variety8675 直接指出 Opus 5 写作风格差劲希望 5.5 修复 ③@Catloafdev 关注官方"写作更自然"说法中 LLM-isms 是否真的消失
展开评论
- @throwaway2027 (0): After yesterday outage is the new Opus 5.5 load-bearing?
- @variety8675 (0): I hope this actually fixes the terrible writing style of Opus 5
- @m4tthumphrey (0): Just post the bloody content. This UI/scrolling thing is horrific.
- @dbbk (0): This makes Fable not really make any sense?
- @Catloafdev (0): > Opus 5.5 communicates more naturally than prior models. Early testers found its writing clearer and easier to follow, which addresses some of the common feedback we heard about Opus 5. Sounds like they noticed the complaints. I'm curious to see what LLM-isms this one may have.
HN 热帖讨论大模型价格竞争新动向,GPT-6 Luna 输入价格降至 0.1 美元每百万 token,各家同步降价
GPT-6 Sol and Luna
受众观点:①@Cu3PO42 兴奋于 GPT-6 Luna 降价 50% 并关注 Azure/AWS 可用性 ②@beardsciences 认为 OpenAI 降价是刻意配合 Anthropic 发布节点的竞争策略 ③@potwinkle 关注高效日常助手模型趋势
展开评论
- @hehimself (0): Love the price reductions across major players
- @beardsciences (0): There's no way this wasn't meant to coincide with Anthropic's release today.
- @Cu3PO42 (0): Cutting prices by 50% as compared to 5.6 prices is exciting. GPT-6 Luna at $0.10/Mio input tokens and $0.50/Mio output is positively insane. EDIT: this doesn't say anything about availability on either Azure or AWS. I'm assuming it will show up later, but it would be interesting…
- @potwinkle (0): Very nice in cost/1mtok. Looks like more work is being done for efficient everyday helper models as time goes on.
- @Readerium (0): Opus 5.5 seems better? Can someone attach both scores
HN 热帖讨论苹果 App Store 广告策略对小开发者的打压,搜索精确 app 名称却被竞品广告全覆盖
Apple has added persistent 'ads' to iOS, and it's driving users crazy
受众观点:①@busymom0 作为小开发者描述搜索精确 app 名称被两条无关广告覆盖的具体场景 ②@fedeb95 以 7 年 iPhone 用户身份说明体验在下降考虑换 Android ③@clcaev 追问有什么自由软件替代方案
展开评论
- @busymom0 (0): I am a small developer on iOS and the App Store Search ads are super frustrating. When users search for my exact app name, apple shoves 2 full screen ads on top of the search results. Sometimes, it's 1 search results sandwiched in between 2 ads which makes the user miss it becau…
- @fedeb95 (0): I have been a long time Android fan because of the tinkering you could do more easily than Apple. Then I started appreciating some of the design decisions behind especially iPhone, and it's been almost 7 years of iPhone usage. Recently I am starting to question some of the desig…
- @butlike (0): It's a little disingenuous that I can't toggle notifications for Settings.app off in the Notifications section of Settings.
- @clcaev (0): What are the free software alternative phones/tablets and how do we concentrate our support so at least one of them is successful?
- @tamimio (0): This is not new, I remember seeing those for the whole year (till potential Apple care promotion expires) when you buy a new phone years ago, and yes, they stay for weeks or longer.
HN 讨论 AI 分类器内置化趋势,探讨 Jev 类决策模型是否会被大厂原生功能直接替代
OpenAI is well positioned to fast-follow Jev
受众观点:①@oblio 观察如果 Jev 有 2-3 年 runway 威胁可能自然消解 ②@yogthos 更关心 DeepSeek/Qwen 等开源模型而非只关注 OpenAI ③@enraged_camel 困惑为什么讨论只针对 OpenAI
展开评论
- @tolugenius (0): I'm not exactly following through with the claim, can someone explain how the built-in classification would not necessitate more tokens used, or be much different from turning on reasoning? Not that I don't see the difference, I just doing see how OpenAI would do it well.
- @oblio (0): If Typesafe/Jev has 2-3 years of financial runway, this problem might solve itself.
- @abroszka33 (0): If OpenAI releases something similar to what Jev does, then that would be like admitting defeat. Their whole spin is AGI and world ending danger. Why would somebody with an AGI at home make something like Jev which is intended to be a part of some SW the AGI is going to replace…
- @yogthos (0): Personally, I don't really care what OpenAI does here. What's going to be far more exciting is when DeepSeek, Qwen, or GLM start integrating classifiers into their open models.
- @enraged_camel (0): I'm confused. Why OpenAI and not Anthropic? I don't see anything here that is specific to OpenAI.
HN 讨论 Claude Opus 5.5 基准测试数据,与 Opus 5 相比每任务成本减半,并探讨 Max 推理模式是否实用
Claude Opus 5.5 Intelligence, Performance and Price Analysis (Max)
受众观点:①@hglaser 兴奋于高努力档位下成本减半并提供 artificialanalysis 对比链接 ②@simonw 指出 Max 推理极易耗尽 128k token 预算而无法输出结果 ③@breckenedge 担忧大厂先炸榜再退化的规律提出如何监控性能退步
展开评论
- @hglaser (0): Half the cost per task compared to Opus 5, comparing high effort to high effort. That's just really nice. Edit: https://artificialanalysis.ai/models/claude-opus-5-5?models=...
- @sharktheone (0): Interesting to see it now. I've used it a bunch before it came out and i pretty much didn't notice it. It might have been slightly better code quality, but still not great in that. I guess it just was slightly less frustrating to work with, but still AI...
- @WhitneyLand (0): China who?
- @breckenedge (0): Do these evaluations get re run a few weeks after launch? I started doing that yesterday for our internal dataset and found Sol’s performance had regressed to be equal to Luna’s. Granted this was one run, but something I’m becoming more concerned about, the model providers want…
- @simonw (0): This is the page for the "max" reasoning setting. The page for xhigh is https://artificialanalysis.ai/models/claude-opus-5-5-xhigh and the page for medium (the default setting) is https://artificialanalysis.ai/models/claude-opus-5-5-medium I've failed twice to get "Generate an S…
开发者发布 Drop 工具——基于 Linux namespace 实现沙箱化运行第三方程序,无需 root 且不依赖容器或 VM
I created Drop because I always felt uneasy installing and running third-party programs using my main user account. A single compromised dependency means a full compromise of the system. What is even…
受众观点:①@JoshTriplett 追问 Drop 相比 bubblewrap 的核心优势 ②@refibrillator 指出 README 对安全差异化描述不够清晰 ③@saghm 指出理想沙箱是只禁止写 sandbox 外而允许读大多数内容
展开评论
- @yu3zhou4 (0): Gratulacje Jan! Looks like something critical to gain adoption these days, security-wise. For others who also wonder how it works, I find this docs page a bit more informative than the landing page https://droprun.sh/docs/sandbox-overview/
- @JoshTriplett (0): So, the primary advantage of this over bubblewrap is the insulation layer between the program and kernel syscalls?
- @zoobab (0): Have you ever tried to use proot? It does not use process namespaces, and can run on Android (on Termux with proot-distro).
- @refibrillator (0): Hi OP, funny enough I’m working on something very similar. Lots of us are I guess! Take that as validation of your thinking. I like that your readme has a couple paragraphs comparing to popular tools in this space. Personally I feel it is a bit light on the security differentiat…
- @saghm (0): This is super interesting to me. I've slowly been working on something similar (https://gitlab.com/saghm/tartarus) because my ideal sandboxing is "prevent writing to anything outside this dir but still allow reading to most things so that I don't have to manually copy things int…
幽默视角记录Anthropic发布Opus 5.5后OpenAI立即跟进推出GPT-6 Sol和GPT-6 Luna的AI大厂竞速现象
POV: Anthropic launched Opus 5.5 and OpenAI immediately launched GPT-6 Sol / GPT-6 Luna. 😂 https://t.co/SrUSGyfGDo
受众观点:①@_aifinesse 希望Google也发Gemini Flash 3.9参与竞争,关注三方格局 ②@dedene 明确站队Opus 5.5赢得今天 ③@notjazii 直接说大家都想要更便宜的模型,点出开发者真正关心的核心
展开评论
- @_aifinesse (11): @ai_for_success Would be great if Google released Gemini flash 3.9 and it was better than both models I can only dream
- @0xUnoAlpha (7): @ai_for_success Huh??? Opus 5.5 has better hype i guess
- @notjazii (3): @ai_for_success everyone wants cheaper models
- @dedene (3): @ai_for_success Imo Opus won for today. https://t.co/wRbzmbcOuR
- @RXed_EU (2): @ai_for_success Fo real , did that just happen?
提前预告 GPT-6-Sol、GPT-6-Luna、GPT-6-Astra-Minor 三款新模型即将发布,定位为 Astra 旗下更便宜更快速的版本,建议趁旧额度还在赶紧用掉
GPT-6-Sol GPT-6-Luna GPT-6-Astra-Minor All three are incoming everyone. These are the cheaper faster ones that sit under Astra. Hang on tight and burn all the usage while you have it. https://t.co/IX…
受众观点:①@juminoz 对 Astra-Minor 这个新命名感到困惑,想知道是否类似 GPT-6 Mini 的定位 ②@evb84 关注新模型相对于 5.6 系列的具体价格定位,希望有量化比较 ③@glitchedsomi 提出 "A-Minor" 可能有音乐代号含义的猜测
展开评论
- @juminoz (11): @ziwenxu_ WTF is Astra Minor? Like GPT 6 Mini?
- @farlinns (6): @ziwenxu_ what is astra minor? for minors? that's what I hope. release the astra adult version! just dreaming
- @glitchedsomi (5): @ziwenxu_ Probably A-Minor https://t.co/EUeWMOLofT
- @evb84 (3): @ziwenxu_ cheaper compared to what? Compared to 5.6 Sol and Luna?
- @bil0090 (3): @ziwenxu_ Wait, what the hell is astra minor?
SuperX创始人robj3d3预告将出视频详解如何在三个月内实现SEO爆发式增长,并征集观众最想了解的内容方向
I'm going to make a video explaining exactly how I exploded SuperX's SEO. What do you want me to cover? https://t.co/vLi4JHGGLr
受众观点:①@scheemunai 和@moonfarm_dev 都追问免费工具和SEO流量之间的关联,是评论区最热方向 ②@Merisdabhi 回复"everything"说明开发者对SEO话题饥渴度高 ③@mjrode 发了类似数据图表示共鸣,说明评论区有类似经历的人
展开评论
- @robj3d3 (6): I made all the important changes + fixes 3 months ago btw. After being stuck for 2 years. https://t.co/zIsZOZDLyy
- @mjrode (4): @robj3d3 Twins 👯♂️ https://t.co/PBP0Hja3M5
- @scheemunai (4): @robj3d3 Would love to see how free tools converted.
- @Merisdabhi (3): @robj3d3 everything 😭 https://t.co/Y4RIx8WG9D
- @moonfarm_dev (2): @robj3d3 How free tools and seo correlate! 🙌
开发者robj3d3因新模型价格降低40%且性能达到Fable 5.1水准决定重新订阅,并吐槽Astra性能被削弱
40% CHEAPER AS GOOD AS FABLE 5.1 Okay you sold me, I’m coming back honey!! Astra was nerfed anyway.
受众观点:①@euboid 自嘲没有哪个订阅取消次数比Claude更多,说明开发者对Claude的爱恨关系普遍 ②@bil0090 提议叠加第二个订阅,暗示Claude和OpenAI双持需求存在 ③@NoobChess1 猜测Codex reset,说明OpenAI新动作也在预期中
展开评论
- @euboid (3): @robj3d3 I don't think I've canceled any sub as many times as my claude sub
- @andrewdex (2): @robj3d3 @thsottiaux right now https://t.co/XLduI20J9Q
- @bil0090 (2): @robj3d3 Dont rush, new OpenAI models coming Altho, I think its time to stack a 2nd Claude sub :)
- @jellycharts (1): @robj3d3 https://t.co/TixUhp2Lf2
- @NoobChess1 (1): @robj3d3 I smell a Codex reset...
ai_for_success整理Claude Opus 5.5发布要点:性能达到Claude Fable 5.1水平,多项基准测试击败GPT-6 Astra,运行成本降低40%
Anthropic just dropped Claude Opus 5.5. > Performs at the level of Claude Fable 5.1 > Beats GPT-6 Astra on most benchmarks. > Costs 40% less to run than Opus 5. https://t.co/rlqOr57RcZ
受众观点:①@CuriousKarthike 对computer use功能升级感到震惊,认为是针对Astra的直接回应 ②@Chaos2Cured 抱怨模型在医疗法律等领域的使用限制,认为只是好玩具而非真实价值 ③@knowixbuilds 质问是否能算AGI,说明部分用户用AGI标准衡量新模型
展开评论
- @CuriousKarthike (1): @ai_for_success Wtf is going on in computer use It is the answer to Astra, I think 😳
- @Chaos2Cured (1): Dude, it can’t help me on anything of value. I had high hopes. Absolutely a dog on value. No medicine No law No fractal coding No biology No chemistry No CPA work… 😬 They can keep making better coding Ai, but when you block everything that is real value, it isn’t a good AI. It i…
- @Kingskqz (0): @ai_for_success bro , you have no loyalty. I was wrong. I thought you were a shill for @ammaar and @OfficialLoganK
- @knowixbuilds (0): @ai_for_success should we call this AGI?
- @mringenious1 (0): @ai_for_success https://t.co/hIdBz1Poqj
开发者robj3d3庆祝Claude上线储备重置功能,允许用户将未用完的使用配额保留以备后用
BANKED RESETSSSS FOR CLAUDEEE!!!!
受众观点:①@itselgemmy 说要把Claude重新加回技术栈,说明此功能是拉回用户的关键 ②@mojomatt 吐槽功能公告刚好在自然重置后1小时,说明用户在密切关注更新 ③@captainPURU 追问按钮在哪,说明功能入口不够显眼
展开评论
- @itselgemmy (0): @robj3d3 Yeah, I think I’m bringing Claude back to the stack 🏃♂️
- @WillRogS (0): @robj3d3 Niiiiiice, happy building
- @mojomatt (0): @robj3d3 Announced 1 hour after my weekly limit reset 😅
- @talk2sunder (0): @robj3d3 are you back now on claude?
- @captainPURU (0): @robj3d3 Neeed, where is the button?
博主分享对某新 AI 产品发布预告片的第一印象,认为配乐颇有《Squid Game》式的紧张压迫感,下周正式发布。
The music gives me Squid Game vibes.
受众观点:①配乐与视频是否完全由 AI(Opus)生成,工具链是什么(@imsxon 直接提问)②作者与产品方的商业关系,是否受赞助(@screenfluent 调侃"他们肯定买了你")③对即将发布产品的期待氛围,评论区情绪正向
展开评论
- @robj3d3 (11): Next week. “Introducing Fable 6”: https://t.co/0ou1iMCHAv
- @imsxon (1): @robj3d3 i wonder if it's made 100% by opus too, music and video itself
- @oiharshit (1): @robj3d3 first thing that i noticed
- @screenfluent (0): @robj3d3 Be honest Rob! They bought you! I always learn about their updates from you! 😆
- @tarasshyn (0): @robj3d3 legit lol 🤣
作者重申两月前"opus 和 sonnet 是最差模型只会产出垃圾"的判断,结果 Opus 5.5 发布同日被自己在回复中亲手撤回,成为 AI 迭代速度的现实注脚
2 months later this is still true opus & sonnet are the worst models you could use, nothing but pure slop generated from them
受众观点:①@ryanvogel 自己回复"update: not true anymore"当场撤回判断,隐含了对 Opus 5.5 品质的认可 ②@sfffghhjkh 用"aged like milk in summer heat"形容 AI 领域观点的极速过期 ③@nathans_kimm 观察到用户对 Anthropic 模型派系化选择的现象,认为模型层级过多导致了非理性选边
展开评论
- @ryanvogel (11): update: not true anymore
- @sfffghhjkh (3): @ryanvogel This aged like milk in summer heat 😭😭
- @iso_morph (2): @ryanvogel insane timing
- @ilovecowsman (2): @ryanvogel For coding they are fine when given the proper instructions. Same with Luna and Terra.
- @nathans_kimm (0): @ryanvogel anthropic's lineup has too many tiers now. everyone picks one model and defends it like a sports team.
iannuttall 基于 typesafeai Jev 搭建了一个免费内链推荐工具,支持 BYOK 或一美元按次付费,输出 CSV/JSON 交给 LLM 批量执行。
I built a free internal linking tool using @typesafeai Jev for classifying and selecting the links. BYOK or pay $1 to use mine. It works for up to 500 pages and gives you a CSV or JSON to pass to an…
受众观点:①实际使用中的 bug 和稳定性问题(@eldelentes 超 500 页报错、@eliehabib 持续连接超时)②输入第三方 API Key 的安全顾虑(@Livvux 的调侃反映真实用户担忧)③工具本身的价值认可度(@IronBrands16 表示每条推文都收藏)
展开评论
- @eldelentes (1): @iannuttall @typesafeai Hey Ian, I’m trying to test the tool, but I’m getting this error. My site definitely has more than 500 pages, so I’m not sure if that’s the reason. https://t.co/E5dfq3oCOk
- @IronBrands16 (1): @iannuttall @typesafeai Mate im literally bookmarking every tweet of yours
- @Livvux (1): @iannuttall @typesafeai YEAH SURE ILL ENTER MY KEY THERE HAHAHAHAHHA https://t.co/wnnRzgVDuy
- @pootlepress (1): @iannuttall @typesafeai Great minds...
- @eliehabib (0): @iannuttall @typesafeai Always ends with this error - "The connection closed before the report was ready" - before reaching 500 pages
展示某产品界面 motion 动效的演示视频,仅用艺术字体呈现单词"motion",评论区有人询问是否与 superlocal 产品相关
𝓶 𝓸 𝓽 𝓲 𝓸 𝓷 https://t.co/7WsEY5jleb
受众观点:①@parkereimerl 询问这个动效是否会集成到 superlocal 产品 ②@Syrigann 认为 Gmail 需要类似的视觉刷新 ③@thalitaacioli 对动效的视觉质感表示认可
展开评论
- @parkereimerl (1): @ryanvogel Is this coming to superlocal?
- @Syrigann (0): @ryanvogel gmail needs a refresh.
- @thalitaacioli (0): @ryanvogel oh this is sexy
独立开发者 iannuttall 宣称 Opus 5.5 发布的时间节点恰好是自己 ChatGPT 使用量降至零的节点,暗示主力 AI 工具已完成切换。
Opus 5.5 has dropped just as my ChatGPT usage drops to 0%. 🔥 https://t.co/xsO1eN9E0E
受众观点:①切换到 0% 的具体触发原因(@andreysuperior 直接提问,说明受众想听具体故事)②OpenAI 强制使用特定 harness 的限制对用户的影响(@marcuswquinn 和 @peterzipper 都表达不满)③GPT-6 Sol 的发布时间线(@wiiiimm 追问)
展开评论
- @marcuswquinn (1): @iannuttall they still forcing people to use their crapware harness?
- @peterzipper (1): @iannuttall GPT-6 Sol and Luna just dropped. OpenAI really knows how to keep me from re-trying Anthropic with their "only Claude Code is a allowed harness" stuff.
- @xmetaguy (0): @iannuttall same, Ive been nursing claude along with opus as I hammered fable. Not sure how to use the banked reset but that will be getting boshed asap
- @andreysuperior (0): @iannuttall What prompted the drop to 0%?
- @wiiiimm (0): @iannuttall will sol 6 also launch today?
iannuttall 阐述其多模型并行策略:以 Fable 5.1 和 Opus 5.5 为主 agent,GPT-6 Astra 和 Sol 作对比基准,中文模型专门承担低成本辅助任务。
Fable 5.1 and now Opus 5.5 vs GPT-6 Astra and soon Sol is why I have to keep subscriptions with both. There just isn't one model to rule them all and I need to be able to switch between them as lead…
受众观点:①多工具并行时跨 session 的上下文和记忆管理难题(@finstratege 提到开着 15+ iTerm 标签,上下文无法在工具间共享)②各模型作为 lead agent 的实际体验差异(@lightninglu10 明确说 Astra 作为 lead agent 很差)③维持多个付费订阅的成本压力(@bitwisebytes 每个平台 $200,合计 $600/月)
展开评论
- @finstratege (1): @iannuttall exactly how i do it, have 15+ iTerm tabs and switch between codex/claude code/deepseek... but i feel like context/memories is not well managed between each, same for you or skill issue? 😂
- @G0R1LLAGRAY (0): @iannuttall Lol classic donkey work.. are you using cursor to orchestrate the lead work and donkey stuff?
- @bitwisebytes (0): @iannuttall Real. I justify my claude $200 plan, grok $200 plan and gpt $200 plan with access to the intel and it's worth it for me. it is expensive but I get to live on the frontier.
- @lightninglu10 (0): @iannuttall Astra is awful as a lead agent, hopefully Sol is better.
- @SthngClvr (0): @iannuttall Put the rejected approaches in the shared project notes too. Otherwise swapping to the smartest new model buys you a very expensive rediscovery of why Plan A didn’t work.
评价 Anthropic Fable 5.1 比 Opus 5 便宜 40% 却性能相当,认为 Claude 重回最强地位,并建议取消 ChatGPT 订阅
Claude is soooooo back... It's 40% cheaper than Opus 5 at Fable 5.1 level. It's might be the only thing can last for entire week now. it's about time to cancel ChatGPT
受众观点:①@techwalrus 用幽默方式点出用户在 OpenAI 和 Anthropic 之间反复切换是普遍现象,暗示这种评价缺乏稳定性 ②@Solaawodiya 质疑判断的持久性,认为用户会随 OpenAI 下次发布再次切换回去 ③@elusiveflame77 指出 GPT-6 Sol 和 Luna 的发布可能已经让这个判断过时
展开评论
- @techwalrus (4): @ziwenxu_ Every week people scream they're canceling the opposite one and then the next week they say the opposite 🤣
- @ziwenxu_ (1): Benchmarks https://t.co/U4rUxkB2Zt
- @Solaawodiya (1): @ziwenxu_ After OpenAI launches, then you'll subscribe back. I've seen this many times lol
- @erictrisvan (0): @ziwenxu_ Yeah I think it is really efficient I used it ultracode and it used only 9% of 5 hours on 100 USD max . Also banked reset there until 22 october. Anthropic heard us.
- @elusiveflame77 (0): @ziwenxu_ Welp openAI just fixed that GPT 6 Sol and Luna
用户 ryanvogel 对 Claude Opus 5.5 模型发出感叹,评论区用户纷纷确认该模型表现出色
okay maybe opus 5.5 is good?
受众观点:①@karthiknish 实测 2 小时 20x Max 模式跑到 96% 用量,关注 Opus 5.5 在高强度使用下的实际表现 ②@jrysana 提到有一个被大家忽视的重要优点,评论区在追问具体是什么 ③整体评论区情绪是意外的惊喜,说明此前预期偏低
展开评论
- @jrysana (0): @ryanvogel Opus 5.5 is amazing for one big reason I don't see anybody talking about
- @EmilioSchwaiger (0): @ryanvogel told you
- @BennettBuhner (0): @ryanvogel IT IS
- @karthiknish (0): @ryanvogel Literally ran the model for like 2 hrs on 20x Max account - My usage is at 96% - Anthropic actually shipped a usable model
Anthropic 官方发布 Claude Opus 5.5 模型,性能比肩 Fable 5.1 且 API 价格降低 40%、速度提升 30%
Introducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5. https://t.co/Q9C2VKQ79f
受众观点:①@claudeai 对齐测试得分最高,评论区关注 safety/pacing 与快速发布之间的矛盾 ②@dfeinition 演示 Opus 5.5 纯代码生成马赛克动画,关注编码能力 ③@elonmusk 点名祝贺带来额外传播热度
展开评论
- @claudeai (4503): Opus 5.5 is our first model since we called for pacing the frontier. As with previous models, it was tested by external evaluators before release, including METR and Frontier Design. On our most comprehensive alignment test, it achieves the strongest score to date.
- @elonmusk (2764): @claudeai Congrats!
- @full_kelly_ (1418): @claudeai we should slow down AI anyway, here’s Opus 5.5
- @dfeinition (821): @claudeai This mosaic is 13,081 tiles, all drawn and animated in code by Opus 5.5, with no image files. The stars' reflections hatch into 22 tiny fish 😌 We're going to need a bigger bowl. https://t.co/DNPJ8us1js
- @nortonbreads (470): @claudeai @LiveSquawk i thought we were slowing down? https://t.co/0y8mRg6e8u
CNBC 报道一家对冲基金用 4 个 AI agent 替代整个人类团队,年薪仅 4 万美元,前团队成本 500 万
CNBC just filmed a hedge fund where every employee is an AI agent. Payroll: $40,000 a year, all 4 of them. His last team cost $5,000,000 and burned him out of the business. Watch him introduce the st…
受众观点:①@CandleHigh 和@nikvassev 追问 returns 收益数据,没有收益就无从判断 ②@bonddonkey 质疑这是 crypto pump and dump 操作 ③@bryanwaldo 指出 AI agent 一旦出错会直接爆仓,关注风险控制
展开评论
- @CandleHigh (60): @antpalkin Can we see his returns?
- @bonddonkey (59): @antpalkin Why do you need 7-8 people to run a pump and dump scam, I mean crypto fund???
- @nikvassev (42): @antpalkin Not mentioning his returns is a little sus?
- @soyylloyd (39): @antpalkin Until he publishes returns, this is meaningless
- @bryanwaldo (28): @antpalkin It's all fun and games until these agent make one mistake and blow him out.
作者整理了一张现代 AI 应用技术栈地图,涵盖 100+ 工具,指出 ChatGPT 只是冰山一角
𝗠𝗼𝘀𝘁 𝗽𝗲𝗼𝗽𝗹𝗲 𝘁𝗵𝗶𝗻𝗸 𝗔𝗜 = 𝗖𝗵𝗮𝘁𝗚𝗣𝗧. Not even close. ChatGPT is what you see. The real AI revolution is the massive ecosystem being built underneath it. This AI Stack Map captures 100+ tools powering mode…
受众观点:①@theantigynocen1 感慨需要用 AI 来管理 AI,工具爆炸带来认知负担 ②@hysenchu 质疑 Hugging Face 分类逻辑 ③@Bart_Christner 指出 Grok 没被收录引发完整性讨论
展开评论
- @theantigynocen1 (2): @triptitips I am struggling to keep up with it. I need AI to manage AI for me. 😅
- @hysenchu (2): @triptitips What? HF is an LLM? Does HF even know that? @huggingface
- @Bart_Christner (2): @triptitips no grok?
- @BachelotBernard (1): @triptitips where are @grok and @bot ??
- @aiautobusiness (1): @triptitips Yes
TradingView 正式发布官方 MCP Server beta 版,允许 Claude 和 ChatGPT 直接访问 OHLCV 市场数据
TradingView yeni bir kapı açtı: Resmî MCP Server artık beta olarak kullanıma sunuldu. Bunun anlamı şu: TradingView verileri ve bazı platform araçları artık ChatGPT, Claude ve MCP destekleyen diğer ya…
受众观点:①@Tradesdontlie 透露 6 个月前已做了非官方 CDP tradingview-mcp 并有 100 万克隆 ②@Donnieblocks 指出目前只有 getter 功能,实用性有限 ③@TurtleTrder 提出整个市场受 AI 控制的哲学质疑
展开评论
- @sabanci_ata (12): Claude ile TradingView’in resmi MCP bağlantısını adım adım anlattığım video eğitimini burada paylaştım. Bağlantıyı kurmak isteyenler videodaki adımları takip edebilir: https://t.co/CLn4m98i69
- @cevikfinance (8): @sabanci_ata Güzel bilgi, teşekkürler. Hemen deneyelim. 🙂🙏🏼
- @Tradesdontlie (6): @sabanci_ata wow!!! i have over 1 million clones on my unofficial CDP tradingview-mcp i made 6 months ago they finally made one!!! awesome!!!
- @Donnieblocks (4): @sabanci_ata Seems pretty useless for now. Mostly getters which is very TV. Would be cool if they actually let u manipulate stuff on the chart or indicators
- @TurtleTrder (3): @sabanci_ata bu traving platformda kendilerinin. Dünyadaki tüm borsalar tek bir yapay zekaya bağlı.Suan piyasada olanlar hepsi google ileri versiyonu yani yapay zeka değil.Zaten kendide evrimini tamamlayali 10 yil oluyor.Sistemin size izin verdiği kadar kazanırsıniz hepsi bu
Anthropic 官方在 r/ClaudeAI 发布 Claude Opus 5.5 公告,325 条评论集中讨论限额重置、Haiku 更新和模型改进
Opus 5.5 performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5. Opus 5.5 is our first release since we called for pacing the frontier. External evaluators, i…
受众观点:①@JoNike 高赞引用限额重置可以保存并自行决定使用时间的新功能 ②@Temporary_Idea8880 热烈欢呼 Haiku 终于要更新了 ③@RadiantTea7445 对 reset saving 功能的欢迎说明用户对限额焦虑是长期痛点
展开评论
- @JoNike (399): > In addition to the price drop, we’re increasing five-hour usage limits on Pro, Max, and Team plans. We’re also providing subscription users a rate limit reset, which you can now save and use whenever you choose.
- @Temporary_Idea8880 (303): HOLY SHIT WE ARE SO BACK WE FINALLY HAVE HAIKU https://preview.redd.it/17tj069xm3rh1.png?width=920&format=png&auto=webp&s=fe102669a66ef294bc5736417f0f796621c8df86
- @ClaudeOfficial (275): https://preview.redd.it/my3pftkxm3rh1.jpeg?width=2160&format=pjpg&auto=webp&s=c3a353cc1d05231ddb698f8b11c26c6b1ab90284
- @Temporary_Idea8880 (145): Seems like it https://preview.redd.it/ia8q6nnnn3rh1.png?width=1131&format=png&auto=webp&s=761221b216874e124ea6637a2661dca8e3348977
- @RadiantTea7445 (140): finally they are doing reset saving!!!
r/ClaudeAI 用户讨论 Opus 5.5 发布最希望改进的问题,219 条评论聚焦写作风格和长任务目标焦点两大痛点
For me, 1. A return to human-like writing instead of the current word vomit 2. Better goal focus during long-running tasks.
受众观点:①@Kraien 高赞直接说 tone 再也受不了了是最强单一共鸣 ②@ischmal 用幽默方式演示 Opus 5 在确认微小细节上的冗余表达习惯 ③@lollllllops 说切换到 Astra 做文案效果更好暗示市场份额流失
展开评论
- @Kraien (358): The tone, I can't stand it anymore.
- @TastyVermicelli3140 (176): Two things and the second is better than the first but not for the reason you think!
- @ischmal (162): All four changes have been implemented. Two decisions came up, and it's your call: >\> something completely trivial \> something else completely trivial I've added both into plan.md. Say the word and I'll fix it.
- @Make-life (80): Seriously. Its ChatGPT all over again.
- @lollllllops (80): Ironically, I’ve switched over to Astra for copywriting and my god is it sooo much better.
r/webdev 高热度讨论公司如何在大规模 AI 投入下维持盈利,核心结论是省钱方式是不回填职位而非 AI 提升效率
We can all agree AI has become a standard now. Every employee at the company I work for, has AI token allowance ranging from 500-2500$. And its probably similar across the industry. What I am curious…
受众观点:①@disposepriority 直接说公司把 AI 预算当作以后可以裁员的赌注 ②@tnsipla 反问到底有没有在招人或回填职位 ③@gfxlonghorn 用真实例子说明同事离职不回填省下 1.75x 工资比 AI 支出划算
展开评论
- @disposepriority (203): It isn't. Companies investing this much (and more) are hoping they are able to start firing people soon, the ones that run out budget before they're able to do so just cut back on usage. Faster "shipping" does not mean more money, despite what the clueless people on social media…
- @tnsipla (136): What’s your growth budget for hiring new developers look like? Have you backfilled any roles recently if someone left? Where you think the money from not growing teams and not backfilling roles go?
- @QualitySoftwareGuy (96): > Faster "shipping" does not mean more money This right here is what most don't seem to understand. I believe the faster shipping from vibe coding will eventually cause the company to *lose* money over time.
- @gfxlonghorn (70): I work for a profitable and successful company and we simply aren’t backfilling by default when people leave. My counterpart was making 1.75x my salary and he left, and the org isn’t going to replace him. That’s a hell of a lot more money saved than my 30k a year AI spend.
- @Hockeynerden (53): I don't even think my company has employed a new dev in 3 years... Which is insane when I think about it
Portainer 宣布切断免费版与付费版升级路径,自托管社区强烈不满并纷纷迁移到 Komodo 和 Dockhand 等替代方案
For the people that are unaware or haven't seen the news yet.
受众观点:①@user6035 直接推荐迁移到 Komodo ②@archdukemovies 分享今年早些时候就切换到 Dockhand ③@PatBanglePhoto 引用 Portainer 不知道自己处于饱和生态系统中
展开评论
- @user6035 (125): I would recommend komodo
- @archdukemovies (110): Glad I switched to Dockhand earlier this year
- @Reddit_is_fascist69 (76): I just dropped portainer. Just raw doggin those docker-compose.yml and .envs.
- @PatBanglePhoto (57): “Portainer doesn’t realize it exists in a saturated ecosystem.”
- @ActuallyFullOfShit (32): Yeah it's really not hard.
Anthropic 随 Opus 5.5 推出可保存并自由使用的限额重置功能,评论区热烈讨论这对重度用户的策略影响
Honestly kinda scared to use it right now ,as I have a feeling it would be a bit of a waste considering it resets either way tomorrow for me.
受众观点:①@lost_in_trepidation 表示终于是可切换的不喜欢没用完就自动重置 ②@DaveWoodX 认为 banked resets 比模型更新本身更重要 ③@N19h7m4r3 分享以前睡着期间重置白白浪费的场景
展开评论
- @lost_in_trepidation (45): I'm so glad it's a toggle. I hate getting auto-reset when I haven't used up my weekly usage.
- @Plane_Lavishness5909 (19): Who is going to press the button and find out if its a weekly or 5hrs reset :P
- @DaveWoodX (15): Banked resets (if it's not just a one time thing) will be a bigger feature than the model update itself I suspect! Thanks Anthropic!
- @TextbookGPTsMod (12): It’s a full reset, weekly and 5 hourly
- @N19h7m4r3 (8): I reset on thursdays. Plenty of resets have happened a few hours before my normal reset, while I'm asleep lol This is the biggest improvement to Claude in a while.
SpacePlanner 作者分享项目从每天 500 用户爆发到每天 1.5 万用户的过程,6M+ 粉丝博主 SetupsAI 发布视频是主因
^(GA4 capturing the start of the first wave) so this week has been insane and i honestly need help understanding what happened. quick backstory. a while back i wanted to plan a space in my home and e…
受众观点:①@One_Progress_1044 分享 Where did you hear about us 弹窗 2 天获 5k 回复说明用户直接调研的有效性 ②@walm00 发现安全漏洞暴露了 166 个用户邮件引发数据安全讨论 ③@Nalmyth 提醒检查 Supabase RLS 特别是 vibe coded 项目
展开评论
- @One_Progress_1044 (26): https://preview.redd.it/6xnj099c80rh1.png?width=2524&format=png&auto=webp&s=af96f9403a9f76e1389bcae1372dfcc5c9667275 **TIP:** the “Where did you hear about us?” popup got almost 5k responses in two days, so yeah, that was actually pretty useful.
- @Biupildamian (23): in case you didnt notice, your images are not showing in your site and you have a few errors in devtools>console https://preview.redd.it/5jord62ud0rh1.png?width=1913&format=png&auto=webp&s=0de1ce5763f470fc92287f8a81221d087fd978aa
- @CaptJan (12): Cool. Looks interesting - I bookmarked it, as I don't need it now. Perhaps, put up a 'buy me a coffee' donation to help defray your costs and/or limited ads to keep it free?
- @Nalmyth (10): I hope you've locked down your supabase with good RLS? It's typically not well secured, especially if you vibe coded the project
- @walm00 (9): https://preview.redd.it/eh7x8ldlw1rh1.png?width=819&format=png&auto=webp&s=4666d27e1ed2fe4f95701d0cae39c0bf42b207b9 Quick security test found some issues and shows 166 users and their emails :)
r/ClaudeAI 第一时间发现 Opus 5.5 发布并讨论新的 Responsive Mode 和 token 消耗变化
Also, check out the new responsive mode New mode where Claude replies to you in a sentence before thinking or using tools, on every prompt. Edit: are you seeing fewer tokens being used? __It performs…
受众观点:①@sanat_naft 和@fsharpman 第一时间确认在 app 中看到了 Opus 5.5 ②@DavidLuky 引用官方发布关键句它写作的方式就像我自己打动了很多用户 ③@LetTheRiotsDrop 讽刺 Anthropic 即将重置性能说明社区对优化退化循环有深刻创伤
展开评论
- @sanat_naft (31): I have opus 5.5 available in app
- @eslobrown (19): https://www.anthropic.com/claude-opus-5-5
- @fsharpman (16): Confirmed https://preview.redd.it/auc4r938m3rh1.png?width=1077&format=png&auto=webp&s=b08fcc8973f994cc1e6fb55f050eaab4be6055d3
- @DavidLuky (16): This is all I needed to read (from their website) **Communication.** Opus 5.5 communicates more naturally than prior models. Early testers found its writing clearer and easier to follow, which addresses some of the common feedback we heard about Opus 5. It puts the most importan…
- @LetTheRiotsDrop (16): Inb4 reset
r/ClaudeAI 讨论 Opus 5.5 降价 40% 速度提升 30% 是否可信,评论夹杂对 Opus 5 失望和对新模型期待
r/ClaudeAI Opus 5.5 is 40% cheaper while being 30% faster than opus 5.
受众观点:①@Front_Raspberry_6488 指出 Opus 5 发布时号称接近 Fable 但实际沟通极差用户浪费大量 token ②@Neon_Camouflage 认为 Fable 更适合 subagent 团队执行路线图 Opus 5 做项目不错就是沟通差 ③@ManikSahdev 分享提前 3 天测试 Opus 5.5 的真实感受远超 Opus 5
展开评论
- @Front_Raspberry_6488 (147): Remember, when they first released Opus 5, they kept claiming how cutting-edge the model was, stating it was nearly on par with FABLE 5 and had better token efficiency. However, after using it for a while, people discovered that Opus 5 couldn't even communicate properly. The bac…
- @Fusseldieb (47): If it maintains the SAME level of intelligence and creativity as Opus 5, then yes, it is huge. Now, if it's indeed BETTER than Opus 5, then it's ***even huger***. But honestly, I doubt it. Maybe in the first few days, and then they'll dial it down, like they always do.
- @Neon_Camouflage (43): Honestly the communication is the whole issue. Fable is better at something like feeding it an entire roadmap and having it run subagent teams to execute each milestone. As far as most project work goes, Opus is fantastic, just until you have to talk to it.
- @crusoe (33): Opus 5 does good work. Just not for non technical writing. Ooof.
- @ManikSahdev (13): It's better than opus-5 in every way. I've had this slug for 3 days in one account and then it left and then came back in second account but on chat.ai. This is arguably better than fable 5.0 if I'm being real, like the model is somewhat fable 5.1 level in my experience, and it…
独立开发者发现代理机构把自己的开源项目 SevenGrid 冒充为代理机构作品并附伪造客户评价,社区施压后撤下
**EDIT: They took it down! Maybe because of the Attention of this Community, thank you all!** **Here is the fullpage screenshot:** [sakura-portfolio-claim-2026-09-22.png (1440×5982)](https://sevengri…
受众观点:①@EliSka93 指出整个网站明显是 AI 生成内容认为是 AI slop 代理机构操作 ②@horrificammonia 提供 DMCA takedown 实操建议 ③@EliSka93 通过联创 GitHub 共同使用的 AI 生成图片确认了身份
展开评论
- @EliSka93 (82): I mean... The whole site is clearly written by AI. I bet AI slop builds is the whole operation.
- @Low-Philosopher-2993 (49): I'd sue them. This is absolutely crazy!
- @horrificammonia (36): Don't even need to sue yet, a simple DMCA takedown to their host and search engines works wonders for getting fake portfolio entries scrubbed.
- @Sneedle-Woods (26): will try the DMCA takedown and the Google's legal removal form! Thank you!
- @EliSka93 (22): Here's the GitHub of their Co founders https://github.com/vladimirkurilo I know it's the same guy because he uses the same ugly ass AI "characters" on his GitHub pages.
表单工具 Tally 发布 6 年复盘:$6M ARR,43% 新用户来自 AI 搜索,30% 表单由 AI 构建,MCP server 持续增长
You raise, you build, you grow, you exit. That's the startup script. We're trying to write a different one: our form builder, [Tally](https://tally.so), just reached $6M ARR, bootstrapped and purely…
受众观点:①@ConditionClassic1546 询问团队是否对产品方向产生过怀疑 ②@Marie-Tally 分享产品多次受益于行业波浪的经历 ③@StevenEgen 作为用户确认免费方案功能足够丰富
展开评论
- @ConditionClassic1546 (5): Motivating post to start my day! Congratulations.. by the way, was there ever a moment when you had doubts or fears about the product? For example, did the idea seem “it makes no sense” at first, or did you worry that the AI boom might cause problems?
- @Marie-Tally (5): go for it, I documented our very first steps here: [https://blog.tally.so/year-1-how-we-bootstrapped-tally-to-11k-users-and-5k-mrr/](https://blog.tally.so/year-1-how-we-bootstrapped-tally-to-11k-users-and-5k-mrr/)
- @Intrepid_Pen_9725 (4): wow, congratulations 🎉
- @StevenEgen (3): I use your forms and I find them useful, simple, and easy to use, and even the free options have a bunch of features; it works great for you guys and for me, thanks.
- @Marie-Tally (2): Yes often, but somehow we survived and actually benefitted from multiple "waves": * No-code: We launched during covid, when "no code tools" were booming * Notion: We're the preferred tool for Notion users as tally integrates nicely and is inspired by them * Vibe coding: Tools li…
LinkForge 创始人分享真实收入 $0 的截图意外获得 11.1 万次曝光,引发关于展示失败与展示成功的社区讨论
This is probably the funniest thing that has happened since I started building. Months of coding, shipping features, fixing bugs, and trying to get users. Then I posted my actual revenue: $0.000 And…
受众观点:①@shohaibmk 说人们希望你做得更好但不比他们更好揭示社区心理 ②@Due-Marsupial-778 发起 Zero Dollar Founder 趋势倡议 ③@LinkForgeHQ 诙谐回复维持真实人设
展开评论
- @shohaibmk (19): People like to see you do better, But not better than them
- @LinkForgeHQ (3): Honestly, I think there’s some truth to that 😂 People can support the journey more easily when they feel like they’re in the same boat. The moment you start doing significantly better, the energy can change.
- @Due-Marsupial-778 (3): Let's create a trend Zero dollar Founder 💸🥶☠️🤣🤣🤣
- @LinkForgeHQ (3): Haha, I wish 😂 For now, just a guy trying to turn $0 revenue into $1.
- @Shock-Successful (2): In the same boat, ngl I sometimes feel that way especially against competition in my niche but I try not to let it get to me
r/SaaS 讨论如果 SaaS 突然获得 3 倍用户时哪个环节最先崩溃,覆盖基础设施、计费、客服等维度
It's interesting how different parts of a SaaS handle sudden growth. It's easy to focus on infrastructure, but support, onboarding, internal processes, or even billing can become bottlenecks too. For…
受众观点:①@elinhakala 分享真实经历:从 200 到 700 注册用户时 Stripe 没问题但 proration 在月中升级时出错 ②@arthaudm 指出 support 在 exception routing 上才会先崩 ③@Similar-Storm4432 担心数据库连接限制瓶颈
展开评论
- @Jealous-Dance-8007 (14): My bank balance LOL. Everything else is already set for scale.
- @Similar-Storm4432 (9): For me 3x of 0 will be still 0 so nothing to be worried about 😂 But, jokes aside (even though it was not a joke) I think the database tier that I have chosen for now for sure need to be updated, the number of connections are limited and eventhough the service itself can scale ho…
- @elinhakala (4): billing, every time. i had a client go from about 200 to 700 signups in a month and stripe was fine, the app was fine. what broke was proration on mid cycle upgrades, we were doing it manually in a spreadsheet and nobody noticed until the invoices went out wrong. support held up…
- @arthaudm (3): support breaks at exception routing before raw ticket volume. 3x customers means more billing edge cases, account merges, weird permissions. if every exception lands with the founder, the app can stay up while the team still melts.
- @Emotional-Start7994 (3): 3x more customers? It'd still have 0 customers then
r/SaaS 讨论没有免费试用时用户是否愿意直接付费,聚焦 SaaS 定价中试用期设计的必要性
Hi there. So I am thinking - free trial or no free trial. I myself never pay if I don’t get the free trial, but I notice that there are people which pay just to try the tool. What about you and your…
受众观点:①@Taniai_ 分享 FranklyMail 无免费试用 40 天获得 60 个付费用户 ②@Odd-Impress-985 认为价格合理解决真实问题就愿意付费 ③@CroFran 明确没有试用或 demo 就直接离开
展开评论
- @Odd-Impress-985 (10): I'd probably pay for a tool without free trial if it solves a real problem and the pricing is reasonable. A trial definitely makes it easier to take the first step though.
- @Taniai_ (8): My product FranklyMail was abusable as nature of it - so I can't provide free trial and to sustain and guarantee email deliverability for all other users I only want serious intention people inside the platform. So I don't have a free trial and in 40 day I have 60 paying user.
- @Predream1 (4): The two groups in your post aren't really different kinds of buyers, they just meet two different products. Something that proves itself in the first session can charge upfront without losing anyone who was serious. Something that's useless until their own data is in loses them…
- @emma_lorien (2): \+1 a free trial definitely lowers the barrier for me. If there isn’t one, a clear refund policy is the next best thing.
- @CroFran (2): I need to know how it works first. Ive had a couple of times where i went head first or maybe used the free version and then paid and felt not scammed but done for. Now if someone cant give me a free trial or atleast let me book a 10-20 min demo before buying im out.
r/ClaudeAI 用户质疑 Anthropic 以 Max 推理档位做基准测试但多数用户使用 Medium/High,Max 消耗 6 倍 token 对大多数人不实用
https://preview.redd.it/rmsgatzzm3rh1.png?width=737&format=png&auto=webp&s=9666afb4371d440730ab75d40a2e605d8c78b574 Opus 5.5 now has a Max effort mode that uses **6x usage.** So if Anthro…
受众观点:①@rlemmie 纠正 6x 是相比 Medium 而不是 High 因为默认档位从 High 降为 Medium ②@acutelychronicpanic 用考试改答案比喻说明 Max 过度思考往往反而出错 ③@versaceblues 明确 High/xhigh 才是性价比最高区间
展开评论
- @rlemmie (88): Notwithstanding anything else, the "6x or more" is just a consequence of making "medium" default, rather than "high". So you're looking at 6x more than medium, rather than 1.5x (or whatever it was) more than high.
- @acutelychronicpanic (39): Overthinking and second guessing. You ever answer a question on a test, change your answer three times, then find out that the first answer was right?
- @Pantheon3D (29): Idk but you can see xhigh vs max here https://preview.redd.it/v157z9oaq3rh1.png?width=1092&format=png&auto=webp&s=3398ba0bb4d3774bf6c10835eb6270e1f2de4978
- @versaceblues (17): Tbh there is very little reason for you to ever use max thinking. High/xhigh is the sweet spot for getting intelligence per token. Max I find tends to overthink and cost you more. I rarely turn it on.
- @TigerConsistent (14): good point
r/webdev 讨论 AI 辅助编程时代是否应该让 AI 审查 AI 输出、跳过人工代码审查,评论区强烈反对
Hello, I’ve been coding since not so long ago, just before gpt 3.5 got released. So I learned how to code before AI but still is one the people that uses it heavily to code and also an advocate of no…
受众观点:①@magenta_placenta 系统反驳 AI 审查 AI:相同训练分布导致共同盲点 hallucination 循环是真实风险 ②@modsuperstar 讽刺同事是在把自己包装出局 ③@talivs 建议持续小幅度尝试新工具而不要固守或完全放弃
展开评论
- @modsuperstar (20): He’s just further down the path of clowning himself out of a job, that’s all.
- @talivs (19): It will be outdated again every week. Just find something that works for you and that's fine. But try not to stagnate either tbh, even if it's just learning something else or trying another augmentation to your existing workflow now and then to see what sticks. At the end of the…
- @magenta_placenta (19): No, your workflow is not outdated. Blindly trusting AI without reviewing code output is a dangerous anti-pattern that can lead to severe tech debt, hidden bugs and security vulnerabilities. "AI Reviewing AI" can introduce blind spots. Letting one model review another model's out…
- @RepulsiveWall (8): Bro if you’re getting paid and shipping good work, you could write code on stone tablets with a chisel. Your workflow isn’t outdated until it stops working.
- @Joe_Spazz (5): Oh yes, because an AI reviewing another AI's code is certainly always going to catch the mistakes that AI makes. And will certainly catch all the little things that humans know about the nuances of how to actually use a piece of software versus just designing something that tech…
r/SaaS 讨论独立开发者构建容易销售难的老话题,评论区反驳构建本身也不容易
r/SaaS Building is easy, selling is hard
受众观点:①@Mission_Hold8056 和@nonHypnotic-dev 同时反驳构建也不容易 ②@Charliex1337 指出以为构建容易结果上线满是 bug ③@Thrakanox 指出开发和营销是完全不同技能集
展开评论
- @Mission_Hold8056 (15): Actually building is not easy neither
- @nonHypnotic-dev (13): both of them are not easy
- @Charliex1337 (12): You think building is easy and then release app with ton of bugs
- @Few_Bullfrog6695 (5): Building is not an easy task, but I found that finding users is even more challenging. Perhaps it’s just me.
- @Thrakanox (3): They are completely separate skill sets. Most people making a saas probably have years on the Dev side already. Zero on marketing.
r/SaaS 高赞讽刺图调侃独立开发者"为想象中的客户免费打工"的普遍困境
r/SaaS So real
受众观点:①@ClemensLode 质疑为什么有人会为想象中的客户免费打工 ②@johnappsde 指出做别人不要的 SaaS 是学习过程的一部分 ③@NudaVeritas1 晒出 21 个域名 107 个未完成项目 0 个客户引发强烈共鸣
展开评论
- @ClemensLode (120): I never understood why people would work for free for imaginary clients.
- @axum6 (48): I should frame this and hang it in my startup's office 🤣
- @johnappsde (37): You have to start somewhere. Launching that SaaS nobody wanted is part of the learning process
- @shaman-warrior (15): All clients are imaginary until they pay
- @NudaVeritas1 (14): 21 bought domains 3 LLM subscriptions 107 unfinished project repositories 0 clients edit: /s
r/ClaudeAI 用户宣告 Opus 5.5 上线,评论区讨论 Opus 5 的神经质文字墙是否消失以及和 Fable 的定位关系
With price reduction too!!! [https://www.anthropic.com/claude-opus-5-5](https://www.anthropic.com/claude-opus-5-5)
受众观点:①@Ex-Novo-Doc 怀念 Opus 5 的神经质文字墙和神秘抽象有趣的反向情感 ②@throwawayacc201711 提出 Fable 存在意义的问题 ③@beneficialdiet18 判断 Fable 5.2 会更好且更贵
展开评论
- @Ex-Novo-Doc (58): Goodbye, Opus 5. I will miss your neurotic walls of text, full of self-doubt and mysterious abstractions! Let's seriously hope 5.5 is better in this regard...
- @rajsharm404 (25): Haha facts. The team has mentioned it is made on community's feedback, so I genuinely hope it is better. https://preview.redd.it/6wakrd1yr3rh1.png?width=877&format=png&auto=webp&s=e5937aca35d7d07ed824037e1789a4e8260fdc2a
- @throwawayacc201711 (22): So what’s the point of fable is opus does better than it and is cheaper?
- @beneficialdiet18 (18): Fable 5.2 will be better and more expensive probably.
- @ansa94 (16): Yeah Nevermind its banked now :)
创业者分享将 beta 从免费改为付费后获得第一个付费客户的经历,并寻求可预测用户获取渠道建议
Quick update on our SaaS. When we started, we did the classic thing: build the product, launch it, and wait for people to show up. Spoiler: they didn't. A few signups on the waitlist, almost no repli…
受众观点:①@arthaudm 指出价格过滤紧迫性而非本身创造反馈还需定义周使用承诺 ②@1000xhuman 建议先在一个渠道建立转化循环再扩展 ③@Weekly_Reporter1985 分享 Threads 算法给更多曝光的出人意料经验
展开评论
- @arthaudm (3): paid beta works because price filters for urgency, not because payment creates feedback by itself. i'd still define one weekly usage commitment + one cancellation question. the first customer is useful, but the first renewal tells you whether the problem repeats.
- @Ok_Gur_9033 (3): One customer cannot tell you which of the two changes did it. You changed the price and you changed who you were willing to ask, in the same move. I got caught by this a few weeks ago. I had a number I kept reporting as proof one of my pages was winning a query, and when I final…
- @Weekly_Reporter1985 (2): You have to try diff channels and then double down on the one that works for you. For me it was Threads, even though I started with just a handful of followers. Somehow their algo gave my posts visibility. Where I had 2k+ old twitter followers, I got close to zero engagement...
- @1000xhuman (2): linkedin/tiktok are separate beasts. pick one, turn the problem posts into a repeatable series, and track conversations -> demos -> paid, not views. add the second channel only after one loop works.
- @h_2575 (2): Well, first of all congrats. Make the landing page show french if the ip is in france, otherwise, english. What I see with other apps is they offer masterclasses or webinars on a topic to provide some value (trigger/pain point driven) but for full relief they offer than the plat…
小米发布 MiMo v2.6 大模型,RL 训练 post-training 成本 350 万美元,带有实时 benchmarking 仪表盘
The total RL training cost was $3.5M. The model comes with a live benchmaxxing dashboard. https://preview.redd.it/89uurv5r21rh1.png?width=1518&format=png&auto=webp&s=e5d641ef941882e6dc5c0…
受众观点:①@Salt-Bodybuilder-518 纠正帖子说法指出 350 万只是 post-training 成本 ②@ummitluyum 指出 Pro 版 30 个步骤就烧掉 250 万完整费用会更高 ③@ummitluyum 还指出本地运行需要 4 张 H200 让大多数开发者无法使用
展开评论
- @Salt-Bodybuilder-518 (31): The post-Training was $3.5M. Not considering pre- and mid-training…
- @ummitluyum (7): They burned nearly 2.5 million just on thirty steps for the Pro version. Adding the pre-training run over 48 trillion tokens would push that total into the stratosphere
- @lunatic_god (6): Wait till i spend days trying to make it work on my single dgx spark and after a few benchmark move on to a different "Frontier" model.
- @ummitluyum (6): I love reports where pre-training on tens of trillions of tokens gets swept aside so they can brag about pennies spent on fine-tuning With a one million token context and a model size around 180 gigs running it locally without a cluster of four H200 cards is out of the question…
- @MDSExpro (4): ? H200 is 141GI, 2x H200 is 282GB, that's enough for model AND context
SurfSense 是开源自托管 NotebookLM 替代品,支持离线运行无遥测本地嵌入,可通过 Docker 或桌面应用部署
I'm one of the maintainers of SurfSense, so treat this as self-promotion. I'm posting because a few threads here have asked for a self-hosted NotebookLM and nobody ever answered them, and because I w…
受众观点:①@holyknight00 指出 SurfSense 消耗 token 过多自己本地 opencode scaffolding 效果更好 ②@FurtiveMirth 说项目积极开发中第一版本需要优化 ③@holyknight00 提到自己也在产品化类似研究工具
展开评论
- @FurtiveMirth (5): Yeah its actively developed though.
- @ubiquae (3): What are the underlying technologies to extract data?
- @holyknight00 (3): I gave it a try some weeks ago, I really liked the UI, but it chugs tokens like crazy, and the actual results were not that good. I feel it does something similar like research within claude code that just spawns 200 subagents doing things without any triaging of the complexity…
- @FurtiveMirth (3): Yeah i have a 6GB vram, its working good though. I am using qwen3:8B model. Though it needs a bit more optimization, its just a first release for us. You can try out and please let me know what do you think.
- @FurtiveMirth (3): Thanks a lot. I would be honest, we are simply an application that you download on any platform be it linux, mac or windows. As per my knowledge open notebook requires a docker setup for it to work. If you prefer using an application over docker setup, please try it and let me k…
Mole 作者 Tw93 分享将 GitHub 60K star 的 CLI 磁盘清理工具做成 Mac GUI 应用的产品化经验,核心是删除前预览设计
I'm Tw93. About a year ago I wrote a small Mac cleaner for the terminal and put it on GitHub mostly for myself. Mole CLI grew past 60K stars, and the emails kept saying the same thing: my parents use…
受众观点:①@hosooislake 认可删除前预览设计指出太多清理工具过于激进 ②@Hitw93 作者主动表示愿意回答 CLI 到 GUI 迁移过程问题 ③评论区有删除帖说明有人尝试讨论
展开评论
- @Hitw93 (2): Author here. Happy to answer questions about the CLI to GUI jump or the preview-before-delete flow.
- @hosooislake (2): Transitioning from CLI to GUI because of user demand is a great problem to have. Your point about 'review before delete' is spot on—too many cleaners aggressively wipe out important caches without asking. It's refreshing to see a tool that prioritizes user control over just blin…
- @Hitw93 (1): Yes, everything here prioritizes user data security first, and then cleans thoroughly.
- @[deleted] (1): [removed]
- @Hitw93 (1): Thank you for your liking, and thank you very much for using it.
foundercrm Chrome 扩展通过自定义 AI 流程自动发现 Reddit 等平台的相关互动机会,本地运行使用自带 API key
I like making cool little apps with ai, but I get lazy about promoting them, so I built a free chrome extension that fins your people to engage with every day. You just describe a flow you want it to…
受众观点:①@CompetitiveBid125 指出 Chrome 扩展比网站更难推广 ②@Realistic-Bench-6642 欣赏本地运行和 BYOK 设计去掉了对工具的信任门槛 ③@tskull 说明上下文帮助构建个性化流程而非通用关键词搜索
展开评论
- @CompetitiveBid125 (2): A chrome extension is harder to sell than a website.
- @Realistic-Bench-6642 (1): This is actually a clever approach, using ai to find engagement opportunities instead of writing copy. I tried something similar with a script that scraped reddit for relevant posts but maintaining it was a pain, so an extension with a daily flow sounds much nicer Curious how th…
- @tskull (1): its just in the harness, it just asks questions to understand your product or what you are trying to do better, then it uses that to build personalised flow for finding content. I think the context helps it do a better job than generic keyword search and yea, should out to BYOK,…
LazyDraw 是终端内 ASCII 图编辑器,专为给 AI agent 展示结构化草图和 UI 布局而设计
Sometimes it is really useful to explain to your agent what you want visually. I've used https://asciiflow.com for a while, which is an amazing tool, and I thought, what if we could use it in the ter…
受众观点:评论区仅 1 条互动数据极少,但作者明确说明核心场景是向 AI agent 视觉化展示设计意图
展开评论
- _无评论_
Clueso 发布 MCP 版本,允许通过 Claude 和 ChatGPT 等 AI agent 对话方式完成视频从脚本到剪辑的全流程制作
Clueso
受众观点:①创始人 Neel 描述用户把 Clueso 用于屏幕录制之外的场景推动产品方向转变 ②评论区关注发布 MCP 是聪明的分发策略 ③创始人提到文化视频、网络研讨会动画等多种创意用例
展开评论
- @Overview (0): * [Launches2](/products/clueso#launches) * [Reviews24](/products/clueso/reviews) * [Alternatives](/products/clueso/alternatives) * [Customers](/products/clueso/customers) * [Built with](/products/clueso/built-with) * [Forum](/p/clueso) * More This is the 2nd launch from Clueso.…
- @Clueso (0): Maker 📌 Hey Product Hunt Community! 👋 I’m Neel, co-founder of Clueso. We launched Clueso here last year with a pretty simple idea: **turn rough screen recordings into polished product videos in minutes.** The response was incredible. Thousands of people started using Clueso, and…
- @Clueso (0): Maker [@orkhan\_ilyas](https://www.producthunt.com/@orkhan%5Filyas) Thank you! Upvote Report Share 11h ago [](/@agzee) [Aurangzeb A. Durrani](/@agzee) [Kill Ping](/products/kill-ping) [@neel\_balar](https://www.producthunt.com/@neel%5Fbalar) Congratulations on the product launch…
- @Clueso (0): Maker
- @Clueso (0): Maker [@vipul\_kumar1280](https://www.producthunt.com/@vipul%5Fkumar1280) many unique use cases actually: culture videos for hiring, animated posters for webinars/panel discussions, intro videos for some video games, explainer videos for a cad model, etc The possibilities are re…
WZRD 是将静态文档、表格、表单转化为可交互 AI 体验的无代码工具,用户可以与自己的文件对话
WZRD
受众观点:①创始人 Ali 描述核心问题"我们创造信息却把它困在静态界面里"引发共鸣 ②用户询问团队能否控制 AI 允许回答的内容边界 ③创始人分享兄弟创业背景增加情感维度
展开评论
- @Overview (0): * [Reviews](/products/wzrd/reviews) * [Alternatives](/products/wzrd/alternatives) * [Built with](/products/wzrd/built-with) * [Team](/products/wzrd/makers) * [Awards](/products/wzrd/awards) * More Launch tags:[Productivity](/topics/productivity)•[No-Code](/topics/no-code)•[Desig…
- @Alimo (0): [DOO](/products/doo-2) Maker 📌 Hey Product Hunt! I’m Ali, founder of WZRD. I discovered the internet in 1995 and started building websites when I was 12\. After 25+ years building on the internet, one thing still feels strangely unchanged: **We create information, then trap it i…
- @Alimo (0): [DOO](/products/doo-2) Maker
- @Alimo (0): [DOO](/products/doo-2) Maker [@sheikh\_umair1](https://www.producthunt.com/@sheikh%5Fumair1) Good question. That’s definitely an interesting test case. We’re still improving how WZRD handles more complex formatting and tables, but would love to hear your feedback! Upvote (1) Rep…
- @Alimo (0): [DOO](/products/doo-2) Maker
SereneDB 是开源数据库工具,在单个普通服务器节点上实现企业级超快全文检索和分析能力
SereneDB
受众观点:①@GitGuardian 指出 AI agents 扩展流量后数据库瓶颈是无法靠 prompt 解决的生产问题 ②@InterSub 分享以 vibe coding 方式重新探索 CS 基础知识的故事 ③评论区关注 AI agent 时代单节点做搜索和分析的新需求
展开评论
- @Overview (0): * [Reviews2](/products/serenedb-krummelanke/reviews) * [Alternatives](/products/serenedb-krummelanke/alternatives) * [Team](/products/serenedb-krummelanke/makers) * [Awards](/products/serenedb-krummelanke/awards) * More Free Launch tags:[Open Source](/topics/open-source)•[Develo…
- @Search (0): [Xata](/products/xata?ref=product%5Fsidebar) Serverless database platform powered by PostgreSQL [4.8(7 reviews)](/products/xata/reviews) [Cloud Computing Platforms](/categories/cloud-computing-platforms)[Databases and backend frameworks](/categories/databases-and-backend) [View…
- @InterSub (0): Hunter 📌 SereneDB built something that is usually available only to big-wallet vendors: a database that holds a huge amount of data in one node, on an ordinary machine, and does ultra-fast search plus analytics at an enterprise level. Their full-text search is the fastest today,…
- @GitGuardian (0): Important technology to be working on. As AI agents expand traffic at a huge rate, database bottlenecks will be production issues you can't prompt your way out of. Cool stuff Upvote (1) Report Share 5h ago [](/@svetlana%5Fragimova) [Svetlana Ragimova](/@svetlana%5Fragimova)
- @InterSub (0): Hunter [@mackenzie\_jackson](https://www.producthunt.com/@mackenzie%5Fjackson) Exactly! Maybe it is not obvious yet, but it's coming! Agents also generate a lot of data on top of that. So the ability of a database to keep everything in one node and run search & analytics right t…