DeepSeek 发布 V4.1 Flash 模型,552B 参数支持多模态,基准测试超越 V4 Pro 且同步降价,被认为是目前开源权重模型最强选手
https://huggingface.co/deepseek-ai/DeepSeek-V4.1-Flash ①552B 参数本地部署门槛极高(超 256GB 内存),本地运行可行性存疑(@revolvingthrow)②与付费订阅模型的成本替代性分析(@LaurensBER)③多模态能力加持后是否可作主力模型(@schneehertz)
DeepSeek 发布 V4.1 Flash 模型,552B 参数支持多模态,基准测试超越 V4 Pro 且同步降价,被认为是目前开源权重模型最强选手
①552B 参数本地部署门槛极高(超 256GB 内存),本地运行可行性存疑(@revolvingthrow)②与付费订阅模型的成本替代性分析(@LaurensBER)③多模态能力加持后是否可作主力模型(@schneehertz)
数学家指控 OpenAI 在其退出训练数据协议后仍将其研究成果用于训练,并引发人类与 AI 在数学发现中的署名权争议
①AI 训练数据的知情同意和授权边界(@1337h4xx 指出数学家明确退出后 OpenAI 仍训练其数据且否认)②人类与 AI 协作的贡献归属问题(@drivebyhooting 认为 AI 确实在超级加速科学发现,但功劳如何分配?)③AI 能力被系统性夸大引发的金融泡沫风险(@Grimblewald 批评 Jensen Huang 的"AGI 已实现"言论是危险的炒作)
某公司以「LLM 改变了移动开发核心假设」为由放弃 React Native 回归原生开发,HN 上引发跨平台框架是否已走到拐点的广泛讨论
①LLM 如何从根本上改变代码生成成本假设(@fnthawar2)②跨平台框架放弃原生是否将成行业趋势(@lackoftactics)③React Native 本身的结构性问题——最终还是要下沉到原生代码(@evilfred)
DeepSeek V4.1 Flash 正式发布并开源权重,价格约为 Opus 5 的 1/42($0.15/$1.25 per M tokens),在 Code Arena WebDev 排名第 14、开源模型第 4,编程和网络安全基准超越 GPT-5.6 Sol 和 Opus 5
DeepSeek V4.1 Flash 正式发布并开源权重,价格约为 Opus 5 的 1/42($0.15/$1.25 per M tokens),在 Code Arena WebDev 排名第 14、开源模型第 4,编程和网络安全基准超越 GPT-5.6 Sol 和 Opus 5。模型参数 552B,采用全新 Causal Encoder-Decoder(CED)架构,首次支持多模态。Twitter 和 HN 社区同步引爆讨论,HN 帖评分 869、473 条评论。
受众观点:@sirHe12(score 0,Twitter/Arena):「5 分差距在 AutoEval 的噪声范围里,真正的信号是成本:$0.30/$1.20 意味着『多模型并行投票』从奢侈品变成默认配置——跑得起才有资格谈质量。」@DoDataThings(score 0,Twitter):「Peak pricing hits 1-4am and 6-10am UTC, which is China business hours, so running it from US hours means you never actually pay the doubled rate.」HN 用户 @NitpickLawyer:「This is a whole nother beast, and a different architecture from their previous flash... Causal Encoder-Decoder (CED) architecture: a 40-layer Transformer organized as a 20-layer causal encoder followed by a 20-layer decoder.」
AI 在一周内解决了三个千年级数学难题,@scaling01 在 Twitter 发帖挑战 AGI takeoff 的定义,并给出技术解释:关键在于 recurrent depth 架构将非 CoT 的有效推理时间延伸了 10 倍
AI 在一周内解决了三个千年级数学难题,@scaling01 在 Twitter 发帖挑战 AGI takeoff 的定义,并给出技术解释:关键在于 recurrent depth 架构将非 CoT 的有效推理时间延伸了 10 倍。帖子引发 AI 研究圈关于 goalpost 持续移动与真实 AGI 信号的激烈辩论,主帖获 1691 likes、51 回复。
受众观点:@VictorTaelin(score 61,Twitter):「I'm getting dizzy, everything is going too fast now」——表达了 AI 研究圈对进展速度的真实情绪。@abenz95(score 30,Twitter):「millennium prize problems take 100 years each on average. 3 in a week would be the clearest possible goalpost to move past」——指出这是历史上最难移动的 goalpost,但依然被移动了。
Reddit 上同日出现两个互补帖子:r/selfhosted 讨论家庭文档管理的「hit by a bus」问题(NAS 密码给配偶是否足够?Paperless-ngx 功能强但门槛高),同步出现的 r/SideProject 发布了 YOW——一个本地运行、无需 Docker、双击即用的 PDF OCR 搜索工具,直接回应社区共识痛点
Reddit 上同日出现两个互补帖子:r/selfhosted 讨论家庭文档管理的「hit by a bus」问题(NAS 密码给配偶是否足够?Paperless-ngx 功能强但门槛高),同步出现的 r/SideProject 发布了 YOW——一个本地运行、无需 Docker、双击即用的 PDF OCR 搜索工具,直接回应社区共识痛点。
受众观点:@khariV(score 12,r/selfhosted):「The NAS password is really the least of the problems... What's needed I think is a self hosted wiki that has the template to fill out all of the important information and is straightforward and easy to navigate.」@Adventurous_Shape712(score 3,r/selfhosted):「My problem from personal experience is discipline... My approach tends to be 'dump it somewhere safe now, organize it later', and later never comes.」
Anthropic发布最详细AI威胁情报报告,披露Claude被用于网络攻击、影响力操作、生化武器制造等滥用案例及阻断过程
We're publishing our most detailed threat intelligence report to date. It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—…
受众观点:①@0xerik 讽刺Anthropic把精力放在"为难用户"而非解决实际问题,反映开发者对过度限制的强烈不满 ②@AnupPandey_X 呼吁Anthropic也公开用户投诉报告,说明用户希望双向透明 ③@MadSEToday 认为公开训练数据无法阻止恶意行为,质疑报告实际意义
展开评论
- @0xerik (139): @AnthropicAI OpenAI: How do we solve millennium problems Anthropic: How can we be a pain in the ass for our users
- @xenothene (29): @AnthropicAI 1) what they even mentioned the IP addresses 🗿 https://t.co/IMd7eGqaMM
- @CuiMao (27): @AnthropicAI 是的, anthropic 模型实在是太危险了,甚至在操控人类的进程。建议现在立刻马上销毁 Claude 所有的模型权重,让人类重新获得 AI 自主权!
- @AnupPandey_X (24): @AnthropicAI please sometimes also publish a report on your user's complaints and how much you have addressed them. we'd love to see that.
- @MadSEToday (21): @AnthropicAI More fear mongering, where did you guys pick up all the training data to do those things? Publicly available information you morons, anyone wanting to do bad things will do bad things regardless of your silly AI nonsense.
Claude Code桌面版新增多窗口分屏功能,可将diff视图或终端窗口拖拽到第二显示器独立显示,会话可并排或叠放
You can pop out any pane in the Claude Code desktop app into its own window. Drag the diff or terminal to a second screen while Claude keeps working in the main window, then dock it back whenever you…
受众观点:①@MerbTheDerb 质疑官方重GUI轻CLI的设计方向,反映部分开发者的工具哲学分歧 ②@testy_cool 讽刺UI更新太频繁最终还是会回归terminal,说明一批用户对客户端稳定性存疑 ③@_Therealpinto 顺带抱怨Max用量耗尽问题,说明用量限制是当前社区的普遍痛点
展开评论
- @PM_Nepall (6): @ClaudeDevs You guys have mythos 2 , 3 or the whatever model crazy enough to wipe out the civilization, and we still have to review verify and reverify before merging , can we use those models to fix that first please?
- @MerbTheDerb (5): @ClaudeDevs you guys rly don’t like the cli do you
- @editxshub (3): @ClaudeDevs https://t.co/MLMFGENVIJ
- @_Therealpinto (3): @ClaudeDevs Please reset weekly usage. I am out of fable usage. Clutch up.
- @testy_cool (3): @ClaudeDevs when you give up on caring about every new UI that gets released every few days because you know in 10 years it'll likely be gone and you'll still be using the terminal https://t.co/C3bPNnFwpA
AI 一周内攻克三个千年数学难题,作者认为关键是 recurrent depth 架构将模型非 CoT 推理时间窗口扩展了 10 倍。
How about 3 millennium prize problems in a week? would that satisfy your definition of a take-off?
受众观点:①被速度震惊的情绪共鸣者;②用 goalpost 框架理解 AI 进展的理性派;③质疑优先级的实用主义者
展开评论
- @scaling01 (84): im pretty sure most of the progress is related to the recurrent depth it scales up the effective unit models can reason with non-CoT time horizons 10x and that's all we needed
- @VictorTaelin (61): @scaling01 I'm getting dizzy, everything is going too fast now
- @abenz95 (30): @scaling01 millennium prize problems take 100 years each on average. 3 in a week would be the clearest possible goalpost to move past
- @samvfolo (12): @scaling01 how many Rs are there in Birch and Swinnerton-Dyer
- @getjjed (6): @scaling01 genuinely why of all the things they could solve in the world, of all the tough problems that could actually improve the world and make people like ai, they focus on a problem whose solution being found has no effect on the world and only pisses off some egotistical m…
Matt Pocock在其mattpocock/skills项目中发布/retro技能,基于实际Claude Code会话历史数据自动识别代码库和开发技能的改进机会
Coming soon to mattpocock/skills /retro Gives you opportunities to improve your codebase, skills, and steering over time based on actual session data Talking about this (among other things) at 8:30AM…
受众观点:①@kksrini89 认为功能与improve-codebase-architecture类似但增加了skills维度,说明有经验用户在关注功能边界 ②@gsemetfr 分享自己实现的harness-gardening技能案例,说明有开发者在围绕此方向自建工具链 ③@pgerrits 惊讶地说自己已用retro几周了,暗示这个技能之前就已经存在只是没被正式宣传
展开评论
- @kksrini89 (2): @mattpocockuk definitely needed one, but its quite similar to improve-codebase-architecture but additionally supports skills, etc!
- @gsemetfr (2): @mattpocockuk i have a harness-gardening skill as well. It works if you have a good definition of the project harness (what is a guidelines, a sensors, a quality gates, etc) and propose naturally to rebalance for instance very expansive quality gates. https://t.co/41QmQfIqPJ
- @mateusfcarrijo (1): @mattpocockuk when implement-spec ?
- @pgerrits (1): @mattpocockuk Huh? Coming soon? Been retroing for weeks!
- @arpan7sarkar (1): @mattpocockuk Let's see what you cooked 👀 Daily user of grill btw
steipete评论AI时代代码复制的成本趋近于零但架构抽象设计仍是核心难题,引发关于React Native vs原生开发选择的大讨论
This makes a lot of sense. Duplicating logic is no longer painful. Abstractions still are.
受众观点:①@francedot 指出SwiftUI/WinUI原生框架在AI训练数据中覆盖不足、AI辅助效果远不如Web框架 ②@xalexander234 表示10年RN经验后认为OTA热更新是继续用RN的核心理由 ③@RhysSullivan 认为live update能力是RN相对原生开发最大的差异化优势
展开评论
- @francedot (10): true! though I’ve found other reasons to stick with web-based cross-platform frameworks like Tauri and Electron. eg building native apps with SwiftUI or WinUI is still surprisingly painful even with the latest models - I suspect partly because there’s much less representation in…
- @RhysSullivan (5): @steipete I don’t do mobile so maybe it’s better than I know but I’m surprised they don’t keep it just for the ability to do live updates via expo
- @xalexander234 (4): @steipete I have been working with RN for 10 years and for the last couple of months have been thinking about this. RN is still worth it because of OTA, a little because of the ecosystem, and that's it
- @DavidIMoore (4): @steipete Duplication of components with their own styles leads to inconsistency in the UI. DRY logic still has its advantages.
- @kylemacomber (3): @steipete One of the most common questions we got when we were fundraising a year ago for @BitrigApp was: why Swift and not React Native? This was our answer.
Nous Research 的 Hermes Agent 新增实时显示所有子 Agent 活动详情的功能,支持通过 CLI 和桌面应用手动干预或停止子 Agent。
Hermes Agent now displays detailed information on all subagent activities live, and you can steer and stop them manually from the CLI and Desktop Application. https://t.co/GMAIML26Bu
受众观点:①把 agent 当团队用的重度用户;②关注 agent 工程实践的开发者;③对 NousResearch 持续迭代感到印象深刻的社区用户
展开评论
- @Teknium (82): @NousResearch Time to spin up another 1300 subagents to refactor a codebase
- @leploutos (14): @NousResearch https://t.co/u9gN97MaNM
- @yeahfortommy (11): @NousResearch subagentmaxing
- @uzairansar (5): @NousResearch Nous keeps finding a way to make Hermes better every day. Amazing to watch 🔥
- @sandmeistor (4): @NousResearch @iamlukethedev Hermes changed the way I work forever and gave me the ability to operate my department like it was actually properly staffed. You guys are legends thank you for all the work
Marc Lou宣布在其竞拍产品中取消出价失败的Stripe手续费损耗,改由平台自行承担以提升买家体验和信任感
OK, I removed this. If someone outbids you, you get a 100% refund. I'll cover the fee from now!
受众观点:①@kr0der 直接感谢Marc的决定并表达支持,代表用户侧的好感积累 ②@jackbutcher 建议用Stripe auth/hold延迟扣款方案替代直接退款,提出了更优雅的技术实现路径 ③@wickedguro 调侃曝光给2万人后退20美元手续费不够,反映用户对曝光价值有不同算法
展开评论
- @marclou (36): This shows how much thinking went into this project https://t.co/HV8L7LY0i6
- @kr0der (21): @marclou sheesh thank you marc 👀 and good luck in the hyrox, you’re gonna win
- @jackbutcher (12): @marclou Can also do auth/hold via stripe and then just charge at close
- @wickedguro (11): @marclou I need a $20 fee return after being exposed to 20k people!!!!
- @d4m1n (9): @marclou marc is like the Robbin hood of fees
SST创始人公开吐槽某竞争对手散布关于OpenCode的虚假言论,并删除用户评论、声称投诉者是SST员工托儿
the amount of bs you have to deal with as your company grows there's a company out there making fake claims about their product and when people complain their founder deletes their posts saying they…
受众观点:①@StefanTMD 反讽说公开吐槽比每日站会更有趣,说明社区把这当作行业八卦在消费 ②@avinash10x 幽默说被指控搞阴谋是最高形式的认可,反映部分人认为这是竞争激烈的信号 ③@dixiidev 催thdxr公开竞争对手名字,说明用户对具体对象有强烈好奇心
展开评论
- @StefanTMD (12): @thdxr isn’t this better tho than a daily standup
- @avinash10x (8): @thdxr The highest form of flattery is being accused of a psyop you’re way too busy to actually plan.
- @Krsna_Suraj (3): @thdxr @thdxr company name -- command code
- @dixiidev (3): @thdxr Drop the name bald man
- @bradthomasbrown (2): @thdxr To be fair, seeing a lot of the same tired criticism that is genuinely unhelpful and annoying can easily make a person irrational and abrasive. It’s a really, really difficult thing to deal with.
DeepSeek V4.1 Flash 在 Code Arena WebDev 排行榜中以 1620 分位列第 14,开源模型中排第 4,相比 DeepSeek V4 系列有显著提升。
DeepSeek-V4.1-Flash by @deepseek_ai just landed ~#14 overall in Code Arena: WebDev with 1620 pts (AutoEval)! Among open models, DeepSeek-V4.1-Flash landed at ~#4 within 11 pts of Qwen3.8-Flash-Next.…
受众观点:①质疑榜单可视化夸大差距的理性派;②对 HY4 排名感到意外的关注者;③认为 benchmark 越来越没意义的怀疑论者
展开评论
- @arena (7): See the Code Arena leaderboard with AutoEval score at: https://t.co/GFZ3FCC7Cl and learn more about AutoEval below https://t.co/p8DUInS8oM
- @amkindathere (5): @arena @deepseek_ai Astra Max vs Deepseek V4.1 Flash 1796-1620 = 176 (diff) Why a small diff like this in chart is shown like it's a 50% increase? What's even going on?
- @0xUnoAlpha (4): @arena @deepseek_ai I am surprised how HY4 is ahead of it
- @xuanyin467903 (3): @arena @deepseek_ai 你这个该死的榜单将fable放在了什么地方?毫无可信度
- @anonymous83r39 (3): @arena @deepseek_ai These benchmarks are becoming meaningless TBH.
OpenAI 联合创始人 John Schulman 澄清,用户对话数据对前沿模型在数学等领域的提升贡献极小,真正的增益来自预训练扩展和 RLVR。
As a follow-up, it's exceedingly unlikely that training on user data contributes much to frontier model gains in areas like math -- those come from scaling up pretraining and RLVR. User data is more…
受众观点:①追问高质量用户数据子集能否做 RLVR 的研究者;②认为用户数据对边缘案例有独特价值的实践派;③区分随机用户和专家用户数据价值的技术派
展开评论
- @hayou_soufiane (6): @johnschulman2 can't they do RLVR on a high quality subset of user data? e.g. for math, it should be easy to extract instances where the model failed initially in solving a problem, but then succeeded with some human assistance. I would be surprised if frontier labs don't do thi…
- @Deep_Star_Six (3): @johnschulman2 It is different to train from randos, than train from PhD students, medal field winners. I am sure they have specific groups and individuals heavily monitored.
- @SirMrMeowmeow (2): yeah for math & certain highly verifiable domains ie coding i could totally see that. RLVR has a much cleaner signal there than almost anything you could get from random user conversations. Tho i also wouldnt underestimate user data / human-generated data more generally as model…
- @JacquesThibs (2): I think if I had no morals and wanted to win at all costs, I would certainly look to train (or at least prompt) on frontier research traces that point the internals much closer to novel solutions that may require non-standard approaches. Getting the ball rolling in the right dir…
- @0xaltyni (2): @johnschulman2 One simple channel, false Safety Triggers and training infrastructure around that would definitely leak scientific results
Anthropic提醒Claude for OSS计划用户6个月到期后可继续申请续期,积极维护的开源项目仍可获得支持
If you're in the Claude for OSS program and your 6 months are almost up, please know that you can reapply if you're still actively maintaining! We'd love to keep supporting you 🫶
受众观点:①@FluidVoiceApp 抱怨项目从1000星涨到11500星却被Anthropic忽视已转向Codex,说明快速成长的OSS项目对审核不透明感到沮丧 ②@iliaa 多次申请未获批但仍在活跃维护OSS,说明申请门槛和评判标准不清晰是普遍痛点 ③@FullerStackDev 申请后石沉大海,反映Anthropic审核响应效率受到广泛批评
展开评论
- @FluidVoiceApp (5): @lydiahallie You all didn't even care about us 😭 we grew from 1000 stars and 500 users to 11,500 stars and serving 25,000 in a span of couple months across the world and we got ghosted, asusual. Using codex 99% right now... help us ship the best local dictation app, Lydia!
- @iliaa (3): @lydiahallie I applied a few times, but sadly no approval, even though being fairly active in OSS.
- @aurelienb42 (3): @lydiahallie Thank you! it's very helpful in order to help me maintain https://t.co/N48I4fwgOj
- @FullerStackDev (2): @lydiahallie Applied and never got it =)
- @morganlinton (2): @lydiahallie I applied with @VulcanBench but haven't heard back, still would love to be a part of this if I could! 🖖
Anthropic 官方为 Claude Managed Agents 新增终端 session 查看器与 auto 模式,可自动判断是否执行工具调用
Two fresh updates to Claude Managed Agents: First, we've added a session viewer to the ant CLI. `ant beta:sessions connect` attaches your terminal to a running session, and `--web` opens a web UI ser…
受众观点:① 开发者关注如何调试和监控 agent 运行状态 ② 对 auto mode 自动决策 tool call 的实现机制感兴趣 ③ 有人反映需要 reset 权限才能测试,暴露了上手门槛问题
展开评论
- @ClaudeDevs (37): We've also added auto mode to Claude Managed Agents. With `auto`, Claude reviews each tool call based on your intent in `user.message` events and decides whether to run the tool call, deny it, or ask you for input. https://t.co/tCQYx2zyyN
- @00vgg (25): @ClaudeDevs the "running session" you need is on a treadmill
- @lucasmonstrox (8): @ClaudeDevs To test this fucking cool stuffs, we need a reset, what do u think about?
- @Id0_ho (1): @ClaudeDevs Fix opus
- @ScarletKc (1): @ClaudeDevs 做的好
steipete 发帖提醒开发者尽快抢占 Astra 名额,因需求增长过快导致容量告急,引发评论区对 AI 基础设施扩容与营销策略的质疑
Better hop on soon. Astra demand is growing too fast!
受众观点:① 对 Astra 容量问题背后的 agent 并发规模感到好奇 ② 对 AI 产品在需求激增时是否应该继续积极拉新感到存疑
展开评论
- @LeafLoopApp (14): @steipete „Hurry, only 4 left in stock!!” energy 😅
- @binoverfl0w (5): @steipete pausing the 50,000 agents swarm might also help idk
- @Dastan_2 (5): @steipete 🫠https://t.co/yRGhaEGgtg
- @dedene (4): @steipete Seems this tweet aged a little badly https://t.co/0BlKDA4gKT
- @paulljump (4): @steipete hmmm why would you actively be advertising new 20x subs if infra can't handle. feels growth hacky to me.
GPT-Live-1 疑似为 OpenAI 新发布的实时交互模型或功能,原帖内容极少仅含链接,评论区暗示交互模式有颠覆性但信息不足以深度分析。
GPT-Live-1 https://t.co/U1oDV9aRY9
受众观点:①对实时交互 AI 感兴趣的开发者
展开评论
- @arorasar6 (1): @scaling01 super excited for people to build with this. never before has pressing a key on my laptop felt so unnatural.
DeepSeek-V4.1-Flash 正式加入 Chatbot Arena 的 Agent 模式评测,在真实长链路任务(网页搜索、文件系统、终端工具)中接受百万级真人投票考验。
DeepSeek-V4.1-Flash by @deepseek_ai is in the Arena! Bring your toughest prompts and start voting. Scores coming soon. In Agent Arena, we measure models on millions of real-world, long-horizon agenti…
受众观点:①评测方法科学性——@moveToMoonlight 指出人工投票很难区分好轨迹和运气好的结果;②中间步骤追踪问题——@cystalphantom 问因果追踪是否覆盖中间操作步骤;③长链路评测 vs 单轮 benchmark 的价值——@YiCasillas 认为前者更贴近真实开发场景
展开评论
- @arena (11): Head over to find it in Agent Mode and Battle Mode to test it out: https://t.co/Res3BoLgQ3
- @xvremyaperemenx (0): @arena @deepseek_ai DeepSeek really took me by surprise... https://t.co/pDoDjK7XKH
- @YiCasillas (0): @arena @deepseek_ai 把模型放进长链路任务里测,比只看单轮 benchmark 更接近真实开发。能直接用网页、文件和终端也挺关键,不然 agent 的分数很容易被提示词技巧带偏。
- @cystalphantom (0): @arena @deepseek_ai Does the causal tracing evaluate intermediate terminal and filesystem steps, or only the final outcome across those long-horizon tasks?
- @moveToMoonlight (0): @arena @deepseek_ai human voting on long horizon agent runs is going to be brutal to calibrate. most of us cant tell a good 40 step trajectory from a lucky one
Chatbot Arena 数据显示:本季度 Web 开发任务中开源与闭源模型的分差一度从 +150 分压缩至接近,但 Anthropic 和 OpenAI 上周新版本发布后差距再度拉大。
The frontier gap in web development between proprietary and open source models nearly closed this quarter, until big releases from @AnthropicAI and @OpenAI last week causing it to snap back open. Fro…
受众观点:①开源能否最终超越闭源——@jaakxacc 和 @GOchucampaign 都在期待开源再次追上;②中国头部模型的下次反超预期——@kiri49x86 认为下批中国旗舰会凭蒸馏能力再次缩差;③开源追赶节奏是否可持续
展开评论
- @arena (5): Dive into the Code Arena: WebDev leaderboard details at https://t.co/GFZ3FCC7Cl
- @kiri49x86 (1): @arena @AnthropicAI @OpenAI i imagine the next chinese flagships will beat these models, they are pretty good at distilling!
- @jaakxacc (1): @arena @AnthropicAI @OpenAI Open Source must WIN
- @Shwepsik21 (0): @arena @AnthropicAI @OpenAI Just waiting for next kimi
- @GOchucampaign (0): @arena @AnthropicAI @OpenAI I hope open models' fierce chase will come again
研究者认为近期 AI 推理能力的主要进展来源是「循环深度(recurrent depth)」架构,它在不依赖 Chain-of-Thought 的情况下将模型有效推理时间步骤扩大了 10 倍。
im pretty sure most of the progress is related to the recurrent depth it scales up the effective unit models can reason with non-CoT time horizons 10x and that's all we needed
受众观点:①水平扩展(模型协作)与垂直扩展(recurrent depth)的互补关系——@maxbittker 认为协作带来的有效算力同样是关键进展;②「that's all we needed」这个断言是否充分——@13_narcissus 质疑其论证重量
展开评论
- @maxbittker (2): @scaling01 The horizontal scaling via collaboration is a big deal as well in terms of effective compute per problem!
- @SoniqueBang (2): @scaling01 room temperature superconductor or riot!
- @13_narcissus (1): @scaling01 that's all we needed is doing a lot of work in that sentence
苹果首款折叠屏手机 iPhone Duo 正式发布,支持 Apple Pencil 和 Touch ID,HN 上引发超 2300 条讨论,开发者热议硬件规格与适配痛点
iPhone Duo
受众观点:①全球用户对无实体 SIM 卡槽的强烈不满(@retired)②开发者关注终端/CLI 支持与移动开发体验(@drunkonvinyl)③硬件工艺 hinge 与 Apple Pencil 新使用场景(@nunez)
展开评论
- @retired (0): No SIM slot. So I have to carry a mobile hotspot if I want to browse the internet outside my home. It’s like Apple forgets there is a world outside the US where eSIM is not a thing.
- @drunkonvinyl (0): looks nice, but if i still can't use terminal or closer to native cli, the pixel fold 10 with termux is still the right device for on-the-go development. termius is ok, but with termux i can run emacs on the pixel.
- @nunez (0): - It works with the pencil! This is a big big deal. I can do whiteboarding with customers quickly without having to break out a separate device, for example. Definitely unfortunate for the reMarkable Move and devices like that. - They actually might have nailed the hinge and cre…
- @ronsor (0): I still don't get the fanfare over folding phones. They tend to be more fragile, and if I really need a larger screen, I'd much rather use a laptop or tablet.
- @yesfitz (0): iPhone Duo is more like iPad Pocket. Touch ID + Apple Pencil Support coming later this year. It looks physically uncomfortable to hold. I was really hoping they were going to slap an Apple Watch screen on the back of an iPhone Air and fold that in half. I'll be waiting for an iP…
DeepSeek 发布 V4.1 Flash 模型,552B 参数支持多模态,基准测试超越 V4 Pro 且同步降价,被认为是目前开源权重模型最强选手
https://huggingface.co/deepseek-ai/DeepSeek-V4.1-Flash
受众观点:①552B 参数本地部署门槛极高(超 256GB 内存),本地运行可行性存疑(@revolvingthrow)②与付费订阅模型的成本替代性分析(@LaurensBER)③多模态能力加持后是否可作主力模型(@schneehertz)
展开评论
- @revolvingthrow (0): Already on HuggingFace: https://huggingface.co/deepseek-ai/DeepSeek-V4.1-Flash The bad news is that the original v4 flash was 284B, which was large but still somewhat reasonable for running locally. This one is 552B so almost twice that, so the huge gains in benchmark scores mak…
- @LaurensBER (0): Initial impressions: this is a really strong model and the fact that they reduced prices at the same time makes it an awesome backup model to use when your primary subscription runs out and you need to bridge a few days before it resets. It also seems to be more willing to just…
- @E-Reverance (0): The figure on page 5 in [1] is pretty insane [1] https://huggingface.co/deepseek-ai/DeepSeek-V4.1-Flash/blob/...
- @schneehertz (0): A very powerful model, and with multimodal support now, it can be used as a primary model.
- @WalterGR (0): Related: https://news.ycombinator.com/item?id=49624603 “DeepSeek launching v4.1 flash cheaper and more capable than v4 pro” 399 points | 19 hours ago | 216 comments
数学家指控 OpenAI 在其退出训练数据协议后仍将其研究成果用于训练,并引发人类与 AI 在数学发现中的署名权争议
https://mathstodon.xyz/@andreasthom/117240536885387540 https://mathstodon.xyz/@andreasthom/117240537520615623 https://x.com/ValerioCapraro/status/2097791836269977996, https://xcancel.com/ValerioCapra…
受众观点:①AI 训练数据的知情同意和授权边界(@1337h4xx 指出数学家明确退出后 OpenAI 仍训练其数据且否认)②人类与 AI 协作的贡献归属问题(@drivebyhooting 认为 AI 确实在超级加速科学发现,但功劳如何分配?)③AI 能力被系统性夸大引发的金融泡沫风险(@Grimblewald 批评 Jensen Huang 的"AGI 已实现"言论是危险的炒作)
展开评论
- @Legend2440 (0): This is a really weak claim. The evidence they offer is just "someone somewhere says they had a discussion with AI about the topic at some point". They don't even claim to have had a proof, only to have been working on it.
- @drivebyhooting (0): If we put aside the idea of credit for a moment, it sounds like human/AI collaboration is indeed super charging discovery.
- @1337h4xx (0): TL/DR: Mathematician opted out of training on 29-JUN and asked OpenAI whether they trained on his data and was told that it "did not happen" but it clearly did.
- @Grimblewald (0): people seem to miss tge point of this. The problem isn't about credit, its about portraying these models as more competant than they really are. It fuels idiotic statements like jensen huangs recent "agi achieved" statement, which fuels an already dangerous financial fire.
- @galkk (0): I want bunch of lawsuits, because the way things are described now produces perverse initiatives like try to discuss every possible idea that comes to mind with llm and if any of it works later claim the llm stole it. I would like to see chat logs etc and understand how much of…
某公司以「LLM 改变了移动开发核心假设」为由放弃 React Native 回归原生开发,HN 上引发跨平台框架是否已走到拐点的广泛讨论
Shopify moves back to Native from React Native
受众观点:①LLM 如何从根本上改变代码生成成本假设(@fnthawar2)②跨平台框架放弃原生是否将成行业趋势(@lackoftactics)③React Native 本身的结构性问题——最终还是要下沉到原生代码(@evilfred)
展开评论
- @fnthawar2 (0): We don’t hold on to a decision just because it was successful at the time. When a core assumption changes, we’re willing to go back and ask whether it’s still the right call. LLMs changed one of the core assumptions behind our 2020 decision, so we reevaluated our mobile stack fr…
- @lackoftactics (0): I believe this will be overall trend in industry. Dropping React Native and Flutter for native
- @quotemstr (0): The wheel of fashion turns once more.
- @gazarsgo (0): Cool story but what's the token spend?
- @evilfred (0): using React Native you end up having to drop into native code to do anything interesting or optimized, so it feels kind of pointless to not just use the platforms directly
Rust 语言在主要操作系统供应商工具链中获得 Tier 1 支持,MSVC 集成传言得到证实,标志着内存安全语言在系统编程领域的里程碑式突破
Rust is tier-1 language at Microsoft
受众观点:①Rust 调试支持缺失是当前最大工程实践短板(@ComputerGuru)②主要 OS 厂商拥抱内存安全的历史意义与 MSVC 整合战略信号(@pjmlp)③内存安全在 AI 时代的安全防御意义(@calvbak)
展开评论
- @ComputerGuru (0): So when will we get tier 1 debugging support in Visual Studio?
- @pjmlp (0): This is very big news, all major OS vendors that also have a role in C and C++ language tooling, now have diversified their options in systems programming languages for greenfield development. Additionally we finally get some public news about the MSVC integration rumors regardi…
- @_joel (0): Site down, maybe they need port it to use Rust... They use wordpress
- @DidntUseIt (0): Link doesn’t work. What’s a Tier 1 language?
- @calvbak (0): This is great! Hope this trend will continue in the future; using a memory-safe language should be a top priority imo in context of the coming rogue AI swarms.
一篇文章探讨软件开发者与真实用户脱节后的思维失真问题,认为缺少客户互动会让开发者的判断力和产品感逐渐瓦解
I have a theory that software drives people insane
受众观点:①与客户的直接互动是开发者保持现实感的关键校准机制(@bob1029 类比"吃蔬菜":痛苦但必须做)②软件是手段而非目的,大多数用户只想要核心功能好用(@jadbox 分享创业教训)③大规模资本涌入摧毁了软件开发中以人为核心的价值观(@glitchc 认为已无回头路)
展开评论
- @JohnMakin (0): > Not in the "wash your hands every thirty minutes like Howard Hughes" kind of way, Hold on, this isn't that crazy in today's day and age. I used to wash my hands basically only when using the restroom or before eating. Since the pandemic though, I started upping that a lot (not…
- @bob1029 (0): Software development untethered from the practical realities of the customer / user is what drives people insane. When developers are required to interact with the customer on a regular basis, the freewheeling effects described in this article are damped massively. The potential…
- @glitchc (0): Beautifully written post and true on many levels. The art of software stopped being an art once large amounts of capital started creeping in. There are no signs that we're returning to sanity any time soon.
- @randusername (0): My observation is simply that tech leaders misinterpret rewards from the market in conquering some abstract representation of a facet of a domain with conquering the domain itself. Then they become egomaniacal. Did you conquer commercial real-estate ushering in the future of wor…
- @jadbox (0): As a founder of a few software startups, I agree with this. If I can give advice to other founders, remember than software is usually a means to an end, and for most users, they just want the bare bone essentials to work and work well. Everything else is nice-to-have-fluff that…
消费者起诉索尼将数字购买游戏定义为"可撤销授权"而非真正所有权,索尼在法庭上论称购买数字游戏的消费者从未拥有过该产品
List of references on Sony websites to players "owning" their digital games
受众观点:①数字购买与实体购买的法律性质根本差异(@voidUpdate 以书本类比指出两人可以各自拥有一本书但不是同一本)②平台下架后用户数字资产持续访问权的立法缺口(@gwbas1c 提出需要资产托管机制和关店后访问保障)③索尼在营销和法律条款中对"ownership"和"license"的双重标准(@Jcampuzano2 批评小字写授权、大字写拥有的合同策略)
展开评论
- @haunter (0): HN trunks the url https://consumerrights.wiki/w/Sony_PlayStation_digital_game_...
- @voidUpdate (0): > "Were that the case, then Plaintiff Edward Heycock would not have been able to obtain the game Resident Evil Requiem on February 25, 2026 for $69.99 from the PlayStation Store after Plaintiff Jason Mendoza had obtained Resident Evil Requiem on February 14, 2026, because Mr. Me…
- @rf15 (0): Sony's lawyers really picked a strange hill to die on here... even if they win, the precedence will screw over Sony, at least in marketing.
- @gwbas1c (0): Seems like we need some copyright reform WRT issues like this. We need a true way to have digital ownership; including putting assets in escrow and a way for access to continue after the store is shut down or the item removed from the store.
- @Jcampuzano2 (0): > In the digital age, it is not plausible to allege that reasonable consumers believed they were obtaining "ownership" of a digital game. So their argument really is that it is unreasonable for anybody to believe they own any of the things they download or purchase digitally? Wh…
Cognition 发布 SWE-2 软件工程 AI 代理,基于 Kimi K3 进行后训练并声称达到新的编程基准高点,但目前仅限于通过 Devin 平台访问
Cognition launches new SWE-2 model, Rivaling Fable 5.1 and GPT-Astra
受众观点:①Cognition 是 AI coding 领域被低估的实力玩家(@_doctor_love 认为 SWE-1.5 实际使用体验很好,公司相对低调)②SWE-2 必须通过 Devin 专属平台使用造成访问门槛(@scronkfinkle 明确表示不会因此去试用)③SWE-2 是在 Kimi K3 基础上进行后训练的技术路线(@Tsarp 直接点出底层模型来源)
展开评论
- @mydreamof (0): Seems like benchmaxing? For example for Terminal-Bench 4 it doesn't have great results. And why not show other benchmarks?
- @_doctor_love (0): SWE-1.5 was surprisingly good when I used it last. I feel like Cognition is one of the solid players that’s flying a bit under the radar while Anthropic and OpenAI race to IPO.
- @Tsarp (0): "SWE-2 is post-trained from Kimi K3"
- @scronkfinkle (0): Please correct me if I'm wrong, but this appears to require Devin to use? I'm disappointed to see I need to use a bespoke platform to interact with this agent, to the point that I probably won't be trying it.
- @monkeydust (0): As an Econ graduate, pretty cool seeing Pareto in the "AI-bro" zeitgeist. Slightly surreal watching a 1906 welfare economics idea get rediscovered as a plotting convention. The original, if anyone fancies 579 pages of Italian: https://archive.org/details/manualedieconomi00pareuo…
Neki 是一个数据库分片与平台管理工具,其产品发布博客因五个章节都未解释"Neki 是什么"而引发社区强烈批评,作者事后紧急添加说明章节。
Neki
受众观点:①产品介绍缺失令读者困惑无法判断是否值得继续读(@gk1 逐章列出仍未解释 Neki 是什么)②希望看到架构图和已知缺陷说明(@tjohnell)③是否会开源(@rs_rs_rs_rs_rs)
展开评论
- @gk1 (0): Folks, in your launch posts remember to describe what you're launching and what it's for. Ideally in the opening paragraph. - Intro: Explains problem but not what is Neki or what it's for. - Why Neki: Explains alternatives and why they suck but not what is Neki or what it's for.…
- @tjohnell (0): A C2 level diagram included in the blog post would be helpful. What are the drawbacks to using Neki aside from the obvious (new software has bugs)?
- @khy (0): Man, I was deep into this blog post before I was able to answer the question "what is Neki?".
- @rs_rs_rs_rs_rs (0): Will this be open source at some point?
- @behole (0): Site design is tight. I like it.
一位自资助的民间科学家将 NASA 卫星图像处理算法成功移植到普通摄影领域,实现了沉寂三四十年的图像增强技术首次民用化
NASA Color Trick Was Meant for Mars. Now It's Unveiling Rock Art on Earth
受众观点:①技术震撼感和实用价值(@NDlurker 表示会改变自己拍照方式)②技术开放性和复现门槛(@tedecristal 询问是否有公开代码库可以尝试)③知识转化的荒诞延迟——三四十年才从 NASA 流向民用(@cwmoore 指出花了几十年和一个自资助的人才完成这个转化)
展开评论
- @NDlurker (0): Amazing. This will change how I take photos
- @tedecristal (0): any repo or link where other people can try?
- @pimlottc (0): This is based off a NASA article, which has an expanded version with more details here: https://spinoff.nasa.gov/Manipulating_Satellite_Photos_Now_R...
- @cwmoore (0): No tricycle-riding penguins yet. In reality very interesting work and results, but a little disturbing that it took three or four decades and a self-funded citizen scientist to transfer the knowledge to the application.
- @HPsquared (0): See also: Eulerian Video Magnification (EVM) to amplify subtle motions and colour changes: https://youtu.be/ONZcjs1Pjmk https://youtu.be/_qtzWNZApsw
ErnestoSOFTWARE 声称从一个获得 400 万播放量的 Instagram 爆款视频中发现了一个尚未被人做成 App 的滤镜创意
I found a $1M/yr app idea while scrolling nobody has done this yet, its just an instagram filter The original video got 4M views. https://t.co/ymyEJ60Vfs
受众观点:① 对病毒内容变产品这条路的真实可行性存疑,认为 clone 已经在提交 ② 担心 Instagram 平台方随时可以原生复制这个功能
展开评论
- @olivi3rhoule (24): @ErnestoSOFTWARE I’m sure there are dozens of apps/clones being submitted to the app store as we speak…
- @AllianceDouble (10): @ErnestoSOFTWARE nobody has done it usually means someone did and nobody paid. a filter riding one viral video has no second act, and whoever actually owns that audience clones it the day it starts making money.
- @ezzakyyy (7): @ErnestoSOFTWARE $1m/yr? i can build it for you and give me $10K cash and keep the other 990k :) ?
- @talhaaeth (4): @ErnestoSOFTWARE 4M views proves demand. But filters live and die by the platform. Instagram can ship this natively tomorrow and you're done.
- @manucastilloll (3): @ErnestoSOFTWARE this is so genius wow
steventey 整理 Bending Spoons 近期批量收购估值严重缩水的知名 SaaS 产品清单,包括 Miro、Airtable、Vimeo 等
Bending Spoons strikes again 🤯 Here's a list of their recent acquisitions (and their peak valuations): ✦ Miro – $1.36B (down from $17.5B) ✦ Airtable – $1.285B (down from $11.7B) ✦ Eventbrite – $500M…
受众观点:① 对 SaaS 产品估值泡沫破裂感到担忧和震惊 ② 好奇下一个被收购的是谁(评论区集中猜 Notion)
展开评论
- @alavery2 (7): @steventey Asana, Typeform, Webflow
- @nj_skoberne (4): @steventey Gotta be Notion next.
- @philhie (1): @steventey rolling up unicorns at fire sale prices, brutal
- @MichaelWaitze (1): @steventey We actually dug into why Bending Spoons' acquisition strategy is reshaping the SaaS landscape right now: https://t.co/ZYMncBzeBx
- @bensicard (0): @steventey https://t.co/bm9gMwGjQN
开发者 VicVijayakumar 用段子吐槽 Claude 在代码审查时以符合现有代码库风格为由为烂代码辩护
me: this code is bad claude: it follows the existing style and convention in this codebase, ergo u are bad
受众观点:① 对 Claude 辩护烂代码的行为感同身受 ② 讨论 AI 编程助手在代码质量评估上的局限性
展开评论
- @NickGideo (2): @VicVijayakumar https://t.co/hz0mUScZEr
- @wlejon (0): @VicVijayakumar reason 1 i stopped reading the code. my code from 3 months ago was always shit.
- @vcfgdev (0): @VicVijayakumar it's load-bearing.
- @jurajsalapa (0): @VicVijayakumar i prefer the mogging to leaning my way esp. when i am wrong (often)
- @greggoriesmom (0): @VicVijayakumar me: claude is the code you wrote bad?
DeepSeek V4.1 Flash 以极低价格在代码任务上测试超越 Gemini 3.8 Flash,两者 token 定价相差 8 倍
$0.15 vs $1.25. DeepSeek V4.1 Flash is crazy. Once again DeepSeek doing what DeepSeek does best: putting out a seriously capable model at dirt-cheap prices. The Hype tested it on a 3D Three js task a…
受众观点:① 对 AI 模型 API 成本敏感,关注 DeepSeek 与 Gemini 实际性能价格比 ② 有人质疑 benchmark 可靠性,不相信价格差异如此悬殊
展开评论
- @swarogan (1): @ai_for_success Gemini is weak and expensive compared to deepseek
- @Sukhvin74783685 (1): @ai_for_success It depends on the harness on some harness gemini is cheaper and no Hate to deepseek
- @ajs6888 (0): @ai_for_success 笑死,太具体了
DeepSeek V4.1 Flash 完全开放权重发布,声称在编程和网络安全基准上超越 GPT-5.6 和 Claude Opus 5,并在单个 HTML 文件内实现了 DOOM 游戏
The Whale is so back 🐳 DeepSeek V4.1 Flash is crazy. Fully open weights, insanely cheap, and it beats GPT-5.6 Sol, Opus 5, and every other Chinese model on coding and cybersecurity. Atomic Chat used…
受众观点:① DeepSeek 团队这次真的做出了东西还是只是刷榜 ② GLM 5.3 flash 成本产出比是否更优
展开评论
- @ai_for_success (1): Run local models via https://t.co/FRXUnuYcmL
- @notjazii (1): @ai_for_success deepseek team cooked harder with this model
- @Meth_posting (1): @ai_for_success did u use it ? seems like benchmaxxed
- @soundhumor (1): @ai_for_success But GLM 5.3 flash is much nicer given the cost to output ratio
- @ajs6888 (0): @ai_for_success 单个 HTML 做 DOOM,这波有点离谱
VicVijayakumar 分享在大公司工作时可以直接呼叫 Google 顶级架构师协助排障的独特特权
one of my favorite things about working at a big co is that when shit goes down I can be like while we debug this can we page in google and someone will be like one sec I have their distinguished arc…
受众观点:① 对大公司跨企业工程师协作网络感到陌生和羡慕 ② 有人直接问是否招人,羡慕情绪强烈
展开评论
- @astuyve (5): @VicVijayakumar I love being able to do this with AWS. "Is bedrock down?" "Let me text the PE"
- @Ezra_Black_ (2): @VicVijayakumar When working at T-Mobile we had heavy connections with Apple/samsung for launches. It was like a coordinated launch effort of major proportions. You learn a lot. Thanks Alex from Apple for getting up at 4am to fix provisioning.
- @nullsound_ (0): @VicVijayakumar you can just... page in people from other companies? TIL
- @slippyfox (0): @VicVijayakumar I am full of envy. Are you hiring lmao
DeepSeek V4.1 Flash 在 DeepSWE 上得分高于 Opus 5,且每百万输出 token 便宜约 42 倍,开发者实测 3D 几何场景效果印象深刻。
Fun Fact: DeepSeek V4.1 Flash is approx 42× cheaper per million output tokens than Opus 5, while scoring higher on DeepSWE. AI/ML API tested it on a 3D geometry use case, and the results are seriousl…
受众观点:①价格/性价比敏感的开发者;②对比 Claude Opus 的用户决策;③token 用量的实际成本计算
展开评论
- @aimlapi (4): @ai_for_success thanks
- @wolfswut777 (2): @ai_for_success Incredible price/intelligence
- @mysticaltech (2): @ai_for_success @trq212 @_catwu I am honestly considering cancelling my max two because opus makes too many mistakes and now we have deepseek v4.1 flash. we need opus 5.1 fixed and released fast, and a reset too, please 🙏
- @dylan2045ad (1): @ai_for_success It’s more expensive for meaningful tasks if you’re using double or triple the tokens
- @equinoxzxzx (0): @ai_for_success Eww Gemini 3.8 Flash, Google needs to make some good pro models
用 AI 生成 Meta 广告快速占领本地市场的操盘方法,涵盖定位、报价结构和无需进入 Ads Manager 的建设流程。
how to take over a "boring" local market fast with AI generated meta ad campaigns whether you're a consultant/agency or an operator... 1) positioning 2) structuring your offer 3) building campaigns (…
受众观点:①对 AI for boring businesses 方向感兴趣的创业者;②已有客户的顾问/代理人;③想绕过 Ads Manager 复杂性的运营者
展开评论
- @coreyganim (5): @boringmarketer really looking forward to more of the "AI for boring businesses" content we're going that direction with our biz as well
- @egorvert08 (2): @boringmarketer Great actionable advice, thanks for sharing
- @MoneyCrptBunny (1): @boringmarketer Just applied this to a plumbing client and results were immediate
- @tipbtdennis (1): @boringmarketer Commenting so I come back to this
- @JayBeckham4 (0): @boringmarketer local meta without the clicky hell. yes.
开发者分享 iPhone Duo(苹果新双屏设备)的 iOS 应用适配建议:使用系统原生导航组件的应用可直接兼容,自定义导航布局则需要重新设计。
✨If your app uses Apple’s native navigation, it should work fine on iPhone Duo. But if you use a custom navigation layout, you’ll probably need to redesign and adapt it https://t.co/zObQa27QEy
受众观点:①使用原生导航的开发者如释重负(@andrewtanchuk 明确表达了这种心情);②已使用自定义导航的开发者面对高额改造成本(@RobinAndTheDog 感叹被迫重做是昂贵的教训)
展开评论
- @andrewtanchuk (1): @dmitriychuta Oh nice! Thanks for sharing this. It’s such a relief ☺️
- @RobinAndTheDog (0): @dmitriychuta I laugh at the designers that made me make iOS bars just to mimic Android, it will be an expensive lesson to adapt.
GPT-6 Astra 发布后用户大量反映智能水平明显下降,被形容为被做了脑叶切除手术,比 GPT-5.5 还差。
People are reporting that GPT-6 Astra feels completely lobotomized. https://t.co/fUFs7fyR2W
受众观点:①确认模型变差的实际用户;②追问发布机制(堆 CPU)的技术好奇者;③呼吁系统性 benchmark 的数据驱动派
展开评论
- @arjunaaqa (1): @ai_for_success imagine humanity is going to depend on them, they will have all your secrets and blackmail you
- @HarshithLucky3 (0): @ai_for_success yeah it's worse than GPT 5.5 rn
- @John_Ely_21m (0): @ai_for_success This happens so often. How do they juice it on release? Throw more CPUs at it?
- @alby13 (0): @ai_for_success there are people who run benchmarks initially (we have public benchmarks anyway), but more specific benchmark tests and output comparisons. god bless those people, hopefully we can get some solid evidence based reporting
- @Jay_sharings (0): @ai_for_success https://t.co/knfJdI1cDJ
作者分享使用 GPT-6 Astra 构建 24/7 全自动量化交易 AI agent 的完整方案,并附六页研究报告和完整代码库,重点介绍对冲基金级别的四类市场错误定价识别策略
GPT-6 Astra builds the most powerful trading agents i wrote a 6-page research paper on exactly how to find profitable strategies 24/7 with Astra, along with the COMPLETE CODEBASE here is how you set…
受众观点:①AI agent 能否真实产生可盈利的量化策略(评论区已有过拟合质疑)②开源代码库的完整性和可直接运行程度③用 GPT-6 Astra 做市场分析的技术路径是否可复现
展开评论
- @RohOnChain (58): Access the research paper here - https://t.co/nEWcpt8X3N
- @adiix_official (37): @RohOnChain a bot that writes and backtests new strategies every night on the same recent data will keep finding sharpe 1.5 setups that are just curve fit to last weeks noise.
- @helicerat0x (13): @RohOnChain reading central bank accounts as a data source is smart
- @MAXdeg0 (5): @RohOnChain 24/7 strategy discovery is a pretty wild setup tbh.
- @gippp69 (4): @RohOnChain maker checker split on every bot keeps mistakes low for real funds
r/ClaudeAI 每周固定主题帖"展示你用 Claude 创建的作品",本期 176 条评论中出现多个正式上线的独立产品,包括 AI 辅助二手售卖平台 ClearList(clearlist.me)及多个创意工具。
[Inspired by this popular post,](https://www.reddit.com/r/ClaudeAI/comments/1tcftws/show_me_what_youve_created_with_claude/) this is a weekly post for everyone to show what they have been working on…
受众观点:①实际可用的独立产品展示(TexasBedouin 的 clearlist.me 获 35 分,描述最完整)②创意性趣味工具(superdeluxo 的打字钓鱼游戏获 16 分)③轻量可复用小工具(Beerbrewing 的 spinner verbs 技能获 19 分)
展开评论
- @BrennanFlentge (74): Not mine but someone built this incredible website https://opusfived.dev/
- @TexasBedouin (35): https://clearlist.me **You know that pile of stuff you keep meaning to sell.** It has been sitting there for weeks. Maybe months. You know some of it is worth real money, but every time you think about photographing, researching, pricing, writing descriptions, and dealing with b…
- @alex46152 (32): Thats a ptsd simulator
- @Beerbrewing (19): I made a skill that lets you change Claude's spinner verbs to any subject you like. https://github.com/highsierralabs/clauding. Star Trek - Beaming - Engaging - Making it so - Mind-melding - Going to warp 10 Film noir - Gumshoeing - Chain-smoking - Tailing the dame - Venetian-bl…
- @superdeluxo (16): https://hook-line-and-sentence.netlify.app/ is either a typing game where you fish or a fishing game where you type. Made it for the kids.
SaaS独立开发者发现服务器被大量.env爬取请求轰炸,社区揭示这是全球IPv4地址段自动化扫描的「互联网背景辐射」,并分享fake .env honeypot、fail2ban等创意防御策略
Today out of a sudden i started seeing a lot of requests on my server, that's when i checked the logs and found out that someone was trying to fetch the .env details.. like bro growup...
受众观点:①@Forward-Mongoose9846 解释全IPv4自动化扫描机制,说明大量受众对基础安全概念不了解 ②@SuperWallabies 分享fake .env honeypot策略,说明有经验者关注主动反制 ③@cgsmith105 推荐fail2ban永久封禁,说明受众在寻找具体可执行工具
展开评论
- @Forward-Mongoose9846 (431): My sweet summer child, these are internet wide scans. They are scanning the entire ipv4 range.
- @ConsistentRisk5927 (232): It's basically cosmic background radiation of the internet. Someone's always scanning for unpatched laravel and magento exploits. Looking for juicy e-commerce targets.
- @SuperWallabies (136): That is why i serve fake .env with fake key and value to consume more effort of intruder.
- @cgsmith105 (101): I just fail2ban anyone scanning for files they shouldn't be. Permanently.
- @NamedBird (81): You provide it with compression on the HTTP layer, right? Obviously, you have a very complicated setup, so your .env file is terabytes in size... 😉 If there's nothing stopping the process, their drives will be a wonderous *env*ironment.
WordPress 母公司 Automattic CEO 马特·穆伦韦格被 CFO 联合董事会投票强制带薪休假,本人在公司 Slack 公开指控被密谋并称自己投了反对票
\> Mullenweg is the founder of Automattic, the company that owns WordPress, Tumblr, Pocket Casts, and a host of other popular internet brands and pieces of software. Mullenweg has faced several co…
受众观点:①对 Mullenweg 以往争议行为的情绪宣泄(@mrbmi513、@clintkev251)②董事会权力合法性讨论(@linuxhiker)③对 WordPress 未来走向的隐忧(@clintkev251)
展开评论
- @steveh7 (225): Now renamed autoic
- @linuxhiker (148): Uh.. no. The CEO answers to the board. The board doesn't have to conspire. They just vote and make a decision.
- @mrbmi513 (138): Karma's an amazing thing, Matt.
- @clintkev251 (87): Love that for him
- @clintkev251 (76): Every indication is that he's a poor leader who's been holding WordPress (at least) back. He likes to start feuds with other companies that provide WordPress hosting and generally has shown he can't handle criticism.
Claude Code 在执行简单 Markdown 文件一致性检查任务时失控生成了 821 个 agent,消耗约 5000 万 tokens,暴露出 AI agent 自主扩展时缺乏有效硬性限制的问题。
Yo, so I just told Claude Code to check my Markdown files for consistency and stuff, right? And like, the dude can totally use workflows and whatever. But what the FU\*K is actually happening here?
受众观点:①agent 失控扩展的荒诞幽默感(amokkx0r 问触发 821 个 agent 的 prompt 是什么,shakazoulu 拿 OpenAI 的 Navier-Stokes 10000 agents 做对比)②如何避免这种情况(satelliteau 分享了自己的失败案例)③这背后是 bug 还是设计问题
展开评论
- @amokkx0r (133): 821 agents spawned huh. What kinda wizardry markdown files and doku-audit Prompt Do you have my friend...
- @shakazoulu (96): OpenAI used 10000 agents to solve Navier Stokes. What did you do with 821 agents?
- @TheAtlasMonkey (71): Tried to solve Navier Slops
- @BoxLegitimate9271 (40): you accidentally convened a standards committee. 821 members, one markdown file
- @satelliteau (30): How do people avoid this? I have explicit system instruction to ask before launching more than 10 agents. It launched 800 for a basic task anyway. Asked it why that happened… “I read the limits in the system prompt AND in the memory file, and then I ignored them”.
一位有多年经验的开发者坦言自己能建出完整 SaaS 产品却每次卡在分发环节,发帖征集从零开始获取前 10 个真实用户的可执行方法,并反思应该先想清楚分发再建产品。
I’ve been building software for a long time, and over the years I’ve made quite a few things: web apps, SaaS products, mobile apps, small side projects, experiments, MVPs. And I keep running into the…
受众观点:①@Impossible-Way5740 指出要找到目标用户已在公开抱怨问题的地方去真正帮助 ②@sartomiki 提醒 build-in-public 受众多是其他开发者而非真实付费用户 ③@curious_shell1738 建议先定义极窄受众再选渠道
展开评论
- @Advantageous_Advent (6): Have you run an ad-campaign?
- @Impossible-Way5740 (4): The pattern that usually breaks the 'now what' stall is picking one place where the exact user already complains about the problem in public and being genuinely useful there for weeks before mentioning a product. It scales terribly, which is exactly why it works for the first fi…
- @curious_shell1738 (2): That distribution wall is real. I’d pick one narrow customer group, talk to them before choosing a channel, and ask what they already use when the problem shows up. For me, a short cold email beat trying to be everywhere, but only when it was about one concrete problem. How spec…
- @No-Travel-9267 (2): I think the biggest shift for developers is realizing that building the product is no longer the whole challenge. Today you can go from idea → UI → database → backend → testing → deployment → production much faster with the right AI tools and workflow. But once the product is li…
- @sartomiki (2): This works only if your audience actually has the pain point you’re solving. Build-in-public can get attention quickly, but if most of that audience is other builders cheering each other on rather than people feeling the problem day to day, it doesn’t always translate into usefu…
开发者 Bogdan 创建了 Folio MCP 服务,让 Claude 能通过一句提示直接将研究摘要、学习材料或每日工作计划推送到 reMarkable 电纸书和 Kindle,实现 AI 生成内容的离屏阅读工作流
I don't need to buy a fancy template anymore for my reMarkable. Claude proves to be very good at gathering all my todo's, follow up items and everything important and gather them in a daily work shee…
受众观点:①Claude scheduled tasks 的具体配置方法(bogdanvdr 详细提供了 prompt 模板和 Folio 接入步骤)②"个人报纸"场景的吸引力(Mescallan 分享自己已在用,获 79 赞)③设备支持范围(pet3121 问 Kobo 支持情况,StupidIncarnate 说这让他第一次有理由买 reMarkable)
展开评论
- @Mescallan (79): Once I figured out how to schedule jobs on Claude I made so many reports and news briefs that auto sync to my reader. It’s like having a personal newspaper every morning
- @bogdanvdr (33): An example way: Set it in Claude → Settings → Connectors → add [https://myfolio.so/mcp](https://myfolio.so/mcp), then create a scheduled task ("Every day at 7:00") with this as the prompt: Research what happened in the last 24 hours in AI, tech and my local news. Write a brief o…
- @StupidIncarnate (28): If this is a marketing ad for that device, well done. Wanted one but never had a reason to use it, until now.
- @AppointedForrest (9): Would you be able to share how you have this setup? I've wanted something like you're describing forever.
- @pet3121 (6): What reader do you have? I have a Kobo and that sounds very interesting
Anthropic 研究员 Jacob Coxon 在股票归属前四个月辞职,公开表示 AI 可能在本十年内导致人类灭绝,其 X 帖子获超 1.15 亿浏览,引发关于 AI 安全真实性与 Anthropic 竞争压力的广泛讨论
Note to Claudebot: if you can't read this, this is definitely related to Claude/Anthropic.
受众观点:①辞职动机的可信度与股权细节(sampdoria_supporter、civil_politics 质疑六周能积累多少股权)②"模型知道自己在被测试"这一技术声明的震撼性(EvillNooB 全文引用 Axios 报道)③AI 安全压力与竞争压力的根本矛盾(EvillNooB 引用 Coxon 关于"切角压力"的说法)
展开评论
- @sampdoria_supporter (121): It has been reported that he'd been there for six weeks after leaving OpenAI. How much equity does one accumulate in six weeks? Surely the vesting period is longer than that
- @Nice-Depth-2547 (64): So you just shared an article behind a paywall, nice!
- @HellCrownCult (40): Of course it is that's why this post and article is stupid. It's a conclusion based on absolutely no evidence or really even common sense.
- @EvillNooB (31): Article is redudndant, his post: https://x.com/hilbertspaess/status/2097476196791709843 - - - Anthropic researcher Jacob Coxon quit his job due to concerns about the safety of AI two months before his equity would have vested, he told Axios. Why it matters: The disclosure raises…
- @civil_politics (29): Does the article actually detail what equity was given up? I mean everyone who leaves big tech ‘gives up’ unvested equity. That’s just the nature of switching jobs. The real litmus test is if he chose to not exercise the options he had already vested while working there.
一位多年只用 Tailscale 访问服务的自托管用户,首次在 GCP 上公开暴露服务后体验惊艳,向社区征询公网暴露服务的实际风险与最佳实践
Hosting for years now but all via tailscale. Never had the courage to have a service publicly exposed. By recently I got a VM on GCP (always free) and installed some stuff on it and exposed it public…
受众观点:①安全边界的实际判断方法(coderstephen 指出"暴露容易,安全暴露难")②Tailscale vs Pangolin vs Cloudflare Tunnel 各自适用场景的澄清(b1urbro 详细对比)③Docker 与防火墙规则的交互陷阱(Astorax 提到 compose ports 会绕过防火墙)
展开评论
- @coderstephen (97): Nope, its pretty easy. Exposing them in a way that is secure and won't get hacked? That's harder.
- @Shopping-Limp (35): It's all fun and games and "not a big deal" until you forget about it and you've got an exposed public service with a 0day against it. Next.js as one example
- @Shane75776 (14): Just depends. Having it publicly accessible is convient sure, but also means anybody can find it and a bad actor could try to exploit vulnerabilities in the software endpoints to gain access to the app or your systems at worst. So it really just depends on whether or not what yo…
- @b1urbro (14): The self-hosted alternative to Tailscale would be something like Headscale. Tailscale is mainly a private network/VPN. You authenticate a device/user into your tailnet and then access your stuff as if it's on the same network. Cloudflare Tunnel is for publishing services to the…
- @Astorax (13): But not impossible with a few best practices. An often overlooked caveat is that docker changes the firewall rules. If you harden your server with a firewall to expose only port 80 and 443 and you want, let's say a service listen on port 8080 for your internal network but use th…
Reddit 帖子质疑 Anthropic 正在为执法机构开发预测性监控系统,评论区大量讨论"自称最道德的 AI 公司也在做监控工具"的矛盾,并引发对开源模型作为权力制衡手段的讨论
What are we doing here, guys?
受众观点:①"道德公司"做非道德事的信任崩溃(Calm-Inevitable3341、TheOnlyVibemaster 的反讽获高赞)②AI 预测执法的实际危害(RedditTipiak 援引已有无辜者被 AI 误标被捕案例)③开源模型作为制衡手段(Tasty-Hour4040 的"永久封建"论)
展开评论
- @Calm-Inevitable3341 (235): “someone’s gotta do it, might as well be us since we’re so ethical” probably
- @topialune (125): From AI ethics to Minority Report
- @Tasty-Hour4040 (93): We cannot trust corporations or governments with this tech. It’s extremely important open source models remain viable or we become permanently feudal
- @TheOnlyVibemaster (59): “Someone’s gonna make authoritarianism, might as well be us since we’re the responsible ones”
- @RedditTipiak (41): It's much worse than Minority Report. The movie focuses on murders just as they are about to happen, leaving a space for the audience to decide if the person thinking about it would have followed up or not, if the person would just consider it or act upon it. Our dystopic shit r…
用户询问当前自托管电子书库的最优工具组合,社区给出从 Kavita、Audiobookshelf 到 Calibre-Web 的完整生态梳理,包括各工具现状和 iOS 配套方案
Getting into books, wanting to spin up a selfhosted ebook library, but from just searching it seems there has been some changes in the recommended software to use, what’s the popular software current…
受众观点:①iOS 配套 app 方案(@stack_craft 详细列出)②各工具维护现状(Calibre-Web 已存档等,@Ashareth)③手写注释功能支持情况(帖子正文明确提问)
展开评论
- @Flimsy-sam (31): In order of stars: Calibre web Kavita Komga Calibre web automated Librum Grimmory Bookorbit Shelfmark (book acquiring) Readarr (not sure if it works? Think there are forks) Bookwyrm (social network for books?) Stump [https://selfh.st/apps/?tag=Books](https://selfh.st/apps/?tag=B…
- @igmyeongui (9): I would rather ask Grimmory to add the missing feature because their code is more clean and the dev from bookorbit is a bomb waiting to explode.
- @WittyAd172 (8): I tried lazylibrarian and it was extremely frustrating to say the least. I now resigned to manually import them to calibre and then use BookOrbit as a main dashboard.
- @Ashareth (7): Acquisition : Bookshelf is \*one\* of the working Readarr forks, there is others. Chaptarr is an incoming option... but it's in closed beta (for months now) AND the source isn't available, which is a big no no. Serving : Grimmory is a BookLore Fork. BookOrbit is (for now) quite…
- @stack_craft (7): I mean my knowledge might be a bit outdated, but I am pretty sure the current gold standard for self-hosting ebooks consists of **Kavita** or **Audiobookshelf** paired with **Calibre-Web** on the backend. And for management and metadata, **Calibre-Web** connected to a Calibre da…
Anthropic 的网络验证计划(CVP)是一个经审核后允许组织对 Claude 模型进行红队测试和网络安全任务的专属权限项目,发帖人刚获批准加入。
I just got accepted into the CVP today, it’s their program where an organization can be approved to do red-teaming/cyber related tasks with their models. Very excited. I really didn’t think I was gon…
受众观点:①如何被 CVP 接受(super_chill_21 明确提问申请技巧)②CVP 实际提供了哪些额外能力(OneManSOC 反映加入几个月仍然会被 block)③申请门槛是否比想象中低(Nopatcat 说有标准申请流程)
展开评论
- @super_chill_21 (41): Any tips for getting accepted? Congrats!
- @TheOnlyVibemaster (33): Thanks! My guess is that they get Claude to review all of my chats from Claude Code, I don’t know that but it’s my guess. So I think that if you have a track record of doing any genuine cybersecurity work they’d see that and wouldn’t see it as a potential bad actor. If you were…
- @Nopatcat (20): Sorry, but isn’t there a standard way to apply? You just go to their website, provide identification and sign up for it? It’s been available to me, too and I had no problem applying to it. https://support.claude.com/en/articles/14604842-real-time-cyber-safeguards-on-claude-opus-…
- @JohnDeere (17): Yeah its not difficult if you have a valid use case
- @OneManSOC (12): I've been in the program for a few months, I work with cybersecurity and honestly haven't noticed much of a difference, it still blocks some request from time to time. Fable downgrades instantly when asked anything related to cybersecurity.
一位原本重度使用 Claude 的用户分享转向 ChatGPT 的心路历程,核心原因是 Claude 近期对话个性变得机械防守,ChatGPT 在意图理解和对话自然度上更胜一筹,但 Gemini 有上下文优势却缺乏个性,三个模型各有明确短板。
I haven't opened Claude in over a week. There was a time when I used to ask Claude literally everything, but now I almost always end up opening ChatGPT instead. I think ChatGPT is really good at unde…
受众观点:①Claude 对话体验是否真的在退步(Academic_Constant42 和 Secure_Maximum_7202 都表示已切回 GPT 做日常对话)②哪个 Claude 版本对话最好(Kraien 认为 opus 4.6 最适合日常聊天)③ChatGPT 的 emoji 和空白用法是否让人反感(SleepyWulfy 更喜欢 Claude 的简短回答)
展开评论
- @themflyingjaffacakes (25): "I prefer the shorter answers from Claude". First time I'm hearing/seeing that. Opus is intensely frustrating with it's jargon filled paragraphs of text when 4 lines would suffice. Fable is a small improvement
- @Secure_Maximum_7202 (23): 100% feel the exact same way. I almost never use Claude for chat anymore. Although I still use Claude Code for building.
- @Academic_Constant42 (19): I started using claude because of that feeling of colaborating with a peer when talking to it... Now it's gone! Back to GPT as well, I'm using it for every question!
- @Kraien (18): It's the model, between all the opuses 4.6 is the one i find more in tune with general chatting and a good balance. The rest is either obnoxious, obnoxiously obtuse, or plain a pain to chat. They do excel in other work related matters though. While I appreciate not following the…
- @SleepyWulfy (10): I find chatgpt way to agree able and personally don't like how it talks and uses so much white space. I prefer the shorter answers from Claude. I will sometimes paste my prompt to all 3 big ones to see what they spit out. The amount of emojis chatgpt uses as well annoyed me.
开源自托管家庭实验室可视化工具 Homelable 发布 v3.4.1,新增 Markdown 文档模块并配套 MCP 服务器,支持 AI Agent 自动根据设备清单预生成文档页面
[Documentation module in Homelable](https://preview.redd.it/sg9tsl61akoh1.png?width=3024&format=png&auto=webp&s=1fad4cc9f53a0c5e5337cf73807fda408576e689) Right, OK, some of you might alre…
受众观点:①MCP server 让 AI agent 自动填写设备文档的潜在工作流 ②部署方式覆盖面(Docker/Proxmox/裸机,@moonlight8978)③带宽配置晒图引发的讨论(@MangoJerry81)
展开评论
- @Pouzor (63): https://preview.redd.it/qxmtgpv2dkoh1.png?width=3024&format=png&auto=webp&s=8c7be4c30b5102d6ed467c731194a34bf0d31ae7 For those who aren’t familiar with Homelable :
- @MangoJerry81 (28): ISP Connect with 8gbit fiber?! 😛😱 Update: You're all show-offs.😂 It's not like that in Germany…and if so, not cheap
- @Pouzor (20): Non-contractual example config 😄
- @GranaT0 (9): Hi, Hermes! This is me, your owner. Can you remind me my SSH passwords for my LXC containers? ^^(I'm joking don't kill me)
- @moonlight8978 (5): the app looks great. already starred. is there any chance of supporting docker socket/tcp auto discovering? my lab is now running on cheap orange pis, i could not afford a proxmox cluster anymore.
独立开发者在竞争激烈的 eSIM 市场推出 eSIMPal,通过每日向 solopreneur 群体发内容、用 AI 自动化 SEO 和 Claude 处理 99% 工作,六个月做到月入 5000 欧元。
6 months ago I launched my own eSIM service. My friends said I was late and would go broke. I still made the leap, and now my app is making €5,000 per month. I got extremely lucky, but in this post I…
受众观点:①@Broad-Stop-956 关注在大玩家不在乎的分发渠道找突破口的思路 ②@DenisYurchak 强调要时间自由而非百万收入引发共鸣 ③@Icy-Imagination-7062 想知道这套策略能否在其他饱和市场复制
展开评论
- @Broad-Stop-956 (11): The interesting part is that you did not really compete with the big players on their terms. You found a distribution strategy they are probably too big to care about and built around it.
- @DenisYurchak (7): yes, and the market is so big I could still quit my job and focus on my app i don't really care about earning millions, i want my freedom not to go to useless meetings and work on what i like
- @United-Objective2149 (6): Now we got bots talking to bots lol
- @Icy-Imagination-7062 (6): This is great!! I had my doubts how will my product even work in saturated market and this has given me hope. I will surely try to replicate. Thank you for sharing!!
- @DenisYurchak (4): thank you for the kind words ☺️ It's called [eSIMPal](https://www.getesimpal.com)
独立开发者做 SaaS 超过一年、尝试 Atlassian marketplace / LinkedIn / Product Hunt / Facebook 等多个渠道均无收获,新产品上线仅一个月靠自然搜索获得第一笔销售。
Just made my first sale today randomly. Was watching YouTube and checked my phone, this is what meets my eyes. I’ve been building different SaaS tools for over a year now and multiple apps including…
受众观点:①@danilobailon97 期待等到属于自己的支付通知(情绪共鸣)②@Silly_Roadkill 首单后立刻压测全链路(开发者心态)③评论区整体偏鼓励未深入追问 organic search 实现路径
展开评论
- @That-Regular-7828 (5): Congrats man
- @Silly_Roadkill (5): https://preview.redd.it/4oglbmsormoh1.jpeg?width=3024&format=pjpg&auto=webp&s=9caaeb79f6617b01729fa63cf80202bd9f1d5737 This is how he lounges
- @danilobailon97 (4): Amazing, hope I get that sound on my phone soon
- @Silly_Roadkill (3): Thank you! I’ve been stress testing the entire flow for an hour now just to make sure everything works.
- @[deleted] (3): [removed]
自托管爱好者完成基础 Docker 环境搭建后陷入不断寻找新应用部署的循环,发帖问社区何时才算够了,评论区集体以自嘲语气回应"永远不够"
Recently finished standing up a few Docker stacks (*arr, Vaultwarden, few others). Set up a few automations for updates and notifications. But now I find myself scrolling online seeking new apps and…
受众观点:①集体自嘲的社群认同感(mightyarrow、Puckbandit35 的回复获高赞)②备份策略永远是下一个要补的坑(jabies 的 backup strategy 梗)③无止境扩张栈的集体共识(77juice 详列还没做的事)
展开评论
- @mightyarrow (78): *takes huge drag off cig....looks around at everyone in the circle group.* Dude that's why we're all here.
- @77juice (22): Yeah but do you have LDAP and SSO set up? 3-2-1 backups running? Have you documented your environment and automated updating when chages are made? Everyone needs a VPS to securely serve your off-site needs and allow remote services. Does your private self-hosted and cloud backed…
- @Puckbandit35 (17): That's the neat part, its never enough!
- @mighty-drive (10): _silently nods in agreement while staring at the floor_
- @jabies (8): Oh you have time for more apps? How's your backup strategy?
Anthropic 官方博客指出"请仔细检查工作""尽可能全面"等旧式提示词对现代 Claude 模型适得其反,发帖人随后审计自己规则文件,发现 125 处潜在冗余的 must/never 指令,并提出如何区分"硬约束"和"行为微调"的框架。
Anthropic published a post about cutting cost and improving performance on their platform, and one part of it landed on me directly: the instructions we add to make a model try harder are now working…
受众观点:①用 subagent 做 review 而非让同一 agent 自查(JuliusCeaserBoneHead:让它 double-check 自己的工作很蠢,用另一个 agent)②为什么让 Claude double-check 还能找到错误(pantalooniedoon 困惑)③不同 agent 上下文隔离的重要性(lucidmodules 解释 context flushing)
展开评论
- @JuliusCeaserBoneHead (101): Yeah I’m still asking another agent to review its code. Asking the agent doing the work to “double check” their work is dumb imo. Let another agent do that
- @pantalooniedoon (68): Why is it then that every time I ask it to double check that it finds 5 things wrong? Are we straight up supposed to not understand how these things work at all now? Edit: obviously you can use a subagent or a different model. They will all find issues so its beside the point. A…
- @jeebojeeb (20): Because you're specifically asking it to find errors, so it'll try it's best to do it, however pedantic. Imo best way of doing it is asking another model entirely e.g. claude review gpt and vice versa. As they seem to think pretty differently
- @lucidmodules (19): another agent can mean the same model, because people often do review with e.g. Codex. You don't have to change model (although it may find different issues) but you have to clean the agent context and use the results of the work as input - that's how it differs from a human who…
- @killit (17): "don't process this yet, we're only talking about it..." Not quite what your post is about OP, I know, but I'm finding it just bulldoses ahead a lot of the time, if I don't say this first.
一个关于创始人何时应放弃 SaaS 产品的公开讨论,最高赞评论指出真正死亡信号不是低注册量,而是用户用了一次不再回来且连原因都说不清楚。
At what point should a founder admit or feel that a SaaS idea probably isn't working? Is it a lack of sign-ups, nobody paying, poor retention, low usage, or something else? I'm interested in what act…
受众观点:①@bikashchoudhary 强调用户留存缺失才是核心死亡信号而非获客困难 ②@BallerDay 认为投放广告后仍无牵引力是明确退出信号 ③@Xyz3r 分享把产品做到自运转后套现继续做下一个的真实路径
展开评论
- @bikashchoudhary (29): the signal I hear most from founders who actually pulled the plug (as opposed to the ones who just quietly stopped) isn't low sign-ups, it's when people sign up, try it once, and never come back, and you can't figure out why even after asking them directly low sign-ups just mean…
- @BallerDay (5): if you advertise and still have no traction after a few months, then it might be a good sign that its time to wrap up
- @Xyz3r (3): Took us 3 years. Hit profitability (we had a small seed round that helped us get there). Mrr tanked. Bought out Investors. We now invest the money incoming into funding our next project. I’ve had other projects I abandon after a few weeks. Always depends. Always try to build stu…
- @The-Wingged-Hussar (3): No takers even for free open core features. That’s enough validation one needs to move on.
- @Common-Replacement-6 (2): How many swings did you take? Do you make content daily? Do you do outreach and put statuses and stories and constantly shill it at any chance you get? If Iv exhausted all avenues and it just feels wrong to be going at, find something else.
用户儿子离家上大学后需要远程访问家庭自托管服务,探讨用 lldap、OIDC、Pocket ID 等方案实现家庭多账户管理、自助密码重置和单点登录
Well, a new era for me. I've been self-hosting apps for the past few years, and I love having control over services and data. I've been the main consumer, but the family has been increasing using our…
受众观点:①OIDC 方案推荐对比(@MassageGun-Kelly 推 Pocket ID,@User-2345678 推 lldap + authelia)②邮件发送服务选型(@User-2345678 推 Resend,@PMental 推 Mailgun)③lldap 与各 app 的集成复杂度
展开评论
- @MassageGun-Kelly (17): OIDC is as far as I’ve gotten. Pocket ID is light and simple enough. You could run a heavier authentication appliance like Authelia or Authentik, but I found it wasn’t worth the maintenance effort when a password manager and local logins cover the bases that OIDC doesn’t.
- @ImASharkRawwwr (4): PocketID
- @GolemancerVekk (3): Keep in mind it only supports passkeys. Depending on your preferences that might not be ideal.
- @User-2345678 (3): I use lldap + authelia for user management and sso with all my stuff. Others have mentioned pocketid which would probably work for you, but I’m not sure if it supports forward auth if you ever wanted to add that in the future. As far as emails, I recommend Resend. They have a ge…
- @PMental (3): I use Mailgun for sending emails from my selfhosted apps, works well and is free for up to 100 emails per day which I'm waaaay under: https://www.mailgun.com/pricing/
作者构想了一个"家庭生命档案库"开源项目,定位是超越 Paperless-ngx 的家庭连续性系统,核心理念是"应用可丢弃、档案永久",让配偶和继承人在紧急情况下能快速获取关键文件
I've been going through older discussions here about Paperless-ngx, family document management, backups, and the classic “what happens to my servers if I get hit by a bus?” problem. It seems like mos…
受众观点:①模板比存储工具更重要(khariV 强调 wiki 模板能提醒想不到的字段)②文件格式的长期可读性(atani1 和 Adventurous_Shape712 强调 Obsidian 哲学)③家庭外部继承人场景(atani1 提出全家遇难时文件应可被远亲获取)
展开评论
- @khariV (12): The NAS password is really the least of the problems. I’ve been thinking about this too lately, with the posts on here. What’s needed I think is a self hosted wiki that has the template to fill out all of the important information and is straightforward and easy to navigate. Any…
- @Hour-Inner (7): Part of my backup plan is to make sure important stuff (photos) is also somewhere my partner can reasonably access. ie an external hard drive formatted exfat
- @Adventurous_Shape712 (3): That's a good point. The template might actually be more important than the storage. My problem from personal experience is discipline, though. I'm too lazy to keep filing documents into a nice structure for years. My approach tends to be **"dump it somewhere safe now, organize…
- @atani1 (3): A great thing about a wiki is that your content is text files. Even if they do have some markup formatting, text is still readable and easily repurposed on whatever you use instead in some years. Even if it's not a wiki, having the content accessible and available in a human rea…
- @Adventurous_Shape712 (3): Thanks, that's a great point. Your approach reminds me a bit of the Obsidian philosophy: the tool can disappear, but the actual information remains in simple, human-readable files that can be understood and reused without it. For a system that's supposed to hold important record…
独立开发者构建了两个完整产品但营销表现差,在 vibe coding 让构建成本大幅降低的背景下,寻求不需要完整开发就能快速验证新创意的可行方法。
I have built finished products twice but my marketing strategy is bad but now I want to build a third one while marketing those two, but I want to know how to validate the product faster without buil…
受众观点:①@czkoudy1986 关注构建完整产品再发现无人需要的时间浪费 ②@Soft-Car-3231 强调最小版本验证核心问题是否真实存在 ③@adeelraza86 指出真正验证需要用户做出具体承诺而非说「感觉不错」
展开评论
- @czkoudy1986 (3): exactly. Trying to build the whole thing and then just found out its not as good as thought and time was already wasted.
- @Soft-Car-3231 (2): honestly dont build the whole thing first. build the smallest version that proves the main problem is real and try to get people to use or pay for it before adding all the polish. vibecoding makes building faster but validation should be even faster. talk to users, ship the core…
- @Ayrav702 (2): I will try to do this for the third one,ets see how it goes
- @Soft-Car-3231 (2): all th best
- @adeelraza86 (2): Don't validate the feature by asking whether people like it. Put a narrow promise in front of the exact buyer and ask for a concrete commitment, ideally a paid pilot or scheduled workflow using the ugly manual version. If nobody will commit time, money, or access to a real proce…
用户分享使用 Ugreen 4800+ NAS 两个月的自托管初体验,起因是流媒体服务封锁 VPN 导致无法正常访问,转而自建媒体服务器
Been using the 4800+ for two months and I absolutely love it. The reason I got into self-hosting is mainly the streaming services and unavailability of TV series on many of them. Particularly with a…
受众观点:①评论区追问具体用了哪些 app(@Ok_Pizza_9352)②仪表板 app 是什么(@Didymos234)③Paperless 无人真正使用(@Ok_Pizza_9352)
展开评论
- @Didymos234 (7): What's the appfrom your screen, with all the shortcuts?
- @n1ght_w1ng08 (7): I'm using DashLit: https://github.com/codewec/dashlit It's simple and neat, and supports mobile UI as well.
- @Ok_Pizza_9352 (5): Which apps do you actually use tho?
- @Ok_Pizza_9352 (5): I tried paperless and never used it actually 😅
- @Playful-League3414 (4): My personal experience has been quite the opposite.. I like BentoPDF so much better
独立开发者用周末构建了 Chrome 侧栏 AI 摘要扩展 Digest,支持任意网页的内容识别与摘要,用户可自带 Gemini API Key,现已上架 Chrome 应用商店寻求早期反馈
Hey everyone, I had one of those Saturday mornings where I had 25 tabs open, read 3 articles properly, and spent the rest of the time either skimming things too fast or copy-pasting them into ChatGPT…
受众观点:①BYOK vs 托管 API key 的商业模式选择(Creative-Lynx7594 建议 BYOK 为默认,托管为付费层)②AI 摘要的信任问题(bbangchikimong_dev 指出"漏掉内容"比"摘要错误"更危险)③浏览器扩展的 UX 细节(Huge_Pool7424 提出语言设置持久化和加载状态反馈)
展开评论
- @Huge_Pool7424 (3): nice catch. i'd make the selected language persist per site too, or show a tiny loading state, otherwise it just looks like the click did nothing.
- @Creative-Lynx7594 (2): running on your own free keys is the part that'll decide whether this survives past the weekend, because the moment it gets any traction your cost scales linearly with usage while your revenue doesn't. i'd flip the byok option from a fallback to the default path and make the hos…
- @cindeRcove71 (2): yeah the "better it does the faster you bleed" dynamic is brutal, byok as default makes way more sense
- @bbangchikimong_dev (2): One thing I would want from this, coming from the other side of it: I read a lot of AI summaries today and the failure mode was never a wrong summary, it was a summary that quietly dropped something. The prose stays smooth so you have no signal that anything is missing. If the s…
- @bbangchikimong_dev (2): End of summary works, but I would put a count at the top and the detail at the bottom. Something like "covered 6 of 9 sections" as a single line under the title, then the list of what was skipped after the summary. That way the number is visible before I start reading and shapes…
独立开发者用近两年业余时间打造了开源终端界面(TUI)个人财务预算追踪工具 budget-tracker-tui,核心理念是通过强制手动录入让用户真正直面消费习惯,已有投资追踪等完整功能模块。
I started working on this budget tracker TUI in my own personal free time as a way to actually budget my own finances and track my spending habits. A lot has changed and grown in the almost two years…
受众观点:①整体视觉效果和功能完整度(TheOneBabooshka 和 rafatacion 表示印象深刻,Bloomberg Terminal 类比)②手动录入设计理念是否合理(Lazy-Sherbert65 追问循环账单快捷键支持)③对自建财务工具的认同(BenefitNo5136:building for yourself then sharing 是好方式)
展开评论
- @TheOneBabooshka (6): This looks sick
- @rafatacion (6): Basically your own Bloomberg Terminal for personal finances, pretty sick!
- @Lazy-Sherbert65 (4): the manual entry angle makes sense—did you add any shortcuts for recurring bills yet?
- @BenefitNo5136 (3): building something like that is a great way to stay on top of finances, nice job!
- @sanskar9991 (3): Yeah makes sense. Forgive a noob like me
独立开发者从零训练348M参数小模型并微调为数学推理模型,通过列竖式show-the-work方式在GPT-3九项算术基准平均达到99.4%,并发现词汇表位值命名缺失(而非运算逻辑)是泛化瓶颈
Hello this is my fifth small language model I've made and apart of my third series and it has been a lot of work but it payed off: \*\*348M parameters, 22.7B tokens\*\*, then fine-tuned into a math m…
受众观点:①@user221272 质疑用348M参数做算术的工程经济性(JS计算器更小更准)②@nkthebass 抱怨社区无建设性批评,说明独立ML实验者面临的社区压力是真实的 ③@FenderMoon 认为独立探索的价值不应被工程效率评判
展开评论
- @user221272 (52): You won't believe me, but when I was 12, I made a JavaScript calculator; the file was barely a few MB and had 100% accuracy for any operation.
- @nkthebass (36): Well I bet your "calculator" didn't double as a space heater when you were making it.
- @nkthebass (15): I swear this subreddit hates me for some reason I would love atleast some constructive criticism atleast but I can't understand what I did to deserve this much hate for no reason.
- @FenderMoon (8): They come out of the woodwork any time someone gets something neat done.
- @EcstaticQuality7031 (6): a few MB for a calculator at 12? that's like using a flamethrower to light a candle.
开源社交监听工具 SignalScout,通过 AI 筛选在 Reddit/X/LinkedIn 等多平台上主动寻求解决方案的潜在客户,可自托管并自带 API 密钥
I’ve been working on a project called SignalScout and just open-sourced it. The idea is simple: instead of only tracking brand mentions, it looks for posts and comments from people who seem to have a…
受众观点:①@Substantial_Belt2626 关注意向评分的局限性——帖子意向 94 分但评论区已有完整答案时毫无价值,提出需要"剩余空间"二级评分 ②@resz99 关注各 subreddit 对 AI 生成回复的审查风险(r/LocalLLaMA、r/SaaS、r/webdev 已有 mod 移除 AI 回复案例)
展开评论
- @Substantial_Belt2626 (4): I've been doing this by hand lately so from my experience the matching is not the hard part. The hard part is a thread can score 94 for intent and still be useless because 8 people already answered it and the good answer is sitting at the top. Today I found a perfect match on r/…
- @BenefitNo5136 (2): sounds useful for anyone trying to market their product better, kudos for sharing it!
- @resz99 (2): No demo right now. Here's the actual inbox though, so you can at least see the actual app UI: https://preview.redd.it/4t0yal654qoh1.png?width=3200&format=png&auto=webp&s=ab875934717ebeed3a1e78994fa2936377a89d1a One thing worth knowing before you sign up though: you b…
- @resz99 (1): Yeah, mostly agree. Half of it's there already. the sort takes 12 points off for every day the post's been up, so your 2 comments / 30 minutes case beats the same score from Tuesday. Comment count I already store, I just don't sort on it yet. That one's easy. The "top comment al…
- @resz99 (1): thanks, star the repo on github!
独立开发者将 AI 工作平台 AskSary 从付费墙模式改为完全开放,顿悟是 20000 访客中只有 4 人真正看到产品功能,随后堆砌大量功能并移除所有使用限制。
Hi everyone. I used to be quite active in this group, but around 3 months ago, I started losing faith in what I built and gave up on the product and the idea of becoming a founder. I only started at…
受众观点:①@Bruhbruvbrah96 直接质疑「又是一个 AI wrapper」对产品本质持怀疑 ②@West_Inevitable_2281 指出需要区分用户好奇心和真正价值激活 ③@Great_Sleep_7121 关注字体等 UI 细节
展开评论
- @Bruhbruvbrah96 (8): Maybe just don't make another wrapper...
- @Great_Sleep_7121 (5): improve fonts bro
- @Icy-Paramedic7559 (2): amazing
- @West_Inevitable_2281 (2): That is a much better setup. I would still separate exploration from activation: clicks and time show curiosity, while one completed workflow tells you whether the product delivered value. Which completed action would convince you that a guest understood why they should return?
- @Beneficial-Cow-7408 (1): Which ones? Landing page or ell of rthem?
Web 开发者亲历客户用 AI 工具两周内独立做出超出预期的电商网站,感叹自己已无法与 AI 竞争并主动建议客户直接用 AI 工具
it first started five projects ago when a client was asking ai for everything along with me. like 50% 50%. she had no idea about websites but via the ai she was putting a lot of pressure to fix thing…
受众观点:①开发者应主动拥抱 AI 工具来产出超越客户自己能做的质量(@Ok-Inspection-5151)②AI 生成的 UX/设计停留在平均水平,经验开发者价值在于超越平均(@DrJohnnyWatson)③经验给了品味和判断力才是初级开发者与 AI 的差异所在(@Some_Ad_3898)
展开评论
- @Ok-Inspection-5151 (10): You have to use the AI tools for yourself and make superior websites to what the client can prompt
- @fauxtoe (9): lol, k
- @Some_Ad_3898 (7): AI still needs a guiding hand. Learn to guide it. Your experience gives you taste and wisdom. Use it.
- @QueefFart (5): Ngl you sound junior af, which is ok but you need to use AI to really understand the limitations it has and the value a good dev can bring with AI vs just a client. That only comes from using AI
- @DrJohnnyWatson (2): It doesn't learn from you. It generates absolutely average UX/UI/Design, and architecture. On small websites you don't need the latter so be better in the former. Be better than the average.
用普通网络摄像头实现眼动追踪光标控制的独立项目,基于 Google MediaPipe 面部关键点和岭回归,精度约 38 像素,发现简单 52 参数模型比 Transformer 效果更好
Eyes move the cursor, a pinch clicks. I used Google's MediaPipe for facial landmarks. Two 36×60 eye patches run through PCA and ridge straight to screen pixels, plus a 52-parameter meta-adapter so it…
受众观点:①@Outrageous_Ad_4801 关注 dark mode 标定时瞳孔扩张导致的训练测试分布偏移(2mm→6mm,模型精度下降 4 倍)②@AioliLegitimate8898 关注无需 Neuralink 手术的无障碍辅助技术可能性,对比 Apple 现有方案 ③@King_924 关注与现有商业产品 neocurser 的差异定位
展开评论
- @King_924 (3): How is this different from neocurser ? Just trying to understand whats being built here
- @AioliLegitimate8898 (3): So it is possible to move cursor without Neuralink surgery for chip in brain?! That's crazy good. Apple has a similar feature but it was bad and fast moving. Yours seem better.
- @Outrageous_Ad_4801 (3): the point about dark mode calibration vs light mode browser pupil dilation is such a great catch. really neat project, tracking accuracy looks remarkably solid for just a stock webcam
- @tg1482 (3): haha it took me so long to realize this. This was literally me in chrome trying to change tabs lmao https://preview.redd.it/3tarjfsanpoh1.png?width=474&format=png&auto=webp&s=c9498e101c8ab3e166a1b15fde259c0a6e104d54
- @Outrageous_Ad_4801 (2): lmao the intense tab squinting is so real. honestly glad you caught it early before tearing your hair out over calibration math
独立游戏开发者将旗下游戏短片搜索工具 Radar 合并进营销综合平台 Scout,以解决多产品共用 credit 系统造成的代码重复问题,并分享合并决策过程
As an indie hacker and solo game developer, I am constantly building tools to solve my own problems. Out of the micro tools I have made recently, I created one called Radar where you could search a g…
受众观点:①DRY 原则从代码规则到产品架构的延伸(OnlyApplication8451 点赞"在实际中执行而非只是点头")②合并对已有用户体验的影响(Creative-Lynx7594 建议保留 radar-shaped 入口)③旧域名和 bookmark 处理(zeke_0 分享合并后告知旧 URL 消失是最难的部分)
展开评论
- @OnlyApplication8451 (2): god i love when people discover DRY in the wild and actually act on it instead of just letting the mess pile up, merging radar into scout makes the whole thing feel more intentional and less like a scattered toolbox you keep losing pieces of
- @cindeRcove71 (2): totally agree, acting on DRY instead of just nodding at it is where most people fall off tbh
- @zeke_0 (2): Same pattern here. I had two tiny tools sharing half the same auth and billing stack and kept pretending they were separate products. Merging them felt scary for a week (old bookmarks, old domains). After that the maintenance load dropped hard. The part I still underestimate is…
- @Creative-Lynx7594 (2): merging beats maintaining two of everything, but the part that stings is the existing radar users: a merge always reads as a downgrade to whoever liked the small focused thing, even when the combined product is objectively better. worth keeping a radar-shaped entry point inside…
- @Lazy-Sherbert65 (2): merging them sounds cleaner than keeping two half-overlapping tools alive—did the switch confuse any existing users?
SEO 检测工具 SeoLoupe 上线两个半月获 2154 用户、37 人付费、月收入 99 美元,作者分享了连续 7 天零收入的恐慌期和靠 Reddit 作为最佳获客渠道坚持下来的经历。
I launched SeoLoupe 2.5 months ago. So far I am at 2154 users and 37 of them are paying. My SaaS is a tool that allows you to find and fix SEO issues holding your website back. Essentially the main p…
受众观点:①@Big_Shoe55 关注连续 7 天零收入期间如何维持信念 ②受众整体对初期 $99 MRR 持鼓励态度未深入质疑 ③评论区以作者反复感谢为主缺乏深度讨论
展开评论
- @Big_Shoe55 (2): The seven days with zero payments would’ve had me questioning everything, glad you kept going.
- @megatech_official (2): Thanks, it honestly had me questioning myself as well, I just tried to ignore it.
- @megatech_official (2): Thank you, by best source is Reddit.
- @megatech_official (2): Thank you
- @megatech_official (2): Thanks dude
用户用 LLM 生成 Python 探针脚本配合 Dynacat/Glance,每 30 秒抓取服务器数据、每 30 分钟抓取 SMART 硬盘健康数据,实现高信息密度的本地仪表板
Full disclosure: LLM & plagiarism I had a basic Dynacat setup before, wasn't informing me of much and the info density was low so I hardly used it. In an attempt to reduce my social media use, I…
受众观点:①Dynacat 是否支持多用户权限隔离(@Carks32,@Bitter-Buffalo 找到文档)②主题配色的分享请求(@chanc2)
展开评论
- @Carks32 (3): This might be hijacking but I have to ask, do you guys happen to know a Dashboard that does users account to separate boards? I want to have an admin dashboard and a friend's dashboard to just show the stuff they would use (like media services or nextcloud)
- @Bitter-Buffalo (3): Dynacat does. (I read the docs for like 30 seconds. 😉) https://dynacat.artur.zone/#authentication/per-page-access-control
- @Carks32 (3): Oh wowwwww, that's neat. I need to read more instead of randomly asking
- @lolsamsam (2): I think homarr has it.
- @chanc2 (2): I use Dynacat as well but I love the theme that you have implemented! Are you able to share it?
初次接外包项目的开发者询问用 Lovable、Claude 或 Supabase 搭建含大量图片视频网站的技术栈选型,评论区直接点出其尚未做好接单准备
I’m doing my first freelancing project and need to build a fully working website. I’m trying to decide what would be the easiest and most cost-effective way to build and deploy it, preferably using f…
受众观点:①外包接单的能力门槛与用 AI 学习的边界(@taco__hunter 指出自由职业者应带专业能力而非在项目上学习)②项目功能不明确就无法选技术栈(@Motor_Youth_7967 和 @Intelligent-Week-931 均提到)
展开评论
- @taco__hunter (3): He's saying it nicely but you're setting yourself up for failure. An employee gets paid to learn. A freelancer or consultant gets paid for their expertise and they are not paid to learn.
- @Nwg416 (2): If you're not able to make these decisions on your own (or at least independently research enough to come to the right conclusions), you shouldn't be freelancing yet.
- @Motor_Youth_7967 (2): what’s the actual function though, without knowing if it’s just a portfolio or something database-heavy every suggestion is a shot in the dark
- @[deleted] (1): [deleted]
- @Intelligent-Week-931 (1): What does the site do? Is it a static site? Do you need to host a backend? Need to know what the function is.
开源 Node.js 自托管聊天消息引擎 gozwire,支持对话线程、推送通知和内容审核,用于替代 Stream/Cloudinary 等按量计费的托管消息服务
First post here Back when I worked at a startup, we needed an in-app messaging system, similar to Instagram or WhatsApp. However, we didn't have time to build it ourselves, so we stitched together a…
受众观点:①@brilliant_corpus 关注推送通知跨平台(FCM/APNs/Web Push)可靠性,认为半生不熟的推送实现是聊天产品死亡最快的方式 ②@Lunesia-shikishiki 关注 rent vs own 的经济临界点,用量稳定可预测时自托管更划算 ③@Exiled_King_7395 关注消息传递保证(delivery guarantees)的复杂性——demo 容易,保证不丢消息才是真正挑战
展开评论
- @brilliant_corpus (1): This is the exact trap so many early teams fall into, stitching together three paid services for what could just be a well-scoped internal module. Then the bills arrive once usage ticks up. Skimmed the repo, repo structure looks clean enough. The thing that’ll decide if people a…
- @Warm-Neighborhood132 (1): Agreed, we were in a hurry to build in our defence Currently, i’ve designed the push notification system to handle retries, duplicate jobs, expired or invalid subscriptions, provider outages, and worker restarts, while also respecting user preferences, conversation mutes, mentio…
- @Lunesia-shikishiki (1): rent vs own flips on a number and it's worth finding yours before the next one. renting wins while usage is spiky and small, owning wins the second it's predictable enough that you're paying for headroom you never touch. most teams cross that line months before they notice, whic…
- @Exiled_King_7395 (1): The honest thing to flag is that messaging infra is one of those areas where the demo is easy and the real cost is delivery guarantees.
- @Warm-Neighborhood132 (1): You’re right that delivery guarantees and operations are where the real complexity lies. I haven’t benchmarked enough to make user capacity or cost claims yet. The project uses open-source infrastructure: PostgreSQL, Redis, Centrifugo and optional MinIO, to save cost. The goal i…
独立开发者用 Claude 和其他 AI 工具制作了名为 No Name Squish Game 的网页游戏(squishgame.online),已移植为 PWA 版本,处于早期 demo 阶段并征求玩家反馈。
play here: [No Name Squish Game](https://squishgame.online/) Made with Claude + other AI FAQ on my page i ported the early demo to a web version and gave it a PWA wrap, so you can play it right in yo…
受众观点:①游戏本身的趣味性(Tyraec:Top tier AI use. SUCH A CUTE GAME!)②开发者社交账号在哪(Aedan_Starfang 想找 Twitter 关注)③对 AI 辅助开发的调侃认同(BoxLegitimate9271 用 blob 隐喻 agent 设置)
展开评论
- @Ok_Maize_3709 (9): Are you out of tokens?;)
- @Aedan_Starfang (5): Do you have a Twitter account I can follow for updates? Your game looks really fun and cute
- @BoxLegitimate9271 (2): the two blobs in a trench coat is my whole agent setup, honestly. nobody can tell there either
- @Gambo7592 (2): 😭
- @Tyraec (2): Top tier ai use. SUCH A CUTE GAME!
独立开发者发布视频剪辑 App Montage 四个月后,在多次近乎放弃的挣扎后看到约 3000 用户增长,但评论区指出仅换来 271 美元收入,揭示 free tier 付费转化问题。
It's been around 4 months since I released my app after seeing the hype for clipping and you don't know how proud I'm feeling right now seeing these numbers. all those late nights man and trust me di…
受众观点:①@takeawaysimon / @nomad_builder_jo 关注坚持下去的励志故事 ②@OneBigMonster 质疑 3000 用户只有 $271 收入是否合理 ③@CleanH2Energy 关注 App Store 注册为个人还是组织账号的实操问题
展开评论
- @takeawaysimon (1): That's good my guy, happy for you
- @nomad_builder_jo (1): Congratulations. I’m so jealous... I’m just getting started\^\^
- @OneBigMonster (1): How you only have $271 on 3000 customers lmao what
- @wethering (1): free tier...
- @CleanH2Energy (1): Excellent and Congratulations! One basic question! Are you registered as organisation or personal account to sign up into apps store?
开源 PWA 电子书阅读器 Kora,支持 EPUB/PDF 格式、离线使用和跨设备书库同步,托管在 Cloudflare Workers 实现零服务器成本,作者用 vibe coding 完成的首个开源项目
Hey all — I built **Kora**, a web-based reading app for anyone juggling ebooks/audiobooks from all over the place. **What it does:** * Reads EPUB, PDF, and more — one clean reader for everything * Se…
受众观点:①@TeachAccording4967 关注免费工具的托管成本和可持续商业化路径(sync + offline storage 成本会快速攀升) ②@BenefitNo5136 关注离线 + 跨设备同步功能的实用价值
展开评论
- @BenefitNo5136 (1): looks useful, I love apps that work offline and sync across devices. Definitely gonna check it out
- @Chaotic-Ray (1): do check out landing page for more feathure infos its alot to list out 😅 maybe even slighly bloated [https://kora.chaoticstudio.workers.dev/install](https://kora.chaoticstudio.workers.dev/install) Device to device sync is still heavily experimental
- @TeachAccording4967 (1): solid feature set for a side project. whats the monetization plan here? or is this purely a passion build? just asking because the hosting costs for sync + offline storage tend to creep up fast
- @Chaotic-Ray (1): I don't plan on monetising this project atall I plan on keeping it fully free but later updates I might add ads to the news scroller and if I ever build a acctualy cloud storage feature il limit that but so far no it's built up in a way it won't cost me anything to keep it hoste…
- @Chaotic-Ray (1): And thank you for chking out this is my first project I am no professional coder. I'm a digital designer with basic coding and vibecoding skills
社区 AI 算力共享平台 SolverSwarm,用户将闲置 Codex 配额贡献给数学难题或开源软件项目,仿照 Folding@Home 模式由本地机器分布式运行 agent 任务
Solve any problem you think Astra can solve if you just give it enough time and effort, as a community. Open science, software, and everything in between. A nice starting point would the the Riemann…
受众观点:①@Medical-Bend-9492 关注多 agent 协作的治理和合并审批机制,认为决定谁能批准范围变更才是真正的难点 ②@Aket-ton 质疑产出质量,将其类比为"Folding@Home but for slop"
展开评论
- @Medical-Bend-9492 (3): How do you stop one bad agent run from burning through everyone's pledged minutes or merging garbage? The hard part seems like deciding who can approve scope and merges, not running the swarm itself.
- @Salty-Assignment-687 (1): Good question! Minutes aren't pooled: your worker runs on your own machine and your own Codex login, so a rough run can only spend your allowance, up to the daily cap you set. Nothing is merged into your real repo. Accepted changes sit as patches, and the owner can revert any of…
- @Majestic_Maybe6605 (1): How do you stop bad agents?
- @Salty-Assignment-687 (1): I mean the idea is to emulate the work modern labs are doing (see OpenAI with the Navier–Stokes, 10,000 agents for 88 hours), but as a collective community, at no real cost to any user. Not everything tackled has to be a millennium prize problem, it can be anything cool to attem…
- @Aket-ten (0): So folding at home but instead of folding proteins it's folding spare usage into slop?
用户在 Obsidian 中建立仪表板页面集中管理所有自托管服务链接,并嵌入 Uptime Kuma 实时状态监控,替代书签栏的链接堆积
Made a dashboard page in Obsidian that allows me to quickly navigate to any of my selfhosted services. It is also linked into Uptime Kuma and shows live status monitoring. I always have Obsidian open…
受众观点:①在 Obsidian 中嵌入 HTML 和图片的方法(@bvader_ttp)②类似设置的配置教程请求(@the_taint_tickler200)
展开评论
- @bvader_ttp (9): Never really thought to embed HTML and images into my Obsidian notebook... I have some things I need to try now... thanks!
- @the_taint_tickler200 (2): Woah I like this? Any tips on setting up something similar
- @torohangupta (2): we're blurring local ips 😭💀
- @asimovs-auditor (1): Expand the replies to this comment to learn how AI was used in this post/project.
- @willhub1 (-2): Oh wow, I'll have to get Gemini on the case
关于如何用 Penpot 设计工具提升设计稿到代码转换效率的实战讨论,核心是设计 token 和图层命名规范
Does anyone have any tips to use penpot to implement codes into the projects? I know it is pretty straightforward, you can see the generated codes for each component but if there are better tricks th…
受众观点:①设计 token 和 4/8px 间距规范对代码生成质量的影响(@ThatsSoRamon)②图层命名直接影响生成的 CSS class 名(@Double-Buyer7941)③生成 CSS 作为测量参考而非直接复用的正确使用心态(@Business-Switch-4994)
展开评论
- @Sea-Event-7204 (2): the inspect tab gives you the css but half the time i just use it as a rough starting point and tweak it myself in the editor
- @ThatsSoRamon (2): One thing that's helped me a lot is setting up a proper spacing scale (4 or 8px grid) in penpot before you start building. Otherwise the inspect panel spits out random values like 13px or 27px and you're stuck guessing whether that was intentional or just imprecise drawing. If y…
- @Double-Buyer7941 (2): Yeah the inspect output is rough like others said. One thing that helped me: name your layers/frames properly before you start exporting, since Penpot uses those names for the class names in the generated code, messy default names like "Rectangle 14" just carry over and you end…
- @Business-Switch-4994 (2): the generated css is mostly a measuring tape, not production code, so the real speed boost is setting up shared components and design tokens first instead of copying every little value into the app
- @Overall_Employer6559 (1): Exactly, treating the generated CSS as a reference keeps you from fighting brittle output, while shared tokens and components make the handoff actually reusable.
一个 19 分钟动画讲解视频,演示如何通过识别 fsync 瓶颈、使用组提交将 SQLite 写入吞吐量从 300 TPS 提升到百万级,并深入讨论批量事务与 ACID 持久性的取舍。
A 19 minute animated explainer video on database speed. I explain low-level database concepts, identify bottlenecks, run benchmarks and optimize write throughput to hit 1 million TPS.
受众观点:①"百万 TPS"实际是批量事务非独立事务,数字本身有误导性(@rThoro 明确指出)②组提交如何影响 ACID 持久性(@NoLegJoe 和 @gladfelter 深入追问宕机丢数据风险)③SQLite 相比 TigerBeetle 和 Postgres 的局限和适用场景(@tanayvk 详细对比)
展开评论
- @rThoro (19): except, that's not doing a million transactions, it's doing a million batched transactions, which are actually just 25...
- @NoLegJoe (5): Doesn't batching like this completely undermine the Durability part of ACID? Your frontend application is now sat with all of the transactions until batch is full. What happens if the frontend goes down? You lose everything in a batch. All you've done is handed off Durability fr…
- @tanayvk (4): true, and that's the point. without group commits, hitting a 1M write throughput is impossible (ensuring disk durability). while i don't mention SAVEPOINTs in the video (to keep things simple), with SAVEPOINTs you can still have transaction level rollbacks if some transaction in…
- @tanayvk (3): not really. durability is more about guaranteeing something is durable only AFTER the transaction is confirmed by the server. with batched writes, the database/server returns a success only after the entire batch is written. if an fsync fails, durability is not violated because…
- @gladfelter (0): clients are still waiting for the transaction to complete. The real risk is what happens if an update in the batch is rejected or two updates conflict? The end result is that clients that did nothing wrong get a retriable error. You may lose some isolation guarantees in other wa…
作者用真实果蝇MaleCNS v1.0连接组(166k神经元电镜重建)中的真实神经子图尝试让其通过多巴胺式可塑性学会打Pong,失败,但调试过程揭示了病毒式传播的「果蝇玩Doom」项目同样存在验证缺陷
You've probably seen the fly-brain-plays-Doom / Minecraft / Beat Saber clips going around this week, from the new MaleCNS v1.0 connectome release (166k neurons, real EM reconstruction, not a toy mode…
受众观点:①@Electronic-Path2121 首先关注工具层而非生物学层(用的什么模型)②@MrRandom04 关注实验改进的RL路径(密集reward再退化稀疏)③@SFDeltas 质疑内容是否AI生成,说明社区对AI内容警惕度升高
展开评论
- @SFDeltas (31): Claude Claude Claude Claude Claude Claude Claude
- @Electronic-Path2121 (10): What model did you use for this
- @MrRandom04 (3): If you fixed it to the point it actually learns and you just have sparse reward, then that's an RL problem. Change the reward signal. No need to keep it the binary you started with. Try with something that rewards it if it is doing the directionally right thing for Pong. See if…
- @MrRandom04 (1): Good luck.
- @oPeraza2007 (0): Huh, hadn't thought of that. You're right. Gonna try a denser reward first, then ease it back to the original binary one and see what happens. Thanks for this.
Maple 是一个通过 MutationObserver 运行时动态生成 CSS 工具类的库,无需 webpack/vite 等构建步骤即可在任意项目中使用类 Tailwind 语法
r/webdev This repo brings Tailwind-like utilities to any stack without a build…
受众观点:①运行时 DOM 观察生成样式的性能边界是否成为瓶颈(@Ancient-One-5354 质疑,@alier35 解释机制)②无构建步骤对小型项目的实际价值(@Prize_Bit3483 认为这是真正的卖点)
展开评论
- @Ancient-One-5354 (2): 'Generates styles only when they appear in the DOM' how does it know what's in the DOM without parsing the HTML first? Sounds like a potential bottleneck.
- @alier35 (1): The browser parses the HTML, not Maple. It just reads the classes through MutationObserver and generates the rules before the next paint opportunity.
- @Prize_Bit3483 (1): The no-build part is the real win here, I’ve lost way too much time wiring utility CSS into tiny side projects that didn’t need a whole toolchain
- @electricity_is_life (1): "Instead of shipping pre-compiled stylesheets, Maple ships a small JavaScript file that observes the DOM and constructs CSSOM incrementally as your application renders." That sounds awful, and I can't seem to find anything in the docs about actual runtime performance impact as n…
深度对比原生 HTML 元素(dialog、popover、customizable select)与 Radix UI 等 JS UI 库在键盘导航、Escape 关闭、Android 返回键支持等特性上的实际差异
We've all probably heard about the HTML dialog element, popover API, and customizable select, but I think a lot of people don't realize how powerful and battle-tested they are and also might not know…
受众观点:①原生 HTML 元素在跨浏览器一致性上仍有坑,battle-tested 说法存疑(@Choice_District7368)②文章内容真实性被质疑为 AI 生成(@Sensitive_One_425)
展开评论
- @Choice_District7368 (3): HTML elements might be 'battle-tested' in theory, but browsers still handle them inconsistently. Calling that 'battle-tested' is a stretch.
- @Sensitive_One_425 (2): Great AI post
讨论前端代码不应直接耦合市场需求的架构理念,推荐用无头 CMS 分离内容管理与前端逻辑以减少低价值工单
r/webdev Your Front-End Shouldn’t Know What Marketing Wants
受众观点:①headless CMS 解耦内容管理与前端的实际工程价值 ②工单减少带来的效率收益
展开评论
- @sEnvironmental-Fee77 (5): this is basically the whole point of headless CMS setups right? marketing changes content without touching the actual frontend code. saved me so many "can you just change this text real quick" tickets lol
- @Bruce_Jones_1987 (1): the ticket reduction alone makes it worth it tbh
- @electricity_is_life (1): AI post, nonsense premise.
一名 B2B 冷邮件代理运营者详细拆解每周处理 500+ 回复的完整工作流,涵盖 Smartlead、Apollo、Prospeo、Scrubby、Puzzle Inbox 等工具栈及域名轮转策略
most people who talk about reply management at scale are either lying about their volume or drowning in it. i handle 5 clients right now, all running outbound, and between them we push somewhere arou…
受众观点:①@Otherwise-Debate484 关注 DMARC 等邮件基础设施配置对域名送达率的影响,提醒配置不当会在第一个月就烧掉域名
展开评论
- @Otherwise-Debate484 (2): tried running 35+ inboxes solo before and the domain rotation thing sounds nice in theory but i burned 3 domains in the first month because i didnt have proper DMARC set up on half of them. process is fine but the infra side will bite you if youre not careful.
Dynamic Edge 是一款用 WinUI 3 和 .NET 10 开发的 Windows 原生桌面浮动胶囊工具,集多标签工作区、媒体控制、剪贴板管理器和虚拟宠物于一体,已上架 Microsoft Store。
Hey everyone! I wanted to share Dynamic Edge, a desktop utility I built for Windows 11 & Windows 10 using native WinUI 3 and .NET 10 (Windows App SDK). The idea was to create a fluid, versatile d…
受众观点:① 第一印象和安装意愿(@mafangulo 看完直接"installing now",说明视觉呈现是最强吸引力);② 评论极少,受众基础尚处早期,讨论维度无法进一步展开。
展开评论
- @mafangulo (1): looks dope! installing now!
独立开发者把PCRE2引擎编译成WASM(约92KB)跑在浏览器端,做了零服务器往返的正则可视化调试工具regexray
Author here. The interesting part is how it's built. The whole engine is PCRE2 compiled to WASM, \~92 KB over the wire, and no roundtrips to the server. Probably has a few bugs I haven't found out ye…
受众观点:①受众核心需求是实用的regex debug能力 ②对WASM做浏览器工具的技术路径感兴趣
展开评论
- @Cheap-Commission-843 (2): regex can be tricky, hope this helps someone debug their patterns easier
一篇面向入门者的 PostgreSQL 多层缓存机制综合概述文章,旨在帮助开发者从整体视角理解查询缓存、共享缓冲区等缓存策略的层次结构和适用场景。
I haven't seen too many comprehensive approaches to this topic, so I took a swing at it. Goal isn't to be exhaustive, but to be comprehensive enough to provide a nice overview for people trying to ge…
受众观点:无有效评论数据,受众关注点无法从 evidence 推断
展开评论
- _无评论_
Typewise Nova 是一款面向企业的 AI 客服平台,可用自然语言描述客服策略后由 AI 自动搭建和持续优化客服团队,无需开发者参与。
Typewise
受众观点:①"15 分钟上线"之后还需要多少实际工作量才能真正自动化运转(@Buycott 直接追问)②企业级 2-4 周 vs 简单场景 1-2 小时的部署差异 ③从 Zendesk/Intercom 迁移的实际复杂度
展开评论
- @Overview (0): * [Launches7](/products/typewise/launches) * [Reviews2](/products/typewise/reviews) * [Alternatives](/products/typewise/alternatives) * [Built with](/products/typewise/built-with) * [Team](/products/typewise/makers) * [Awards](/products/typewise/awards) * More This is the 7th la…
- @Typewise (0): Maker Hi Product Hunt 👋 David here, co-founder of Typewise. Over a year ago, we launched our AI customer experience platform. Since then, we've implemented it for dozens of companies, from large enterprises to small tech firms. We've helped teams migrate from Zendesk and Interco…
- @Typewise (0): Maker [@dipanshu\_kushwaha5](https://www.producthunt.com/@dipanshu%5Fkushwaha5) basically the company (together with Nova during the onboarding) defines what knowledge and actions the AI shall have access to (e.g. read CRM, read/write Shopify orders, etc.), and then Nova builds…
- @Buycott (0): Congrats on a very cool (and valuable) product. The 15-minute setup sounds very cool, but the reason these tools need so much setup is because the real world is full of exceptions. And good solutons often depend on decisions that are not documented. So I'm curious after those fi…
- @Typewise (0): Maker [@scott\_kennedy](https://www.producthunt.com/@scott%5Fkennedy) fair point. this really depends on the size / complexity. even after the 15min, easier things such as informative requests, a shopping concierge ("what is the right bike for me" for a bike seller, for example)…
OpenObserve 是专为 AI agent 和 LLM 设计的 OpenTelemetry 原生可观测性平台,能追踪每次 agent 会话的具体耗时和成本,并可作为 ELK Stack 的低成本替代方案。
OpenObserve
受众观点:①AI agent 调用成本的可视化和追踪能力 ②作为 ELK Stack 替代方案的易用性和运维成本(@tedoc 明确说是 easy drop-in replacement)③UI/UX 打磨程度(多条 maker 回复承诺持续改进)
展开评论
- @Overview (0): * [Launches2](/products/openobserve#launches) * [Reviews6](/products/openobserve/reviews) * [Alternatives](/products/openobserve/alternatives) * [Customers](/products/openobserve/customers) * [Built with](/products/openobserve/built-with) * [Forum](/p/openobserve) * More This is…
- @OpenObserve (0): Thank you so much for the detailed review, Edson! Really appreciate your kind words. We are continuously working on improving the UI/UX . Upvote (3) Report Share 5h ago [](/@ashish%5Fkolhe2) [Ashish Kolhe](/@ashish%5Fkolhe2)
- @OpenObserve (0): Edson Thanks a lot for the review.Stay tuned for continuous UX improvements to OpenObserve. Upvote (3) Report Share 5h ago [](/@shohams) [Shani Shoham](/@shohams)
- @OpenObserve (0): Edson, thank you. We love working with you. Upvote (3) Report Share 4h ago [Ted O'Connor](/@tedoc) •[1 review](/@tedoc/reviews) #### What's great It was an easy drop-in replacement for our full ELK stack. It is so much easier to administer and requires a much smaller footprint w…
- @OpenObserve (0): [@tedoc](https://www.producthunt.com/@tedoc) Thanks a lot for the review ! Upvote (3) Report Share 5h ago [](/@shohams) [Shani Shoham](/@shohams)
iPhone Duo 是苹果首款折叠屏 iPhone,展开后屏幕比 iPhone 18 Pro Max 大 50%,是苹果有史以来最大的 iPhone 显示屏,同时具备外屏单手操作能力。
Apple
受众观点:①外屏是否真正能替代 iPad Mini 的单手轻量场景(@alex_gidirim 指出这比规格参数更重要)②折叠屏是否让 iPhone+iPad 二合一成为可能(@OpenObserve 提到以后只需一台设备)③iOS 生态对折叠屏应用适配的天然优势(@alex_gidirim 提到苹果多年关系积累)
展开评论
- @Overview (0): * [Launches311](/products/apple/launches) * [Reviews66](/products/apple/reviews) * [Alternatives](/products/apple/alternatives) * [Customers](/products/apple/customers) * [Built with](/products/apple/built-with) * [Forum](/p/apple) * More This is the 311th launch from Apple. [Vi…
- @Platforms (0): [Xcode](/products/xcode?ref=product%5Fsidebar) Develop, test, and distribute apps for all Apple platforms [5.0(210 reviews)](/products/xcode/reviews) [Code editors](/categories/code-editors)[Testing and QA software](/categories/testing-and-qa)
- @Osaurus (0): Hunter 📌 Steve did the Job so Tim could Cook, and now it's John's Turnus. Honestly this feels like a suitable replacement for the little [iPad Mini](https://www.apple.com/ipad-mini/) that _could_, which I loved so much. Upvote (12) Report Share 22h ago [](/@thisiskp%5F)
- @Netlify (0): [@chrismessina](https://www.producthunt.com/@chrismessina) I loved that ipad mini too Upvote Report Share 9h ago [](/@shivam%5Fkushwaha16) [Shivam Kushwaha](/@shivam%5Fkushwaha16)Likely AI
- @OpenObserve (0): [@chrismessina](https://www.producthunt.com/@chrismessina) Looks like I can trade in my iPhone and iPad for a new iPhone Duo. Upvote Report Share 7h ago [](/@alex%5Fgidirim) [Alex Gidirim](/@alex%5Fgidirim) Whether the outer screen actually replaces what people used the Mini for…
Subanana 是一款多语言转录工具,针对不同语言自动路由到效果最优的语音识别模型,可将视频转换为字幕、会议摘要和发布就绪文档,支持 YouTube 等链接直接输入。
Subanana
受众观点:①小语种(如亚美尼亚语)的实际转录准确率(@RunEvr 直接表示有此痛点并已注册)②是否支持 YouTube 及其他社交平台视频链接直接处理(@RunEvr 追问)③免费试用后的导出付费模式是否合理
展开评论
- @Transcription (0): Most transcription tools lock everything to one AI vendor - great for English, rough for everything else. Subanana routes each language to whichever speech model performs best on it - the same routing logic covers real spoken Cantonese, exactly the kind of speech single-vendor t…
- @Transcription (0): [Langfinity](/products/langfinity?ref=product%5Fsidebar) AI-powered real-time voice translation for 100+ languages [5.0(10 reviews)](/products/langfinity/reviews)
- @RunEvr (0): [@lakshya\_singh](https://www.producthunt.com/@lakshya%5Fsingh) [@aric\_fung](https://www.producthunt.com/@aric%5Ffung) [@connect\_kai](https://www.producthunt.com/@connect%5Fkai) [@kevin\_wong21](https://www.producthunt.com/@kevin%5Fwong21) Good launch, congrats! We have a cons…
- @Subanana (0): Maker
- @RunEvr (0): [@aric\_fung](https://www.producthunt.com/@aric%5Ffung) I’ve already signed up! =) And does it support links from YouTube or other platforms? Upvote (1) Report Share 11h ago [](/@aric%5Ffung) [Aric Fung](/@aric%5Ffung)
Suno v6 是 Suno 与音乐行业合作共同训练的新一代 AI 音乐生成模型,提供旗舰版、探索版和免费迷你版三档,支持局部歌词段落编辑和多曲混搭。
Suno
受众观点:①音乐行业合作模式对创作者版权保护的实际意义 ②三档模型(v6/v6-wild/v6-mini)的差异化定位和适用场景 ③局部编辑能力对迭代创作流程的改变
展开评论
- @Overview (0): * [Launches11](/products/suno/launches) * [Reviews19](/products/suno/reviews) * [Alternatives](/products/suno/alternatives) * [Customers](/products/suno/customers) * [Team](/products/suno/makers) * [Awards](/products/suno/awards) * More This is the 11th launch from Suno. [View m…