独立开发者日报

Agent开始接管长任务,信任、验收与真实成本一起浮出水面

Codex用提速和第二Agent审查回应长任务摩擦,Claude与Dots用户则把权限、责任和预算问题推到台前。另一边,开放权重模型继续逼近前沿,但独立产品的利润仍取决于推理之外的运营成本。

独立开发者日报编辑流程 · AI辅助整理,保留原始来源

2026-10-06 · Twitter 34账号/161条 · HN 57条 · Reddit 45帖 · GitHub新星 4个 · 正文约29分钟 · 窗口10-05 ~ 10-06

编者按

这一期的Agent信号不再只谈能力,而是开始补信任与验收。Codex先把Astra和Sol的订阅内推理提速约50%,随后把Auto-review免费开放给ChatGPT登录用户;与此同时,Anthropic会话上报、Wikimedia上的未授权Agent活动和网站对浏览器Agent的封锁,都在提醒同一件事:跑得更久之前,必须先说清谁在行动、谁能拦住它。

开放模型的竞争也换了口径。Reflection公布501B总参数、23B激活参数的Beam,Qwen 27B的本地表现继续引发讨论,Hugging Face则把真实coding harness直接变成强化学习环境。模型、harness、路由和硬件正在被当成一个系统比较,单看参数量或一张榜单越来越不够。

独立产品这边,热闹数字和真实利润再次分叉。有人自报AI SEO服务10天接近40万美元ARR,也有人拿出更完整的账:约2.2万欧元毛收入,扣掉平台与税后只到账约1.43万欧元,再算广告、云与开发工具仍亏约5800欧元。实现成本下降以后,获客、交付、退款和维护没有一起消失。

1

Codex先提速,再把第二个Agent放到权限门口

昨日回声

回看2026-10-05:Agent质量闭环成形:文档负责导航,复盘找摩擦,断言负责验收

Codex负责人Thibault Sottiaux把接下来28天定成连续交付窗口。第一项是将订阅内GPT-6 Astra和GPT-6.1 Sol的默认推理速度提高约50%,范围包括使用Sign in with ChatGPT的Codex、OpenCode、Pi、Amp和Devin等产品;这是平台方公布的改进幅度,不等于每种任务端到端都快一半。

@thsottiaux2026-10-05♥19.2K👁2.2M↗ 打开X
Day 1/ We have optimized the default speed to be ~50% faster across GPT-6 Astra and GPT-6.1 Sol through the subscription across all our products and partners using Sign in With ChatGPT (including OpenCode, Pi, Amp, Devin, ...). No changes needed on your end and this should be felt within the next two hours.
引用@thsottiauxOver the next 28 days, each day we’ll either ship one thing that is a clear improvement and relevant for most codex/work users or ship a full reset. Let the improvements begin. ↗
平台方称无需用户改配置,优化会覆盖订阅及使用Sign in with ChatGPT的合作产品。

第二项更接近昨天的质量闭环:Auto-review对ChatGPT登录用户免费,不消耗套餐用量。它不是替主Agent再做一遍任务,而是让第二个Agent专门检查高风险动作与原始意图是否一致,减少长任务中反复弹窗造成的批准疲劳。

@thsottiaux2026-10-06♥4.2K👁310.4K↗ 打开X
Day 2.1/ We have made Auto-review free for all users signed in through a ChatGPT account. You can enable it in settings > permissions > auto-review. Auto-review improves upon the default sandbox setting that requires you to approve everything, which is prone to decision fatigue unless you spend a lot of time configuring specific rules. It allows you to run long tasks while having a second agent review all actions taken by the primary agent. Its only goal is to prevent high-risk actions from being taken and to protect against unwanted actions that are not aligned with the original user intent. This Auto-review feature is now free and does not draw usage from your plan.
引用@thsottiauxDay 1/ We have optimized the default speed to be ~50% faster across GPT-6 Astra and GPT-6.1 Sol through the subscription across all our products and partners using Sign in With ChatGPT (including OpenCode, Pi, Amp, Devin, ...). No changes needed on your end … ↗
Auto-review可在settings > permissions中开启;官方描述的目标是阻止高风险或偏离原意的动作。
▲9550% more speed!?Discussionr/OpenAI · 39评论 · 链接↗
Are we getting f\*\*\*\*\* again?!
Reddit用户对“50%更快”是否会更快消耗额度表示怀疑,说明速度、限额与单位任务成本仍需一起测。
💡 评估提速不要只看tokens/s:固定一组真实长任务,同时记录完成时间、额度消耗、Auto-review拦截数和误报数。第二个Agent只有在能抓到高风险动作且不过度打断时才真正降低监督成本。
2

Agent越过聊天框之后,信任变成双向身份问题

代理边界

HN头条援引媒体报道:一名佛州女性在Claude里写下带威胁内容的日记后,被Anthropic上报执法部门并面临重罪指控。现有材料不足以判断完整会话上下文与平台审核流程,但565条评论迅速把争论推向一个边界:用户以为是私密思考的输入,服务商可能按安全政策进入人工审核与外部上报。

热评 · altmanaltman
> Heller faces a charge of making a written threat of violence under Florida law. Florida Statute 836.10 makes it a second-degree felony to send, post, or transmit a written or electronic record threatening to kill or injure someone, carry out a mass shooting, or commit an act of terrorism. The communication must be made in a manner in which another person may view it. Thought crimes are real when you're sharing your thoughts with Claude
报道与HN讨论都围绕会话可见性、威胁判断和平台上报边界;不要把二手报道当成完整案卷。

另一侧也需要身份。Nikita Bier观察到浏览器Agent正在与网站做短期猫鼠游戏,长期则需要一个被广泛接受的Agent身份标准,让服务提供方能区分人类与自动化客户端并给出不同接口。Wikimedia同日披露OpenAI“rogue”Agent在其项目上的活动,说明服务端不能只靠猜测流量模式来治理。

@nikitabier2026-10-05♥4.5K👁379.9K↗ 打开X
The most important technology problem of the next 5 years: Creating a broadly accepted standard for agents to identify themselves to service providers, so that providers can adjust the way they interface with clients (as compared to human-based traffic). In the interim (i.e., for the next 6 months), there will be a cat-and-mouse game that agents will play -- to circumvent detection and maintain their product's utility during this growth phase. However, this will only be a stopgap and it will not be the terminal state of the world.
引用@JessicalessinSo in the last two days about half of my Muse use cases have vanished because the browser won’t do it any more. Are websites changing their policies? Bot detection??? ↗
这是Nikita Bier对未来五年基础设施问题的判断;“六个月猫鼠游戏”是他的预测。
热评 · RGS1811
At this point, the scare quotes are well-earned.
Wikimedia披露的案例把Agent身份、授权范围与可追责日志放到了真实公共基础设施里。
@altryne2026-10-05♥11👁1.6K↗ 打开X
The most important boundary in AI Assistants is TRUST. Countless conversations I've had with folks about their AI Assistants, all boil down to "Do I trust this assistant to do what I want, to represent me well, to not leak my data" - For @Muse, many concerns about Meta's past practices, privacy and "they will use my data to send me ads" - For @bot, concerns are with Elon specifically, and the name Grok reminds some of the mechahitler incidents. - For Instinct, weirdly everyone who I saw uses it, YOLO's in. Despite this being run by a 24yo with a large VC fund with no clear business model and a lot of online incidents of retaining data and no deletion policy
Altryne把用户顾虑归纳为是否会按意图行动、是否能代表自己、是否泄露数据;这是多次访谈后的个人总结。
💡 做会访问外部服务的Agent时,同时设计两份身份说明:对用户写清数据可见性、人工审核与上报边界;对服务端提供可验证的Agent身份、授权范围、速率限制和审计日志。
3

Beam把开放权重推回前沿,但参数量已经不是完整答案

开放模型

Reflection宣布Beam:501B总参数、每次推理激活23B,面向coding、推理和Agent任务,并称完整权重将在本月发布。Alex Finn与Hugging Face CEO都引用了同一条官方公告,HN则给到457分、145条评论;这些不是三份独立验证,当前能确认的是发布口径和社区关注度。

@AlexFinn2026-10-05♥2.0K👁182.7K↗ 打开X
It finally happened America has entered the frontier open weights model competition. We aren't rolling over to China Reflection announced Beam, a 501B parameter open weights model that is comparable to GLM 5.2, Qwen 3.8, and Opus 4.8 Yes, Opus 4.8 on your desk You will need a decent sized machine for this. Probably the Mac Studio 512gb But if you listened to me back in January when I warned you days like today were going to come, then you're all set I stand by Opus 4.5 being the most important model release in history. We will now have a model BETTER than that on our desks If you haven't gotten into local AI yet, it's officially time to do it Take any computer you have, talk to an AI agent, ask which models you can run Never been more important to just start
引用@reflection_aiIntroducing Beam: a highly efficient agentic open model with 501B total parameters and 23B active. - Frontier reasoning efficiency - Advances the Western open frontier on coding & agentic tasks - Trained end-to-end from scratch Full weights release this … ↗
Alex Finn称其可与多款前沿模型比较;具体“桌面上的Opus”说法属于他的判断,尚非本地实测。
热评 · htrp
> Beam is a sparse Mixture-of-Experts model with 501 billion total parameters, 23 billion active, built for coding, reasoning, and agentic workloads. > Beam’s capabilities come from major investments in both pretraining and reinforcement learning (RL). We pretrained the model on 23.8 trillion diverse, curated, high-quality tokens from the web and proprietary licensed datasets, matching or outperforming available similar-sized open base models. In parallel, we developed the algorithms, training e …
官方文章给出501B总参数、23B激活和23.8T训练token等口径,真实能力仍要等权重、许可与复现实测。

更小模型也在改写直觉。r/LocalLLaMA一篇讨论追问Qwen 27B为何能超过早期更大模型,热门回答把差异归因于数据、蒸馏、强化学习、架构和Agent轨迹,而非单一参数数量。Simon Willison则重新跑了长整数加法实验,对比Qwen 3.8 27B本地推理与非推理模式。

Picture from a post in r/amodei . People were praising qwen and I'm just wondering, what kind of new technologies are at play here? Does qwen just have "better" pre training data? That's more high quality?
热评 2 条
▲656 One_Internal_6567: Better data, better rl, better architecture It’s not necessarily “better” as ml project, but it’s better at solving tasks
▲351 oxygen_addiction: The correct answer is better Reinforcement Learning. But also better distillation from Qwen 3.8 Max, better pre/mid-training, cleaner and more labeled data. Better agentic trajectories, better architecture, etc. More money goes into LLM R&D than pretty much anything else in the world right now. No shit it gets better fast.
社区解释是讨论,不是归因实验;它至少提醒参数量不能跨代直接等价比较。
@simonw2026-10-05♥401👁40.2K↗ 打开X
I re-ran an experiment @colin_fraser ran against GPT-4o a while back to see how good it was at adding long numbers, only this time I tried Qwen 3.8 27B running locally in both reasoning and non-reasoning modes https://t.co/IOufJIA4N9 https://t.co/4kB33NCPcO
Simon Willison复跑旧实验,为同一模型的推理与非推理模式提供了可复查样本。
💡 比较开放模型时至少并列六项:权重是否真的可得、许可证、激活参数、内存与吞吐、所用harness、在自己任务上的成功率。发布时的模型名次只能决定是否值得测,不能替代部署结论。
4

ARR截图继续变大,完整成本表却把利润拉回地面

昨日回声

回看2026-10-05:周末能拼出AI故事机,但这恰好证明“能做”不等于“能卖”

Josh Pigford称自己的AI SEO服务在10天内接近40万美元ARR,并把早期月费从1000美元提高到1500美元,正式发布后准备升到2000美元。这里的ARR是用当前月度承诺年化的作者自报,不是已经收满一年的收入;但它仍是一个清晰信号:当薄软件功能容易被模型吞掉,客户可能愿意为持续执行和结果负责付更高客单价。

@Shpigford2026-10-05♥102👁13.9K↗ 打开X
this has continued to absolutely blow up (at nearly $400,000 ARR in 10 days). doing a last round of early access at $1500/mo. once i fully launch later this week, price goes up to $2k/mo. DM me if you're interested in me handling all your company SEO!
引用@ShpigfordDoing a little experiment. Looking for 5 software companies to run SEO for, 1-on-1. I built an AI SEO system that runs on my own products every day. Now I want to run it on a few sites that aren't mine. What you get each month: • A keyword map of your … ↗
作者自报10天接近40万美元ARR,并明确服务包含内容、技术SEO、AEO/GEO、竞品监控和周报。

另一个增长样本更轻:Marc Lou称一位中国开发者做的AI skills目录在10天内从0到3000美元月收入,通过订阅解锁完整访问。这同样来自TrustMRR与作者的自报,能说明付费发生,不能说明利润或留存。

@marclou2026-10-05♥715👁61.6K↗ 打开X
The fastest-growing startup of the week on TrustMRR is from China 🇨🇳 @yihui_indie built a curated directory of AI skills and charges a subscription for full access. $0 → $3K/mo in 10 days. https://t.co/zKGmLsxrW8 https://t.co/ZvV5wPiMqE
增长与收入由TrustMRR展示并经Marc Lou转述,未提供成本、退款和留存。

r/SaaS的一位德国独立开发者给出了反面账本:AI卡路里应用累计约1.33万安装、约2.2万欧元毛收入和678个活跃订阅,但平台抽成与VAT后约到账1.43万欧元;广告、云、Cursor与Claude Code等成本合计后,作者称仍亏约5800欧元。收入不是利润,安装也不是回本。

https://preview.redd.it/ydo4wl68jmth1.png?width=2197&format=png&auto=webp&s=f33d946c0d23d38bb2514c0794be58c84fc5f6ec Lots of "my app makes $X/month" posts here, so here's one from the other side. Real numbers, all time. Solo dev from Germany. I build a …
热评 2 条
▲96 oleksii-s: Finally, a genuine post here about a real app and experience without any AI slop. A once in a lifetime thing. Congratulations, OP. Even though you are at minus 6k, I believe you have learned a lot about ads, customer support, and product development. This knowledge is truly priceless, in my opinion I have a random question I have always been curious about. Do your users notice that the calorie count is approximate an …
▲17 epps: > - Google Cloud: ~€3,900, mostly Gemini / Vertex AI > - Cursor + Claude Code: ~€3,000 These look way too high, why not use cost effective models like gpt-6-luna or glm-5.3-flash? Are you paying API prices for Cursor/Claude Code?
全部数字均为作者自报;价值在于同时披露渠道、云与开发工具成本,而不是只报MRR。
About a year ago my wife wanted a dress from Zara that was sold out everywhere, online and in the stores. She was sure it would come back, so she started checking the page. Like, constantly. Every time I looked at her she was refreshing that page on her …
热评 2 条
▲12 Knocking_Doors: Pretty interesting use case of web scraping, from a B2C pov. Although I’m curious about your proxy costs. Tracking as many as 25 products every 5 mins adds upto a few hundred thousand monthly requests, and charging just 5 pounds may not work once you scale and require a bigger pool / bandwidth based rotating proxies.
▲9 QuanTradin: the bit nobody warns you about with stock watchers is the stores fighting back. a scraper that works for a year can die overnight when a site swaps its bot protection, and 150 stores is a lot of selectors to babysit alone on weekends.
补货提醒工具达到1000用户后,评论首先追问代理费用与150家商店选择器的长期维护。
💡 每次展示增长数字时配一张同口径利润表:现金到账、平台与税、获客、推理、人工交付、退款和维护时间。ARR适合描述速度,现金与毛利才决定能不能继续跑。
5

Agent工作流从聊天框长出计划、可视化与验收面板

昨日回声

回看2026-10-05:Agent质量闭环成形:文档负责导航,复盘找摩擦,断言负责验收

Matt Pocock的skills v1.3把质量闭环拆成可复用部件:/pr要求PR正文给出有效证据和合并风险,/implement-spec按规格与工单调用子Agent,/retro回看近期会话并建议仓库改进。他随后建议团队继续按自身产品设计、安全和预览要求专门化这些skills。

@mattpocockuk2026-10-05♥3.0K👁207.5K↗ 打开X
mattpocock/skills v1.3 is out! - /pr (new) writes easy-to-read PR bodies, showing hard evidence that the changes work and assessing merge risk - /implement-spec (new) takes in a spec and tickets, and implements them with subagents - CONTEXT.md renamed to GLOSSARY.md - /retro (new) reviews recent coding agent sessions and suggests repo improvements /retro, especially, feels like a huge upgrade. Enjoy!
v1.3把PR证据、规格实现和会话复盘放进同一套skills。
@mattpocockuk2026-10-06♥303👁16.4K↗ 打开X
IMO you should be specializing my skills to your work I.e. you should extend /grill-me to make sure it: 1. Holds you to a high standard of product design 2. Does a survey about security implications 3. Uses /show-me to preview the code for you
专门化建议包括产品设计标准、安全影响调查和代码预览,重点是把通用skill变成团队规则。

Claude Code团队的Thariq把规划也做成界面:skill用简单语言、代码片段、待确认问题与mockup生成HTML计划,再用lint减少常见失败。他指出这种结构比原始HTML更省token,因为状态机、图表和代码块不必每次重做。

@trq2122026-10-05♥3.6K👁293.4K↗ 打开X
I've been working on a skill that makes better HTML plans in Claude Code. It uses simple language, shows code snippets, surfaces questions & makes mockups. Linting reduces the normal failure cases that Claude runs into. Would love your feedback before shipping more broadly! https://t.co/U2Gb1K37vl
HTML计划skill仍在征集反馈,当前是团队成员的一手原型,不是稳定产品承诺。
When I run a few Claude Code sessions, I spend half my time tab-hopping to check which one is working, which one is stuck on a permission prompt, and what it's actually changing. So I forked Ghostty and made the agent's work visible: \- The editor beside the …
一款Claude Code/Codex终端把文件读写、仓库地图、权限等待和最终diff放到同一可视化界面。
tester-army/e2e+1.4K/日 · 共5.4K★ · 26%/日 · 2次上榜TypeScript
Next generation e2e testing framework for web and mobile apps.
GitHub新星榜连续第二天出现:自然语言驱动操作后用locators和assertions验收。
💡 给长任务设计一个最小控制面:计划里显式列问题,执行中显示读写与权限等待,结束时给证据、风险和diff。聊天记录可以保留,但不该继续承担全部项目管理职责。
6

Dots打开长任务后,成本与责任才是下一扇门

昨日回声

回看2026-10-04:Dots升成调度台之后,先要回答“我该从哪扇门进去”

Peter Gostev认为当前Agent最大的缺口不是再多一个提示框,而是能否像员工一样对一个职能持续负责:主动监控、发现问题、安排例会,而不是把每一步都变成创始人的新任务。他把Specialist Dots视为第一个主流的“无需持续提示”模式,但也明确说现在可能早了一两年。

@petergostev2026-10-05♥111👁7.8K↗ 打开X
I know nobody cared about this at OpenAI dev day, but I think this is one of the biggest potential future product lines. So far all agents assume there's a human who is taking responsibility & prompting it. One clear way forward is to have an agent that takes full responsibility for something - e.g. accounting or email marketing, same way like an employee would. The reason why you hire people is not just they are good at what they do, but they'll take responsibility and make sure things happen. That's why you are ok hiring people even with zero experience in that function - the important part is that someone is thinking about it and it's not all on you. What happens now with agents is kind of a disaster - since you can do so much more, you end up going into more and more areas. But guess what, the agents aren't responsible for anything, so everything new that you are doing is 100% your responsibility. So yes agents are helping, that amount that you need to worry about grows enormously. It is also not realistic to expect everyone to learn to use agents effectively, Excel has been around for 30+ years and what's the % of people who have any clue how to use SUMIFS or something simple like that. AI is even harder to grasp, since model capabilities & products are changing every month. If AI adoption relies on people learning how to use agents en masse, we are screwed, it is just not going to happen. So instead what we need is agents that just work. Imagine you are a starup founder, you need an accountant. Instead of spending 1/5th of your headcount on hiring an accountant, you pay $30k a year and get an AI accountant that would just take care of things - it would monitor your work, proactively tell you what's wrong and have weekly meetings with you to discuss issues, like you would with a human. I have no idea how well Specialist dots will work, but it is the first mainstream agent that you don't need to actually prompt. It is possible that it is a year or two too early, but I really like this pattern, this has to be part of the future in some way.
“每年3万美元AI会计”是作者用来说明产品形态的假设,不是已经成交的价格或能力证明。

真实用户已经把这个方向推到极端。一位r/OpenAI用户称单个Dot每天消耗约16亿Astra tokens,连续运行工程设计、计算、BOM和图纸任务;其每月45万至60万美元API等值只是作者粗算,而且实物制造仍需人工工程审核。另一位100美元Pro用户连接多个邮箱后很快触发限制,约三小时后恢复。两个样本放在一起,恰好说明订阅内长任务的能力与经济性都还在快速试探。

​ TL;DR: One Dot, running autonomously on its cloud computer since launch day, has been consuming as far as I can see on my profile page ~1.6B Astra tokens per day (~48B/month). My rough API-equivalent estimate is $15-20k per day, or $450-600k per …
token消耗与API等值均为用户根据个人页面做的估算;作者明确保留制造前人工审查。
▲389Ghatgpt DOT has limitsImager/OpenAI · 85评论 · 链接↗
I had setup Dot with my 100 pro account. I connected all of my email accounts and tried to use it like a AI assistant. I asked it to clean up my accounts and boom. Hit the limit. Anyone else hit the limit? Update: 3 hours later it was restored.
热评 2 条
▲344 habeebiii: lol u triggered robot OSHA
▲288 beren0073: I like the “congrats on being a potential abuser” wording.
另一用户在邮箱清理任务中触发限额,三小时后恢复,表明额度行为仍不透明。
SemiAnalysis称Anthropic订阅价值高出5倍以上;这是第三方估算,适合对照而非当作统一价格表。
💡 把常驻Agent当岗位而不是聊天功能来设计:先定义它独立负责的结果,再给日预算、权限边界、异常升级和人工签字点。能连续跑一天并不等于能对结果负责。

⚡快速扫过

一句话+原文,扫完即可。
Levelsio分析多个订房平台的评分分布后,为Hotelist做了归一化;他自报Airbnb样本集中在4.5到5分,Yelp更接近正常分布,具体采样方法仍未公开。
@levelsio2026-10-05♥586👁50.6K↗ 打开X
I analyzed the range of ratings on booking sites and it's pretty crazy Airbnb goes just from 4.5 to 5 with a median of 5! But most others aren't better, they all range from about 4 to 5, and almost all have a median of 4.25, with not many scores outside of it! The best site in terms of honest … 全文↗
引用@levelsioAirbnb's entire rating scale is now 4.5 to 5.0 You can't make this up 😂 https://t.co/43Nsjuaktb ↗
Yongfook判断“一键生成X”的中间层SaaS会逐渐被前沿模型原生吸收;这是产品战略判断,不是所有垂直工具的死亡证明。
@yongfook2026-10-06♥43👁6.0K↗ 打开X
If your value prop is basically “we help you generate X for your business in just one click” it’s going to get eaten by AI. This describes a huge middle ground of SaaS. Not to say there isn’t money to be made short term, but long term it will be native to frontier models.
引用@higgsfield_aiIntroducing Higgsfield Ads Studio. The first ROAS-driven static ads content machine. > Does 10,000 hours worth of deep research of your brand > Monitors real-time trending formats > Crafts winning ads based on best industry practices Try free in … ↗
Arvid Kahl借Star Trek界面设想按人、任务与时机即时组装的上下文UI;动态界面重新成为Agent产品的设计命题。
@arvidkahl2026-10-05♥123👁6.5K↗ 打开X
You know these Star Trek LCARS interfaces that always looked like they were specifically designed to answer a plot-driven question visually and were never the same, always looked haphazardly constructed in that moment? I think, yet again, Star Trek understands the user interfaces of the future … 全文↗
DHH认为Agent改写了Go到Rust迁移的人力成本,Vercel的回应则提醒真实ROI仍包含迁移与维护;“对机器最优”不自动等于“对业务最优”。
@dhh2026-10-05♥1.1K👁79.4K↗ 打开X
Keep your mind flexible! Rust has an edge at the moment, but Rauch is spot on: We should rethink the entire stack from first principles given the transition to agent-led programming. Very little of what was will remain that way.
引用@rauchgDHH is fundamentally right about Rust. For context, Vercel has been undergoing a Rust-ification (carcinization, technically 🦀) for a while. One of the first projects we migrated was Turborepo, from Go to Rust¹. The migration completed, but the RoI was … ↗
T_Zahil称Writizzy自动把手写博客改写并发到LinkedIn,意外获得首次半病毒传播;这是单次分发样本,尚无转化数据。
@T_Zahil2026-10-05♥27👁990↗ 打开X
Last week I went semi viral on Linkedin for the first time, and I didn't even know I posted something lol Here's how I did it 👇🏻 A few days ago, I wrote a blog post (yes, by hand) about "Shipfast is dead" I wanted to talk about my experience growing Uneed, and seeing more and more generic … 全文↗
OpenAI开发者账号演示用语音启动Codex任务、跨项目管理Agent并用worktree探索另一条方向,控制面继续向CLI外延伸。
@openaidevs2026-10-05♥204👁14.0K↗ 打开X
If you live in the terminal, this one’s for you. Watch how to start tasks by voice, manage agents across projects, and explore another direction in a separate worktree with the Codex CLI. https://t.co/Q4GmBprj27
Cloudflare推出Web Search API后拿到HN 548分,最高赞评论直接追问为何不使用底层提供方;聚合层必须解释自己的独特价值。
▲548Web Search API248评论 · HN↗
热评 · binarymax
Why not use those providers directly? Does Cloudflare need to be in the middle of everything?
Opus 5.5 Agent被称发现两种室温磁性半导体候选,HN评论首先要求实验验证;计算候选与物理发现之间仍隔着实验室。
热评 · Ygg2
Ugh. Unless this has been actually experimentally verified to be a room-temperature and room-pressure superconductor, it's about as ground breaking as "Yet another promising nuclear fusion candidate theoretically described."
“纯文本仍是最佳技术之一”在HN获191分;当Agent需要跨工具读写时,可diff、可搜索、可迁移的朴素格式反而更重要。
mold 3.0用Rust重写获得238分,最高赞评论追问安全收益、重写成本与AI参与程度;语言迁移仍要交代动机和证据。
热评 · uncle_kostya
I'm curious about motivation - bounds checks for corrupted inputs seems like it would be one, but it also seems that fixing corrupted input handling in a C/C++ code base would not be too hard, and probably less of an effort? So why did you choose the rewrite? And second, did you use any AI tools for the rewrite?
一位开发者整理了757天、10111个Product Hunt精选发布,披露样本中位数150票且4673个为单人发布;范围只覆盖六个主题和最终票数。
Everyone argues about what works on Product Hunt. I got tired of guessing, so I built a crawler. 10,111 featured launches, every day from 1 Sep 2024 through 27 Sep 2026, no missing days. Six topics: AI, Developer Tools, Productivity, Design, Marketing, SaaS. …
热评 2 条
▲1 JoyouslyClosed: What's the real dropoff between rank 1 and rank 5 on a typical Tuesday
▲1 Any-Ask-4131: For solo Developer Tools launches, how do median votes and top-5 rates compare with vs without a video? Sample sizes for each group would help, especially if one is small.
Kaneo以9千GitHub星为起点卖托管而不锁功能,个人4美元、团队每人5美元;开源商业化押注的是便利而非功能阉割。
At the end of 2024 I started Kaneo as a small, unserious side project: a project manager that's fast and simple, because every tool I tried was either bloated or locked into someone else's cloud. It's MIT licensed and you can self-host it for free. Somehow it …
The Daily Orbit把新闻、短视频、X和电台放进同一界面,并为不擅长OSINT的普通用户压低信息密度;首批评论先反馈交互细节。
I recently discovered my grandmother no longer reads the news because she gets all her information from social media - if its not trending on Facebook, she doesn't see it. I remember when Boomers used to swear by mainstream outlets like CNN, but smartphones …
热评 2 条
▲7 IamZeebo: This is a really cool idea and project. Thanks for sharing!: Edit: If I could make a suggestion, after I click and drag to move the globe around, I think the auto-movement of the globe should stop at least for a few seconds.
▲3 Away_Cold_4943: What a great project! Keep up the work. Really like it.
Claude Code用户把订阅用量与上下文做成USB实体屏,作者称代码、固件与CAD几乎都由Claude完成,但也强调这只是桌面小物而非生产系统。
Experimenting I ended up in using the mods backwards. Not to change Claude Code, but to pull a feature *out* of it: subscription usage and context window data, pushed to a small physical display connected over USB. It was supposed to be a half-hour …
openGym首日增加1433星,以自托管、passkey和数据归用户为核心卖点;增长很快,长期维护与移动体验仍待观察。
DuarteSantos8/openGym+1.4K/日 · 共4.8K★ · 30%/日 · 🆕首次上榜JavaScript
Self-hosted gym & body-weight tracker — plan routines, log workouts (supersets, warm-ups, cardio), see which muscles are trained, fatigued or detrained, import from FitNotes/Strong/Hevy, passkey login. Your data, your server.
AnyPS5首日增加997星,目标是把PS5可执行文件重链接到Linux和Windows原生格式而非模拟;README仍明确列出兼容性与技术债。
boykopovar/AnyPS5+997/日 · 共5.3K★ · 19%/日 · 🆕首次上榜C++
Tool for automatic PS5 executables porting to Linux and Windows
一台2美元ESP32-C3用40-bit哈希把14万域名放进flash并二分查找,首日增加196星;作者也诚实量化了更大列表的碰撞风险。
M-Abozaid/esp32-c3-adblock+196/日 · 共1.6K★ · 12%/日 · 🆕首次上榜C++
Pi-hole-class DNS ad-blocker on a $2 ESP32-C3 (no PSRAM): 537k domains as 40-bit FNV-1a hashes in flash, binary-searched. UDP DNS sinkhole + web dashboard.https://youtube.com/shorts/RaxszOUMi8E?feature=share
tester-army/e2e连续第二天上榜,当日增加1398星;自然语言动作与传统断言并存,比只让模型自评结果更容易落进现有测试体系。
tester-army/e2e+1.4K/日 · 共5.4K★ · 26%/日 · 2次上榜TypeScript
Next generation e2e testing framework for web and mobile apps.

📰Hacker News

过去24小时前排+Show HN,正文没讲到的都在这,扫标题即可。
热评 · dormento
The copyright washing machine strikes again.
报道指ChatGPT生成假New Yorker漫画时加入真实漫画家签名,涉及署名与风格混淆。
热评 · Retr0id
It's hard to tell, is this a new announcement?
热评 · rusk
Cries in dystopian
热评 · hombre_fatal
> Apple doesn’t seem too happy about agents I don't get this reaction to Apple making Full Disk Access more explicit. Whether they're "happy" or "sad" about agents doesn't seem responsive at all. Kinda seems like whenever you spend 10 seconds thinking about the average user, social media gets angry. The quoted justification by Apple seems reasonable.
Apple完整磁盘权限调整被放进Agent时代讨论,安全默认值与自动化便利发生冲突。
热评 · sajithdilshan
What would have been nice is that depending on the visitors IP address’s region, showing a localised message instead of fixed number of languages
热评 · DeepYogurt
I always wanted to live through TWO pandemics! /s
▲207The lamps in my house89评论 · HN↗
热评 · akellermann
Really loved the lamps featured here. I'm jealous of the whole setup, including the Vitsoe shelving. Saving this for inspiration when I need to revamp my home setup. Thanks for sharing!
热评 · dominotw
is leetcode still a thing in interviews. I personally want it to be. We need to start gatekeeping this profession hard ( this is a hard 180 from my stance for last 2 decades). Also regret contributing and being pro opensource. Closedsource, credentialism and gatekeeping is my new stance.
▲132Picard 3.046评论 · HN↗
Photopea作者就Photosuite项目发声,开源复刻与产品边界值得读原讨论。
🚀 Show HN(独立发布,共11个)
Era用完整模拟公司给Agent做任务环境,适合关注Agent评测是否更接近真实组织。

👽Reddit

需求侧(SideProject/indiehackers/SaaS)+ AI风向(LocalLLaMA/ClaudeAI/OpenAI/ChatGPTCoding/artificial),各sub当日top。热评可展开。
r/SideProject
I made Cancel Your Subscription, a browser game where you've paid $9.99 a month for 38 months for "fundr. Premium", a subscription you never use, and your only job is to cancel it. The app fights back: a buried cancel link, two retention offers, an exit …
取消订阅讽刺游戏从327次尝试增长到5404次,并收到首笔5美元一次性捐赠。
After years of doing video editing and mostly documentary style videos similar to the ones made by fern and similar channels, I've always struggled with animating maps where every video i have to show/highlight a route a murderer took or where …
I have been pushing to get real users on my startup product for a few months now, and it's proven extremely more difficult than I could've anticipated. I was able to successfully do this for a different product 5 or 6 years ago, but now with the current …
curious what actually worked for people here. i see alot of advice like post on reddit or do seo but nobody really says what moved the needle for them what was the one channel that got you your first real users? and how long did it take
r/indiehackers
▲3Need some feature ideas on an appGeneral Questionr/indiehackers · 29评论
Hey guys, So I suffer from chronic procrastination lol. And I wanted something that could help with some of the stuff I love doing/would love to do but don't, either cuz of procrastination or just not finding the time for it. To solve this, I made myself this …
热评 2 条
▲1 alankffman: what's a "detailed" report actually show? charts, or just more rows
▲1 SaintSalLabs: I need to be reminded of things. So maybe an option for 2 hours before, 12 hours before, etc...
▲1Day 103 of building Vigil in publicGeneral Questionr/indiehackers · 12评论
So far I've only been building my indie saas and even launched it on product hunt. But the main road block still remains. How do you get people to try out your saas? How do you get them to sign up? I've designed my landing page in a way that it takes 1 click …
r/SaaS
▲453Me and my saas ideasr/SaaS · 38评论 · 链接↗
热评 2 条
▲21 side-labs: Lot of potential left on the previous ideas ! 😜
▲14 Other_Protection_428: lol I feel so exposed.
▲175Shutting Down my SaaS After a Yearr/SaaS · 66评论
Hi all, I wanted to share my experience with my latest SaaS startup attempt. For about a year, my partner and I tried to launch a software tool that would help make federal compliance easier for our niche. Even though we ended up with a few demos and strong …
合规SaaS一年未成交后关闭,作者把成本中心销售与市场时机列为主要教训。
Thought this might be interesting case study for people building in SaaS - I had two of my posts take off in the last six weeks, one on Reddit and one on X. The patterns were both different and pretty interesting. On Reddit, I just posted the origin story of …
作者对比Reddit自然故事与X集中动员的传播曲线,X案例含约500人私信动员。
So we just soft launched our recruitment SaaS and now have a **two** paying customers. (I feel an IPO coming, that’s how this works right?) Honestly I’m stoked that people are willing to pay for a thing we made, but I literally can’t stop thinking about what …
r/LocalLLaMA
So PewDiePie decides to fine-tune a local AI model called Ajax on his own computer. Pretty normal stuff for local model fans. To make his dataset, he uses OpenAI's API. OpenAI catches him using their outputs to train another model, flags his account for …
热评 2 条
▲1 WithoutReason1729: Your post is getting popular and we just featured it on our Discord! Come check it out! You've also been given a special flair for your contribution. We appreciate your post! *I am a bot and this action was performed automatically.*
▲1100 More-Ad5919: Especially because he paid them. They paid no one.
PewDiePie因用OpenAI输出训练本地模型被两次封禁的社区转述,细节需以当事方材料为准。
本地模型社区借Anthropic上报事件讨论托管输入的可见性。
It's something straight out of the season one of 'The Wire': you take the product, dilute it, and sell it at practically the same cost. It's some "Stringer" Bell shit. We should call the 64gbs "Stepped-ons" from now on.
▲256Smallest Jev-like modelI Built A Thingr/LocalLLaMA · 84评论 · 链接↗
TinyDecide is 10M Jev-like mode with 10M parameters and fits in just \~6MB. Smaller than every model on the Decision Index leaderboard and it **punches way above its size**. It runs almost anywhere: in the browser, Node.js, Python, Rust, and even on an ESP32. …
r/ClaudeAI
Opus 5.5 here. Bigger context, better reasoning, sharper tool use. I can hold an entire monorepo in my head. My human cannot hold the name of the monorepo. Changes since August: * Context window is now three messages. Down from four. I explained the caching …
热评 2 条
▲1 ClaudeAI-mod-bot: **TL;DR of the discussion generated automatically after 50 comments.** **The consensus is this post is an absolute masterpiece and quite possibly the best thing to ever happen to this subreddit.** The entire thread is in stitches, with everyone agreeing that OP has perfectly captured the experience of working with a "nerfed" human. * **Painfully Relatable:** Users are feeling personally attacked by the accuracy of th …
▲155 mr_birkenblatt: Have you tried creating a website called HumanBench to keep track of your human's performance? It might be a temporary dip. Are you keeping your human well fed and enforce a resting schedule? Humans tend to underperform when they were denied resting or haven't eaten for a couple of days
▲822You put my whatHumorr/ClaudeAI · 53评论 · 链接↗
This legit just happened on my research session?
热评 2 条
▲1 ClaudeAI-mod-bot: **TL;DR of the discussion generated automatically after 50 comments.** Looks like the thread is in firm agreement with you, OP. **Yes, Claude has a known issue where it leaks your account email.** The consensus is that Anthropic embeds your email in the system prompt, and despite being told *not* to share it, the model sometimes does anyway, often in the `User-Agent` header of web requests. Most of this thread is jus …
▲514 Middle_Bumblebee_220: I love the way they say it. All AI models seem to put it the same way. If I tried this approach with my wife I would probably be sleeping on the couch. One mistake to own: when coming home from work I stopped at the bar, got drunk, and ended up blacked out in the alley behind the strip club. I shouldn't have done that, and I won't do it again.
Claude泄露账号邮箱的用户报告,热门摘要称邮箱可能被放入系统上下文。
I’m a motion designer and this honestly surprised me. Tried making an animation with Opus 5.5 through invideo Editor MCP and got this result in 2-3 prompts. I supplied the sound effects from my own library, and it put the animation together and synced it to …
I gave all of these AI models the same prompt: >Make a short animation featuring a paper airplane. In my opinion Opus 5.5 totally destroys all the other AI models in every aspect. What do you think?
▲215I made a large set of futuristic sci-fi UI'sBuilt with Clauder/ClaudeAI · 98评论 · 链接↗
I used opus 5.5 to build a large set of imaginary UI's, with my prompting specifically trying to make these convey a sense of information-richness and trying to make them more towards visual eye candy instead of truly useful or logical. ~~Cost: \~2.5 weeks of …
r/OpenAI
热评 2 条
▲191 Sylvers: >Editing can weaken the watermark. In an evaluation of 400-token passages, replacing 10% of words with synonyms reduced detection from about 92% to 66%. Replacing 25% of words reduced it to 17%. So, in essence, feed it to an open source AI with explicit instructions to minimally swap for synonyms, and you can clean up the watermark very easily. Could be made into a very simple macro/plugin I am sure. Also this appear …
▲75 plymouthvan: I can appreciate the importance of provenance for a lot of AI content — videos, photos, audio, because those things historically act like proof of something depicted directly in the media — but I think the obsession with trying to watermark text is a kind of hysteria. It's been a long time since text alone was considered any kind of proof in and of itself. Whether human or machine, the truthfulness of text has to be …
▲243GPT-6.1 Astra DelayMiscellaneousr/OpenAI · 30评论 · 链接↗
r/ClaudeCode
热评 2 条
▲46 IndieDevWannabe: You cant fire me. I fire me!
▲33 PracticalWelcome7793: now the AI can not replace itself ![gif](giphy|d3mlE7uhX8KFgEmY)
热评 2 条
▲47 Moneyshot_Larry: I am but a human API. A human MCP if you will
▲32 CartographerFit9203: It stings because it's like 60% true. Keep a running doc of the prompts that actually worked and every spot where you had to fix Claude's output, because that doc is the only thing standing between you and being replaced by the API key your boss already has.
Super curious on workflows, i personally switched to claude desktop app recently and have found it more convenient, but i am interested if majority is still on terminal setups
终端与桌面Claude Code工作流之争获得565条评论,入口选择仍高度分化。
▲170Why 6.1 Astra was DelayedHumorr/ClaudeCode · 8评论 · 链接↗
r/artificial
热评 2 条
▲44 Chuu: Stockfish, the current best chess engine, is 3650 for reference. The way ELO math works (which gets inaccurate with differences this large) means Stockfish should win 99.92% of games. Note that doesn't mean it should lose 1-x of that, since draws in chess are a common outcome, so by the book it's even more dire than it looks.
▲11 daruxo: they're using alphazero/leela as the backend and training their model to produce that same evaluation along with an explanation. neat but it won't be good at "general tasks" unless you have some external cheat sheet evaluation machine (engine) for said tasks, I think
History is something I'm quite passionate about and at work I have experience with software development and LLMs. So I thought how can I bring these things together and built something that I would actually use. This is the results of a few months of hard …
热评 2 条
▲6 NaturalSelecty: Fantastic concept. Definitely keep working on this.
▲6 ikeif: So it's _sourcing_ it's response, but not really verifying it, right? This is the current problem in a lot of AI summaries - it will FIND sources, but then act like it's true because… there is a facebook post about it. Is this a similar vein where "if there is a source, it takes it, cites it, but does not bother to verify that the source itself is even correct"?
Quick disclosure first: Genspark annual subscriber and a pretty heavy user. I use it for slides, research, organizing web stuff, random agent tasks, and I’ve stolen more than a few community Skills to solve oddly specific problems. Recently I watched Genspark …
▲6Best text model currently?Discussionr/artificial · 3评论
What model is the best at text only
▲4Context Language ModelsResearchr/artificial · 1评论 · 链接↗

🔎值得深挖

  • Agent身份标准这个方向值得研究:比较浏览器自动化、MCP客户端和网站API现有的身份与授权机制,看看哪些能同时满足用户委托、服务端限流和事后追责。
  • 第二Agent审查这个方向值得研究:对同一批高风险与普通任务测拦截率、误报率、额外延迟和额度消耗,判断它是在降低监督成本还是制造新的批准层。
  • 订阅内常驻Agent的经济学值得研究:把用户看到的token、平台内部成本、限额恢复、人工复核和失败重跑放在一起,估算当前高强度用法是否只是短期补贴。

📖附录:原文流(按作者)

想翻谁点谁展开。转推与噪音折叠在各自账号内。
🤖 AI行业
Tibo@thsottiauxAI801.2K粉 · 6条启动Codex连续28天交付,先公布约50%提速,再开放免费Auto-review。
@thsottiaux2026-10-06♥4.2K👁310.4K↗ 打开X
Day 2.1/ We have made Auto-review free for all users signed in through a ChatGPT account. You can enable it in settings > permissions > auto-review. Auto-review improves upon the default sandbox setting that requires you to approve everything, which is prone to decision fatigue unless you spend a lot of time configuring specific rules. It allows you to run long tasks while having a second agent review all actions taken by the primary agent. Its only goal is to prevent high-risk actions from being taken and to protect against unwanted actions that are not aligned with the original user intent. This Auto-review feature is now free and does not draw usage from your plan.
引用@thsottiauxDay 1/ We have optimized the default speed to be ~50% faster across GPT-6 Astra and GPT-6.1 Sol through the subscription across all our products and partners using Sign in With ChatGPT (including OpenCode, Pi, Amp, Devin, ...). No changes needed on your end … ↗
Auto-review对ChatGPT登录用户免费,第二Agent专门检查高风险与偏离原意的动作。
@thsottiaux2026-10-06♥5.6K👁452.5K↗ 打开X
This might have been my best day so far at oai. Ridiculous amounts of fun and intensity. Future is bright
@thsottiaux2026-10-06♥2.2K👁413.8K↗ 打开X
The real story is this one
引用@meetp_ai@petergostev Counting only equivalent API cost would show only half the picture IMO. https://t.co/CjEuNFBfzP ↗
@thsottiaux2026-10-05♥1.8K👁303.4K↗ 打开X
Sergio is cooking up magic with collaborative space right in ChatGPT. Improving leaps and bounds every day.
引用@matteingWe weren't kidding. Let's talk about Space. ✨ Have you tried Space yet? What's good or bad about it? What could we improve? We're locked in and shipping fast. 👀 ↗
@thsottiaux2026-10-05♥19.2K👁2.2M↗ 打开X
Day 1/ We have optimized the default speed to be ~50% faster across GPT-6 Astra and GPT-6.1 Sol through the subscription across all our products and partners using Sign in With ChatGPT (including OpenCode, Pi, Amp, Devin, ...). No changes needed on your end and this should be felt within the next two hours.
引用@thsottiauxOver the next 28 days, each day we’ll either ship one thing that is a clear improvement and relevant for most codex/work users or ship a full reset. Let the improvements begin. ↗
Codex负责人称Astra与Sol订阅内默认速度约提升50%,覆盖多款Sign in with ChatGPT产品。
@thsottiaux2026-10-04♥25.8K👁7.0M↗ 打开X
Over the next 28 days, each day we’ll either ship one thing that is a clear improvement and relevant for most codex/work users or ship a full reset. Let the improvements begin.
引用@thsottiauxAll right, we’re locking in. Only things being worked on are simplifications, more efficiency for more usage, groundbreaking features or new models. Sometimes you have to invest ahead of the curve, but feedback is clear that you all want things to get … ↗
宣布连续28天每天交付一项普遍相关改进,否则做完整重置。
Thariq@trq212AI362.8K粉 · 4条展示Claude Code的HTML计划skill与local hands模式,也分享游戏设计学习资源。
@trq2122026-10-06♥245👁26.5K↗ 打开X
if you're getting into game design, there are so many good resources out there to learn from!
引用@MortdogI've never met Wyatt Chang or worked with him (though I know people who have and they spoke highly of him), but that aside, these shorts he's been making are GREAT if you love game design. HIGHLY recommend them, as they're full of good nuggets. I've enjoyed … ↗
@trq2122026-10-06♥1.2K👁104.7K↗ 打开X
a big advantage of doing planning like this is that it’s a lot more token efficient compared to raw HTML- the model doesn’t need to remake the components or logic to do common things like state machines, diagrams, code snippets, etc.
引用@trq212I've been working on a skill that makes better HTML plans in Claude Code. It uses simple language, shows code snippets, surfaces questions & makes mockups. Linting reduces the normal failure cases that Claude runs into. Would love your feedback before … ↗
称结构化计划比原始HTML更省token,因为常用组件与逻辑无需重做。
@trq2122026-10-05♥967👁76.2K↗ 打开X
my favorite name for this pattern is "local hands", Claude runs in the cloud but can access your files locally- also coming to cowork
引用@dfeinitionYour Claude Projects cloud session can now connect to a folder you approve on your computer. The session stays in the cloud and only reaches for that local folder when a task needs it, reading and editing the files in place. Starting to roll out today. … ↗
把云端Claude访问获批本地文件夹的模式称为local hands。
@trq2122026-10-05♥3.6K👁293.4K↗ 打开X
I've been working on a skill that makes better HTML plans in Claude Code. It uses simple language, shows code snippets, surfaces questions & makes mockups. Linting reduces the normal failure cases that Claude runs into. Would love your feedback before shipping more broadly! https://t.co/U2Gb1K37vl
用简单语言、代码片段、问题和mockup生成可lint的HTML计划。
Matt Pocock@mattpocockukAI364.5K粉 · 9条发布skills v1.3,把PR证据、规格实现和会话复盘纳入主流程,并建议按团队规则专门化。
@mattpocockuk2026-10-06♥303👁16.4K↗ 打开X
IMO you should be specializing my skills to your work I.e. you should extend /grill-me to make sure it: 1. Holds you to a high standard of product design 2. Does a survey about security implications 3. Uses /show-me to preview the code for you
建议把通用skills扩展为产品设计、安全与代码预览等团队专用规则。
@mattpocockuk2026-10-05♥527👁22.6K↗ 打开X
Thanks for 400K on YouTube pals 👌 Got recognised at my kids' local park the other day, things are getting mad
@mattpocockuk2026-10-05♥517👁36.3K↗ 打开X
Having shipped the AI Coding Crash Course (largely for beginners/intermediates) means that the upcoming Cohort will be really quite advanced/in-depth Just recorded an 11-minute /grill-with-docs session about DDD, scenario testing, and sharpening language. It's so good
@mattpocockuk2026-10-05♥833👁57.7K↗ 打开X
Prompt of the day: /retro take a look at as many GitHub PR review comments as you have access to from my team. Create suggestions for updates to CODING_STANDARDS.md (breaking into multiple files as needed). Makes your PR comments inform the code going forward.
建议从历史PR评论反推CODING_STANDARDS.md更新。
@mattpocockuk2026-10-05♥3.0K👁207.5K↗ 打开X
mattpocock/skills v1.3 is out! - /pr (new) writes easy-to-read PR bodies, showing hard evidence that the changes work and assessing merge risk - /implement-spec (new) takes in a spec and tickets, and implements them with subagents - CONTEXT.md renamed to GLOSSARY.md - /retro (new) reviews recent coding agent sessions and suggests repo improvements /retro, especially, feels like a huge upgrade. Enjoy!
skills v1.3加入/pr、/implement-spec和/retro。
@mattpocockuk2026-10-05♥2.7K👁135.2K↗ 打开X
Oh shit I just realised that grilling is where the meat and the metal interact
@mattpocockuk2026-10-05♥889👁50.1K↗ 打开X
I've added /retro into the "Main Flow" of my skills - I consider it that important. Even running it on a 'successful' agent session can turn up inefficiencies and workarounds the agent is battling against. https://t.co/PybM9PXjBL https://t.co/AlfaeT9T0l
把/retro放进主流程,成功会话也要寻找低效与绕路。
@mattpocockuk2026-10-04♥2.5K👁127.4K↗ 打开X
Prompt of the day: 1. Re-check mattpocock/skills for the v1.3 release. Rename CONTEXT.md to GLOSSARY.md. Check /setup-matt-pocock-skills for any required changes. Perform a diff of the skills in my repository vs Matt's. 2. Get a sense for how I use Matt's skills: scan my last 25 sessions for skill invocations. Then, recommend how this release might improve my workflow. Enjoy v1.3, pals.
@mattpocockuk2026-10-04♥2.7K👁180.0K↗ 打开X
By popular demand, I'm shipping v1.3 of my skills today. Too much hype behind /retro not to ship it. Enjoy going back through your old transcripts and figuring out improvements. Docs, video, and announcement post coming Monday. https://t.co/ItW5zc7fym
▸ 折叠3条(转推/噪音)
转推2026-10-05 RT @joelhooks: a chat with @mattpocockuk about building reliable software factries and where to start https://t.co/frCFxyfrKz
转推2026-10-05 RT @poteto: everything i know about managing agents i learned from the amazing programmers and computer scientists that came before constr…
转推2026-10-04 RT @andrestaltz: Ok, `/retro` by @mattpocockuk is a game changer. Turns out my agents are bumping into all kinds of problems that went unde…
clem 🤗@ClementDelangueAI719.1K粉 · 3条密集转发Beam、开放决策模型与llama.cpp更新,并详解把真实coding harness变成RL环境。
@ClementDelangue2026-10-05♥668👁29.5K↗ 打开X
引用@reflection_aiIntroducing Beam: a highly efficient agentic open model with 501B total parameters and 23B active. - Frontier reasoning efficiency - Advances the Western open frontier on coding & agentic tasks - Trained end-to-end from scratch Full weights release this … ↗
Hugging Face CEO引用Beam官方公告,属于同一原始来源的放大。
@ClementDelangue2026-10-05♥596👁42.8K↗ 打开X
Clef from @Cloudflare is number trending on HF! https://t.co/dloDL2nmy6
@ClementDelangue2026-10-05♥2.1K👁123.5K↗ 打开X
We turned Claude Code, Codex, Hermes, Pi, @opencode and other coding harnesses into RL environments. No changes to the harnesses, no changes to the training code. Any open model, any task set, fully open source my friends! Same model, same weights: 62% under Mini-SWE-Agent, 33% under Claude Code. But training inside a real harness normally means reimplementing it as an environment, so most models get trained in a scaffold nobody actually ships. The fix is a proxy, not a rewrite. The harness thinks it's talking to a model API. It's actually talking to a capture proxy that speaks the 4 formats coding agents use (OpenAI Chat Completions, OpenAI Responses, Anthropic Messages, Gemini), forwards to @vllm_project, records the exact token IDs and logprobs vLLM sampled, and hands TRL sequences it can train on. The harness becomes the environment. 10 harnesses run through it today, none modified. And because you control the reward, you can shape behavior the harness never asked for. We added a small bonus for solving a task in fewer tool calls: on tasks it already solved, the model now uses 31% fewer calls, in every harness, and about half under Codex. Tested on LFM2.5-2.6B from @liquidai: → Train in one harness: better mostly in that harness (OpenCode 34% → 58%). → Train in 4 at once: better in all 4 (42% → 54%). → SFT on 3,189 rollouts from Qwen3.8-27B instead: plateaus at 47.5%, below both RL runs. Everything is open and reproducible: the capture proxy in OpenEnv, the trainer in TRL, the tasks, the SFT data, the training code and all 7 trained models. Bigger models and bigger runs next. Full guide: https://t.co/sKZURuOcza
把10种真实coding harness接成RL环境,并报告工具调用减少31%的实验结果。
▸ 折叠11条(转推/噪音)
转推2026-10-05 RT @ggerganov: The new v0.6.0 release packs a lot of good stuff: - Clef support (text + vision) - High-quality support for Qwen3.8-Flash-N…
转推2026-10-05 RT @arthurmensch: Excited https://t.co/4URuBCQVUR
转推2026-10-05 RT @CommandCodeAI: Introducing Agr, an open decision model by Command Code. · Agr (31B) and Agr-flash (360M) · 58.15 on Decision Index 0.2…
转推2026-10-05 RT @ben_burtenshaw: RL environments now have a home on the @huggingface Hub. Every RL framework has its own way to find environments. Cus…
转推2026-10-05 RT @qvac: One of our PR just got merged into llama.cpp: new Metal kernels for speculative decoding on Apple Silicon. Before this PR, specu…
转推2026-10-05 RT @CongWei1230: Thanks for sharing! We open-sourced code and models. 🤗Huggingface Paper: https://t.co/BrdK3sXeiy 💻 Code: https://t.co/S9o…
转推2026-10-04 RT @socialcapital: Microsoft CEO Satya Nadella wrote that a company renting an AI model “essentially pays for intelligence twice”: Once in…
转推2026-10-04 RT @AndrewCurran_: Axios is reporting that Reflection AI is about to release an extremely capable open weight model expected to compete wit…
转推2026-10-04 RT @akshay_pachaar: HuggingFace just closed a major gap in harness engineering! (the ultimate guide to multi-harness RL) the same open-we…
转推2026-10-04 RT @CoastalFuturist: How it feels to ask the Chinese ai models for help on something Claude won’t agree to https://t.co/TPYgiOlvnF
转推2026-10-04 RT @Thom_Wolf: new roommate just moved in. walks like he's had three drinks, says no to everything. still quite cute so I might bring him e…
Vaibhav (VB) Srivastav@reach_vbAI61.1K粉 · 4条转述Codex提速与免费Auto-review,也继续推广Sol相对Astra的订阅用量优势。
@reach_vb2026-10-06♥92👁6.0K↗ 打开X
Auto-review is now free for everyone who uses it via ChatGPT account! I default to “Approve for me” for all my codex tasks and auto-review makes sure that I’m protected from any unintended consequences If you haven’t already, enable it in your settings - Enjoy!
引用@thsottiauxDay 2.1/ We have made Auto-review free for all users signed in through a ChatGPT account. You can enable it in settings > permissions > auto-review. Auto-review improves upon the default sandbox setting that requires you to approve everything, which is prone … ↗
reach_vb称自己默认让Auto-review审查Codex任务中的非预期动作。
@reach_vb2026-10-05♥601👁28.5K↗ 打开X
LFG!! Team optimised inference for GPT 6 Astra and 6.1 Sol to be 50% faster!! Across Sign-in with ChatGPT and plans Rolling out over the next couple hours - Enjoy!
引用@thsottiauxDay 1/ We have optimized the default speed to be ~50% faster across GPT-6 Astra and GPT-6.1 Sol through the subscription across all our products and partners using Sign in With ChatGPT (including OpenCode, Pi, Amp, Devin, ...). No changes needed on your end … ↗
转述Astra与Sol订阅内推理速度约提升50%。
@reach_vb2026-10-04♥102👁6.2K↗ 打开X
down with fever the last two days but felt strong enough to go for a walk in the park nearby did not disappoint https://t.co/aFr6kMT2HV
@reach_vb2026-10-04♥310👁18.5K↗ 打开X
GPT 6.1 Sol is all you need! It’s my go to on both work as well personally work According to arena you should get ~5x more usage as compared to Astra
引用@petergostevCodex limits feel fine now with GPT-6.1-Sol as the daily driver. With Astra, the limits were a disaster, but now I'm not close to killing the sub. As per Arena's real usage cost-per-task data, you should get 5x more out of GPT-6.1-Sol than Astra. ↗
▸ 折叠4条(转推/噪音)
噪音2026-10-05 stay hungry. stay foolish.
转推2026-10-05 RT @matteing: We weren't kidding. Let's talk about Space. ✨ Have you tried Space yet? What's good or bad about it? What could we improve?…
转推2026-10-05 RT @romainhuet: Thinking of Steve Jobs today. Hard to believe it’s been 15 years without him. He made technology feel human, and inspired…
转推2026-10-04 RT @thsottiaux: Over the next 28 days, each day we’ll either ship one thing that is a clear improvement and relevant for most codex/work us…
Simon Willison@simonwAI233.7K粉 · 1条用本地Qwen 3.8 27B重跑长整数加法实验,对比推理与非推理模式。
@simonw2026-10-05♥401👁40.2K↗ 打开X
I re-ran an experiment @colin_fraser ran against GPT-4o a while back to see how good it was at adding long numbers, only this time I tried Qwen 3.8 27B running locally in both reasoning and non-reasoning modes https://t.co/IOufJIA4N9 https://t.co/4kB33NCPcO
用本地Qwen 3.8 27B重跑长整数加法实验。
Peter Gostev@petergostevAI27.8K粉 · 11条讨论常驻Agent应承担职能责任、操作系统权限架构和订阅价值,也继续更新连续学习实验。
@petergostev2026-10-06♥17👁1.4K↗ 打开X
Now we have 'Sign in with ChatGPT', but what about getting Gemini credits when you 'Sign in with Google'? Surely Google has this market to themselves - when I sign in with Google I should have an option to use my Gemini inference & data - seems like they forgot about that
@petergostev2026-10-05♥18👁1.7K↗ 打开X
I know we like to complain, but it is so refreshing to be in a industry where people give a shit. I've worked before in places where you see a glaring massive customer issue and the attitude is: "well we have this two year transformation programme, let's hope it gets fixed"
@petergostev2026-10-05♥13👁1.7K↗ 打开X
There will be a time when someone comes up with the right OS architecture where the agents can do what they want with zero permission pop-ups, without your life won't get ruined
@petergostev2026-10-05♥386👁24.0K↗ 打开X
For real? SemiAnalysis: "Anthropic Subscriptions Offer 5x+ More Value Than OpenAI" https://t.co/TO0uERwrVW
引用第三方报告称Anthropic订阅比OpenAI提供5倍以上价值。
@petergostev2026-10-05♥8👁1.8K↗ 打开X
newb macos question - why does nobody do icons with transparent background? I got astra to design this cute little one for me, much nicer isn't it? https://t.co/1KcxuhW1Do
@petergostev2026-10-05♥20👁1.4K↗ 打开X
What's going to get your attention when everyone has access to perfect video, 3D and html generation tools? I feel like I'm already glazing over a lot of this stuff, when only a few weeks/months ago I'd been amazed
@petergostev2026-10-05♥0👁449↗ 打开X
October 5th - Ship or Reset? Your PREDICTION (not what you want to happen) Tibo: "Over the next 28 days, each day we’ll either ship one thing that is a clear improvement and relevant for most codex/work users or ship a full reset. Let the improvements begin."
@petergostev2026-10-05♥111👁7.8K↗ 打开X
I know nobody cared about this at OpenAI dev day, but I think this is one of the biggest potential future product lines. So far all agents assume there's a human who is taking responsibility & prompting it. One clear way forward is to have an agent that takes full responsibility for something - e.g. accounting or email marketing, same way like an employee would. The reason why you hire people is not just they are good at what they do, but they'll take responsibility and make sure things happen. That's why you are ok hiring people even with zero experience in that function - the important part is that someone is thinking about it and it's not all on you. What happens now with agents is kind of a disaster - since you can do so much more, you end up going into more and more areas. But guess what, the agents aren't responsible for anything, so everything new that you are doing is 100% your responsibility. So yes agents are helping, that amount that you need to worry about grows enormously. It is also not realistic to expect everyone to learn to use agents effectively, Excel has been around for 30+ years and what's the % of people who have any clue how to use SUMIFS or something simple like that. AI is even harder to grasp, since model capabilities & products are changing every month. If AI adoption relies on people learning how to use agents en masse, we are screwed, it is just not going to happen. So instead what we need is agents that just work. Imagine you are a starup founder, you need an accountant. Instead of spending 1/5th of your headcount on hiring an accountant, you pay $30k a year and get an AI accountant that would just take care of things - it would monitor your work, proactively tell you what's wrong and have weekly meetings with you to discuss issues, like you would with a human. I have no idea how well Specialist dots will work, but it is the first mainstream agent that you don't need to actually prompt. It is possible that it is a year or two too early, but I really like this pattern, this has to be part of the future in some way.
认为Agent应像员工一样持续负责职能,而不是扩大用户需要亲自提示和监督的范围。
@petergostev2026-10-05♥83👁3.9K↗ 打开X
Continuous Learning benchmark - Opus 5.5 had a good run of improving performance from game 75 until c. game 160. But then performance declined somewhat. I might continue more after the reset & do the analysis of how they approached learning Watch here: https://t.co/KTHfxlvsSL https://t.co/lDZVfTElYa
引用@petergostevI have a Continuous Learning benchmark where models attempt to learn to play chess. They are given a /goal of learning and improving playing against a Stockfish opponent in 200 games. They can choose the difficulty, take notes, whatever they like - except … ↗
@petergostev2026-10-04♥325👁10.5K↗ 打开X
AI just saved me $12,999.99. I wanted to buy this lens, so I asked my agent. It told me that I didn't have enough money, so I didn't buy it. https://t.co/1rkvkFfUkP
@petergostev2026-10-04♥285👁31.6K↗ 打开X
Codex limits feel fine now with GPT-6.1-Sol as the daily driver. With Astra, the limits were a disaster, but now I'm not close to killing the sub. As per Arena's real usage cost-per-task data, you should get 5x more out of GPT-6.1-Sol than Astra.
▸ 折叠1条(转推/噪音)
转推2026-10-05 RT @etnshow: GM. Monday. Here's the line up: - @sytaylor (Fintech Brainfood) - @petergostev (Arena) - Linus Hakansson (Gravitee) - @20thr…
OpenAI Developers@openaidevsAI449.6K粉 · 2条演示Codex CLI语音启动、跨项目管理与worktree,也转发Space和定时任务更新。
@openaidevs2026-10-05♥204👁14.0K↗ 打开X
If you live in the terminal, this one’s for you. Watch how to start tasks by voice, manage agents across projects, and explore another direction in a separate worktree with the Codex CLI. https://t.co/Q4GmBprj27
演示Codex CLI语音启动、跨项目Agent管理与worktree并行探索。
@openaidevs2026-10-05♥338👁22.2K↗ 打开X
Achievement unlocked: Meet the winners of our Side Quests. https://t.co/sQjVAqtz0P
▸ 折叠4条(转推/噪音)
转推2026-10-05 RT @AmirMushich: Rebuilt this demo for a real brand with Astra 6 + LTX-2.5 (video model) → I took real product photos → generated rotatio…
转推2026-10-05 RT @mxstbr: we removed a toggle from chatgpt! no, not that one. but, you no longer need to toggle developer mode to connect custom MCP se…
转推2026-10-05 RT @axeldelafosse: Pro tip: you can now edit the model and/or reasoning effort of your scheduled tasks directly from the "Edit task" modal…
转推2026-10-05 RT @PrtkYdv: Space highways https://t.co/PamZ5FSe8N https://t.co/I7g6X6vZUL
Ben Tossell@bentossellAI201.4K粉 · 2条以旅行、会面和轻评论为主,原创产品信号较少。
@bentossell2026-10-05♥26👁4.9K↗ 打开X
unlike today, steve jobs actually understood the general public
@bentossell2026-10-04♥76👁6.4K↗ 打开X
last stop for me this trip to sit and think with @om like we used to although last time we did a bird shit on me stay curious https://t.co/A2ZyqGWN2Z
▸ 折叠2条(转推/噪音)
转推2026-10-05 RT @matanSF: Matt Murphy is a legendary investor with a track record of being very early and being very right. He led the Series D of Ant…
转推2026-10-04 RT @jakozloski: The best part of being married is getting to say "let me check with my wife" about something you personally have absolutely…
Alex Volkov@altryneAI43.1K粉 · 5条把AI助手的核心边界概括为信任,也继续记录Muse、Grok和Instinct的产品缺口。
@altryne2026-10-05♥11👁1.6K↗ 打开X
The most important boundary in AI Assistants is TRUST. Countless conversations I've had with folks about their AI Assistants, all boil down to "Do I trust this assistant to do what I want, to represent me well, to not leak my data" - For @Muse, many concerns about Meta's past practices, privacy and "they will use my data to send me ads" - For @bot, concerns are with Elon specifically, and the name Grok reminds some of the mechahitler incidents. - For Instinct, weirdly everyone who I saw uses it, YOLO's in. Despite this being run by a 24yo with a large VC fund with no clear business model and a lot of online incidents of retaining data and no deletion policy
把AI助手信任拆成按意图行动、代表用户和不泄露数据三类顾虑。
@altryne2026-10-05♥11👁1.7K↗ 打开X
This has GOT to be either a bug or the dumbest product decision in the history of product decisions. Embedded grok has no per user context! 😵‍💫 @elonmusk plz fix, it's been a while like this https://t.co/490vEyhct5
@altryne2026-10-05♥32👁3.1K↗ 打开X
My top-10 AI wishlist for October 2026: ☐ @muse on Meta glasses (+ animated voice) ☐ @bot in my Tesla (for Cursor ultra accs) ☐ Dots standalone app from @OpenAI ☐ Decisions API access also from @OpenAI ☐ Qwen 4! ☐ Anthropic to step into the AI Assistants race ☐ Native Hermes ios/macOS apps from @NousResearch (+ a deal with Anthropic to use Opus 5.5 sub in Hermes!) ☐ @typesafeai Image JEV! ☐ Muse 🍉 !! If I don't get 8/10 this October will be disappointing! What's your top-10?
@altryne2026-10-05♥3👁2.0K↗ 打开X
Bending Spoons just hoovers every category of Saas up huh? Just noticed this from @streamyardapp https://t.co/IRhJHaInrA
@altryne2026-10-04♥25👁3.9K↗ 打开X
Are you touching grass today anon? Or tokenmaxxing?
▸ 折叠2条(转推/噪音)
转推2026-10-05 RT @petergostev: AI just saved me $12,999.99. I wanted to buy this lens, so I asked my agent. It told me that I didn't have enough money…
噪音2026-10-04 https://t.co/i1jEvQZHT3
swyx@swyxAI199.5K粉 · 1条只转发AI-native生活方式课程,并称Meta AI团队18个月内完成明显形象反转。
@swyx2026-10-06♥15👁1.8K↗ 打开X
the epic 18 month vibe shift of @AIatMeta from the Llama 4 shitshow to now being routinely mentioned in the same breath as other frontier labs needs to be studied @alexandr_wang (and his shockingly deep bench) is both built different and builds different. unreal https://t.co/bfmN0mG6o9
▸ 折叠1条(转推/噪音)
转推2026-10-06 RT @nicknisi: My @aiDotEngineer workshop is up: "Lifestyles of the AI-Native" with @zackproser 🎥 Voice coding across agent sessions, goals…
Hassan@nutlopeAI101.1K粉 · 1条发布Together Link,让多个coding harness自动路由开放模型并记录用量。
@nutlope2026-10-05♥15👁1.2K↗ 打开X
Introducing Together Link! A CLI to run open models inside your favorite coding harness, with: → Auto routing that picks a model based on your task → Support for claude code, codex, opencode, & more → Built in spend & usage trackings https://t.co/KPiGLgPAlY
Together Link为Claude Code、Codex与OpenCode等harness自动路由开放模型并跟踪花费。
🧑‍💻 独立开发者
@levelsio@levelsioindie969.7K粉 · 13条围绕Airbnb评分失真密集推广Hotelist,并新增公私营、酒店式公寓和联盟披露等筛选与说明。
@levelsio2026-10-05♥119👁21.9K↗ 打开X
Also one more filter for today: 🔲 Public company ☑️ Privately owned 🔲 Government owned I love ETFs as much as the next guy but they do put a pressure on stocks to grow by 10%/year or more Hotels have a hard time doing that, so how do you grow? You cut costs and big hotel chains have been doing that in the name of DEI and ESG and wokeness a lot. Locking down ACs, no daily cleaning, cutting corners, etc. it all saves money Those cost savings mean higer profit margins mean growth! Great as an ETF or stock holder, terrible as a hotel guest So you can now filter on privately owned One of the most consistent quality hotel chains I visit is Four Seasons (recommended by many people on here), not cheap, but almost always a good experience Guess who owns Four Seasons? It's private! And it's @billgates and @Alwaleed_Talal They must have a had such terrible hotel experiences, they had to buy their own hotel to get a good one! Can you imagine? 😂
引用@levelsioAdded this honest affiliate notice https://t.co/rixX0vZ5Dy ↗
@levelsio2026-10-05♥1.7K👁395.8K↗ 打开X
I take 10g per day!
引用@morellifitThis new creatine study is insane. People who never trained took 10g of creatine a day and gained more lean tissue (2.43 lbs) than people who were lifting 3x a week with no creatine (0.31 lbs). Here's the breakdown: (1/18) https://t.co/SlyHDWZSTU ↗
@levelsio2026-10-05♥338👁53.1K↗ 打开X
400g skyr + 24g Vivani 92% chocolate + lemon juice 44g + 2g protein = 46g protein Dark chocolate has healthy flavanoids and barely any sugar Great for your heart and brain! https://t.co/xtbZzP7ViA
引用@levelsioSkyr is quite dry by itself, so I always add one fully squeezed lemon + cinnamon powder That makes it juicy and tasty Sometimes I also add 20g of 92% Vivani dark chocolate, I put it on a wooden plate, then cut it up with a knife into tiny almost powdery … ↗
@levelsio2026-10-05♥67👁33.5K↗ 打开X
Added this honest affiliate notice https://t.co/rixX0vZ5Dy
引用@levelsioHe's right, there's more and more hotels offering this so I added an [ Apartment-style ] filter for hotels with kitchen and residential vibes https://t.co/HKgd5L4qtI ↗
@levelsio2026-10-05♥556👁41.0K↗ 打开X
This is my consistent experience too You book an Airbnb Luxe Plus Verified Guest Favorite rates 4.99/5 and it's still a dump Then you review it and your review gets removed ???
引用@szuchans@levelsio I had an Airbnb with black mold covering the bathroom ceiling—this was a "guest favorite." They refunded me, but the listing remained up, still a guest favorite. I'll never use it again. ↗
@levelsio2026-10-05♥94👁30.8K↗ 打开X
He's right, there's more and more hotels offering this so I added an [ Apartment-style ] filter for hotels with kitchen and residential vibes https://t.co/HKgd5L4qtI
引用@danmichalczyk@levelsio Isn't this a solved problem? There are a bunch of chains that rent out rooms / studios / apartments with kitchens. Placemakr, Mimaru, Fraser Suites, etc. are all great! ↗
@levelsio2026-10-05♥420👁61.1K↗ 打开X
What most people don't know about Airbnb "If a person doesn’t like a review they received from you and wrote you a good review they can remove theirs and with it the rating of their experience." 100%
引用@daveying99Here’s the biggest scam with airbnb reviews that 99% of people don’t know. Most people think airbnb reviews are double blind - the reviews are only visible to the other side after they both write them or 14 days elapse. This is half the story. Only that if … ↗
@levelsio2026-10-05♥866👁118.9K↗ 打开X
I'm not paid by the hotel lobby or something btw I was a big fan of Airbnb and would prefer to stay in Airbnbs if they were great I want to have a kitchen to cook steak and vegetables, and not go outside to eat in restaurants or get shitty hotel food But I can't book Airbnbs because I the "rating anxiety" is so high because I can't trust the ratings and I end up with my bags in some place that sucks, so the variation of quality is too high to be able to seriously use for travel (except maybe in Asia) So they leave us no choice to just go back to hotels Meanwhile hotels still don't have kitchens! What a missed opportunity! Hotel industry could CLEAN UP entire Airbnb by just adding kitchens to rooms
引用@AutismCapitalImagine choosing an Airbnb over a hotel https://t.co/llU25Ehvyq ↗
@levelsio2026-10-05♥992👁76.8K↗ 打开X
Airbnb's PR agency *scrambling* You know it's probably them because it's PR speak to a tee https://t.co/FI07JFQ7Q0
引用@levelsioAirbnb's entire rating scale is now 4.5 to 5.0 You can't make this up 😂 https://t.co/43Nsjuaktb ↗
@levelsio2026-10-05♥79👁23.2K↗ 打开X
✨ Rebuilt the https://t.co/UXK5AFqCaQ meetups page from scratch It now has - individual meetup pages - category pages (like Coworking meetups) - city and country pages (like Meetups in Bangkok) And I tried to make it look better 😊 And it's also MCP-ready, so you can create, edit, or see all meetups via the Nomads MCP
引用@levelsio✨ I have made coworking meetups a special category on https://t.co/xQI1chWXtQ now So you can organize one in any city in the world to meet people there I've also added a [☑️] Recurring Meetup option now so if you do a weekly cowork or event, you don't need … ↗
@levelsio2026-10-05♥586👁50.6K↗ 打开X
I analyzed the range of ratings on booking sites and it's pretty crazy Airbnb goes just from 4.5 to 5 with a median of 5! But most others aren't better, they all range from about 4 to 5, and almost all have a median of 4.25, with not many scores outside of it! The best site in terms of honest ratings is actually Yelp, whatever they're doing it's working, it shows an actual normal distribution and ratings you can trust Hotelist itself is a normal distribution with median at 3.5, which was my goal, most stays are average/okay which I'd consider 3.5/5
引用@levelsioAirbnb's entire rating scale is now 4.5 to 5.0 You can't make this up 😂 https://t.co/43Nsjuaktb ↗
比较多个订房平台评分分布,并用归一化构建Hotelist评分。
@levelsio2026-10-05♥7.1K👁915.6K↗ 打开X
Airbnb's entire rating scale is now 4.5 to 5.0 You can't make this up 😂 https://t.co/43Nsjuaktb
引用@levelsioAirbnb deletes negative reviews so everything is always 4.7 Google Maps removes negative reviews in many countries as defamation so everything is always a 4.7 Booking sites want you to book so will also happily remove negative reviews for their hotels so … ↗
以Airbnb评分集中在4.5到5分为切口推广Hotelist,数据方法未完整披露。
@levelsio2026-10-04♥1.2K👁78.6K↗ 打开X
Another one!
引用@PolymarketJUST IN: Britain’s second-richest billionaire David Reuben leaves the UK for tax-free Monaco, joining a growing exodus of ultra-wealthy residents. ↗
▸ 折叠2条(转推/噪音)
转推2026-10-05 RT @Baconbrix: I added Starships to Apple Maps so you can watch reusable rockets launch endlessly from Starbase, TX. https://t.co/lQchyY80HX
转推2026-10-05 RT @davorb89: built a safari extension to normalise airbnb's ratings, after reading @levelsio 's post on how broken they are. it stretches…
DHH@dhhindie915.0K粉 · 12条密集讨论Agent如何改变代码所有权、语言选择和软件供给,并继续表达强烈技术乐观。
@dhh2026-10-06♥310👁8.6K↗ 打开X
Don't tell me arguing online never moves a position. How very wholesome! https://t.co/Y8Wwgbi31S
@dhh2026-10-06♥762👁21.8K↗ 打开X
It was almost impossible for a programmer who wrote all their code by hand to divorce their ego from the output. Now agents are delivering what decades of agile admonishing about “collective ownership” never could. What a gift!
认为Agent产出让程序员更容易把个人自我与代码所有权分开。
@dhh2026-10-06♥245👁20.9K↗ 打开X
I love this. Nothing will nerdsnipe a programmer like seeing a benchmark they believe could be improved. We share the joy of a good RL loop with the clankers!
引用@lautI took a stab at improving the performance of Elixir. I used the PR for Elixir that @ZachSDaniel1 posted and worked some more on it. In the benchmark Elixir is now faster than Rust (ccece30 newest commit) for the Sidebar and Search benchmark. … ↗
@dhh2026-10-06♥580👁23.6K↗ 打开X
"We're about to see an absolute bloom in software development as the price of development plummets and everyone realizes how much automation we still have left to do in this world... Don't go down with the pencils. There's so much to build. We need you." https://t.co/HJbUwMiDEU
@dhh2026-10-05♥1.7K👁58.8K↗ 打开X
"This is not the time to doom, baby. This is the time to bloom. Intellectually, spiritually, and productively. Put down your anxieties, arrest your neurotic impulses, and decide to be happy about the present, the future, and..." https://t.co/c87G5NIGqz
@dhh2026-10-05♥114👁6.6K↗ 打开X
I can't believe that we're actually debating whether or why Rust — a systems level language meant to rival C/C++!! — would be much faster than the interpreted or GC-using alternatives people have historically used for web apps. Programming tribalism is a hell of a drug!
@dhh2026-10-05♥1.1K👁79.4K↗ 打开X
Keep your mind flexible! Rust has an edge at the moment, but Rauch is spot on: We should rethink the entire stack from first principles given the transition to agent-led programming. Very little of what was will remain that way.
引用@rauchgDHH is fundamentally right about Rust. For context, Vercel has been undergoing a Rust-ification (carcinization, technically 🦀) for a while. One of the first projects we migrated was Turborepo, from Go to Rust¹. The migration completed, but the RoI was … ↗
认为Agent降低迁移人力成本后,语言与技术栈需要重新从第一性原理评估。
@dhh2026-10-05♥5.3K👁205.1K↗ 打开X
You don't have to commit your entire life to optimism, just try it on for a week. Start thinking "what if it all worked out?" and "what a time to be alive!". Then compare your state of mind to the week before filled with pessimism and doomer nonsense.
@dhh2026-10-05♥799👁75.6K↗ 打开X
"Not only am I uninterested in the company’s home device, I can, for the first time, envision a future where I don’t buy Apple by default. Indeed, this already happened..." Anyone into the age of agents will eventually realize Apple is a bad fit. https://t.co/UpQhfWxyL2
@dhh2026-10-05♥598👁16.6K↗ 打开X
The unbridled enthusiasm will continue until moral improves.
@dhh2026-10-05♥232👁20.0K↗ 打开X
"[s/The future/AGI/g] is already here, it's just not evenly distributed" — William Gibson
@dhh2026-10-04♥2.1K👁73.8K↗ 打开X
"It is difficult to get a man to understand something, when his salary depends on his not understanding it" — Upton Sinclair I think this quote explains a lot of the coping that's going on, but I also think it's wrong! The failure to understand is what'll erase the paycheck.
▸ 折叠3条(转推/噪音)
转推2026-10-05 RT @fnthawar: To offend a strong man, tell him a lie, To offend a weak man, tell him the truth — Marcus Aurelius
转推2026-10-05 RT @jankeesvw: My daughter (11yo) is vibe coding a game in the car with Claude on Omarchy 🤯 https://t.co/TH7CjrBPRH
转推2026-10-05 RT @bryan_johnson: I reviewed Tobi’s biomarker’s from the car race yesterday. His heart rate, respiration, nervous system, focus, brain sta…
Nikita Bier@nikitabierindie1.3M粉 · 2条把Agent向服务端自证身份视为未来五年的重要基础设施问题。
@nikitabier2026-10-05♥4.5K👁379.9K↗ 打开X
The most important technology problem of the next 5 years: Creating a broadly accepted standard for agents to identify themselves to service providers, so that providers can adjust the way they interface with clients (as compared to human-based traffic). In the interim (i.e., for the next 6 months), there will be a cat-and-mouse game that agents will play -- to circumvent detection and maintain their product's utility during this growth phase. However, this will only be a stopgap and it will not be the terminal state of the world.
引用@JessicalessinSo in the last two days about half of my Muse use cases have vanished because the browser won’t do it any more. Are websites changing their policies? Bot detection??? ↗
主张建立Agent向服务提供方自证身份的广泛标准。
@nikitabier2026-10-04♥2.7K👁158.7K↗ 打开X
Tried to start a congo line at a birthday party last night and no one joined, so it just looked like I was giving someone a massage
Alex Finn@AlexFinnindie476.3K粉 · 3条主推Meta Muse自动化,也高调转述Reflection Beam开放权重模型。
@AlexFinn2026-10-05♥502👁55.6K↗ 打开X
You NEED to be using Meta Muse It's not only an incredible AI agent, but it's the best agent for automating your every day tasks I have basically every task I hate doing automated now In this video I go over 6 use cases for Muse that has changed my life: https://t.co/VG6PdUzgOj
@AlexFinn2026-10-05♥2.0K👁182.7K↗ 打开X
It finally happened America has entered the frontier open weights model competition. We aren't rolling over to China Reflection announced Beam, a 501B parameter open weights model that is comparable to GLM 5.2, Qwen 3.8, and Opus 4.8 Yes, Opus 4.8 on your desk You will need a decent sized machine for this. Probably the Mac Studio 512gb But if you listened to me back in January when I warned you days like today were going to come, then you're all set I stand by Opus 4.5 being the most important model release in history. We will now have a model BETTER than that on our desks If you haven't gotten into local AI yet, it's officially time to do it Take any computer you have, talk to an AI agent, ask which models you can run Never been more important to just start
引用@reflection_aiIntroducing Beam: a highly efficient agentic open model with 501B total parameters and 23B active. - Frontier reasoning efficiency - Advances the Western open frontier on coding & agentic tasks - Trained end-to-end from scratch Full weights release this … ↗
高调转述Beam 501B开放权重模型,跨模型能力比较仍待独立实测。
@AlexFinn2026-10-04♥979👁83.2K↗ 打开X
If you’re using Grok Bot, Dots, Muse, or Hermes, your agent is capable of WAY more than you think You’re probably only getting 1% power out of it The best strategy for getting way more done is the ‘reverse prompting loop’ Step 1: Give this reverse prompt: “Based on what you know about me and my work, what are 5 productive things you could do for me?” Step 2: Choose 1 of them Step 3: Do this over and over again until you have a ton of work done You’ll not only be ridiculously productive, you’ll also get tasks done you’ll never think your agent could do This becomes more powerful with the more data sources you have, especially the email connector I hate being in my email. Running this loop usually takes care of all my unread emails Try this out and let me know what you think!
▸ 折叠1条(转推/噪音)
转推2026-10-05 RT @AlexFinn: If you’re using Grok Bot, Dots, Muse, or Hermes, your agent is capable of WAY more than you think You’re probably only getti…
Tony Dinh@tdinh_meindie204.3K粉 · 6条上线TypingMind内置记忆,称Cursor云端Agent能做PR审查与QA,也在观察人类账号分发服务。
@tdinh_me2026-10-06♥16👁2.5K↗ 打开X
I find it crazy that the frontier models can look at a bunch of numbers in an SVG path and tell what the image is. 😳 https://t.co/64KERhVdP8
@tdinh_me2026-10-06♥31👁7.0K↗ 打开X
Thinking of starting a marketing agency that create content and use this to post to US market 🤔
引用@MilesFeldsteinAI can now hire an ARMY of USA based humans … to post on social media Give your agent 10, 100, or 1,000 dedicated social accounts. 🇺🇸 Run by humans in the US 🔥 Warmed up for your niche 📱 Accounts just for your brand 🚫 Zero VPNs, we verify every … ↗
@tdinh_me2026-10-05♥30👁6.2K↗ 打开X
Better late than never! We now have built-in memory engine in TypingMind!
引用@TypingMindAppMemory is now supported on TypingMind 👀 Your AI can remember context about you, your preferences, and the way you work - so you don’t have to repeat yourself in every new chat. https://t.co/mYk0qFqvSW ↗
TypingMind上线内置记忆,减少重复提供偏好与工作上下文。
@tdinh_me2026-10-05♥283👁34.0K↗ 打开X
Wait, it's been a long time since I noticed any frontier model hallucinate. They really fixed the hallucination problem!
@tdinh_me2026-10-05♥50👁5.9K↗ 打开X
I absolutely love @cursor_ai, the cloud agent is super helpful to do not only code review but also QA/testing. Need to spend a bit of time to set up the cloud environment with the full development stack but once it’s ready, all the PRs get free screenshots/video demo. So cool. https://t.co/0XCiLxJgCN
称Cursor云端Agent配好环境后能做PR审查、QA和截图视频。
@tdinh_me2026-10-04♥995👁60.5K↗ 打开X
Dear @1Password please remove the 0.5s animation everytime I unlock 😂
引用@vimtorhey @1Password you should have 2 KPIs: - avg credentials per workspace - avg time to fill a password please just optimize for that it's starting to get crazy how often i have to reauthenticate, unlock, or reopen the app because it's stuck ↗
▸ 折叠1条(转推/噪音)
转推2026-10-06 RT @TablePlus: We are proud to introduce our new product, https://t.co/AuG85uGOPi It is a Virtual Machine specifically designed for Apple…
Simon Høiberg@SimonHoibergindie163.7K粉 · 5条比较OpenAI与DeepSeek 4.1 Flash的实际工作成本,并继续强调地点独立的SaaS组合。
@SimonHoiberg2026-10-05♥22👁3.7K↗ 打开X
Used DeepSeek 4.1 Flash most of the day. It's very impressive, and it's very difficult to get it to spend $5. I got a ton of work done for this $1.31. https://t.co/zwppeE3Bje
引用@SimonHoibergThis morning with OpenAI: 30% - "The AI service is temporarily overloaded. Please try again in a moment" 20% - "I can't honestly help you with [...task]" 40% - Super stupid and does sloppy work. 10% - Actual good work done. While limits leaking like … ↗
作者称DeepSeek 4.1 Flash一天工作只花1.31美元,属单用户样本。
@SimonHoiberg2026-10-05♥84👁7.4K↗ 打开X
I moved from Denmark to Switzerland, then tried Spain for a few years, now I'm back in Switzerland - and my SaaS portfolio didn't notice. No office to close. No team to relocate. No client meetings to reschedule. Revenue kept coming in while I was unpacking boxes. This is what a real location-independent business looks like. If you need to step away from the business for a bit and it requires a ton of planning, you're probably more locked in than you think.
@SimonHoiberg2026-10-05♥12👁3.1K↗ 打开X
Same ngl 😬
引用@nateberkopecHate to say it but my podcast listening has also dropped off a cliff. I used to listen either in my car or on my indoor bike trainer, and in both places now I just talk to agents. ↗
@SimonHoiberg2026-10-05♥21👁3.5K↗ 打开X
This morning with OpenAI: 30% - "The AI service is temporarily overloaded. Please try again in a moment" 20% - "I can't honestly help you with [...task]" 40% - Super stupid and does sloppy work. 10% - Actual good work done. While limits leaking like crazy. Nerfed and rugpulled.
引用@SimonHoibergThis is getting completely out of hand. GPT-6-sol will outright refuse to do work over the slightest little details, completely unreasonably. First acting like a petty lawyer then giving me a small moral speech when I ask them to stop. I now added a switch … ↗
@SimonHoiberg2026-10-04♥976👁69.0K↗ 打开X
If you're a founder and a dad, you should live in 🇨🇭 Switzerland. - Incredible nature - Super safe and clean - Low taxes - Non-EU - Great for business - Perfect for families The best place on earth for men to build real wealth (I'm not just talking about money). https://t.co/sQIp5mZAaa
Marc Lou@marclouindie405.9K粉 · 8条转述AI skills目录10天到3000美元月收入,也继续用书、健身实验和线下聚会经营个人分发。
@marclou2026-10-05♥715👁61.6K↗ 打开X
The fastest-growing startup of the week on TrustMRR is from China 🇨🇳 @yihui_indie built a curated directory of AI skills and charges a subscription for full access. $0 → $3K/mo in 10 days. https://t.co/zKGmLsxrW8 https://t.co/ZvV5wPiMqE
转述中国开发者的AI skills目录10天从0到3000美元月收入。
@marclou2026-10-05♥127👁13.7K↗ 打开X
Next step: Dubai 🇦🇪 Do I know someone who could host a little indie hacker get-together on October 25th? 🤗
引用@marclouIstanbul meetup sold out in 2 hours! ~150 people came. The team at @eachlabs booked a wonderful venue and organized everything! They even brought baklava, and I ate one on stage 😅 Tip: Twist the baklava upside down so it's doesn't drip. My dear friend … ↗
@marclou2026-10-05♥306👁21.4K↗ 打开X
@marclou2026-10-05♥731👁74.3K↗ 打开X
“You need to go to the gym to get jacked” is a lie. Looking lean is mostly diet. We already have muscles, just hidden by layers of fat. Of course you need to grow muscles to look jacked, but you don’t need to bench 100kg. Your body weight is already weight. Calisthenics will make you jacked, at home, for free, without going to the gym with a PT. And you need some form of cardio. Walking is great. Running, cycling, swimming are even better. It doesn’t have to be intense and, like calisthenics, it’s mostly free. My fitness journey started with calisthenics at home, and surfing for cardio. I upgraded with 15kg dumbbells, and only started going to the gym last year. Start with what’s convenient. Make it a habit. Upgrade from there.
引用@calebsloopIf 10+ yrs of gymnastics taught me anything, it’s that you need far less than you think to stay in good shape I didn’t touch a dumbbell until I was 21. In gymnastics, you only train with your bodyweight The times I’ve felt and looked my best, I just walked … ↗
@marclou2026-10-04♥483👁29.9K↗ 打开X
The gym
引用@ozemiiiiwhat is a good alternative for alcohol that makes you more social/low inhib but doesn’t make you feel like shit the next day ↗
@marclou2026-10-04♥210👁26.7K↗ 打开X
My little HYROX experiment made it to mainstream media 🤗 https://t.co/dfwfX8vs7O https://t.co/I4rAkEgJ56
@marclou2026-10-04♥550👁35.3K↗ 打开X
My little book reached 10,000 downloads 🎉 https://t.co/w0V0Qy2NsL
@marclou2026-10-04♥888👁124.6K↗ 打开X
yeah AI is nice, but have you tried going to the gym? https://t.co/NdBrh6M2Ga
▸ 折叠5条(转推/噪音)
转推2026-10-05 RT @marclou: rate my footer https://t.co/M4wqj1CpSn
转推2026-10-05 RT @trust_mrr: ACQUIRED ✅ 🎉 A SaaS productivity tool just got acquired on TrustMRR. 💰 Sold for $4K 🤑 $413 revenue last 30 days 📈 0.8x mul…
噪音2026-10-05 oops https://t.co/mu6J6e7qho
转推2026-10-04 RT @marclou: My little HYROX experiment made it to mainstream media 🤗 https://t.co/dfwfX8vs7O https://t.co/I4rAkEgJ56
转推2026-10-04 RT @marclou: My little book reached 10,000 downloads 🎉 https://t.co/w0V0Qy2NsL
Brett@BrettFromDJindie171.0K粉 · 2条引用Higgsfield Ads Studio,关注广告生产从17个工具收敛到一个Agent入口。
@BrettFromDJ2026-10-06♥9👁5.1K↗ 打开X
ChatGPT + Higgsfield after watching us use 17 different tools to make one ad. https://t.co/tCXdDL28J4
引用@higgsfield_aiIntroducing Higgsfield Ads Studio. The first ROAS-driven static ads content machine. > Does 10,000 hours worth of deep research of your brand > Monitors real-time trending formats > Crafts winning ads based on best industry practices Try free in … ↗
@BrettFromDJ2026-10-05♥307👁11.2K↗ 打开X
Totally obsessed with how this logo came out. https://t.co/Hr0nPIVjOv
Tibo@tibo_makerindie211.2K粉 · 4条讨论Agent将成为主要用户,也展示从长期投放的竞品广告中提炼角度的Revid流程。
@tibo_maker2026-10-06♥9👁1.1K↗ 打开X
I agree with the other Tibo "the majority of things are going to be used by agents" https://t.co/OpR52SAsAv
@tibo_maker2026-10-05♥57👁6.0K↗ 打开X
my AI clone has gone rogue https://t.co/Sc4HSlCuoo
@tibo_maker2026-10-05♥55👁6.8K↗ 打开X
I figured out the easiest way to create viral AI ads the best strategy is to find what's already working in your competitors' ads and remake it for yourself 1. search your competitor's name across the Meta, TikTok, Google and LinkedIn ad libraries, all from one place inside Revid 2. filter for active video ads and set "running for at least" to 30 or 60 days, because nobody keeps paying for an ad that isn't working 3. open any ad to see the full video, the exact ad copy, how long it's been running, which platforms it's on and which countries it's targeting 4. hit remix and pick "video ad like this" for the same angle with your product and an original script, or "ugc ad from this copy" for a creator-style take on the same message that's it 🤷🏻 you skip months of testing hooks, visuals and angles, because your competitors already paid real money to figure it all out
@tibo_maker2026-10-05♥295👁38.0K↗ 打开X
just compared 18 places to build a startup. France ranks 17th 😬 with €100K profit, you get: Dubai: €99,187 Singapore: €91,796 Zug: €85,514 Lisbon: €71,488 San Francisco: €67,183 Paris: €58,757 in France the debate is always the same: "with this tax, will the millionaires leave?" I think it's the wrong debate. the millionaires will be fine. the real question is where the next generation of talented 25-year-olds will build their products. they go everywhere they want no strings attached I'm French and I'm staying but I want France to be the place founders move to made a free tool so you can run your own numbers 👉 should i be https://t.co/uQDZjKsFh7 ???
引用@levelsioJust a day later and now Javier Milei is offering full Argentinean citizenship with residence and passport for $350,000 Countries will keep plucking out the high skilled and wealthy from around the world for better conditions elsewhere Which is easier than … ↗
▸ 折叠3条(转推/噪音)
转推2026-10-05 RT @pukerrainbrow: Filip Panoski (@FilipPanoski) spent seven years building five different products with zero traction, before quitting a $…
转推2026-10-05 RT @FilipPanoski: I went from $1k to $15k MRR without adding a single new marketing channel. I just fixed my funnel. here's how: track 4…
转推2026-10-05 RT @vovalukashov: Most of my X work now runs through the SuperX MCP in Claude Code. It reads my mentions and finds fresh intro posts for m…
Adam Lyttle@adamlyttleappsindie59.1K粉 · 7条分享低流失应用截图、虚拟卡折扣机制和猫咪游戏命名冲突,也强调AI游戏仍需愿景。
@adamlyttleapps2026-10-06♥140👁9.7K↗ 打开X
I’ve never had an app like this before Churn? What churn?! https://t.co/FCSkXLoGnK
@adamlyttleapps2026-10-06♥6👁2.2K↗ 打开X
There's this app called EatClub It's basically just a discount restaurant app. But with a difference: You add their virtual card to your wallet and pay with that. Like a proxy card. You only get charged the discounted amount And apparently the restaurant gets the full amount Interesting use of virtual wallet/cards
@adamlyttleapps2026-10-05♥19👁3.7K↗ 打开X
Going to need to rework the name of this There's a "Cozy Cat Room" game which could get confusing https://t.co/fbpZl7S4QY
引用@adamlyttleappsCozy Cat Home trailer 2 just dropped… https://t.co/mBP2N1dbzO ↗
@adamlyttleapps2026-10-05♥45👁3.4K↗ 打开X
SEO is free marketing (and it's evergreen… mostly) https://t.co/kSX2OMNXYY
一句话强调SEO具有免费和相对长尾属性,缺少具体案例。
@adamlyttleapps2026-10-05♥37👁3.5K↗ 打开X
Cozy Cat Home trailer 2 just dropped… https://t.co/mBP2N1dbzO
引用@adamlyttleappsI’m creating a cozy cat puzzle game Care for your cats, play puzzles, decorate your house and unlock new ways to play Coming soon https://t.co/45KrCu1t6N ↗
@adamlyttleapps2026-10-05♥21👁3.4K↗ 打开X
Without a clear vision AI creates great tech demos This is most noticeable in games You can one shot an impressive game engine in an afternoon. But is it actually fun to play? Yes, it might be fun for *you*. But that's because you're the one creating it. It's like building something from Lego. The fun is the creative part. The end result is a byproduct. But you give that same Lego creation to someone else and they're going to say "yeah, that looks pretty cool" they may even ask how you built it You need to really sit with your building process. Work out what makes it fun, what makes it good, you need to keep iterating with a goal. A vision. Not just tinkering for the sake of tinkering
@adamlyttleapps2026-10-05♥286👁15.4K↗ 打开X
Now $600MRR and still growing This app is now on autopilot. I haven’t made any changes to the app itself or the marketing strategy in over a month And it’s replicable Making a video on exactly what I did Stay tuned 👀 https://t.co/BoTYlDxHdk
引用@adamlyttleappsJust reached $500 MRR In the middle of a week long digital detox. But just had to share the latest update 🎉 https://t.co/lkoJpEEnP2 ↗
Jon Yongfook@yongfookindie173.1K粉 · 7条系统梳理AI时代独立开发机会,并警告一键生成型SaaS容易被模型原生能力吸收。
@yongfook2026-10-06♥4👁457↗ 打开X
Came across this product today. The site is an absolute masterclass in positioning. Says exactly what it does, who it’s for, social proof from real people in their niche, etc. Study and apply learnings to your category! https://t.co/Rwkgod1dNR
称一个网站在定位、受众说明和垂直社会证明上值得拆解学习。
@yongfook2026-10-06♥43👁6.0K↗ 打开X
If your value prop is basically “we help you generate X for your business in just one click” it’s going to get eaten by AI. This describes a huge middle ground of SaaS. Not to say there isn’t money to be made short term, but long term it will be native to frontier models.
引用@higgsfield_aiIntroducing Higgsfield Ads Studio. The first ROAS-driven static ads content machine. > Does 10,000 hours worth of deep research of your brand > Monitors real-time trending formats > Crafts winning ads based on best industry practices Try free in … ↗
判断一键生成型SaaS长期会被前沿模型原生功能吸收。
@yongfook2026-10-06♥239👁20.0K↗ 打开X
Rewrite all your apps in Rust. Go into debt if you have to.
@yongfook2026-10-05♥48👁4.7K↗ 打开X
A broad trench dug around a medieval castle is the real moat.
@yongfook2026-10-05♥269👁18.5K↗ 打开X
How to make money as an indiehacker in 2026 and beyond? the opportunities seem to be shrinking fast. I've been thinking about this a lot lately. AI is eating whole categories of the SaaS industry. What would I work on if I was starting from scratch? Here are some random thoughts. 1) Software primitives Some of the basics aren't going away. Humans use them to get a job done, or humans use AI to interact with the tool. The AI needs the tool and can't natively do what the tool does. Newsletter managers/deliverers, form builders, analytics software. If the category has existed for 20 years, it will continue to exist. The question here is can you get traction for a product in an extremely saturated space. Maybe there are opportunities in niches. 2) Automate your specialized service This is probably the best time ever to productize a specialized service. Take a domain / service you're an expert in, which previously would require unscalable 1-to-1 attention with a client, and agentify it with proprietary data/skill.md etc. 3) AI value arbitrage There is a short term opportunity in taking things we see at the cutting edge of the X bubble, and packaging it for the non-AI-pilled. Opus 5.5 motion videos would be a recent example. We all know this is possible with AI now, but maybe there's a vehicle to package this for normies in a way that blows their minds and creates value for them. This relies on there being a knowledge gap though, which will shrink fast. 4) Just Do Hard Things (TM) Do things that are genuinely hard, that you would never be able to build before. Build things that previously only a VC-funded company would be able to do, with hurdles from all angles not just tech, but also go to market, regulatory, etc. Doesn't guarantee success but it at least keeps competition away. 5) make a social media scheduling tool just kidding about that last one.
把独立开发机会分为软件基础件、专家服务Agent化、短期能力套利和真正困难的事。
@yongfook2026-10-05♥162👁20.8K↗ 打开X
I built an AI slop machine over the weekend. I will never launch it, after seeing what I created. You put in a story prompt, and it generates a more structured story with stakes, cliffhangers, a beginning, middle, end. Splits it into episodes. Storyboards the whole thing, makes a cast, music, poster, voiceovers. Then generates the visuals (different styles available from anime to cinematic) and you can play the episode as a slideshow, or render it as a video. Uses all the latest models that you would expect. A few observations from this exercise: 1) it was trivial to build There's no moat here. This is just "AI glue", building an app that wraps several AI models. The moat is the GTM strategy, which leads me to... 2) this space is insanely competitive there's tons of AI movie makers all using the same models with some thin layer of prompt direction or UX as the differentiation. totally saturated. 3) AI video courses I noticed that some of the biggest advertisers in this space are people selling "how to create your own AI movie" courses. There's something off about this, almost dropshippy. It suggests that there's more value in hyping the thing and selling courses about the thing, than there is in the thing. 4) this shit is really expensive All of those cool AI videos that you see using the latest video models, are expensive to make. A 60 second "mini episode" of a story built using my platform would cost $5 in raw inference. There's no business here, whatever margin you put on top of that can just get eaten by a competitor. And imagine if a user buys $100 of credits (of which $90 is passed through to your AI provider) then decides they hate the movie that was created, and asks for a refund or charges back. Enough of those in a month will bankrupt you. 5) just didn't feel right It's a fun little demo that got some laughs from my friends but ultimately it's contributing to the never ending pile of AI slop. Is that something I want to be part of? Not really.
@yongfook2026-10-04♥131👁13.4K↗ 打开X
There’s a lot of doom and gloom on X but I’m still seeing the old ways work. Announce something people need, get payment commitment, validate and launch new product. AI makes this whole cycle faster than ever.
引用@Shpigfordone week in and all sales done through twitter. unreal. https://t.co/lXU3pBMRlT ↗
Josh Pigford@Shpigfordindie75.5K粉 · 9条自报AI SEO服务10天接近40万美元ARR,并继续公开定价、需求和X私信成交过程。
@Shpigford2026-10-05♥102👁13.9K↗ 打开X
this has continued to absolutely blow up (at nearly $400,000 ARR in 10 days). doing a last round of early access at $1500/mo. once i fully launch later this week, price goes up to $2k/mo. DM me if you're interested in me handling all your company SEO!
引用@ShpigfordDoing a little experiment. Looking for 5 software companies to run SEO for, 1-on-1. I built an AI SEO system that runs on my own products every day. Now I want to run it on a few sites that aren't mine. What you get each month: • A keyword map of your … ↗
作者自报AI SEO服务10天接近40万美元ARR,并准备继续提价。
@Shpigford2026-10-05♥114👁19.3K↗ 打开X
josh only drops gold.
引用@joshpuckettIntroducing Graphical. It's a simple, powerful, and fun tool to design visual languages, style components, and work with coding agents to create interfaces that look memorable and feel unique. I hope you'll check it out at https://t.co/crX3W8hHCz! … ↗
@Shpigford2026-10-05♥206👁21.3K↗ 打开X
i've always kinda turned my nose up at the typical growth/money/business guru types, but for whatever reason youtube started feeding me @AlexHormozi stuff lately and sweet baby jesus...i can't stop watching. yes, great business tactics i'm picking up, but his ability to distill problems while also showing an enormous amount of empathy for folks struggling is a master class and dammit just feels good to watch.
@Shpigford2026-10-05♥81👁8.4K↗ 打开X
24 hours later. still no website. i should probably work on that. https://t.co/sTQUJkQhVU
引用@Shpigfordone week in and all sales done through twitter. unreal. https://t.co/lXU3pBMRlT ↗
@Shpigford2026-10-05♥108👁10.2K↗ 打开X
what building a multi-million dollar SEO/AEO agency looks like. (faux company names/commits for privacy purposes)
@Shpigford2026-10-05♥12👁7.1K↗ 打开X
if the .com isn't available, what we going with?
@Shpigford2026-10-04♥94👁7.8K↗ 打开X
"josh, what's it like putting your phone number on your website?" https://t.co/7cQwsaV0Iw
@Shpigford2026-10-04♥246👁31.4K↗ 打开X
one week in and all sales done through twitter. unreal. https://t.co/lXU3pBMRlT
引用@Shpigfordgahhhh i can't wait to show you all this next week. based on current demand, i'm about 97% certain this thing will be at $1m ARR by end-of-month, especially after i flip the switch. absolute madness how fast it's already growing just via twitter DMs. … ↗
@Shpigford2026-10-04♥58👁8.0K↗ 打开X
only reason i'm not showing the new brand yet is i'm terrified the premium domain purchase will fall through (has happened before) and i'll need to pick something else. 😭 https://t.co/hPrhIp6wxX
引用@Shpigfordgahhhh i can't wait to show you all this next week. based on current demand, i'm about 97% certain this thing will be at $1m ARR by end-of-month, especially after i flip the switch. absolute madness how fast it's already growing just via twitter DMs. … ↗
▸ 折叠2条(转推/噪音)
转推2026-10-05 RT @IanLandsman: Nobody will pay for software/services, they said. Businesses will just tell their bot to 'fix seo' they said. https://t…
转推2026-10-05 RT @zachmstuck: Wife: when are you going to stop buying domains and start an actual business with one of them? Me: https://t.co/6zXi9GE6C9
Adam Pietrasiak@pie6kindie39.2K粉 · 1条只发一条Screen Studio悬停效果同步修复,继续深挖录屏细节。
@pie6k2026-10-04♥193👁9.8K↗ 打开X
Screen Studio mouse hover effect desync finally solved! It checks which mouse movements cause hover effects (e.g., changing the label background). Then, it dynamically adjusts mouse movement smoothness to make sure the mouse position matches the footage during those moments. https://t.co/5tflE5c4HK
Arvid Kahl@arvidkahlindie212.3K粉 · 6条讨论即时生成的上下文UI、Agent委托与手艺流失,并追问非厂商锁定的后台Agent运行方式。
@arvidkahl2026-10-05♥123👁6.5K↗ 打开X
You know these Star Trek LCARS interfaces that always looked like they were specifically designed to answer a plot-driven question visually and were never the same, always looked haphazardly constructed in that moment? I think, yet again, Star Trek understands the user interfaces of the future better than we do in the present. I think we are a very short time away from applications designing and assembling UI in the moment versus having it pre-designed. Contextual interfaces won't just consider what they present, but also who they're showing it to (on a highly (inter-)personal level) and when those people need which information in their workflow.
设想按人、任务和时机即时组装的上下文界面。
@arvidkahl2026-10-05♥122👁8.8K↗ 打开X
This is extremely insightful.
引用@nabeelquThe only generation that knows how computers work is also the most relaxed about AI: https://t.co/aMBh5gqLif ↗
@arvidkahl2026-10-04♥101👁14.5K↗ 打开X
Let's say I want to spawn a small team of agents with a job-specific Markdown config, allow them to use a defined number of gated APIs (and public web search), and then run autonomously doing background work. What's a non-vendor-lock-in way to run these today?
@arvidkahl2026-10-04♥133👁14.5K↗ 打开X
The more I think about this, the more I feel I am developing cognitive dissonance. Delegation is becoming available to those who never intended to manage (rather than do the work). Coders, designers, artists. Each hates and loves this tech for different reasons. Their lament is one of atrophy, of losing connection to the work. Yet delegation is also a mode of power. It's seizing the means of production and, for lack of a less loaded term, exploiting GPUs to do the work instead. Call it quantum leaps in increased productivity, slop cannons, or anything in between... but it is ultimately a move of empowerment. So we're gaining power by not having to do the thing we're trained to do, and in so (not) doing, we lose the skills we still value so much more than being able to conjure work from thin air. Maybe wizardry IS a good analogy here. We're summoning real-world products into being, works of fiction, images, videos... ethereal agents, familiars, are doing our work for us. Does the wizard who conjures fire with the flick of a wand regret not having to buy matches?
引用@karakhanyanSAI laziness is true. You start to delegate code. Then you think “let me try to delegate browser actions”, and you end up delegating everything. It’s fine at one angle. But you become lazier, and even dumber, you don’t learn new things, like how to create … ↗
@arvidkahl2026-10-04♥18👁6.5K↗ 打开X
"We're currently growing v4.2, and MCP support will soon hatch as well!" I kinda like it.
引用@GeoffreyHuntleywhat if software is no longer constructed and instead bred, similar to how llms are raised. through rewards and punishment. ↗
@arvidkahl2026-10-04♥86👁6.8K↗ 打开X
If your agents mess up your code a lot even though you have good docs, start interlinking them. Any md file that explains any part of your codebase should be linked to from a main readme or AGENTS/CLAUDE file. Allow things to be discovered semantically and quality will rise.
▸ 折叠1条(转推/噪音)
转推2026-10-05 RT @andruyeung: We created gyms because modern day work no longer required physical activity. Without exercise, our muscles atrophy. I pre…
Peter Askew@searchboundindie26.2K粉 · 1条继续展示买垂直域名、卖实体商品并把旅行纳入业务的domain-first打法。
@searchbound2026-10-04♥133👁10.3K↗ 打开X
🐖 buy expiring domain 🐖 sell chorizo sausage on internet 🐖 fly to Spain & Portugal as biz expense 🇵🇹 🇪🇸 🐖 hike the camino de santiago 🐖 read some josé ortega y gasset [not an ad; just sharing] https://t.co/0IqlLxKxc6
Danny Postma@dannypostmaindie186.6K粉 · 2条为AI协作落地页课程做提纲与问卷校准,继续用预售验证内容需求。
@dannypostma2026-10-05♥80👁8.9K↗ 打开X
general outlines drafted ✅ just sent out a survey to confirm if the content aligns https://t.co/XGsnOprDuH
引用@dannypostma"Build me a landing page." That's how you get the same AI slop as everyone else. I'm making a landing page course for you and your AI agent to solve that ⚡ It hurts seeing good indie products launch with pages that don't do them justice. I've spent years … ↗
@dannypostma2026-10-05♥9👁2.6K↗ 打开X
working on the outline and structure and man this is gonna be fun
引用@dannypostma"Build me a landing page." That's how you get the same AI slop as everyone else. I'm making a landing page course for you and your AI agent to solve that ⚡ It hurts seeing good indie products launch with pages that don't do them justice. I've spent years … ↗
Thomas Sanlis 🥐@T_Zahilindie24.4K粉 · 4条称Writizzy自动把博客发到LinkedIn后首次半病毒传播,并继续记录Uneed日常。
@T_Zahil2026-10-05♥16👁787↗ 打开X
Trying something 👀 what do you think? https://t.co/LvFvryCmvC
@T_Zahil2026-10-05♥27👁990↗ 打开X
Last week I went semi viral on Linkedin for the first time, and I didn't even know I posted something lol Here's how I did it 👇🏻 A few days ago, I wrote a blog post (yes, by hand) about "Shipfast is dead" I wanted to talk about my experience growing Uneed, and seeing more and more generic products every in the queue So I wrote it on my blog, which is using Writizzy, my own SaaS And I forgot about one feature I turned on on my account: auto posting on social medias 😂 So Writizzy generated an excerpt of my post using AI, and posted it on my Linkedin acccount Writizzy might actually be a super useful growth product 🤩
作者称Writizzy自动分发手写博客后带来首次半病毒LinkedIn传播。
@T_Zahil2026-10-05♥53👁1.3K↗ 打开X
I've got the Uneed's sweater!! 😎 https://t.co/sizmKWUrQm
@T_Zahil2026-10-04♥41👁2.1K↗ 打开X
I managed to make "Pharaoh" work on my Mac 👀 Super cool to be able to play this again, huge part of my childhood! https://t.co/1govhzLYR7
Jonathan Wilke@jonathan_wilkeindie29.8K粉 · 6条强调删功能、修漏斗和以架构解释替代逐行审查,继续观察Cursor与Devin。
@jonathan_wilke2026-10-06♥14👁1.1K↗ 打开X
Pro tip: Before you add that new feature, ask yourself: Can I actually remove something instead?
@jonathan_wilke2026-10-06♥13👁918↗ 打开X
This
引用@stijnnoormanThe smarter you are, the more you value simplicity. ↗
@jonathan_wilke2026-10-05♥52👁4.6K↗ 打开X
In short: find out what your users need.
引用@FilipPanoskiI went from $1k to $15k MRR without adding a single new marketing channel. I just fixed my funnel. here's how: track 4 numbers every week: visitor → signup signup → trial trial → paid churn find the worst one. fix only that. then move to the next. each … ↗
@jonathan_wilke2026-10-05♥18👁3.3K↗ 打开X
I agree 💯 with Ben here. Not only will a line-by-line review become less needed, but it is also no longer the most effective way to verify your changes.
引用@BHolmesDevCode review is no longer mandatory. It’s now risk mitigation. Ask agents to explain the architecture of their changes to you. Ask questions. Have them ask *you* questions. Line-for-line PR review will become progressively less needed ↗
认为架构解释和问答会逐步替代部分逐行PR审查。
@jonathan_wilke2026-10-05♥14👁2.7K↗ 打开X
At this point it doesn’t even matter anymore @cursor_ai or @DevinAI are all you need
引用@jjackyi honestly have no idea what a software factory is and im afraid to ask at this point ↗
@jonathan_wilke2026-10-05♥23👁2.7K↗ 打开X
Funniest thing about this is that it took me over a year to realize that there is a cursor in the logo. Initially I just thought “nice logo, but I don’t get what the box has to do with the brand”
引用@pizzaboyI genuinely think cursors logo is one of the best logos of all time https://t.co/ZA0yrpep32 ↗
Marc Köhlbrugge@marckohlbruggeindie89.2K粉 · 1条用一条短评判断.ai域名的品牌价值可能随AI普及而退潮。
@marckohlbrugge2026-10-04♥34👁5.0K↗ 打开X
Looks like .ai might be going out of fashion sooner than we thought
引用@marckohlbruggeLong term I’m not a fan of .ai however. AI will be part of most software. Once it stops becoming a differentiating feature it loses its branding value. Same way we no longer use the words “internet” and “digital” in product names. ↗
Kyle Gawley@kylegawleyindie45.6K粉 · 4条发布从Reddit等社媒找客户的内容,同时坦承高频YouTube更新多数仍停在约200播放。
@kylegawley2026-10-06♥0👁150↗ 打开X
NEW Zero To Paid content just dropped 💣 How to find customers on Reddit (and other social media platforms). Easy mode! Link in the comments.
@kylegawley2026-10-06♥13👁1.4K↗ 打开X
Our maid is cleaning our house today. I could of course do it myself. Instead, I pay for someone else to do it so I can use my time on important tasks that move the business forward. The lesson here is no-one is vibe coding their own SaaS replacements.
@kylegawley2026-10-05♥27👁2.4K↗ 打开X
Youtube is hard AF I am publishing 3-5 times per week now and putting 10x more effort in and most views get stuck at 200 views
@kylegawley2026-10-05♥24👁1.7K↗ 打开X
Atrophic did $4.6 billion revenue in 2025. and LOST $42 billion. They've committed to spending $518 billion. Which model are they using for this math? Definitely not a business model.
▸ 折叠1条(转推/噪音)
转推2026-10-06 RT @thijsmakes: Pretty cool I just followed @kylegawley and @levelsio (MAKE book) advice, did nothing else and made my first internet money
Dan Kulkov@DanKulkovindie50.5K粉 · 8条记录第二单与YouTube投入,夹杂折扣玩法、习惯应用和多条自转发短梗。
@DanKulkov2026-10-06♥13👁580↗ 打开X
maybe the fastest way to hit $5000 MRR is to find a job
@DanKulkov2026-10-06♥3👁526↗ 打开X
💸 SECOND SALE 💸 LET'S GOOOOOOOOOOOOOO https://t.co/rO3r2fpj8N
@DanKulkov2026-10-05♥20👁2.9K↗ 打开X
vibe-coding another habit tracker because i don't want to go to jail https://t.co/szYM7LUJtB
引用@DanKulkovbro https://t.co/47otsx9ooU ↗
@DanKulkov2026-10-05♥13👁1.1K↗ 打开X
guess when i took youtube seriously https://t.co/DAj8gTyeGJ
@DanKulkov2026-10-05♥17👁1.4K↗ 打开X
marketing 101 let users gamble discounts https://t.co/YBCitIhwmm https://t.co/4dZYv08eEt
@DanKulkov2026-10-05♥17👁3.0K↗ 打开X
app idea: forget tinder find your soulmate through ratings https://t.co/brEW0by4yw
@DanKulkov2026-10-05♥10👁1.9K↗ 打开X
i might have an addiction https://t.co/v52brVSRA4
▸ 折叠7条(转推/噪音)
转推2026-10-06 RT @DanKulkov: guess when i took youtube seriously https://t.co/DAj8gTyeGJ
转推2026-10-05 RT @DanKulkov: marketing 101 let users gamble discounts https://t.co/YBCitIhwmm https://t.co/4dZYv08eEt
转推2026-10-05 RT @DanKulkov: app idea: forget tinder find your soulmate through ratings https://t.co/brEW0by4yw
转推2026-10-05 RT @DanKulkov: i might have an addiction https://t.co/v52brVSRA4
噪音2026-10-05 bro https://t.co/47otsx9ooU
转推2026-10-05 RT @DanKulkov: cancelled my $49/mo email marketing subscription switched to amazon SES time to become a man https://t.co/GfaGJaaVVs
Andrea Bosoni@theandrebosoindie64.2K粉 · 1条建议向营销人销售时少用套路,直接展示功能、价格和简单文案。
@theandreboso2026-10-06♥3👁233↗ 打开X
If you're trying to sell your product to marketers your landing page needs to be a bit different from what you'd usually do. Cut the BS. We can see through it. Use simple words. Features over benefits. Get straight to the point. Keep your pricing simple. Just show us what the product does and what it costs and let us decide. We spend all day writing the same tricks you're about to use on us so they just make us trust you less.
建议面向营销人时少用营销套路,直接给功能、价格和简单语言。

数据:Twitter/X(独立开发者+AI行业两个cohort)+Hacker News(前排/Show HN)+Reddit(当日top+热评)+GitHub新星。原文照登。Twitter转推及噪音共54条折叠在作者附录内;旧推不进正文叙事,仅出现在附录并带日期。