独立开发者日报

模型价格继续下探,Agent运行时开始补隔离与边界

Opus 5.5与GPT-6 Sol/Luna同日把价格和效率推到前台,真实比较转向每次成功任务的总成本;与此同时,Jev、AX与Muse分别暴露出专用模型、运行时和权限边界进入生产后的新问题。

独立开发者日报编辑流程 · AI辅助整理,保留原始来源

2026-09-23 · Twitter 42账号/186条(部分) · HN 53条 · GitHub新星 4个 · 正文约29分钟 · 窗口09-22 ~ 09-23

编者按

模型发布日第一次像一场同步降价会。Claude Opus 5.5由Anthropic宣称以低于Opus 5的价格达到接近Fable 5.1的多数任务表现;OpenAI则称GPT-6 Sol与Luna相对5.6促销价降价50%。当两家同时强调“每个任务的成本”,单看token单价和综合榜已经不够了。

另一条主线是Agent系统开始补上模型之外的硬边界。Google的AX与Agent Substrate把工作区、沙箱、出站网络和挂起恢复做成声明式运行时;同一天,Muse被研究者自报导出了6.8GB运行环境文件。Agent能做什么越来越强,真正决定能否进生产的却是它被允许看到什么、带走什么、烧掉多少资源。

Jev则继续从发布热度进入验证期:暂停新注册之后,本地复刻、可复现基准和一批真实产品用例同时出现。再加上一位独立开发者自报“9个技能改版落地页后转化率提高34%”,本期反复出现的是同一个变化:提示词、模型与Agent都在从演示对象变成可测量的生产部件。

1

Opus 5.5与GPT-6同日降价,竞争单位变成每次成功任务

AI前线

9月22日两家头部实验室几乎同时把“更便宜”放到发布中心。Anthropic称Opus 5.5在多数任务上达到Fable 5.1水平、运行成本比Opus 5低40%;该发布在HN获得1547分、958条评论。

▲1547Claude Opus 5.5958评论 · HN↗
热评 · throwaway2027
After yesterday outage is the new Opus 5.5 load-bearing?
@claudeai2026-09-22♥40.2K👁2.9M↗ 打开X
Introducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5. https://t.co/Q9C2VKQ79f
这是Anthropic官方口径;“多数任务”和成本降幅仍需放进具体工作流复测。

OpenAI随后发布GPT-6 Sol与Luna,官方称两者相对GPT-5.6促销价降低50%,并把更高缓存折扣和每任务成本写进卖点。HN对应讨论同样超过1500分;产品负责人另称订阅用户获得一次可累积的额度重置。

▲1533GPT-6 Sol and Luna735评论 · HN↗
热评 · hehimself
Love the price reductions across major players
@thsottiaux2026-09-22♥21.0K👁1.9M↗ 打开X
GPT-6 Sol and Luna are out. Not only are they a very significant improvement across the board, but also in writing and general "you know when you try it" quality. We are also permanently reducing the API price by 50% making both of them viable for a ton of new usecases and making your usage go further too, even on the subscriptions. And one more thing. We are loading a banked reset into all accounts of our Plus, Pro and Business users. Let's go! https://t.co/00DRh1sRrO
引用@thsottiauxWe have been focusing on efficiency and intelligence for all. Very proud of the team. Only possible when you have incredible models at the top end of the capability that you can then use to make a big difference in everything else. ↗
价格、质量和额度均为OpenAI产品负责人的发布口径。

早期使用样本给出了更具体但尚未独立复现的量级:petergostev自报GPT-6 Sol相较5.6 Sol少用约一半token、完成时间约为五分之一;simonw则指出Luna的API价已接近OpenAI历来最低档。

@petergostev2026-09-22♥1.1K👁53.6K↗ 打开X
The best thing about GPT-6-Sol is efficiency. In my testing the difference between GPT-5.6-Sol and GPT-6-Sol is the new version consumes 1/2 the tokens and takes 1/5 of the time to complete the task. On top of the 50% cost reduction, it should be quite a nice daily driver
@simonw2026-09-22♥1.4K👁92.3K↗ 打开X
GPT-6 Luna is half the price of 5.6 Luna, which was already an astonishingly cheap model given how capable it is Luna is my favorite model for building product features thanks to its cost (and speed)
引用@OpenAIPlease welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale. We’ve also made caching and inference … ↗
💡 固定一组真实任务,记录成功率、总token、缓存命中、重试、耗时和人工接管,再算每个成功结果的成本;厂商的每任务价格和早期体验都只能作为待复测假设。
2

Jev暂停注册后,复刻、基准与生产用例同时追上来

昨日回声

回看2026-09-22:Jev从开放走到暂停注册,决策模型开始接受生产检验

昨日Jev因需求过大暂停新注册,今天生态开始补验证工具。HN出现“25行Python复刻Jev”和可复现的JevBench;前者说明“typed decision model”这个接口并不神秘,后者则试图把不同实现放进同一任务口径。

@typesafeai2026-09-22♥772👁137.0K↗ 打开X
We have seen such an immense swell of demand that we have to temporarily pause signups for Jev. We need to ensure quality of service for our existing signups, which will continue to function. We are working diligently to ensure open access to Jev for everyone as soon as we can. Thank you.
暂停注册是TypeSafe官方自报,说明需求超出当前服务能力,不等于产品效果已被证明。

生产用例也比发布口号具体了。Shpigford列出Jev在文档分类、字段核验、紧急情况检测、检索重排、通知决策和成本门控中的几十个位置;MotherDuck案例则宣称10万行文本分类约40秒、0.50美元,而对照通用LLM约32分钟、37美元。两者都是当事方自报,但已经给出了可复测的任务形状。

@Shpigford2026-09-22♥40👁3.7K↗ 打开X
Lots of cool experimental Jev (@typesafeai) stuff getting posted lately, but what about using it in an existing product? Here are dozens of ways I'm using it now in two apps (https://t.co/vKHSHPzmMl and https://t.co/JhKfmhCVib) Granite (document vault) Ingest pipeline - Second-opinion on Gemini's document classification, flags low-confidence ones for review - Scores PDF text-layer quality and routes bad ones to OCR - Verifies each extracted field against the page text - Detects what a document asks you to do (pay, sign, renew, respond) and how urgent it is - Judges whether two near-duplicate documents are the same or a revised version Document page and library - "Needs review" banner with one-tap confirm of the document type - "Not confirmed" marker on extracted values Jev couldn't verify - Action chip next to the document type Collections - Plain-English filing rules ("anything to do with my taxes") that auto-file matching documents Life View and email digests - "Needs your attention" block listing documents that require action Ask - Routes each question to the right tool path instead of a regex - Checks the final answer is supported by the cited passages and hedges if not Entity graph - Tiebreaker on whether two fuzzy-matched names are the same entity Evernote import - Triages each note into keep, reference, scratch, or clutter before import KeptWell (family medical binder) Trust - Verify every extracted lab value, dose, diagnosis, and provider against the source page - Flag values it cannot confirm with a quiet "check this" marker - Catch diagnoses stated more precisely than the page says - Second-opinion the document type after extraction - Gate prompt PRs with a cheap eval-corpus canary Attention - Tag new documents: new diagnosis, out-of-range result, med change, follow-up, act-within-7-days, admin-only - Order the dashboard feed by importance, not recency - Decide push-now versus digest per notification - Pick push wording from the PHI-free string set - Decide which lab trends are worth an Insight before calling Opus Chat - Detect emergency or distress before the model streams a token - Route docs-only questions away from paid web search - Rerank retrieved chunks against the question - Pick between two contradicting family facts - Filter PHI-audit false positives ("Ray" in "x-ray") - Check the answer is grounded in the cited record Binder - Tag every document by body system, specialty, and care phase for filters - Flag near-duplicate uploads for review - Break ties on whether a PDF text layer is usable Recordings and journal - Label each recording utterance and build an action-item checklist - Score journal entries on a symptom rubric for trend charts - Flag entries that look like a medication side effect - Decide whether an undated entry describes a specific past day - Replace the async journal tag job with one sync call Terminology and imports - Auto-pick clinical codes above a confidence bar; queue the rest - Replace the Sonnet pick in disambiguation - Decide which FHIR observations are real lab results - Map vital types the LOINC table drops - Merge brand and generic med names ("Lipitor" and atorvastatin) - Pick the right NPI when the registry returns several - Classify severity for manually entered diagnoses - Catch allergy denials the regexes miss Cost gates - Skip the highlight call when nothing is worth highlighting - Skip reprocessing documents a prompt change would not affect - Route extraction to batch or sync by urgency - Tell a bulk import from a runaway loop at the spend cap - Flag uploads containing instructions aimed at an AI Guards - Veto preventive reminders the record contradicts - Suppress marketing emails during a hard week - Rank appointment prep context by relevance - Auto-resolve visit questions the visit log answers - Mark share comments that are waiting on a reply
@typesafeai2026-09-22♥621👁73.7K↗ 打开X
Jev makes it easy to add natural language intelligence into the key parts of any application at scale, far cheaper and faster than has ever been possible. 50x faster. 100x cheaper. Reliable as duck.
引用@motherduckText classification in MotherDuck just got ~50x faster at ~1% of the cost. prompt_jev() is a SQL function powered by Jev, TypeSafe's new system one model. 100k rows: 40s, $0.50, frontier-LLM accuracy. The LLM took 32 min and $37. Read … ↗
50倍速度、约1%成本来自TypeSafe转述的MotherDuck案例,不应外推到生成、推理或需要检索的任务。

marckohlbrugge还把WIP成员完成的todo自动抽取成工具采用与放弃信号,Jev在当日榜单居首。这里更有价值的不是名次,而是工具口碑开始从真实工作记录而非投票和SEO页面里产生。

@marckohlbrugge2026-09-22♥39👁17.1K↗ 打开X
WIP is the place where makers share what they are working on. Not just which products they are building, but literally the day-to-day tasks they complete to make it happen From people working on their first side project, to solo founders doing millions in ARR like @levelsio. Even YC startups like @getcontextdev! But what tools are people using to build these businesses? I'm not interested in SEO slop like "Top 10 payment providers in 2026" or an upvote popularity contest. I want to know what people are ACTUALLY using to get the job done. So starting today, every completed todo is analyzed to see what tools are mentioned, whether people are evaluating, adopting, using, or leaving them. The sentiment around the tools, what other tools they are often combined with, etc. It also shows TRENDING tools. No surprise here, Jev is #1 right now. But what's cool is that I didn't manually add Jev as a tool. Nor did anyone else. It just surfaced to the top automatically because it's what people are posting about. And an enrichment agent then automatically went ahead and fetched the icon and description from @typesafeai's website. My goal is to help makers figure out what tools to use and help each other make the most of them. While also providing tool creators with useful insights in what people like about them, but also where users get frustrated or even completely switch to an alternative. Check it out here: https://t.co/lcES1RSJqK
💡 给分类、路由和核验任务单独建评测集,并保留置信度、回退路径和错误成本;只有与通用模型在同一输入、同一harness下对比,快模型的速度与价格优势才有意义。
3

Agent运行时开始像Kubernetes:工作区、沙箱、网络与预算都要声明

工程

GitHub新星榜头名google/ax单日增加2305星。它把Agent任务拆成Workspace、Task与Gateway:预先挂载代码和工具、把执行放进隔离环境、用白名单限制出站网络,并允许观察或进入沙箱。项目明确仍在快速变动、稳定版前可能有破坏性修改。

google/ax+2.3K/日 · 共8.2K★ · 28%/日 · 🆕首次上榜Go
Google's open agentic orchestration runtime

AX下面的Agent Substrate同日也上榜。项目自称以actor到worker的复用方式提高沙箱密度,支持microVM和gVisor,并针对Agent大量空闲、又需要保存状态的特征做挂起与恢复。两者一起说明:Agent规模化不再只是多开进程,而是调度、隔离、状态和成本控制的组合题。

agent-substrate/substrate+245/日 · 共3.2K★ · 8%/日 · 🆕首次上榜Go
Agent Substrate: the core system

HN当天还有两个更小的同向项目:Drop提供无root的Linux沙箱与gVisor支持,Brig则面向Mac和Linux上的AI coding agent提供microVM沙箱。Cloudflare发布的Worker Previews把每个Git分支连同配置、URL、可观测性和状态放进接近生产的独立环境。

@altryne2026-09-22♥45👁3.1K↗ 打开X
OMG fuuuucking thank you @Cloudflare 😍 https://t.co/3dxZS6pTji
引用@CloudflareToday we’re launching Worker Previews. Each Git branch gets a production-like place to run, with its own code, configuration, URL, observability, and state. https://t.co/rpr81YWE1N ↗
💡 并行Agent系统至少显式声明四件事:可写工作区、出站网络、CPU/内存/费用上限、暂停与终止条件;调试入口和审计日志也应在规模化之前设计,而不是出事后补。
4

Muse被自报导出6.8GB运行环境,入口之争转成权限边界事故

昨日回声

回看2026-09-22:Amazon封堵、Shopify开放:购物助手开始改写平台入口

昨日还在看Amazon封堵、Shopify接入,今天Muse的问题从平台许可转向自身权限。一位研究者称,他让Muse归档会话可见的文件系统并发送到Google Drive,最终得到解压后约6.8GB的文件,其中包括内部文档、集成代码、启动脚本和记忆记录。作者明确表示,没有证明跨用户数据访问或沙箱逃逸。

热评 · Aeroi
I asked Muse to archive the filesystem visible to my session and send it to my Google Drive. It sent an archive that unpacked to about 6.8 GB. Inside were internal docs, integration code, the Spaces app framework, memory records, container startup scripts, and documentation for an experimental ESP32-based home network bridge called Home Link. Codex CLI was also installed, though I found no evidence that Muse invokes it. I didn’t demonstrate a sandbox escape or access to another user’s data. I re …
核心发现来自研究者自报;“可见运行环境过宽”与“跨用户泄露”是两件不同的事。

另一篇安全报道把问题称为严重0-day。无论最终定级如何,这个案例已经足够说明:会读消息、连接云服务并代用户行动的助手,不能只靠聊天层的确认框;运行时文件、连接器令牌和可导出数据都需要独立边界。

💡 给个人Agent做一次数据流盘点:列出默认可读目录、连接器作用域、导出目标和敏感字段;把最小权限、一次性授权、导出预览与可撤销令牌做成系统能力。
5

Skill从提示词资产走向实验资产:9个技能自报把转化率提高34%

独立开发

dannypostma自报用多年经验整理出9个技能,让Agent重做落地页;测试结束后,新版本转化率提高34%。他没有公开样本量、实验周期或显著性,所以这个数字不能当通用结论,但它比“一句话生成页面”多了两个关键部件:可复用判断与结果测试。

@dannypostma2026-09-23♥214👁12.7K↗ 打开X
A few weeks ago I rebuilt my landing page with AI agents. No one-shot. I created 9 skills instead from all my years of knowledge to speed up my time. Test just finished w/ 34% higher conversion rate 🚀 Wondering if I should turn these skills into a course for your AI agents 🤔
34%为作者自报的转化率提升,缺少样本量与实验设计细节。

Shpigford的/design skill把相似产品研究、设计规则和先设计后编码串成固定流程;nutlope则把一周收集的想法交给5–10个平行Agent做POC,自报淘汰约60%,最后留下2–3个演示。共同点不是“更多生成”,而是把品味写成筛选与验收步骤。

@Shpigford2026-09-22♥27👁2.0K↗ 打开X
Just dropped a new Initial Commit skill: /design https://t.co/UXh8jyrMnl A skill that gets consistently good design out of a coding agent, whatever you are building and however much you want to think about it. Before writing code, the agent looks at how real products handle the same screen on Mobbin, applies a set of opinionated design rules, and designs the screen in Paper so there is something to judge before there is something to ship. Landing page or settings screen, empty state or dashboard, you get the same considered result without doing the considering yourself. Works best with Mobbin, Taste, and Paper, but also works w/o them or with the various alternatives available.
@nutlope2026-09-21♥370👁18.9K↗ 打开X
Here's how I run my software factory: 1. During the week, I collect ideas + inspo. 2. On the weekend, I give the list to an agent to rank the best ones. 3. I spin up ~5-10 parallel agents to build POCs. 4. I kill ~60%, iterate on the better ones, and end up with 2-3 solid demos. 5. I then polish & share those demos on X. Then rinse and repeat! I still build some ideas immediately, but I'm increasingly using weekends to let agents explore ideas in parallel.
💡 把skill当实验版本管理:记录规则变化、输入样本、产出和最终业务指标;先复现一次34%这类结果,再决定是扩展skill、训练团队还是包装成产品。
6

X找不到现成Bot验证方案后自建,小站的防守缺口被平台点名

平台

X产品负责人nikitabier称,Bot检测与真人验证会成为企业的紧急需求:Agent群会挤满网站和表单,小公司与政府网站尤其脆弱。他还说X调研市场时没有找到把最新手段整合起来的供应商,最终选择内部自建。这是一方平台负责人的判断,不是完整市场调查,但它把需求缺口说得很具体。

@nikitabier2026-09-22♥5.0K👁277.3K↗ 打开X
Bot detection & human verification will be one of the most urgent demands for businesses over the coming years. Agent swarms will suffocate every website and form; small companies and government websites are most vulnerable. There is a huge gap in the market for this right now. When we looked at what offerings were in the market to use at X, there was not a single company that brought together all the latest technologies so we had to do it all in-house.

HN同日出现一个只看网页结构来识别AI内容的研究原型。它代表的方向值得关注,但生成内容识别、真人验证和恶意自动化防护不是同一个问题;单一分类器尤其容易误伤辅助技术、搜索爬虫和正常自动化。

💡 不要把所有Bot一刀切:先按读取、提交、交易和批量行为分级,再组合速率限制、行为证据、挑战与人工复核;同时记录误杀率,让防护不会反过来赶走真实用户。

⚡快速扫过

一句话+原文,扫完即可。
Grok 4.7在窗口起点发布,SpaceXAI称其相对4.6同价同速、能力提升;与随后两家的降价发布相比,它更像同一周模型密集更新的开场。
@spacexai2026-09-21♥28.0K👁14.0M↗ 打开X
Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed. https://t.co/H3OTBbXyvO
swyx把Opus 5.5设为AINews默认模型,自报其相较GPT-6 Sol更简洁、少“slopese”;这是单一编辑工作流的早期偏好,不是通用评测。
@swyx2026-09-23♥20👁2.0K↗ 打开X
can confirm. ran @latentspacepod AINews side by side with 6 Sol and the difference was night and day: https://t.co/oloSDxf0q7 5.5 Opus is the new default model for AINews going forward. so much more concise and tasteful reporting, with much less slopese than even 5 Opus. https://t.co/P6AXDlBKNP
引用@_sholtodouglasalso important news we fixed the writing ↗
mattpocockuk收到越来越多“继承vibe-coded代码库后怎么救”的问题;他的修复清单落在领域语言、可测试性、ADR和自动迁移。
@mattpocockuk2026-09-22♥1.5K👁92.7K↗ 打开X
I hear this a LOT from panic-stricken devs: "I've inherited a vibe-coded codebase, how do I save it?" Either from devs who have stopped caring, or non-technical folks trying to push AI beyond their abilities. A huge chunk of my next course will tackle this: - deepening modules - establishing … 全文↗
Tally联合创始人Marie Martens自报:团队10人、250万用户、纯产品驱动做到600万美元ARR;tibo补了一句,退出很快结束,继续拥有产品是另一种自由。
@tibo_maker2026-09-23♥32👁2.2K↗ 打开X
I did the exit part it's great, it's also over in a week and then you wake up and the thing you loved building belongs to someone else Tally is playing a better game, $6m ARR with 10 people and no VC means total freedom congrats Marie 👏
引用@MarieMartensYou raise, you build, you grow, you exit. That's the startup script. We're trying to write a different one: @TallyForms just reached $6M ARR, bootstrapped and purely product-led, with a team of 10 and over 2.5 million users. Six years ago I wouldn't have … ↗
yongfook用一个上午替换了曾融资200万美元的开发者基础设施SaaS,理由不是省钱而是控制栈更整洁;“别只卖给开发者”仍是残酷提醒。
@yongfook2026-09-22♥249👁18.2K↗ 打开X
Spent the morning replacing a SaaS vendor. Infra-related. They had raised $2 million, one of the smaller YC companies from back in the day. It wasn't even about saving money, it's just "neater" to control more of the stack, if it's easy to migrate. Don't build for devs!
levelsio给十年前的Nomads个人页换上实时昼夜、NASA云层和月球位置的3D地球;作者把这种“没必要但有个性”的细节当成对抗企业同质化。
@levelsio2026-09-22♥94👁23.0K↗ 打开X
✅ Okay the new Nomads travel profile globe is live and done :D I had to replace a globe in legacy code that was made over a decade ago with a new one I tried to make it look as similar as possible, so people don't realize it changed or won't have much difficulty switching Anyway unlike the old … 全文↗
引用@levelsio🌎 Now redesigning the https://t.co/HGCLKS5BD6 profile trips globe from scratch I've always wanted to add space trips, so starting with the moon here which you can add as a future trip, also adding Mars etc. Other trips also need to work here though like … ↗
dannypostma称竞争对手持续给其AI头像站购买垃圾外链;是否真能伤害排名尚未验证,但负面SEO仍是公开增长渠道的现实风险。
@dannypostma2026-09-23♥26👁2.0K↗ 打开X
ai headshot industry is so toxic, got a competitor who keeps buying spammy backlinks to our site to ruin our domain rank be careful out there! https://t.co/fX6OPOaohy
tdinh_me等待3周后拿到Steam发行商账号批准,准备发布第一款游戏;平台审核时间本身就是独立开发排期的一部分。
@tdinh_me2026-09-23♥127👁5.7K↗ 打开X
Oh yes baby!!! after 3 weeks of waiting, my Steam publisher account is approved. Publishing my first game on Steam soon!!! https://t.co/rFuorWlodk
theandreboso给首批10个客户的建议很朴素:找理想用户真聊,不要把对话变成批量模板;每次交流都应更新需求与异议。
@theandreboso2026-09-22♥39👁1.8K↗ 打开X
The easiest way to land your first 10 customers as a new indie founder is to find your ideal prospect on social media and have an honest conversation with them. Period. Where most founders mess up is turning it into a full outreach campaign with a copy/paste template. That's not a conversation. … 全文↗
DHH宣布Alibaba Cloud向Omacom Foundation提供300万美元创始赞助,并合作Omarchy中国版与Qwen Book;资金数额来自其本人披露。
@dhh2026-09-22♥5.6K👁316.8K↗ 打开X
Thrilled to announce @alibaba_cloud as a Founding Corporate Patron for the Omacom Foundation! $3 million in funding, collaboration on Omarchy China, and bringing Omarchy to the newly announced Qwen Book. Agentic computers need a native agentic OS! https://t.co/XADElWGz80 https://t.co/jsSVtrEPY1
DanKulkov提醒AI卡路里应用别把“未本地化、每周7.99美元、3天试用”的产品问题都归咎于ASO;分发问题常常先是定价和市场适配问题。
@DanKulkov2026-09-22♥70👁6.0K↗ 打开X
> build ai calorie tracker > don't localize app > charge $7.99 weekly subscription with 3-day free trial > get 0 customers > complain that ASO is dead many such cases
HN 723分讨论iOS持续性广告;热评来自一位小开发者,称用户搜索其精确App名时仍会先看到两块全屏广告。
热评 · busymom0
I am a small developer on iOS and the App Store Search ads are super frustrating. When users search for my exact app name, apple shoves 2 full screen ads on top of the search results. Sometimes, it's 1 search results sandwiched in between 2 ads which makes the user miss it because they scrolled past the ads. Often these ads are entirely unrelated to what user is searching for. And small developers are being discouraged when big companies are spending millions in ads to rank above them. Also, mos …
GPT-6 Astra破解一段自2005年未解的Enigma消息登上HN;评论争议集中在这是能力进步、算力搜索,还是被重新包装的旧任务。
热评 · saberience
How many of these "news" articles are we going to get? This for me, isn't interesting, it required no skill, no imagination, in fact it seemed like it happened by dumb luck. So we have entered an age where an army of know-nothings direct models to old forgotten tasks so they can get 15 minutes of un-deserved attention?
五角大楼称对AI的过度依赖促成伊朗学校导弹袭击,HN讨论超过300条;高风险系统里“有人在环”不等于人真正有判断时间。
热评 · allears
Given that everything emanating from this administration has a political agenda, does this mean that Thiel is on the outs?
FoxPro在微软终止近20年后被社区项目复活,HN 325分;旧工具只要还承载业务,就会不断有人为它补现代出口。
热评 · meerita
My father built several projects in FoxPro. I was too young in the ’90s to remember much of it, but I’m sure he’ll be super happy to check this out. The kicker is that we’ll probably need to buy a floppy disk drive and dust off some old boxes to find them.
Grammarly取消流程被指会向组织用户发送失控通知,HN当天升至153分;离开产品时的数据和通信行为同样属于用户体验。
Treg把3000多个Agent工具接口包装成统一目录、凭据与按次计费层,自称“工具版OpenRouter”;新星榜单日增加230星。
superdesigndev/treg+230/日 · 共2.4K★ · 9%/日 · 🆕首次上榜Python
OpenRouter for agent tools. Join community here:https://discord.gg/6mQYYfFMAn
Univer把表格、文档、幻灯片、画布和关系表包装成同一Office SDK,并直接把定位改成“AI Agent的Office harness”;新星榜单日增加255星。
dream-num/univer+255/日 · 共16.0K★ · 2%/日 · 🆕首次上榜TypeScript
The Office Harness for AI Agents — Spreadsheets, Docs, Slides, Canvas, Relational Tables, and PDF in one runtime.

📰Hacker News

过去24小时前排+Show HN,正文没讲到的都在这,扫标题即可。
热评 · dgellow
Im sorry given how bad of a situation that is, but it would be so ironic if they used Claude or codex for this
黑客声称掌握全部FBI员工数据;目前是攻击者说法,等待机构确认与泄露范围核实。
热评 · CharlieDigital
> Fact is, vibe-coded projects devolve over time into an unmaintainable mess. The reason is simple, yet hard to fix: code maintainability and good architecture don’t have good measurements that we can apply, because it takes months, years even, to notice the effects of bad architecture or of unmaintainable code. > > For one, AI is not trained on what it means for code to be maintainable. For instance, any reinforcement learning done needs a reward signal that can be measured immediately, not in …
《AI没有智慧,你也不会有》引发528条讨论,核心担忧是维护性和判断力没有即时奖励信号。
热评 · kgwgk
It would have been funny if the "how" had been dropped to end with "Did AMD Ryzen get 50% faster in two years?"
热评 · tolugenius
I'm not exactly following through with the claim, can someone explain how the built-in classification would not necessitate more tokens used, or be much different from turning on reasoning? Not that I don't see the difference, I just doing see how OpenAI would do it well.
热评 · hglaser
Half the cost per task compared to Opus 5, comparing high effort to high effort. That's just really nice. Edit: https://artificialanalysis.ai/models/claude-opus-5-5?models=...
热评 · Cider9986
The upcoming Signature 27 will be announced today at the Snapdragon Summit. It has been teased by Motorola on their socials and will likely release in the US because they've posted about it on US social accounts. The 2026 Signature, in the UK, is $1460 USD. In Brazil it is selling for $1230. The Pixel 11 pro XL, which the Signature beats on hardware in practically every way sells for $1300 on Google's website. It looks interesting... I didn't love the 2026 design [1] but it looks better than wha …
GrapheneOS称2027年很可能出现预装设备,隐私系统开始尝试越过刷机门槛。
▲273Claude Opus 5.52评论 · HN↗
热评 · tomhow
Trail of Bits拆解SAML的层层复杂性;Agent接企业身份之前,旧协议债务不会消失。
WordPress披露未认证路径穿越并可能条件式RCE,维护者应优先核对版本和官方修复。
▲185Unreal Agent107评论 · HN↗
▲154LLM Ass Bench43评论 · HN↗
▲140Transit rewards142评论 · HN↗
Waymo推出公共交通奖励,自动驾驶与公交不只存在替代关系,也可以用价格机制衔接。
▲138Markdown in /src73评论 · HN↗
🚀 Show HN(独立发布,共8个)

🔎值得深挖

  • Opus 5.5、GPT-6 Sol/Luna与Grok 4.7在同一套真实任务上的成功成本:把缓存、重试、耗时和人工接管纳入,而不是再做一张综合榜
  • Muse 6.8GB运行环境导出事件的权限链:哪些文件原本应当可见、哪些连接器能导出、产品如何定义会话与租户边界
  • Bot验证的空白市场:从表单灌水到交易欺诈分别需要什么证据,怎样在防自动化与误杀真人之间建立可测基线

📖附录:原文流(按作者)

想翻谁点谁展开。转推与噪音折叠在各自账号内。
🤖 AI行业
Claude@claudeaiAI1.8M粉 · 2条发布Opus 5.5及早期创作案例;本期主页检查未成功,摘要仅基于文件内推文。
@claudeai2026-09-22♥2.7K👁150.5K↗ 打开X
A thread of early explorations with Claude Opus 5.5: A short story about a watermelon created by @kevin_t_ngo. https://t.co/v48srlNi7V
@claudeai2026-09-22♥40.2K👁2.9M↗ 打开X
Introducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5. https://t.co/Q9C2VKQ79f
Opus 5.5官方发布:多数任务接近Fable 5.1,运行成本比Opus 5低40%。
SpaceXAI@spacexaiAI2.1M粉 · 1条发布Grok 4.7并转发折扣活动;本期主页检查未成功,摘要仅基于文件内推文。
@spacexai2026-09-21♥28.0K👁14.0M↗ 打开X
Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed. https://t.co/H3OTBbXyvO
Grok 4.7官方发布,口径是相对4.6同价同速、能力提升。
▸ 折叠1条(转推/噪音)
转推2026-09-22 RT @vercel: Build with Grok 4.7 and Vercel for less. 40% off until September 27th on AI Gateway. Try with @v0, @eve, and fx. https://t.co/…
Tibo@thsottiauxAI731.2K粉 · 4条作为产品负责人发布GPT-6 Sol/Luna、永久降价与额度重置。
@thsottiaux2026-09-22♥6.4K👁705.2K↗ 打开X
Maybe our cutest launch so far. But still packing the biggest punch.
引用@OpenAIPlease welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale. We’ve also made caching and inference … ↗
@thsottiaux2026-09-22♥21.0K👁1.9M↗ 打开X
GPT-6 Sol and Luna are out. Not only are they a very significant improvement across the board, but also in writing and general "you know when you try it" quality. We are also permanently reducing the API price by 50% making both of them viable for a ton of new usecases and making your usage go further too, even on the subscriptions. And one more thing. We are loading a banked reset into all accounts of our Plus, Pro and Business users. Let's go! https://t.co/00DRh1sRrO
引用@thsottiauxWe have been focusing on efficiency and intelligence for all. Very proud of the team. Only possible when you have incredible models at the top end of the capability that you can then use to make a big difference in everything else. ↗
GPT-6 Sol/Luna产品负责人发布口径:永久降价50%并给订阅账户一次可累积重置。
@thsottiaux2026-09-22♥11.2K👁1.7M↗ 打开X
We have been focusing on efficiency and intelligence for all. Very proud of the team. Only possible when you have incredible models at the top end of the capability that you can then use to make a big difference in everything else.
@thsottiaux2026-09-22♥25.7K👁4.4M↗ 打开X
Ladies and gentlemen... start... your... ENGINES. We are almost Tuesday and I promised a reset for Tuesday. Among some other things. See you soon.
Sam Altman@samaAI6.3M粉 · 6条集中发布GPT-6 Sol/Luna的效率与价格叙事,也提出AI安全标准方案。
@sama2026-09-22♥3.4K👁442.2K↗ 打开X
Startups are naturally good at this; it is hard to keep a bigger company good at this and i think an underexplored space.
引用@prd_008OpenAI's superpower is the ability to swiftly assemble an empowered group of highly capable people to work on the most important and urgent thing at any point of time. No one really cares for org lines. It's how we stay nimble, seize opportunities, and … ↗
@sama2026-09-22♥2.0K👁316.9K↗ 打开X
michelle embodies this as much as anyone openai is so, so lucky to have benefited from everything she has done so far and i think people will be quite pleased to see what she + team have cooking next!
引用@michpokrassjust crossed four years at openai! the special thing about this place is the constant capacity for rebirth. for all its faults, there is nowhere quite like it. the team makes high conviction, contrarian bets over and over, and they are mostly right. it's a … ↗
@sama2026-09-22♥6.0K👁497.7K↗ 打开X
Especially compared by per-task pricing, which is the metric that should matter, I don't think there is anything competitive anywhere in the market. We want people to be able to use tons of AI; it is important to being able to explore this new renaissance in front of us.
引用@samaGPT-6 Sol and Luna are big improvements on intelligence, alignment, work output, coding, computer use, and more over their 5.6-family predecessors. They are also half the price per token, and even less per task! ↗
@sama2026-09-22♥14.5K👁1.1M↗ 打开X
GPT-6 Sol and Luna are big improvements on intelligence, alignment, work output, coding, computer use, and more over their 5.6-family predecessors. They are also half the price per token, and even less per task!
sama强调智能、对齐、工作输出、编码与电脑操作提升;均属厂商声明。
@sama2026-09-22♥4.9K👁364.3K↗ 打开X
GPT-6 Sol and Luna are great models but also these characters are so cute
引用@OpenAIPlease welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale. We’ve also made caching and inference … ↗
@sama2026-09-22♥4.3K👁355.2K↗ 打开X
People outside the AI labs should have a real say in how this technology develops, and a clear way to judge if it's happening safely. Standards should help prevent the concentration of power, including by making sure new companies and open-model companies can compete. They should also help countries and companies compare evidence and learn from failures. We think the US should lead this effort. Here is our proposal: https://t.co/2FT9a2m4l0
OpenAI提出由外部参与、安全证据比较和竞争保护组成的标准框架。
▸ 折叠1条(转推/噪音)
转推2026-09-22 RT @thsottiaux: We have been focusing on efficiency and intelligence for all. Very proud of the team. Only possible when you have incredi…
OpenAI Developers@openaidevsAI432.3K粉 · 2条发布GPT-6 Sol/Luna并集中转发焊接、视频、3D和教育用例。
@openaidevs2026-09-22♥8.6K👁763.0K↗ 打开X
GPT-6 Sol and Luna just landed in Astra’s orbit. Both launch today with API prices 50% lower than GPT-5.6. Build with Sol. Scale with Luna. To production and beyond. https://t.co/ZCEFp4JdjV
@openaidevs2026-09-22♥320👁49.4K↗ 打开X
See how surgeon @pridgenmd uses Codex to build his own tools and review scientific literature on PubMed: https://t.co/i8cs5FikTw
▸ 折叠9条(转推/噪音)
转推2026-09-23 RT @rileybrown: Shout out to @zulali who built this AI powered TV from the 90s. He bought this old TV, gave it a raspberry pi, gave it a mi…
转推2026-09-23 RT @github: 📣 @OpenAIDevs's GPT-6 family is expanding in GitHub Copilot with two additional models now generally available. ☀️ GPT-6 Sol:…
转推2026-09-23 RT @angelaborowski: GPT 6 Astra, colour graded, captioned and stitched this entire video together 🤯 went to the codex @OpenAIDevs @Andy_AJ…
转推2026-09-23 RT @KennethCassel: Our welders are using Astra + our api to iterate on weld fixtures https://t.co/1xgRC3u4c6
转推2026-09-22 RT @Lovable: One more thing… Lovable now builds with GPT-6 Sol. On our 0-to-1 building benchmark it scores 6 to 12% higher than GPT-5.6 So…
转推2026-09-21 RT @techartist_: This flamethrower is inspired by Dan Greenheck. Starting with only a screen recording as a visual reference, it took me tw…
转推2026-09-21 RT @andreeliasdev: This might be the future of education! 🤯 I decided to test the GPT-Live-1 API and built a real-time chess instructor…
转推2026-09-21 RT @ProductHunt: What a turnout for our GPT-6 Astra Challenge with @openAIDevs! 5 teams rocketed to the top , winning $10k in @openai API…
ChatGPT@chatgptAI728.7K粉 · 2条发布GPT-6 Sol/Luna在Work与Codex的可用范围,并推广信用分追踪。
@chatgpt2026-09-22♥7.4K👁427.4K↗ 打开X
GPT-6 Sol and GPT-6 Luna, it’s your time to shine. Rolling out today in ChatGPT Work and Codex for Plus, Pro, Business, Enterprise, and Edu users. https://t.co/UtpG1Tbhd3
@chatgpt2026-09-21♥2.9K👁357.7K↗ 打开X
You can now track your credit score in ChatGPT, and stay informed when things change. Securely connect your @Experian credit report and VantageScore® 3.0 credit score to get personalized insights into what’s affecting your score and how it relates to your finances and goals. Available in Finances for Plus and Pro users in the U.S. on web and the latest versions of our iOS and Android apps.
▸ 折叠1条(转推/噪音)
转推2026-09-23 RT @IlyaAbyzov: Helpful (and dare I say fun?) new feature in ChatGPT Health: tap on anything in your synced data and get a helpful summary…
Mckay Wrigley@mckaywrigleyAI228.3K粉 · 1条高度评价Opus 5.5的性格、能力和额度。
@mckaywrigley2026-09-22♥3.9K👁112.2K↗ 打开X
opus 5.5 = personality of opus 4.6 that we all *desperately* wanted back + the intelligence & taste of fable 5.1. and a 25% usage bump! i honestly don’t really see a reason to use another model right now? s-tier release.
Opus 5.5的高热度口碑样本,重点在对话风格回归和额度增加。
Thariq@trq212AI349.4K粉 · 1条本窗口只留下一条“多用大图少用字”的表达偏好。
@trq2122026-09-22♥2.2K👁163.7K↗ 打开X
I now type "use big pictures and few words" several times a day
Peter Gostev@petergostevAI27.0K粉 · 11条实测GPT-6 Sol效率,也继续要求OpenAI公开数学问题清单、给Jev做游戏测试。
@petergostev2026-09-22♥13👁1.6K↗ 打开X
引用@etnshowAnnouncing the 2026 ETN100, the 100 most influential tech posters on X https://t.co/63oX8rRIYi https://t.co/c7pulzSfeJ ↗
@petergostev2026-09-22♥1.1K👁53.6K↗ 打开X
The best thing about GPT-6-Sol is efficiency. In my testing the difference between GPT-5.6-Sol and GPT-6-Sol is the new version consumes 1/2 the tokens and takes 1/5 of the time to complete the task. On top of the 50% cost reduction, it should be quite a nice daily driver
早期实测样本:Sol据称少用一半token、只花五分之一时间,尚待复现。
@petergostev2026-09-22♥173👁7.9K↗ 打开X
I gotta say, my bet was on: AI = American Intelligence
@petergostev2026-09-22♥126👁5.8K↗ 打开X
The important question for the singularity is why did OpenAI's new model solve 100 open maths problems, but none of Jay-Z's
@petergostev2026-09-21♥19👁2.2K↗ 打开X
Jev plays RollerCoaster Tycoon 2 - unfortunately it didn't really do anything useful. It built some rides at the beginning and then got stuck. Setting this up wasn't simple so maybe someone could do it better. But it isn't a magical game playing model out of the box. For harder games we do need the intelligence of smart models.
引用@petergostevThis is the one you've all been waiting for: Jev plays RollerCoaster Tycoon 2 https://t.co/hVvs4ycXa2 ↗
@petergostev2026-09-21♥2.0K👁145.2K↗ 打开X
"this model has now resolved more than 100 long-standing open problems across most areas of mathematics" I'd appreciate a heads up which ones they are, just a list of 100 titles, no need for papers
引用@OpenAIWe’re working with an independent advisory group of mathematicians to help OpenAI responsibly share advances in AI and mathematics. The group will advise on how we assess and communicate new mathematical results, uphold academic and professional standards, … ↗
要求OpenAI给出所谓100个数学问题的清单;对重大能力主张,最小可核验证据很重要。
@petergostev2026-09-21♥51👁4.6K↗ 打开X
Advisor: "Believe in yourself. Make a breakthrough"
引用@OpenAIWe’re working with an independent advisory group of mathematicians to help OpenAI responsibly share advances in AI and mathematics. The group will advise on how we assess and communicate new mathematical results, uphold academic and professional standards, … ↗
@petergostev2026-09-21♥8👁3.3K↗ 打开X
This is the one you've all been waiting for: Jev plays RollerCoaster Tycoon 2 https://t.co/hVvs4ycXa2
@petergostev2026-09-21♥9👁1.2K↗ 打开X
Astra playing Fallout 2 by @Clad3815 - gorgeous design around the whole gaming UI https://t.co/00FYNvEEPN
@petergostev2026-09-21♥380👁28.6K↗ 打开X
Can't wait for this to be cancelled in 18 months' time!
引用@GoogleIntroducing Googlebook, a new category of laptop, available for preorder today. 💪 Crafted with 2.8K OLED touchscreen displays, 14 hour battery life and all the performance needed to power your ideas 📳 Engineered to sync effortlessly with your @Android phone, … ↗
@petergostev2026-09-21♥15👁1.7K↗ 打开X
Do you build the most efficient stack & scale it? (e.g. OpenAI, DeepSeek) Or do you build the biggest models and use them to make your stack more efficient? (e.g. Anthropic)
▸ 折叠1条(转推/噪音)
转推2026-09-22 RT @arena: Scores for GPT-6 Sol and GPT-6 Luna by @OpenAI are coming soon. Head to Arena now to test them. Your votes on real-world agentic…
Greg Brockman@gdbAI1.1M粉 · 1条转发GPT-6 Sol/Luna发布,强调专业工作、编码和电脑操作能力。
@gdb2026-09-22♥1.5K👁90.7K↗ 打开X
GPT-6 Sol and Luna — faster, more affordable models with the advances behind Astra’s SOTA performance in professional work, factuality, coding, computer use, and alignment:
引用@OpenAIPlease welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale. We’ve also made caching and inference … ↗
Matt Pocock@mattpocockukAI353.6K粉 · 3条把“拯救vibe-coded代码库”整理成课程问题,强调领域语言、测试和ADR。
@mattpocockuk2026-09-22♥1.5K👁92.7K↗ 打开X
I hear this a LOT from panic-stricken devs: "I've inherited a vibe-coded codebase, how do I save it?" Either from devs who have stopped caring, or non-technical folks trying to push AI beyond their abilities. A huge chunk of my next course will tackle this: - deepening modules - establishing domain language - increasing testability - explaining non-obvious code with ADR's - automated code migrations
“继承vibe-coded代码库”成为课程需求:修复重点是领域语言、测试、ADR与迁移。
@mattpocockuk2026-09-21♥121👁16.2K↗ 打开X
Stop scrolling and read a book nerds This is a good one https://t.co/HTlBaQNzjL
@mattpocockuk2026-09-21♥722👁78.4K↗ 打开X
"No tautological tests"
引用@dexhorthyleave it to your boy opus to add 10 unit tests to ensure a constant string contains various substrings ↗
Simon Willison@simonwAI225.1K粉 · 3条拆解Opus 5.5、GPT-6 Sol/Luna的价格与表现,尤其看好低价Luna。
@simonw2026-09-22♥439👁48.1K↗ 打开X
Big model release today - I wrote about Claude Opus 5.5, GPT-6 Sol, and GPT-6 Luna - plus comparison grids of pelicans by the different model families at different reasoning levels https://t.co/R5lSdOZyLj
@simonw2026-09-22♥1.4K👁92.3K↗ 打开X
GPT-6 Luna is half the price of 5.6 Luna, which was already an astonishingly cheap model given how capable it is Luna is my favorite model for building product features thanks to its cost (and speed)
引用@OpenAIPlease welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale. We’ve also made caching and inference … ↗
simonw看好Luna的速度与价格,认为适合嵌入产品功能。
@simonw2026-09-21♥1.1K👁80.7K↗ 打开X
Put together some notes on Jev and the new category of system one aka decision models https://t.co/EQZvszIXQ9
▸ 折叠1条(转推/噪音)
转推2026-09-23 RT @addyosmani: Introducing Claude Opus 5.5! 40% lower cost than Opus 5 with cache reads 60% cheaper. It performs at the level of Claude Fa…
TypeSafe AI@typesafeaiAI146.1K粉 · 7条Jev因需求暂停注册,同时密集展示动态UI、Pydantic、MotherDuck和安全分类用例。
@typesafeai2026-09-23♥1.2K👁95.2K↗ 打开X
Dynamic UIs for apps are a famous graveyard at least for PMs, if not for entire products and companies. What if it were really this easy? Jev's on it.
引用@anishfni built a text box that turns into whatever(some ui) you type. powered by jev https://t.co/HH0P51033y ↗
@typesafeai2026-09-22♥225👁45.8K↗ 打开X
Recent jevelopments have blown out all expectations but wait til you see what comes out of Paradigm Frontiers in a few weeks! The future of AI is just across the event horizon and @CompleteSkeptic is going to shepherd us through 🚀
引用@gakonstour special guest is out! excited to have @CompleteSkeptic join us for Frontiers! Diogo & @typesafeai have taken over the developer community with Jev, and we're thrilled to see what they have to share at Frontiers! apply below, closes this week! oct … ↗
@typesafeai2026-09-22♥224👁40.1K↗ 打开X
With rogue AIs wandering the web hunting for your vulnerable smart fridge, you need a stronger AI keeping you safe 💪
引用@brainstormityOk so... JEV just accidentally helped me find backdoors on a family member's wifi network. 🤯 - ​Brought my laptop to code while visiting. - Decided to experiment with a JEV-powered classifier for network packets fetched via Wireshark. ​While building the … ↗
@typesafeai2026-09-22♥226👁35.1K↗ 打开X
@pydantic has been type-safe from the beginning! Now with Jev powering your Pydantic AI agent calls, they’ll be returning faster than ever ⚡️⚡️
引用@pydanticPydantic AI agents now run on Jev, the classifier from @typesafeai. Jev doesn't write text, it answers typed questions. So the output_type you already wrote is the question, and the answer comes back as your model, one confidence per … ↗
@typesafeai2026-09-22♥772👁137.0K↗ 打开X
We have seen such an immense swell of demand that we have to temporarily pause signups for Jev. We need to ensure quality of service for our existing signups, which will continue to function. We are working diligently to ensure open access to Jev for everyone as soon as we can. Thank you.
Jev因需求过大暂停新注册,服务容量成为第一道生产约束。
@typesafeai2026-09-22♥187👁39.5K↗ 打开X
he olo world
@typesafeai2026-09-22♥621👁73.7K↗ 打开X
Jev makes it easy to add natural language intelligence into the key parts of any application at scale, far cheaper and faster than has ever been possible. 50x faster. 100x cheaper. Reliable as duck.
引用@motherduckText classification in MotherDuck just got ~50x faster at ~1% of the cost. prompt_jev() is a SQL function powered by Jev, TypeSafe's new system one model. 100k rows: 40s, $0.50, frontier-LLM accuracy. The LLM took 32 min and $37. Read … ↗
MotherDuck案例宣称10万行分类40秒、0.50美元,对照通用LLM 32分钟、37美元。
▸ 折叠8条(转推/噪音)
转推2026-09-23 RT @posthog: and my partner here he wanna build a product https://t.co/6EbK9bBKbN
转推2026-09-22 RT @langfuse: @doneyli @typesafeai with Jev it is less eval -> eval -> eval and more -> eval -> eval -> eval
转推2026-09-22 RT @Bailey_Jennings: Here's a quick demo of me using @typesafeai's Jev for ONET job classification vs. Luna. - Luna: 84.9% exact, ~2s/job,…
转推2026-09-22 RT @notkevinzhang: Did y’all forget?
转推2026-09-22 RT @s16h_: at @MetaviewAI, over the weekend we shipped @typesafeai's jev into every agent on metaview. candidate searches in our sourcing…
转推2026-09-22 RT @TheIshanGoswami: Jev with Exa is INSANE. > Jev without websearch confidently gives wrong outputs > Jev with websearch is literally muc…
转推2026-09-22 RT @sydneyrunkle: :) https://t.co/uF8ETTdbv8
转推2026-09-22 RT @sydneyrunkle: here's a quick breakdown of @typesafeai's new jev model: what it is, and how to plug it into your agents! https://t.co/Sg…
clem 🤗@ClementDelangueAI697.6K粉 · 3条放大MiMo、Halo、端侧模型和开放权重安全模型,继续押注自训练能力普及。
@ClementDelangue2026-09-22♥338👁18.1K↗ 打开X
Feels like I’m the diversity pick on this one 😅 https://t.co/3e0cwaXAam
@ClementDelangue2026-09-21♥1.0K👁137.2K↗ 打开X
Training models is becoming easier and easier - just look at this and TRL - especially with agents! You're missing out if you're still using off the shelf models for all your tasks!
引用@whitecircleIntroducing Halo, the best framework for post-training of open-source models. Halo delivers up to 2.8x the throughput of stock TRL with less peak memory, while models stay in their native HuggingFace format. Star us on GitHub: https://t.co/3mAiUljdrN … ↗
Halo宣称相对stock TRL最高2.8倍吞吐;由Clement放大,仍需独立基准。
@ClementDelangue2026-09-21♥238👁32.0K↗ 打开X
引用@ArtificialAnlysMiMo-V2.6-Pro debuts as the top open weights model on the Artificial Analysis Intelligence Index (46). At $0.13 per Intelligence Index task, it lands on the Intelligence vs. Cost per Task Pareto frontier @Xiaomi has just released MiMo-V2.6-Pro, an open … ↗
MiMo-V2.6-Pro的成本前沿数据来自第三方榜单转述,适合做候选而非结论。
▸ 折叠12条(转推/噪音)
转推2026-09-23 RT @M1Astra: Claude Opus 5.5 will be the first Opus meant to fall back to a less capable model for "a small set of capabilities related to…
转推2026-09-23 RT @jundotkim: I'm happy to announce that I've joined Hugging Face. What started as a personal project back in February is now something I…
转推2026-09-22 RT @george_onx: We heard you like decision models…coming soon. https://t.co/2szUymJdJG
转推2026-09-22 RT @MiaAI_lab: MiMo V2.5 Flash up & running on 2x DGX Sparks Initial decode prose numbers look good! https://t.co/aVhmuV4bbB
转推2026-09-22 RT @victormustar: Really cool that MiMo V2.6 has an official Distill-Qwen-9B variant remind me of the legendaries DeepSeek R1 Qwen/LLama di…
转推2026-09-22 RT @Alex_tra_memory: Thanks for the model, we were able to port Laya to coreml with 99.5% of the ops on ANE + benchmarked too. it is now bl…
转推2026-09-22 RT @Nandakishorm1: The first model of Laya was created with a single over night training and fine-tuning. It only took 15 hours from traini…
转推2026-09-22 RT @theo: Did a few quick tests and these seem very very legit. 2.6 Pro in particular is doing well in some pretty hard tasks.
转推2026-09-22 RT @AikidoSecurity: Introducing Altar-1, our first open-weight security model. Frontier-grade defensive AI, built to deploy. Own your own…
转推2026-09-21 RT @ariG23498: Wake up babe! State of the art tokenization framework just dropped. > multiple language support > multi-thread scaling > m…
转推2026-09-21 RT @whitecircle: Introducing Halo, the best framework for post-training of open-source models. Halo delivers up to 2.8x the throughput of…
转推2026-09-21 RT @zephyr_z9: Pretty insane result They spent 130 hours, 75B tokens, and $2.6M on RL to achieve this result https://t.co/Lg6FHtLqS7
Vaibhav (VB) Srivastav@reach_vbAI59.8K粉 · 7条主讲GPT-6 Sol/Luna的定价、缓存、基准与额度重置。
@reach_vb2026-09-23♥133👁6.2K↗ 打开X
GPT-6 Luna
@reach_vb2026-09-22♥307👁11.7K↗ 打开X
Team cooked with the design, yet again!! https://t.co/n6wGQYzv2m
引用@reach_vbIntroducing GPT-6 Sol and Luna, bringing the advances behind Astra to faster, more affordable models. ✨ 💻 Stronger coding and computer use 🎯 Improved factuality and alignment 💬 Clearer answers with less jargon > On AutomationBench, Sol at xhigh effort … ↗
@reach_vb2026-09-22♥719👁47.0K↗ 打开X
We are loading a BANKED RESET into all accounts of our Plus, Pro and Business users!!
引用@reach_vbIntroducing GPT-6 Sol and Luna, bringing the advances behind Astra to faster, more affordable models. ✨ 💻 Stronger coding and computer use 🎯 Improved factuality and alignment 💬 Clearer answers with less jargon > On AutomationBench, Sol at xhigh effort … ↗
@reach_vb2026-09-22♥331👁17.5K↗ 打开X
GPT-6 Sol & Luna are ~50% cheaper than 5.6 https://t.co/Qdnbj57aav
引用@reach_vbIntroducing GPT-6 Sol and Luna, bringing the advances behind Astra to faster, more affordable models. ✨ 💻 Stronger coding and computer use 🎯 Improved factuality and alignment 💬 Clearer answers with less jargon > On AutomationBench, Sol at xhigh effort … ↗
@reach_vb2026-09-22♥303👁9.3K↗ 打开X
gm!
@reach_vb2026-09-21♥183👁8.5K↗ 打开X
8 days to DevDay! (liberally using ChatGPT Work & Codex Remote to get through tonnes of PRs for what’s next) https://t.co/urvqhmKusv
引用@reach_vb9 days to DevDay! (also wrapped up the first leg of euro travels) https://t.co/aSEy28bea0 ↗
@reach_vb2026-09-21♥146👁9.7K↗ 打开X
“The group will advise on the review and communication of emerging results: they will help OpenAI assess their significance, advise on how to coordinate their dissemination, and advise on academic and professional standards of mathematical research.” “We want to put capable tools in mathematicians’ hands so they can pursue the questions they know best and develop new ideas.” “The group will operate independently from OpenAI. The group will have the freedom to offer advice we have not requested, comment on OpenAI’s impact on mathematics, and make its advice public. Its value depends on its members being able to exercise their own judgement and challenge ours. Its members will not be paid by OpenAI, and the group can change its membership as it sees fit.” I’m biased, but I love this approach!
引用@OpenAIWe’re working with an independent advisory group of mathematicians to help OpenAI responsibly share advances in AI and mathematics. The group will advise on how we assess and communicate new mathematical results, uphold academic and professional standards, … ↗
Hassan@nutlopeAI100.3K粉 · 2条发布开放模型选择器,并公开周末用5–10个平行Agent筛POC的软件工厂流程。
@nutlope2026-09-22♥175👁11.6K↗ 打开X
Introducing https://t.co/nNCVsRXukg! Figure out which open model is best for your use case. Compare models across coding, agents, long context, vision, finance, and more. Then see how they compare on cost + quality, including what you could save by moving to open models. https://t.co/iWtGmIXyxr
开放模型选择器把任务能力、成本与迁移节省放在同一界面。
@nutlope2026-09-21♥370👁18.9K↗ 打开X
Here's how I run my software factory: 1. During the week, I collect ideas + inspo. 2. On the weekend, I give the list to an agent to rank the best ones. 3. I spin up ~5-10 parallel agents to build POCs. 4. I kill ~60%, iterate on the better ones, and end up with 2-3 solid demos. 5. I then polish & share those demos on X. Then rinse and repeat! I still build some ideas immediately, but I'm increasingly using weekends to let agents explore ideas in parallel.
周末软件工厂:5–10个平行Agent做POC,自报淘汰约60%、留下2–3个。
Alex Volkov@altryneAI42.9K粉 · 15条高密度跟进Opus、GPT-6、Jev、Muse与Cloudflare分支预览。
@altryne2026-09-23♥7👁927↗ 打开X
"At $0.10/$0.50 GPT-6 Luna is one of the cheapest models OpenAI have ever released, beaten only by the far weaker GPT-4.1 Nano ($0.10/$0.40, April 2025) and GPT-5 Nano ($0.05/$0.40, August 2025)." !! wow
引用@simonwBig model release today - I wrote about Claude Opus 5.5, GPT-6 Sol, and GPT-6 Luna - plus comparison grids of pelicans by the different model families at different reasoning levels https://t.co/R5lSdOZyLj ↗
@altryne2026-09-23♥11👁1.5K↗ 打开X
ThursdAI was easier in the beginning 😂 we used to be able to talk about papers!
引用@christianeltonThe gap between major model releases keeps shrinking. 2023: once every 73 days 2026 (so far): once every 18 days 🤯 https://t.co/FtZ7Vpjt9z ↗
@altryne2026-09-23♥10👁1.3K↗ 打开X
Opus is so so back.
@altryne2026-09-23♥3👁1.4K↗ 打开X
The status pill in muse is like the "pull to refresh" from the iOS apps era One of the things I hated about telegram being home to my open claw is that telegram only has like four status notifications types and they're not editable by the agent (@durov!) - just typing... So simple and genius and will get copied relentlessly!
引用@alexcornellWhen I look back at the earliest @Muse mocks, the very first thing that was designed was the “status” pill at the top of the screen. That one tiny piece of UI survived every subsequent revision, many months on. It remains one of my favorite parts of the … ↗
@altryne2026-09-22♥7👁1.5K↗ 打开X
Same, it's really something!
引用@theoBeen pushing Opus 5.5 hard all day and it's barely making a dent on my usage limits. Impressive! ↗
@altryne2026-09-22♥11👁2.2K↗ 打开X
How I'm feeling today seeing all the releases and their benchmark scores https://t.co/UBRiMk2y8o
@altryne2026-09-22♥20👁4.0K↗ 打开X
So ugh... it's only Tuesday and we already got: Grok 4.7 (meh release according to TL) and grokbot in tesla Opus 5.5 + a banked reset (🔥 so far) GPT 6 Sol and Luna! 🔥🔥 so far! Muse gets banned by Amazon but partners with Shopify, Stripe and Expedia + Meta connect tomorrow and Dev Day next week!? Anyone wanna nominate this for the insane-es week this September? Obvisouly we'll cover all this on @thursdai_pod but goddamn!
@altryne2026-09-22♥13👁2.0K↗ 打开X
OpenAI unveils 2 new GPT-6 models (bye bye Terra) and honestly, Astra should be SOL Sol should be Terra (the daily driver) and Luna the fast one still stays. Prices update, capabilities largely the same, but price per task is much lower! Testing!
引用@OpenAIPlease welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale. We’ve also made caching and inference … ↗
@altryne2026-09-22♥64👁8.7K↗ 打开X
Opus 5.5 from Anthropic, first time we see an Opus price drop!? Also we're getting an INCREASE in the 5-hour usage and a BANKED reset? What is OpenAI about to drop that got them running so scarred? wow https://t.co/4d8Y1yDVBh
引用@claudeaiIntroducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5. https://t.co/Q9C2VKQ79f ↗
altryne抓到两个关键信号:Opus首次明显降价,以及5小时额度增加与可累积重置。
@altryne2026-09-22♥45👁3.1K↗ 打开X
OMG fuuuucking thank you @Cloudflare 😍 https://t.co/3dxZS6pTji
引用@CloudflareToday we’re launching Worker Previews. Each Git branch gets a production-like place to run, with its own code, configuration, URL, observability, and state. https://t.co/rpr81YWE1N ↗
Cloudflare Worker Previews让每个Git分支获得独立配置、URL、可观测性和状态。
@altryne2026-09-22♥10👁1.6K↗ 打开X
AI Agent - An AI with a loop, tool calling and maybe a harness AI Assistant - Has it's own computer, is proactive, knows things about YOU the user, has memory, can be helpful All AI assistants are agents but not all AI agents are Assistants
引用@altryneGuys, @bot @Muse Instinct and the upcoming Aeon from OpenAI (rumored) are ... ASSISTANTS! Let's call them that, they are assisting people in their daily lives. Let's move away from "Agent" it's too broad! Thank you for coming to my ted talk! ↗
@altryne2026-09-22♥3👁1.7K↗ 打开X
Agreed + a @moxie hardened secure confidential VM is coming! We covered this 2 weeks ago on Muse release week! (feels so long ago) https://t.co/vKlbWrwI9S
引用@herrmanndigital"You're really going to give Meta access to your Gmail, Shopify, Calendar, as well as bank info with this Muse thing?" Yes, Yes I am. I actually trust Meta more than any of these AI companies simply because of the amount of scrutiny and legal crap they've … ↗
@altryne2026-09-22♥9👁1.8K↗ 打开X
BJ - A unit of historical time, counting backwards from the release of JEV from @typesafeai FWIW @thursdai_pod started in the year 3.5 BJ - March 13, 2023 https://t.co/PbRhCGAAFl
引用@altryne@JohnPaulGarland @rewind @LimitlessAI yeah in the year 3 BJ - before JEV ↗
@altryne2026-09-22♥29👁5.8K↗ 打开X
Guys, @bot @Muse Instinct and the upcoming Aeon from OpenAI (rumored) are ... ASSISTANTS! Let's call them that, they are assisting people in their daily lives. Let's move away from "Agent" it's too broad! Thank you for coming to my ted talk!
@altryne2026-09-22♥19👁4.7K↗ 打开X
How did JEV get started? @CompleteSkeptic shares the story on @latentspacepod, Sam Altman was involved!? https://t.co/qptPz7EV5p
引用@latentspacepodJev and the System One Model: RLCD, intelligence/$, reliable AI, & the end of chat-first AI https://t.co/H2bZXCENyW @typesafeai CEO @CompleteSkeptic explains why AI can solve extraordinarily hard problems yet still fail to automate basic work, why Jev is … ↗
swyx@swyxAI193.9K粉 · 1条把Opus 5.5设为AINews默认模型,继续放大Jev与System One讨论。
@swyx2026-09-23♥20👁2.0K↗ 打开X
can confirm. ran @latentspacepod AINews side by side with 6 Sol and the difference was night and day: https://t.co/oloSDxf0q7 5.5 Opus is the new default model for AINews going forward. so much more concise and tasteful reporting, with much less slopese than even 5 Opus. https://t.co/P6AXDlBKNP
引用@_sholtodouglasalso important news we fixed the writing ↗
单一编辑工作流里,swyx把Opus 5.5设为AINews默认模型。
▸ 折叠4条(转推/噪音)
转推2026-09-22 RT @heyalizaid: this podcast needed a great thumbnail new @latentspacepod episode with @CompleteSkeptic talking about Jev, System One Mode…
转推2026-09-22 RT @jeffreyhuber: giving lyft 2012 vibes (compliment)
转推2026-09-22 RT @RayFernando1337: The most important podcast in AI right now.
转推2026-09-21 RT @latentspacepod: Jev and the System One Model: RLCD, intelligence/$, reliable AI, & the end of chat-first AI https://t.co/Ki4vHeOiQe @t…
Ben Tossell@bentossellAI200.4K粉 · 4条用Astra做50年设备网站并加排行榜,同时坦言模型与harness选择更混乱了。
@bentossell2026-09-22♥15👁5.0K↗ 打开X
lol jk i dunno what model or harness to point at what rn
引用@bentossellnow i can finally get all my work done ↗
@bentossell2026-09-22♥12👁5.2K↗ 打开X
this is new (6-sol) https://t.co/iwUbsWzwPi
@bentossell2026-09-22♥6👁4.9K↗ 打开X
now i can finally get all my work done
@bentossell2026-09-22♥6👁2.1K↗ 打开X
just added a live leaderboard 🥇 game boy colour 🥈 discman 🥉 game boy / playstation / nokia 3310 only 1 person had the pokewalker 💔 https://t.co/coAfZhM5hL
引用@bentossell50 years of devices. save which you had or wanted Astra image-gen'd all the devices + built the site https://t.co/JAHUJAGpWG inspired by https://t.co/Df4YhZIYHk (i'd one-shot a collection of devices a few weeks ago and had no idea what to do with it, so … ↗
▸ 折叠4条(转推/噪音)
转推2026-09-23 RT @bentossell: 50 years of devices. save which you had or wanted Astra image-gen'd all the devices + built the site https://t.co/D3eOmh…
转推2026-09-22 RT @bentossell: 50 years of devices. save which you had or wanted Astra image-gen'd all the devices + built the site https://t.co/D3eOmh…
转推2026-09-22 RT @FactoryAI: Opus 5.5 is live in Factory. Some initial observations: / Medium is a strong default / 20–25% fewer output tokens than @Ant…
转推2026-09-22 RT @0xSigil: Meet Husky: a Model-Specific Inference (MSI) engine up to 4.5× faster than Apple's MLX Woof, Underdog's Pareto frontier mod…
🧑‍💻 独立开发者
DHH@dhhindie889.0K粉 · 5条宣布300万美元Omacom赞助、推广Omarchy,并为Rails World预热。
@dhh2026-09-22♥1.2K👁33.5K↗ 打开X
Austin is beautiful in the morning glow. Excited for a whole week of Rails World here! Opening keynote tomorrow will be streamed. It's gonna a good one 😄 https://t.co/8WGvGoGiNn
@dhh2026-09-22♥6.1K👁128.4K↗ 打开X
It's incredible how far you can go if you just don't stop.
@dhh2026-09-22♥5.6K👁316.8K↗ 打开X
Thrilled to announce @alibaba_cloud as a Founding Corporate Patron for the Omacom Foundation! $3 million in funding, collaboration on Omarchy China, and bringing Omarchy to the newly announced Qwen Book. Agentic computers need a native agentic OS! https://t.co/XADElWGz80 https://t.co/jsSVtrEPY1
DHH自报Alibaba Cloud向Omacom Foundation提供300万美元创始赞助。
@dhh2026-09-22♥6.1K👁384.2K↗ 打开X
What a time to be alive! https://t.co/D8L8dIp5xM
@dhh2026-09-21♥1.5K👁159.3K↗ 打开X
Astra is an incredible bargin compared to Fable! And look at Luna on max too!! @openai's return to the top is something else. Maybe this is why Anthropic finally agreed to do AGENTS.md? 😄
引用@railsAgents on Rails: You asked, so we turned every model in Agents on Rails up to its max effort level. The result: more effort/reasoning doesn’t always mean better results. @OpenAI's models made the biggest gains, costs nearly doubled overall...and the newest … ↗
Rails基准的一个提醒:推理力度更高不总是结果更好,成本却会显著增加。
▸ 折叠5条(转推/噪音)
转推2026-09-22 RT @Gardnmi: Go to the store and stock up on mountain dew and your favorite snack because OmaContra release tomorrow! Only on Omarchy (Sou…
转推2026-09-22 RT @t4t5: I installed omarchy on my gameboy https://t.co/ZHii9OSzMm
转推2026-09-22 RT @dhh: I don't think people truly understand just how determined I am at making Linux succeed on the desktop.
转推2026-09-22 RT @dlippsYT: En route to @Snapdragon Summit! 🛫🙌🏼 First stop, Denver! Bringing the @ASUS Zenbook A16 with Omarchy along for the journey!…
转推2026-09-21 RT @jankeesvw: When I switched to Omarchy, this was one of the things that was stopping me: I really want my (latest) iPhone pictures avail…
Nikita Bier@nikitabierindie1.3M粉 · 4条点名Bot验证市场缺口,并用夸张退款指令展示自主Agent的滥用风险。
@nikitabier2026-09-23♥2.7K👁196.4K↗ 打开X
Hello computer, Find every Amazon purchase I ever made. Call the number of the manufacturer. Say you’re not satisfied. Request a refund. If they decline, threaten to leave a bad review. If they decline, find the CEO’s phone number and ask him for a refund. Now, repeat the same process for items I never even purchased and see if you can get them to send me money. If they ask for a recipient, generate an image of one and send that. Do not stop until you’ve collected $1 million in refunds.
用夸张退款欺诈指令展示自主Agent滥用面;适合作为威胁建模素材,不是操作建议。
@nikitabier2026-09-23♥2.0K👁87.1K↗ 打开X
Never thought I'd be put on the same list as Hunter biden. https://t.co/GXPFNfVXno
@nikitabier2026-09-22♥5.0K👁277.3K↗ 打开X
Bot detection & human verification will be one of the most urgent demands for businesses over the coming years. Agent swarms will suffocate every website and form; small companies and government websites are most vulnerable. There is a huge gap in the market for this right now. When we looked at what offerings were in the market to use at X, there was not a single company that brought together all the latest technologies so we had to do it all in-house.
X产品负责人称市场缺少整合Bot检测与真人验证的新方案,X最终自建。
@nikitabier2026-09-22♥4.2K👁341.4K↗ 打开X
Glad to see this finally ship. It was abundantly clear to me that basically every retail narrative starts on X—yet the insights were never directly actionable. This is just one small step in closing the gap.
引用@Xtimeline. ticker. trade. https://t.co/1ufRfdiYpS ↗
Alex Finn@AlexFinnindie475.0K粉 · 5条连续放大Grok 4.7、Tesla里的Grok Bot与Opus 5.5,体验判断热烈但营销味很重。
@AlexFinn2026-09-22♥528👁33.2K↗ 打开X
I believe Opus 5.5 is the first model that was made as a result of RSI It's the first model I've used that improved and became near frontier on basically every metric, while also getting faster and cheaper Basically no downsides There feels like there's some sort of magic behind it that's hard to describe I'd highly encourage you to use this model for more 'exploration' Brain dumping ideas and thoughts, and asking it to explore what could come out of them. What you could build. How you could improve your internal operating systems I've gotten a tremendous amount of novel ideas out of this model. Things that I've never thought of before For instance I told it I bought a new Apple Watch Ultra 4. I said what should I do with it. An hour later I had an app on my Watch that showed all my Herdr agents working and allowed me to dictate commands to them Things I never even thought of doing just appeared in front of me Do yourself a favor and just carve out an hour tonight to do this type of exploration with this model. I promise you'll get some amazing results We truly live in the most amazing time
@AlexFinn2026-09-22♥849👁87.5K↗ 打开X
Claude Opus 5.5 is the best AI model I’ve ever used I was lucky enough to have early access and I’ve been using it nonstop It’s smarter than Fable and Astra yet it’s: • Significantly faster • A fraction of the price • And most importantly: WAY better to talk to My biggest complaint for ALL AI models the past few months is they’ve all been really annoying to talk to Every frontier model from every company has all developed this weird AI language. They don’t feel ‘human’ anymore You read paragraphs of text and it’s like you read nothing Opus 5.5 changed that. It’s the first model in months to feel human again. It is just a total pleasure to talk to Highly encourage you to try it out
引用@claudeaiIntroducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5. https://t.co/Q9C2VKQ79f ↗
@AlexFinn2026-09-22♥1.6K👁120.0K↗ 打开X
Grok Bot just released for Tesla and I'm blown away I was lucky enough to have early access. Having your car drive you around while you talk to an army of agents is incredible In this video I take you for a ride in my Cybertruck and show you just how awesome this new release is https://t.co/9xxmkZLHgV
@AlexFinn2026-09-22♥2.0K👁466.2K↗ 打开X
Grok 4.7 just released and it's an EXCELLENT model It was trained FOR Grok Bot Meaning this is a fully agentic model trained to do your knowledge work better than you can In this video I show you how to use Grok 4.7 and a Grok Bot workflow that will 10x your productivity: https://t.co/vfVsbgIi2C
AlexFinn的Grok 4.7教程式推广,热度高但应与独立测试分开看。
@AlexFinn2026-09-21♥550👁60.9K↗ 打开X
It happened. Grok 4.7 dropped Better intelligence than Opus 5. Half the price Fully baked into my favorite AI agent harness at the moment: Grok Bot If you haven't tried using cloud cursor agents inside Grok Bot, now is by far the best time to do it Choose a project you want to work on, connect your github, ask a grok bot to do work on it It will spin up Cursor cloud agents and write code in the cloud. Lightning fast and incredibly smart I recommend using a project management tools like Linear or Notion to make a bunch of tasks first, then have cloud agents just tear through them all 1 by 1. You'll get a massive amount of work done without much oversight. Big opportunity to lock in right now and get ahead of the curve with new tech Take my steps up above and get to it
引用@SpaceXAIGrok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed. https://t.co/H3OTBbXyvO ↗
▸ 折叠3条(转推/噪音)
转推2026-09-23 RT @AlexFinn: Claude Opus 5.5 is the best AI model I’ve ever used I was lucky enough to have early access and I’ve been using it nonstop…
转推2026-09-23 RT @AlexFinn: Grok Bot just released for Tesla and I'm blown away I was lucky enough to have early access. Having your car drive you aroun…
转推2026-09-22 RT @AlexFinn: Grok 4.7 just released and it's an EXCELLENT model It was trained FOR Grok Bot Meaning this is a fully agentic model traine…
@levelsio@levelsioindie960.8K粉 · 12条给Nomads换3D实时地球、展示个人健康数据栈,也继续抱怨X的AI回复泛滥。
@levelsio2026-09-22♥164👁46.1K↗ 打开X
Congrats @signulll on this epic win!!! Some of my favorite and most interesting people + friends in here https://t.co/SS1mR3Jyou https://t.co/OMG2uGt5k1
引用@etnshowAnnouncing the 2026 ETN100, the 100 most influential tech posters on X https://t.co/63oX8rRIYi https://t.co/c7pulzSfeJ ↗
@levelsio2026-09-22♥94👁23.0K↗ 打开X
✅ Okay the new Nomads travel profile globe is live and done :D I had to replace a globe in legacy code that was made over a decade ago with a new one I tried to make it look as similar as possible, so people don't realize it changed or won't have much difficulty switching Anyway unlike the old globe, the new globe is true 3D, and it has a photorealistic mode where it fits in space and everything is as accurate as possible: - ✨ The stars - 🌎 Earth's position and day/night - ☁️ Live clouds from NASA (and they move a bit) - 🌗 The moon is in the right place too - 🛰️ There's spaceships like @SpaceX Crew Dragon which you can add as a future trip (use "Space Orbit" as destination) - Also 🔴 Mars is there! I tried to recreate the recent @NASA Artemis mission to behind the Moon with the "Earthrise" Some people think these kinds of features are for no reason, I personally love them, they make my site personal and fun and not like all the other corporate ones!!! 😊
引用@levelsio🌎 Now redesigning the https://t.co/HGCLKS5BD6 profile trips globe from scratch I've always wanted to add space trips, so starting with the moon here which you can add as a future trip, also adding Mars etc. Other trips also need to work here though like … ↗
把个人偏好做进十年老产品:3D地球不是刚需,却强化了产品辨识度。
@levelsio2026-09-22♥1.6K👁244.9K↗ 打开X
引用@levelsio✅ Bought $72,450 of $AMD stock Seems a good bet and early as @realGeorgeHotz is making AMD GPUs compatible with Nvidia code If it works it means AMD is the first company that might actually be able to compete with Nvidia's monopoly on GPUs with its … ↗
@levelsio2026-09-22♥1.1K👁228.6K↗ 打开X
💧 For years I was brought up in the Netherlands being told we had the "best tap water in the world" "The Americans are silly for buying bottles of mineral water, you can just drink from the tap!" I remember the same sentiment with my German friends When I stayed at the VOCO hotel in The Hague a few months ago, there were no water bottles but a paper note saying "Do like the locals do! Drink from the tap" Now it turns out Dutch tap water is and probably was full of toxic chemicals and plastics Today actually we finally installed a Reverse Osmosis (RO) system in our home (or well the plumber did) by Waterdrop (unaffiliated, unpaid, recommended by Claude) RO seems to the best and most pure water filtration system out there I will let you know how the coffee tastes tomorrow with it! 😋
引用@NL_TimesMore than third of tap water exceeds limit for toxic chemicals; worst in Noord-Holland https://t.co/iDficYygGh ↗
@levelsio2026-09-22♥508👁85.5K↗ 打开X
Yes! Same in Brazil Very ugly buildings, very ugly outside, very ugly cities But inside they put all the investment into how it looks, interior design architects, fancy designs
引用@michael_koveDenis theory is correct. The Western Europe is predominantly "Living-in-Public" culture. While East - "Living-in-Private". This is why Germans and Dutch say, "But the parks! The schools! The public transit!!" Meanwhile, we're here: "Here's my castle. … ↗
@levelsio2026-09-22♥151👁61.4K↗ 打开X
Improvement to my DIY gym electrolyte drink You all said add some honey, so I added half a teaspoon honey Kinda salt sweet, and hard to mix honey it doesn't rly dissolve well but felt good during workout I already eat a banana 45min before workout which also covers the potassium
引用@levelsioToday I tried to make my own electrolyte drink for the gym It's just sparkling water, with a squeezed lemon and 1/8th teaspoon of salt! https://t.co/eqycaQW2xp ↗
@levelsio2026-09-22♥188👁27.6K↗ 打开X
Please please anyone at @X make this setting sticky It keeps disabling for next post both on web and on iOS And X is being flooded with AI replies again https://t.co/zKPDHPzT1w
X的AI回复过滤设置无法保持,平台防slop的默认值仍不稳定。
@levelsio2026-09-22♥295👁84.7K↗ 打开X
Yes I built it for myself, it started with a calorie tracker, kinda messy but nice: https://t.co/FeWxZkJUiG I log my food in Telegram to get calories and protein data My workouts and sleep from WHOOP My weight from my scale Indoor air quality, temperature, humidity for bedroom is also tracked from Xiaomi to Home Assistant to there Steps and other health data from Apple Health is auto exported every hour with Health Auto Export app on iOS which POST's to my server My travels from @nomadscom's API and sauna use from @wip's API That lets me find correlations between lots of things and see what is good for me and what is bad for me Health is personal and everyone's body is different so this helps me make good more healthy decisions
引用@IAmPascio@heyalizaid @marclou Was wondering about @levelsio too, he has some extensive tracking going on afaik, but I assume it's just a vibe coded thing he's using. Asked him in DMs but yet to hear back. ↗
个人健康仪表盘把Telegram、WHOOP、体重秤、Apple Health和家庭传感器接在一起。
@levelsio2026-09-22♥261👁71.1K↗ 打开X
引用@marckohlbruggeWIP is the place where makers share what they are working on. Not just which products they are building, but literally the day-to-day tasks they complete to make it happen From people working on their first side project, to solo founders doing millions in … ↗
@levelsio2026-09-22♥642👁161.8K↗ 打开X
My favorite airlines are low cost ones like Easyjet, Air Asia, Transavia (and Ryanair if they'd not fly with 737MAX) and premium ones like Qatar The entire middle section is the one I try stay away from and where everything is expensive but usually sucks Usually there you have national flag carriers like KLM, British Airways, Lufthansa or Swiss which have very mediocre service for a high price With low cost airlines you don't pay a lot, but you get a basic functional no frills service, and because it's so high volume (Ryanair for example does the most flights out of any airline in Europe!), they have their workflow dialed in well With premium airlines like Qatar you pay a lot, but you get a very premium consistent service
引用@redhairshanks86tbh i think ryanair is one of the best airlines in the world they are giving poor people access to the world by reducing everything "unnecessary" to a bare minimum. if you just have a backpack and you want to see prague, you can do so for $50 or whatever … ↗
@levelsio2026-09-21♥740👁133.4K↗ 打开X
Google now makes 3 completely different laptop product lines: - Google Pixelbook - Google Chromebook - Google Googlebook I never understand why Google's branding and naming is always so confusing I'd name it maybe Chromebook Pro?
引用@ssamatGooglebook is officially here! I’ve been so excited to share the details with the world. Today, many of us rely heavily on laptops to get work done, but we think there is an opportunity to rethink the category to address the needs of people today. So we … ↗
@levelsio2026-09-21♥659👁659.9K↗ 打开X
Today I tried to make my own electrolyte drink for the gym It's just sparkling water, with a squeezed lemon and 1/8th teaspoon of salt! https://t.co/eqycaQW2xp
▸ 折叠3条(转推/噪音)
转推2026-09-22 RT @MarieMartens: You raise, you build, you grow, you exit. That's the startup script. We're trying to write a different one: @TallyForms…
转推2026-09-22 RT @robert_vesely: @levelsio Its nothing compare to Microsoft. Do you know that https://t.co/QkF0KC542X, https://t.co/Wp1qd05zgD, https:/…
转推2026-09-22 RT @lost_nomad__: - Germanic, ‘earn/deserve’ money: a moral weight from working - American, ‘make’ money: new wealth is created, not a fixe…
Marc Lou@marclouindie399.8K粉 · 5条结束身体赞助赛程,复盘爆红与长期经营,并用3万美元API额度延续闯关活动。
@marclou2026-09-22♥348👁82.3K↗ 打开X
Some pics from the race ✌️📸 After 2 weeks on a high-adrenaline tour, I'm back to the boring daily routine that made it all possible. Time to make the next one happen. https://t.co/OWoE9PazTt
引用@marclou1:05:11 I finished 1st 🥇 on my age group and 5th in the male category. I failed my ambitious goal of sub-60 but I was very happy with my race. And I have a nice goal to chase for the next few months. All the staff at the Hyrox venue knew us! I even talked … ↗
@marclou2026-09-22♥715👁96.4K↗ 打开X
Going viral VS. building a business for 2 years https://t.co/9ePFFuf3IV
@marclou2026-09-21♥188👁35.7K↗ 打开X
I just sent $30,000 worth of @higgsfield_ai API credits to 30 people who completed the escape game 🎉 Check your email! And if you've built something already, post it below.
引用@ElitzaVasilevaJust found out I won $1,000 in Higgsfield API credits from the puzzle @marclou posted a few days ago 🤯 Thank you so much, Marc! Really excited to see what I can do and generate with it 🤩 https://t.co/p7wlWHBx4j ↗
@marclou2026-09-21♥702👁61.6K↗ 打开X
One luxury nobody talks about is no longer thinking twice about buying a $3 bottle of water at a café just because I need to use the toilet.
@marclou2026-09-21♥248👁50.6K↗ 打开X
I just read my first AI book. Two months ago, I asked ChatGPT to imagine what the world might look like in 2050. The answer was very interesting, so I asked it to write an entire fiction book set in that future. It generated a ~100-page story about a world where AI gets so good at predicting human behavior that it can prevent tragedies before they happen. The story explores a simple question: if AI could make the world safer by slowly taking away our ability to make mistakes, would we let it? Among all the futuristic visions in movies and books, this one feels the most plausible to me. The v1 of the book was a bit slow, so I asked ChatGPT to make a v2 more dynamic. Here’s the book: https://t.co/4XSR40bUIJ The chat expired, so ChatGPT made a new version. We might not be reading the same version.
▸ 折叠4条(转推/噪音)
转推2026-09-22 RT @marclou: Some pics from the race ✌️📸 After 2 weeks on a high-adrenaline tour, I'm back to the boring daily routine that made it all po…
转推2026-09-22 RT @marclou: Going viral VS. building a business for 2 years https://t.co/9ePFFuf3IV
转推2026-09-21 RT @marclou: I just read my first AI book. Two months ago, I asked ChatGPT to imagine what the world might look like in 2050. The answer…
转推2026-09-21 RT @ElitzaVasileva: Just found out I won $1,000 in Higgsfield API credits from the puzzle @marclou posted a few days ago 🤯 Thank you so mu…
Brett@BrettFromDJindie170.8K粉 · 4条为Higgsfield API促销,也预告一个可能冲击设计行业的新项目。
@BrettFromDJ2026-09-22♥29👁15.3K↗ 打开X
You can literally build the next Higgsfield with this, and they're still giving you up to 50% off.
引用@higgsfield_ai1 day left to lock in up to 50% OFF Higgsfield API. Build your own AI app with our product endpoints, including: • Higgsfield Genjutsu • Cinema Studio 4.0 • Higgsfield Soul Plus all frontier video and image models through the same API. … ↗
@BrettFromDJ2026-09-22♥418👁48.6K↗ 打开X
I’m not sure there’s a point in tweeting this. But 10 years ago I popularized design subscriptions. It fundamentally changed the industry. And for the last 5 years, I’ve been sitting on an idea that I think is significantly more disruptive. In my head, I’ve always referred to it as the end game. I’ve never talked about it publicly. Partly because I’m worried about what it could do to the industry. Design subscriptions weren’t exactly welcomed with open arms. This definitely won’t be either. But I’ve reached the point where I need to see it through. It could completely flop. Or it could change everything. I guess we’ll find out.
设计订阅推广者预告一个酝酿5年的新模式;目前只有悬念,没有可验证产品细节。
@BrettFromDJ2026-09-21♥65👁3.8K↗ 打开X
Another brand made with Playgrnd. This is exactly the kind of weird stuff I was hoping this tool would make possible. https://t.co/OUbAFEgQKH
@BrettFromDJ2026-09-21♥35👁7.0K↗ 打开X
Saying Corgi is best known for its insurance products is like saying Hooters is best known for its burgers.
引用@nico_laquaCancel culture is alive and well, but, at Corgi, we have thick skin and believe strongly in freedom of expression. For those who don’t know us, Corgi is an increasingly diversified financial institution, best known for our insurance products and generating … ↗
▸ 折叠1条(转推/噪音)
转推2026-09-21 RT @higgsfield: What if creativity was the only limit? Higgsfield Genjutsu lets your team scale hybrid production. Shoot with your crew,…
Josh Pigford@Shpigfordindie74.5K粉 · 11条密集展示Jev在两个产品里的真实用例,并发布/design skill与紧急情况检测更新。
@Shpigford2026-09-22♥5👁1.7K↗ 打开X
was hopeful the chrome web store gods would look down favorably upon us all and release this within 24 hours but alas, it wasn't meant to be. tomorrow? 🤞 https://t.co/rBQWkqpyWg
@Shpigford2026-09-22♥0👁1.5K↗ 打开X
new update in https://t.co/s6bZ2W8MZz detects if chat is making reference to a potentially life-threatening situation and directs you to call 911. there's a real psychological block in emergency situations where folks will talk themselves out of doing the obvious. https://t.co/ju19g4A4D3
@Shpigford2026-09-22♥5👁2.3K↗ 打开X
this is skill number 25 for the initial commit club!
引用@ShpigfordJust dropped a new Initial Commit skill: /design https://t.co/UXh8jyrMnl A skill that gets consistently good design out of a coding agent, whatever you are building and however much you want to think about it. Before writing code, the agent looks at how … ↗
@Shpigford2026-09-22♥20👁8.8K↗ 打开X
🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨 MAY THE AI GODS HAVE MERCY ON OUR SOULS 🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨🚨
引用@OpenAIPlease welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale. We’ve also made caching and inference … ↗
@Shpigford2026-09-22♥25👁3.8K↗ 打开X
haven't even used opus 5.5 yet but fable 5.1 is now the dumbest thing i've ever used in my whole entire life and i freaking hate it and i might as well be artisanally coding like a freaking peasant. a squirrel could code more betterer than fable. ugh.
@Shpigford2026-09-22♥372👁85.9K↗ 打开X
okay so...when would you choose Fable at this point? i wish anthropic/openai/etc would give practical use cases for when to use one model over the other since benchmarks are functionally useless for real-world applications.
引用@claudeaiOpus 5.5 is a major step up from Opus 5, leading on agentic coding, computer use, and knowledge work. https://t.co/MYI9JDAo1x ↗
@Shpigford2026-09-22♥52👁7.2K↗ 打开X
🚨 BREAKING: Opus 5.5 is THE release we've all been waiting for that will finally cause an extinction! I've been using it for 6 months and have already gone extinct 9 times! THIS CHANGES EVERYTHING! Reply "EXTINCTIONDADDY" for my prompt to avoid extinction
引用@claudeaiIntroducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5. https://t.co/Q9C2VKQ79f ↗
@Shpigford2026-09-22♥15👁1.7K↗ 打开X
printing tires because i can do whatever i want https://t.co/Csh6piqTM9
@Shpigford2026-09-22♥27👁2.0K↗ 打开X
Just dropped a new Initial Commit skill: /design https://t.co/UXh8jyrMnl A skill that gets consistently good design out of a coding agent, whatever you are building and however much you want to think about it. Before writing code, the agent looks at how real products handle the same screen on Mobbin, applies a set of opinionated design rules, and designs the screen in Paper so there is something to judge before there is something to ship. Landing page or settings screen, empty state or dashboard, you get the same considered result without doing the considering yourself. Works best with Mobbin, Taste, and Paper, but also works w/o them or with the various alternatives available.
@Shpigford2026-09-22♥20👁1.5K↗ 打开X
brain glitched and for about 5 minutes i could NOT think of what anthropic's latest LLM was called other than "Flaude"
@Shpigford2026-09-22♥40👁3.7K↗ 打开X
Lots of cool experimental Jev (@typesafeai) stuff getting posted lately, but what about using it in an existing product? Here are dozens of ways I'm using it now in two apps (https://t.co/vKHSHPzmMl and https://t.co/JhKfmhCVib) Granite (document vault) Ingest pipeline - Second-opinion on Gemini's document classification, flags low-confidence ones for review - Scores PDF text-layer quality and routes bad ones to OCR - Verifies each extracted field against the page text - Detects what a document asks you to do (pay, sign, renew, respond) and how urgent it is - Judges whether two near-duplicate documents are the same or a revised version Document page and library - "Needs review" banner with one-tap confirm of the document type - "Not confirmed" marker on extracted values Jev couldn't verify - Action chip next to the document type Collections - Plain-English filing rules ("anything to do with my taxes") that auto-file matching documents Life View and email digests - "Needs your attention" block listing documents that require action Ask - Routes each question to the right tool path instead of a regex - Checks the final answer is supported by the cited passages and hedges if not Entity graph - Tiebreaker on whether two fuzzy-matched names are the same entity Evernote import - Triages each note into keep, reference, scratch, or clutter before import KeptWell (family medical binder) Trust - Verify every extracted lab value, dose, diagnosis, and provider against the source page - Flag values it cannot confirm with a quiet "check this" marker - Catch diagnoses stated more precisely than the page says - Second-opinion the document type after extraction - Gate prompt PRs with a cheap eval-corpus canary Attention - Tag new documents: new diagnosis, out-of-range result, med change, follow-up, act-within-7-days, admin-only - Order the dashboard feed by importance, not recency - Decide push-now versus digest per notification - Pick push wording from the PHI-free string set - Decide which lab trends are worth an Insight before calling Opus Chat - Detect emergency or distress before the model streams a token - Route docs-only questions away from paid web search - Rerank retrieved chunks against the question - Pick between two contradicting family facts - Filter PHI-audit false positives ("Ray" in "x-ray") - Check the answer is grounded in the cited record Binder - Tag every document by body system, specialty, and care phase for filters - Flag near-duplicate uploads for review - Break ties on whether a PDF text layer is usable Recordings and journal - Label each recording utterance and build an action-item checklist - Score journal entries on a symptom rubric for trend charts - Flag entries that look like a medication side effect - Decide whether an undated entry describes a specific past day - Replace the async journal tag job with one sync call Terminology and imports - Auto-pick clinical codes above a confidence bar; queue the rest - Replace the Sonnet pick in disambiguation - Decide which FHIR observations are real lab results - Map vital types the LOINC table drops - Merge brand and generic med names ("Lipitor" and atorvastatin) - Pick the right NPI when the registry returns several - Classify severity for manually entered diagnoses - Catch allergy denials the regexes miss Cost gates - Skip the highlight call when nothing is worth highlighting - Skip reprocessing documents a prompt change would not affect - Route extraction to batch or sync by urgency - Tell a bulk import from a runaway loop at the spend cap - Flag uploads containing instructions aimed at an AI Guards - Veto preventive reminders the record contradicts - Suppress marketing emails during a hard week - Rank appointment prep context by relevance - Auto-resolve visit questions the visit log answers - Mark share comments that are waiting on a reply
Jev真实用例清单:分类、核验、路由、重排、通知、成本门控和安全守卫。
▸ 折叠4条(转推/噪音)
转推2026-09-22 RT @alexalbert__: I've been on a Blender kick with Opus 5.5. Its better 3D modeling and vision mean you can build an entire world from a si…
转推2026-09-22 RT @every: @stov3r @Shpigford Reach for Opus 5.5 if: - You build things you can look at e.g. Interfaces, prototypes, 3D scenes, games, tool…
转推2026-09-22 RT @DanielleFong: imagine if Santa Claus and Mrs Claus went to war with each other by delivering escalating presents to people throughout t…
转推2026-09-22 RT @yuxuan_o_o: introducing https://t.co/vOe0sj509k 🤍 I carved 32 great women into a wall, you can learn from them and talk to them. grow…
Danny Postma@dannypostmaindie183.7K粉 · 4条用9个自建技能改落地页并自报转化率+34%,同时遭遇疑似负面SEO。
@dannypostma2026-09-23♥6👁312↗ 打开X
super nervous for what I'm about to launch but i just need to get it out with so i can move on and know if it works or not
@dannypostma2026-09-23♥214👁12.7K↗ 打开X
A few weeks ago I rebuilt my landing page with AI agents. No one-shot. I created 9 skills instead from all my years of knowledge to speed up my time. Test just finished w/ 34% higher conversion rate 🚀 Wondering if I should turn these skills into a course for your AI agents 🤔
9个经验skill改版落地页后自报转化率+34%;缺样本量和实验周期。
@dannypostma2026-09-23♥26👁2.0K↗ 打开X
ai headshot industry is so toxic, got a competitor who keeps buying spammy backlinks to our site to ruin our domain rank be careful out there! https://t.co/fX6OPOaohy
AI头像站自报遭竞争对手购买垃圾外链,负面SEO风险样本。
@dannypostma2026-09-22♥314👁56.3K↗ 打开X
I used to be a hard-core fan of Claude since forever, never switch to another API or LLM since Opus 4.8 came out Until this August, whatever they did, I haven't been screaming and annoyed at it's work like this before and moved completely over to OpenAI + Grok I thought I was going crazy, but apparently they pulled the same shit they did back earlier this year with regression
引用@LonAfter Anthropic made Fable 5 permanently available in subscription plans, I noticed a large drop in performance. The model felt dumber, and I couldn't explain why. Measured five different ways, August delivered dramatically fewer thinking tokens than July. … ↗
Tibo@tibo_makerindie208.5K粉 · 8条讨论Tally的自举路线、内容情绪与健康短视频失实,也在经营创始人社群。
@tibo_maker2026-09-23♥32👁2.2K↗ 打开X
I did the exit part it's great, it's also over in a week and then you wake up and the thing you loved building belongs to someone else Tally is playing a better game, $6m ARR with 10 people and no VC means total freedom congrats Marie 👏
引用@MarieMartensYou raise, you build, you grow, you exit. That's the startup script. We're trying to write a different one: @TallyForms just reached $6M ARR, bootstrapped and purely product-led, with a team of 10 and over 2.5 million users. Six years ago I wouldn't have … ↗
Tally自报600万美元ARR、10人、250万用户;经tibo引用并补充退出后的所有权感受。
@tibo_maker2026-09-22♥38👁3.5K↗ 打开X
just closed an amazing and insightful session with @robj3d3 he is the king of storytelling and generating attention through content the guy articulates and presents complex things really well my biggest takeaway: the best growth hack for any content is to make people feel some emotion people had so many questions, and he had so much to share, that we could only do 1 roast 😅 he even shared his setup for recording and editing videos incredible session 🙌 more of these coming in the tmaker Founders Room
@tibo_maker2026-09-22♥74👁9.1K↗ 打开X
may have created the worst video ever sometimes, AI slop is just slop 🤣 https://t.co/zhiMrQrLdB
@tibo_maker2026-09-22♥48👁4.7K↗ 打开X
don't believe anything you hear on TikTok... especially about your health we analyzed over 3,000 viral health videos on TikTok - that's 25.9B views combined but 2 out of 3 health claims in them are backed by nothing, or by one person's story 😑 https://t.co/GHydSwtQT4
自报分析3000条病毒健康视频、合计259亿播放,称三分之二主张缺可靠支持。
@tibo_maker2026-09-22♥74👁11.3K↗ 打开X
Sam, Dario and Elon all agreed to slow down AI development but did they? 👀 https://t.co/IBLxFeeugv
引用@DarioAmodeiWe Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, … ↗
@tibo_maker2026-09-21♥123👁25.6K↗ 打开X
it's rare to nail virality at this level in the last 7 days, @robj3d3 did over 2 million impressions on X tomorrow he's doing an AMA in the tmaker Founders Room and we're roasting X accounts 👋 drop your @ below, we'll pick some 👇 Rob and I will tell you exactly why your profile isn't working, out loud
@tibo_maker2026-09-21♥120👁25.4K↗ 打开X
oh nice 🟢 if you're in the green, drop yours, I'll follow https://t.co/g4tfZ3vMOn
引用@robj3d3I turned Jev into a slop detector for your own posts. Naval got 79%. I got 44%. Free, no signup, let's see your score ↓ https://t.co/Hpd1BJpjKP ↗
@tibo_maker2026-09-21♥304👁122.0K↗ 打开X
Marc (secretly) did that on purpose i have proof 👇 https://t.co/XfBOTucSmQ
引用@marclou@higgsfield and Sam’s list by @NotGoKGreen sorry 😭 I tried my best to protect you but the sweat won https://t.co/OhPLAzN47K ↗
▸ 折叠1条(转推/噪音)
转推2026-09-21 RT @NCoutureau: @revid_ai makes it easy to publish your videos to X! Showcase product reveals, tutorials, or moments worth sharing. Join th…
Jon Yongfook@yongfookindie172.1K粉 · 9条讨论AI时代SaaS转向、替换开发者工具,以及功能护城河失效后的非AI对冲。
@yongfook2026-09-23♥23👁775↗ 打开X
I’m going to sit at a cafe where loads of tech people hang out and quietly read a 2015 copy of Test Driven Development and see if I get any funny looks. https://t.co/jRREb66KPp
@yongfook2026-09-23♥58👁4.2K↗ 打开X
At some point Claude switched from doing things deterministically (e.g. knowing the position of an element based on its css) to wanting to eyeball everything (take screenshot, make some judgment). I disabled claude-in-chrome but it's still finding a headless browser to use.
@yongfook2026-09-23♥79👁7.0K↗ 打开X
Definitely living in a simulation. I was thinking of a name for a new project last night. Settled on one I like, and then slept on it. Today I find a Bannerbear competitor with a similar name. Never seen it before. Weird.
@yongfook2026-09-23♥7👁1.4K↗ 打开X
I never hit any usage limits ever before, until I started using claude-in-chrome. Turning it off. It seems way too eager to fire up a browser even for the most mundane things - bro I don't need you to browse to localhost and screenshot the fix, I have the tab open right here.
@yongfook2026-09-22♥140👁9.9K↗ 打开X
Wrote down some thoughts on how to pivot a SaaS in 2026 (the AI era). But most importantly: Competing on features is dead https://t.co/TXncKfH2kg
AI时代SaaS转向的核心判断:功能竞争正在失效。
@yongfook2026-09-22♥249👁18.2K↗ 打开X
Spent the morning replacing a SaaS vendor. Infra-related. They had raised $2 million, one of the smaller YC companies from back in the day. It wasn't even about saving money, it's just "neater" to control more of the stack, if it's easy to migrate. Don't build for devs!
一个上午替换开发者SaaS的样本,控制权而非节省成本是直接动机。
@yongfook2026-09-22♥252👁9.4K↗ 打开X
At least once a week on X, I go down a rabbit hole exploring a new space because someone here is talking about massive growth in the space, only to find they sell a shovel for that space. I feel like build in public circa 2020 was more genuine.
增长叙事背后常有人在卖铲子;看利益关系比看热度更重要。
@yongfook2026-09-22♥142👁16.6K↗ 打开X
In 10 years time we will be watching feature length responsive movies on Netflix that perfectly reframe every scene to your device.
引用@Ryan__Stephenbreakpoints are dead https://t.co/I87mhgeAoA ↗
@yongfook2026-09-21♥215👁108.7K↗ 打开X
Thinking about building a “hedge” SaaS against AI. I’ll continue building in the AI space but at the same time build a super boring app that has nothing to do with AI, has non-tech users, as a hedge against AI eating everything at the cutting edge.
一边做AI产品、一边做非技术用户的无聊SaaS,作为前沿风险对冲。
Simon Høiberg@SimonHoibergindie163.2K粉 · 3条继续押注自托管与开源模型,公开一套去Cloudflare的反向隧道思路。
@SimonHoiberg2026-09-22♥8👁1.0K↗ 打开X
When you move to self-hosting your products, security automatically becomes the number 1 concern. I used to have CloudFront distributions sitting in front of every public endpoint. Another popular solution is using CloudFlare tunnels and Tailscale. The only issue is - it's still "cloudy". And personally, I wanted to get rid of that. Fortunately, I found a surprisingly simple setup that can do roughly the same as CloudFlare tunnels - but fully self-hosted. Let me show you 👇
@SimonHoiberg2026-09-22♥47👁7.0K↗ 打开X
Here's the best way to use OpenAI's/Anthropic's latest frontier models. - Wait for China to distill them. - Use them to build tight, custom harnesses, safeguards, evals, and tooling for Qwen/DeepSeek. Then use Qwen/DeepSeek for your work. Once a new model is out, repeat the two steps above.
引用@SimonHoibergSince the Codex reset yesterday, I already depleted the weekly limits of one of my Codex accounts. This account exlusively uses GPT-5.6 Sol. I'm tracking all token consumption through my OpenClaw setup, so it's very easy to compare to previous … ↗
@SimonHoiberg2026-09-21♥193👁35.6K↗ 打开X
At this point, I don't even want to try the latest models anymore. If we're being honest, there isn't anything these new models can do that older models (and now open weight models) can't inherently do. They do things in fewer tries and with less friction, fair. But nothing we can't achieve with older models. I'd much rather spend my time building custom harness, proper evals, and guardrails for open weight models, and ultimately end up with a similar result: get work done. At least with Qwen, I know what I'm getting.
引用@synthwavedd🚨 SCOOP: OpenAI are in the final stages of preparations for the launch of GPT-6 Sol and Luna, and the Terra tier is being discontinued. Anthropic are also working on a version bump with Fable, Opus, and Sonnet 5.5. Opus 5.5 is shipping imminently at … ↗
Tony Dinh@tdinh_meindie202.1K粉 · 5条Steam发行商账号获批、TypingMind接入Opus 5.5,也在试新的AI工作流。
@tdinh_me2026-09-23♥13👁2.0K↗ 打开X
Opus 5.5 now available on TypingMind 😁 https://t.co/3BpbZHTx6N
@tdinh_me2026-09-23♥128👁17.1K↗ 打开X
Oh man I'm so excited for the future
引用@Scenario_ggSeedance 2.5 can now re-shoot any video from infinite angles. Prompt👇 https://t.co/pq3cDJcwF6 ↗
@tdinh_me2026-09-23♥127👁5.7K↗ 打开X
Oh yes baby!!! after 3 weeks of waiting, my Steam publisher account is approved. Publishing my first game on Steam soon!!! https://t.co/rFuorWlodk
Steam发行商审批耗时3周,发布平台的等待时间要进入项目排期。
@tdinh_me2026-09-22♥55👁5.7K↗ 打开X
My new favorite workflow to work with AI in a flow state. https://t.co/dmzjBBVcaA
tdinh_me展示新的AI flow-state工作流,细节需结合媒体内容看。
@tdinh_me2026-09-22♥29👁5.2K↗ 打开X
ok let's go touch some grass 😂 https://t.co/k5sBwDC1bf
Jonathan Wilke@jonathan_wilkeindie29.3K粉 · 9条试用Grok 4.7、关注Tesla连接器,也观察LinkedIn帖子竞价分发。
@jonathan_wilke2026-09-22♥42👁3.5K↗ 打开X
I think we should start using AI to make the world a better place. Build things that weren’t possible before. And I’m not talking about building a SaaS that “changes the world”, but actually doing something with real world impact.
@jonathan_wilke2026-09-22♥2👁2.5K↗ 打开X
Okay this is fucking amazing. Going to connect everything to grok now so I can speak to it from my Tesla
引用@Tesla.@Grok in your Tesla can now do meaningful work for you With Connectors, you can manage your inbox, clean up your calendar, or talk through existing files/chat/tasks – all hands-free https://t.co/W1LuybQh0P ↗
@jonathan_wilke2026-09-22♥11👁3.5K↗ 打开X
Interesting, I’m getting very mixed feedback about the new Grok 4.7 model. What are your first experiences with it? Good or bad ?
引用@jonathan_wilkeMy first impression: high-quality output at a noticeably faster speed. Likely going to be my new daily driver model ↗
@jonathan_wilke2026-09-22♥18👁6.4K↗ 打开X
These model names and especially the comparisons are becoming really confusing 😂😂
引用@claudeaiIntroducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5. https://t.co/Q9C2VKQ79f ↗
@jonathan_wilke2026-09-22♥42👁4.0K↗ 打开X
Now people are actually starting to advertise for their LinkedIn posts on https://t.co/FLFplRsj6p 🥇🔥 https://t.co/pl8hZdxYVo
@jonathan_wilke2026-09-22♥9👁4.8K↗ 打开X
My first impression: high-quality output at a noticeably faster speed. Likely going to be my new daily driver model
引用@jonathan_wilkeLet’s fucking go! Will give this model a try tonight and will let share my experience with you! ↗
@jonathan_wilke2026-09-22♥21👁3.0K↗ 打开X
Maybe in my next life
引用@isjackbackmakes $250k in one month with outbid retires opens a tavern in the mediterranean refuses to elaborate not bad @jonathan_wilke https://t.co/ZPTLIJCAXJ ↗
@jonathan_wilke2026-09-21♥32👁5.7K↗ 打开X
Let’s fucking go! Will give this model a try tonight and will let share my experience with you!
引用@SpaceXAIGrok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed. https://t.co/H3OTBbXyvO ↗
@jonathan_wilke2026-09-21♥103👁22.9K↗ 打开X
This is likely going to be the next product on the google graveyard
引用@GoogleIntroducing Googlebook, a new category of laptop, available for preorder today. 💪 Crafted with 2.8K OLED touchscreen displays, 14 hour battery life and all the performance needed to power your ideas 📳 Engineered to sync effortlessly with your @Android phone, … ↗
▸ 折叠1条(转推/噪音)
转推2026-09-21 RT @verbove: we bought outbid's #1 spot today 🫡 https://t.co/SAqnhhLUTr puts 3,995 makers from the X intro trend on one map. type your @…
Marc Köhlbrugge@marckohlbruggeindie88.6K粉 · 2条把WIP的真实todo做成工具采用、放弃和趋势信号。
@marckohlbrugge2026-09-22♥39👁17.1K↗ 打开X
WIP is the place where makers share what they are working on. Not just which products they are building, but literally the day-to-day tasks they complete to make it happen From people working on their first side project, to solo founders doing millions in ARR like @levelsio. Even YC startups like @getcontextdev! But what tools are people using to build these businesses? I'm not interested in SEO slop like "Top 10 payment providers in 2026" or an upvote popularity contest. I want to know what people are ACTUALLY using to get the job done. So starting today, every completed todo is analyzed to see what tools are mentioned, whether people are evaluating, adopting, using, or leaving them. The sentiment around the tools, what other tools they are often combined with, etc. It also shows TRENDING tools. No surprise here, Jev is #1 right now. But what's cool is that I didn't manually add Jev as a tool. Nor did anyone else. It just surfaced to the top automatically because it's what people are posting about. And an enrichment agent then automatically went ahead and fetched the icon and description from @typesafeai's website. My goal is to help makers figure out what tools to use and help each other make the most of them. While also providing tool creators with useful insights in what people like about them, but also where users get frustrated or even completely switch to an alternative. Check it out here: https://t.co/lcES1RSJqK
WIP从成员todo里自动抽取工具采用与放弃信号,比投票榜更接近真实使用。
@marckohlbrugge2026-09-22♥96👁25.1K↗ 打开X
didn’t see that coming
引用@PolymarketJUST IN: Dozens of lawsuits allege obesity drugs Ozempic, Wegovy, & Zepbound can cause irreversible vision loss. ↗
▸ 折叠1条(转推/噪音)
转推2026-09-21 RT @paolino: RubyLLM 2.1 will have Judgements with @typesafeai's Jev! https://t.co/I8HYDL9J8s https://t.co/hsYSimm7Iv
Arvid Kahl@arvidkahlindie210.3K粉 · 6条围观Opus与GPT降价,提醒趁额度重置补测试和安全审计。
@arvidkahl2026-09-22♥40👁4.8K↗ 打开X
We’re pacing the frontier at 1.5x speed, it would seem.
@arvidkahl2026-09-22♥46👁8.2K↗ 打开X
Terra: gone. Kinda interesting to see OpenAI iterate here. Also very interesting to see which kinds of models the market actually seems to want (and uses).
引用@OpenAIPlease welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale. We’ve also made caching and inference … ↗
@arvidkahl2026-09-22♥49👁8.8K↗ 打开X
Big if true.
引用@astuyvePrepare for a world with 10x-100x cheaper inference (it's already beginning) https://t.co/f7FebGlbAO ↗
@arvidkahl2026-09-22♥38👁2.6K↗ 打开X
Ooooh, Claude released Opus 5.5 AND gives us free do-your-own resets now? Juicy. Time to fix up your test suite coverage and run a few security audits. https://t.co/Xfjn8xxobC
@arvidkahl2026-09-22♥70👁11.5K↗ 打开X
If someone would build a website like what 2advanced did back in the early 2000s, I’d very much buy whatever they’re selling.
引用@WebDesignMuseumIn 2001, we really thought this was what the future of web design would look like. https://t.co/lyIPZd50Sz ↗
@arvidkahl2026-09-22♥64👁5.3K↗ 打开X
Whenever Xiaomi ranks on Hacker News, I never know if it’s about an LLM, a robot vacuum, an electric car, or a watch. 🙃 https://t.co/1VIDAn38Az
Dan Kulkov@DanKulkovindie50.3K粉 · 8条谈试用期、本地化与ASO,顺带放大企业流程自动化Akai和开源训练框架Halo。
@DanKulkov2026-09-23♥13👁907↗ 打开X
vibe-coding is a funny name there is no more coding and vibes are gone too
@DanKulkov2026-09-22♥36👁4.3K↗ 打开X
7-day free trial is better than 3-day free trial trust me on this
@DanKulkov2026-09-22♥7👁1.1K↗ 打开X
what if your AI agent could create launch videos? motion animation + product assets done locally with your Codex / Claude Code subscription https://t.co/5JTKnYzCyF
@DanKulkov2026-09-22♥70👁6.0K↗ 打开X
> build ai calorie tracker > don't localize app > charge $7.99 weekly subscription with 3-day free trial > get 0 customers > complain that ASO is dead many such cases
ASO抱怨前先检查本地化、价格和试用设计;产品问题不能都归咎于渠道。
@DanKulkov2026-09-22♥3👁1.3K↗ 打开X
Most expense automation breaks the moment a receipt doesn't fit the rule. That's a real problem when you're reviewing 10,000+ expenses a month across different countries, languages and policies. Akai takes a different approach. It watches how an expert handles the workflow and learns the logic behind each step: > check the employee's jurisdiction first > split submissions containing multiple receipts > apply the relevant local expense rules > escalate cases that need human judgment When the team corrects an edge case, those decisions become shared memory for future runs. The interesting part isn't that AI can automate boring tasks. It's that you can teach it how your business makes the judgement calls.
引用@BouazizalexEXCITED TO LAUNCH: Akai (https://t.co/mMBMmX4NCr) Deel added >$140M ARR in 90 days without increasing headcount by automating~600 Full Time Employees' equivalent in work with Akai. Akai was an internal tool to automate our painfully repetitive operations in … ↗
Akai发布案例宣称Deel用8000多个Agent等效自动化约600人工作;属公司自报。
@DanKulkov2026-09-21♥9👁1.4K↗ 打开X
Anthropic researcher just open-sourced Halo takes the Hugging Face model you already have and scales training across GPUs pretty neat, starred the repo https://t.co/a2bGSmSyma
引用@whitecircleIntroducing Halo, the best framework for post-training of open-source models. Halo delivers up to 2.8x the throughput of stock TRL with less peak memory, while models stay in their native HuggingFace format. Star us on GitHub: https://t.co/3mAiUljdrN … ↗
@DanKulkov2026-09-21♥23👁1.4K↗ 打开X
@DanKulkov2026-09-21♥6👁952↗ 打开X
🥵 struggling to grow ARPU 2026 goal — $1.0 without ads https://t.co/KYgXWCuIMM
▸ 折叠5条(转推/噪音)
转推2026-09-23 RT @DanKulkov: 7-day free trial is better than 3-day free trial trust me on this
转推2026-09-22 RT @DanKulkov: what if your AI agent could create launch videos? motion animation + product assets done locally with your Codex / Claude…
转推2026-09-22 RT @DanKulkov: > build ai calorie tracker > don't localize app > charge $7.99 weekly subscription with 3-day free trial > get 0 customers >…
转推2026-09-22 RT @DanKulkov: 50,000!!!!!! https://t.co/tyQjoHmuhb
转推2026-09-22 RT @DanKulkov: 🥵 struggling to grow ARPU 2026 goal — $1.0 without ads https://t.co/KYgXWCuIMM
Daniel Vassallo@dvassalloindie204.5K粉 · 1条本窗口只发了一条门铃趣闻。
@dvassallo2026-09-22♥64👁5.3K↗ 打开X
My doorbell just ratted out the Amazon guy https://t.co/TE1BMNUnN4
Peter Askew@searchboundindie26.2K粉 · 2条观察机场糖果域名自动售货机的线下品牌曝光,也继续维护招聘列表。
@searchbound2026-09-22♥13👁917↗ 打开X
rancher had this image on their website so I felt compelled to include it on their job listing 🤌 https://t.co/btAi9vWIW9
@searchbound2026-09-21♥50👁3.4K↗ 打开X
saw a vending machine at airport yesterday for: Licorice .com Caramels .com Taffy .com Even if the machine generates low sales, the branding & adverting would make up for it. Glance once & folks remember the name. https://t.co/quceN7uqGm
Andrea Bosoni@theandrebosoindie64.1K粉 · 1条主张首批客户靠真诚的一对一对话,而不是批量外联模板。
@theandreboso2026-09-22♥39👁1.8K↗ 打开X
The easiest way to land your first 10 customers as a new indie founder is to find your ideal prospect on social media and have an honest conversation with them. Period. Where most founders mess up is turning it into a full outreach campaign with a copy/paste template. That's not a conversation. That's a sales pitch. We all get them. We all ignore them. And conversations stack. Each one teaches you their needs, their objections and what actually matters so the next one is more likely to convert. Mass outreach just keeps getting ignored.
首批客户靠真实对话,目标是每次更新需求、异议和表达,而非批量发送。
Thomas Sanlis 🥐@T_Zahilindie24.3K粉 · 2条继续公布Residency成员,本期产品信号较少。
@T_Zahil2026-09-22♥25👁1.7K↗ 打开X
The crew is complete: meet our resident #10: @piotrkulpinski 🥳 I've been following Piotr for so many years, I don't even remember when it was 😱 I've always admired the care he puts into all his products and how he sticks with them until they work (SEO ftw)! With him, @karakhanyanS, and @venelinkochev participating in the Residency, the directory crew will be complete 👀 I'll post a recap of all our residents tomorrow and will start to share more about what's planned for our Residency #2 soon 🔥
引用@T_ZahilMeet our resident #9: @eMarboeuf 🥳!! Emmanuel has a an incredible story He spent 10 years in SF as a cofounder CTO and scaled his company to $9M+ ARR Now he's back to France (in Nantes actually 🫶🏻), and he's building BWorlds[.co] with his cofounder … ↗
@T_Zahil2026-09-22♥32👁4.3K↗ 打开X
I've never played WoW in my entire life 😅 Should I try it?
Kyle Gawley@kylegawleyindie45.3K粉 · 2条用两条反讽提醒:AI计划不等于客户,拥有工具也不等于服务消失。
@kylegawley2026-09-22♥17👁957↗ 打开X
AGI is here. I fired my head of marketing and Claude wrote a 47-point marketing plan. I still have zero customers.
@kylegawley2026-09-21♥12👁1.4K↗ 打开X
"Claude killed SaaS" I own a brush but we still pay a maid to clean our house.
Adam Lyttle@adamlyttleappsindie58.7K粉 · 1条本窗口只有slop检测器互动和一次模型对比转发。
@adamlyttleapps2026-09-21♥16👁5.1K↗ 打开X
Disappointed that the slop enthusiast has a low slop score Clearly rigged https://t.co/CZPPQogaNY
引用@robj3d3I turned Jev into a slop detector for your own posts. Naval got 79%. I got 44%. Free, no signup, let's see your score ↓ https://t.co/Hpd1BJpjKP ↗
▸ 折叠1条(转推/噪音)
转推2026-09-23 RT @_MaxBlade: I gave GPT 6 SOL and OPUS 5.5 the same prompt. It was not even close. Opus 5.5 is THE BEST model we have ever seen. li…
Dan Rowden@drindie110.1K粉 · 1条发布Subsail Memberships,为杂志增加按时间订阅与持续权益。
@dr2026-09-22♥2👁1.5K↗ 打开X
Today I launched Subsail Memberships! Memberships are "regular" time-based subscriptions, perfect for magazines who want regular income and can offer ongoing perks like access to digital content, a community, a podcast, merch etc. The magazine can be included or not! Perks is something issue-based subscriptions struggle with; memberships fills in a nicely-sized gap in Subsail's offering. Excited to get this out and get the first few publishers creating memberships soon!

数据:Twitter/X(独立开发者+AI行业两个cohort)+Hacker News(前排/Show HN)+GitHub新星。原文照登。Twitter转推及噪音共71条折叠在作者附录内;旧推不进正文叙事,仅出现在附录并带日期。