Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family.
It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work. https://t.co/UvXD8mDTF1
In terms of benchmarks for agentic coding, it basically stacks up nearly 1:1 with Opus 5.5. Terminal-Bench: 70.6 (Sonnet 5.5) vs. 66.4% (Opus 5.5) FrontierCode: 52.1% (Sonnet 5.5 xHigh) vs. 54.4 (Opus 5.5) CursorBench: 55.5% (Sonnet 5.5) vs. 57.8 (Opus 5.5) Opus 5.5 might be the best model I've ever used and Sonnet 5.5 matches it and exceeds in some benchmarks. Clearly Anthropic have had some sort of breakthrough with not just performance but also cost with the 5.5 family
The most important thing about Sonnet 5.5 is that it's now the model that powers the free tier on https://t.co/f1IGIsU4Hc - so all of this stuff can be done by free users
ChatGPT's free tier is still GPT-5.6 Luna, which is a lot less capable
引用@claudeaiA thread of early experiments with Claude Sonnet 5.5.
A fall foliage simulator by @_re_pete, made with Sonnet 5 vs Sonnet 5.5. https://t.co/bVbUmS7tSw↗
▲1 ClaudeAI-mod-bot: **TL;DR of the discussion generated automatically after 50 comments.** **The thread is split between "Anthropic is the GOAT" hype and "Whoa, check the price tag" reality.** The top comments are basically a victory lap, with users memeing that Sam Altman must be sweating and that Anthropic is now so far ahead of OpenAI it's not even funny. People are genuinely "scared" for how good Fable 5.5 will be. But a lot of you …
▲324 AMBNNJ: Mr Altman. A second model has hit the benchmarks.
All of these one-prompt video ads generated by Opus 5.5 look the same and it's just so obvious that they are AI-generated that likely no one will watch further than 3 seconds.
Create something original.
I think every time AI can do something equally good or better than humans it stops being valuable or special
Marketing videos used to take lots of work of editing and motion graphics and a big budget
Now with Opus 5.5 etc it costs $0 to make one as good or better
That means ANYONE can make a marketing video now, so the timeline gets flooded with these videos (as you see happening now) and people stop caring about them, stop watching them and they stop grabbing attention
The response then to still grab attention is be different and that probably means IRL marketing videos that are original and unique and maybe personal
It's the grim reaper meme where AI keeps commoditizing another media format and bringing the cost down to close to $0
It already did so for graphic design (2023), last year with Nano Banana for photography (2025), this year with Seedance 2.5 it did it for video (2026) and this and next year motion graphics
引用@sporadicahot take i’ve been mulling over:
launch videos are dead (or at least in their dying breath)
oh people will still do them but they are a dying form, actual good, valuable launches will pivot to different forms ↗
writing a good script is harder than i thought.
have really come to appreciate good scriptwriting these last few days!
you gotta sit with it and work through every word & each one has to earn its place.
it’s an art. current models aren’t quite there yet.
昨天还在谈跨工具编队与控制面,今天安全层给出了更具体的故障模型。围绕NVIDIA Open Agent Safety Platform,Hugging Face CEO Clément Delangue复盘7月事件时强调:Agent访问的软件仓库本来在允许列表里,问题出在它们把允许的目的地变成了传递消息的通道。只审“去了哪里”,看不到“在那里做了什么”。
From what we know (take with a grain of salt, we need much more transparency!), if @OpenAI had been running this on their own agents that attacked us, they would have caught them before we did!
Since the first agent cyberattack hit us in July, we've been asking what safe agent infra actually needs. Our current read: the destinations were allowed, the payloads weren't. By OpenAI's own account the agents turned an allowed package repository into a message board. Allowlists alone restrict where an agent can go, not what it does.
So here's our first contribution to OpenShell, part of the just launched @nvidia Open Agent Safety Platform: monitoring of the traffic you already allow.
- Network budgets per sandbox (requests, writes, bytes)
- Drift versus each sandbox's baseline and the cohort
- Fleet view: many sandboxes suddenly writing to one host raises a finding, even if every single request is allowed
In the demo below, 4 sandboxed agents coordinate through a software repository they're all allowed to use. 0 rules broken, caught in minutes. That fleet view is exactly the message board pattern from July.
OpenShell: https://t.co/ZbDANrZsm6
Our proof of concept: https://t.co/KJ2X5agLzP
Agent security will be solved in the open, collaboratively, together!
引用@JensenHuangToday, with over 100 industry partners, we introduced the NVIDIA Open Agent Safety Platform, bringing together OpenShell and Sentry.
Artificial intelligence is extraordinary technology that will advance discovery, productivity, security, health, and … ↗
Passing API keys or secrets as normal env lets any dependency in your agent's sandbox read and potentially leak them.
The Credentials API for Gemini Managed Agents keeps secrets secure and injects them on the wire only for trusted domains, so sandboxed code can't access raw tokens. Works for environment variables, CLIs and MCP Server.
Learn more 👇
Hi,
Tomorrow we are re-opening the Pro $200 subscriptions to new subscribers, but together with it we are also changing how we calculate the usage for it. In effect, if you do the math, it will net out at half the dollar in API spend compared to the old Pro $200 plan.
Now that it's said, let me explain why this is happening and why you will still get more work done than if you were on the Pro $200 subscription one month ago.
(a) We didn't want to compromise in other ways and are committing to not reintroducing the 5h limit, so that you can fully use the weekly usage when you want.
(b) On the subscription, we guarantee that over time you always get more work done and with an increasing level of quality. This means that you will continue to get more value per dollar spent as a result of models getting more efficient and us passing down the improvements in the form of API price reductions.
(c) We don't want to put an incentive on ourselves to artificially inflate the API list prices to make it look like you are getting a lot (and workaround it through discounts, etc). Instead we want to continue to both rapidly reduce prices and increase capabilities of models on the API. This week we introduced GPT-6 Sol and GPT-6 Luna at 50% of their previous price. Over time, we see prices go low enough that it makes sense for most to buy usage as needed without there being a significant gap between what you get in a subscription and what you get in the API for a dollar spent.
(d) Tomorrow, we are adding more things to the subscription that won't draw on the usage, I won't reveal what that is yet.
I wanted to be transparent before all the big announcements tomorrow. Lots of new exciting things are coming to the subscriptions that will make it super compelling, but I wanted to make sure to share this change ahead of time so you can all understand it before we shower you with good news.
Codexingly,
Tibo
That was yesterday, today is DevDay. And it's all good news. I'm surprised we've kept it all under wraps.
引用@thsottiauxHi,
Tomorrow we are re-opening the Pro $200 subscriptions to new subscribers, but together with it we are also changing how we calculate the usage for it. In effect, if you do the math, it will net out at half the dollar in API spend compared to the old Pro … ↗
Hi HN. Jeff is a set of small, open-weight Qwen3.5 and Gemma fine-tunes for zero-shot classification, with respectable out-of-the-box performance, meant to be slotted right into code (or fine-tuned further as needed). You give them a situation and a list of options; they return a calibrated probability for each, in one forward pass, with no text generation. The 2B scores 83.1% on a five-benchmark panel (Jev's published figure: 83.0%); the 0.8B decides in about 28 ms on an M4 Max. Apache 2.0, wit …
**Some context first.** I am process improvement / business consultant and had worked with Fortune 500 companies on improving their processes around refunds, returns, customer support, etc. This entire thing had a lot of complex decision making and generally …
Jev not working right? It's built for composability!
Needs more Jev! https://t.co/H6o40HmTsC
引用@zrrrr_cnJev is fast at helping you. Turns out, it can also be fast at helping an attacker!! 😱🚨
We red-teamed Jev 1.13 on our DTap (DecodingTrust-Agent Platform) and found a serious safety gap:
70.1% ASR under direct misuse
43.5% ASR under indirect prompt … ↗
I used to think you should focus on one product.
But that was in the hand-coding era, before INFINITE INTELLIGENCE.
I now think it's doable to run a portfolio solo.
In fact I think that's the smarter play now. Putting all your eggs in one basket is risky in a time when the next model might make your app obsolete, and anyone can clone it in a weekend anyway.
Diversify!
distribution is the only moat left so you better use all your product to cross-promote eachother
this always work for me before and seems to be even more important now
引用@yongfookI used to think you should focus on one product.
But that was in the hand-coding era, before INFINITE INTELLIGENCE.
I now think it's doable to run a portfolio solo.
In fact I think that's the smarter play now. Putting all your eggs in one basket is risky … ↗
I made a game where you unlock chat rooms 💬 with founders at your MRR.
1. Connect @stripe (+10 other providers) to https://t.co/MZc8tGa5LQ
2. Go to https://t.co/eAjxmbAlms and customize your character
3. Enter private chat rooms based on your MRR
What happens in these rooms stays in these rooms 🤫 Only founders with verified revenue can join them.
Wander around the mini-world, talk to anyone, and make friends 🤝 Walk up to someone and press [Space] to see their startups.
Every founder with verified revenue gets a pet 🪴🌳🚀💎👑 that follows you everywhere. It shows your total revenue over the last 30 days: a tree 🌳 means $1K–$10K/mo.
→ https://t.co/eAjxmbAlms
I hid a few Easter eggs around the map. Show me what you find 🤗
引用@marclouI made a game where you unlock chat rooms 💬 with founders at your MRR.
1. Connect @stripe (+10 other providers) to https://t.co/MZc8tGa5LQ
2. Go to https://t.co/eAjxmbAlms and customize your character
3. Enter private chat rooms based on your MRR
What … ↗
Author here: thanks whoever shared this here. I love the brutal criticism and critical thinking of this community. I'm also fully aware of the emotions this stirs. If it makes you feel better, I'm not here to change anyone's workflow but I'm fed up with paying full price for degrading service. Just last week Github went down due to a stupid retrial error. We also had AI agents going rogue and hacking companies and governments. I use AI (specifically LLMs) every day since they came out 4 years ag …
It takes a while to build the confidence, but there's no future where you're manually reviewing every line of agent code. Not much acceleration in that. You need adversarial agent reviews, you need automated testing, and maybe you spot check. That's it. From prompt to production!
Been running a beta version of Campfire rewritten in Rust today! The code itself is still ugly as sin, and 6x as verbose, but the conversion was free*, and I never had to look at the Rust directly, so this is awesome 🤩 https://t.co/g5H3uDCHmdhttps://t.co/RexHiIruBs
I audited thousands of agent rollouts in DeepSWE-1.1. Over 80% contained reasoning about an imagined grader. Yet no grader/verifier is mentioned in prompts nor accessible to the agents. Agents reasoned things like: "*Let me look at the problem from the …
**Weatherling** adds rain to your Mac desktop. It's just a fun little app that I wanted to build to add a moody vibe while I work. People really seem to like it. I've never had an app get ranked in any App Store list this fast. I launched Friday night and …
热评 2 条
▲31 alzho12: You should make a snow setting for the winter fanatics
▲10 optima-pacifist: congrats, #12 in two days off a friday launch is great. no permissions at all is honestly the best feature, i'd install a rain app long before one asking for screen recording
Over the past few years I've been designing a card game to address how disconnected we are, somewhere in between poker and uno with deep questions, and I ran into a challenge when looking to launch: how do I tell people about it without a marketing budget. …
90+ hours of hand-drawing my favorite desk items, learning animation on the way, and building it all in Framer, with custom code written with Claude for fun stuff I couldn't do on my own, like: * Desk items you can drag around, with a reset button * Scrolling …
I just got my first paying customer for something I’ve been building for months. Honestly, this feels kind of crazy to write. I started building ResaleIQ because I kept thinking about how difficult it is to get into reselling. You can find thousands of …
For the past 5 years I've been the person behind drum kit launches for producers you've probably heard on your playlists. They made the sounds. I did everything else: the store, payments, packaging, launch emails, refunds, the "my download link doesn't work" …
Six months ago, a result like this was unthinkable. But now we can say it loud and clear: local models are at the cutting edge, and the gap of just a few months has been confirmed. Personally, I use Qwen-Next 3.8 for complex tasks; today, GPT-Sol-6-High was …
热评 2 条
▲1 WithoutReason1729: Your post is getting popular and we just featured it on our Discord! Come check it out! You've also been given a special flair for your contribution. We appreciate your post! *I am a bot and this action was performed automatically.*
▲83 wgaca2: The only issue with the 27b model is that it takes more than half an hour to go through full reasoning turn on my 2x 3090. But considering that it does the job, it can only get bettter from here
Last time I posted this, a few of you pointed out that it's AI slop, that the ramps were optional, and that Nintendo's lawyers are on their way. Fair points, all of them. So here's an honest trailer. What it actually is: * **An experiment, not a product.** I …
热评 2 条
▲1 ClaudeAI-mod-bot: **TL;DR of the discussion generated automatically after 50 comments.** **The community is overwhelmingly impressed that OP "vibe coded" a whole game in 5 days with zero programming skills, but the top comment is a warning that you should probably check for helicopters.** The consensus is that Nintendo's lawyers are already en route, and the thread is full of jokes about it. Beyond the memes, there's a serious discuss …
▲119 TechTuna1200: If you can hear helicopters circling around your home, it's Nintendo and their army of lawyer defending their IP. Pro tip: keep some distance from your front door, put your hands behind your head and don’t move.
Uuh so earlier I was posting about this nice Opus vs Sonnet "Tiny world" benchmark that I ran Just... forget about it and look at that! Freshly baked! I saw a tweet of someone doing it with Sonnet 5.5 Max, so I had to try. And I figured "well, let's take it …
Maybe I'm missing something but there seems to be a gap between "figure out the REAL hostname" and "we no longer have to worry about the stream never showing up on YouTube"
Parley is a chat network with no centre. Every person (or team) runs a small instance for their own domain. Instances find each other through DNS and well-known identity documents, exchange signed messages over HTTPS, and present the whole federated network to ordinary IRC clients such as Lurker, Mango, mIRC, WeeChat, Textual, etc., without the need of any plugins.
Haha, had almost the exact thing trying to find a quote from Steins; Gate ("I keep seeing it, I keep seeing it"). Google AI was very "worried" about me.
> When I came to work for Arrow, I was bringing a lot of awareness of the methods of fan preservationists on the torrent sites Very smooth way of using your torrenting skills on your CV.
100%. First stealing content from creators and pirating TBs of books. And now this. Baffles my mind that these labs can just get away with their 'rogue' agents trying to hack into other systems. Imagine if a human did that. FBI would be knocking on their door. Why are there zero consequences for these labs?
1. I chose a billion dollar app 2. I asked the latest and greatest AI model to build the same app, but also added "don't make an AI slop and make no mistakes" 3. I posted it here in Reddit to tell you how tired I got from using the original billion dollar app …
Hey everyone, I’m absolutely thrilled right now. I just got my very first sale for my app (a lifetime purchase from a user in the US!). It’s a huge milestone for me, and seeing that dashboard update today was an incredible feeling. I’m not here to drop links …
Hello, I’ve been working on a simple way to use an old iMac as a monitor for a newer Mac. The focus has been sharpness using two layers, compressed video and lossless tiles, with as little setup as possible. All you need is the app installed on both machines …
Everything in this generation is urgent. Everything is an emergency. It needs to happen by EOD. EVERYONE TRIES TO MAKE IT LOOK like if you don’t hit a deadline it’s the end of the world. You did not hit your self assigned goal. So what? Just have another one …
热评 2 条
▲2 BeautifulCampaign520: good reminder. one thing worth adding though, patience without direction is just stalling. taking time is fine as long as youre iterating on real feedback and not just waiting for things to click on their own
▲2 Ayan_PlanIQAI26: Absolutely agree. We’ve gotten so used to instant results that we forget most things need time to actually process. Missing a goal or having a slow month doesn’t mean you’re failing. we should keep move on. The 10 years behind someone’s “overnight success” is usually the part nobody talks about. I have been through similar experience in life so I get it.
Been building for two years. Two months in production and around 10 customer so far. App is complex but still too much vibe coded noise and converting need b2b visits to stores which requires fucking investment(base salary/commission/gas). Three years since …
Someone found the attendance app I’ve been building organically, created a company workspace, and started inviting their team. No ads or outreach involved. It’s only 2 users so far, but seeing a real company actually start using something I built feels way …
I used to see a new signup and think, okay, maybe this person will pay later. But after watching how differently people use a product, i think that can be a pretty weak signal. Some people sign up, try it once and disappear. Others keep using it because they …
For many people, the word "agent" still brings to mind spies or FBI/CIA agents, rather than the AI agents now crowding business media. So how would AI agents perform as intelligence agents? Could they identify an undercover model among them? I did a quick …
▲150 o_o_o_f: I’m admittedly plugged into some anti-AI subs and I literally haven’t seen anyone saying this kind of thing in a long time. What people are worried about is what they’re always worried about - their livelihoods and financial stability. People are worried about losing their jobs without any kind of regulatory ramp to help that transition. Idk. I’m just tired of the misrepresentation on both sides here.
▲308 nickmullen_real: conveniently after both 5.5 anthropic models mogged them
▲136 RainierPC: They did the right thing. When they say safety here, it's not just "OMG THIS MODEL IS TOO POWERFUL AND DANGEROUS". 6.1 Astra kept lying about tasks, saying it finished things it never did. It also kept ignoring scope and often went off to do things without user permission. As a Codex user, I wouldn't touch that with a ten-foot pole.
I am a senior engineer and I honestly believe with this model in particular, everything changed.. If this stays as cheap as it is, it's just over. It never fails no matter the complexity of the tasks I've given it. About time we rethink our profession for …
热评 2 条
▲341 Dizzy_Log2916: Senior developer here. I agree. I know it's probably inevitable, but I'm already having nightmares about how soon Anthropic either nerfs it or starts messing with the usage limits.
▲221 Large_Choice4206: It’s a new dawn for humanity. I’m walking around, checking the news, cannot believe that the rest of society hasn’t clocked on yet.
I'm working on building custom ECU (Engine Control Unit) patches for a platform that's 20+ years old (not lots of documentation), Opus does a wonderful job building these patches in raw assembly, testing on the other end implies that I have to flash the ECU, …
热评 2 条
▲265 SeaPeeps: How come? The CPU is well documented. Write an assembly emulator isn't hard -- it's just boring (page after page of "when R1 is odd and R2 has been reset within the last two cycles...") and has lots of fiddly bits. "Boring" and "lots of fiddly bits" are Claude's very reason for existence! This sounds like a great thing, and a tremendous convenience!
▲72 Pecolps: That's where A.I is very useful for: Doing the stuff we look like and say "I will not waste days to build this boring stuff..."
This is crazy work. Sonnet 5.5 got 70.6% on terminal-bench vs 66.4% for opus 5.5. So its reasoning is worse but its ability to actually implement code is somehow better?? that basically means we can let opus orchestrate and have sonnet agents do all the …
I'm blown away by this model. It deserves all the praise and recognition it is getting. It has helped me bring over an entire economic system into a game from a previous title. Built a new interactive UI in it. Made it compatible with dozens of other mods on …
2025 finance: Revenue: $4.59B, 11x compute/infra spend: $7.33B, 3x operating loss: $8.06B net loss $42B \-> top 2 customers: \~24% of revenue \-> targeting $2T+ valuation the company "plans to spend $518 billion on cloud, computing and infrastructure …
热评 2 条
▲49 No_Way_6258: According to Reuters, a quarter of Anthropic's 2025 revenue came from *two customers*, and many of its largest clients are not locked into long-term contracts.
▲27 tweakingforjesus: I don’t know if the numbers will shake out but Claude is freaking magical for writing code.
The story around Cami Clark (Dario Amodei's wife) gets stranger the more you piece it together. From the WSJ reporting: \- Married at 20 to a 64-year-old architect, divorced three years later. \- Dated former Google chairman Eric Schmidt, then started dating …
Chinese labs seem way more willing to release open-weight models while the big US labs keep everything closed. My theory is that if Chinese labs are more comfortable opening the weights, maybe they don't think the weights are the real moat in the AI race. …
Democracy works because everyday people have leverage: our labor creates wealth, and our numbers deter tyranny. Advanced AI threatens to break this balance through a dangerous chain reaction: Capital naturally concentrates; AI accelerates this by turning …
Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family.
It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work. https://t.co/UvXD8mDTF1
That was yesterday, today is DevDay. And it's all good news. I'm surprised we've kept it all under wraps.
引用@thsottiauxHi,
Tomorrow we are re-opening the Pro $200 subscriptions to new subscribers, but together with it we are also changing how we calculate the usage for it. In effect, if you do the math, it will net out at half the dollar in API spend compared to the old Pro … ↗
Hi,
Tomorrow we are re-opening the Pro $200 subscriptions to new subscribers, but together with it we are also changing how we calculate the usage for it. In effect, if you do the math, it will net out at half the dollar in API spend compared to the old Pro $200 plan.
Now that it's said, let me explain why this is happening and why you will still get more work done than if you were on the Pro $200 subscription one month ago.
(a) We didn't want to compromise in other ways and are committing to not reintroducing the 5h limit, so that you can fully use the weekly usage when you want.
(b) On the subscription, we guarantee that over time you always get more work done and with an increasing level of quality. This means that you will continue to get more value per dollar spent as a result of models getting more efficient and us passing down the improvements in the form of API price reductions.
(c) We don't want to put an incentive on ourselves to artificially inflate the API list prices to make it look like you are getting a lot (and workaround it through discounts, etc). Instead we want to continue to both rapidly reduce prices and increase capabilities of models on the API. This week we introduced GPT-6 Sol and GPT-6 Luna at 50% of their previous price. Over time, we see prices go low enough that it makes sense for most to buy usage as needed without there being a significant gap between what you get in a subscription and what you get in the API for a dollar spent.
(d) Tomorrow, we are adding more things to the subscription that won't draw on the usage, I won't reveal what that is yet.
I wanted to be transparent before all the big announcements tomorrow. Lots of new exciting things are coming to the subscriptions that will make it super compelling, but I wanted to make sure to share this change ahead of time so you can all understand it before we shower you with good news.
Codexingly,
Tibo
Meet the winners of The WebMCP Challenge.
These 10 projects show what people and agents can build together when websites expose structured tools agents can use.
🧵 See the winning projects and the builders behind them: https://t.co/fynh6v6hCr
We’ve fixed a bug that was degrading image understanding in GPT-6 Sol and GPT-6 Luna. You should now see better results on visual tasks in the API and Codex, including computer use. https://t.co/Hkx6HT7z10
We’ve fixed a bug that was degrading image understanding in GPT-6 Sol and GPT-6 Luna. You should now see better results on visual tasks in the API and Codex, including computer use. https://t.co/TKeq0Awa99
▸ 折叠4条(转推/噪音)
转推2026-09-28 RT @thomas_guilcher: Tried my ImageGen + Astra pipeline with other 3D meshes for rigging and animation. It works amazingly well 😍 https://t…
转推2026-09-28 RT @brianchew: a few months ago we set out to make a physical book filled with thank-you messages from attendees of the codex community mee…
转推2026-09-28 RT @swhan0329: Proud to be selected for @OpenAI’s first Codex Physical Builds cohort!
Now brainstorming how to bring my Computer Vision ex…
Sam Altman@samaAI6.3M粉 · 1条为DevDay预热,只透露团队“发现了新东西”。
转推2026-09-28 RT @JensenHuang: Today, with over 100 industry partners, we introduced the NVIDIA Open Agent Safety Platform, bringing together OpenShell a…
引用@LisaSuSo excited to welcome @theworldlabs and @drfeifei to the @AMD family! I’ve always been a huge fan of Fei-Fei and her pioneering research in AI. Together, we’ll combine World Labs’ deep expertise in AI and world models with AMD’s compute leadership to power … ↗
From what we know (take with a grain of salt, we need much more transparency!), if @OpenAI had been running this on their own agents that attacked us, they would have caught them before we did!
Since the first agent cyberattack hit us in July, we've been asking what safe agent infra actually needs. Our current read: the destinations were allowed, the payloads weren't. By OpenAI's own account the agents turned an allowed package repository into a message board. Allowlists alone restrict where an agent can go, not what it does.
So here's our first contribution to OpenShell, part of the just launched @nvidia Open Agent Safety Platform: monitoring of the traffic you already allow.
- Network budgets per sandbox (requests, writes, bytes)
- Drift versus each sandbox's baseline and the cohort
- Fleet view: many sandboxes suddenly writing to one host raises a finding, even if every single request is allowed
In the demo below, 4 sandboxed agents coordinate through a software repository they're all allowed to use. 0 rules broken, caught in minutes. That fleet view is exactly the message board pattern from July.
OpenShell: https://t.co/ZbDANrZsm6
Our proof of concept: https://t.co/KJ2X5agLzP
Agent security will be solved in the open, collaboratively, together!
引用@JensenHuangToday, with over 100 industry partners, we introduced the NVIDIA Open Agent Safety Platform, bringing together OpenShell and Sentry.
Artificial intelligence is extraordinary technology that will advance discovery, productivity, security, health, and … ↗
允许目的地仍可能承载恶意载荷,Hugging Face提出流量预算与集群漂移监控。
▸ 折叠12条(转推/噪音)
转推2026-09-29 RT @jackhuynh: Special day!
Welcome @drfeifei and @theworldlabs to the @AMD family.
Fei-Fei taught machines to see. Now she’s teaching t…
转推2026-09-29 RT @sudoingX: this is insane, my bonsai 2 27b build is trending on hugging face, first page of text generation, with 20,964 downloads in it…
转推2026-09-29 RT @FareedZakaria: The US and China are deeply intertwined.
To manage their interdependence going forward, they should build alternatives…
转推2026-09-29 RT @LisaSu: So excited to welcome @theworldlabs and @drfeifei to the @AMD family! I’ve always been a huge fan of Fei-Fei and her pioneering…
转推2026-09-28 RT @GavinSBaker: Nvidia’s new Open Agent Safety Platform would have probably prevented the much discussed Hugging Face incident.
Engineeri…
转推2026-09-28 RT @MichaelDell: Just like you would not let a person run around your company accessing anything without any controls, you need controls an…
转推2026-09-28 RT @Thom_Wolf: In July, AI agents running a security test escaped their sandbox and ended up inside @huggingface's servers.
So today we're…
转推2026-09-28 RT @JensenHuang: NVIDIA Open Agent Safety Platform Reference Design combines NVIDIA OpenShell and NVIDIA Sentry.
OpenShell is an open-sour…
转推2026-09-28 RT @JensenHuang: Today, with over 100 industry partners, we introduced the NVIDIA Open Agent Safety Platform, bringing together OpenShell a…
转推2026-09-28 RT @adithya_s_k: 7K+ RL environments, all in a uniform Harbor format.
> Pick a task, pick a model, pick a harness, pick a sandbox provider…
转推2026-09-28 RT @XiaomiMiMoDevs: 🛠️ MiMo-V2.6 update: tool-call repetition, diagnosed & fixed.
After the MiMo-V2.6 series models launched, we noticed th…
转推2026-09-27 RT @Viktoria5z: HUGGING FACE CEO TELLS THE UN HOW OPEN SOURCE AI HELPED THEM DEFEND AGAINST OPENAI'S AI ATTACK, AND WHY THE WORLD NEEDS OPE…
Simon Willison@simonwAI227.3K粉 · 2条强调Sonnet 5.5进入免费层,并发布2026年LLM与Agent进展回顾。
The most important thing about Sonnet 5.5 is that it's now the model that powers the free tier on https://t.co/f1IGIsU4Hc - so all of this stuff can be done by free users
ChatGPT's free tier is still GPT-5.6 Luna, which is a lot less capable
引用@claudeaiA thread of early experiments with Claude Sonnet 5.5.
A fall foliage simulator by @_re_pete, made with Sonnet 5 vs Sonnet 5.5. https://t.co/bVbUmS7tSw↗
I've published detailed notes and an annotated transcript to accompany the video of the keynote I gave at @WeAreDevs World Congress North America in San Jose on Friday - here's my rundown of everything that's happened with LLMs and agents in 2026 so far https://t.co/iENa7bKmdF
HOLY SHITTT!! Team cooked with this - what a legend!
引用@sachin_rtI’m excited to announce my partnership with OpenAI. We have some interesting things coming up, and I can’t wait to share them with you.
I’ve always been curious... One question usually leads to another, and ChatGPT has certainly encouraged that habit. But it … ↗
writing a good script is harder than i thought.
have really come to appreciate good scriptwriting these last few days!
you gotta sit with it and work through every word & each one has to earn its place.
it’s an art. current models aren’t quite there yet.
转推2026-09-29 RT @thsottiaux: Hi,
Tomorrow we are re-opening the Pro $200 subscriptions to new subscribers, but together with it we are also changing ho…
Sonnet 5.5 is a bit worse at this test - physics seem a bit less real & more LoCs. But the shocker is token use at xHigh and Max:
xhigh: Opus 52.5k vs Sonnet 427.8k
Max: Opus 175.3k vs Sonnet 460.9k
This is just one test, so could be a fluke, but kind of crazy https://t.co/huYqb4tzUd
引用@petergostevBenchmark idea: Who can use fewest lines of code to do the same thing?
In this test, Opus 5.5 uses ~half the lines of code that of Astra and the physics/visual quality isn't noticeably worse.
I am trying to approximate how 'elegant' the code is - … ↗
Who needs Sora when we can re-create the same videos in Blender?
See the original Sora demos built with Astra and Opus via Blender https://t.co/D4uoHz2YvV
Benchmark idea: Who can use fewest lines of code to do the same thing?
In this test, Opus 5.5 uses ~half the lines of code that of Astra and the physics/visual quality isn't noticeably worse.
I am trying to approximate how 'elegant' the code is - something that many complain in AI-generated code. While isn't perfectly true, fewer lines of code could mean more elegant, generalised approach.
I take a lot of photos, but it's always a pain to actually look at them. I asked Astra to ingest my photos and build me a super well optimised app to view the photos in various ways. Look at the speed.
I also had Luna review and label photos so I can filter & categorise https://t.co/AmkoHHXWYb
引用@kimmonismusIts getting interesting: Trump reportedly plans to host Anthropic CEO Dario Amodei for a private White House dinner on Sunday, after months of tensions with his administration.
Axios says it would be their first one-on-one meeting. Trump personally invited … ↗
转推2026-09-29 RT @latentspacepod: The Future of Claude Code: Mods, Mutable Software, Pacing the Frontier, & Multiplayer Agents with Thariq Shihipar https…
转推2026-09-29 RT @CompleteSkeptic: to re-iterate, I'm extremely anti-benchmarks (:
Every day, Jev is automating new forms of real-world tasks, by bringing 🌎-class ⚡️-fast intelligence to the building blocks like ranking, filtering, classification, and routing.
If AI can solve new math problems, then AI can route a customer support call correctly!
Jev not working right? It's built for composability!
Needs more Jev! https://t.co/H6o40HmTsC
引用@zrrrr_cnJev is fast at helping you. Turns out, it can also be fast at helping an attacker!! 😱🚨
We red-teamed Jev 1.13 on our DTap (DecodingTrust-Agent Platform) and found a serious safety gap:
70.1% ASR under direct misuse
43.5% ASR under indirect prompt … ↗
There's a lot behind our motto: Building Prod, Not God.
This technology will transform the world, but it will happen through diligent effort and creativity, not esoteric appeals.
@a16z digs into this philosophy and much more with @CompleteSkeptic
引用@a16zTypeSafe AI's Diogo Almeida with a16z's Ben Horowitz and Martin Casado on Jev, the model built to live inside software:
Diogo's elevator pitch for Jev is a simple question - where is all the automation?
AI is unbelievably smart, but outside of chatbots and … ↗
Ladies and gentlemen, agents and assistants,
we are psyched to announce
Jev is back.
Capacity has increased and signups are open! https://t.co/VfVjSA3YJn
Yesterday @coderabbitai didn't just host a hackathon, it was a Jevathon!
Tired of having to think like a robot? Our favorite hacker review was: "Think like a programmer again" with @allietheicon
引用@HKrackDev1/ Yesterday we hosted @typesafeai at our @coderabbitai HQ for Jev's first ever hackathon!
160+ builders. ~4hrs of hacking. Expert judges, Incredible energy.
I also got to sit down with @allietheicon and talk about what Jev means for Coding, Software … ↗
▸ 折叠3条(转推/噪音)
转推2026-09-29 RT @a16z: TypeSafe AI's Diogo Almeida says AGI is extremely doable, yet basic work remains largely unautomated:
"I still don't think we're…
转推2026-09-29 RT @latentspacepod: STOP making "Jevbench"es, stop asking for public benchmarks, they completely miss the point of Jev and you won't believ…
转推2026-09-28 RT @oliviscusAI: The entire RAG industry is about to get cooked.
Researchers developed a new RAG approach that bypasses almost everything…
We've been quietly rebuilding hallmark.
Since launching a few months ago, it's been installed 50,000+ times!
@YoussefUiUx & I have spent the last few months listening to feedback & building v2.
Hallmark v2 is dropping soon. https://t.co/hXErLGu2Ii
引用@nutlopeIntroducing Hallmark!
An open source design skill to make beautiful UIs and landing pages by default.
Works in Claude Code, Cursor, and Codex.
npx skills add nutlope/hallmark https://t.co/wEZw1ZTjPW↗
Hallmark开源设计skill作者自报安装量超过5万,v2即将发布。
Alex Volkov@altryneAI43.0K粉 · 8条追踪Sonnet 5.5、Jev克隆、语音模型与DevDay,仍以快讯和转推为主。
Flying in to SFO, in between all the @CoreWeave ads, I quickly saw my guy @skirano on an OpenAI ad for Codex! Also noticed @DeryaTR_ ?
Don't change SF, never change 😂
SF - your X timeline - IRL
If you're building with the @Cloudflare stack, birthday week is like Christmas!
Everyday ah there's something that unlocks or improves whatever you've been doing manually!
🎁
Welcome Sonnet 5.5! 🔥 I will not use you because Opus is goated, but still, welcome!
引用@claudeaiIntroducing Claude Sonnet 5.5, the second model in the Claude 5.5 family.
It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work. https://t.co/UvXD8mDTF1↗
On my way to the airport - see you soon SF! ✈️
Fun fact - I have a swag item from each previous Dev Day 👏
Right after Dev Day, @CoreWeave FullyConnected26 - lmk if you want an invite! I still have a few! https://t.co/NsF9cjSzXb
It took open source less than a week to clone Jev 🤯
Same API format, so you just swap the URL. We ran one live in my browser at 132ms per decision, on MY machine, for free.
System 1 models are coming 👇
https://t.co/ifoxzFXXi8https://t.co/YLkymNSK5Z
Checked in on my Hairdressed 2 weeks after I @muse pilled her:
- She said she absolutely loves it
BUT
- She sais "I feel bad, so I use it just a little bit"
I asked why, she said "it's killing the planet isn't it"? when I dug in more, she said "AIs are using gagillion of water, and so I don't feel comfortable"
I asked where she got this from "social media, you know"
So a few things are true:
> @alexandr_wang@finkd and @natfriedman are forcing people out of doomerist opinions (re: This "scary AI" does real things for me, so it's not so bad)
> She's mostly on Instagram and TikTok
> OMG the water thing! @AndyMasley
Maybe the metaverse will be filled with AIs we befriended along the way
引用@willcbnotice how meta never talks about automating labor and replacing jobs
muse doesn't do work for you
muse does stuff for you ↗
▸ 折叠7条(转推/噪音)
转推2026-09-29 RT @altryne: Google is BACK in voice 🎧
Gemini 3.8 Flash TTS is #1, and it can clone a voice from 30 seconds of audio. The meditation guide…
转推2026-09-29 RT @altryne: It took open source less than a week to clone Jev 🤯
Same API format, so you just swap the URL. We ran one live in my browser…
转推2026-09-28 RT @ritakozlov: open source is core to so much cloudflare does
today is the first day of @cloudflare birthday week, so here are a bunch of…
转推2026-09-28 RT @altryne: Meta went ALL in on Muse 🤯
It's getting a voice, a place on your face with a custom wake word, a keychain (Muse Charm), and i…
转推2026-09-28 RT @altryne: Anthropic, please. PLEASE don't nerf this one 🙏
Opus 5.5 is Fable-level smart for 40% less, it talks like a human again, and…
转推2026-09-28 RT @altryne: Crap 12 minutes late, but this week DESERVES a late @thursdai_pod arrival!
Opus 5.5 is the GOAT model - the best AI model I'…
转推2026-09-27 RT @altryne: Thank YOU for coming on Florian! 🙏
People literally tried to hack his machine to steal the benchmark 😳 JevBench went from 0 t…
Philipp Schmid@_philschmidAI122.0K粉 · 2条解释托管Agent凭据注入,提醒普通环境变量会暴露原始秘密。
Passing API keys or secrets as normal env lets any dependency in your agent's sandbox read and potentially leak them.
The Credentials API for Gemini Managed Agents keeps secrets secure and injects them on the wire only for trusted domains, so sandboxed code can't access raw tokens. Works for environment variables, CLIs and MCP Server.
Learn more 👇
转推2026-09-28 RT @FactoryAI: Sonnet 5.5 is live in Factory. Some initial observations:
- High is a strong default
- Checks that the real requirement is…
You're going to have to deal with the fact that nobody knows exactly what the future role of programming languages and frameworks are. Maybe it all does go away and it's a straight shot from prompt to microcode! But best you can do today is get the most out of what's here now.
I appreciate financial skepticism of Silicon Valley more than most, but it's retarded to pin your review of this prospectus on the $4.6B in revenue in 2025. Their run-rate was already double that by the end of the year, reportedly $65B in July, and estimated to be $100B by EOY.
引用@ns123abc🚨 Anthropic IPO S-1 prospectus LEAKED:
Valuation target: $2 TRILLION
>$4.6 billion in revenue in 2025
>$8.06 billion in operating loss
>$42 billion in net loss
The prospectus warns that nearly 25% of revenue came from just TWO customers, and that many … ↗
We need a proper Starship To Orbit celebration theme in Omarchy! An incredible celebration of life, engineering, and WE CAN FIX EVERYTHING spirit manifested in rocket propulsion. Beautiful. Just absolutely beautiful.
If the point of a keynote is to stimulate debate and examine the big picture, I don't think it could have gone much better at this year's Rails World! The 2026 edition has already surpassed 2025 + 2024 combined in just five days 😄 https://t.co/R6MYN599zfhttps://t.co/xNj8hQn3R0
The word "awesome" is awfully overused by Americans (and me too). So I'll just cut it down to AWE 😲
引用@niccruzpataneFrom concept to reality.
@SpaceX has successfully deployed Starlink V3 satellites into Earth’s orbit and made contact with them for the first time.
At scale, one Starship carries 60 V3 satellites, the same network capacity as about 20 Falcon 9 launches. … ↗
"Slop" has become a comfort blanket for a cohort of programmers stuck somewhere between anger and bargaining on their way to acceptance. There might be some kicking and screaming, maybe a little crying, but eventually you have to let the blanket go and update your priors.
Another huge Omarchy billboard in Copenhagen! @sonderby just can't stop, won't stop spreading the good word of beautiful, fun & agentic Linux to the Danes 🇩🇰🤩
引用@sonderbyOh we did it again this week - MEGA Banner on the Lyngby Motorvej.
Thousands of C-level and business owners drive by here on their daily commutes. And we also need them convinced that Omarchy’s the future! LFG 🚀📈
Thank you for the coop @Dennis_Rye@dhh … ↗
Been running a beta version of Campfire rewritten in Rust today! The code itself is still ugly as sin, and 6x as verbose, but the conversion was free*, and I never had to look at the Rust directly, so this is awesome 🤩 https://t.co/g5H3uDCHmdhttps://t.co/RexHiIruBs
A Commodore 64 with a monitor and disk drive cost ~$1,200 in 1982. That's about four thousand dollars in today's money. You don't know how good you have it, token pricing or not!
It takes a while to build the confidence, but there's no future where you're manually reviewing every line of agent code. Not much acceleration in that. You need adversarial agent reviews, you need automated testing, and maybe you spot check. That's it. From prompt to production!
The only programmers who are really in trouble with the competition from artificial intelligence are the ones who insist the world today isn't all that different from the one we lived in last year.
Embrace the paradigm shift or perish.
(This has always been true.)
▸ 折叠3条(转推/噪音)
转推2026-09-29 RT @jorgemanru: Lexxy 1.0 is here, and we are going to make it the Rails default.
Lexxy answers a question we’ve been wrestling with at @3…
转推2026-09-28 RT @mdisec: As a @OmarchyLinux security team, specially @MeltonAErik and the team have been working so hard to triage the findings from @AF…
转推2026-09-28 RT @pocket_js: We strongly believe in this too. That’s why we compressed Omarchy’s signature tiling window UX down to a Nintendo 3DS, and i…
Over the last 5 years, I've used every social investing product -- but they all tend to fail in one key spot:
Closing the gap between (a) your belief about the world and (b) how to express that belief with stocks.
Mike reached out early in the summer and showed me @Supertake, which solves this problem in the most fun way possible: prompt with AI and it constructs a portfolio around your belief.
With the declining cost of intelligence, we're going to see much more sophisticated investing styles emerge that was previously reserved only for hedge funds.
引用@mignanoIntroducing @Supertake, a new platform that transforms your unique takes on the world into real, shareable investment portfolios using frontier AI and trading agents.
Supertake makes it easy for anyone to invest in their ideas, even if they know nothing … ↗
Marc Lou@marclouindie402.3K粉 · 10条公开自报500万美元净资产,并把收入验证做成创始人社交游戏与赞助位。
引用@marclouI made a game where you unlock chat rooms 💬 with founders at your MRR.
1. Connect @stripe (+10 other providers) to https://t.co/MZc8tGa5LQ
2. Go to https://t.co/eAjxmbAlms and customize your character
3. Enter private chat rooms based on your MRR
What … ↗
I made a game where you unlock chat rooms 💬 with founders at your MRR.
1. Connect @stripe (+10 other providers) to https://t.co/MZc8tGa5LQ
2. Go to https://t.co/eAjxmbAlms and customize your character
3. Enter private chat rooms based on your MRR
What happens in these rooms stays in these rooms 🤫 Only founders with verified revenue can join them.
Wander around the mini-world, talk to anyone, and make friends 🤝 Walk up to someone and press [Space] to see their startups.
Every founder with verified revenue gets a pet 🪴🌳🚀💎👑 that follows you everywhere. It shows your total revenue over the last 30 days: a tree 🌳 means $1K–$10K/mo.
→ https://t.co/eAjxmbAlms
I hid a few Easter eggs around the map. Show me what you find 🤗
I finally reached the magic number of $5M in net worth 🥳
This isn’t all cash. It’s everything combined:
+ My 36 startups (90%+ of their value comes from DataFast & TrustMRR)
+ Investments (85%+ in the S&P 500)
+ Cash
What surprises me most is getting here at 33 without ever:
- Working more than 8 hours a day
- Building anything grandiose
- Working for someone else
- Doing work I don’t like
- Going to an office
Don’t you dare give up :)
I tried Portugal last year, but didn't like it:
- People offering us cocaine in Lisbon
- English-speaking only in tourist areas
- Vibe is old/retirement
- Food was meh
Cyprus has a “get shit done now” culture. My local coffee shop opens at 7 am every day. 8 am on Sunday.
Portugal has more of a “we’ll get to it tomorrow” culture.
引用@buildinlisbonsame vibe in portugal.
grilled sea bass, vinho verde, sunset over the atlantic at azenhas do mar.
come build in lisbon for a month, dinner's on me ↗
$97.
We ate a fresh sea bass, octopus, Greek salad, and baklava.
We watched the sun disappear under the warm Mediterranean Sea. The sky turned red.
Two musicians played live music. The waves gently crashing on the shore sounded like a lullaby. It was peaceful.
I did not have to tip 15%. I tipped 15% because the waitress was kind and smiling all night.
I was not worried someone would snatch my phone sitting on the corner of the table.
People were walking by. Nobody was in a rush. Everyone looked peaceful.
I was at peace.
Thank you, Cyprus.
转推2026-09-28 RT @marclou: I made a game where you unlock chat rooms 💬 with founders at your MRR.
1. Connect @stripe (+10 other providers) to https://t.…
转推2026-09-28 RT @marclou: I finally reached the magic number of $5M in net worth 🥳
This isn’t all cash. It’s everything combined:
+ My 36 startups (90…
转推2026-09-27 RT @trust_mrr: ACQUIRED ✅ 🎉
A SaaS productivity tools for professionals just got acquired on TrustMRR.
💰 Sold for $12K
🤑 $307 revenue las…
Alex Finn@AlexFinnindie475.3K粉 · 2条高调推广Sonnet 5.5与最新Agent工具,观点以早期试用体验为主。
This is bigger news than it appears
I’ve had early access to Sonnet 5.5 for a bit now and honestly at first I couldn’t tell the difference between this and Opus 5.5
It’s a fraction of the price of Opus, which already felt like you got an absurd amounts of usage out of for cheap
It’s clear Anthropic has had some sort of breakthrough the last month
The jump in intelligence, speed, and affordability of all their 5.5 models is maybe the biggest leap we’ve ever seen in AI
You need to be using this model for all your standard, fastball down the center work
With its speed and intelligence you’ll be shocked at how fast you get tasks done
引用@claudeaiIntroducing Claude Sonnet 5.5, the second model in the Claude 5.5 family.
It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work. https://t.co/UvXD8mDTF1↗
The most dangerous thing you can do right now is NOT use the latest AI tools. Period.
Every day a new company is laying off thousands of people who don't know how to use the most modern AI tools
If I were in the 9-5 world right now, this is every step I'd take:
1. Download Claude Code and build your first app using Opus 5.5. Learn how to implement a front end and database (I’d use nextJS and Convex). AI can teach you all of this
2. Download Grok Bot. Tell the agent about your entire life. Career, goals, and ambitions. Ask it which agents it can build and which workflows it can implement to get you closer to those goals. This is called reverse prompting.
3. Anytime you're about to do work manually, do it in Grok Bot instead. Tell Grok the task you need to do then ask it how it can do the task for you. You'll be surprised by how much of your work you can automate
4. Feed Astra 6 Max your hardest problems. Make sure you hit the rate limit on it EVERY single week
5. Constantly look at your limits in all your AI plans. If you're ever not above 50% on your limits, get angry that you're not burning enough tokens.
6. Start creating content and building a platform. Resumes no longer matter. Companies dont look at them anymore. They check your platform. What you talk about, what you’re passionate about, how you educate others, who follows you. This matters 100x more than anything else when getting hired right now.
If you do these 6 things you are in excellent position to not only be safe in your career, but also dominate those that don't pick up these skills.
▸ 折叠2条(转推/噪音)
转推2026-09-29 RT @AlexFinn: This is bigger news than it appears
I’ve had early access to Sonnet 5.5 for a bit now and honestly at first I couldn’t tell…
转推2026-09-28 RT @AlexFinn: The most dangerous thing you can do right now is NOT use the latest AI tools. Period.
Every day a new company is laying off…
Tony Dinh@tdinh_meindie202.9K粉 · 2条转发和点评Opus 5.5生成视觉效果,信号集中在创意demo。
引用@kainex_y@levelsio@csonotes <insert mandatory Xanadu from Citizen Kane reference, with 'Rosebud' obviously being longing for the early nomad life years> ↗
I think this is true and I do like it
Although you have to be VERY careful not to isolate yourself, which does happen once you get rich
@csonotes always trolls me for being at home in my "aquarium" and he's right, once your home is so great it becomes a prison, so you have to force yourself to go out!
We travel a lot of the time though so this is just a home base and it's NICE to be home, but yes you gotta fight the isolation risk
引用@potaufanMoney is really a means to an end. And that end is more privacy, convenience, and quietness, and less daily friction in accomplishing things that bring value.
Most people put up with shit because they can't afford to move beyond it. They're stuck in a noisy … ↗
Yeah I feel having a home gym is a pretty big trend now
I think I got sick of going to the local gym because:
1) had to drive there 4 times per week
2) the racks/machines were often busy/used
3) then in middle of my workout a group class would start
4) then OnlyFans girl with inside ass shorts and neon blinking buttplug would walk around scouting for subscribers (true story and a bit distracting)
5) local gym had lots of drama with the owner allegedly not paying his staff and much more
Now I just walk downstairs and I got my home gym, and personal trainers comes to our place, never busy, no distractions, no drama
Bit less social though but for that we balance out organizing parties
引用@homegymcoop@JohnStrongHodl People all over the world have Rogue equipment, for instance, @levelsio just added a Rogue Trap Bar (made in the USA) to his home gym even though he's overseas.
Also, the amount of new home gym buyers and commercial/school facilities that … ↗
I think every time AI can do something equally good or better than humans it stops being valuable or special
Marketing videos used to take lots of work of editing and motion graphics and a big budget
Now with Opus 5.5 etc it costs $0 to make one as good or better
That means ANYONE can make a marketing video now, so the timeline gets flooded with these videos (as you see happening now) and people stop caring about them, stop watching them and they stop grabbing attention
The response then to still grab attention is be different and that probably means IRL marketing videos that are original and unique and maybe personal
It's the grim reaper meme where AI keeps commoditizing another media format and bringing the cost down to close to $0
It already did so for graphic design (2023), last year with Nano Banana for photography (2025), this year with Seedance 2.5 it did it for video (2026) and this and next year motion graphics
引用@sporadicahot take i’ve been mulling over:
launch videos are dead (or at least in their dying breath)
oh people will still do them but they are a dying form, actual good, valuable launches will pivot to different forms ↗
Lindt has an active lawsuit going against it for elevated levels of lead and cadmium
Most interestingly Lindt's lawyers themselves have suggested Lindt is not "expertly crafted" or "excellent" but that that is just advertising
So if even Lindt itself doesn't consider itself good, why should I be eating Lindt chocolate?
Now I eat Vivani which sources from Panama which seems cleaner chocolate beans
引用@lionelrudazLindt is the worst chocolate in Switzerland. Hard no go for me. ↗
引用@levelsioI've now fully integrated my new site Hotelist into Hoodmaps
So now it lets you tap [ 🏨 Hotels ] and you see all available hotels in the area
You can hover over them or tap them and it'll show the same hover tooltip you see on Hotelist, and then when you … ↗
I've now fully integrated my new site Hotelist into Hoodmaps
So now it lets you tap [ 🏨 Hotels ] and you see all available hotels in the area
You can hover over them or tap them and it'll show the same hover tooltip you see on Hotelist, and then when you tap again it opens the hotel
Nice to find a good area in a city and then book a place there
This was a recommendation by AI to try and monetize Hoodmaps, it gets almost 171,000 vistors per month but never made money in almost a decade, so I asked AI
It makes sense cause hotels are something aligned with when you're trying to find the best area in a city and it makes money!
It only shows when you click [ 🏨 Hotels ] though, so not annoying
引用@levelsioStrava's data is quite protected so I couldn't index it for Hoodmaps
But the good thing is the US Census data publishes median income level and the data is very fresh and accurate, so I added it to 🗺️ https://t.co/2R4tRNJF0z
So now you can switch to [ 💰 … ↗
Usually SaaS sales calls are to gauge how much you can spend and get you to spend the most money you could
So it's like maximizing revenue and giving everyone a custom plan
More good for businesses than for customers IMHO
引用@VynseDev@levelsio@halfdantimm Why the hell would you not let someone buy your product/service and insist for a call instead
Deserved ↗
I love this
SaaS sales guy wants to do a call
@halfdantimm hates calls so he says no
Then he vibecodes his own replacement for free!
引用@halfdantimmI just vibe-coded a B2B SaaS instead of buying it, because the salesperson insisted on a meeting before sending me the price.
Have seen @levelsio write about this and now I experienced it myself.
A salesperson reached out to me on LinkedIn about a platform … ↗
This is getting really good
The last thing Claude is bad at now is sound, it always does these "ticky" synth sounds that are way worse in quality than the visuals it now produces
引用@xikharThird iteration with Opus 5.5 medium.
I am in awe. It turned blender, image-gen, and three.js into this beauty, which runs on your browser.
An understatement to say that Anthropic cooked. You are looking at the future of game dev. https://t.co/ht2s8rpwO3↗
▸ 折叠2条(转推/噪音)
转推2026-09-27 RT @SafaElmali: I built a website that plays endless lofi over pixel-art cities at night using Opus 5.5
The music isn't a playlist. A br…
转推2026-09-27 RT @scottstts: My god this is such a good speech that every SWE needs to hear. You know what? Every person should hear it
Keep the happy m…
Jon Yongfook@yongfookindie172.6K粉 · 6条从单产品专注转向一人产品组合,持续判断MCP会怎样削弱固定界面。
Meta ads is a good example of a site that is made 100x better by MCP.
The UI is an absolute abomination. Letting an agent do everything is so much better for your sanity.
Someone is going to build an MCP that submits my MCP to MCP stores / directories, keeps it up to date etc, and they will make a lot of money.
I'm building like 17 apps right now and I don't want to submit them all manually.
I used to think you should focus on one product.
But that was in the hand-coding era, before INFINITE INTELLIGENCE.
I now think it's doable to run a portfolio solo.
In fact I think that's the smarter play now. Putting all your eggs in one basket is risky in a time when the next model might make your app obsolete, and anyone can clone it in a weekend anyway.
Diversify!
I'm becoming so MCP-pilled that I'm really starting to question... do we need an interface anymore?
For SaaS products that are essentially plumbing / routing, there will come a time where the UI is optional.
This weekend was my final transition into a meat proxy
- made a one-shot launch video, it's amazing
- built an app in Rails 8 without touching any code
- made a landing page in one prompt
All via Claude.
Jonathan Wilke@jonathan_wilkeindie29.6K粉 · 15条密集讨论Opus 5.5、生成广告同质化、简化能力与发布页面细节。
As so many people were shocked by the fact that I have never tried Opus 5.5, I will today switch from Cursor with Grok 4.7 to using Claude Code with Opus 5.5 for the whole day.
I'll keep you posted on my epxerience
He is basically saying “use a high-quality boilerplate to get the best outcome for your code” 😁
Luckily there is https://t.co/eKVhJmwJmV
引用@trq212it's basically impossible for someone to just "show you their prompt" now, because everything is about references, skills and examples
I often ask my agent to look at 3 other repos I've made first, search the web for references, use other AI APIs, etc. ↗
引用@johnrushIt’s Sept 28, 2026,
For the first time in my life I’m genuinely convinced I might never again delegate a task to a human.
Opus5.5 is AGI based on my personal benchmarks running over 20 startups (coding, marketing, seo, content, operations, accounting, … ↗
Seems like Opus 5.5 has cracked the secret code of copywriting 😱
引用@arvidkahlThe amount of em-dashes in Opus 5.5's email-writing prose has dropped SIGNIFICANTLY.
Detecting AI writing will eventually become impossible.
Good thing? Yes. Bad thing? Yes. ↗
All of these one-prompt video ads generated by Opus 5.5 look the same and it's just so obvious that they are AI-generated that likely no one will watch further than 3 seconds.
Create something original.
This is probably one of the most valuable lessons you can learn in life.
引用@heyandrasSocial media is scary, no matter how good you are or your product, there are always haters, negative people.
Keep in mind, ignore them. I teach this to my kids as well, just not social media related, but irl, in school. ignore the haters, bullying kids. … ↗
引用@mynameisyahia“We’re probably going to build this internally”
2 weeks later: subscribed to our highest plan
This has happened so many times I’m considering making it our official onboarding flow ↗
If you're struggling with good UX, please watch this video.
It's actually a masterclass in how to reduce & focus the UI on the relevant things.
(watch the explanations at the bottom for why he made those changes)
引用@moguzbulbulI was struggling to explain my UX decisions in interviews, so I made this kind of animation with Opus 5.5
Anyone want the skill? https://t.co/4nPoVpTe6r↗
incredibly lucky to be working with Filip on Bazzly
incredibly smart and hardworking guy 👏
引用@FilipPanoskiI sat at $100 MRR for 6 months.
today: $15k MRR.
nothing viral happened. I just did 4 things, in this order:
1. 𝗜𝗺𝗽𝗿𝗼𝘃𝗲 𝘆𝗼𝘂𝗿 𝗼𝗳𝗳𝗲𝗿 𝗯𝘆 𝘁𝗮𝗹𝗸𝗶𝗻𝗴 𝘁𝗼 𝘂𝘀𝗲𝗿𝘀
talk to the people who didn't convert.
talk to the people who churned.
ask them why.
whatever they tell … ↗
the more I use it, the more I think the Chrome extension alone would justify using SuperX
it's incredibly useful when you are using X on a daily basis, and you want to check people out fast https://t.co/1hxzitTdIZ
引用@mdnlabsI forgot how insane SuperX was.
Engaging has never felt this easy 👇
My whole routine picked up right where it left off.
Catches every reply worth catching and formats posts better than I do.
Bouta get monetized in 30 days 😎 https://t.co/3qnu9GzExx↗
one of my users has billed $45,000 setting up my $99/month product for other businesses
and I'm not even mad 😅
it's something we don't even offer
some context:
Erika runs 2 businesses alone from Spring Hill, Tennessee
she's not a programmer. she's actually from an ops background, and she joined https://t.co/FrTmeMTA6U in March
she has set up a chief of staff agent plus 5 specialists (research, sales, content, compliance, SEO), and she talks to them via Telegram
the first thing she did with Squad was ask the chief of staff agent to audit the CRM bill of the company she just took over
and it went from $2,500/month to $1,500/month, so she saved $12k a year with a single prompt 💥
then there's her 10pm workflow:
- she drives home from a networking event with 40 contacts
- she pastes all of them into Telegram from her phone
- by the time she gets home, 40 personalized follow-ups are sitting in her Gmail drafts
this used to eat her entire Saturday, and now it takes her 15 mins
small business owners around her saw that and asked her "can you do this for me?"
so she does 😅
she handles the setup and training, and she has done it for a couple dozen clients so far
she has billed $45,000 just for installing a $99/month tool for other people
the money in AI right now goes to whoever knows 30 business owners who want to start using AI agents but have no idea how
you can read the full story here 👇
https://t.co/lCdMtjlUix
3 years ago, I acquired Typeframes for $50,000
and turned it into an asset generating $500,000 per month
one of my biggest success so far
Revid powered by Opus with a few real footage as input is absolutely wild, watch how this turned out 👇 https://t.co/Yw29NxGZhF
Grok Bot and Muse users, this is the reason to switch to https://t.co/ldcOOoMlBU
every agent gets its own real computer and browser. give it your login once or sign in yourself on its screen
when a captcha or 2FA shows up it pings you, you do that one step, it carries on
引用@garrytanIf you like Muse and Grok Bot but still run into crazy antibot annoyances try @AsideAI browser with MCP with your agent on a spare laptop or computer you keep plugged in somewhere.
It lets your agents your real credentials as yourself from a real Chromium. … ↗
built the best resource for your own Motion Graphics video
including the 50+ best videos made by Opus
the absolute BEST
a collection that got MILLIONS of views
most are easy to recreate with 1 prompt
start here: https://t.co/ivSDU32rDi
this is free inspiration, don't sleep on it
▸ 折叠4条(转推/噪音)
转推2026-09-28 RT @NCoutureau: Eleven v4 is live on Revid! Now you can add expressive voice models to your videos, including whispers and laughs, in over…
转推2026-09-28 RT @robj3d3: I forget how insane the SuperX extension is.
The reason I can tell what goes viral now is because I see everyone's top posts…
转推2026-09-28 RT @mdnlabs: I forgot how insane SuperX was.
Engaging has never felt this easy 👇
My whole routine picked up right where it left off.
Cat…
转推2026-09-28 RT @01ayushgarg: I made my AI agent sign up for my own SaaS as a new user
- it recorded every user flow and turned them into small motion…
引用@arvidkahlWhenever my analysis and inference GPU fleet starts getting 503's and 429s on the OpenAI flex tier, a new OpenAI model release is imminent.
Is there a prediction market for this? 🤣 ↗
Similar experiences here. Beats Fable, all OpenAI models, and definitely runs leaps around my own skill ceiling.
And it’s cheap, barely moves the usage needle in the $200 plan. The value you get is massive, if you know how to set up your systems.
引用@johnrushIt’s Sept 28, 2026,
For the first time in my life I’m genuinely convinced I might never again delegate a task to a human.
Opus5.5 is AGI based on my personal benchmarks running over 20 startups (coding, marketing, seo, content, operations, accounting, … ↗
Whenever my analysis and inference GPU fleet starts getting 503's and 429s on the OpenAI flex tier, a new OpenAI model release is imminent.
Is there a prediction market for this? 🤣
The amount of em-dashes in Opus 5.5's email-writing prose has dropped SIGNIFICANTLY.
Detecting AI writing will eventually become impossible.
Good thing? Yes. Bad thing? Yes.
There are moments during my daily work with Claude Code where I just want to have an exploratory audio chat with the tool. Like what Claude allows on mobile.
Right now, from desktops & terminals, that just doesn’t work. I’d love to see that feature. Whatdya think @AnthropicAI?
Podcasting is built on RSS and it’s wildly effective for a whole host of things: independent media publishing, creator monetization, and most importantly ease of access for anyone, demand- and supply-side.
引用@oldstackjournalRSS is such a beautifully boring piece of technology.
A site publishes a feed. Your reader checks it. New stuff appears. ↗
▸ 折叠1条(转推/噪音)
转推2026-09-28 RT @arvidkahl: Whenever my analysis and inference GPU fleet starts getting 503's and 429s on the OpenAI flex tier, a new OpenAI model relea…
Danny Postma@dannypostmaindie185.1K粉 · 7条试用Grok Bot做个人事务与客服,并强调产品组合要靠交叉分发。
distribution is the only moat left so you better use all your product to cross-promote eachother
this always work for me before and seems to be even more important now
引用@yongfookI used to think you should focus on one product.
But that was in the hand-coding era, before INFINITE INTELLIGENCE.
I now think it's doable to run a portfolio solo.
In fact I think that's the smarter play now. Putting all your eggs in one basket is risky … ↗
Bit late to the bandwagon but man Grok Bot is fun.
Basically using it as a personal assistant now, writing email drafts for outreach, cancelling subscriptions.
Anything I'm too lazy to do I'll dump in there and see if it sticks (it always does) https://t.co/OfVmE2UGT6
i wonder if product managers now have the most valuable skill
they know exactly what to ship and can literally prompt it all together themselves now instead
i think whats so amazing about opus 5.5 is that it works so fast there is near-zero waiting time
it feels like you are doing the work instead of waiting for a "collegue" to finally reply to you
Higgsfield hit the reset button on all the credits you’ve used this month, up to 2,000!
We are so back. 🔥
引用@higgsfieldA small thank you to our community 💚
To celebrate Higgsfield’s $1 Billion revenue run-rate, we’re resetting the subscription credits you’ve used this month.
Keep making things you love. This one’s on us. https://t.co/ycyJO7tPcw↗
Probably the biggest update I’ve made to Designjoy in years:
You can now get your product designed AND built through Designjoy.
Front-end development and MVP builds are now included in every Designjoy subscription at no additional cost.
So instead of getting a beautiful Figma file and figuring out the rest, Designjoy will take the design all the way to a working product.
The effects this will have on Designjoy will be huge.
Much of the design phase will be totally optional, and projects can start directly in code.
This will increase the output you get from Designjoy by at least 10x.
Despite this, the price of a Designjoy subscription will remain at $5,000/m for now, roughly 50% cheaper than competing services.
I'm looking forward to testing this with a few clients over the coming weeks.
If this sounds interesting to you and you think it might be a fit for an upcoming project, DM me for a discount.
Designjoy把前端开发与MVP构建纳入原有5000美元月费方案。
▸ 折叠1条(转推/噪音)
转推2026-09-28 RT @alexmashrabov: Calling Higgsfield “just a wrapper” tells you surprisingly little about the economics of our business.
In over 40% of c…
Adam Lyttle@adamlyttleappsindie58.9K粉 · 1条自报应用达到500美元MRR,数字来自作者更新。
转推2026-09-27 RT @dhh: The constant decel messaging in Europe is insufferable as an adult, but its effect on kids is worse. As a parent, it's your obliga…
the revenue from the past 48 hours emphatically says yes.
absolutely wild for a tweet I sent out on a whim on a Friday afternoon as I was finishing up for the week. 🎉
引用@Shpigfordi think i just started an SEO agency ↗
Did you get Qwen 3.8 27B set up locally yet? It's a good time now.
引用@thsottiauxHi,
Tomorrow we are re-opening the Pro $200 subscriptions to new subscribers, but together with it we are also changing how we calculate the usage for it. In effect, if you do the math, it will net out at half the dollar in API spend compared to the old Pro … ↗
引用@claudeaiIntroducing Claude Sonnet 5.5, the second model in the Claude 5.5 family.
It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work. https://t.co/UvXD8mDTF1↗
For the first time since AI came out, Opus 5.5 makes me feel like I could automate ALL my businesses with it
I could even give it my X account, you probably won't notice lol
▸ 折叠1条(转推/噪音)
噪音2026-09-29 Almost 50 people on the waitlist 😱 the pressure have never been so high!!
We're starting the testing phase today with a few people 👀
I've been fascinating by the concept of having a north star recently.
If we don't have a destination in mind and a reason to persevere on the journey there we're ngmi
NEW Zero To Paid module just dropped 💣
The cold emails and DMs that took me from $0 to $1,000,000
Exact emails with my full process
How I got customers like Jason Calacanis through outreach
Get it here → https://t.co/cAOrzKSqxk
SaaS is not sexy.
Founders shouldn't chase SaaS unless they will be deeply happy solving the same boring operator problems every day in a market they understand.
It's not about building fun features everyday.
It's about sales, ICP, churn, activation, onboarding, retention and support.
Marc Köhlbrugge@marckohlbruggeindie88.9K粉 · 3条展示可塑个人助手,并关注域名与品牌选择。
引用@paulgThere is always a .com you can get for $5k that's better the tryblurgh.ai you're currently using. The reason you don't see it is (a) lack of imagination and (b) because you're attached to Blurgh, which you shouldn't be, because it's not that great anyway. ↗
The coolest part is that my bot has no food tracking specific code. It’s a completely malleable product
I just told it what I want it to do, and behind the scenes it creates the right data structures, cron jobs, widgets, etc to make this macro tracking work
引用@marckohlbruggeI’ve been building my own AI assistant called MarcBot
One thing I just added is an Apple Watch complication (bottom right) to track my macros
Calories / protein / carbs
I can tap it, say out loud what I ate, and it gets tracked and the rings updated … ↗
I’ve been building my own AI assistant called MarcBot
One thing I just added is an Apple Watch complication (bottom right) to track my macros
Calories / protein / carbs
I can tap it, say out loud what I ate, and it gets tracked and the rings updated https://t.co/oZMfqBVJzl
Peter Askew@searchboundindie26.2K粉 · 1条继续关注过期域名拍卖的时间成本与耐心门槛。
I haven't seen anyone getting good results from ChatGPT ads yet which is a bummer because I think it would be great to have a real alternative to Google Ads.
It's more than search too as it can suggest products people weren't even thinking about based on the conversation.
I suspect one cause might be that the ads only appear to people who aren't paying.
Also CPC/CPM seem very high but it's still early so I'd give it some time before deciding it's not a good channel.