Lost and Found

Things I stumbled upon that caught my attention

[TWITTER]
Sep 23
Cursor@cursor_ai

We've reduced token costs in Cursor by 7% with no drop in agent quality. Savings came from tighter prompts, selective tool loading, better caching, and compressed file reads.

Go to source
[TWITTER]
Sep 23
Cursor@cursor_ai

We've reduced token costs in Cursor by 7% with no drop in agent quality. Savings came from tighter prompts, selective tool loading, better caching, and compressed file reads.

Go to source
[TWITTER]

Cursor

We've reduced token costs in Cursor by 7% with no drop in agent quality. Savings came from tighter prompts, selective tool loading, better caching, and compressed file reads.

Go to source
[GITHUB]REPO
Sep 23

A coordinator conversation, parallel worker threads, shared memory and an overview of what needs you. A Herdr plugin.

eliasstravik/herdr-projectseliasstravik
Go to source
[GITHUB]eliasstravik/herdr-projects
Loading README…
Go to source
[GITHUB]

A coordinator conversation, parallel worker threads, shared memory and an overview of what needs you. A Herdr plugin.

Go to source
[TWITTER]
Sep 23
Guillermo Rauch@rauchg

Muse, Instinct, OpenClaw, Claude Code… All successful agents have 3 key components: 🧠 Brain → model, harness (logic) 👐 Hands → tools, computer, browser 🗃️ Files → memories, skills, repos The 'easy' way is to throw all these in 1 stateful computer (a Mac Mini) Like, you run 𝚌𝚕𝚊𝚞𝚍𝚎 or 𝚏𝚡 in your mac, you keep it running all day with 𝚌𝚊𝚏𝚏𝚎𝚒𝚗𝚊𝚝𝚎, it has storage, and CLIs and apps installed. But if you want to cost-efficiently run agents in the cloud, you actually start breaking down these parts. 🧠 The harness can run in Fluid compute. To make it reliable across restarts,…

Vercel Developers@vercel_dev

Vercel Sandbox now has persistent storage with Drives, in public beta on every plan. ▪︎ Store agent workspaces, data, models, deps ▪︎ Read snapshots across parallel sandboxes ▪︎ Mount up to four Drives per sandbox ▪︎ Up to 16 TiB per Drive https://vercel.com/changelog/drives-for-vercel-sandbox-are-now-in-public-beta

[TWITTER]
Sep 23
Guillermo Rauch@rauchg

Muse, Instinct, OpenClaw, Claude Code… All successful agents have 3 key components: 🧠 Brain → model, harness (logic) 👐 Hands → tools, computer, browser 🗃️ Files → memories, skills, repos The 'easy' way is to throw all these in 1 stateful computer (a Mac Mini) Like, you run 𝚌𝚕𝚊𝚞𝚍𝚎 or 𝚏𝚡 in your mac, you keep it running all day with 𝚌𝚊𝚏𝚏𝚎𝚒𝚗𝚊𝚝𝚎, it has storage, and CLIs and apps installed. But if you want to cost-efficiently run agents in the cloud, you actually start breaking down these parts. 🧠 The harness can run in Fluid compute. To make it reliable across restarts, rollouts, crashes, you make its event log durable using Workflow. 👐 The hands can be a dedicated browser fleet like Browserbase/Kernel, a computer like Sandbox, and even more efficient lightweight tools like just-bash. 🗃️ 🆕 What was missing was a way to also decouple storage. Imagine you want to run a memory consolidation cron job every night ("dreaming"). You can read/write to the files directly without 'booting up' the agent's full computer. Today we're introducing the perfect companion to Sandbox: Drives. We shipped the computer for agents, now we're giving you the 'external disk' you can attach at will. It's early, and we'll be expanding capabilities here quickly. Btw, breaking apart the agent into these independent parts not only optimizes costs in a big way, it also *massively* improves security and auditability. I'd argue you can't even run a secure agent otherwise!

Vercel Developers@vercel_dev

Vercel Sandbox now has persistent storage with Drives, in public beta on every plan. ▪︎ Store agent workspaces, data, models, deps ▪︎ Read snapshots across parallel sandboxes ▪︎ Mount up to four Drives per sandbox ▪︎ Up to 16 TiB per Drive https://vercel.com/changelog/drives-for-vercel-sandbox-are-now-in-public-beta

Go to source
[TWITTER]

Guillermo Rauch

Muse, Instinct, OpenClaw, Claude Code… All successful agents have 3 key components: 🧠 Brain → model, harness (logic) 👐 Hands → tools, computer, browser 🗃️ Files → memories, skills, repos The 'easy' way is to throw all these in 1 stateful computer (a Mac Mini) Like, you run 𝚌𝚕𝚊𝚞𝚍𝚎 or 𝚏𝚡 in your mac, you keep it running all day with 𝚌𝚊𝚏𝚏𝚎𝚒𝚗𝚊𝚝𝚎, it has storage, and CLIs and apps installed. But if you want to cost-efficiently run agents in the cloud, you actually start breaking down these parts. 🧠 The harness can run in Fluid compute. To make it reliable across restarts,…

Go to source
[TWITTER]
Sep 23
QuiverAI@QuiverAI

Introducing Arrow 2 Our latest and most advanced models for generating precise, editable vector graphics. Higher quality. Faster outputs. Available now in App and API.

[TWITTER]
Sep 23
QuiverAI@QuiverAI

Introducing Arrow 2 Our latest and most advanced models for generating precise, editable vector graphics. Higher quality. Faster outputs. Available now in App and API.

Go to source
[TWITTER]

QuiverAI

Introducing Arrow 2 Our latest and most advanced models for generating precise, editable vector graphics. Higher quality. Faster outputs. Available now in App and API.

Go to source

Weekly picks

The best links I found this week, with context.

My (human) thoughts on what I think matters and why. No AI slop.

No ads. No bullshit. Unsubscribe anytime.

[OTHERS]
Sep 23

Introducing transit rewards

Today, Waymo is excited to announce our transit rewards program, which we will initially offer in the San Francisco Bay Area, with future cities to follow. This first-of-its-kind initiative rewards riders with Waymo Cash when they connect their Waymo rides with public transit and use their Visa card.

waymo.comWaymo
[OTHERS]

Introducing transit rewards

Today, Waymo is excited to announce our transit rewards program, which we will initially offer in the San Francisco Bay Area, with future cities to follow. This first-of-its-kind initiative rewards riders with Waymo Cash when they connect their Waymo rides with public transit and use their Visa card.

Go to source
[TWITTER]
Sep 23
Peter Steinberger 🦞@steipete

Next version of OpenClaw uses a decision model to automatically decide between steer or queue. (Lab feature, we support Jef* and API-compat (e.g. local models like Kef...) and ONNX variants)

[TWITTER]
Sep 23
Peter Steinberger 🦞@steipete

Next version of OpenClaw uses a decision model to automatically decide between steer or queue. (Lab feature, we support Jef* and API-compat (e.g. local models like Kef...) and ONNX variants)

Go to source
[TWITTER]

Peter Steinberger 🦞

Next version of OpenClaw uses a decision model to automatically decide between steer or queue. (Lab feature, we support Jef* and API-compat (e.g. local models like Kef...) and ONNX variants)

Go to source
[TWITTER]
Sep 23
Epoch AI@EpochAIResearch

AI is getting cheaper more quickly than any other transformative tech in history. At a given level of performance, cost has fallen ~47%/quarter since 2023. That’s 4× faster than DNA sequencing, 6× faster than compute, 18× faster than lithium batteries, and (up to 1973) 54× faster than electricity.

[TWITTER]
Sep 23
Epoch AI@EpochAIResearch

AI is getting cheaper more quickly than any other transformative tech in history. At a given level of performance, cost has fallen ~47%/quarter since 2023. That’s 4× faster than DNA sequencing, 6× faster than compute, 18× faster than lithium batteries, and (up to 1973) 54× faster than electricity.

Go to source
[TWITTER]

Epoch AI

AI is getting cheaper more quickly than any other transformative tech in history. At a given level of performance, cost has fallen ~47%/quarter since 2023. That’s 4× faster than DNA sequencing, 6× faster than compute, 18× faster than lithium batteries, and (up to 1973) 54× faster than electricity.

Go to source
[GITHUB]REPO
Sep 23

Workflow engine, JSON control-flow tool, and live terminal viewer for the pi coding agent

osolmaz/pi-workflowsosolmaz
[GITHUB]osolmaz/pi-workflows
Loading README…
[GITHUB]

Workflow engine, JSON control-flow tool, and live terminal viewer for the pi coding agent

Go to source
[TWITTER]
Sep 22
Fireside Alpha@firesidealpha

Anthropic Technical Staff Thariq Shihipar says a smarter model needs a heavier harness, not a lighter one "The harness is super important. Sometimes there's this idea that the harness doesn't matter, because the models will get better and better." "And if the model just does everything perfectly, then why do you need a harness at all?" "And in practice, what we see is the models get better and better. And so the harness needs to become more and more complicated to allow the model to do more things." "And an example of this is auto mode. Auto mode is a classifier that runs after every task…

Fireside Alpha@firesidealpha

Anthropic Technical Staff Jess Yan says the harness and the model cannot be pulled apart without giving up performance, though she flags her own bias. "I think I am quite biased but I also think that it is impossible to get the maximum possible performance without tying together the harness and the model." "Now the components of the harness and maybe the thickness of the harness will change over time as models get more and more capable." "However, when we test our models and when we are assessing their performance we always have to test it in conjunction with a harness." "And are we going to test it with all the different harnesses of the world? We're going to select the harnesses that we have built." "And so there is an aspect of the necessity of building models is that you have to be testing them with harnesses, and that sort of keeps them paired together."

[TWITTER]
Sep 22
Fireside Alpha@firesidealpha

Anthropic Technical Staff Thariq Shihipar says a smarter model needs a heavier harness, not a lighter one "The harness is super important. Sometimes there's this idea that the harness doesn't matter, because the models will get better and better." "And if the model just does everything perfectly, then why do you need a harness at all?" "And in practice, what we see is the models get better and better. And so the harness needs to become more and more complicated to allow the model to do more things." "And an example of this is auto mode. Auto mode is a classifier that runs after every task that you normally have to ask a permission prompt for Claude." "And back when we were Opus 4 or even Opus 4.5, it was not so bad to hit enter on the permission prompts because the turns would only last a few minutes anyways." "And now Claude is running, it can run for hours, and so you really need that ability for it to do work safely, stick to your instructions. And auto mode is really complicated software. Sandboxing is really complicated software." ______ More takeaways from Thariq's conversation: https://firesidealpha.substack.com/p/the-race-off-the-raw-model-openai

Fireside Alpha@firesidealpha

Anthropic Technical Staff Jess Yan says the harness and the model cannot be pulled apart without giving up performance, though she flags her own bias. "I think I am quite biased but I also think that it is impossible to get the maximum possible performance without tying together the harness and the model." "Now the components of the harness and maybe the thickness of the harness will change over time as models get more and more capable." "However, when we test our models and when we are assessing their performance we always have to test it in conjunction with a harness." "And are we going to test it with all the different harnesses of the world? We're going to select the harnesses that we have built." "And so there is an aspect of the necessity of building models is that you have to be testing them with harnesses, and that sort of keeps them paired together."

Go to source
[TWITTER]

Fireside Alpha

Anthropic Technical Staff Thariq Shihipar says a smarter model needs a heavier harness, not a lighter one "The harness is super important. Sometimes there's this idea that the harness doesn't matter, because the models will get better and better." "And if the model just does everything perfectly, then why do you need a harness at all?" "And in practice, what we see is the models get better and better. And so the harness needs to become more and more complicated to allow the model to do more things." "And an example of this is auto mode. Auto mode is a classifier that runs after every task…

Go to source
[OTHERS]
Sep 22

How will AI change operating systems? Part 2: Windows

Deepdive into the Windows team’s efforts to make the OS “AI agent-friendly” and win back developers by going all-in on Linux on Windows, local models, GPUs, & more

newsletter.pragmaticengineer.comGergely Orosz
[OTHERS]

How will AI change operating systems? Part 2: Windows

Deepdive into the Windows team’s efforts to make the OS “AI agent-friendly” and win back developers by going all-in on Linux on Windows, local models, GPUs, & more

Go to source
[OTHERS]
Sep 22

Introducing Claude Opus 5.5

Claude Opus 5.5 leads in agentic coding and knowledge work, and costs 40% less to run than Opus 5 on typical workloads.

anthropic.com
[OTHERS]

Introducing Claude Opus 5.5

Claude Opus 5.5 leads in agentic coding and knowledge work, and costs 40% less to run than Opus 5 on typical workloads.

Go to source
[OTHERS]
Sep 22

Pi Coding Agent

A terminal-based coding agent

pi.dev
[OTHERS]

Pi Coding Agent

A terminal-based coding agent

Go to source
[TWITTER]
Sep 22
Kirill Skrygan@kskrygan

Today we introduce JetBrains Air - the product system for software development orgs in the age of agents. Built for developers, team leads & engineering directors. Open and flexible: steer and control almost any LLM & harness. Read more: https://jb.gg/air-announce

[TWITTER]
Sep 22
Kirill Skrygan@kskrygan

Today we introduce JetBrains Air - the product system for software development orgs in the age of agents. Built for developers, team leads & engineering directors. Open and flexible: steer and control almost any LLM & harness. Read more: https://jb.gg/air-announce

Go to source
[TWITTER]

Kirill Skrygan

Today we introduce JetBrains Air - the product system for software development orgs in the age of agents. Built for developers, team leads & engineering directors. Open and flexible: steer and control almost any LLM & harness. Read more: https://jb.gg/air-announce

Go to source
[TWITTER]
Sep 22
Arnav Gupta@championswimmer

Most models of major.minor version means the major version is a pre train checkpoint and the minor is a post train. The post training makes the model more “agentic” but not truly more intelligent. It has just been tortured to “think out loud” for longer. But that said no one ones when you’ve “saturated” the raw intelligence of the pre trained base model until you actually reach there. The point where the the + 0.1 releases increase very little on benchmark but a lot on tokens means it is time to pretrain a new model.

Shantanu Goel@shantanugoel

The diff in intelligence v/s output tokens for Grok 4.6 to Grok 4.7

[TWITTER]
Sep 22
Arnav Gupta@championswimmer

Most models of major.minor version means the major version is a pre train checkpoint and the minor is a post train. The post training makes the model more “agentic” but not truly more intelligent. It has just been tortured to “think out loud” for longer. But that said no one ones when you’ve “saturated” the raw intelligence of the pre trained base model until you actually reach there. The point where the the + 0.1 releases increase very little on benchmark but a lot on tokens means it is time to pretrain a new model.

Shantanu Goel@shantanugoel

The diff in intelligence v/s output tokens for Grok 4.6 to Grok 4.7

Go to source
[TWITTER]

Arnav Gupta

Most models of major.minor version means the major version is a pre train checkpoint and the minor is a post train. The post training makes the model more “agentic” but not truly more intelligent. It has just been tortured to “think out loud” for longer. But that said no one ones when you’ve “saturated” the raw intelligence of the pre trained base model until you actually reach there. The point where the the + 0.1 releases increase very little on benchmark but a lot on tokens means it is time to pretrain a new model.

Go to source
[OTHERS]
Sep 22

OpenCode Reloaded

The subtle pleasure of hot reloading, and how OpenCode keeps its environment changing while the agent keeps working.

anoma.ly
[OTHERS]

OpenCode Reloaded

The subtle pleasure of hot reloading, and how OpenCode keeps its environment changing while the agent keeps working.

Go to source
[TWITTER]
Sep 22
Will Eastcott@willeastcott

A pivotal moment for real estate! 🏡 Capturing a property in 3D used to mean a $5,000 LiDAR scanner. This house was scanned with a $500 @insta360 X5 - a consumer 360 camera. And best of all? All of the software is free and open source! It's based on a technique called 3D Gaussian splatting: 🪄 Splat creation: Spirula Studio 🖼️ Splat rendering: @PlayCanvas So for $500, you can present any property online and let prospective purchasers freely explore as if it was a videogame. 🎮 [1/3]

[TWITTER]
Sep 22
Will Eastcott@willeastcott

A pivotal moment for real estate! 🏡 Capturing a property in 3D used to mean a $5,000 LiDAR scanner. This house was scanned with a $500 @insta360 X5 - a consumer 360 camera. And best of all? All of the software is free and open source! It's based on a technique called 3D Gaussian splatting: 🪄 Splat creation: Spirula Studio 🖼️ Splat rendering: @PlayCanvas So for $500, you can present any property online and let prospective purchasers freely explore as if it was a videogame. 🎮 [1/3]

Go to source
[TWITTER]

Will Eastcott

A pivotal moment for real estate! 🏡 Capturing a property in 3D used to mean a $5,000 LiDAR scanner. This house was scanned with a $500 @insta360 X5 - a consumer 360 camera. And best of all? All of the software is free and open source! It's based on a technique called 3D Gaussian splatting: 🪄 Splat creation: Spirula Studio 🖼️ Splat rendering: @PlayCanvas So for $500, you can present any property online and let prospective purchasers freely explore as if it was a videogame. 🎮 [1/3]

Go to source
[TWITTER]
Sep 22
clem 🤗@ClementDelangue

Most companies that shut down give up their impact. This startup did something much better: they open-sourced 1,274 hours of egocentric robotics data. 13,451 recordings of humans doing everyday tasks, as a gift to the robotics community. Thank you @eidon_ai! Open source has a superpower: work can outlive the organization that created it. More startups should do this! https://huggingface.co/eidon-ai

[TWITTER]
Sep 22
clem 🤗@ClementDelangue

Most companies that shut down give up their impact. This startup did something much better: they open-sourced 1,274 hours of egocentric robotics data. 13,451 recordings of humans doing everyday tasks, as a gift to the robotics community. Thank you @eidon_ai! Open source has a superpower: work can outlive the organization that created it. More startups should do this! https://huggingface.co/eidon-ai

Go to source
[TWITTER]

clem 🤗

Most companies that shut down give up their impact. This startup did something much better: they open-sourced 1,274 hours of egocentric robotics data. 13,451 recordings of humans doing everyday tasks, as a gift to the robotics community. Thank you @eidon_ai! Open source has a superpower: work can outlive the organization that created it. More startups should do this! https://huggingface.co/eidon-ai

Go to source
[TWITTER]
Sep 22
Alex Kotliarskyi 🇺🇦@alex_frantic

Code reviews today are mostly about gut checking if the complexity is worth the benefits.

[TWITTER]
Sep 22
Alex Kotliarskyi 🇺🇦@alex_frantic

Code reviews today are mostly about gut checking if the complexity is worth the benefits.

Go to source
[TWITTER]

Alex Kotliarskyi 🇺🇦

Code reviews today are mostly about gut checking if the complexity is worth the benefits.

Go to source
[TWITTER]
Sep 21
In reply to
brain function collapse@brainFnCl

This is NOT Jev. Open source. Runs on your laptop. Decides in ~27 ms, about 200× faster than waiting on a hosted LLM. Here it is playing Tetris by itself 👇 https://brainfunctioncollapse.com/laya

Alejandro Fanjul 🛸@alejandrofanjul

Self-hosted an open decision model on my Proxmox box, then ran it head to head against the paid one on the same 520 labelled examples from our invoicing SaaS. ▎Jev won 19 of 20 questions. Laya scored at chance on 6 of them — including the tax treatment question, where being wrong has legal consequences. The most useful finding wasn't the winner. With option keys named R1…R5, Laya scored 15.7% — below chance — and answered R5 to 48 of 51 examples. Renaming them to credito_incobrable and friends took it to 43%. Reorder the options and the answer changes entirely. ▎ The key name isn't metadat…

[TWITTER]
Sep 21
In reply to
brain function collapse@brainFnCl

This is NOT Jev. Open source. Runs on your laptop. Decides in ~27 ms, about 200× faster than waiting on a hosted LLM. Here it is playing Tetris by itself 👇 https://brainfunctioncollapse.com/laya

View parent post
Reply
Alejandro Fanjul 🛸@alejandrofanjul

Self-hosted an open decision model on my Proxmox box, then ran it head to head against the paid one on the same 520 labelled examples from our invoicing SaaS. ▎Jev won 19 of 20 questions. Laya scored at chance on 6 of them — including the tax treatment question, where being wrong has legal consequences. The most useful finding wasn't the winner. With option keys named R1…R5, Laya scored 15.7% — below chance — and answered R5 to 48 of 51 examples. Renaming them to credito_incobrable and friends took it to 43%. Reorder the options and the answer changes entirely. ▎ The key name isn't metadata. The model reads it. ▎ Caveat: labels are synthetic, written by Opus and screened by a rubric. The high scores need real data to confirm. The ones at chance already don't.

Go to source
[TWITTER]

Alejandro Fanjul 🛸

Self-hosted an open decision model on my Proxmox box, then ran it head to head against the paid one on the same 520 labelled examples from our invoicing SaaS. ▎Jev won 19 of 20 questions. Laya scored at chance on 6 of them — including the tax treatment question, where being wrong has legal consequences. The most useful finding wasn't the winner. With option keys named R1…R5, Laya scored 15.7% — below chance — and answered R5 to 48 of 51 examples. Renaming them to credito_incobrable and friends took it to 43%. Reorder the options and the answer changes entirely. ▎ The key name isn't metadat…

Go to source
[OTHERS]
Sep 21

Cloudflare Quick Tunnels

Put localhost on the Internet with a free, encrypted Cloudflare Quick Tunnel. No account, DNS, or open ports required.

try.cloudflare.com
[OTHERS]

Cloudflare Quick Tunnels

Put localhost on the Internet with a free, encrypted Cloudflare Quick Tunnel. No account, DNS, or open ports required.

Go to source
[TWITTER]
Sep 21
In reply to
SpaceXAI@SpaceXAI

Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed.

SpaceXAI@SpaceXAI

Grok 4.7 works longer on difficult tasks, checks its work more carefully, and comes with our strongest safeguards to date.

[TWITTER]
Sep 21
In reply to
SpaceXAI@SpaceXAI

Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed.

View parent post
Reply
SpaceXAI@SpaceXAI

Grok 4.7 works longer on difficult tasks, checks its work more carefully, and comes with our strongest safeguards to date.

Go to source
[TWITTER]

SpaceXAI

Grok 4.7 works longer on difficult tasks, checks its work more carefully, and comes with our strongest safeguards to date.

Go to source
[TWITTER]
Sep 21
Obie Fernandez@obie

You're Already a Meat Proxy

I’m excited about a future that I suspect will be very hard on many long-time friends and certain people that I love. The same technology that is allowing me to be more successful than ever is

[TWITTER]
Sep 21
Obie Fernandez@obie

Your job may be surviving on inertia. Your resentment won’t extend the runway. What do you actually want to do with the intelligence now at your disposal? https://x.com/i/article/2101743989305315329

Go to source
[TWITTER]
Sep 21
Obie Fernandez@obie

You're Already a Meat Proxy

I’m excited about a future that I suspect will be very hard on many long-time friends and certain people that I love. The same technology that is allowing me to be more successful than ever is dismantling an arrangement they depended on to live a comfortable life. I can tell them to use the current runway they have. I can encourage them to take more initiative. I can’t honestly promise that everyone will discover that they enjoy what happens next in their life.

“But I want to use my brain, man!”

Hung out last night with a long time friend who's a staff engineer at a large successful startup with hundreds of engineers.

My friend has been despairing because, in his words, he “doesn't know how to do his job anymore.” From time to time whenever we discuss the subject, I suggest some variation on, "Why don't you just ask Claude to tell you how to do your job?"

His response, in turn, is always some variation on, "No you don't get it."

Last night we were going round and round and so eventually I got fed up and was like, "Dude, just let me see a terminal."

We go over to his work computer. He brings up a few dozen browser tabs and half a dozen terminals, with an instance of Claude running in each one of them. I remark something along the lines of okay so it looks like you're using the thing to do the thing!

"Yeah but you don't understand. I feel like I can't make any progress on anything," he replies.

"But I don't understand how you cannot be making any progress. Let me fire up a new Claude. Do you mind?" And I do. The first thing I do is to switch to Fable because he was on Opus and thus my first piece of concrete advice was to always use Fable whenever you're orchestrating work because it's an order of magnitude smarter than Opus.

"But I don't want to waste my token budget."

Frustrated, I admonish him “No man, it’s not wasting budget. Opus sucks. Just. Use. Fable. Ehrm, you're allowed to use Fable, right?"

"Yeah I think so."

"Great, just use Fable for everything. It's smart and won't let you down," I tell him.

"But I want to use my brain, man!"

This friend is reasonably intelligent. He's also that one friend that despairs over what generative AI is doing to humanity.

"Man, fuck your brain. You're telling me for six months that you don't know how to do your job. Fable is smarter than you. It’s smarter than me, smarter than all of us, I’m not joking."

"It's a machine! I refuse to believe it’s smarter..." he protests.

"Oh it's a stochastic parrot??" I interrupt.

He winces, "No I'm not saying that. I'm saying it has no soul."

We argue a bit, with him telling me that he wants to stay in control and me telling him he has to let that shit go.

He's a bit frustrated, but watches as I try to demonstrate the power of prompting for outcomes instead of fine grained tasks. I type in the main monolith directory:

This code is shit and the architecture is all fucked.

I hit enter and my friend starts laughing. He's in disbelief.

"That's your prompt???"

I'm like, "Well I don't know anything about your system, man, so I got to start somewhere, right?"

We both laugh as Fable does its thing and starts trying to figure out my prompt. Apparently there's a lot of context in this system, so it has a lot to go on. The list of failures includes a lot of missing test coverage, God objects with thousands of lines of code, plenty of obviously dead code, in other words, a lot of issues that constitute real technical debt.

To someone like me, that's a gobsmacking amount of low-hanging fruit.

Trying to make a point, I hit enter to accept Fable's first suggestion. Minutes later I hit enter again, and then again, choosing to delete some dead code. We start making a PR. My friend asks me to make sure it's set to draft. Sure, whatever.

Fable does its thing. My friend checks the diff. It's a simple deletion of dead code and associated unit tests. I want to push on, but my friend begs me to stop.

"You don't understand Obie, I can't just do what you're doing, man."

I challenge him to explain why not.

He explains that he has a boss and teammates and that he can't just make changes like that, he has to present plans and execute on them.

"Great, then have Claude make your plans," I retort.

"I've done that!"

My friend proceeds to click over to Confluence and show me scads of obviously machine-generated plans discussing the architecture changes that need to be made to this system to clean it up.

Fuck's sake, he has plans already! I wonder out loud why he's not just executing on them, and ask him if it's a matter of his boss signing off.

"No, he doesn't read that shit."

WTF?!?

"I have to work with the team. They read it."

I am somewhat incredulous, and press him on that point. "Your team has to read that?"

I don't believe that his team reads all that stuff, not by any stretch of the imagination. If anything, they ask Claude to read it for them.

He continues, "Well, yeah, and then I have to turn it all into epics and stories in Jira."

As an aside, some of you reading this fully understand what the use of Jira and Confluence represents at a company. In a nutshell, it means they are not exactly interested in developer productivity above all else. But in the age of pervasive intelligence on tap, that's not the curse it used to be, at least I don't think that it is anymore.

"Here man, your Claude is connected to Jira via MCP, right?" as I turn back to the terminal and type the following:

we can't push this any of this work yet, let's put it into Jira as epics and stories first

Claude immediately publishes 5 epics to Jira and indicates it's about to start creating associated tickets under them.

My friend freaks out and hits ESC to stop.

Fine. I guess my demonstration is concluded for the night.

I press him on why he doesn't just try to make incremental improvements every day. His answer is because nobody is thinking.

I am puzzled by the non-sequitur.

He goes on to complain about how everyone is just letting Claude do the thinking, like I just did. That he doesn't want to let his brain rot, just sitting there hitting enter all day.

Ah. I'm starting to understand, maybe.

"But then why isn't any progress being made?" I propose.

"I don't know, it's like nobody's thinking about anything anymore."

My dude has a high paying job, with (apparently) low expectations from his management, and is complaining about not being able to work? Even worse, he's depressed about it and anxious about his career prospects. Presumably my friend is at least somewhat smarter and more effective than his lower-ranked colleagues (otherwise he wouldn't be a staff engineer), and yet he's in this state? I care about this dude like a brother. I'm puzzled, and concerned.

I'm like "wtf is wrong with you man, just get shit done, it doesn't matter how you get it done. Literally anything is better than just spinning your wheels everyday."

Doesn't he know how this ends? High paid people that don't do anything of value to a company eventually become not paid people.

"But I don't want to just be a meat proxy!" he protests.

Ha!

My friend is already a meat proxy. He's been a meat proxy before it was a thing. I know it. Pretty sure he knows it too, he just doesn't want to accept it. I shake my head.

"Brother, you need to just hit enter all day and play with your kids while Claude does its thing. Fable is fucking smart, you can trust it."

He's not listening. He thinks my point of view is biased. It is indeed biased, by success. He accuses me of not reading every line of code I produce. He's right, and it's something I'm proud of.

At my job, the systems I primarily work on are greenfield, well-designed, well-tested Ruby on Rails internal systems. I don't need to read every line of code Claude produces. Over the last couple months, I don't even supervise the writing of the code using Claude Code anymore, because I have autonomous agents that write 75-80% of the code. I only get involved in the details when I want to add significant new functionality or make big architectural changes.

My friend is clearly in a different situation. He works on a part of his production system that involves customer payments. So I tell him he should just go ahead and read every line of code that his Claude Code changes. Why not?

After much heated discussion, it's clear that he doesn't trust Claude because he doesn't want to trust it. At some point, he's going to have to put an accurate name on that fear. He's afraid of being obsolete.

The thing is, with that attitude, he's already obsolete. As are probably 80%+ of all the people involved with creating software. Their career is dead in the water, they just haven't realized it yet.

Frustration boiling over

Coincidentally, the following post went viral yesterday, with over 1.5 million views and tons of engagement as I write this. Clearly, it hit a nerve in the industry.

No need to click away, it's a short post so let reproduce a little more past the preview above, for context.

...everything is made by Claude Code. Nobody on my team likes this. They are being forced to ship as much as they can. I have heard multiple times from higher management that pushing code is not a bottleneck, so why are we slow? People are working 12 to 13 hours a day just to press enter. Nobody is reading anything. Humans in corporate are doing nothing on their own. Everyone, literally everyone, from an L1 to an L7 engineer here is doing the same thing. Talk to Claude. There is no sense of victory. Nobody is resolving bugs. In reality, nobody is thinking anymore. Everything is done by LLMs. It is so soul-sucking. I would not mind it, to be honest, if we were at least given the time to check out the code and see what is going where. But no, the goal is to just ship. No matter what happens.

That post and its sentiment struck a nerve with me as well, but not for what I suspect is the typical reason. In a career spanning more than three decades, I was never really been a meat proxy. Even for the first five years, when most of what I remember was impostor syndrome and not knowing what the fuck I was doing, I'm pretty sure I always questioned what I was being told to do and tried to understand things before doing them.

Then around 2000 I became involved in the earliest stages of the eXtreme Programming movement, got promoted to technical leadership at work, and the rest is history. If I'm in a position of taking direction from my superiors, as I often am, the way that I implement those directions is generally at my discretion. That is the vaunted "taste and judgment" that people talk about as being so valuable now.

Most software developers don't have that.

Most software developers never even try to have that, because it was always good enough to just be a consistent team player, follow the process, whatever it might be, and deliver results according to plans made by other people above their pay grade.

Most software developers don't try simply because they don't have it in them, or because it's easier and they'd rather give their surplus energy to some other thing: family, video games, travel, you name it... as long as their boss is happy, they get a relatively big paycheck and everything works out.

Well, it's obvious to me and many others that system is in the process of falling apart. And the aftermath is not going to be pretty. If all you do for your paycheck is to fit into a process where other people decide what to do and then you transform those directions into software, then congrats you're a "meat proxy". If all you do is take direction from higher level bosses, and transform those into instructions for other people, you're also a "meat proxy".

The only reason meat proxies still have jobs is inertia. Current AI technology is already good enough to replace all meat proxies, it hasn't simply because that level of disruptive change takes time.

What do you actually want?

Which brings me back to telling my friend to "hit enter and go play with his kids" in-between prompting Claude Code.

I meant it. My friend isn't going to magically turn into me. I want him to finish something useful every day, keep his job, and have some energy left for his beautiful baby daughters. He's at a large, successful company. I suspect he has some runway to figure things out. A lot of people do. Organizational inertia can keep a paycheck coming long after the economic rationale for a particular job starts to evaporate.

So use that time. For fuck's sake, enjoy some of it. But understand what you're spending.

Because I can't tell you how much time you have. All I can tell you is that arguing with friends about whether a computer-based intelligence has a soul seems like a terrible use of that time.

I'm not going to offer the comforting advice that you should become an architect because AI will always need someone to do the high-level thinking. Architecture has been a big part of my value proposition for more than twenty years, but I fully expect AI to eat more and more of that work too. I expect it to eat literally all of it, until the only humans left are the ones paying for whatever is being built and/or maintained. Moving one box higher on the org chart doesn't get you outside that process.

My long-time friend @chadfowler is writing a book about regenerative software: systems whose implementations can be replaced while preserving the behavior people depend on. Getting there involves recovering knowledge buried in existing software, making its obligations explicit, and establishing how to tell whether a replacement works. Today, that takes considerable engineering judgment and experience, but I see no reason to assume that figuring out the boundaries, extracting the requirements, or designing the checks will remain exclusively human work.

What excites me personally about this historical shift is how much more value I can add as those capabilities become available. Going back to ThoughtWorks and then Hashrocket, clients brought me into situations where figuring out what to do was a substantial part of the assignment. I love that responsibility. I still do. Give me more capability to act on a decision and I immediately start thinking about what more I can accomplish for whoever is paying me.

That is of course a preference, not a guarantee of permanent employment. I don't have an AI-proof certificate hidden in a drawer somewhere. I expect to need to keep evolving what I do and how I do it on a daily basis, just like I've done for 30 years at this point.

I say that with the knowledge that plenty of people find my own preferred relationship to work exhausting. All they want is a reasonably clear assignment, a good paycheck, and a life outside it. I understand the appeal. Sadly, I have serious doubts that the software industry will keep offering that arrangement to everyone who wants it.

So you, dear reader, what do you actually want?

If it's a fat paycheck and plenty of free time with your family, be honest about that. Use AI tools to meet your obligations to a satisfactory level and use the resulting breathing room to figure out what comes next. You don't owe this industry a lifelong love affair. In fact, you might want to consider doing something else entirely.

But if you can muster the interest in wanting more say in what gets built, start exercising that judgment now. Do it at work, if possible. If not, then find some problem worth solving in your free time. Talk to whoever lives with it. Use AI to investigate it, challenge your assumptions, and work out what you could try. Decide what result would make the effort involved worthwhile. Then carry one of those ideas through far enough to discover whether you were right. There's seriously never been a better time than now.

My little demonstration for my friend produced a draft PR and some Jira epics in a matter of 20 minutes. We never got to the point of establishing that anything important had improved. After sleeping on it and working on this essay, I realized something: following through on that question of importance is precisely the work I wanted my friend to stop retreating from. If the obstacle is getting five people to agree to do something, then getting them to agree is an essential part of the job. Claude can help you prepare, but you still have to engage with the people whose cooperation you need, right?

To my friend who I will be sending this to, I say pick something this week. Something small enough to finish and consequential enough that another person will notice. Let the machine do as much of the work as it can. Pay attention to what happens. If the result is wrong, find out why and keep going. You don't become less of a meat proxy by personally typing more of the code.

I can't promise it will save your career as a software engineer. But it gives you a chance to find out what you can do with capabilities you didn't have before, while you still have time to experiment.

For those of you reading this and saying "no, no, no!" I only have one response. Standing in front of this AI tidal wave and having a temper tantrum about it isn't going to keep you from being swept away.

You really want to keep using your brain? You don't want it to rot? Figure out what you actually want, and put Claude to work on it.

Read on X
[TWITTER]

You're Already a Meat Proxy

Your job may be surviving on inertia. Your resentment won’t extend the runway. What do you actually want to do with the intelligence now at your disposal? https://x.com/i/article/2101743989305315329

Go to source
[TWITTER]
Sep 21
lauren@poteto

here's how i shipped 2,500 PRs last month to production this was originally supposed to be for Cursor Compile in London. i couldn't make it since i was livestreaming for Grok @Bot Galaxy so i'm making it available for free here on X! watch it on 2x speed, i talk slowly

[TWITTER]
Sep 21
lauren@poteto

here's how i shipped 2,500 PRs last month to production this was originally supposed to be for Cursor Compile in London. i couldn't make it since i was livestreaming for Grok @Bot Galaxy so i'm making it available for free here on X! watch it on 2x speed, i talk slowly

Go to source
[TWITTER]

lauren

here's how i shipped 2,500 PRs last month to production this was originally supposed to be for Cursor Compile in London. i couldn't make it since i was livestreaming for Grok @Bot Galaxy so i'm making it available for free here on X! watch it on 2x speed, i talk slowly

Go to source
[TWITTER]
Sep 21
Guillermo Rauch@rauchg

Looks like today may be a record day for token volume % of open models on Vercel AI Gateway: 🟦 Open 78.4% 🟨 Closed 21.6% While spend 💲 usually tells a different story, #3 and #4 today are Moonshot AI & DeepSeek. Adding Z⁠.ai, their combined spend surpasses OpenAI (#2). (Do note that's the spend for inference of the model across providers (mostly in the US), not revenue going directly to the open weight labs.)

[TWITTER]
Sep 21
Guillermo Rauch@rauchg

Looks like today may be a record day for token volume % of open models on Vercel AI Gateway: 🟦 Open 78.4% 🟨 Closed 21.6% While spend 💲 usually tells a different story, #3 and #4 today are Moonshot AI & DeepSeek. Adding Z⁠.ai, their combined spend surpasses OpenAI (#2). (Do note that's the spend for inference of the model across providers (mostly in the US), not revenue going directly to the open weight labs.)

Go to source
[TWITTER]

Guillermo Rauch

Looks like today may be a record day for token volume % of open models on Vercel AI Gateway: 🟦 Open 78.4% 🟨 Closed 21.6% While spend 💲 usually tells a different story, #3 and #4 today are Moonshot AI & DeepSeek. Adding Z⁠.ai, their combined spend surpasses OpenAI (#2). (Do note that's the spend for inference of the model across providers (mostly in the US), not revenue going directly to the open weight labs.)

Go to source
[TWITTER]
Sep 21
JSFILMZ@JSFILMZ0412

Why Hollywood is using Seedance 2.5 and Seedance 2.0 Part 3. #DreaminaCPP #DreaminaAi #Seedance25 @dreamina_ai

JSFILMZ@JSFILMZ0412

This is exactly why Hollywood didn’t ban Seedance 2.0 and Seedance 2.5. They’re using it instead. #DreaminaCPP #DreaminaAi

[TWITTER]
Sep 21
JSFILMZ@JSFILMZ0412

Why Hollywood is using Seedance 2.5 and Seedance 2.0 Part 3. #DreaminaCPP #DreaminaAi #Seedance25 @dreamina_ai

JSFILMZ@JSFILMZ0412

This is exactly why Hollywood didn’t ban Seedance 2.0 and Seedance 2.5. They’re using it instead. #DreaminaCPP #DreaminaAi

Go to source
[TWITTER]

JSFILMZ

Why Hollywood is using Seedance 2.5 and Seedance 2.0 Part 3. #DreaminaCPP #DreaminaAi #Seedance25 @dreamina_ai

Go to source
[TWITTER]
Sep 21
In reply to
Armin Ronacher ⇌@mitsuhiko

What when it comes to AI in software engineering are you struggling with the most?

iury souza@IurySza

Multi tasking. Flow state is kinda gone and it seems like the best I can get now feels more like playing a late game RTS match where you need to make sure you dont get idle villagers while you're microing other 5 groups of units in different places of the map. An RTS match only lasts 30min though.

[TWITTER]
Sep 21
In reply to
Armin Ronacher ⇌@mitsuhiko

What when it comes to AI in software engineering are you struggling with the most?

View parent post
Reply
iury souza@IurySza

Multi tasking. Flow state is kinda gone and it seems like the best I can get now feels more like playing a late game RTS match where you need to make sure you dont get idle villagers while you're microing other 5 groups of units in different places of the map. An RTS match only lasts 30min though.

Go to source
[TWITTER]

iury souza

Multi tasking. Flow state is kinda gone and it seems like the best I can get now feels more like playing a late game RTS match where you need to make sure you dont get idle villagers while you're microing other 5 groups of units in different places of the map. An RTS match only lasts 30min though.

Go to source
[TWITTER]
Sep 21
Kaushik Gopal@kaushikgopal

Self-learning harnesses are here. After @deepseek_ai released their new harness, even the authors of Pi & OpenCode said there were some super compelling ideas in there. But what exactly? @IurySza and I take a shot at explaining it in @FragmentedCast

[TWITTER]
Sep 21
Kaushik Gopal@kaushikgopal

Self-learning harnesses are here. After @deepseek_ai released their new harness, even the authors of Pi & OpenCode said there were some super compelling ideas in there. But what exactly? @IurySza and I take a shot at explaining it in @FragmentedCast

Go to source
[TWITTER]

Kaushik Gopal

Self-learning harnesses are here. After @deepseek_ai released their new harness, even the authors of Pi & OpenCode said there were some super compelling ideas in there. But what exactly? @IurySza and I take a shot at explaining it in @FragmentedCast

Go to source
[TWITTER]
Sep 21
Armin Ronacher ⇌@mitsuhiko

What when it comes to AI in software engineering are you struggling with the most?

[TWITTER]
Sep 21
Armin Ronacher ⇌@mitsuhiko

What when it comes to AI in software engineering are you struggling with the most?

Go to source
[TWITTER]

Armin Ronacher ⇌

What when it comes to AI in software engineering are you struggling with the most?

Go to source
[TWITTER]
Sep 21
iury souza@IurySza

First experiment with @typesafeai's Jev A semantic log filter. (with a CLI that lets my agent use it too) Ofc, I had to overdoit and build a TUI around it :)

[TWITTER]
Sep 21
iury souza@IurySza

First experiment with @typesafeai's Jev A semantic log filter. (with a CLI that lets my agent use it too) Ofc, I had to overdoit and build a TUI around it :)

Go to source
[TWITTER]

iury souza

First experiment with @typesafeai's Jev A semantic log filter. (with a CLI that lets my agent use it too) Ofc, I had to overdoit and build a TUI around it :)

Go to source
[TWITTER]
Sep 21
Shane Parrish@shaneparrish

Lazy work used to mean too little output. Now, with AI, it often means too much and more work for everyone else. @tobi call it "slop grenades." A "Slop Grenade" is when you let AI produce the work and pass it on without adding any value (including checking it). Someone else has to wade through it, catch the mistakes, and clean up the mess. You save time and look productive but someone else pays for it.

Shane Parrish@shaneparrish

My third conversation with Shopify co-founder and CEO @tobi. 0:00 How Shopify Uses AI 7:18 River: Shopify's Internal AI 8:55 How to Encourage Osmosis Learning 10:52 AI Dreaming and Self-Reflection 11:53 How to Use AI for Strategic Decision Making 14:11 The One Thing AI Cannot Do 16:04 What AI is Making Worse at Shopify 19:46 Predictions: Where AI is Headed Next 21:55 The Future of AI-Powered Software 24:40 Will CEOs Be Replaced with AI? 27:54 Can Superintelligence Be Controlled? 31:13 Critical Skills in AI Age 34:22 Why Complex Solutions are Usually Wrong 36:33 Conditions Needed for True Intuition 38:02 The Best Path Doesn't Have Instant Feedback 44:51 How Affirmations Can Shift Your Behavior 50:44 The Inobvious Thing Hurting Companies 52:37 Relationship Between Beauty and Creation 56:23 How SpaceX Moves Forward By Subtraction 1:00:24 Why Companies Need Refounding Events 1:01:50 Books as Cheat Codes 1:02:39 Three Books to Change Your Thinking Enjoy! (Includes paid promotions.)

[TWITTER]
Sep 21
Shane Parrish@shaneparrish

Lazy work used to mean too little output. Now, with AI, it often means too much and more work for everyone else. @tobi call it "slop grenades." A "Slop Grenade" is when you let AI produce the work and pass it on without adding any value (including checking it). Someone else has to wade through it, catch the mistakes, and clean up the mess. You save time and look productive but someone else pays for it.

Shane Parrish@shaneparrish

My third conversation with Shopify co-founder and CEO @tobi. 0:00 How Shopify Uses AI 7:18 River: Shopify's Internal AI 8:55 How to Encourage Osmosis Learning 10:52 AI Dreaming and Self-Reflection 11:53 How to Use AI for Strategic Decision Making 14:11 The One Thing AI Cannot Do 16:04 What AI is Making Worse at Shopify 19:46 Predictions: Where AI is Headed Next 21:55 The Future of AI-Powered Software 24:40 Will CEOs Be Replaced with AI? 27:54 Can Superintelligence Be Controlled? 31:13 Critical Skills in AI Age 34:22 Why Complex Solutions are Usually Wrong 36:33 Conditions Needed for True Intuition 38:02 The Best Path Doesn't Have Instant Feedback 44:51 How Affirmations Can Shift Your Behavior 50:44 The Inobvious Thing Hurting Companies 52:37 Relationship Between Beauty and Creation 56:23 How SpaceX Moves Forward By Subtraction 1:00:24 Why Companies Need Refounding Events 1:01:50 Books as Cheat Codes 1:02:39 Three Books to Change Your Thinking Enjoy! (Includes paid promotions.)

Go to source
[TWITTER]

Shane Parrish

Lazy work used to mean too little output. Now, with AI, it often means too much and more work for everyone else. @tobi call it "slop grenades." A "Slop Grenade" is when you let AI produce the work and pass it on without adding any value (including checking it). Someone else has to wade through it, catch the mistakes, and clean up the mess. You save time and look productive but someone else pays for it.

Go to source
[TWITTER]
Sep 21
Thorsten Ball@thorstenball

Most predictions I see are still way too conservative. Here's mine

[TWITTER]
Sep 21
Thorsten Ball@thorstenball

Most predictions I see are still way too conservative. Here's mine

Go to source
[TWITTER]

Thorsten Ball

Most predictions I see are still way too conservative. Here's mine

Go to source
[TWITTER]
Sep 21
Robin Bilgil@RBilgil

Made a real-time slop detector with jev as you scroll

[TWITTER]
Sep 21
Robin Bilgil@RBilgil

Made a real-time slop detector with jev as you scroll

Go to source
[TWITTER]

Robin Bilgil

Made a real-time slop detector with jev as you scroll

Go to source
[TWITTER]
Sep 21
tobi lutke@tobi

MCP vs CLI for LLMs is the wrong discussion. People are debating at the wrong layer. Both work incredibly well as long as they are run through a repl like environment that can persist state. The funny thing about CLI is that CLI tools are usually accessed through BASH which happens to be a repl with persistent state (the file system), therefore cli works much better than mcp. But this isn't even close to being an intrinsic property of MCP. Just means that we need better harnesses. Right now the best repl for this are: - bash + fs - jupyter kernels - codemode type repls, usually quickjs LLM…

[TWITTER]
Sep 21
tobi lutke@tobi

MCP vs CLI for LLMs is the wrong discussion. People are debating at the wrong layer. Both work incredibly well as long as they are run through a repl like environment that can persist state. The funny thing about CLI is that CLI tools are usually accessed through BASH which happens to be a repl with persistent state (the file system), therefore cli works much better than mcp. But this isn't even close to being an intrinsic property of MCP. Just means that we need better harnesses. Right now the best repl for this are: - bash + fs - jupyter kernels - codemode type repls, usually quickjs LLMs understand the concept of forward evolving a system to solve a need very well. This comes from the agentic RL. They understand how to change the state of a codebase or acquire data from databases and then operate on it similarly to humans. But without an execution environment, they actually cannot do this properly. My bet? Sometime soon someone will (or has already?) create a embeddable, sqlite style mini execution environment that parses bash, typescript, or tool calls into a common IL execution plan that's easy to security check before executing. And design this specifically for durable execution environments. Then we will just connect cli, mcp, webmcp, whatever to that and it accepts any of the input modalities as they can all be represented as each other.

Go to source
[TWITTER]

tobi lutke

MCP vs CLI for LLMs is the wrong discussion. People are debating at the wrong layer. Both work incredibly well as long as they are run through a repl like environment that can persist state. The funny thing about CLI is that CLI tools are usually accessed through BASH which happens to be a repl with persistent state (the file system), therefore cli works much better than mcp. But this isn't even close to being an intrinsic property of MCP. Just means that we need better harnesses. Right now the best repl for this are: - bash + fs - jupyter kernels - codemode type repls, usually quickjs LLM…

Go to source
[TWITTER]
Sep 21
Nico Bailon@nicopreme

pi-subagents now has code mode. 🥳 Orchestrate subagents with plain JavaScript: loop, fan out, await, branch on real child output. Mix parallel and sequential phases, and isolate every child in its own git worktree, all in one script. https://github.com/nicobailon/pi-subagents pi install npm:pi-subagents

[TWITTER]
Sep 21
Nico Bailon@nicopreme

pi-subagents now has code mode. 🥳 Orchestrate subagents with plain JavaScript: loop, fan out, await, branch on real child output. Mix parallel and sequential phases, and isolate every child in its own git worktree, all in one script. https://github.com/nicobailon/pi-subagents pi install npm:pi-subagents

Go to source
[TWITTER]

Nico Bailon

pi-subagents now has code mode. 🥳 Orchestrate subagents with plain JavaScript: loop, fan out, await, branch on real child output. Mix parallel and sequential phases, and isolate every child in its own git worktree, all in one script. https://github.com/nicobailon/pi-subagents pi install npm:pi-subagents

Go to source
[TWITTER]
Sep 21
iury souza@IurySza

Summoning People of @pidotdev! Been thinking about building this for a while and now it's here! Powered by the new @ChatGPT Live API (released last week) You talk to one skipper. That person delegates to the other agents, and you can keep up with the threads. For brainstorming, bouncing ideas, getting a status update while you walk around, it is just a nice interface. You dont even have to sit at the computer. You can interrupt and brainstorm live. Voice mode on its own has always super limited. But this last generation is actually amazing! These models are still not super smart, and they…

[TWITTER]
Sep 21
iury souza@IurySza

Summoning People of @pidotdev! Been thinking about building this for a while and now it's here! Powered by the new @ChatGPT Live API (released last week) You talk to one skipper. That person delegates to the other agents, and you can keep up with the threads. For brainstorming, bouncing ideas, getting a status update while you walk around, it is just a nice interface. You dont even have to sit at the computer. You can interrupt and brainstorm live. Voice mode on its own has always super limited. But this last generation is actually amazing! These models are still not super smart, and they cannot be. But when you pair it with an agent that can drive the specialists in the codebase, then it can really get useful This is just an experiment I built it this evening. A couple of hours to get it somewhat right. My idea is to have a voice mode skipper agent which is nice to talk to and than can efficiently delegate and give me status updates of whatever is going on.

Go to source
[TWITTER]

iury souza

Summoning People of @pidotdev! Been thinking about building this for a while and now it's here! Powered by the new @ChatGPT Live API (released last week) You talk to one skipper. That person delegates to the other agents, and you can keep up with the threads. For brainstorming, bouncing ideas, getting a status update while you walk around, it is just a nice interface. You dont even have to sit at the computer. You can interrupt and brainstorm live. Voice mode on its own has always super limited. But this last generation is actually amazing! These models are still not super smart, and they…

Go to source
[TWITTER]
Sep 21
OpenClaw🦞@openclaw

OpenClaw 2.0 has arrived https://openclaw.ai/blog/openclaw-2-accidentally

[TWITTER]
Sep 21
OpenClaw🦞@openclaw

OpenClaw 2.0 has arrived https://openclaw.ai/blog/openclaw-2-accidentally

Go to source
[TWITTER]

OpenClaw🦞

OpenClaw 2.0 has arrived https://openclaw.ai/blog/openclaw-2-accidentally

Go to source
[TWITTER]
Sep 21
The Developers' Bakery@thebakerydev

Episode 101 is live! 🎙️ We're diving into Coding Agents with mobile dev & platform engineer @IurySza. We explore agentic dev setups, shaping context with AGENTS.md, and the reality of running local models. Listen to it here: https://thebakery.dev/101/

[TWITTER]
Sep 21
The Developers' Bakery@thebakerydev

Episode 101 is live! 🎙️ We're diving into Coding Agents with mobile dev & platform engineer @IurySza. We explore agentic dev setups, shaping context with AGENTS.md, and the reality of running local models. Listen to it here: https://thebakery.dev/101/

Go to source
[TWITTER]

The Developers' Bakery

Episode 101 is live! 🎙️ We're diving into Coding Agents with mobile dev & platform engineer @IurySza. We explore agentic dev setups, shaping context with AGENTS.md, and the reality of running local models. Listen to it here: https://thebakery.dev/101/

Go to source
[OTHERS]
Sep 21

Proposed KDE LLM guidelines (#187) · Issues · Plasma / Plasma Workspace

This is a continuation of this mailing list thread about "fully LLM-generated merge requests, preserved at https://mail.kde.org/pipermail/kde-devel/2026-September/004496.html There, I proposed...

invent.kde.org
[OTHERS]

Proposed KDE LLM guidelines (#187) · Issues · Plasma / Plasma Workspace

This is a continuation of this mailing list thread about "fully LLM-generated merge requests, preserved at https://mail.kde.org/pipermail/kde-devel/2026-September/004496.html There, I proposed...

Go to source
[TWITTER]
Sep 21
The Pragmatic Engineer@Pragmatic_Eng

From Tibo @thsottiaux, “It just goes out to a billion users and it's fine”. How OpenAI engineers can still ship the same day: "Even though ChatGPT goes out to a billion active users, you can ship a PR, you can make a change and get it shipped the next day or even the same day. And it just goes out to a billion users and it's fine. We just really instill a sense of ownership and care. So people are very empowered to make changes, even large changes. The general thing that is being asked is evidence that it's going to be well received, evidence that it's a worthy addition, evidence that it is…

[TWITTER]
Sep 21
The Pragmatic Engineer@Pragmatic_Eng

From Tibo @thsottiaux, “It just goes out to a billion users and it's fine”. How OpenAI engineers can still ship the same day: "Even though ChatGPT goes out to a billion active users, you can ship a PR, you can make a change and get it shipped the next day or even the same day. And it just goes out to a billion users and it's fine. We just really instill a sense of ownership and care. So people are very empowered to make changes, even large changes. The general thing that is being asked is evidence that it's going to be well received, evidence that it's a worthy addition, evidence that it is worth maintaining over time. The cost of maintenance has gone down significantly as well. So we think about these things slightly differently than say two or three years ago. And then the other thing is we automate as much as possible. So a lot of the process of code review and deploys and catching regressions, all of that is pretty much automated. You get to focus on just really the idea and how it's going to help our users."

Go to source
[TWITTER]

The Pragmatic Engineer

From Tibo @thsottiaux, “It just goes out to a billion users and it's fine”. How OpenAI engineers can still ship the same day: "Even though ChatGPT goes out to a billion active users, you can ship a PR, you can make a change and get it shipped the next day or even the same day. And it just goes out to a billion users and it's fine. We just really instill a sense of ownership and care. So people are very empowered to make changes, even large changes. The general thing that is being asked is evidence that it's going to be well received, evidence that it's a worthy addition, evidence that it is…

Go to source
[OTHERS]
Sep 21

Tyler Cowen: A Doomsday Scenario for American AI

The U.S. is slouching toward heavy regulation while China is poised to rush ahead, writes Tyler Cowen. Beijing will then dominate everything from global arms sales to healthcare.

thefp.comTyler Cowen
[OTHERS]

Tyler Cowen: A Doomsday Scenario for American AI

The U.S. is slouching toward heavy regulation while China is poised to rush ahead, writes Tyler Cowen. Beijing will then dominate everything from global arms sales to healthcare.

Go to source
[TWITTER]
Sep 21
voxium@v0xium

I am done with this shit. It is over. The state of engineering right now is horrible. It has been half a month since I started a new role at a big company. Nobody knows anything here. The specs, code, tests, PRDs, tickets, resolution of those tickets, reports, etc., everything is made by Claude Code. Nobody on my team likes this. They are being forced to ship as much as they can. I have heard multiple times from higher management that pushing code is not a bottleneck, so why are we slow? People are working 12 to 13 hours a day just to press enter. Nobody is reading anything. Humans in corporat…

[TWITTER]
Sep 21
voxium@v0xium

I am done with this shit. It is over. The state of engineering right now is horrible. It has been half a month since I started a new role at a big company. Nobody knows anything here. The specs, code, tests, PRDs, tickets, resolution of those tickets, reports, etc., everything is made by Claude Code. Nobody on my team likes this. They are being forced to ship as much as they can. I have heard multiple times from higher management that pushing code is not a bottleneck, so why are we slow? People are working 12 to 13 hours a day just to press enter. Nobody is reading anything. Humans in corporate are doing nothing on their own. Everyone, literally everyone, from an L1 to an L7 engineer here is doing the same thing. Talk to Claude. There is no sense of victory. Nobody is resolving bugs. In reality, nobody is thinking anymore. Everything is done by LLMs. It is so soul-sucking. I would not mind it, to be honest, if we were at least given the time to check out the code and see what is going where. But no, the goal is to just ship. No matter what happens.

Go to source
[TWITTER]

voxium

I am done with this shit. It is over. The state of engineering right now is horrible. It has been half a month since I started a new role at a big company. Nobody knows anything here. The specs, code, tests, PRDs, tickets, resolution of those tickets, reports, etc., everything is made by Claude Code. Nobody on my team likes this. They are being forced to ship as much as they can. I have heard multiple times from higher management that pushing code is not a bottleneck, so why are we slow? People are working 12 to 13 hours a day just to press enter. Nobody is reading anything. Humans in corporat…

Go to source
[TWITTER]
Sep 21
Reid Jackson@reidjjackson

*taps the sign*

staysaasy@staysaasy

The problem with all of these AI personal assistant things is that people don’t actually want to do stuff. You fucking get that right? Do you understand that? Just like AI tools in the workforce, AI assistants don’t actually abstract decisions for you, they let you accelerate decision making. So instead of making one decision a day and then lamenting you can’t do more, AI assistants are going to let you make ten decisions a day and act on all of them. A miracle? Nah dog, a fucking nightmare. Nobody except type A strivers who need Twitter fodder actually want that. Let me tell you a story. When I got back from college I was inspired. If I could take on an elite university with grit and determination, so could everyone I know. I was all razzed up to get people I know to achieve their potential. What I learned after many years of trying is that absolutely nobody wants to realize their potential. People want to hang out with friends and complain. That’s how the world has worked for ten thousand years. This weather fucks. Can you believe the cows walked off. People don’t want enablement. People don’t want you to take all of their excuses. TV is mindless passive activity. AI assistants are not. They’re telling people that they can achieve more. And people absolutely do not want that. Let me tell you another story. Once upon a time I lead a big process overhaul at work and it was great. Two years later I hired a contractor to do something similar. But I personally didn’t want to have to do a bunch of work. Within a week I knew I fucked in. This contractor kept coming to me with decisions I needed to make. Dude I wanted to not have to think about this. Again, a helper when you don’t actually want help is a disaster. None of these AI assistants will take off because people don’t actually want assistants.

[TWITTER]
Sep 21
Reid Jackson@reidjjackson

*taps the sign*

staysaasy@staysaasy

The problem with all of these AI personal assistant things is that people don’t actually want to do stuff. You fucking get that right? Do you understand that? Just like AI tools in the workforce, AI assistants don’t actually abstract decisions for you, they let you accelerate decision making. So instead of making one decision a day and then lamenting you can’t do more, AI assistants are going to let you make ten decisions a day and act on all of them. A miracle? Nah dog, a fucking nightmare. Nobody except type A strivers who need Twitter fodder actually want that. Let me tell you a story. When I got back from college I was inspired. If I could take on an elite university with grit and determination, so could everyone I know. I was all razzed up to get people I know to achieve their potential. What I learned after many years of trying is that absolutely nobody wants to realize their potential. People want to hang out with friends and complain. That’s how the world has worked for ten thousand years. This weather fucks. Can you believe the cows walked off. People don’t want enablement. People don’t want you to take all of their excuses. TV is mindless passive activity. AI assistants are not. They’re telling people that they can achieve more. And people absolutely do not want that. Let me tell you another story. Once upon a time I lead a big process overhaul at work and it was great. Two years later I hired a contractor to do something similar. But I personally didn’t want to have to do a bunch of work. Within a week I knew I fucked in. This contractor kept coming to me with decisions I needed to make. Dude I wanted to not have to think about this. Again, a helper when you don’t actually want help is a disaster. None of these AI assistants will take off because people don’t actually want assistants.

Go to source
[TWITTER]

Reid Jackson

*taps the sign*

Go to source
[OTHERS]
Sep 21

Hugging Face

We’re on a journey to advance and democratize artificial intelligence through open source and open science.

huggingface.co
[OTHERS]

Hugging Face

We’re on a journey to advance and democratize artificial intelligence through open source and open science.

Go to source

One more useful thing

If this feed helped, you will like the weekly digest.

More context on what I found, and better takeaways.

No ads. No bullshit. Unsubscribe anytime.

Get weekly picks