HeadlinesBriefing HeadlinesBriefing 12 languages

AI Made Me 5x Faster. It Also Made Me 5x Worse at My Job.

Towards Data Science ·

🇬🇧 English

There is a moment happening in offices and bedrooms all over the world right now, and it looks like this. Three AI agent sessions are running. One is refactoring something.

One is writing tests. One is halfway through a migration nobody wanted to do by hand. The person in front of them is not typing.

They are watching. Their eyes move between panes like someone who put chips on three tables and cannot decide which one to be nervous about. Then the thought arrives.

I could start a fourth one. I call it the fourth terminal, and I think it is the defining mistake of the AI coding era. Not because running agents in parallel is bad.

Because of what the reflex reveals. AI handed us spare capacity, and our first instinct was to fill it with more AI, instead of asking what the spare capacity was actually for. I made that mistake for about four months.

Nearly everyone I know made it too. This is what it cost, what the research now says about why it happens, and what the people who came out the other side are doing instead. The AI honeymoon is real, and you should enjoy it Let me be fair to the tools first, because the backlash has gotten lazy.

When agentic coding properly landed in my workflow, it felt like someone lifted a weight off my chest I had not known I was carrying. All that configuration written by hand. All that run, squint, fix the typo, run again.

Suddenly optional. Work that used to take a day was done before lunch. A migration I had avoided for a quarter got drafted in an afternoon.

My manager noticed. My team noticed. This is not vibes.

In a controlled study of developers building a simple HTTP server, the ones with an AI assistant finished noticeably faster. In a field experiment across thousands of developers, merged pull requests rose by roughly a quarter. If your work involves a lot of greenfield code or a lot of boilerplate, the AI speedup is real and it is not small.

So we did the obvious thing. We got faster, so we took on more. Bug report that smells like infrastructure? I am on it.

Someone needs a dashboard by Friday? Sure. Ticket from March rotting in the backlog? Why not. My open pull request count started to look like a typo.

The bill AI quietly runs up Here is what nobody tells you about being five times faster. You can also get lost five times faster. The first sign was easy to ignore.

A colleague asked about one of my open PRs, and I had to read my own description to remember what it was for. The description had been written by AI. I was reading a machine's summary of a decision I had apparently made, in order to find out what I thought.

That is not productivity. That is a queue with your name on it. The second sign was not small.

An agent produced a change adding a new permission set for a service. Clean diff. Sensible naming.

Tests green. I reviewed it the way I had started reviewing everything by then, which is to say I scrolled, nodded, approved. It had been open six days and I wanted it gone.

Two things saved me. A teammate who actually reads policy documents left one comment: "is this wildcard on purpose?" And luck, in that the comment landed before the merge did. The AI had done exactly what I asked.

I had asked for the wrong thing, vaguely, and it filled the gap with the most permissive option available. The tests passed because the tests checked that the permission existed, not that it was safe. That is the sentence that now governs my working day:"All tests pass" is not the same as "this does what I meant.".

View original article →


🇸🇦 العربية

الذكاء الاصطناعي جعلني أسرع 5 مرات، وأسوأ 5 مرات في العمل

There is a moment happening in offices and bedrooms all over the world right now, and it looks like this. Three AI agent sessions are running. One is refactoring something.

One is writing tests. One is halfway through a migration nobody wanted to do by hand. The person in front of them is not typing.

They are watching. Their eyes move between panes like someone who put chips on three tables and cannot decide which one to be nervous about. Then the thought arrives.

I could start a fourth one. I call it the fourth terminal, and I think it is the defining mistake of the AI coding era. Not because running agents in parallel is bad.

Because of what the reflex reveals. AI handed us spare capacity, and our first instinct was to fill it with more AI, instead of asking what the spare capacity was actually for. I made that mistake for about four months.

Nearly everyone I know made it too. This is what it cost, what the research now says about why it happens, and what the people who came out the other side are doing instead. The AI honeymoon is real, and you should enjoy it Let me be fair to the tools first, because the backlash has gotten lazy.

When agentic coding properly landed in my workflow, it felt like someone lifted a weight off my chest I had not known I was carrying. All that configuration written by hand. All that run, squint, fix the typo, run again.

Suddenly optional. Work that used to take a day was done before lunch. A migration I had avoided for a quarter got drafted in an afternoon.

My manager noticed. My team noticed. This is not vibes.

In a controlled study of developers building a simple HTTP server, the ones with an AI assistant finished noticeably faster. In a field experiment across thousands of developers, merged pull requests rose by roughly a quarter. If your work involves a lot of greenfield code or a lot of boilerplate, the AI speedup is real and it is not small.

So we did the obvious thing. We got faster, so we took on more. Bug report that smells like infrastructure? I am on it.

Someone needs a dashboard by Friday? Sure. Ticket from March rotting in the backlog? Why not. My open pull request count started to look like a typo.

The bill AI quietly runs up Here is what nobody tells you about being five times faster. You can also get lost five times faster. The first sign was easy to ignore.

A colleague asked about one of my open PRs, and I had to read my own description to remember what it was for. The description had been written by AI. I was reading a machine's summary of a decision I had apparently made, in order to find out what I thought.

That is not productivity. That is a queue with your name on it. The second sign was not small.

An agent produced a change adding a new permission set for a service. Clean diff. Sensible naming.

Tests green. I reviewed it the way I had started reviewing everything by then, which is to say I scrolled, nodded, approved. It had been open six days and I wanted it gone.

Two things saved me. A teammate who actually reads policy documents left one comment: "is this wildcard on purpose?" And luck, in that the comment landed before the merge did. The AI had done exactly what I asked.

I had asked for the wrong thing, vaguely, and it filled the gap with the most permissive option available. The tests passed because the tests checked that the permission existed, not that it was safe. That is the sentence that now governs my working day:"All tests pass" is not the same as "this does what I meant.".

ما هي المشكلة الأساسية في الاعتماد المفرط على وكلاء ترميز الذكاء الاصطناعي؟

المشكلة الأساسية هي أن الذكاء الاصطناعي يمكنه إنتاج رمز يجتاز الاختبارات لكنه لا يعكس نية المستخدم الفعلية، مما يؤدي إلى شعور خاطئ بالإنتاجية ومخاطر أمان محتملة عندما يتوقف المستخدمون عن فهم عملهم الخاص.

العربية version →


🇧🇩 বাংলা

AI আমাকে কাজে 5 গুণ দ্রুত, 5 গুণ খারাপ করে দিয়েছে

There is a moment happening in offices and bedrooms all over the world right now, and it looks like this. Three AI agent sessions are running. One is refactoring something.

One is writing tests. One is halfway through a migration nobody wanted to do by hand. The person in front of them is not typing.

They are watching. Their eyes move between panes like someone who put chips on three tables and cannot decide which one to be nervous about. Then the thought arrives.

I could start a fourth one. I call it the fourth terminal, and I think it is the defining mistake of the AI coding era. Not because running agents in parallel is bad.

Because of what the reflex reveals. AI handed us spare capacity, and our first instinct was to fill it with more AI, instead of asking what the spare capacity was actually for. I made that mistake for about four months.

Nearly everyone I know made it too. This is what it cost, what the research now says about why it happens, and what the people who came out the other side are doing instead. The AI honeymoon is real, and you should enjoy it Let me be fair to the tools first, because the backlash has gotten lazy.

When agentic coding properly landed in my workflow, it felt like someone lifted a weight off my chest I had not known I was carrying. All that configuration written by hand. All that run, squint, fix the typo, run again.

Suddenly optional. Work that used to take a day was done before lunch. A migration I had avoided for a quarter got drafted in an afternoon.

My manager noticed. My team noticed. This is not vibes.

In a controlled study of developers building a simple HTTP server, the ones with an AI assistant finished noticeably faster. In a field experiment across thousands of developers, merged pull requests rose by roughly a quarter. If your work involves a lot of greenfield code or a lot of boilerplate, the AI speedup is real and it is not small.

So we did the obvious thing. We got faster, so we took on more. Bug report that smells like infrastructure? I am on it.

Someone needs a dashboard by Friday? Sure. Ticket from March rotting in the backlog? Why not. My open pull request count started to look like a typo.

The bill AI quietly runs up Here is what nobody tells you about being five times faster. You can also get lost five times faster. The first sign was easy to ignore.

A colleague asked about one of my open PRs, and I had to read my own description to remember what it was for. The description had been written by AI. I was reading a machine's summary of a decision I had apparently made, in order to find out what I thought.

That is not productivity. That is a queue with your name on it. The second sign was not small.

An agent produced a change adding a new permission set for a service. Clean diff. Sensible naming.

Tests green. I reviewed it the way I had started reviewing everything by then, which is to say I scrolled, nodded, approved. It had been open six days and I wanted it gone.

Two things saved me. A teammate who actually reads policy documents left one comment: "is this wildcard on purpose?" And luck, in that the comment landed before the merge did. The AI had done exactly what I asked.

I had asked for the wrong thing, vaguely, and it filled the gap with the most permissive option available. The tests passed because the tests checked that the permission existed, not that it was safe. That is the sentence that now governs my working day:"All tests pass" is not the same as "this does what I meant.".

AI কোডিং এজেন্টে অতিরিক্ত নির্ভরশীলতার মূল সমস্যা কী?

মূল সমস্যা হল যে AI পরীক্ষা পাস করতে পারে কোড তৈরি করতে পারে কিন্তু ব্যবহারকারীর প্রকৃত intenciónকে প্রতিফলিত করে না, যা কুতूहলপ্রদ উত্পাদনশীলতার ভ্রম এবং ব্যবহারকারীর নিজের কাজ বুঝতে বন্ধ দিলে সম্ভাব্য নিরাপত্তা जोখিম তৈরি করে।

বাংলা version →


🇩🇪 Deutsch

KI hat mich bei der Arbeit 5x schneller gemacht, aber auch 5x schlechter

There is a moment happening in offices and bedrooms all over the world right now, and it looks like this. Three AI agent sessions are running. One is refactoring something.

One is writing tests. One is halfway through a migration nobody wanted to do by hand. The person in front of them is not typing.

They are watching. Their eyes move between panes like someone who put chips on three tables and cannot decide which one to be nervous about. Then the thought arrives.

I could start a fourth one. I call it the fourth terminal, and I think it is the defining mistake of the AI coding era. Not because running agents in parallel is bad.

Because of what the reflex reveals. AI handed us spare capacity, and our first instinct was to fill it with more AI, instead of asking what the spare capacity was actually for. I made that mistake for about four months.

Nearly everyone I know made it too. This is what it cost, what the research now says about why it happens, and what the people who came out the other side are doing instead. The AI honeymoon is real, and you should enjoy it Let me be fair to the tools first, because the backlash has gotten lazy.

When agentic coding properly landed in my workflow, it felt like someone lifted a weight off my chest I had not known I was carrying. All that configuration written by hand. All that run, squint, fix the typo, run again.

Suddenly optional. Work that used to take a day was done before lunch. A migration I had avoided for a quarter got drafted in an afternoon.

My manager noticed. My team noticed. This is not vibes.

In a controlled study of developers building a simple HTTP server, the ones with an AI assistant finished noticeably faster. In a field experiment across thousands of developers, merged pull requests rose by roughly a quarter. If your work involves a lot of greenfield code or a lot of boilerplate, the AI speedup is real and it is not small.

So we did the obvious thing. We got faster, so we took on more. Bug report that smells like infrastructure? I am on it.

Someone needs a dashboard by Friday? Sure. Ticket from March rotting in the backlog? Why not. My open pull request count started to look like a typo.

The bill AI quietly runs up Here is what nobody tells you about being five times faster. You can also get lost five times faster. The first sign was easy to ignore.

A colleague asked about one of my open PRs, and I had to read my own description to remember what it was for. The description had been written by AI. I was reading a machine's summary of a decision I had apparently made, in order to find out what I thought.

That is not productivity. That is a queue with your name on it. The second sign was not small.

An agent produced a change adding a new permission set for a service. Clean diff. Sensible naming.

Tests green. I reviewed it the way I had started reviewing everything by then, which is to say I scrolled, nodded, approved. It had been open six days and I wanted it gone.

Two things saved me. A teammate who actually reads policy documents left one comment: "is this wildcard on purpose?" And luck, in that the comment landed before the merge did. The AI had done exactly what I asked.

I had asked for the wrong thing, vaguely, and it filled the gap with the most permissive option available. The tests passed because the tests checked that the permission existed, not that it was safe. That is the sentence that now governs my working day:"All tests pass" is not the same as "this does what I meant.".

Was ist das Hauptproblem bei einer zu starken Abhängigkeit von KI-Coding-Agenten?

Das Hauptproblem besteht darin, dass KI Code erzeugen kann, der Tests besteht, aber nicht die tatsächliche Absicht des Benutzers widerspiegelt, was zu einem falschen Produktivitätsgefühl und potenziellen Sicherheitsrisiken führt, wenn Benutzer ihre eigene Arbeit nicht mehr verstehen.

Deutsch version →


🇪🇸 Español

La IA me hizo 5 veces más rápido, 5 veces peor en el trabajo

There is a moment happening in offices and bedrooms all over the world right now, and it looks like this. Three AI agent sessions are running. One is refactoring something.

One is writing tests. One is halfway through a migration nobody wanted to do by hand. The person in front of them is not typing.

They are watching. Their eyes move between panes like someone who put chips on three tables and cannot decide which one to be nervous about. Then the thought arrives.

I could start a fourth one. I call it the fourth terminal, and I think it is the defining mistake of the AI coding era. Not because running agents in parallel is bad.

Because of what the reflex reveals. AI handed us spare capacity, and our first instinct was to fill it with more AI, instead of asking what the spare capacity was actually for. I made that mistake for about four months.

Nearly everyone I know made it too. This is what it cost, what the research now says about why it happens, and what the people who came out the other side are doing instead. The AI honeymoon is real, and you should enjoy it Let me be fair to the tools first, because the backlash has gotten lazy.

When agentic coding properly landed in my workflow, it felt like someone lifted a weight off my chest I had not known I was carrying. All that configuration written by hand. All that run, squint, fix the typo, run again.

Suddenly optional. Work that used to take a day was done before lunch. A migration I had avoided for a quarter got drafted in an afternoon.

My manager noticed. My team noticed. This is not vibes.

In a controlled study of developers building a simple HTTP server, the ones with an AI assistant finished noticeably faster. In a field experiment across thousands of developers, merged pull requests rose by roughly a quarter. If your work involves a lot of greenfield code or a lot of boilerplate, the AI speedup is real and it is not small.

So we did the obvious thing. We got faster, so we took on more. Bug report that smells like infrastructure? I am on it.

Someone needs a dashboard by Friday? Sure. Ticket from March rotting in the backlog? Why not. My open pull request count started to look like a typo.

The bill AI quietly runs up Here is what nobody tells you about being five times faster. You can also get lost five times faster. The first sign was easy to ignore.

A colleague asked about one of my open PRs, and I had to read my own description to remember what it was for. The description had been written by AI. I was reading a machine's summary of a decision I had apparently made, in order to find out what I thought.

That is not productivity. That is a queue with your name on it. The second sign was not small.

An agent produced a change adding a new permission set for a service. Clean diff. Sensible naming.

Tests green. I reviewed it the way I had started reviewing everything by then, which is to say I scrolled, nodded, approved. It had been open six days and I wanted it gone.

Two things saved me. A teammate who actually reads policy documents left one comment: "is this wildcard on purpose?" And luck, in that the comment landed before the merge did. The AI had done exactly what I asked.

I had asked for the wrong thing, vaguely, and it filled the gap with the most permissive option available. The tests passed because the tests checked that the permission existed, not that it was safe. That is the sentence that now governs my working day:"All tests pass" is not the same as "this does what I meant.".

¿Cuál es el problema central de depender demasiado de los agentes de codificación de IA?

El problema central es que la IA puede producir código que pasa las pruebas pero que no refleja la intención real del usuario, lo que genera una falsa sensación de productividad y riesgos de seguridad potenciales cuando los usuarios dejan de entender su propio trabajo.

Español version →


🇫🇷 Français

L'IA m'a rendu 5 fois plus rapide, 5 fois pire au travail

There is a moment happening in offices and bedrooms all over the world right now, and it looks like this. Three AI agent sessions are running. One is refactoring something.

One is writing tests. One is halfway through a migration nobody wanted to do by hand. The person in front of them is not typing.

They are watching. Their eyes move between panes like someone who put chips on three tables and cannot decide which one to be nervous about. Then the thought arrives.

I could start a fourth one. I call it the fourth terminal, and I think it is the defining mistake of the AI coding era. Not because running agents in parallel is bad.

Because of what the reflex reveals. AI handed us spare capacity, and our first instinct was to fill it with more AI, instead of asking what the spare capacity was actually for. I made that mistake for about four months.

Nearly everyone I know made it too. This is what it cost, what the research now says about why it happens, and what the people who came out the other side are doing instead. The AI honeymoon is real, and you should enjoy it Let me be fair to the tools first, because the backlash has gotten lazy.

When agentic coding properly landed in my workflow, it felt like someone lifted a weight off my chest I had not known I was carrying. All that configuration written by hand. All that run, squint, fix the typo, run again.

Suddenly optional. Work that used to take a day was done before lunch. A migration I had avoided for a quarter got drafted in an afternoon.

My manager noticed. My team noticed. This is not vibes.

In a controlled study of developers building a simple HTTP server, the ones with an AI assistant finished noticeably faster. In a field experiment across thousands of developers, merged pull requests rose by roughly a quarter. If your work involves a lot of greenfield code or a lot of boilerplate, the AI speedup is real and it is not small.

So we did the obvious thing. We got faster, so we took on more. Bug report that smells like infrastructure? I am on it.

Someone needs a dashboard by Friday? Sure. Ticket from March rotting in the backlog? Why not. My open pull request count started to look like a typo.

The bill AI quietly runs up Here is what nobody tells you about being five times faster. You can also get lost five times faster. The first sign was easy to ignore.

A colleague asked about one of my open PRs, and I had to read my own description to remember what it was for. The description had been written by AI. I was reading a machine's summary of a decision I had apparently made, in order to find out what I thought.

That is not productivity. That is a queue with your name on it. The second sign was not small.

An agent produced a change adding a new permission set for a service. Clean diff. Sensible naming.

Tests green. I reviewed it the way I had started reviewing everything by then, which is to say I scrolled, nodded, approved. It had been open six days and I wanted it gone.

Two things saved me. A teammate who actually reads policy documents left one comment: "is this wildcard on purpose?" And luck, in that the comment landed before the merge did. The AI had done exactly what I asked.

I had asked for the wrong thing, vaguely, and it filled the gap with the most permissive option available. The tests passed because the tests checked that the permission existed, not that it was safe. That is the sentence that now governs my working day:"All tests pass" is not the same as "this does what I meant.".

Quel est le problème fondamental lié à une dépendance excessive aux agents de codage IA ?

Le problème fondamental est que l'IA peut produire du code qui passe les tests mais qui ne reflète pas l'intention réelle de l'utilisateur, ce qui entraîne un faux sentiment de productivité et des risques de sécurité potentiels lorsque les utilisateurs cessent de comprendre leur propre travail.

Français version →


🇮🇳 हिन्दी

AI ने मुझे काम में 5 गुना तेज़ किया, लेकिन 5 गुना बुरा भी बना दिया

There is a moment happening in offices and bedrooms all over the world right now, and it looks like this. Three AI agent sessions are running. One is refactoring something.

One is writing tests. One is halfway through a migration nobody wanted to do by hand. The person in front of them is not typing.

They are watching. Their eyes move between panes like someone who put chips on three tables and cannot decide which one to be nervous about. Then the thought arrives.

I could start a fourth one. I call it the fourth terminal, and I think it is the defining mistake of the AI coding era. Not because running agents in parallel is bad.

Because of what the reflex reveals. AI handed us spare capacity, and our first instinct was to fill it with more AI, instead of asking what the spare capacity was actually for. I made that mistake for about four months.

Nearly everyone I know made it too. This is what it cost, what the research now says about why it happens, and what the people who came out the other side are doing instead. The AI honeymoon is real, and you should enjoy it Let me be fair to the tools first, because the backlash has gotten lazy.

When agentic coding properly landed in my workflow, it felt like someone lifted a weight off my chest I had not known I was carrying. All that configuration written by hand. All that run, squint, fix the typo, run again.

Suddenly optional. Work that used to take a day was done before lunch. A migration I had avoided for a quarter got drafted in an afternoon.

My manager noticed. My team noticed. This is not vibes.

In a controlled study of developers building a simple HTTP server, the ones with an AI assistant finished noticeably faster. In a field experiment across thousands of developers, merged pull requests rose by roughly a quarter. If your work involves a lot of greenfield code or a lot of boilerplate, the AI speedup is real and it is not small.

So we did the obvious thing. We got faster, so we took on more. Bug report that smells like infrastructure? I am on it.

Someone needs a dashboard by Friday? Sure. Ticket from March rotting in the backlog? Why not. My open pull request count started to look like a typo.

The bill AI quietly runs up Here is what nobody tells you about being five times faster. You can also get lost five times faster. The first sign was easy to ignore.

A colleague asked about one of my open PRs, and I had to read my own description to remember what it was for. The description had been written by AI. I was reading a machine's summary of a decision I had apparently made, in order to find out what I thought.

That is not productivity. That is a queue with your name on it. The second sign was not small.

An agent produced a change adding a new permission set for a service. Clean diff. Sensible naming.

Tests green. I reviewed it the way I had started reviewing everything by then, which is to say I scrolled, nodded, approved. It had been open six days and I wanted it gone.

Two things saved me. A teammate who actually reads policy documents left one comment: "is this wildcard on purpose?" And luck, in that the comment landed before the merge did. The AI had done exactly what I asked.

I had asked for the wrong thing, vaguely, and it filled the gap with the most permissive option available. The tests passed because the tests checked that the permission existed, not that it was safe. That is the sentence that now governs my working day:"All tests pass" is not the same as "this does what I meant.".

AI कोडिंग एजेंट्स पर अत्यधिक निर्भर रहने की मुख्य समस्या क्या है?

मुख्य समस्या यह है कि AI ऐसा कोड उत्पन्न कर सकता है जो परीक्षण पास कर देता है लेकिन उपयोगकर्ता के वास्तविक इरादे को नहीं दर्शाता, जिससे उत्पादकता की झूठी भावना पैदा होती है और जब उपयोगकर्ता अपना काम समझना बंद कर देते हैं तो संभावित सुरक्षा जोखिम पैदा होते हैं।

हिन्दी version →


🇮🇩 Bahasa Indonesia

AI Membuat Saya 5x Lebih Cepat, 5x Lebih Buruk di Kerja

There is a moment happening in offices and bedrooms all over the world right now, and it looks like this. Three AI agent sessions are running. One is refactoring something.

One is writing tests. One is halfway through a migration nobody wanted to do by hand. The person in front of them is not typing.

They are watching. Their eyes move between panes like someone who put chips on three tables and cannot decide which one to be nervous about. Then the thought arrives.

I could start a fourth one. I call it the fourth terminal, and I think it is the defining mistake of the AI coding era. Not because running agents in parallel is bad.

Because of what the reflex reveals. AI handed us spare capacity, and our first instinct was to fill it with more AI, instead of asking what the spare capacity was actually for. I made that mistake for about four months.

Nearly everyone I know made it too. This is what it cost, what the research now says about why it happens, and what the people who came out the other side are doing instead. The AI honeymoon is real, and you should enjoy it Let me be fair to the tools first, because the backlash has gotten lazy.

When agentic coding properly landed in my workflow, it felt like someone lifted a weight off my chest I had not known I was carrying. All that configuration written by hand. All that run, squint, fix the typo, run again.

Suddenly optional. Work that used to take a day was done before lunch. A migration I had avoided for a quarter got drafted in an afternoon.

My manager noticed. My team noticed. This is not vibes.

In a controlled study of developers building a simple HTTP server, the ones with an AI assistant finished noticeably faster. In a field experiment across thousands of developers, merged pull requests rose by roughly a quarter. If your work involves a lot of greenfield code or a lot of boilerplate, the AI speedup is real and it is not small.

So we did the obvious thing. We got faster, so we took on more. Bug report that smells like infrastructure? I am on it.

Someone needs a dashboard by Friday? Sure. Ticket from March rotting in the backlog? Why not. My open pull request count started to look like a typo.

The bill AI quietly runs up Here is what nobody tells you about being five times faster. You can also get lost five times faster. The first sign was easy to ignore.

A colleague asked about one of my open PRs, and I had to read my own description to remember what it was for. The description had been written by AI. I was reading a machine's summary of a decision I had apparently made, in order to find out what I thought.

That is not productivity. That is a queue with your name on it. The second sign was not small.

An agent produced a change adding a new permission set for a service. Clean diff. Sensible naming.

Tests green. I reviewed it the way I had started reviewing everything by then, which is to say I scrolled, nodded, approved. It had been open six days and I wanted it gone.

Two things saved me. A teammate who actually reads policy documents left one comment: "is this wildcard on purpose?" And luck, in that the comment landed before the merge did. The AI had done exactly what I asked.

I had asked for the wrong thing, vaguely, and it filled the gap with the most permissive option available. The tests passed because the tests checked that the permission existed, not that it was safe. That is the sentence that now governs my working day:"All tests pass" is not the same as "this does what I meant.".

Apa masalah inti dari ketergantungan berlebihan pada agen coding AI?

Masalah inti adalah bahwa AI dapat menghasilkan kode yang lolos tes tetapi tidak mencerminkan niat sebenarnya pengguna, yang mengakibatkan rasa produktivitas yang salah dan risiko keamanan potensial ketika pengguna berhenti memahami pekerjaan mereka sendiri.

Bahasa Indonesia version →


🇯🇵 日本語

AIは私を仕事で5倍速くしたが、5倍悪くもした

There is a moment happening in offices and bedrooms all over the world right now, and it looks like this. Three AI agent sessions are running. One is refactoring something.

One is writing tests. One is halfway through a migration nobody wanted to do by hand. The person in front of them is not typing.

They are watching. Their eyes move between panes like someone who put chips on three tables and cannot decide which one to be nervous about. Then the thought arrives.

I could start a fourth one. I call it the fourth terminal, and I think it is the defining mistake of the AI coding era. Not because running agents in parallel is bad.

Because of what the reflex reveals. AI handed us spare capacity, and our first instinct was to fill it with more AI, instead of asking what the spare capacity was actually for. I made that mistake for about four months.

Nearly everyone I know made it too. This is what it cost, what the research now says about why it happens, and what the people who came out the other side are doing instead. The AI honeymoon is real, and you should enjoy it Let me be fair to the tools first, because the backlash has gotten lazy.

When agentic coding properly landed in my workflow, it felt like someone lifted a weight off my chest I had not known I was carrying. All that configuration written by hand. All that run, squint, fix the typo, run again.

Suddenly optional. Work that used to take a day was done before lunch. A migration I had avoided for a quarter got drafted in an afternoon.

My manager noticed. My team noticed. This is not vibes.

In a controlled study of developers building a simple HTTP server, the ones with an AI assistant finished noticeably faster. In a field experiment across thousands of developers, merged pull requests rose by roughly a quarter. If your work involves a lot of greenfield code or a lot of boilerplate, the AI speedup is real and it is not small.

So we did the obvious thing. We got faster, so we took on more. Bug report that smells like infrastructure? I am on it.

Someone needs a dashboard by Friday? Sure. Ticket from March rotting in the backlog? Why not. My open pull request count started to look like a typo.

The bill AI quietly runs up Here is what nobody tells you about being five times faster. You can also get lost five times faster. The first sign was easy to ignore.

A colleague asked about one of my open PRs, and I had to read my own description to remember what it was for. The description had been written by AI. I was reading a machine's summary of a decision I had apparently made, in order to find out what I thought.

That is not productivity. That is a queue with your name on it. The second sign was not small.

An agent produced a change adding a new permission set for a service. Clean diff. Sensible naming.

Tests green. I reviewed it the way I had started reviewing everything by then, which is to say I scrolled, nodded, approved. It had been open six days and I wanted it gone.

Two things saved me. A teammate who actually reads policy documents left one comment: "is this wildcard on purpose?" And luck, in that the comment landed before the merge did. The AI had done exactly what I asked.

I had asked for the wrong thing, vaguely, and it filled the gap with the most permissive option available. The tests passed because the tests checked that the permission existed, not that it was safe. That is the sentence that now governs my working day:"All tests pass" is not the same as "this does what I meant.".

AIコーディングエージェントに過度に依存することの核心的な問題とは何ですか?

核心的な問題は、AIがテストをパスするコードを生成できるが、ユーザーの実際の意図を反映していないことです。これにより、誤った生産性の感覚が生じ、ユーザーが自分の仕事を理解しなくなると、潜在的なセキュリティリスクが生じます。

日本語 version →


🇧🇷 Português

A IA me deixou 5 vezes mais rápido, 5 vezes pior no trabalho

There is a moment happening in offices and bedrooms all over the world right now, and it looks like this. Three AI agent sessions are running. One is refactoring something.

One is writing tests. One is halfway through a migration nobody wanted to do by hand. The person in front of them is not typing.

They are watching. Their eyes move between panes like someone who put chips on three tables and cannot decide which one to be nervous about. Then the thought arrives.

I could start a fourth one. I call it the fourth terminal, and I think it is the defining mistake of the AI coding era. Not because running agents in parallel is bad.

Because of what the reflex reveals. AI handed us spare capacity, and our first instinct was to fill it with more AI, instead of asking what the spare capacity was actually for. I made that mistake for about four months.

Nearly everyone I know made it too. This is what it cost, what the research now says about why it happens, and what the people who came out the other side are doing instead. The AI honeymoon is real, and you should enjoy it Let me be fair to the tools first, because the backlash has gotten lazy.

When agentic coding properly landed in my workflow, it felt like someone lifted a weight off my chest I had not known I was carrying. All that configuration written by hand. All that run, squint, fix the typo, run again.

Suddenly optional. Work that used to take a day was done before lunch. A migration I had avoided for a quarter got drafted in an afternoon.

My manager noticed. My team noticed. This is not vibes.

In a controlled study of developers building a simple HTTP server, the ones with an AI assistant finished noticeably faster. In a field experiment across thousands of developers, merged pull requests rose by roughly a quarter. If your work involves a lot of greenfield code or a lot of boilerplate, the AI speedup is real and it is not small.

So we did the obvious thing. We got faster, so we took on more. Bug report that smells like infrastructure? I am on it.

Someone needs a dashboard by Friday? Sure. Ticket from March rotting in the backlog? Why not. My open pull request count started to look like a typo.

The bill AI quietly runs up Here is what nobody tells you about being five times faster. You can also get lost five times faster. The first sign was easy to ignore.

A colleague asked about one of my open PRs, and I had to read my own description to remember what it was for. The description had been written by AI. I was reading a machine's summary of a decision I had apparently made, in order to find out what I thought.

That is not productivity. That is a queue with your name on it. The second sign was not small.

An agent produced a change adding a new permission set for a service. Clean diff. Sensible naming.

Tests green. I reviewed it the way I had started reviewing everything by then, which is to say I scrolled, nodded, approved. It had been open six days and I wanted it gone.

Two things saved me. A teammate who actually reads policy documents left one comment: "is this wildcard on purpose?" And luck, in that the comment landed before the merge did. The AI had done exactly what I asked.

I had asked for the wrong thing, vaguely, and it filled the gap with the most permissive option available. The tests passed because the tests checked that the permission existed, not that it was safe. That is the sentence that now governs my working day:"All tests pass" is not the same as "this does what I meant.".

Qual é o problema central de depender excessivamente de agentes de codificação de IA?

O problema central é que a IA pode produzir código que passa nos testes, mas que não reflete a intenção real do usuário, levando a uma falsa sensação de produtividade e riscos de segurança potenciais quando os usuários param de entender seu próprio trabalho.

Português version →


🇷🇺 Русский

ИИ сделал меня в 5 раз быстрее, но и в 5 раз хуже на работе

There is a moment happening in offices and bedrooms all over the world right now, and it looks like this. Three AI agent sessions are running. One is refactoring something.

One is writing tests. One is halfway through a migration nobody wanted to do by hand. The person in front of them is not typing.

They are watching. Their eyes move between panes like someone who put chips on three tables and cannot decide which one to be nervous about. Then the thought arrives.

I could start a fourth one. I call it the fourth terminal, and I think it is the defining mistake of the AI coding era. Not because running agents in parallel is bad.

Because of what the reflex reveals. AI handed us spare capacity, and our first instinct was to fill it with more AI, instead of asking what the spare capacity was actually for. I made that mistake for about four months.

Nearly everyone I know made it too. This is what it cost, what the research now says about why it happens, and what the people who came out the other side are doing instead. The AI honeymoon is real, and you should enjoy it Let me be fair to the tools first, because the backlash has gotten lazy.

When agentic coding properly landed in my workflow, it felt like someone lifted a weight off my chest I had not known I was carrying. All that configuration written by hand. All that run, squint, fix the typo, run again.

Suddenly optional. Work that used to take a day was done before lunch. A migration I had avoided for a quarter got drafted in an afternoon.

My manager noticed. My team noticed. This is not vibes.

In a controlled study of developers building a simple HTTP server, the ones with an AI assistant finished noticeably faster. In a field experiment across thousands of developers, merged pull requests rose by roughly a quarter. If your work involves a lot of greenfield code or a lot of boilerplate, the AI speedup is real and it is not small.

So we did the obvious thing. We got faster, so we took on more. Bug report that smells like infrastructure? I am on it.

Someone needs a dashboard by Friday? Sure. Ticket from March rotting in the backlog? Why not. My open pull request count started to look like a typo.

The bill AI quietly runs up Here is what nobody tells you about being five times faster. You can also get lost five times faster. The first sign was easy to ignore.

A colleague asked about one of my open PRs, and I had to read my own description to remember what it was for. The description had been written by AI. I was reading a machine's summary of a decision I had apparently made, in order to find out what I thought.

That is not productivity. That is a queue with your name on it. The second sign was not small.

An agent produced a change adding a new permission set for a service. Clean diff. Sensible naming.

Tests green. I reviewed it the way I had started reviewing everything by then, which is to say I scrolled, nodded, approved. It had been open six days and I wanted it gone.

Two things saved me. A teammate who actually reads policy documents left one comment: "is this wildcard on purpose?" And luck, in that the comment landed before the merge did. The AI had done exactly what I asked.

I had asked for the wrong thing, vaguely, and it filled the gap with the most permissive option available. The tests passed because the tests checked that the permission existed, not that it was safe. That is the sentence that now governs my working day:"All tests pass" is not the same as "this does what I meant.".

Какая основная проблема возникает при чрезмерной зависимости от ИИ-агентов для написания кода?

Основная проблема заключается в том, что ИИ может генерировать код, который проходит тесты, но не отражает реальное намерение пользователя, что приводит к ложному ощущению продуктивности и потенциальным рискам безопасности, когда пользователи перестают понимать свою собственную работу.

Русский version →


🇨🇳 简体中文

AI 让我在工作中效率提高了 5 倍,但也变得糟糕了 5 倍

There is a moment happening in offices and bedrooms all over the world right now, and it looks like this. Three AI agent sessions are running. One is refactoring something.

One is writing tests. One is halfway through a migration nobody wanted to do by hand. The person in front of them is not typing.

They are watching. Their eyes move between panes like someone who put chips on three tables and cannot decide which one to be nervous about. Then the thought arrives.

I could start a fourth one. I call it the fourth terminal, and I think it is the defining mistake of the AI coding era. Not because running agents in parallel is bad.

Because of what the reflex reveals. AI handed us spare capacity, and our first instinct was to fill it with more AI, instead of asking what the spare capacity was actually for. I made that mistake for about four months.

Nearly everyone I know made it too. This is what it cost, what the research now says about why it happens, and what the people who came out the other side are doing instead. The AI honeymoon is real, and you should enjoy it Let me be fair to the tools first, because the backlash has gotten lazy.

When agentic coding properly landed in my workflow, it felt like someone lifted a weight off my chest I had not known I was carrying. All that configuration written by hand. All that run, squint, fix the typo, run again.

Suddenly optional. Work that used to take a day was done before lunch. A migration I had avoided for a quarter got drafted in an afternoon.

My manager noticed. My team noticed. This is not vibes.

In a controlled study of developers building a simple HTTP server, the ones with an AI assistant finished noticeably faster. In a field experiment across thousands of developers, merged pull requests rose by roughly a quarter. If your work involves a lot of greenfield code or a lot of boilerplate, the AI speedup is real and it is not small.

So we did the obvious thing. We got faster, so we took on more. Bug report that smells like infrastructure? I am on it.

Someone needs a dashboard by Friday? Sure. Ticket from March rotting in the backlog? Why not. My open pull request count started to look like a typo.

The bill AI quietly runs up Here is what nobody tells you about being five times faster. You can also get lost five times faster. The first sign was easy to ignore.

A colleague asked about one of my open PRs, and I had to read my own description to remember what it was for. The description had been written by AI. I was reading a machine's summary of a decision I had apparently made, in order to find out what I thought.

That is not productivity. That is a queue with your name on it. The second sign was not small.

An agent produced a change adding a new permission set for a service. Clean diff. Sensible naming.

Tests green. I reviewed it the way I had started reviewing everything by then, which is to say I scrolled, nodded, approved. It had been open six days and I wanted it gone.

Two things saved me. A teammate who actually reads policy documents left one comment: "is this wildcard on purpose?" And luck, in that the comment landed before the merge did. The AI had done exactly what I asked.

I had asked for the wrong thing, vaguely, and it filled the gap with the most permissive option available. The tests passed because the tests checked that the permission existed, not that it was safe. That is the sentence that now governs my working day:"All tests pass" is not the same as "this does what I meant.".

过度依赖 AI 编码代理的核心问题是什么?

核心问题是 AI 可以生成通过测试的代码,但并不反映用户的实际意图,导致错误的生产力感和潜在的安全风险,当用户停止理解自己的工作时。

简体中文 version →