{"id":19990,"date":"2026-07-14T08:00:53","date_gmt":"2026-07-14T06:00:53","guid":{"rendered":"https:\/\/www.gtb.de\/?p=19990"},"modified":"2026-07-14T17:28:21","modified_gmt":"2026-07-14T15:28:21","slug":"when-the-agents-write-the-code-the-human-checks-themselves-2-2","status":"publish","type":"post","link":"https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/","title":{"rendered":"# When the Agents Write the Code, the Human Checks Themselves [2\/2]"},"content":{"rendered":"<div class=\"fusion-fullwidth fullwidth-box fusion-builder-row-1 fusion-flex-container has-pattern-background has-mask-background nonhundred-percent-fullwidth non-hundred-percent-height-scrolling\" style=\"--awb-border-radius-top-left:0px;--awb-border-radius-top-right:0px;--awb-border-radius-bottom-right:0px;--awb-border-radius-bottom-left:0px;--awb-flex-wrap:wrap;\" ><div class=\"fusion-builder-row fusion-row fusion-flex-align-items-flex-start fusion-flex-content-wrap\" style=\"max-width:calc( 1920px + 30px );margin-left: calc(-30px \/ 2 );margin-right: calc(-30px \/ 2 );\"><div class=\"fusion-layout-column fusion_builder_column fusion-builder-column-0 fusion_builder_column_1_1 1_1 fusion-flex-column\" style=\"--awb-bg-size:cover;--awb-width-large:100%;--awb-margin-top-large:0px;--awb-spacing-right-large:15px;--awb-margin-bottom-large:20px;--awb-spacing-left-large:15px;--awb-width-medium:100%;--awb-order-medium:0;--awb-spacing-right-medium:15px;--awb-spacing-left-medium:15px;--awb-width-small:100%;--awb-order-small:0;--awb-spacing-right-small:15px;--awb-spacing-left-small:15px;\"><div class=\"fusion-column-wrapper fusion-column-has-shadow fusion-flex-justify-content-flex-start fusion-content-layout-column\"><div class=\"fusion-builder-row fusion-builder-row-inner fusion-row fusion-flex-align-items-flex-start fusion-flex-content-wrap\" style=\"--awb-flex-grow:0;--awb-flex-grow-medium:0;--awb-flex-grow-small:0;--awb-flex-shrink:0;--awb-flex-shrink-medium:0;--awb-flex-shrink-small:0;width:calc( 100% + 30px ) !important;max-width:calc( 100% + 30px ) !important;margin-left: calc(-30px \/ 2 );margin-right: calc(-30px \/ 2 );\"><div class=\"fusion-layout-column fusion_builder_column_inner fusion-builder-nested-column-0 fusion_builder_column_inner_1_1 1_1 fusion-flex-column blog-headline-box\" style=\"--awb-bg-size:cover;--awb-width-large:100%;--awb-margin-top-large:0px;--awb-spacing-right-large:15px;--awb-margin-bottom-large:20px;--awb-spacing-left-large:15px;--awb-width-medium:100%;--awb-order-medium:0;--awb-spacing-right-medium:15px;--awb-spacing-left-medium:15px;--awb-width-small:100%;--awb-order-small:0;--awb-spacing-right-small:15px;--awb-spacing-left-small:15px;\" data-scroll-devices=\"small-visibility,medium-visibility,large-visibility\"><div class=\"fusion-column-wrapper fusion-column-has-shadow fusion-flex-justify-content-flex-start fusion-content-layout-column\"><div class=\"fusion-title title fusion-title-1 fusion-sep-none fusion-title-center fusion-title-text fusion-title-size-two\" style=\"--awb-text-color:var(--awb-color6);--awb-margin-bottom:56px;--awb-margin-top-small:0px;--awb-margin-right-small:0px;--awb-margin-bottom-small:56px;--awb-margin-left-small:0px;--awb-font-size:var(--awb-custom_typography_12-font-size);\"><h2 class=\"fusion-title-heading title-heading-center fusion-responsive-typography-calculated\" style=\"font-family:var(--awb-custom_typography_12-font-family);font-weight:var(--awb-custom_typography_12-font-weight);font-style:var(--awb-custom_typography_12-font-style);margin:0;letter-spacing:var(--awb-custom_typography_12-letter-spacing);text-transform:var(--awb-custom_typography_12-text-transform);font-size:1em;--fontSize:40;line-height:var(--awb-custom_typography_12-line-height);\">## Part 2: The Price, or Three Cognitive Bottlenecks No Demo Ever Shows<\/h2><\/div><div class=\"fusion-title title fusion-title-2 fusion-sep-none fusion-title-center fusion-title-text fusion-title-size-paragraph\" style=\"--awb-text-color:var(--awb-color6);--awb-margin-bottom:56px;--awb-margin-top-small:0px;--awb-margin-right-small:0px;--awb-margin-bottom-small:56px;--awb-margin-left-small:0px;--awb-font-size:32px;\"><p class=\"fusion-title-heading title-heading-center title-heading-tag fusion-responsive-typography-calculated\" style=\"font-family:var(--awb-custom_typography_12-font-family);font-weight:var(--awb-custom_typography_12-font-weight);font-style:var(--awb-custom_typography_12-font-style);margin:0;letter-spacing:var(--awb-custom_typography_12-letter-spacing);text-transform:var(--awb-custom_typography_12-text-transform);font-size:1em;--fontSize:32;line-height:var(--awb-custom_typography_12-line-height);\">*(A two-part series. Part 1 describes what a working agentic setup looks like. Part 2 shows the price the human pays for it and how to keep that price under control. Both parts can be read on their own.)*<\/p><\/div><\/div><\/div><\/div><div class=\"fusion-builder-row fusion-builder-row-inner fusion-row fusion-flex-align-items-flex-start fusion-flex-content-wrap\" style=\"--awb-flex-grow:0;--awb-flex-grow-medium:0;--awb-flex-grow-small:0;--awb-flex-shrink:0;--awb-flex-shrink-medium:0;--awb-flex-shrink-small:0;width:calc( 100% + 30px ) !important;max-width:calc( 100% + 30px ) !important;margin-left: calc(-30px \/ 2 );margin-right: calc(-30px \/ 2 );\"><div class=\"fusion-layout-column fusion_builder_column_inner fusion-builder-nested-column-1 fusion_builder_column_inner_1_2 1_2 fusion-flex-column\" style=\"--awb-bg-size:cover;--awb-width-large:50%;--awb-margin-top-large:0px;--awb-spacing-right-large:15px;--awb-margin-bottom-large:20px;--awb-spacing-left-large:15px;--awb-width-medium:50%;--awb-order-medium:0;--awb-spacing-right-medium:15px;--awb-spacing-left-medium:15px;--awb-width-small:100%;--awb-order-small:0;--awb-spacing-right-small:15px;--awb-spacing-left-small:15px;\"><div class=\"fusion-column-wrapper fusion-column-has-shadow fusion-flex-justify-content-flex-start fusion-content-layout-column\"><div class=\"fusion-text fusion-text-1\" style=\"--awb-text-color:var(--awb-color6);\"><p>Picture a normal morning. The setup is warm, several agents run in parallel. One asks whether it may commit. The second has reported a red test and waits for a diagnosis. The third proposes an architecture switch. The fourth wants a decision on logging. Four decisions before the coffee goes cold. And that is the quiet part of the morning.<\/p>\n<\/div><\/div><\/div><div class=\"fusion-layout-column fusion_builder_column_inner fusion-builder-nested-column-2 fusion_builder_column_inner_1_2 1_2 fusion-flex-column\" style=\"--awb-bg-size:cover;--awb-width-large:50%;--awb-margin-top-large:0px;--awb-spacing-right-large:15px;--awb-margin-bottom-large:20px;--awb-spacing-left-large:15px;--awb-width-medium:50%;--awb-order-medium:0;--awb-spacing-right-medium:15px;--awb-spacing-left-medium:15px;--awb-width-small:100%;--awb-order-small:0;--awb-spacing-right-small:15px;--awb-spacing-left-small:15px;\"><div class=\"fusion-column-wrapper fusion-column-has-shadow fusion-flex-justify-content-flex-start fusion-content-layout-column\"><div class=\"fusion-text fusion-text-2\" style=\"--awb-text-color:var(--awb-color6);\"><p>Agentic development here does not mean vibe coding, that is &#8220;the AI builds on the fly, I keep whatever sounds right&#8221;. It means an orchestrated setup of several specialized subagents, each with a clear mandate, running against a repository base with good test coverage, and a human who holds the strategic layer. How that setup looks in detail, why several models run side by side, and what makes it economically viable is covered in Part 1. This part is about the other half, the one that rarely gets told to the end: what lands on the human who supervises the agents.<\/p>\n<p>Most texts about agentic development sell speed. Ship faster, more output per day, less typing. All true. Except typing was never the exhausting part. Judging is. And that is exactly where three bottlenecks appear that I have come to take seriously.<\/p>\n<\/div><\/div><\/div><\/div><div class=\"fusion-builder-row fusion-builder-row-inner fusion-row fusion-flex-align-items-flex-start fusion-flex-justify-content-space-between fusion-flex-content-wrap\" style=\"--awb-flex-grow:0;--awb-flex-grow-medium:0;--awb-flex-grow-small:0;--awb-flex-shrink:0;--awb-flex-shrink-medium:0;--awb-flex-shrink-small:0;width:calc( 100% + 30px ) !important;max-width:calc( 100% + 30px ) !important;margin-left: calc(-30px \/ 2 );margin-right: calc(-30px \/ 2 );\"><div class=\"fusion-layout-column fusion_builder_column_inner fusion-builder-nested-column-3 awb-sticky awb-sticky-medium awb-sticky-large fusion_builder_column_inner_2_5 2_5 fusion-flex-column blue-box person-box\" style=\"--awb-padding-top:24px;--awb-padding-right:24px;--awb-padding-bottom:24px;--awb-padding-left:24px;--awb-padding-top-medium:24px;--awb-padding-right-medium:24px;--awb-padding-bottom-medium:24px;--awb-padding-left-medium:24px;--awb-padding-top-small:24px;--awb-padding-right-small:24px;--awb-padding-bottom-small:24px;--awb-padding-left-small:24px;--awb-overflow:hidden;--awb-bg-color:#28A0DC33;--awb-bg-color-hover:#28A0DC33;--awb-bg-size:cover;--awb-border-radius:16px 16px 16px 16px;--awb-width-large:40%;--awb-margin-top-large:0px;--awb-spacing-right-large:15px;--awb-margin-bottom-large:187px;--awb-spacing-left-large:15px;--awb-width-medium:40%;--awb-order-medium:0;--awb-spacing-right-medium:15px;--awb-spacing-left-medium:15px;--awb-width-small:100%;--awb-order-small:0;--awb-spacing-right-small:15px;--awb-spacing-left-small:15px;--awb-sticky-offset:150px;\" data-scroll-devices=\"small-visibility,medium-visibility,large-visibility\"><div class=\"fusion-column-wrapper fusion-column-has-shadow fusion-flex-justify-content-flex-start fusion-content-layout-column\"><div class=\"fusion-image-element\" style=\"text-align:center;--awb-margin-bottom:8px;--awb-caption-title-font-family:var(--h2_typography-font-family);--awb-caption-title-font-weight:var(--h2_typography-font-weight);--awb-caption-title-font-style:var(--h2_typography-font-style);--awb-caption-title-size:var(--h2_typography-font-size);--awb-caption-title-transform:var(--h2_typography-text-transform);--awb-caption-title-line-height:var(--h2_typography-line-height);--awb-caption-title-letter-spacing:var(--h2_typography-letter-spacing);\"><span class=\" fusion-imageframe imageframe-none imageframe-1 hover-type-none\" style=\"border-radius:50%;\"><img decoding=\"async\" width=\"1600\" height=\"1610\" alt=\"J\u00f6rg Amelunxen\" title=\"J\u00f6rg Amelunxen\" src=\"https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Joerg-Amelunxen.png\" data-orig-src=\"https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Joerg-Amelunxen.png\" class=\"lazyload img-responsive wp-image-19977\" srcset=\"data:image\/svg+xml,%3Csvg%20xmlns%3D%27http%3A%2F%2Fwww.w3.org%2F2000%2Fsvg%27%20width%3D%271600%27%20height%3D%271610%27%20viewBox%3D%270%200%201600%201610%27%3E%3Crect%20width%3D%271600%27%20height%3D%271610%27%20fill-opacity%3D%220%22%2F%3E%3C%2Fsvg%3E\" data-srcset=\"https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Joerg-Amelunxen-200x201.png 200w, https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Joerg-Amelunxen-400x403.png 400w, https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Joerg-Amelunxen-600x604.png 600w, https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Joerg-Amelunxen-800x805.png 800w, https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Joerg-Amelunxen-1200x1208.png 1200w, https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Joerg-Amelunxen.png 1600w\" data-sizes=\"auto\" data-orig-sizes=\"(max-width: 800px) 100vw, 800px\" \/><\/span><\/div><div class=\"fusion-text fusion-text-3 md-text-align-right fusion-text-no-margin\" style=\"--awb-content-alignment:center;--awb-font-size:14px;--awb-text-color:var(--awb-color6);--awb-margin-bottom:24px;\"><p>J\u00f6rg Amelunxen<\/p>\n<\/div><div class=\"fusion-text fusion-text-4 fusion-text-no-margin\" style=\"--awb-text-color:var(--awb-color6);--awb-margin-bottom:24px;\"><p>J\u00f6rg Amelunxen (M.Sc. Computer Science, University of Paderborn) is an AI Solutions Architect at mgm, with a path from lead developer through technical project lead to his current focus on AI- and agent-driven software development. In mgm&#8217;s cross-project AI &amp; Data Engineering team he works on making agentic development production-ready in the enterprise: from multi-agent workflows and RAG systems, to MCP-based agent interfaces, to quality assurance in agent-driven processes. His project experience spans public sector, e-commerce, insurance, and finance. Most recently he designed the end-to-end architecture of a voice agent with its own intent classifier. He enjoys talking through these topics; if you would like to, you can find him easily on LinkedIn.<\/p>\n<\/div><div class=\"fusion-text fusion-text-5 fusion-text-no-margin\" style=\"--awb-content-alignment:center;--awb-font-size:14px;--awb-text-color:var(--awb-color6);--awb-margin-bottom:0px;\"><p>Ver\u00f6ffentlicht: 07.2026<\/p>\n<\/div><\/div><\/div><div class=\"fusion-layout-column fusion_builder_column_inner fusion-builder-nested-column-4 fusion_builder_column_inner_3_5 3_5 fusion-flex-column\" style=\"--awb-bg-size:cover;--awb-width-large:60%;--awb-margin-top-large:0px;--awb-spacing-right-large:15px;--awb-margin-bottom-large:20px;--awb-spacing-left-large:15px;--awb-width-medium:60%;--awb-order-medium:0;--awb-spacing-right-medium:15px;--awb-spacing-left-medium:15px;--awb-width-small:100%;--awb-order-small:0;--awb-spacing-right-small:15px;--awb-spacing-left-small:15px;\"><div class=\"fusion-column-wrapper fusion-column-has-shadow fusion-flex-justify-content-flex-start fusion-content-layout-column\"><div class=\"fusion-title title fusion-title-3 fusion-sep-none fusion-title-text fusion-title-size-three\" style=\"--awb-text-color:var(--awb-color6);--awb-margin-bottom:32px;--awb-margin-top-small:0px;--awb-margin-right-small:0px;--awb-margin-bottom-small:32px;--awb-margin-left-small:0px;--awb-font-size:var(--awb-custom_typography_13-font-size);\"><h3 class=\"fusion-title-heading title-heading-left fusion-responsive-typography-calculated\" style=\"font-family:var(--awb-custom_typography_13-font-family);font-weight:var(--awb-custom_typography_13-font-weight);font-style:var(--awb-custom_typography_13-font-style);margin:0;letter-spacing:var(--awb-custom_typography_13-letter-spacing);text-transform:var(--awb-custom_typography_13-text-transform);font-size:1em;--fontSize:32;line-height:var(--awb-custom_typography_13-line-height);\">## Bottleneck One: Decision Fatigue<\/h3><\/div><div class=\"fusion-text fusion-text-6\" style=\"--awb-text-color:var(--awb-color6);\"><p>In social psychology, decision fatigue has been described since Baumeister and colleagues in 1998 <span class=\"fusion-tooltip tooltip-shortcode\" data-animation=\"\" data-delay=\"0\" data-placement=\"top\" data-title=\"Baumeister, R.F., Bratslavsky, E., Muraven, M., &amp; Tice, D.M. (1998). Ego depletion: Is the active self a limited resource? Journal of Personality and Social Psychology, 74(5), 1252-1265.\" title=\"Baumeister, R.F., Bratslavsky, E., Muraven, M., &amp; Tice, D.M. (1998). Ego depletion: Is the active self a limited resource? Journal of Personality and Social Psychology, 74(5), 1252-1265.\" data-toggle=\"tooltip\" data-trigger=\"hover\">[1]<\/span>. The exact mechanism (ego depletion as a central reservoir of willpower) has been strongly called into question by a multi-lab preregistered replication with 23 labs and over 2000 participants <span class=\"fusion-tooltip tooltip-shortcode\" data-animation=\"\" data-delay=\"0\" data-placement=\"top\" data-title=\"Hagger, M.S., Chatzisarantis, N.L.D., Alberts, H., et al. (2016). A multilab preregistered replication of the ego-depletion effect. Perspectives on Psychological Science, 11(4), 546-573.\" title=\"Hagger, M.S., Chatzisarantis, N.L.D., Alberts, H., et al. (2016). A multilab preregistered replication of the ego-depletion effect. Perspectives on Psychological Science, 11(4), 546-573.\" data-toggle=\"tooltip\" data-trigger=\"hover\">[2]<\/span>. The phenomenon itself has not been: the quality of a decision measurably drops over the course of a sequence of decisions. The most robust evidence comes from a PNAS study of more than 1000 real parole decisions by eight experienced judges, observed over 50 court days <span class=\"fusion-tooltip tooltip-shortcode\" data-animation=\"\" data-delay=\"0\" data-placement=\"top\" data-title=\"Danziger, S., Levav, J., &amp; Avnaim-Pesso, L. (2011). Extraneous factors in judicial decisions. Proceedings of the National Academy of Sciences, 108(17), 6889-6892.\" title=\"Danziger, S., Levav, J., &amp; Avnaim-Pesso, L. (2011). Extraneous factors in judicial decisions. Proceedings of the National Academy of Sciences, 108(17), 6889-6892.\" data-toggle=\"tooltip\" data-trigger=\"hover\">[3]<\/span>. Right after a break or lunch, roughly 65 percent of applications are granted. Over the course of the following session, the rate falls gradually, in individual sessions down to near zero, and jumps back up after the next break. A conceptual review from health psychology confirms that the phenomenon shows up across professions <span class=\"fusion-tooltip tooltip-shortcode\" data-animation=\"\" data-delay=\"0\" data-placement=\"top\" data-title=\"Pignatiello, G.A., Martin, R.J., &amp; Hickman, R.L. (2020). Decision fatigue: A conceptual analysis. Journal of Health Psychology, 25(1), 123-135.\" title=\"Pignatiello, G.A., Martin, R.J., &amp; Hickman, R.L. (2020). Decision fatigue: A conceptual analysis. Journal of Health Psychology, 25(1), 123-135.\" data-toggle=\"tooltip\" data-trigger=\"hover\">[4]<\/span>. Clinical decision research provides a second, independent confirmation: a study of more than 21,000 outpatient visits in primary care practices shows that physicians increasingly prescribe antibiotics inappropriately over the course of each clinic session, peaking in the last hours before the lunch break <span class=\"fusion-tooltip tooltip-shortcode\" data-animation=\"\" data-delay=\"0\" data-placement=\"top\" data-title=\"Linder, J.A., Doctor, J.N., Friedberg, M.W., Reyes Nieva, H., Birks, C., Meeker, D., &amp; Fox, C.R. (2014). Time of day and the decision to prescribe antibiotics. JAMA Internal Medicine, 174(12), 2029-2031.\" title=\"Linder, J.A., Doctor, J.N., Friedberg, M.W., Reyes Nieva, H., Birks, C., Meeker, D., &amp; Fox, C.R. (2014). Time of day and the decision to prescribe antibiotics. JAMA Internal Medicine, 174(12), 2029-2031.\" data-toggle=\"tooltip\" data-trigger=\"hover\">[5]<\/span>. Same mechanism, different context, same direction.<\/p>\n<p>What happens in an agentic setup is the rapid translation of this mechanism into a developer&#8217;s day. Where I used to make maybe eight conscious decisions in two hours (&#8220;do I take this library, do I refactor this method, do I write this test&#8221;), it now feels like forty. In the same amount of time, the agents produce a multiple of the material that needs deciding. Typing was never the exhausting part. Judging is.<\/p>\n<p>For testers this is doubly interesting. Anyone who in the classic setup already held the role of &#8220;reviewer of code someone else wrote&#8221; knows this fatigue already. In the agentic world it is not reinvented but generalized: suddenly everyone working with agents is constantly in the reviewer position. What used to be a time-boxed task in classic code review becomes a permanent state.<\/p>\n<p>The everyday effect sounds harmless. You start clicking &#8220;OK&#8221; faster. You skim diffs you should really be reading line by line. You accept architecture proposals where yesterday you would still have asked &#8220;why, actually&#8221;. Nobody says it out loud. But almost everyone who works honestly with the setup knows that moment in the late afternoon.<\/p>\n<\/div><div class=\"fusion-image-element\" style=\"--awb-margin-bottom:32px;--awb-caption-title-font-family:var(--h2_typography-font-family);--awb-caption-title-font-weight:var(--h2_typography-font-weight);--awb-caption-title-font-style:var(--h2_typography-font-style);--awb-caption-title-size:var(--h2_typography-font-size);--awb-caption-title-transform:var(--h2_typography-text-transform);--awb-caption-title-line-height:var(--h2_typography-line-height);--awb-caption-title-letter-spacing:var(--h2_typography-letter-spacing);\"><span class=\" fusion-imageframe imageframe-none imageframe-2 hover-type-none\"><img decoding=\"async\" width=\"2560\" height=\"1440\" alt=\"Decision Fatigue\" title=\"Decision Fatigue\" src=\"https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Post08-IMG01-scaled.png\" data-orig-src=\"https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Post08-IMG01-scaled.png\" class=\"lazyload img-responsive wp-image-19988\" srcset=\"data:image\/svg+xml,%3Csvg%20xmlns%3D%27http%3A%2F%2Fwww.w3.org%2F2000%2Fsvg%27%20width%3D%272560%27%20height%3D%271440%27%20viewBox%3D%270%200%202560%201440%27%3E%3Crect%20width%3D%272560%27%20height%3D%271440%27%20fill-opacity%3D%220%22%2F%3E%3C%2Fsvg%3E\" data-srcset=\"https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Post08-IMG01-200x113.png 200w, https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Post08-IMG01-400x225.png 400w, https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Post08-IMG01-600x338.png 600w, https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Post08-IMG01-800x450.png 800w, https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Post08-IMG01-1200x675.png 1200w, https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Post08-IMG01-scaled.png 2560w\" data-sizes=\"auto\" data-orig-sizes=\"(max-width: 800px) 100vw, 1200px\" \/><\/span><\/div><div class=\"fusion-separator fusion-full-width-sep\" style=\"align-self: center;margin-left: auto;margin-right: auto;margin-top:12px;margin-bottom:32px;width:100%;\"><div class=\"fusion-separator-border sep-single sep-solid\" style=\"--awb-height:20px;--awb-amount:20px;--awb-sep-color:rgba(40,160,220,0.5);border-color:rgba(40,160,220,0.5);border-top-width:1px;\"><\/div><\/div><div class=\"fusion-title title fusion-title-4 fusion-sep-none fusion-title-text fusion-title-size-three\" style=\"--awb-text-color:var(--awb-color6);--awb-margin-bottom:32px;--awb-margin-top-small:0px;--awb-margin-right-small:0px;--awb-margin-bottom-small:32px;--awb-margin-left-small:0px;--awb-font-size:var(--awb-custom_typography_13-font-size);\"><h3 class=\"fusion-title-heading title-heading-left fusion-responsive-typography-calculated\" style=\"font-family:var(--awb-custom_typography_13-font-family);font-weight:var(--awb-custom_typography_13-font-weight);font-style:var(--awb-custom_typography_13-font-style);margin:0;letter-spacing:var(--awb-custom_typography_13-letter-spacing);text-transform:var(--awb-custom_typography_13-text-transform);font-size:1em;--fontSize:32;line-height:var(--awb-custom_typography_13-line-height);\">## Bottleneck Two: Context Switching<\/h3><\/div><div class=\"fusion-text fusion-text-7\" style=\"--awb-text-color:var(--awb-color6);\"><p>The second cost factor is context switching. The cost of context switches has been folklore in software engineering practice for decades and is well documented in CHI research. A field study of knowledge workers finds an average of roughly 25 minutes until an interrupted task is resumed, typically after more than two other tasks in between <span class=\"fusion-tooltip tooltip-shortcode\" data-animation=\"\" data-delay=\"0\" data-placement=\"top\" data-title=\"Mark, G., Gonzalez, V.M., &amp; Harris, J. (2005). No task left behind? Examining the nature of fragmented work. In Proceedings of the SIGCHI Conference on Human Factors in Computing Systems (CHI \u201905), 321-330.\" title=\"Mark, G., Gonzalez, V.M., &amp; Harris, J. (2005). No task left behind? Examining the nature of fragmented work. In Proceedings of the SIGCHI Conference on Human Factors in Computing Systems (CHI \u201905), 321-330.\" data-toggle=\"tooltip\" data-trigger=\"hover\">[6]<\/span>. A follow-up study shows that interrupted tasks do get finished measurably faster, but at a measurable price of more stress, frustration, and effort <span class=\"fusion-tooltip tooltip-shortcode\" data-animation=\"\" data-delay=\"0\" data-placement=\"top\" data-title=\"Mark, G., Gudith, D., &amp; Klocke, U. (2008). The cost of interrupted work: more speed and stress. In Proceedings of the SIGCHI Conference on Human Factors in Computing Systems (CHI \u201908), 107-110.\" title=\"Mark, G., Gudith, D., &amp; Klocke, U. (2008). The cost of interrupted work: more speed and stress. In Proceedings of the SIGCHI Conference on Human Factors in Computing Systems (CHI \u201908), 107-110.\" data-toggle=\"tooltip\" data-trigger=\"hover\">[7]<\/span>. For programming work specifically, the picture is even clearer. Parnin &amp; Rugaber in 2011 empirically examine how developers find their way back into the code after an interruption, and find that they do not simply pick up where they left off; they visibly invest time in recovery strategies such as code notes, deliberately left compile errors as bookmarks, or rereading the last commit, because the mental stack is deeper than in generic knowledge work <span class=\"fusion-tooltip tooltip-shortcode\" data-animation=\"\" data-delay=\"0\" data-placement=\"top\" data-title=\"Parnin, C., &amp; Rugaber, S. (2011). Resumption strategies for interrupted programming tasks. Software Quality Journal, 19(1), 5-34.\" title=\"Parnin, C., &amp; Rugaber, S. (2011). Resumption strategies for interrupted programming tasks. Software Quality Journal, 19(1), 5-34.\" data-toggle=\"tooltip\" data-trigger=\"hover\">[8]<\/span>.<\/p>\n<p>In agentic supervision, the context switch does not sit between the hours but inside the block of work. You jump from the refactor review to the E2E diagnosis to the architecture decision to the logging detail. Four different contexts, four different mental models, all in thirty minutes. Every jump costs setup time. And because the agents work in parallel and do not wait, the temptation is strong to take each one as it comes in.<\/p>\n<p>What helps is classic engineering discipline in new clothes: class-batched reviews. All refactor reviews in one block. All architecture decisions in a second. Test diagnoses in a third. This sounds obvious and is not; the agents behave like an open-plan office where someone is constantly standing at your desk.<\/p>\n<p>The trick is not to run more agents at once. The trick is that only one kind of question reaches you at a time.<\/p>\n<\/div><div class=\"fusion-separator fusion-full-width-sep\" style=\"align-self: center;margin-left: auto;margin-right: auto;margin-top:12px;margin-bottom:32px;width:100%;\"><div class=\"fusion-separator-border sep-single sep-solid\" style=\"--awb-height:20px;--awb-amount:20px;--awb-sep-color:rgba(40,160,220,0.5);border-color:rgba(40,160,220,0.5);border-top-width:1px;\"><\/div><\/div><div class=\"fusion-title title fusion-title-5 fusion-sep-none fusion-title-text fusion-title-size-three\" style=\"--awb-text-color:var(--awb-color6);--awb-margin-bottom:32px;--awb-margin-top-small:0px;--awb-margin-right-small:0px;--awb-margin-bottom-small:32px;--awb-margin-left-small:0px;--awb-font-size:var(--awb-custom_typography_13-font-size);\"><h3 class=\"fusion-title-heading title-heading-left fusion-responsive-typography-calculated\" style=\"font-family:var(--awb-custom_typography_13-font-family);font-weight:var(--awb-custom_typography_13-font-weight);font-style:var(--awb-custom_typography_13-font-style);margin:0;letter-spacing:var(--awb-custom_typography_13-letter-spacing);text-transform:var(--awb-custom_typography_13-text-transform);font-size:1em;--fontSize:32;line-height:var(--awb-custom_typography_13-line-height);\">## Bottleneck Three: Attention Span<\/h3><\/div><div class=\"fusion-text fusion-text-8\" style=\"--awb-text-color:var(--awb-color6);\"><p>The third bottleneck is the most insidious, because it does not feel like fatigue but like routine. Attention span under continuous review.<\/p>\n<p>Cognitive psychology knows the phenomenon as the vigilance decrement, documented since Mackworth&#8217;s classic study of radar monitoring in the Second World War <span class=\"fusion-tooltip tooltip-shortcode\" data-animation=\"\" data-delay=\"0\" data-placement=\"top\" data-title=\"Mackworth, N.H. (1948). The breakdown of vigilance during prolonged visual search. Quarterly Journal of Experimental Psychology, 1(1), 6-21.\" title=\"Mackworth, N.H. (1948). The breakdown of vigilance during prolonged visual search. Quarterly Journal of Experimental Psychology, 1(1), 6-21.\" data-toggle=\"tooltip\" data-trigger=\"hover\">[9]<\/span>: someone who has to attentively monitor something over a long time, where an error appears only rarely, loses the ability to detect it. In the lab, Mackworth found a 10 to 15 percent drop in detection rate within the first thirty minutes, with further deterioration afterward. A much-cited review from human factors carries the pattern over to modern automation: when a human monitors an automated pipeline while also having their own tasks, complacency rises, and simple practice does not help against it <span class=\"fusion-tooltip tooltip-shortcode\" data-animation=\"\" data-delay=\"0\" data-placement=\"top\" data-title=\"Parasuraman, R., &amp; Manzey, D.H. (2010). Complacency and bias in human use of automation: An attentional integration. Human Factors, 52(3), 381-410.\" title=\"Parasuraman, R., &amp; Manzey, D.H. (2010). Complacency and bias in human use of automation: An attentional integration. Human Factors, 52(3), 381-410.\" data-toggle=\"tooltip\" data-trigger=\"hover\">[10]<\/span>. More recent research specifically on human oversight of AI suggestions shows the pattern in its present-day form. Vasconcelos and colleagues demonstrate empirically in 2023 that users tend to accept AI suggestions without checking them themselves as soon as they perceive the cognitive effort of verification as too high; even well-constructed explanations only partly reduce this over-reliance, because the human runs a simple cost-benefit calculation <span class=\"fusion-tooltip tooltip-shortcode\" data-animation=\"\" data-delay=\"0\" data-placement=\"top\" data-title=\"Vasconcelos, H., J\u00f6rke, M., Grunde-McLaughlin, M., Gerstenberg, T., Bernstein, M.S., &amp; Krishna, R. (2023). Explanations can reduce overreliance on AI systems during decision-making. Proceedings of the ACM on Human-Computer Interaction, 7(CSCW1), 1-38.\" title=\"Vasconcelos, H., J\u00f6rke, M., Grunde-McLaughlin, M., Gerstenberg, T., Bernstein, M.S., &amp; Krishna, R. (2023). Explanations can reduce overreliance on AI systems during decision-making. Proceedings of the ACM on Human-Computer Interaction, 7(CSCW1), 1-38.\" data-toggle=\"tooltip\" data-trigger=\"hover\">[11]<\/span>. In agentic supervision this is exactly the situation: many suggestions, each individually plausible, real verification more expensive than acceptance.<\/p>\n<p>In practice this means: after thirty minutes of continuous review work, the detection rate drops. You stop seeing bugs you would have caught immediately when fresh. And because the agents keep producing, this effect adds up fast.<\/p>\n<p>From my own practice, a single observation, not a study finding: when I check agent output for longer than half an hour at a stretch, I find things on rereading after a break that I did not see on the first pass. This is the same mechanism that code review research has known since the beginnings of formal inspections; as early as 1976, Fagan recommends limiting an inspection session to at most two hours and reviewing no more than 150 lines per hour, because otherwise the defect detection rate drops significantly <span class=\"fusion-tooltip tooltip-shortcode\" data-animation=\"\" data-delay=\"0\" data-placement=\"top\" data-title=\"Fagan, M.E. (1976). Design and code inspections to reduce errors in program development. IBM Systems Journal, 15(3), 182-211.\" title=\"Fagan, M.E. (1976). Design and code inspections to reduce errors in program development. IBM Systems Journal, 15(3), 182-211.\" data-toggle=\"tooltip\" data-trigger=\"hover\">[12]<\/span>. A modern empirical study of code review practice at Microsoft extends the picture: in today&#8217;s asynchronous practice, reviews revolve less around pure defect detection than expected; a large share of the documented reviewer comments aims at code understanding and knowledge transfer within the team <span class=\"fusion-tooltip tooltip-shortcode\" data-animation=\"\" data-delay=\"0\" data-placement=\"top\" data-title=\"Bacchelli, A., &amp; Bird, C. (2013). Expectations, outcomes, and challenges of modern code review. In Proceedings of the 2013 International Conference on Software Engineering (ICSE \u201913), 712-721.\" title=\"Bacchelli, A., &amp; Bird, C. (2013). Expectations, outcomes, and challenges of modern code review. In Proceedings of the 2013 International Conference on Software Engineering (ICSE \u201913), 712-721.\" data-toggle=\"tooltip\" data-trigger=\"hover\">[13]<\/span>. That does not make the attention demand on the reviewer smaller, it shifts it. In the agentic world the same effect just becomes permanent instead of situational.<\/p>\n<\/div><div class=\"fusion-separator fusion-full-width-sep\" style=\"align-self: center;margin-left: auto;margin-right: auto;margin-top:12px;margin-bottom:32px;width:100%;\"><div class=\"fusion-separator-border sep-single sep-solid\" style=\"--awb-height:20px;--awb-amount:20px;--awb-sep-color:rgba(40,160,220,0.5);border-color:rgba(40,160,220,0.5);border-top-width:1px;\"><\/div><\/div><div class=\"fusion-title title fusion-title-6 fusion-sep-none fusion-title-text fusion-title-size-three\" style=\"--awb-text-color:var(--awb-color6);--awb-margin-bottom:32px;--awb-margin-top-small:0px;--awb-margin-right-small:0px;--awb-margin-bottom-small:32px;--awb-margin-left-small:0px;--awb-font-size:var(--awb-custom_typography_13-font-size);\"><h3 class=\"fusion-title-heading title-heading-left fusion-responsive-typography-calculated\" style=\"font-family:var(--awb-custom_typography_13-font-family);font-weight:var(--awb-custom_typography_13-font-weight);font-style:var(--awb-custom_typography_13-font-style);margin:0;letter-spacing:var(--awb-custom_typography_13-letter-spacing);text-transform:var(--awb-custom_typography_13-text-transform);font-size:1em;--fontSize:32;line-height:var(--awb-custom_typography_13-line-height);\">## What the Setup Needs Against This<\/h3><\/div><div class=\"fusion-text fusion-text-9\" style=\"--awb-text-color:var(--awb-color6);\"><p>When the three bottlenecks come together, they form a pattern I have come to take seriously. Three levers that work in my daily practice:<\/p>\n<p>First, a decision-budget logic. A long-term study of 27 CEOs of large companies, tracked around the clock over 13 weeks, shows the pattern cleanly: top executives spend the largest part of their working time in meetings, where the main activities are prioritizing, sorting, and handing off; the real strategic decisions of their own are a clear minority <span class=\"fusion-tooltip tooltip-shortcode\" data-animation=\"\" data-delay=\"0\" data-placement=\"top\" data-title=\"Porter, M.E., &amp; Nohria, N. (2018). How CEOs manage time. Harvard Business Review, 96(4), 42-51.\" title=\"Porter, M.E., &amp; Nohria, N. (2018). How CEOs manage time. Harvard Business Review, 96(4), 42-51.\" data-toggle=\"tooltip\" data-trigger=\"hover\">[14]<\/span>. In an agentic setup, by contrast, you first decide everything yourself. You have to correct that explicitly: which decisions are reversible and cheap (a code comment, a small refactor in an internal library) and pass without review? Which are reversible and expensive (a larger refactor, an architecture detail)? Which are not reversible (a schema change in production, an external interface)? The first category moves to auto-approve, the third stays strictly with the human.<\/p>\n<p>Second, lead in groups instead of darting between individual agents. Three fixed slots a day are the clean ideal and in practice rarely feasible; in a warm agent fleet you have to hook in somewhere all the time. What does work, once you are no longer running just four agents in parallel but rather twelve or sixteen, is bundling them into thematic groups with their own focus. One group works on the backend, one on the tests, one on documentation and small bug fixes. You actively devote yourself to one group at a time, instead of jumping between the agents of different groups.<\/p>\n<p>This works like a dog walker with several groups of dogs at once. He actively leads one group; the others stand around, look about, or run toward the goal where they are already playing anyway. When he switches, he switches at the group level, not between individual dogs of two parallel groups. That is exactly the difference between functioning parallel supervision and chronic darting about.<\/p>\n<p>Third, a thirty-minute cap on continuous review work. With a break afterward that is not a screen again. Movement, coffee, staring at a wall. The vigilance research is clear enough that you do not need to optimize further.<\/p>\n<\/div><div class=\"fusion-separator fusion-full-width-sep\" style=\"align-self: center;margin-left: auto;margin-right: auto;margin-top:12px;margin-bottom:32px;width:100%;\"><div class=\"fusion-separator-border sep-single sep-solid\" style=\"--awb-height:20px;--awb-amount:20px;--awb-sep-color:rgba(40,160,220,0.5);border-color:rgba(40,160,220,0.5);border-top-width:1px;\"><\/div><\/div><div class=\"fusion-title title fusion-title-7 fusion-sep-none fusion-title-text fusion-title-size-three\" style=\"--awb-text-color:var(--awb-color6);--awb-margin-bottom:32px;--awb-margin-top-small:0px;--awb-margin-right-small:0px;--awb-margin-bottom-small:32px;--awb-margin-left-small:0px;--awb-font-size:var(--awb-custom_typography_13-font-size);\"><h3 class=\"fusion-title-heading title-heading-left fusion-responsive-typography-calculated\" style=\"font-family:var(--awb-custom_typography_13-font-family);font-weight:var(--awb-custom_typography_13-font-weight);font-style:var(--awb-custom_typography_13-font-style);margin:0;letter-spacing:var(--awb-custom_typography_13-letter-spacing);text-transform:var(--awb-custom_typography_13-text-transform);font-size:1em;--fontSize:32;line-height:var(--awb-custom_typography_13-line-height);\">## The Question That Counts in the End<\/h3><\/div><div class=\"fusion-text fusion-text-10\" style=\"--awb-text-color:var(--awb-color6);\"><p>Anyone who measures agentic development only along the axis of &#8220;faster and cheaper&#8221; is measuring the wrong thing. Speed has become a given, not a differentiator. What sets software apart today is no longer that it gets shipped, but that it is unique. That it does something that is not also sitting in every second repository. That the human who is responsible for it is genuinely still in the loop, with the three bottlenecks above under control.<\/p>\n<p>For the testing community, I think, this means two things.<\/p>\n<p>One: the role of the tester does not disappear in an agentic world. It becomes more central. What the agents produce has to be checked by someone who holds the context. That is exactly the skill good testers have always had. It is just now being called upon under new conditions.<\/p>\n<p>The other: the three bottlenecks above apply especially to people in review roles. Whoever takes the setup seriously builds guardrails. Whoever does not runs the risk of scrapping the AI in half a year with the argument, &#8220;it was not for us&#8221;. The difference does not lie in the AI. It lies in the architecture around it.<\/p>\n<\/div><div class=\"fusion-separator fusion-full-width-sep\" style=\"align-self: center;margin-left: auto;margin-right: auto;margin-top:12px;margin-bottom:32px;width:100%;\"><div class=\"fusion-separator-border sep-single sep-solid\" style=\"--awb-height:20px;--awb-amount:20px;--awb-sep-color:rgba(40,160,220,0.5);border-color:rgba(40,160,220,0.5);border-top-width:1px;\"><\/div><\/div><\/div><\/div><div class=\"fusion-layout-column fusion_builder_column_inner fusion-builder-nested-column-5 fusion_builder_column_inner_2_5 2_5 fusion-flex-column\" style=\"--awb-bg-size:cover;--awb-width-large:40%;--awb-margin-top-large:0px;--awb-spacing-right-large:15px;--awb-margin-bottom-large:20px;--awb-spacing-left-large:15px;--awb-width-medium:40%;--awb-order-medium:0;--awb-spacing-right-medium:15px;--awb-spacing-left-medium:15px;--awb-width-small:100%;--awb-order-small:0;--awb-spacing-right-small:15px;--awb-spacing-left-small:15px;\"><div class=\"fusion-column-wrapper fusion-column-has-shadow fusion-flex-justify-content-flex-start fusion-content-layout-column\"><\/div><\/div><div class=\"fusion-layout-column fusion_builder_column_inner fusion-builder-nested-column-6 fusion_builder_column_inner_3_5 3_5 fusion-flex-column blue-box\" style=\"--awb-padding-top:40px;--awb-padding-right:32px;--awb-padding-bottom:40px;--awb-padding-left:32px;--awb-padding-top-medium:32px;--awb-padding-right-medium:32px;--awb-padding-bottom-medium:32px;--awb-padding-left-medium:32px;--awb-padding-top-small:24px;--awb-padding-right-small:24px;--awb-padding-bottom-small:24px;--awb-padding-left-small:24px;--awb-bg-color:#28A0DC33;--awb-bg-color-hover:#28A0DC33;--awb-bg-size:cover;--awb-width-large:60%;--awb-margin-top-large:0px;--awb-spacing-right-large:15px;--awb-margin-bottom-large:20px;--awb-spacing-left-large:15px;--awb-width-medium:60%;--awb-order-medium:0;--awb-spacing-right-medium:15px;--awb-spacing-left-medium:15px;--awb-width-small:100%;--awb-order-small:0;--awb-spacing-right-small:15px;--awb-spacing-left-small:15px;\" data-scroll-devices=\"small-visibility,medium-visibility,large-visibility\"><div class=\"fusion-column-wrapper fusion-column-has-shadow fusion-flex-justify-content-flex-start fusion-content-layout-column\"><div class=\"fusion-title title fusion-title-8 fusion-sep-none fusion-title-text fusion-title-size-three\" style=\"--awb-text-color:var(--awb-color6);--awb-margin-bottom:32px;--awb-margin-top-small:0px;--awb-margin-right-small:0px;--awb-margin-bottom-small:32px;--awb-margin-left-small:0px;--awb-font-size:64px;\"><h3 class=\"fusion-title-heading title-heading-left fusion-responsive-typography-calculated\" style=\"font-family:&quot;Bai Jamjuree&quot;;font-style:normal;font-weight:700;margin:0;letter-spacing:var(--awb-custom_typography_13-letter-spacing);text-transform:var(--awb-custom_typography_13-text-transform);font-size:1em;--fontSize:64;line-height:var(--awb-custom_typography_13-line-height);\">## References<\/h3><\/div><div class=\"fusion-text fusion-text-11 simple-table\" style=\"--awb-text-color:var(--awb-color6);\"><table>\n<tbody>\n<tr>\n<td>[1]<\/td>\n<td>Baumeister, R.F., Bratslavsky, E., Muraven, M., &amp; Tice, D.M. (1998). Ego depletion: Is the active self a limited resource? Journal of Personality and Social Psychology, 74(5), 1252-1265. <a href=\"https:\/\/doi.org\/10.1037\/0022-3514.74.5.1252\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1037\/0022-3514.74.5.1252<\/a><\/td>\n<\/tr>\n<tr>\n<td>[2]<\/td>\n<td>Hagger, M.S., Chatzisarantis, N.L.D., Alberts, H., et al. (2016). A multilab preregistered replication of the ego-depletion effect. Perspectives on Psychological Science, 11(4), 546-573. <a href=\"https:\/\/doi.org\/10.1177\/1745691616652873\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1177\/1745691616652873<\/a><\/td>\n<\/tr>\n<tr>\n<td>[3]<\/td>\n<td>Danziger, S., Levav, J., &amp; Avnaim-Pesso, L. (2011). Extraneous factors in judicial decisions. Proceedings of the National Academy of Sciences, 108(17), 6889-6892. <a href=\"https:\/\/doi.org\/10.1073\/pnas.1018033108\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1073\/pnas.1018033108<\/a><\/td>\n<\/tr>\n<tr>\n<td>[4]<\/td>\n<td>Pignatiello, G.A., Martin, R.J., &amp; Hickman, R.L. (2020). Decision fatigue: A conceptual analysis. Journal of Health Psychology, 25(1), 123-135. <a href=\"https:\/\/doi.org\/10.1177\/1359105318763510\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1177\/1359105318763510<\/a><\/td>\n<\/tr>\n<tr>\n<td>[5]<\/td>\n<td>Linder, J.A., Doctor, J.N., Friedberg, M.W., Reyes Nieva, H., Birks, C., Meeker, D., &amp; Fox, C.R. (2014). Time of day and the decision to prescribe antibiotics. JAMA Internal Medicine, 174(12), 2029-2031. <a href=\"https:\/\/doi.org\/10.1001\/jamainternmed.2014.5225\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1001\/jamainternmed.2014.5225<\/a><\/td>\n<\/tr>\n<tr>\n<td>[6]<\/td>\n<td>Mark, G., Gonzalez, V.M., &amp; Harris, J. (2005). No task left behind? Examining the nature of fragmented work. In Proceedings of the SIGCHI Conference on Human Factors in Computing Systems (CHI &#8217;05), 321-330. <a href=\"https:\/\/doi.org\/10.1145\/1054972.1055017\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1145\/1054972.1055017<\/a><\/td>\n<\/tr>\n<tr>\n<td>[7]<\/td>\n<td>Mark, G., Gudith, D., &amp; Klocke, U. (2008). The cost of interrupted work: more speed and stress. In Proceedings of the SIGCHI Conference on Human Factors in Computing Systems (CHI &#8217;08), 107-110. <a href=\"https:\/\/doi.org\/10.1145\/1357054.1357072\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1145\/1357054.1357072<\/a><\/td>\n<\/tr>\n<tr>\n<td>[8]<\/td>\n<td>Parnin, C., &amp; Rugaber, S. (2011). Resumption strategies for interrupted programming tasks. Software Quality Journal, 19(1), 5-34. <a href=\"https:\/\/doi.org\/10.1007\/s11219-010-9104-9\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1007\/s11219-010-9104-9<\/a><\/td>\n<\/tr>\n<tr>\n<td>[9]<\/td>\n<td>Mackworth, N.H. (1948). The breakdown of vigilance during prolonged visual search. Quarterly Journal of Experimental Psychology, 1(1), 6-21. <a href=\"https:\/\/doi.org\/10.1080\/17470214808416738\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1080\/17470214808416738<\/a><\/td>\n<\/tr>\n<tr>\n<td>[10]<\/td>\n<td>Parasuraman, R., &amp; Manzey, D.H. (2010). Complacency and bias in human use of automation: An attentional integration. Human Factors, 52(3), 381-410. <a href=\"https:\/\/doi.org\/10.1177\/0018720810376055\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1177\/0018720810376055<\/a><\/td>\n<\/tr>\n<tr>\n<td>[11]<\/td>\n<td>Vasconcelos, H., J\u00f6rke, M., Grunde-McLaughlin, M., Gerstenberg, T., Bernstein, M.S., &amp; Krishna, R. (2023). Explanations can reduce overreliance on AI systems during decision-making. Proceedings of the ACM on Human-Computer Interaction, 7(CSCW1), 1-38. <a href=\"https:\/\/doi.org\/10.1145\/3579605\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1145\/3579605<\/a><\/td>\n<\/tr>\n<tr>\n<td>[12]<\/td>\n<td>Fagan, M.E. (1976). Design and code inspections to reduce errors in program development. IBM Systems Journal, 15(3), 182-211. <a href=\"https:\/\/doi.org\/10.1147\/sj.153.0182\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1147\/sj.153.0182<\/a><\/td>\n<\/tr>\n<tr>\n<td>[13]<\/td>\n<td>Bacchelli, A., &amp; Bird, C. (2013). Expectations, outcomes, and challenges of modern code review. In Proceedings of the 2013 International Conference on Software Engineering (ICSE &#8217;13), 712-721. <a href=\"https:\/\/doi.org\/10.1109\/ICSE.2013.6606617\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1109\/ICSE.2013.6606617<\/a><\/td>\n<\/tr>\n<tr>\n<td>[14]<\/td>\n<td>Porter, M.E., &amp; Nohria, N. (2018). How CEOs manage time. Harvard Business Review, 96(4), 42-51. <a href=\"https:\/\/hbr.org\/2018\/07\/how-ceos-manage-time\" target=\"_blank\" rel=\"noopener\">https:\/\/hbr.org\/2018\/07\/how-ceos-manage-time<\/a><\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<\/div><\/div><\/div><\/div><\/div><\/div><\/div><\/div>\n","protected":false},"excerpt":{"rendered":"","protected":false},"author":6,"featured_media":19989,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"_acf_changed":false,"footnotes":""},"categories":[81],"tags":[91,84],"class_list":["post-19990","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-blog","tag-quality-engineering","tag-test"],"acf":[],"yoast_head":"<!-- This site is optimized with the Yoast SEO plugin v28.1 - https:\/\/yoast.com\/product\/yoast-seo-wordpress\/ -->\n<title># When the Agents Write the Code, the Human Checks Themselves [2\/2] - German Testing Board<\/title>\n<meta name=\"description\" content=\"Decision fatigue, context switching, and fading attention: three cognitive bottlenecks of supervising agents, with evidence and concrete countermeasures.\" \/>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/\" \/>\n<meta property=\"og:locale\" content=\"en_US\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"# When the Agents Write the Code, the Human Checks Themselves [2\/2] - German Testing Board\" \/>\n<meta property=\"og:description\" content=\"Decision fatigue, context switching, and fading attention: three cognitive bottlenecks of supervising agents, with evidence and concrete countermeasures.\" \/>\n<meta property=\"og:url\" content=\"https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/\" \/>\n<meta property=\"og:site_name\" content=\"German Testing Board\" \/>\n<meta property=\"article:published_time\" content=\"2026-07-14T06:00:53+00:00\" \/>\n<meta property=\"article:modified_time\" content=\"2026-07-14T15:28:21+00:00\" \/>\n<meta property=\"og:image\" content=\"https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Heroimage-Post08-1024x462.png\" \/>\n\t<meta property=\"og:image:width\" content=\"1024\" \/>\n\t<meta property=\"og:image:height\" content=\"462\" \/>\n\t<meta property=\"og:image:type\" content=\"image\/png\" \/>\n<meta name=\"author\" content=\"Dr. Armin Metzger\" \/>\n<meta name=\"twitter:card\" content=\"summary_large_image\" \/>\n<meta name=\"twitter:creator\" content=\"@GTB_ISTQB\" \/>\n<meta name=\"twitter:site\" content=\"@GTB_ISTQB\" \/>\n<meta name=\"twitter:label1\" content=\"Written by\" \/>\n\t<meta name=\"twitter:data1\" content=\"Dr. Armin Metzger\" \/>\n\t<meta name=\"twitter:label2\" content=\"Est. reading time\" \/>\n\t<meta name=\"twitter:data2\" content=\"12 minutes\" \/>\n<script type=\"application\/ld+json\" class=\"yoast-schema-graph\">{\"@context\":\"https:\\\/\\\/schema.org\",\"@graph\":[{\"@type\":\"Article\",\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/blog\\\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\\\/#article\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/blog\\\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\\\/\"},\"author\":{\"name\":\"Dr. Armin Metzger\",\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/#\\\/schema\\\/person\\\/580154dcda34782d68a1503d906ba00a\"},\"headline\":\"# When the Agents Write the Code, the Human Checks Themselves [2\\\/2]\",\"datePublished\":\"2026-07-14T06:00:53+00:00\",\"dateModified\":\"2026-07-14T15:28:21+00:00\",\"mainEntityOfPage\":{\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/blog\\\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\\\/\"},\"wordCount\":12411,\"commentCount\":0,\"publisher\":{\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/#organization\"},\"image\":{\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/blog\\\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\\\/#primaryimage\"},\"thumbnailUrl\":\"https:\\\/\\\/www.gtb.de\\\/wp-content\\\/uploads\\\/2026\\\/07\\\/Heroimage-Post08.png\",\"keywords\":[\"Quality Engineering\",\"Test\"],\"articleSection\":[\"Blog\"],\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"CommentAction\",\"name\":\"Comment\",\"target\":[\"https:\\\/\\\/www.gtb.de\\\/en\\\/blog\\\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\\\/#respond\"]}]},{\"@type\":\"WebPage\",\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/blog\\\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\\\/\",\"url\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/blog\\\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\\\/\",\"name\":\"# When the Agents Write the Code, the Human Checks Themselves [2\\\/2] - German Testing Board\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/#website\"},\"primaryImageOfPage\":{\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/blog\\\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\\\/#primaryimage\"},\"image\":{\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/blog\\\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\\\/#primaryimage\"},\"thumbnailUrl\":\"https:\\\/\\\/www.gtb.de\\\/wp-content\\\/uploads\\\/2026\\\/07\\\/Heroimage-Post08.png\",\"datePublished\":\"2026-07-14T06:00:53+00:00\",\"dateModified\":\"2026-07-14T15:28:21+00:00\",\"description\":\"Decision fatigue, context switching, and fading attention: three cognitive bottlenecks of supervising agents, with evidence and concrete countermeasures.\",\"breadcrumb\":{\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/blog\\\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\\\/#breadcrumb\"},\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"ReadAction\",\"target\":[\"https:\\\/\\\/www.gtb.de\\\/en\\\/blog\\\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\\\/\"]}]},{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/blog\\\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\\\/#primaryimage\",\"url\":\"https:\\\/\\\/www.gtb.de\\\/wp-content\\\/uploads\\\/2026\\\/07\\\/Heroimage-Post08.png\",\"contentUrl\":\"https:\\\/\\\/www.gtb.de\\\/wp-content\\\/uploads\\\/2026\\\/07\\\/Heroimage-Post08.png\",\"width\":2392,\"height\":1080,\"caption\":\"Wenn die Agenten den Code schreiben, pr\u00fcft der Mensch sich selbst\"},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/blog\\\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\\\/#breadcrumb\",\"itemListElement\":[{\"@type\":\"ListItem\",\"position\":1,\"name\":\"Home\",\"item\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/\"},{\"@type\":\"ListItem\",\"position\":2,\"name\":\"# When the Agents Write the Code, the Human Checks Themselves [2\\\/2]\"}]},{\"@type\":\"WebSite\",\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/#website\",\"url\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/\",\"name\":\"German Testing Board\",\"description\":\"Software.Testing.Excellence\",\"publisher\":{\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/#organization\"},\"potentialAction\":[{\"@type\":\"SearchAction\",\"target\":{\"@type\":\"EntryPoint\",\"urlTemplate\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/?s={search_term_string}\"},\"query-input\":{\"@type\":\"PropertyValueSpecification\",\"valueRequired\":true,\"valueName\":\"search_term_string\"}}],\"inLanguage\":\"en-US\"},{\"@type\":\"Organization\",\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/#organization\",\"name\":\"German Testing Board e. V.\",\"url\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/\",\"logo\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/#\\\/schema\\\/logo\\\/image\\\/\",\"url\":\"https:\\\/\\\/www.gtb.de\\\/wp-content\\\/uploads\\\/2023\\\/10\\\/gtb-logo.png\",\"contentUrl\":\"https:\\\/\\\/www.gtb.de\\\/wp-content\\\/uploads\\\/2023\\\/10\\\/gtb-logo.png\",\"width\":224,\"height\":183,\"caption\":\"German Testing Board e. V.\"},\"image\":{\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/#\\\/schema\\\/logo\\\/image\\\/\"},\"sameAs\":[\"https:\\\/\\\/x.com\\\/GTB_ISTQB\",\"https:\\\/\\\/de.linkedin.com\\\/company\\\/german-testing-board\"]},{\"@type\":\"Person\",\"@id\":\"https:\\\/\\\/www.gtb.de\\\/en\\\/#\\\/schema\\\/person\\\/580154dcda34782d68a1503d906ba00a\",\"name\":\"Dr. Armin Metzger\",\"image\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/70d5768f24f60d4aa0012915fb37c2937223f49270944aa45222f8a20583d028?s=96&d=mm&r=g\",\"url\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/70d5768f24f60d4aa0012915fb37c2937223f49270944aa45222f8a20583d028?s=96&d=mm&r=g\",\"contentUrl\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/70d5768f24f60d4aa0012915fb37c2937223f49270944aa45222f8a20583d028?s=96&d=mm&r=g\",\"caption\":\"Dr. Armin Metzger\"}}]}<\/script>\n<!-- \/ Yoast SEO plugin. -->","yoast_head_json":{"title":"# When the Agents Write the Code, the Human Checks Themselves [2\/2] - German Testing Board","description":"Decision fatigue, context switching, and fading attention: three cognitive bottlenecks of supervising agents, with evidence and concrete countermeasures.","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/","og_locale":"en_US","og_type":"article","og_title":"# When the Agents Write the Code, the Human Checks Themselves [2\/2] - German Testing Board","og_description":"Decision fatigue, context switching, and fading attention: three cognitive bottlenecks of supervising agents, with evidence and concrete countermeasures.","og_url":"https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/","og_site_name":"German Testing Board","article_published_time":"2026-07-14T06:00:53+00:00","article_modified_time":"2026-07-14T15:28:21+00:00","og_image":[{"width":1024,"height":462,"url":"https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Heroimage-Post08-1024x462.png","type":"image\/png"}],"author":"Dr. Armin Metzger","twitter_card":"summary_large_image","twitter_creator":"@GTB_ISTQB","twitter_site":"@GTB_ISTQB","twitter_misc":{"Written by":"Dr. Armin Metzger","Est. reading time":"12 minutes"},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"Article","@id":"https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/#article","isPartOf":{"@id":"https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/"},"author":{"name":"Dr. Armin Metzger","@id":"https:\/\/www.gtb.de\/en\/#\/schema\/person\/580154dcda34782d68a1503d906ba00a"},"headline":"# When the Agents Write the Code, the Human Checks Themselves [2\/2]","datePublished":"2026-07-14T06:00:53+00:00","dateModified":"2026-07-14T15:28:21+00:00","mainEntityOfPage":{"@id":"https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/"},"wordCount":12411,"commentCount":0,"publisher":{"@id":"https:\/\/www.gtb.de\/en\/#organization"},"image":{"@id":"https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/#primaryimage"},"thumbnailUrl":"https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Heroimage-Post08.png","keywords":["Quality Engineering","Test"],"articleSection":["Blog"],"inLanguage":"en-US","potentialAction":[{"@type":"CommentAction","name":"Comment","target":["https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/#respond"]}]},{"@type":"WebPage","@id":"https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/","url":"https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/","name":"# When the Agents Write the Code, the Human Checks Themselves [2\/2] - German Testing Board","isPartOf":{"@id":"https:\/\/www.gtb.de\/en\/#website"},"primaryImageOfPage":{"@id":"https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/#primaryimage"},"image":{"@id":"https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/#primaryimage"},"thumbnailUrl":"https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Heroimage-Post08.png","datePublished":"2026-07-14T06:00:53+00:00","dateModified":"2026-07-14T15:28:21+00:00","description":"Decision fatigue, context switching, and fading attention: three cognitive bottlenecks of supervising agents, with evidence and concrete countermeasures.","breadcrumb":{"@id":"https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/#breadcrumb"},"inLanguage":"en-US","potentialAction":[{"@type":"ReadAction","target":["https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/"]}]},{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/#primaryimage","url":"https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Heroimage-Post08.png","contentUrl":"https:\/\/www.gtb.de\/wp-content\/uploads\/2026\/07\/Heroimage-Post08.png","width":2392,"height":1080,"caption":"Wenn die Agenten den Code schreiben, pr\u00fcft der Mensch sich selbst"},{"@type":"BreadcrumbList","@id":"https:\/\/www.gtb.de\/en\/blog\/when-the-agents-write-the-code-the-human-checks-themselves-2-2\/#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https:\/\/www.gtb.de\/en\/"},{"@type":"ListItem","position":2,"name":"# When the Agents Write the Code, the Human Checks Themselves [2\/2]"}]},{"@type":"WebSite","@id":"https:\/\/www.gtb.de\/en\/#website","url":"https:\/\/www.gtb.de\/en\/","name":"German Testing Board","description":"Software.Testing.Excellence","publisher":{"@id":"https:\/\/www.gtb.de\/en\/#organization"},"potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/www.gtb.de\/en\/?s={search_term_string}"},"query-input":{"@type":"PropertyValueSpecification","valueRequired":true,"valueName":"search_term_string"}}],"inLanguage":"en-US"},{"@type":"Organization","@id":"https:\/\/www.gtb.de\/en\/#organization","name":"German Testing Board e. V.","url":"https:\/\/www.gtb.de\/en\/","logo":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/www.gtb.de\/en\/#\/schema\/logo\/image\/","url":"https:\/\/www.gtb.de\/wp-content\/uploads\/2023\/10\/gtb-logo.png","contentUrl":"https:\/\/www.gtb.de\/wp-content\/uploads\/2023\/10\/gtb-logo.png","width":224,"height":183,"caption":"German Testing Board e. V."},"image":{"@id":"https:\/\/www.gtb.de\/en\/#\/schema\/logo\/image\/"},"sameAs":["https:\/\/x.com\/GTB_ISTQB","https:\/\/de.linkedin.com\/company\/german-testing-board"]},{"@type":"Person","@id":"https:\/\/www.gtb.de\/en\/#\/schema\/person\/580154dcda34782d68a1503d906ba00a","name":"Dr. Armin Metzger","image":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/secure.gravatar.com\/avatar\/70d5768f24f60d4aa0012915fb37c2937223f49270944aa45222f8a20583d028?s=96&d=mm&r=g","url":"https:\/\/secure.gravatar.com\/avatar\/70d5768f24f60d4aa0012915fb37c2937223f49270944aa45222f8a20583d028?s=96&d=mm&r=g","contentUrl":"https:\/\/secure.gravatar.com\/avatar\/70d5768f24f60d4aa0012915fb37c2937223f49270944aa45222f8a20583d028?s=96&d=mm&r=g","caption":"Dr. Armin Metzger"}}]}},"_links":{"self":[{"href":"https:\/\/www.gtb.de\/en\/wp-json\/wp\/v2\/posts\/19990","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.gtb.de\/en\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.gtb.de\/en\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.gtb.de\/en\/wp-json\/wp\/v2\/users\/6"}],"replies":[{"embeddable":true,"href":"https:\/\/www.gtb.de\/en\/wp-json\/wp\/v2\/comments?post=19990"}],"version-history":[{"count":8,"href":"https:\/\/www.gtb.de\/en\/wp-json\/wp\/v2\/posts\/19990\/revisions"}],"predecessor-version":[{"id":20069,"href":"https:\/\/www.gtb.de\/en\/wp-json\/wp\/v2\/posts\/19990\/revisions\/20069"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.gtb.de\/en\/wp-json\/wp\/v2\/media\/19989"}],"wp:attachment":[{"href":"https:\/\/www.gtb.de\/en\/wp-json\/wp\/v2\/media?parent=19990"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.gtb.de\/en\/wp-json\/wp\/v2\/categories?post=19990"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.gtb.de\/en\/wp-json\/wp\/v2\/tags?post=19990"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}