Saturday, August 8, 2026

How one can enhance understanding with Claude Code


Paul Goldsmith-Pinkham posted on his substack feed the opposite day a hyperlink to this fascinating weblog put up by Geoffrey Litt referred to as “Understanding is the brand new bottleneck”. I encourage you to learn it. It’s quick however candy because the outdated children used to say.

I assumed it was fascinating as a result of it edged me a bit in direction of a bumpersticker-style factor I’ve been saying the final six or so months, which is that “verification is the brand new bottleneck”.

Closing tabs: Tuesday version

I awoke at 2am so like several rational individual grabbed my telephone reasonably than attempt to battle again to sleep and as a substitute attempt to empty round 70 or so open hyperlinks off my telephone, a number of of which had been about our blizzard in Boston yesterday. Thanks once more everybody to your assist! Should you get pleasure from this substack, think about changing into a paying subscriber! It’s solely $5/month …

“Verification is the brand new bottleneck” is a sentence and sort of rhetoric you see in all places through which a author posits that the demand for analysis output is downward sloping on the particular person stage. Earlier than AI brokers, the marginal value of manufacturing analysis was some quantity that decided analysis amount demanded. I’ve drawn under two photos representing how I consider this shift in analysis output, which is self-explanatory and intuitive to economists, however could also be useful to stroll by means of for non-economists.

The primary graph is the outdated equilibrium decided by the marginal value of manufacturing analysis which I symbolize as a horizontal “fixed marginal value” for simplicity. At this outdated marginal value of manufacturing, MC_b for “earlier than AI”, we produce an quantity of analysis as much as the purpose the place the marginal advantages are equal to the marginal value, most probably fastened by time and ability. To do extra would require extra human capital and/or extra time, and given the shortage of each, the researcher wouldn’t do it.

However now the second graph. Now with AI brokers, a number of issues occur. First, I and others argue that the manufacturing of analysis will increase beneath AI brokers as a result of the marginal value of manufacturing analysis fell from MC*_b (earlier than marginal value) to the a lot decrease MC^hat_AI which is at some arbitrarily small quantity above 0 equal to epsilon.

I additionally put surpluses on this graphic although I’m obscure about simply what “analysis surplus” means exactly, largely as a result of I’m probably not clear myself. The outdated surplus was B, and it constituted I believe the subjective web advantages of the analysis. However beneath the decrease marginal value, we achieve not solely extra analysis (the elasticity of which depends upon the scale of the demand curve on this case), but additionally further surplus too. Particularly, the outdated work we’d’ve executed is completed with fewer time inputs, and thus the outdated work now has web advantages of B+AI_1, which is the world beneath the demand curve as much as Q*b. However as analysis output rose from Q*b to Q*_AI, there’s further surplus generated from the new work equal to AI*2.

The factor that I’m unclear about is that this, although. In precept, all of us have further time from this shift that we now get to reallocate to different duties in our lives. What that’s or shall be is formed by heterogenous preferences.

However one factor that existed, at the very least in my life, was that the subjective high quality of the outdated work has modified. By subjective high quality I imply that the outdated work, Q*b, had two qualities inside that triangle B. The primary high quality, was that it was produced in any respect, which is revealed by Q*b. That it existed in any respect beneath the human mode of manufacturing implies that it was created — a press release that admittedly does appear tautological, however I believe it’s value saying anyway. Q*b was created by people as a result of people created it, which leads on to the second level which is that the online advantages of B was that it was the (web) worth of that output to the researcher consisting each of its existence (manufacturing) but additionally one thing a bit slippery which I believe is simply the innate understanding of the work itself that occurs nearly by chance. I consider it as a type of Hobbs-like squatting by employees whereby the sheer act of working to show the pure earth into one thing else created possession. There was, in different phrases, nearly like possession to the output. Not a lot possession as in nobody else might have it a lot as a sort of psychological possession. I made it, and subsequently I knew it. One way or the other the manufacturing of the factor and the understanding of the factor existed concurrently, as you can not make cognitive output with out understanding it too because the understanding of it was often a prerequisite to doing it, and likewise the byproduct of constructing it in any respect.

However now, with AI brokers producing, what even is B? Is it the identical B as earlier than or has even the brand new B modified to one thing else — like a B’? That is the place the geometry of this begins to interrupt down. I don’t suppose it’s the identical B as earlier than personally as a result of there’s not there the identical type of Hobbs-like squatter rights over it since we weren’t those who made it. By which I imply the epistemological perform of labor itself just isn’t current. The latent understanding of the work has gone away as a result of understanding will not be produced not directly as a byproduct when we aren’t doing it.

Now it may be, 100%. Lab managers perceive by proxy the work when interrogating and dealing intently withe the coauthor and RA who made it, however usually that’s not the case. Major authors shall be held accountable, as an example, when their RAs make errors and they didn’t do the due diligence to be on high of it. That has occurred earlier than. We are able to all consider well-known instances over the past 10+ years or so through which individuals had been roughly deceived by duplicitous coauthors who had fabricated knowledge outright. I believe B was in different phrases extremely contingent on belief with a coauthor, traditionally anyway, even with managers. They might roughly free journey on the belief of the coauthor or RA’s personal subjective understanding of the work in order that B might exist even within the minds of the non-producer just because they trusted the producer.

However put that apart, as a result of now my level is that we’ve two new surpluses and that’s AI1 and AI2. And I believe indirectly, these two new surpluses symbolize the extra positive aspects subjectively that occurs because of taking their arms off the wheel to some extent. It’s gained subjective web advantages and it’s gained subjective web advantages coming from manager-like belief that the work executed was executed effectively even when they didn’t do it, and even when they don’t perceive it.

I believe for some, it’s going to be the principal-agency drawback, although. It may very well be the blind main the blind. I belief the agent did the work effectively, however because the agent just isn’t actual, we don’t truly realize it was executed effectively, and if our expertise atrophy, we could not even possess the identical comprehension of what it means to be executed effectively.

Nicely, the query I’ve in my thoughts is that couldn’t we get there? I imply we’ve a lot extra time, don’t we? Couldn’t we simply reallocate the time financial savings again to the verification job of that new and outdated materials created and it grow to be actual surplus, one thing we each expertise as optimistic web advantages and which is precise optimistic web advantages? That’s, we expertise it as beneficial as a result of we predict it’s appropriate, and it’s truly socially beneficial as a result of it’s appropriate. I believe these may not be the identical factor. I believe they might be completely different.

See, the factor I preserve considering is that it’s an assumption that the time financial savings we now have in our possession are even able to getting used to fill in these areas in any respect. We don’t know that it’s. We don’t know, I don’t suppose, if the time financial savings we get from AI brokers may be allotted to fill in that whole B+AI1+AI2 with subjective understanding. What if it’s not? What if the outdated approach of understanding one thing was primarily an unintentional consequence of the manufacturing course of. Air pollution is like that, as an example. Air pollution is the byproduct of manufacturing. It’s an externality. What if understanding was all the time an externality — one thing which was created nearly by chance when one got down to make some cognitive output utilizing time inputs?

So I believe I agree with this concept that it’s not actually probably the most correct method to describe the state of affairs we at the moment are in to say that “verification is the brand new bottleneck”. I believe it’s extra correct to say “understanding is the brand new bottleneck” as a result of it’s fully doable that AI brokers will have the ability to confirm in addition to produce, however that doesn’t subsequently imply we will perceive it.

So I’ve been considering so much about this essay by Geoffrey Litt, in addition to this concept Paul had put forth which is the unit of verification is the git diff operation. I now use the git diff operations like Paul advised, however I’m noticing that I’m skimming and accepting them creating what Paul and others name cognitive debt — solely it’s imagined to be that the cognitive debt is cleared by the sheer act of checking off these diffs. However I believe I’m truly simply silently accepting them with out full understanding. That doesn’t imply everybody else is a lot because it means I’m, and if you happen to sense generally that you’re doing the identical factor time and again habitually, it’s time to think about if maybe you need to strive one thing completely different. Possibly not as a substitute of the diffs, however as well as, the place the objective is now two issues, not one.

  1. Zero Errors Stays the Constraint. I stay satisfied that we should say to ourselves and each other that “zero errors is the constraint”, and never the target perform. Our objective shouldn’t be to reduce errors, as a result of if you happen to say that, then you’ll quietly come to just accept the thought of “optimum errors”. And that can not be what we transfer in direction of given the mixture provide elasticities of analysis manufacturing with respect to costs might be higher than 1. I believe in the long term it’s anyway, even when within the shortrun, resulting from fastened inputs, it’s small.

  2. Understanding is the constraint. Is knowing now the constraint or is it additionally the objective? I’m considering a threshold of understanding the work should grow to be the constraint too. Paul often quotes IBM thus far, at the very least not directly: since machines can’t be held accountable for errors, they can’t be trusted to make selections both. They both work to assist us perceive the work they’ve executed, or we don’t use them in any respect. I believe it’s doable these are the 2 nook options.

  3. Analysis output is the target perform. So then what’s the goal perform? It’s no matter we predict is perfect in equilibrium however not perfected, and I believe that may very well be the concept that we are attempting to succeed in optimum analysis output, and optimum analysis output is the purpose the place the social marginal advantages equal the social marginal value, which requires verification and understanding.

So, what am I doing in a different way than git diffs. I’m experimenting with a brand new ability referred to as /quiz. It’s an extension of issues I used to be already doing which had been interviews used to extract my beliefs about which covariates to make use of and what the goal parameter ought to be in conditions the place I might probably not make sure, however I suspected I did know deep down sufficient to say one thing. However now what I’m engaged on is a little bit stranger of an interview idea.

My /quiz ability has two parts. First, I’m engaged on Claude Code making a set of slides when invoked that teaches me one thing. After which I’ve Claude Code ask me a number of selection questions, one by one, based mostly on it. You realize the place I obtained the thought? The coaching movies that HR makes you do on a regular basis at work. That’s how they do it — they offer you one thing to learn, oftentimes which features a video, after which they offer you a quiz. That’s the place I’m going. I’m going in direction of roughly “coaching video quizzes” to continuously create in me steady understanding. On high of verification. On high of my detailed staged checklists.

We’ll see how this goes. I don’t know if I’m over-engineering the manufacturing course of utilizing AI brokers, or whether it is optimum, however I’m positively transferring in direction of one thing that’s taping my arms to the wheels. I’m tying myself to the mast to float by means of the strait in order that I don’t go insane by the screaming of the harpies within the water, so to talk. There is just too many issues that may go fallacious, and I for one don’t get excited by the concept that a sniper will come alongside after me and discover my evaluation is simply riddled with issues. The present surroundings through which issues in each other’s work is handled as scalps to be taken and kilos of flesh to be eliminated is much too antagonistic for my tastes. I’m not desirous about changing into serving to somebody get their quarter-hour of fame on a podcast as they focus on the issues with my work to be frank. However I additionally simply wish to be my greatest self. I wish to constantly be taught. Studying is why I obtained into this. I’m, deep down, probably not a aggressive individual a lot as a romantic about studying, and if brokers may help me be taught in my artistic work, I need them to sit down within the cockpit with me. But when they’ll be auto piloting whereas I sleep, then my truest self would like I not do this as till the lights exit, I stay in love with creativity and artistic duties. I get pleasure from it, it motivates me, it fills my coronary heart with lightness and pleasure, and my psychological well being desperately desires it.

In order that’s all. I simply needed to share these ideas. Now again to work attempting to wrangle these brokers into doing what I say even when they, in a really good fashion, stumbling and mumbling as they do it.

Related Articles

Latest Articles