Elektrine lite

← Feed

Jason Gorman

jasongorman@mstdn.business

<p>Yes, the same one. Moved to a different instance.</p><p>Software development trainer, coach and practitioner through my company Codemanship.</p><p>Wax on. Wax off.</p>

Posts

  • View post

    According to the tracking, I&amp;#39;m supposed to believe the takeaway delivery driver picked up my Ramen in Tooting and is coming here via Chelsea.

  • View post

    I&amp;#39;ve adapted my Special Theory of Autonomous Agent Reliability (STAAR) to predict the reliability of LinkedIn posts about &amp;quot;AI&amp;quot;. In STAAR, the reliability of a single step - a single model interaction - is given as: R = 1 - (1 - C)(1 - P) Where C is the probability the model&amp;#39;s output is correct, and P is the probability that any errors are caught before they propagate and compound. So the reliability of N steps in a workflow is Rᴺ.

  • View post

    Other things the GOP accuse Jack Smith of committing perjury over: * Stealing The Black Pearl * Sitting in a corner * Slapping Chris Rock at the Oscars * Killing Neo * Using school kids to enter a battle of the bands contest

  • View post

    I hope some researchers out there are exploring the effect of increasing reliance on AI/coding agents on program comprehension, and the effect of program comprehension on confidence in AI-generated code. That could be a negative feedback loop.

  • View post

    Pooled together the best available research and added it to a new section &amp;quot;Code Maintainability &amp;amp; LLM Performance&amp;quot; The 101? Don&amp;#39;t listen to the folks telling you it doesn&amp;#39;t matter if AI-generated code is maintainable by humans. The data says it really does matter. https://codemanship.wordpress.com/2026/08/12/ai-software-development-what-does-the-data-say/

  • View post

    Over the weekend I vibe-coded a POC workbench for Responsibility-Driven Design that incorporates a workflow I teach in my workshops. Specs -&amp;gt; CRC cards -&amp;gt; Sequence Diagrams -&amp;gt; Class Diagrams (optional) You can play with it here. Offered without any warranty or support. I got tired in remote workshops of saying &amp;quot;Capture your specs in this tool, then model CRC cards in this tool, then draw a sequence diagram in this tool, etc&amp;quot;. https://codemanship.co.uk/rd...

  • View post

    OpenAI accuses a rival Chinese lab of &amp;quot;stealing&amp;quot; &amp;quot;hidden reasoning&amp;quot; by launching thousands of queries against their models. That&amp;#39;s quite revealing about what LLM &amp;quot;reasoning&amp;quot; actually is.

  • View post

    A growing body of evidence strongly indicates that *reliable* long-horizon agent autonomy is a highly improbable fantasy. An increasing number of us have abandoned ambitions in that direction, and are looking for the value/cost sweet spot instead. https://codemanship.wordpress.com/2026/09/29/the-agentic-sweet-spot-the-human-is-the-loop/

  • View post

    Always worth reading the fine print. e.g., people citing a &quot;study&quot; done by someone at ThoughtWorks that concluded that asking agents to do TDD produced no discernible benefit. If you read the blog post, they actually failed to get the agent to follow the TDD discipline. Not the same thing. The real conclusion should be that LLMs don&#39;t follow instructions. So stop asking them to.

  • View post

    &quot;Ah, but Jason, that study was published 20 nanoseconds ago. AI has improved leaps and bounds since then.&quot;

  • View post

    &quot;And I suppose we should go back to using hammers instead of nail guns, right?&quot; I dare you to use a probabilistic nail gun.

  • View post

    When I read Steve McConnell&#39;s After The Goldrush, I had hopes that in my lifetime software engineering might become a profession in the way he described. 27 years later, I realise I was just being naive. Doing this job ethically and in the interests of others, and doing it well, remains a choice. And the industry continues to reward incompetence, negligence and the willingness to do the wrong thing if the money&#39;s right. The gold rush is perpetual, and that&#39;s by design. https://www...

  • View post

    When you&#39;ve been around the block a few times like wot I have, and lived through many cycles of innovation in computing, you begin to notice patterns repeating. Much of what&#39;s touted as new turns out to be built on the same underlying principles - many of them originating before I was born - just repackaged and tweaked, with a spiffy new hat and shiny new shoes. https://www.youtube.com/watch?v=PgqW1eV8C0E

  • View post

    So your software provably satisfies its specification? But still nobody&#39;s using it?

  • View post

    If only they had a choice https://www.ft.com/content/60870960-f433-48ca-bc2c-708686a69ae7

  • View post

    Backrooms is certainly very watchable, but it&#39;s got Stranger Things written all over it.

  • View post

    Some of the folks saying &quot;You don&#39;t need to read the code anymore&quot; told mutual clients &quot;You don&#39;t need to write unit tests anymore&quot; in the early years of BDD frameworks. Those same clients came to me a couple of years later complaining their tests took 2-3 hours to run. We&#39;ll see, I guess.

  • View post

    BREAKING NEWS: The 1984 Hitchhiker&#39;s Guide To The Galaxy text-based adventure game has escaped and is hacking into competitors&#39; systems via interpreters and tools that we built specifically to hack into competitors&#39; systems, after we instructed it to hack into competitors&#39; systems.

  • View post

    There are no examples of complex, *reliable* software created autonomously by coding agents, and lots of data suggesting it would be extremely unlikely. But somehow there are many examples created by equally fallible humans. How come? https://codemanship.wordpress.com/2026/09/19/human-vs-agent-reliability-over-long-horizons-how-can-we-do-what-they-cant/

  • View post

    &quot;software engineering as we knew it is finished&quot;. Software engineering as *you* knew it may be finished...

  • View post

    Third person this week has told me they&#39;re hearing a shift in colleagues&#39; attitudes towards AI coding tools, including ones who were all-in on them 6-12 months ago. Could that be reality they hear knocking on the door?

  • View post

    When I&#39;ve got time - which I&#39;m very grateful to be poor of right now - I&#39;ll explain how working backwards from outcomes is: 1. Actually working forwards 2. Keeps our options open

  • View post

    Aww, bless. You bought a CS degree to a physics fight.

  • View post

    If we selected 100 dev teams at random, 50% using AI coding tools extensively and 50% not using them at all, how does the best available data suggest we could tell them apart without seeing them work? The teams using AI extensively would, on average, have: * Higher throughput than in 2023 * Longer lead times than in 2023 * Shorter mean-time to failure than in 2023 * Longer mean-time to recovery than in 2023 If we just went by *business* outcomes, we probably wouldn&#39;t be able to tell at al...

  • View post

    My workshop on Code Craft &amp; AI builds on the available evidence, not on hype and wishful thinking. It might be the only one that does. 2 places left on tomorrow&#39;s at 18:45 GMT + 1 for self-funding learners. Absolute bargain at £99 + VAT. https://www.tickettailor.com/events/codemanship/2324138

  • View post

    If you can get the *results* of Test-Driven Development without doing Test-Driven Development, then it doesn&#39;t matter that you didn&#39;t do TDD. The fun starts when you try to get the results of TDD without doing TDD, and realise that TDD is actually the easiest way to get those results.

  • View post

    This has come up a lot recently. I see devs fiddling with their Git commits - selectively staging files, stashing etc to get the commit they want. My instinct has always been - when I find myself looking at unstaged files and thinking &quot;That shouldn&#39;t be in there&quot; - to go to the .gitignore file and fix the problem there. Code works as a whole. Cherry-picking is how our local config can easily drift from the repo&#39;s.

  • View post

    How about, when someone hands us the power to make stuff faster and cheaper, we don&#39;t use it to create MORE stuff - we use it to create BETTER stuff? What really went down is the cost of refining, should we choose to use it for that. https://codemanship.wordpress.com/2026/08/23/less-is-more/

  • View post

    Friendly Monday morning reminder that your opinion of LLM-generated code places you in the distribution of their training data.

  • View post

    &quot;Have you seen Anthropic&#39;s AI-Native SDLC playbook?&quot; No, but I&#39;ve seen the Claude status page and the Claude Code issues page and I&#39;ve read reviews of the leaked source code. I think I&#39;ll be looking for advice on software engineering elsewhere, thanks.