AI Feed

September 2026

11 entries from September 2026. The latest entries are on the AI Feed itself.

  • 11 Sep 2026
    The version was current and the behaviour had changed

    Two hundred and three upgrades carrying a change nobody announced, and a disclosure that one line of instruction switches off.

  • 10 Sep 2026
    The record stayed true and the work moved

    A reviewer's role, a metric's definition and a tool's description each stayed accurate while the work behind them changed.

  • 9 Sep 2026
    Thirty times the cost for the same task

    What an agent costs to run, how well one model marks another's work, and what a vendor has tested. Three numbers people quote as settled.

  • 8 Sep 2026
    The limit existed and the money went anyway

    Sixty-three catalogued incidents in which an agent ran past its budget, and 31,073 review comments of which developers threw away more than half.

  • 7 Sep 2026
    What the record did not reach

    Three measurements of things a firm records: what an agent costs once someone has to review it, what a register of tool interfaces can see, and what keeping instructions in a se...

  • 6 Sep 2026
    Nothing was enforcing it

    Three studies about arrangements a firm believes it has: a route from junior to senior, a delivery loop that damps rework, and a credential that narrows when an agent hands work...

  • 5 Sep 2026
    The number was right and the thing it stood for was not

    AI spend rose twenty-eight times against a budget set once a year, a retrieval score sat on a set that could not tell right code from nearly-right, and a revocation completed wh...

  • 4 Sep 2026
    Everything on the list was reviewed

    Twenty-one interviews name what a year of AI adoption cost the people doing the work, a third of functionally-correct patches fail the review constraints they were written under...

  • 3 Sep 2026
    The change arrived without a release

    Product managers began shipping code with nothing released to them, faithfully stored memory made answers worse, and the floor for the least-served users is set before a firm bu...

  • 2 Sep 2026
    It ran, and it could not keep up

    The smallest model was the most expensive to run, forty-eight coding agents failed together, and a review that worked properly still let defects through.

  • 1 Sep 2026
    On both sides of the test

    Enterprise usage that is heaviest among the youngest workers, a skill that works better for the model it was not built for, and an automated researcher that closed the safety ga...