IBM wants to sell the review queue article image

IBM wants to sell the review queue

Friday afternoon, the code is done. Monday morning, the review queue is not.

That gap is where the market moved. GitLab's new AI Accountability Report says "85% agree AI has shifted the bottleneck from writing code to reviewing and validating it." IBM is answering that shift with IBM Bob V2, a release built around multi-agent execution, cost visibility, and modernization workflows for old systems that still run the business.

The reason this matters is simple. We already know how to make code faster. The hard part now is deciding what is trustworthy enough to ship. GitLab says 92% of organizations report governance challenges with AI-generated code, and 80% say they adopted AI tools faster than they developed policies to govern them. That is not a tooling problem. That is an operating problem.

IBM's bet

IBM said Bob now has new multi-agent capabilities, built-in AI cost and use analytics, and pre-built workflows for modernizing enterprise systems. The company framed the update around something sharper than a coding assistant. Its own line is blunt: "The bar for enterprise AI is no longer a better coding assistant." That is a useful line because it tells you what IBM thinks the customer actually wants.

IBM wants to sell the review queue
The queue gets longer when the code still depends on old systems.

Bob V2 runs on a single agent and a shared harness, with subagents handling self-contained work so the main context stays clean. It also adds parallel, native tool calling, background tasks, rollback, and structured workflows. That sounds like architecture talk until you map it to the work most enterprises actually pay for. A lot of the value is in keeping the process from drifting while the model is doing the work.

IBM's premium packages make the target even clearer. The release adds opinionated workflows for IBM Z, IBM i, and Java modernization. In plain English, IBM is trying to help them touch code they already depend on without turning every change into a one-off rescue operation.

The customer examples make that less abstract. IBM says Jack Henry used Bob to accelerate RPG development workflows and gain deeper insight into decades of system knowledge. Blue Pearl says a legacy modernization effort that had been projected to take nine months with 14 engineers was completed in three days. Those are not benchmark claims. They are operations claims.

IBM Bob's own blog is even more direct about the shape of the problem. It says "AI is good at open-ended problem-solving and bad at doing the same thing twice." That is the whole reason workflows matter. One-off cleverness is useful in a demo. Repeatability is what gets a modernization program through the quarter.

Free HubSpot workshopBring one HubSpot problem to a free 30-minute callA screen-share walkthrough of your portal with me, not a salesperson, and a short roadmap at the end. No contract or credit card.Book the free workshop

What buyers are buying

The extra piece here is visibility. IBM added Bobalytics, its cost and usage layer, because enterprise teams do not just want output. They want to know what the output cost, what it touched, and how the result fits into a wider development process. That is the difference between a neat assistant and something a platform team can actually govern.

GitLab's report gets at the same point from the other side. It says 91% of organizations have two or more AI coding tools in active use, but 78% report faster code output while the overall delivery process has not accelerated at the same pace. That is the AI Paradox in one sentence. Speed went up. Control did not.

That is why IBM's release reads less like a model announcement and more like a product for the middle of the pipeline. The people buying this do not need another shiny demo of code generation. They need less uncertainty between the first draft and the thing they are willing to put in production.

The practical result is easy to see. A subagent can read, search, and figure out a local problem without polluting the main thread. A background task can keep running while somebody else keeps working. A workflow can split automation, AI, and human approval into the right steps instead of pretending every problem should be handled the same way.

That is where IBM is placing its bet. Not on better autocomplete. On making review, modernization, and governance boring enough that the queue stops running the business. For the companies still carrying mainframe, IBM i, and Java systems, that is not a small promise. It is the product.

IBM wants to sell the review queue supporting image
IBM is trying to make review feel like a process instead of a pileup.
Sam C BarthAI sticks when the CRM underneath it is cleanI help teams get HubSpot, data, and handoffs in shape so new tools have something solid to run on.Visit samcbarth.com