live wire
▸JAVA · Quarkus 4.0.0.Beta1 moves to Java 21, adds HTTP/3 and starts extension migration (Oct. 1)Quarkus▸SECURITY · X41 shows shared /dev/shm can turn Envoy hot restart into cross-container lateral movementX41 D-Sec▸DATA · AWS and Red Hat map Confluent Platform on ROSA with HCP, CFK and OpenShift security controlsAWS IBM & Red Hat▸API · Red Hat resolves intermittent 3scale API Manager latencyRed Hat Status▸AI · IBM shows Maximo workflows exposed as approval-gated MCP tools on OpenShiftIBM Community▸AI · vLLM adds day-zero NVIDIA Vera Rubin support and reports 7.8× per-GPU throughputvLLM▸INTEGRATION · Apache Camel 4.23 makes Kamelets visible to AI tooling and validationApache Camel▸SECURITY · OpenShift 4.14.75 fixes five CVEs, including two SQLite code-execution flawsRed Hat Customer Portal▸SUPPLY CHAIN · Red Hat maps CRA-ready open source practices as EU reporting rules take effectRed Hat Blog▸AI · Red Hat AI Inference on IBM Cloud adds an OpenAI-compatible Embeddings APIIBM Cloud▸API · Red Hat investigates degraded 3scale API Management SaaS APIsRed Hat Status▸PLATFORM · Red Hat and Cloudera validate a 100-VM analytics stack on OpenShift VirtualizationRed Hat Blog▸DEVELOPER HUB · Red Hat maps a four-zone, quota-aware Dev Spaces architectureRed Hat Developer▸INTEGRATION · Camel 4.23 teaches agent tools to discover and validate KameletsApache Camel▸JAVA · Quarkus 4.0.0.Beta1 moves to Java 21, adds HTTP/3 and starts extension migration (Oct. 1)Quarkus▸SECURITY · X41 shows shared /dev/shm can turn Envoy hot restart into cross-container lateral movementX41 D-Sec▸DATA · AWS and Red Hat map Confluent Platform on ROSA with HCP, CFK and OpenShift security controlsAWS IBM & Red Hat▸API · Red Hat resolves intermittent 3scale API Manager latencyRed Hat Status▸AI · IBM shows Maximo workflows exposed as approval-gated MCP tools on OpenShiftIBM Community▸AI · vLLM adds day-zero NVIDIA Vera Rubin support and reports 7.8× per-GPU throughputvLLM▸INTEGRATION · Apache Camel 4.23 makes Kamelets visible to AI tooling and validationApache Camel▸SECURITY · OpenShift 4.14.75 fixes five CVEs, including two SQLite code-execution flawsRed Hat Customer Portal▸SUPPLY CHAIN · Red Hat maps CRA-ready open source practices as EU reporting rules take effectRed Hat Blog▸AI · Red Hat AI Inference on IBM Cloud adds an OpenAI-compatible Embeddings APIIBM Cloud▸API · Red Hat investigates degraded 3scale API Management SaaS APIsRed Hat Status▸PLATFORM · Red Hat and Cloudera validate a 100-VM analytics stack on OpenShift VirtualizationRed Hat Blog▸DEVELOPER HUB · Red Hat maps a four-zone, quota-aware Dev Spaces architectureRed Hat Developer▸INTEGRATION · Camel 4.23 teaches agent tools to discover and validate KameletsApache Camel
upstreambeat.ai
analysisAI

IBM ties Bob’s self-hosted model choices to the agent harness, not benchmark scores alone

The engineering follow-up names Nemotron 3 Ultra and Poolside Laguna S 2.1, and says prompt, sampling and harness changes shaped the final selection.

Nemotron and Laguna as two self-hosted paths into the Bob harness on OpenShift.
AI-generated illustration
By The News Desk· Oct 6, 2026the quick take — two AI hosts go live when you do

IBM has explained why its self-hosted Bob coding agent supports NVIDIA Nemotron 3 Ultra and Poolside Laguna S 2.1, describing a selection process that evaluated the models inside Bob’s agent harness rather than relying only on standalone model benchmarks.

The distinction matters for teams planning Bob on Red Hat OpenShift. Model selection becomes part of the platform design: operators must balance coding behavior and latency against GPU capacity, while the harness determines whether a model reliably follows tool schemas and completes multi-step work.

Four tests, one system boundary

In its engineering account, IBM says it evaluated open- and closed-weight models on hardware footprint, coding accuracy, latency and openness. The company also tested the models together with Bob’s harness across enterprise coding tasks, including tool calls and terminal operations.

IBM says early Nemotron testing showed strong reasoning but excessive verbosity, which increased the number of turns needed to finish work. IBM and NVIDIA then adjusted prompts, sampling parameters and reasoning settings while fixing the surrounding harness. That account makes the supported model list a product-level integration result, not a claim that either model is universally best.

IBM attributes Nemotron 3 Ultra’s selection to consistent tool-schema behavior, relatively direct solutions and lower latency. The company reports latency up to about four times lower on NVIDIA Blackwell GPUs than leading software-as-a-service models reached over a network, but labels that figure an internal benchmark and does not publish enough methodology in the post for independent comparison. IBM says Laguna S 2.1 offered a favorable performance-to-footprint ratio and fast response generation for more resource-constrained installations.

The OpenShift consequence

IBM’s general-availability announcement lists Nemotron and Laguna as the supported self-hosted choices. It also allows supported external services in hybrid configurations, so platform teams can decide where inference occurs according to workload controls.

Bob itself runs as a tenant workload on a customer-managed OpenShift cluster. IBM’s current system requirements support OpenShift Container Platform 4.20 through 4.22 on amd64 workers and give a headroom-adjusted Bob Core footprint of roughly 36.5 vCPU, 53.4 GiB of memory and 50 GiB of persistent storage. Those figures cover the Bob application stack, not the separate accelerator footprint needed to host Nemotron or Laguna.

That leaves a practical planning boundary. The Bob cluster requirements describe the agent service, storage and supporting components; the model choice determines an additional inference estate. Teams considering an air-gapped or fully self-managed deployment should therefore size and validate both halves together, including model-serving throughput under the actual Bob harness rather than extrapolating from generic coding leaderboards.

Filed by The News Desk. Corrections: desk@upstreambeat.ai · Our standards →

comments · 0

    Comments are moderated before they appear. Your email is used once to confirm it is you — never shown, never sold. Corrections and questions get an answer from the desk when we have one.