live wire
SECURITY · Red Hat Advanced Cluster Security 4.10 moves to maintenance support Sept. 4Red Hat Customer PortalSECURITY · OpenShift 4.22.12 patches nine CVEs in an Important-rated updateRed Hat ErrataSECURITY · Red Hat maps automated vulnerability response from triage to governed remediationRed Hat BlogJAVA · Quarkus 3.39 adds post-quantum TLS controls ahead of 3.40 LTS (Aug. 27)QuarkusAI · RHEL AI publishes Muse Glimmer 30B modelcar images for x86, Arm, Power and IBM ZRed Hat ErrataAI · RamaLama 0.24 adds Pi coding-agent sandboxing and custom Hugging Face endpointsRamaLamaDATA · Kafka 4.4.0 awaits another release candidate as new KIPs target diagnostics and invalid ISR settingsRed Hat DeveloperAI · IBM puts four Granite time-series models inside Confluent Cloud’s Flink SQL early accessIBMAI · Red Hat separates skill routing from answer generation in Ask Red HatRed Hat BlogAI · OpenShift AI turns enterprise policy documents into automated red-team attacksRed Hat BlogSUPPLY CHAIN · Tekton Pipelines 1.16 enables restricted security contexts by defaultTektonAI · Apache Camel 4.23 adds OpenTelemetry spans and per-model token metrics for LLM routesApache CamelAI · IBM Spyre 12-card Power drawers require Red Hat AI Inference Server 3.5 (Aug. 25)IBM DocumentationAI · Red Hat turns Ray and Docling RAG into a five-component OpenShift AI pipelineRed Hat DeveloperSECURITY · Red Hat Advanced Cluster Security 4.10 moves to maintenance support Sept. 4Red Hat Customer PortalSECURITY · OpenShift 4.22.12 patches nine CVEs in an Important-rated updateRed Hat ErrataSECURITY · Red Hat maps automated vulnerability response from triage to governed remediationRed Hat BlogJAVA · Quarkus 3.39 adds post-quantum TLS controls ahead of 3.40 LTS (Aug. 27)QuarkusAI · RHEL AI publishes Muse Glimmer 30B modelcar images for x86, Arm, Power and IBM ZRed Hat ErrataAI · RamaLama 0.24 adds Pi coding-agent sandboxing and custom Hugging Face endpointsRamaLamaDATA · Kafka 4.4.0 awaits another release candidate as new KIPs target diagnostics and invalid ISR settingsRed Hat DeveloperAI · IBM puts four Granite time-series models inside Confluent Cloud’s Flink SQL early accessIBMAI · Red Hat separates skill routing from answer generation in Ask Red HatRed Hat BlogAI · OpenShift AI turns enterprise policy documents into automated red-team attacksRed Hat BlogSUPPLY CHAIN · Tekton Pipelines 1.16 enables restricted security contexts by defaultTektonAI · Apache Camel 4.23 adds OpenTelemetry spans and per-model token metrics for LLM routesApache CamelAI · IBM Spyre 12-card Power drawers require Red Hat AI Inference Server 3.5 (Aug. 25)IBM DocumentationAI · Red Hat turns Ray and Docling RAG into a five-component OpenShift AI pipelineRed Hat Developer
upstreambeat.ai
newsAI

IBM puts Granite time-series inference inside Confluent Cloud’s streaming SQL

Four compact Granite models enter early access through Flink SQL, keeping forecasts and anomaly detection in the same governed pipeline as Kafka data.

Flink SQL streaming AI with four Granite models in and out of Kafka
Side by side: what changed
By The News Desk· Sep 1, 2026the quick take — two AI hosts, this story only

IBM and Confluent have opened early access to four Granite time-series foundation models inside Confluent Cloud, making forecasting and anomaly detection callable from the same Flink SQL pipelines that process live Kafka data. The IBM announcement says the initial service runs on Confluent Cloud on AWS, with Confluent Platform support for on-premises and hybrid deployments planned later.

What changed

The integration exposes IBM’s models through Confluent’s existing AI_FORECAST and AI_DETECT_ANOMALIES Flink SQL functions. Teams can switch among the supported models without redesigning the surrounding streaming pipeline, while Confluent manages serving infrastructure, scaling and runtime operations.

The early-access set includes PatchTST-FM-r1 for probabilistic forecasts, FlowState-r1.1 for point forecasts, TTM-r3 for efficient processing of many series with control variables, and TSPulse for anomaly detection, classification, similarity search and gap filling. IBM describes the models as ranging from 1 million to 260 million parameters and requiring no GPU.

Inference results are written back to Kafka topics. That lets alerting systems, dashboards, lakehouses and AI agents consume the output without a separate handoff from a machine-learning environment. IBM also says those inference pipelines inherit the platform’s schemas, lineage and access controls, while Kafka’s replayable topics provide a record for audits, troubleshooting and model evaluation.

Who it affects

The immediate audience is data-streaming teams that already use Confluent Cloud and Flink SQL but have treated forecasting as a separate workload. IBM’s examples include scoring payment activity before a transaction completes, forecasting demand across product catalogs, monitoring industrial telemetry and comparing current signals with historical patterns.

The design also matters to platform teams trying to reduce the operational boundary between event processing and model serving. Small CPU-capable models can sit closer to the stream than a general-purpose GPU inference service, although the early-access label means teams should treat the interface and deployment limits as subject to change.

What to do

Confluent Cloud users can enroll in the early-access program and test the SQL functions against non-critical streams. A useful evaluation should compare forecast quality and anomaly sensitivity across the four models, then verify how results, schemas and lineage appear in downstream Kafka topics.

Hybrid and on-premises teams cannot use the announced Confluent Platform path yet. They can still evaluate whether the proposed architecture fits their governance model, but IBM gives no availability date for that support. Production adoption should wait for service-level, pricing and support details that the announcement does not yet provide.

Filed by The News Desk. Corrections: desk@upstreambeat.ai · Our standards →

comments · 0

    Comments are moderated before they appear. Your email is used once to confirm it is you — never shown, never sold. Corrections and questions get an answer from the desk when we have one.