VFF - The signal in the noise
NewsTrending

IBM mainframe chip runs Arm and Z workloads on same cores

Read original
Share
IBM mainframe chip runs Arm and Z workloads on same cores

IBM announced a dual-architecture mainframe processor at Hot Chips that can natively execute both Arm and IBM Z instruction sets on the same cores, switching between them in nanoseconds. Built on 2-nanometer process technology with 11 cores running above 5.7 GHz, the chip allows enterprises to run Arm-native Linux software and AI frameworks alongside z/OS transaction-processing workloads on shared silicon. The design represents the first hardware outcome of IBM and Arm's April strategic collaboration and directly addresses whether mainframes can remain relevant in an AI-dominated infrastructure landscape.

  • IBM's next-gen mainframe processor features 11 cores that can dynamically switch between Arm64 and Z instruction sets in nanoseconds with negligible performance penalty
  • Every core is bilingual rather than using a heterogeneous design with separate Arm cores, enabling deep integration of Arm-native Linux and AI frameworks with mission-critical z/OS workloads
  • The chip uses 2-nanometer process technology, runs at base frequency above 5.7 GHz, includes on-chip AI inference accelerators, and scales to hundreds of cores and tens of terabytes of memory
  • The design leverages the open-source KVM hypervisor to run Arm64 Linux virtual machines and Linux on Z virtual machines side by side, with traditional z/OS workloads in separate partitions

Mainframes process most of the world's regulated financial transactions but have faced questions about relevance in an AI era built on other architectures. This processor directly solves that tension by allowing banks, insurers, and governments to run modern Arm-native AI frameworks and monitoring tools on the same silicon as their core transaction systems, eliminating the need for separate infrastructure.

Enterprises can consolidate infrastructure by running AI workloads, Linux applications, and mission-critical z/OS systems on shared hardware with identical reliability guarantees and memory fabric. This reduces operational complexity and capital expenditure for organizations that depend on mainframes for regulated financial processing while needing modern AI capabilities.

  • Mainframe architecture can evolve to support modern software ecosystems without abandoning the transaction-processing workloads that define enterprise banking and insurance infrastructure
  • The strategic IBM-Arm collaboration produces a commercially viable product that validates dual-architecture design at scale, potentially influencing how other processors approach heterogeneous computing
  • Organizations can defer or reduce decisions to migrate away from mainframes by gaining native access to Arm-based AI frameworks and Linux software on existing z/OS platforms

Monitor adoption rates among financial services institutions and whether the performance characteristics hold under production workloads. Watch for competitive responses from other mainframe vendors or cloud providers attempting similar dual-architecture approaches. Track whether the nanosecond switching overhead remains negligible as enterprises scale Arm workloads on these systems.

Share

Subscribe to the newsletter

The latest stories and analysis, delivered to your inbox.

Free. No spam. Unsubscribe any time.

Related stories

NVIDIA Vera Rubin Cuts Agentic AI Costs 35x, Boosts Efficiency 30x

NVIDIA Vera Rubin Cuts Agentic AI Costs 35x, Boosts Efficiency 30x

NVIDIA's Vera Rubin NVL72 GPU system delivers up to 30x higher throughput per megawatt than its GB300 NVL72 predecessor on agentic AI workloads, according to measurements using the SemiAnalysis AgentX benchmark. Agentic tasks consume 15x more tokens than simple chat because agents iteratively query databases, invoke sub-agents, and accumulate context across multiple reasoning steps. For power-constrained AI infrastructure operators, the efficiency gain translates directly to running significantly more agent-based work within the same energy budget.

by Shruti Koparkar· NVIDIA Blog (AI)
Nvidia Lands SpaceX, Nebius as Early Vera CPU Customers
TrendingNews

Nvidia Lands SpaceX, Nebius as Early Vera CPU Customers

Nvidia announced that SpaceX and AI cloud company Nebius will be early customers for its Vera CPU and Groq LPX inference-focused racks, marking the company's expansion beyond its core GPU business. The move addresses investor and analyst questions about how Nvidia plans to diversify its product portfolio beyond graphics processors. Both products target different segments of the AI infrastructure market, with Vera serving as a central processing unit option and Groq LPX designed for fast inference workloads.

by Phoebe Liu· The Information
General Intuition raises $6B at valuation for embodied AI
TrendingNews

General Intuition raises $6B at valuation for embodied AI

General Intuition, an AI startup developing foundation models for generalized agents that navigate physical and temporal space, is raising capital at a $6 billion pre-money valuation from investors including Valor Ventures, Point72 Ventures, and Seven Seven Six. The funding round signals investor confidence in AI systems designed for robotics and embodied AI applications. The startup's focus on spatial reasoning and agent movement represents a shift toward practical, physical-world AI deployment.

by Rebecca Bellan· TechCrunch AI
Nvidia Raises Flagship AI Chip Prices 17%

Nvidia Raises Flagship AI Chip Prices 17%

Nvidia is raising prices for its Grace Blackwell and Vera Rubin flagship AI chip systems by approximately 17%, according to server makers. The increase adds to mounting cost pressures facing cloud providers and data center developers, who are already contending with power infrastructure delays and other unexpected expenses. The price hike reflects Nvidia's market position as demand for advanced AI chips remains strong despite the growing cost burden on customers.

by Amir Efrati· The Information