Topic
Infrastructure
AI compute, cloud infrastructure, MLOps, and deployment tooling
Featured
All Stories
NVIDIA Moves Memory Controller to Cut Power, Boost Bandwidth
NVIDIA expanded its NVLink Fusion platform with NVHBM, a custom high-bandwidth memory technology that integrates the…
Nvidia Heads Toward $100B Quarterly Revenue Milestone
Nvidia projects quarterly revenue of $108 billion in its next earnings report, up from a record $96.2 billion in the…

Chinese AI Model Undercuts US Rivals by 7x on Cost
Zhipu's GLM-5.3-Flash model launched on OpenRouter at 7.5 to 25 cents per million tokens (promotional pricing),…

Atlassian Bets on Knowledge Graphs for AI Agent Infrastructure
Atlassian and other software firms are promoting knowledge graphs, or graph databases, to help AI agents analyze…
OpenAI CFO: AI Scaling Requires Full-Stack Advances
OpenAI CFO Sarah Friar outlined how the company views intelligence scaling as a function of advances across four…

ClickHouse Hits $350M ARR on AI Agent Demand
ClickHouse, a database company specializing in high-speed analysis of large data volumes, has crossed $350 million in…
Accel-backed Keenable builds web index for AI agents
Keenable, backed by Accel, has exited stealth mode with a $26 million seed round to build a web search index designed…
NVIDIA Vera Rubin Cuts Agentic AI Costs 35x, Boosts Efficiency 30x
NVIDIA's Vera Rubin NVL72 GPU system delivers up to 30x higher throughput per megawatt than its GB300 NVL72 predecessor…
Nvidia Lands SpaceX, Nebius as Early Vera CPU Customers
Nvidia announced that SpaceX and AI cloud company Nebius will be early customers for its Vera CPU and Groq LPX…

IBM mainframe chip runs Arm and Z workloads on same cores
IBM announced a dual-architecture mainframe processor at Hot Chips that can natively execute both Arm and IBM Z…

Nvidia cuts model handoff costs with linear math KV cache transfer
Nvidia researchers have developed a technique that uses linear math to transfer key-value caches between different AI…

Nvidia Raises Flagship AI Chip Prices 17%
Nvidia is raising prices for its Grace Blackwell and Vera Rubin flagship AI chip systems by approximately 17%,…
Starcloud's $250M bet reflects orbital launch crunch
Starcloud has raised $250 million to develop orbital data centers as launch capacity becomes constrained. The funding…
Ramp launches Router, an AI model routing service
Ramp, a financial operations platform, has launched Router, an AI model routing service that allows users and companies…

Stripe Acquires OpenRouter, Pivots to AI Infrastructure
Stripe confirmed its acquisition of OpenRouter, a platform that aggregates access to hundreds of AI models including…

OpenAI Competes for OpenRouter Developer Spending
OpenAI is competing with Anthropic for developer spending on OpenRouter, a multi-provider API service that lets…

Snowflake adds auto-routing to cut AI query costs up to 3x
Snowflake has launched dynamic model routing in its Cortex AI Gateway, automatically selecting the most cost-effective…

Nvidia's Upgrade Dilemma: Push New Chips or Protect Old Ones
Nvidia faces a strategic tension between pushing customers to upgrade to new chips annually and assuring them that…
Groq pivots to neocloud with $350M funding round
Groq, a former AI chipmaker, raised $350 million at a $3.5 billion valuation while shifting its business model toward…

