VFF - The signal in the noise
Research

Data Infrastructure, Not AI Models, Limits Agent Success

Read original
Share
Data Infrastructure, Not AI Models, Limits Agent Success

A MIT Technology Review Insights report based on a survey of 300 data and technology executives finds that legacy data systems are a major blocker to AI agent adoption and effectiveness. Organizations with mature data infrastructure, termed 'data leaders,' report significantly higher trust in agent decisions and fewer scaling constraints than 'data laggards.' The research suggests that without modernizing data systems, enterprises will struggle to realize ROI from agentic AI despite widespread adoption plans.

  • AI agents currently have access to only 45% of company data on average, dropping to 30% or less in lagging organizations but exceeding 70% in data leaders
  • Only half of surveyed organizations trust their AI agents' decisions, while 100% of data leaders report full trust, indicating data quality directly correlates with agent reliability
  • Two-thirds of data laggards cite legacy systems as limiting agent scaling and decision speed, compared to just 8% of data leaders
  • 100% of respondents plan to deploy agentic AI within two years, with 69% expecting widespread use, creating urgent pressure to modernize data infrastructure

As organizations race to deploy AI agents, a critical gap is emerging between technical capability and operational readiness. The survey reveals that data infrastructure, not AI models, is the limiting factor for most enterprises. Without addressing legacy system constraints, companies risk deploying agents that lack the data access and context needed to make reliable decisions at scale.

For executives evaluating AI agent investments, the research shows that ROI depends primarily on data modernization, not agent sophistication. Organizations that delay data system upgrades will struggle with agent accuracy, speed, and scaling, potentially wasting significant capital on agent deployments. Data leaders are already capturing competitive advantage through faster, more trustworthy agent operations.

  • Data infrastructure modernization should precede or accompany AI agent deployment, not follow it, to avoid costly rework and underperformance
  • Organizations need to prioritize unified access to both structured and unstructured data across enterprise systems, including supply chain, point-of-sale, and HR data
  • Data governance frameworks that embed business context are becoming operational requirements, not optional compliance measures, as agents make autonomous decisions

Monitor how quickly enterprises move to modernize legacy data systems in response to agent deployment timelines. Watch for emerging data infrastructure vendors targeting the agent-readiness gap, and track whether organizations that delay data modernization report lower agent adoption success rates or higher failure rates in production deployments.

Share

Subscribe to the newsletter

The latest stories and analysis, delivered to your inbox.

Free. No spam. Unsubscribe any time.

Related stories

Meta Deploys Thousands of Engineers to Train Coding AI
TrendingNews

Meta Deploys Thousands of Engineers to Train Coding AI

Meta is deploying its in-house coding agent MetaCode to thousands of engineers to improve the coding capabilities of its AI models and close the gap with Anthropic and OpenAI. VP Maher Saba has asked engineers to submit at least one code change per week for review and integration. The feedback loop has already improved Meta's latest model, Muse Spark 1.1, and will be used to train an upcoming model called Watermelon.

by Jyoti Mann· The Information
AI Drug Discovery Hits a Data Wall
TrendingNews

AI Drug Discovery Hits a Data Wall

AI is accelerating drug discovery by enabling predictive design of candidates and hit identification at scale, but the technology is exposing critical gaps in data quality and lab infrastructure. Drug companies are hitting a 'data wall' where publicly available datasets lack the structure and diversity needed to train accurate models, while lab teams struggle to validate the growing volume of AI-generated compounds. Success depends on closing the loop between computational prediction and experimental validation through better data collection and integration.

by MIT Technology Review Insights· MIT Technology Review
Brain Waves Join Video as Physical AI Training Data
TrendingNews

Brain Waves Join Video as Physical AI Training Data

Frontier physical AI models are moving beyond video training data to incorporate multiple camera angles, dense annotation, and brain wave readings as training inputs. The shift reflects growing recognition that traditional video datasets alone are insufficient for training AI systems that interact with the physical world. Brain wave data represents an emerging frontier in multimodal training approaches for robotics and embodied AI.

by Tim Fernholz· TechCrunch AI
Mercor's $614M Revenue Surge Hinges on AI Lab Spending

Mercor's $614M Revenue Surge Hinges on AI Lab Spending

Mercor, a three-year-old data startup that trains AI models through contractor networks, generated $614 million in gross revenue in the first half of 2026, up 70% from all of 2025. The company's growth is heavily concentrated among AI foundation model makers, with 91% of first-half revenue coming from customers like OpenAI, Anthropic, and Google DeepMind. This revenue concentration reveals both the startup's market traction and its dependency on a narrow customer base.

by Julia Hornstein· The Information