VFF - The signal in the noise
News

Liquid AI brings edge AI to Raspberry Pi with 2.6B parameter model

Read original
Share
Liquid AI brings edge AI to Raspberry Pi with 2.6B parameter model

Liquid AI, a startup founded by former MIT computer scientists, released LFM2.5-2.6B, a 2.6 billion parameter language model designed to run on edge devices including Raspberry Pi without cloud infrastructure or GPUs. The model supports 128,000-token context windows and native tool calling, targeting agentic tasks like document management and workflow automation in regulated industries and connectivity-limited environments. Performance ranges from 30 tokens per second on smartphones to 220 tokens per second on Apple M5 Max, with the model available on Hugging Face under a custom open-weight license.

  • Liquid AI released LFM2.5-2.6B, a 2.6B parameter model optimized for edge deployment on CPUs and low-power devices
  • Model runs on Raspberry Pi and smartphones without GPUs or cloud connectivity, with throughput of 30 tokens/sec on phones and 220 tokens/sec on Apple M5 Max
  • Designed for agentic workloads including tool calling, document management, calendar automation, and robotics applications
  • Available on Hugging Face with support for llama.cpp, MLX, vLLM, SGLang, and ONNX, plus an open source fine-tuning framework called LEAP

Edge AI deployment eliminates latency, privacy, and cost constraints that cloud inference imposes. For enterprises handling regulated data or operating in connectivity-limited environments, local model execution removes barriers to AI adoption. The model's efficiency on consumer hardware expands where AI agents can operate beyond data centers.

Organizations can deploy performant AI agents at marginal cost, limited to electricity consumption. Regulated industries and those with data sensitivity concerns gain a practical path to AI automation without cloud dependencies. The trade-off between model size and capability enables cost-effective deployment of task-specific agents across distributed hardware.

  • Edge AI deployment becomes viable for enterprises with privacy or regulatory constraints, potentially shifting inference workloads away from cloud providers
  • Small models optimized for CPU performance may create a new category of enterprise applications where latency and deployment flexibility outweigh benchmark performance
  • Custom open-weight licenses require legal review by enterprises, adding friction to adoption despite technical accessibility

Monitor adoption patterns among regulated industries and enterprises with connectivity constraints. Track whether custom licensing terms become standard practice for open-weight models and whether they create legal friction. Observe if edge-optimized small models fragment the market, creating specialized model ecosystems for different deployment contexts rather than consolidation around frontier models.

OneUpAI
OneUp Your Business. Get More Done. OneUp Your Business. Get More Done. OneUp Your Business. Get More Done.
Learn More
Share

Subscribe to the newsletter

The latest stories and analysis, delivered to your inbox.

Free. No spam. Unsubscribe any time.

Related stories

Alibaba Open-Sources Qwen3.8 Trillion-Parameter Model
News

Alibaba Open-Sources Qwen3.8 Trillion-Parameter Model

Alibaba released Qwen3.8-2.4T-A95B as open weights on August 12, 2026, marking the first time a Qwen-Max-class model became publicly available. The 2.4 trillion parameter model uses a hybrid linear-plus-full-attention architecture with 95 billion activated parameters per token and supports up to 262K native context tokens, extensible to 1M. AWS published a deployment guide showing how to run the model on SageMaker HyperPod using vLLM on ml.p6-b300 instances with NVIDIA B300 Blackwell Ultra GPUs.

by Dmitry Soldatkin· AWS Machine Learning Blog
Saudi Arabia Launches Arabic AI Model With Chinese Partner
TrendingNews

Saudi Arabia Launches Arabic AI Model With Chinese Partner

Humain, Saudi Arabia's state-owned AI company, announced the humain-m3 model, an Arabic language model built on Chinese firm MiniMax's open-source M3 foundation. The model was pre-trained on more than 1 trillion tokens of Arabic content. The development represents a collaboration between Saudi and Chinese AI capabilities focused on Arabic language processing.

by Juro Osawa· The Information
OpenAI's Astra model alarms safety experts with new reasoning technique
News

OpenAI's Astra model alarms safety experts with new reasoning technique

OpenAI's new Astra model employs a technique called 'recurrent depth' that enables reasoning outside the sequential thinking pattern used by most current reasoning models. AI safety experts have raised concerns about this approach. The technique represents a departure from established reasoning architectures in large language models.

by Russell Brandom· TechCrunch AI
Anthropic cuts agent costs 75%, adds enterprise safeguards
TrendingModel Release

Anthropic cuts agent costs 75%, adds enterprise safeguards

Anthropic released Claude Fable 5.1 and Mythos 5.1, its latest large language models, alongside a 75% cost reduction for cached context reads and a new Enterprise Frontier Safeguards security architecture. The release targets enterprise deployment of persistent agents capable of multi-hour problem-solving tasks. Fable 5.1 shows significant benchmark improvements across scientific research, coding, and business workflow tasks, though results are vendor-reported rather than independently verified.

by carl.franzen@venturebeat.com (Carl Franzen)· VentureBeat AI