edge-ai
3 articles
Gemma 4 Is Here: Google Says It Uses the Same Breakthrough Tech as Gemini 3
Google releases the Gemma 4 open-source model family, including 31B Dense, 26B MoE, and E2B/E4B edge models under Apache 2.0. With 256K context, function calling, and multimodality, Google claims it beats models 20x its size on Arena.
Paweł Huryn Claims: Holo3 with 3B Active Parameters Beats GPT-5.4 and Opus 4.6 at Computer Use
Paweł Huryn posted on X claiming H Company's Holo3 beat GPT-5.4 and Opus 4.6 at computer use tasks with just 3B active parameters. He says it's a sparse MoE fine-tuned from Qwen3.5 and could theoretically run on a single GPU.
Reasoning Model on Your Phone? Liquid AI Fits LFM2.5-1.2B Into ~900MB — Edge Agents Are Getting Real
Liquid AI's LFM2.5-1.2B-Thinking (1.17B param, 32K context) runs on-device (<1GB mem). Claims to match/beat Qwen3-1.7B on reasoning, with faster decoding & fewer tokens. Strong for tool-calling/data extraction, but weaker on knowledge-heavy tasks.