Tagged "distillation"
- Gemma 4 Turns Ancient Laptops Into Dedicated Local LLM Inference Stations
- Liquid AI Releases LFM2.5 Q4_0 Checkpoints from Quantization-Aware Distillation
- Chrome and Edge Browsers Quietly Deploy Up to 20GB AI Models on Windows 11
- LFM2.5-2.6B: On-Device Agentic Model With 128K Context and Tool Calling
- Your Smartwatch Now Detects a Heart Irregularity in Milliseconds – Without Ever Touching the Cloud
- Samsung's Newest Foldable Phones Use Google's Gemini Nano 4 On-Device AI Model
- OPPO Launches Xiaobu Next Beta, Debuts On-Device Multi-Agent System on Smartphones
- On-Device AI vs Cloud AI: Which One Should Power Your Next Phone?
- Sunday Reboot: Shrinking Models and an On-Device AI Future
- Samsung Galaxy Watch 9 to Feature Snapdragon Wear Elite Chip: Report
- Apple in Talks with PrismML to Shrink AI Models 15x for iPhone Deployment
- Apple Boosts On-Device AI, Partners With PrismML to Enable Running Large Models Locally on iPhone
- CEO Calls for Lower AI Pricing to Enable Practical Labor Automation Deployment
- Local LLM Performance Gap With Frontier Models Smaller Than Expected
- Hermes MoA Virtual Models: 8% Higher Than Opus 4.8, 11% Higher Than GPT 5.5
- Brilliant Labs Halo: Open-Source AI Glasses for On-Device Intelligence
- Show HN: Lowfat – Pluggable CLI Filter Saving 91.8% of LLM Tokens
- On-Device AI to Be in 80% of Wearables by 2032
- Google Limits Gemini Intelligence to New Flagships—Hardware Requirements for Local Deployment
- Small On-Device AI Model Beats Claude Sonnet 4.5 and GPT-5
- Perplexity Brings On-Device AI Workflow to Macs with 'Personal Computer' Feature
- Agentic AI Community Focus: Building Local Agents in 2026
- Major Smartphone Brands Introduce Advanced On-Device AI Features
- NVIDIA Nemotron 3 Nano Omni Powers Multimodal Agent Reasoning in a Single Efficient Open Model
- Google's Gemma 4: Powerful AI Models Optimized for Your Phone and Laptop
- Economic Implications of AI Adoption: Why Local Deployment Matters for Cost Control
- Anker Unveils 'Thus' Chip to Bring On-Device AI Across Product Line
- Llama 4 Scout on MLX: The Complete Apple Silicon Guide (2026)
- Researchers Achieve 1-Bit Quantization of OLMo-3 7B Using Distillation
- Apple Research Shows Self-Distillation Significantly Improves Local Code Generation
- Apple Gets Full Gemini Access and Uses Distillation to Build Lightweight On-Device AI
- Coding Implementation to Run Qwen3.5 Reasoning Models Distilled With Claude-Style Thinking Using GGUF and 4-Bit Quantization
- Samsung Galaxy A37 and A57 5G Launch with On-Device AI Capabilities in India
- Apple Plans Slimmed-Down Gemini Models for Local iPhone AI Features
- Ultra-Large 400B-Class LLM Runs on iPhone in Test
- LLM Neuroanatomy II: Modern LLM Hacking and Hints of a Universal Language
- Ultra-Compact 28M Parameter Models Show Promise for Specialized Domain Tasks
- India's Mobile-First AI Strategy Could Accelerate Local Inference Adoption in Emerging Markets
- Ex-Manus Backend Lead Shares: Moving Beyond Function Calling in Agent Design
- Qwen 3.5 Ultra-Compact Models Enable On-Device AI from Watches to Gaming
- Qwen 3.5 Small Expands On-Device AI to Phones and IoT with Offline Support
- Show HN: TLDR – Free Chrome Extension for AI-Powered Article Summarization
- Apple Intelligence, Galaxy AI, Gemini: Why Your AI-Powered Phone Is Worth Repairing
- How to Run High-Performance LLMs Locally on the Arduino UNO Q
- Meta Reveals AI-Packed Smartwatch In 2026 – Why Wearables Shift Now
- The Real AI Competition Is Closed-Source vs Open-Source, Not America vs China
- Anthropic Reveals Industrial-Scale Distillation Attacks by Chinese AI Labs
- Future of Mobile AI: What On-Device Intelligence Means for App Developers
- Taalas Etches AI Models onto Transistors to Rocket Boost Inference
- Mirai Secures $10M to Optimize On-Device AI Amid Cloud Cost Surge
- Sarvam Brings AI to Feature Phones, Cars, and Smart Glasses