Tagged "gemma"
- Gemma 4 Turns Ancient Laptops Into Dedicated Local LLM Inference Stations
- Phi-4 Mini vs Gemma 3 vs Llama 3.2: 128K vs 32K Context Window Comparison
- Kioxia's UFS 5.0 Embedded Flash Enables Practical On-Device AI
- Gemma 4's Quantized Models Finally Made Local AI Practical in Homelab
- Google's Gemma AI Runs Locally on a $300 Mini PC, and It Replaced ChatGPT
- On-Device AI vs Cloud AI: Which One Should Power Your Next Phone?
- Google Gemma 4 Debuts for Pixel 10 With Powerful On-Device AI Features
- Google expands on-device AI for Pixel phones with Gemma 4
- Google Rolls Out Android 17 and Gemma 4 with Advanced On-Device AI
- Google's Gemma AI Runs Locally on a $300 Mini PC, and It Replaced ChatGPT for More Than Expected
- Build Your Own Local AI Coding Agent with Gemma 4 and OpenCode
- Getting Started With NVIDIA DGX Spark: Unboxing, First Boot, Dashboard, and Running Gemma Locally
- Google's DiffusionGemma Achieves 4x Faster Text Generation for Local Deployment
- DiffusionGemma: The Developer Guide for Local Deployment
- Google Releases Gemma 4 QAT Models with Reduced Memory Requirements for Mobile and Laptop Deployment
- Apple Enhances Siri With On-Device AI for Faster, Private Voice Responses
- Google Introduces Gemma 4 QAT for Ultra-Low Memory Local Inference
- Google's New Gemma 4 12B AI Model Is Built for Laptops
- Google Releases Gemma 4 QAT Models for Local AI Deployment
- Google Releases Gemma 4 12B Model for Local Inference on 16GB Enterprise Laptops
- Bosgame Launches VTA-439 Mini PC with 86 TOPS for Practical Local AI
- Google Releases Gemma 4 12B: Encoder-Free Multimodal Model for 16GB Laptops
- Google Launches Tiny Board for Running Gemma 3 Locally
- Gemma 4: A New Budget-Focused Model in Posit AI
- BT Explainer: Google's Gemma 4 Could Put Powerful AI on Your Phone and Laptop
- Gemma 4 Replaces Entire Local LLM Stack for Many Practitioners
- Google Releases Gemma 4 Multi-Token Prediction Drafters To Accelerate AI Inference
- Airplane AI – Local NDA Safe AI Powered by Gemma
- Perplexity Brings On-Device AI Workflow to Macs with 'Personal Computer' Feature
- Google Accelerates Gemma 4 Inference Speed 3x With Multi-Token Prediction Drafters
- Google's Gemma 4 Could Put Powerful AI on Your Phone and Laptop
- Gemma 4 Just Replaced My Whole Local LLM Stack
- Google's Gemma 4 Brings Powerful AI Capabilities to Phones and Laptops
- Google's Gemma 4: Powerful AI Models Optimized for Your Phone and Laptop
- Google's Gemma 4 Could Put Powerful AI on Your Phone and Laptop
- Google's Gemma 4 Could Put Powerful AI on Your Phone and Laptop
- Google's Gemma 4 Brings Powerful On-Device AI to Phones and Laptops
- Google's Gemma 4 Finally Makes Local LLM Deployment Compelling for Practitioners
- 16 Ways to Make a Small Language Model Think Bigger
- Gemma 4 Just Replaced My Whole Local LLM Stack
- Gemma 4 Just Replaced My Whole Local LLM Stack
- Google's Gemma 4: The Most Practical Local LLM Despite Not Being The Smartest
- Running Gemma 4 on an iPhone 13 Pro
- Google's Gemma 4 Brings Game-Changing Performance to Local Laptop Inference
- Speculative Decoding Achieves 29% Speed Boost for Gemma-4 31B
- Audio Processing Support Lands in llama.cpp with Gemma-4
- Google Gemma 4 Delivers Exceptional Speed and Accuracy for Local Inference
- Google's Gemma 4 Brings Free Agentic AI to Your Phone With Zero Data Leaving the Device
- Critical Unsloth Gemma-4 Chat Template Updates for Tool Calling
- Gemma 4 31B vs Qwen 3.5 27B: Comprehensive Long Context Benchmark
- Community Reverse Engineers Gemma 4 Multi-Token Prediction Capability
- Gemma 4 Template Improvements Enhance Tool Use and Dialog Compliance
- Gemma 4 Support Stabilized in Llama.cpp
- Gemma 4 GGUF Models Updated with Critical Quantization Fixes
- Google AI Edge Gallery Showcases Offline Inference with Gemma 4
- Google's Gemma 4 Brings Powerful On-Device AI to Android and iOS
- Gemma 4 Achieves Top Multilingual Performance Across European Languages
- Google Launches Offline AI Dictation App for iOS with Gemma
- Gemma 4 26B Achieves Impressive Local Performance With Proper Configuration
- AMD Announces Day 0 Support for Google Gemma 4 Across Processors and GPUs
- TurboQuant-Optimized llama.cpp Fork Delivers GFX906 GPU Acceleration
- Google AI Edge Gallery Tops App Store Charts with On-Device Gemma 4
- Real-time Multimodal AI on Apple Silicon: Gemma E2B Demo Shows Practical Edge Deployment
- Gemma 4 31B Achieves Exceptional Performance on Local Hardware
- Context Window Optimization: Extending Gemma 4 Context Length Through Efficient Projection Quantization
- Gemma 4 31B Achieves Third Place on FoodTruck Bench, Beating Larger Models
- Gemma 4 26B MoE Emerges as Optimal All-Around Local Model for Consumer Hardware
- Apple Research Shows Self-Distillation Significantly Improves Local Code Generation
- NVIDIA and Google Optimize Gemma 4 AI Models for Local RTX Deployment
- Google Launches Gemma 4 For Advanced On-Device AI
- Gemma 4 31B Outperforms GLM 5.1 in Real-World Testing
- Gemma 4 KV Cache Memory Issues Fixed in llama.cpp
- AMD Rolls Out Gemma 4 Model Support Across Full Range of GPUs & CPUs
- Gemma 4 26B A4B Outperforms Qwen 3.5 35B on Apple Silicon
- Gemma 4 2B Successfully Runs on Raspberry Pi 5
- NVIDIA Accelerates Gemma 4 for Local Agentic AI on RTX GPUs
- VRAM Optimization Technique Cuts Gemma 4 Memory Usage by 3x
- Google Gemma 4 Released with GGUF Quantizations
- Google Launches Gemma 4 Open Models for Local On-Device AI
- Gemma 4 Makes Local AI Agents Practical
- AMD Provides Day 0 Support for Gemma 4 on Ryzen AI Processors and GPUs
- Gemma 4 Shows Strong Reasoning Performance with Thinking Tokens
- Gemma 4 on Arm: Optimized On-Device AI for Mobile and Edge Deployment
- April 2026 TLDR Setup for Ollama and Gemma 4 26B on a Mac mini
- O-TITANS: Orthogonal LoRA Framework for Gemma 3 with Google TITANS Memory Architecture