Tagged "marktechpost"
- Thinking Machines Lab Releases Inkling-Small: A 276B Total, 12B Active Open Weights Multimodal MoE Model
- NVIDIA AI Releases Molt: A PyTorch-Native Agentic Reinforcement Learning Framework
- Deploying 1-Bit Bonsai-27B with PrismML and llama.cpp for Local Inference
- WebBrain: Open-Source Local AI Browser Agent for Task Automation
- Meet EverOS: An Open Source Markdown-First Agent Memory Runtime With Hybrid BM25 + Vector Retrieval
- Liquid AI Ships LFM2.5-230M with Broad Framework Support for On-Device Inference
- NVIDIA Dynamo Snapshot Accelerates AI Inference Startup on Kubernetes
- Meet Memory OS: A 6-Layer Open-Source Memory Stack Built on Hermes Agent
- Meet EAGLE 3.1: The Speculative Decoding Algorithm That Fixes Attention Drift in LLM Inference
- Elastic KV Cache Memory Breakthrough Enables Efficient Bursty LLM Serving and GPU Sharing
- Coding Implementation to Run Qwen3.5 Reasoning Models Distilled With Claude-Style Thinking Using GGUF and 4-Bit Quantization
- NVIDIA Releases Dynamo v0.9.0: Infrastructure Overhaul With FlashIndexer and Multi-Modal Support