Local AI

Technical Report Demonstrates Qwen3.8-27B Running on RTX 4070 Ti Super
A technical report confirms the Qwen3.8-27B model runs on a 16GB RTX 4070 Ti Super GPU, achieving up to 50 tokens per second across 100,000 context tokens.

Kimi Work Desktop Application Launches with Local Parallel AI Agent Swarm
Kimi Work desktop launches as a local AI agent running up to 300 parallel sub-agents on macOS and Windows to automate browser and financial tasks.

Microsoft Build Details High-Performance Local AI Hardware Specifications
Microsoft Build has revealed high-performance hardware specifications that enable developers to run trillion-parameter models locally on Windows devices.

NVIDIA and Microsoft Partner on Agentic AI PC
NVIDIA and Microsoft are reportedly collaborating on a new class of personal computers designed to run artificial intelligence models locally.

Dell Deskside Agentic AI Enables Local AI Workflows from Desk to Data Center
Dell Technologies introduces Dell Deskside Agentic AI, a new solution for deploying agentic AI locally with predictable costs, supporting models from 30 billion to 1 trillion parameters.

Ollama 0.19 Powered by MLX Brings Fastest AI Performance on Apple Silicon
Ollama 0.19 is now powered by Apple’s MLX framework, delivering up to 57% faster prefill speeds and nearly double the decode performance on Apple Silicon. The update also introduces NVFP4 quantization support and smarter caching for coding agents.

Perplexity Personal Computer Brings AI Agents to Mac Mini Files
Perplexity Personal Computer allows AI agents to access files on a Mac mini by coordinating multiple AI models to complete user tasks. The product is currently available via waitlist only.
