#local-models

4 articles

NVIDIA Just Dropped a 30B Model That Runs on a Single GPU — And It's Built for Agents

Nemotron 3.5 Lightning hit August 11: 30B MoE with 3B active params, 1M context, 4x throughput. Kilo Code already integrated it. H…

GitHub Copilot Just Got Memory — and It Runs Free Local Models Too

GitHub Copilot for JetBrains now remembers your preferences across sessions and runs free local models through Ollama. Beginner-fr…

Pi 0.81–0.84: Fullscreen TUI, Mermaid Diagrams, Local Models via llama.cpp, and Opus 5 on Copilot

Four Pi releases in three weeks add a fullscreen TUI mode with Mermaid/LaTeX rendering, per-directory context overrides, local LLM…

How to Run Coding Agents with Ollama: The Complete Local Setup Guide

Stop paying per token. This guide covers installing Ollama, picking the right local model, and wiring it into Claude Code, OpenCod…