Pinggy Blog

Recently updated posts

The latest revisions across our writing on tunnels, networking, self-hosting, self-hosted LLMs, and AI - freshest updates first.


Small LLMs That Fit in 8GB: The Best Models to Self-Host in 2026

Local LLM Self-Hosted AI Ollama

Which open-weight LLMs actually fit in 8GB of VRAM or RAM in 2026, with measured file sizes, KV cache math from published configs, and Ollama commands for Qwen3.5, Gemma 4, Ministral 3, Granite 4.1, Nemotron 3 Nano, and Phi-4-mini.

JetKVM Mini: A $39 KVM Over IP, and What to Do When Its Cloud Can't Reach You

KVM Over IP Remote Access Self-Hosted

JetKVM Mini packs a full KVM-over-IP into a 42mm aluminum shell for $39, shipping October 26, 2026. Its remote access runs through a cloud dashboard and Tailscale - here is what happens when WebRTC gets blocked, and how to open a backup path with a plain SSH tunnel.

Best AI Tools for Coding in 2026

AI Coding Tools AI Coding Agents Cursor

The AI coding tools worth installing in late 2026: Cursor after the SpaceX deal, Google Antigravity, GitHub Copilot's credit billing, Cline, Kilo Code, Zed, Kiro, and terminal agents like Claude Code, OpenAI Codex, DeepSeek Harness, OpenCode, Pi, Amp, Grok Build, and Droid.

Best Hardware to Self-Host LLMs for Coding and Agentic Work in 2026

AI Hardware Local LLM AI Coding Agents

A buying guide for running coding agents on your own hardware. Why prompt caching makes generation speed the number that matters, where a cold cache costs you minutes instead, a comparison table of Mac Studio M5 Ultra, RTX 5090, Radeon AI PRO R9700, Strix Halo and DGX Spark with September 2026 prices, and the memory every open-weight coding model actually needs.