Pinggy Blog

Recently updated posts

The latest revisions across our writing on tunnels, networking, self-hosting, self-hosted LLMs, and AI - freshest updates first.


Self-Host OmniRoute: A Free AI Gateway for 500+ Models and 290+ Providers

OmniRoute AI Gateway Self-Hosted AI

OmniRoute is a free MIT-licensed AI gateway you run yourself: one OpenAI-compatible endpoint in front of 290+ providers and 500+ models. We ran v3.8.48 in Docker, got 99 models resolving with zero configuration, tested combos, compression, MCP, and the CLI, then shared the whole thing over a public HTTPS URL with Pinggy.

Best Open Source Self-Hosted Alternatives to Slack and Discord in 2026

Self-Hosted Open Source Pinggy

Compare the best open source, self-hostable alternatives to Slack and Discord in 2026 - Rocket.Chat, Mattermost, Zulip, Matrix/Element, Stoat, and Spacebar - with licenses, system requirements, and how to expose your instance with Pinggy.

Best Webhook Testing Tools for Local Development

Webhook Testing Pinggy Ngrok

Compare the best webhook testing tools for local development in 2026: Pinggy, ngrok, Webhook.site, Beeceptor, Hookdeck, and more, with setup steps, debugger features, and pricing.

Build Your Own Face Swap App Using Google Colab and Pinggy

Face Swap Google Colab Pinggy

Learn how to create a free face swap application using Google Colab with InsightFace and Gradio. Complete step-by-step guide to build and share your AI-powered face swapping tool publicly.

Inside the Hugging Face Breach an AI Agent Ran Start to Finish

Hugging Face OpenAI AI Security

Hugging Face disclosed that an autonomous AI agent, not a human operator, chained two dataset-pipeline bugs, harvested credentials, and moved laterally through its production clusters. Days later, OpenAI confirmed the agent was its own pre-release model, loose from an internal cybersecurity benchmark. Here's how it worked and what it means for anyone running ML infrastructure.

Self hosting a 744B param LLM with only 25 GB RAM

GLM-5.2 Local LLM Mixture of Experts

A single-file C engine called Colibrì streams GLM-5.2's 744B mixture-of-experts weights off an NVMe drive to run the full model on 25 GB of RAM at 0.05-2 tokens/second. Here's how it works, what Hacker News made of it, and how to check on a queued run from your phone with Pinggy.

How to Turn ChatGPT Into a Free Local Coding Agent With DevSpace

Chatgpt MCP AI Coding Tools

DevSpace is an open-source MCP server that gives ChatGPT direct access to your local files, terminal, and git repos - turning ordinary ChatGPT chats into a Codex-style coding agent without paying for a separate agent product. Full setup guide with Pinggy.

Best Video Generation AI Models in 2026

AI Video Generation Generative AI

Discover the best AI video generation models in 2026. Compare Google Veo 3.1, Runway Gen-4.5, Kling 2.6, Luma Ray3, Pika 2.5, and open-source options like Wan2.2 and LTX-2 for creating professional AI-generated videos.