Ollama

Blog posts tagged “Ollama”


Small LLMs That Fit in 8GB: The Best Models to Self-Host in 2026

Local LLM Self-Hosted AI Ollama

Which open-weight LLMs actually fit in 8GB of VRAM or RAM in 2026, with measured file sizes, KV cache math from published configs, and Ollama commands for Qwen3.5, Gemma 4, Ministral 3, Granite 4.1, Nemotron 3 Nano, and Phi-4-mini.

Picking the Right Hardware to Run LLMs Locally in 2026

Local LLM Self-Hosted AI Apple Silicon

A practical hardware guide for self-hosting LLMs in 2026. Compare consumer GPUs, Apple Silicon, enterprise cards, and pre-built AI workstations. Find the right setup for 7B to 405B models at every budget.

How to Host Your AI App on Google Colab for Free

Google Colab Ollama Pinggy

Learn how to build and host a complete AI-powered Flask application on Google Colab with free GPU access. Use Ollama for LLM inference and Pinggy for public access - no server costs required.

Running Ollama on Google Colab Through Pinggy

Ollama Google Colab Pinggy

Learn how to run Ollama models on Google Colab and access them remotely using Pinggy tunneling. Complete setup guide with OpenWebUI integration for free AI model hosting.

How to Self-Host Any LLM – Step by Step Guide

Self-Hosted AI Ollama Open WebUI

Complete guide to self-hosting large language models locally using Ollama and Open WebUI with Docker. Learn to run AI models privately with full control over your data.

Self-Host AI Agents Using n8n and Pinggy

N8n Ai Self-Hosted

Learn how to self-host n8n's AI Starter Kit and access your AI workflows remotely with Pinggy. Step-by-step guide to privacy-focused AI solutions with local LLMs.