← Back to Categories
Local AI
Running and self-hosting local LLMs and AI tools: Ollama, llama.cpp, vLLM, LM Studio, RamaLama, quantization, and the hardware to run them.

Guide
LM Studio Guide: Run Local LLMs on Your Mac Fast
Learn to install LM Studio, pick the right model, and chat with local LLMs on macOS—privacy-first, no cloud required.
May 10, 202618 min read

Guide
LM Studio Windows Guide: Run Local LLMs Effortlessly
Learn how to install LM Studio on Windows, choose the best model, download quantized LLMs, and run your first local AI chat fast.
Jun 11, 202623 min read

Guide
VS Code's GitHub Copilot Chat + LM Studio Local API for Offline Coding
Keep Copilot Chat prompts private with an LM Studio local OpenAI-compatible backend—fully offline, no token costs, and easy setup steps.
Jun 13, 202621 min read

Guide
Running Local LLMs with RamaLama and Docker on a Mac: A Hands-On Guide
I ran RamaLama on an Apple Silicon Mac with Docker to see how it handles local LLMs: install, first model, an OpenAI-compatible API, and the one macOS GPU gotcha to know.
Aug 25, 20269 min read