DEV Community

#localllm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
How Much RAM Do You Need for Local LLMs on a Mac?

How Much RAM Do You Need for Local LLMs on a Mac?

Comments
3 min read
Claude Code MCP setup and usage rules

Claude Code MCP setup and usage rules

Comments
3 min read
Clone your own voice with a 17-second recording

Clone your own voice with a 17-second recording

Comments
2 min read
Prompt engineering: How to fix instructions that AI ignores

Prompt engineering: How to fix instructions that AI ignores

Comments
3 min read
How I upgrade xiaoai speaker local llm: Sub-200ms AI

How I upgrade xiaoai speaker local llm: Sub-200ms AI

Comments 1
10 min read
What Happens When You Ask an LLM a Question

What Happens When You Ask an LLM a Question

Comments
8 min read
Run vLLM on Kubernetes with Minikube, WSL2 and NVIDIA GPU

Run vLLM on Kubernetes with Minikube, WSL2 and NVIDIA GPU

Comments
13 min read
I Ran DeepSeek V4 Flash Across Two DGX Sparks Over Ethernet

I Ran DeepSeek V4 Flash Across Two DGX Sparks Over Ethernet

Comments
11 min read
VRAM and RAM for local LLMs — honest planning bands, not a GPU tier list

VRAM and RAM for local LLMs — honest planning bands, not a GPU tier list

Comments
4 min read
Running Ollama on a 32 GB MacBook Air: A Practical First Setup

Running Ollama on a 32 GB MacBook Air: A Practical First Setup

Comments
6 min read
Running llama.cpp on a 32 GB MacBook Air: A Direct Comparison with Ollama

Running llama.cpp on a 32 GB MacBook Air: A Direct Comparison with Ollama

Comments 1
9 min read
Ollama vs LM Studio: Which Local LLM Tool Should You Use?

Ollama vs LM Studio: Which Local LLM Tool Should You Use?

Comments
4 min read
llama.cpp vs Ollama: Which Should You Run in 2026?

llama.cpp vs Ollama: Which Should You Run in 2026?

Comments
6 min read
Best Local LLM for Coding: 8GB to 24GB VRAM Picks

Best Local LLM for Coding: 8GB to 24GB VRAM Picks

Comments
5 min read
Running a 35B MoE Model on an 8 GB Laptop GPU: Testing FreeToken

Running a 35B MoE Model on an 8 GB Laptop GPU: Testing FreeToken

Comments 3
7 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.