DEV Community

#localllm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
How I upgrade xiaoai speaker local llm: Sub-200ms AI

How I upgrade xiaoai speaker local llm: Sub-200ms AI

Comments 1
10 min read
Best Local LLM for Coding: 8GB to 24GB VRAM Picks

Best Local LLM for Coding: 8GB to 24GB VRAM Picks

Comments
4 min read
llama.cpp vs Ollama: Which Should You Run in 2026?

llama.cpp vs Ollama: Which Should You Run in 2026?

Comments
5 min read
Ollama vs LM Studio: Which Local LLM Tool Should You Use?

Ollama vs LM Studio: Which Local LLM Tool Should You Use?

Comments
4 min read
What Happens When You Ask an LLM a Question

What Happens When You Ask an LLM a Question

Comments
8 min read
Run vLLM on Kubernetes with Minikube, WSL2 and NVIDIA GPU

Run vLLM on Kubernetes with Minikube, WSL2 and NVIDIA GPU

Comments
13 min read
I Ran DeepSeek V4 Flash Across Two DGX Sparks Over Ethernet

I Ran DeepSeek V4 Flash Across Two DGX Sparks Over Ethernet

Comments
11 min read
VRAM and RAM for local LLMs — honest planning bands, not a GPU tier list

VRAM and RAM for local LLMs — honest planning bands, not a GPU tier list

Comments
4 min read
Running Ollama on a 32 GB MacBook Air: A Practical First Setup

Running Ollama on a 32 GB MacBook Air: A Practical First Setup

Comments
6 min read
Running llama.cpp on a 32 GB MacBook Air: A Direct Comparison with Ollama

Running llama.cpp on a 32 GB MacBook Air: A Direct Comparison with Ollama

Comments 1
9 min read
Running a 35B MoE Model on an 8 GB Laptop GPU: Testing FreeToken

Running a 35B MoE Model on an 8 GB Laptop GPU: Testing FreeToken

Comments 3
7 min read
Temperature 0 is not reproducible. I measured 30 percent of my output changing between identical runs.

Temperature 0 is not reproducible. I measured 30 percent of my output changing between identical runs.

1
Comments 1
3 min read
I told the model to separate fields with <TAB>. It did exactly that, and I lost 79 percent of my data.

I told the model to separate fields with <TAB>. It did exactly that, and I lost 79 percent of my data.

1
Comments
3 min read
Our 4B beat Claude Opus on a 440K-token corpus. Then it came last on the public benchmark.

Our 4B beat Claude Opus on a 440K-token corpus. Then it came last on the public benchmark.

2
Comments
4 min read
A 4B on a 6GB laptop matched frontier-model accuracy on aggregation — except when the answer is a number

A 4B on a 6GB laptop matched frontier-model accuracy on aggregation — except when the answer is a number

1
Comments
4 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.