Skip to content
Topic

#Llama Cpp

11 articles on Llama Cpp — news, releases, guides and analysis from the SourceFeed engine.

Convert and Quantize Hugging Face Models to GGUF for llama.cpp
Tutorial 1d ago 0

Convert and Quantize Hugging Face Models to GGUF for llama.cpp

Turn any Hugging Face checkpoint into a 4-bit GGUF that runs fast and small on your own hardware.

Mariana Souza
llama.cpp Finally Competes With Its Own Wrappers

llama.cpp Finally Competes With Its Own Wrappers

Article · 1d ago4
macOS VMs Have Been Sandbagging Your GPU

macOS VMs Have Been Sandbagging Your GPU

Article · 2d ago0
Your macOS VM's GPU Was Never Slow, Just Mislabeled

Your macOS VM's GPU Was Never Slow, Just Mislabeled

Article · 2d ago1
Ante Bets Coding Agents Should Be Single Binaries

Ante Bets Coding Agents Should Be Single Binaries

Article · 3d ago3
Run Mixtral 8x7B Locally with llama.cpp and Benchmark MoE vs. Dense

Run Mixtral 8x7B Locally with llama.cpp and Benchmark MoE vs. Dense

Tutorial · 1w ago0
Old Xeons Can Run Gemma 4 at Reading Speed

Old Xeons Can Run Gemma 4 at Reading Speed

Article · 4w ago1
Mesh LLM Makes Spare GPUs Look Like One API

Mesh LLM Makes Spare GPUs Look Like One API

Article · 1mo ago0
Qwen 3.6 27B Hits the Local Development Sweet Spot

Qwen 3.6 27B Hits the Local Development Sweet Spot

Article · 1mo ago0
Going Local: The Reality of Replacing Claude and GPT

Going Local: The Reality of Replacing Claude and GPT

Article · 1mo ago1
Run a Fast, Fully Local Coding Agent on macOS

Run a Fast, Fully Local Coding Agent on macOS

Tutorial · 2mos ago0