Checked for new stories 12m ago

Updates on Llama Cpp

Every AI story we track on Llama Cpp — 10 stories so far, each summarized in our own words and linked back to the publisher that reported it.

Pulled from 124 sources

This week

Machine Learning4 min read

Nous Research Adds One-Click Local Model Setup to Hermes Desktop

MarkTechPost

This month

Machine Learning9 min read

Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps

Hugging Face

Speed Up LLM Inference with DSpark Speculative Decoding

KDnuggets
Machine Learning5 min read

Liquid AI Open-Sources Pipette: A Reproducible Benchmarking Suite That Measures On-Device Models, Quantization, Runtime and Hardware Together

MarkTechPost
Machine Learning6 min read

Up to 3.2x Faster Inference with LFM2.5-DSpark

Covered by 3 sources
Dev5 min read

Run Muse Glimmer for Local Vibe Coding with llama.cpp, DFlash, and Pi

KDnuggets
Machine Learning3 min read

LFM2.5 Q4\_0 Checkpoints from Quantization-Aware Distillation

Hugging Face
Agents4 min read

AI Agents that are Portable

Hacker News
Dev3 min read

AI;dr – AI wrote it, you shouldn't have to read it

Hacker News

Homebench – Benchmark local LLMs for speed, memory, and quality

Hacker News
That's everything we have on Llama Cpp right now