Hackernews posts about Llama.cpp
Related:
M2 Max
- llama.cpp (llama.app)
- Building a Rust Inference Engine That Matches Llama.cpp (www.fratepietro.com)
- Release v0.2.0 · ggml-org/llama.cpp (github.com)
- Show HN: FEDERaiDE, a TUI harness with P2P multi-agent routing and built in IDE (federaide.rocklab.in)
- Show HN: VelocityNote – A tiny Markdown notebook with local AI (velocitynote.app)
- Show HN: Otaku – A Roleplay Terminal Client (github.com)
- Show HN: Anjadhe – privacy first AI assistant, no account, no server DB (www.anjadhe.com)
- Show HN: Gainz.fast – Local Inference, Faster (gainz.fast)
- Vision Now Available in Llama.cpp (github.com)
- Llama.cpp guide – Running LLMs locally on any hardware, from scratch (steelph0enix.github.io)
- Heap-overflowing Llama.cpp to RCE (retr0.blog)
- Llama.cpp supports Vulkan. why doesn't Ollama? (github.com)
- Ollama violating llama.cpp license for over a year (github.com)
- Llama.cpp AI Performance with the GeForce RTX 5090 Review (www.phoronix.com)
- Mistral Integration Improved in Llama.cpp (github.com)
- Llama.cpp: Add GPT-OSS (github.com)
- Llama.cpp now has an official website: llama.app (twitter.com)
- Llama.cpp Now Part of the Nvidia RTX AI Toolkit (developer.nvidia.com)
- LLama.cpp Got Screwd (github.com)
- Llama.cpp now has an official website: llama.app (llama.app)
- DeepSeek-R1 speeds up llama.cpp code by x2 (github.com)
- Llama.cpp AI Performance with the GeForce RTX 5090 (www.phoronix.com)
- Ollama vs. Llama.cpp – Quick Benchmark (blawg.pages.dev)
- Show HN: Run Llama.cpp In-Process from Java with Project Panama FFM (deemwar-products.github.io)
- WebGPU support in llama.cpp (reeselevine.github.io)