Hackernews posts about gguf
- GLM-5.3-Flash-GGUF (huggingface.co)
- Unsloth/Qwen3.8-Flash-Next-GGUF (huggingface.co)
- Axera AX8850 LLM running ggufs (github.com)
- GLM5.3 unsloth GGUFs are now up (huggingface.co)
- Per-tensor layout maps for GGUF quantization (huggingface.co)
- Ternary-Bonsai-8B-Gguf (huggingface.co)
- Show HN: Single-File GGUF Inference (wasm-gguf.netlify.app)
- Show HN: PicoLM v1.0-rc1 (github.com)
- Qwen3.8-Flash-Next non-uniform quantization runs on 2 RTX3090s (huggingface.co)
- Show HN: Find the cheapest GPU to run your model (www.gpufind.cloud)
- Unsloth Dynamic 3.0 GGUFs (unsloth.ai)
- Unsloth Dynamic 2.0 GGUFs (unsloth.ai)
- Ollama and gguf (github.com)
- Unsloth Qwen3.8-27B GGUF files (huggingface.co)
- Show HN: Quant Picker – which GGUF file fits your model and machine (vettedconsumer.com)
- Ollama can run any GGUF Model on Hugging Face Hub now (huggingface.co)
- Benchmark GGUF model with ONE line of code (github.com)
- Qwen3.5 GGUF Benchmarks (unsloth.ai)
- GGUF Quantization Compared: Q4_K_M vs. IQ4_XS vs. IQ4_NL (kaitchup.substack.com)
- Choosing a GGUF Model: K-Quants, IQ Variants, and Legacy Formats (kaitchup.substack.com)
- GGUF vs. MLX when to choose what when running Local LLMs (muhammadraza.me)
- Choosing a GGUF Model: K-Quants, IQ Variants, and Legacy Formats (kaitchup.substack.com)
- Unsloth/DeepSeek-V4-Pro-0813-GGUF (huggingface.co)
- Show HN: Host any GGUF model in one command (github.com)
- GGML GGUF File Format Vulnerabilities (www.databricks.com)
- Unsloth Dynamic GGUFs DeepSeek (671B) outperforms SOTA models (docs.unsloth.ai)
- Unsloth Dynamic 2.0 GGUFs (docs.unsloth.ai)
- DeepSeek-R1 on iPhone? (DeepSeek-R1-Distill-Qwen-1.5B-GGUF) (huggingface.co)