Hackernews posts about FP16
- Show HN: Cactus Needle 3: 8-29MB automation models can match DeepSeek V4 Flash (cactuscompute.com)
- Historic Business Machines Collection (cdm16471.contentdm.oclc.org)
- ONNX Runtime and CoreML May Silently Convert Your Model to FP16 (ym2132.github.io)
- Running the Deepseek-R1 671B Model at FP16 Fidelity on AMD EPYC CPUs (www.servethehome.com)
- 90T/s on my iPhone llama3.2-1B-fp16 (www.reddit.com)
- Number to GPU Float Converter: FP32, FP16, BF16, FP8, and FP4 Explained (www.bestgpusforai.com)
- TTS engines: WebSocket vs. sync is 5.5x, INT8 slower than fp16 on M4 (ai.gopubby.com)
- PyTorch 2.6 Delivers FP16 Support for x86 CPUs, Better Intel GPU Experience (www.phoronix.com)
- PyTorch 2.6 Delivers FP16 Support for x86 CPUs, Better Intel GPU Experience (www.phoronix.com)
- Show HN: OpenGraviton – Run 500B+ parameter models on a consumer Mac Mini (opengraviton.github.io)
- Show HN: I made Qwen3.5-4B 13% smarter by compressing it to 4-bit (huggingface.co)