Hackernews posts about Llama 3
Llama 3 is an AI-powered chatbot that uses pure NumPy to generate human-like conversations in a browser-based interface.
- Show HN: Willow Voice – Free AI Dictation (willowvoice.com)
- 6.4x faster than llama.cpp, 3.9x faster than MLX (www.basecompute.co)
- Show HN: Gainz.fast – Local Inference, Faster (gainz.fast)
- TurboPrefill: 3.27× Prefill Speedup in Llama.cpp (devpost.com)
- Llama-3.3-70B-Instruct (huggingface.co)
- Llama 3.1 Omni Model (github.com)
- Show HN: Llama 3.3 70B Sparse Autoencoders with API access (www.goodfire.ai)
- Meta's Llama 3.1 can recall 42 percent of the first Harry Potter book (www.understandingai.org)
- Show HN: Tune LLaMa3.1 on Google Cloud TPUs (github.com)
- Implementing LLaMA3 in 100 Lines of Pure Jax (saurabhalone.com)
- Longwriter – Increase llama3.1 output to 10k words (github.com)
- Hermes 3: The First Fine-Tuned Llama 3.1 405B Model (lambdalabs.com)
- Llama 3.2 released: Multimodal, 1B to 90B sizes (www.llama.com)
- Meta's Llama 3.1 can recall 42 percent of the first Harry Potter book (www.understandingai.org)
- Llama-3.2-3B-Instruct-uncensored (huggingface.co)
- Llama can now see and run on your device – welcome Llama 3.2 (huggingface.co)
- Llama 3.2 (huggingface.co)
- Nvidia presents Llama-Mesh: Generating 3D Mesh with Llama 3.1 8B (research.nvidia.com)
- Reflection-Llama-3.1-70B is Llama-3 with LoRA (old.reddit.com)