Hackernews posts about Llama 3 405B
- We fine-tuned Llama 405B on AMD GPUs (publish.obsidian.md)
- Llama 405B 506 tokens/second on an H200 (developer.nvidia.com)
- The Future of AI: Synthetic Data Gen with Llama 3.1 405B and Raft (techcommunity.microsoft.com)
- Create a free Llama 3.1 405B-powered chatbot on a GitHub repo in <1 min (blog.stephenturner.us)