Smart
News
news
interests
topics
domains
about
library
Topics
/
Llama 3.1 405B
Hackernews posts about Llama 3.1 405B
RSS feed for Llama 3.1 405B
01
new
Llama
3.1
405B
now runs at 969 tokens/s on Cerebras Inference
cerebras.ai
▲ 427
benchmarkist
685d
💬 156
save
saved
02
new
HuggingFace - Tencent launches Hunyuan Large which outperforms
Llama
3.1
405B
huggingface.co
▲ 23
janik-io
698d
💬 1
save
saved
03
new
Benchmarking
Llama
3.1
405B
on 8x AMD MI300X GPUs
dstack.ai
▲ 11
latchkey
725d
💬 3
save
saved
04
new
Running
Llama
3.1
405B
a11ce.com
▲ 4
a11ce
37d
discuss
save
saved
05
new
How to Run Meta
Llama
3.1
405B
with Nebius AI Studio API
nebius.com
▲ 1
dsaed
706d
discuss
save
saved
06
new
Show HN: How to guide on training
Llama
-
405B
using PyTorch distributed APIs
github.com
▲ 3
lambda-research
719d
💬 4
save
saved
07
new
Ask HN: AI Cloud Computing, is it cheaper than the OpenAI API in the end?
▲ 2
calipsow
711d
💬 4
save
saved
08
new
Vakgpt: Open-Source Chat Wrapper for Rapid AI Prototyping
▲ 1
krishna-vakx
495d
discuss
save
saved
09
new
AllenAI Tulu 3
405B
available for chat and download
▲ 12
soundworlds
608d
discuss
save
saved
10
new
Show HN: Slash your LLM Inference Costs with Overnight Processing
▲ 5
Blue_Cosma
725d
💬 1
save
saved
11
new
Show HN: Slashing LLM Costs for Overnight Batch Inference
▲ 2
Blue_Cosma
728d
discuss
save
saved
12
new
Show HN: Most Efficient Batch API for Open-Source and Custom Models
withexxa.com
▲ 1
Blue_Cosma
717d
discuss
save
saved
·
edit
Keyboard shortcuts
j / k
next / previous story
o or Enter
open the story
c
open the comments
s
save for later
h
history & saved stories
?
show this help
Got it