Hackernews posts about HumanEval
- Maincoder-1B – an open 1B-parameter coding model with 76% HumanEval (huggingface.co)
- Show HN: I built the LLM Comparison Tool I wish existed (llm-stats.com)
- Show HN: European Swallow AI – Sonnet-quality coding at $2.60/M tokens (www.europeanswallowai.com)
- Show HN: AIBenchy – Independent AI Leaderboard (aibenchy.com)
- Show HN: Atlas: Independent Evals and Benchmarking for Generative AI Models (app.layerlens.ai)