Hacker News new | ask | show | jobs
user: Danau5tin
created: 2020-06-18
karma: 107

www.danaustin.ai

submissions:

0 points | 0 comments
0 points | 0 comments
0 points | 0 comments
0 points | 0 comments
0 points | 0 comments
0 points | 0 comments
0 points | 0 comments
0 points | 0 comments
0 points | 0 comments
0 points | 0 comments
0 points | 0 comments
0 points | 0 comments
0 points | 0 comments
Show HN: I RL-trained an agent that trains models with RL (for ~$1.3k)
107 points | 49 comments
0 points | 0 comments
Scaling Coding-Agent RL to 32x H100s. 160% Improvement on Stanford's TBench
2 points | 1 comments
Show HN: Multi-Agent-Coder Is #12 on Stanford's TBench. Beats Claude Code
5 points | 1 comments
0 points | 0 comments
My weekend project accidentally beat Claude Code – #12 on Stanford's TBench
2 points | 2 comments
0 points | 0 comments
Show HN: Terminal-Bench-RL: Training long-horizon terminal agents with RL
125 points | 12 comments