Hacker Newsnew | past | comments | ask | show | jobs | submit | natolambert's submissionslogin
1.Learning to solve hard problems in RL for LLMs by never giving up (mnoukhov.github.io)
117 points by natolambert 2 days ago | past | 9 comments
2.Show HN: Colloquium – a Markdown-native slide tool for academics (github.com/natolambert)
2 points by natolambert 6 months ago | past
3.The Atom Project (atomproject.ai)
13 points by natolambert on Aug 4, 2025 | past | 4 comments
4.How RLHF Works (interconnects.ai)
165 points by natolambert on June 21, 2023 | past | 32 comments
5.Different Development Paths of LLMs (interconnects.ai)
1 point by natolambert on June 15, 2023 | past
6.Evaluating and Uncovering Open LLMs (interconnects.ai)
1 point by natolambert on May 31, 2023 | past
7.Unfortunately, OpenAI and Google have moats (interconnects.ai)
2 points by natolambert on May 17, 2023 | past
8.Specifying Hallucinations of LLMs (interconnects.ai)
1 point by natolambert on May 3, 2023 | past
9.The last reliable path into AI research (natolambert.com)
1 point by natolambert on Jan 30, 2022 | past
10.Remote Robotic-Data Farms (robotic.substack.com)
1 point by natolambert on Aug 9, 2021 | past
11.Reward Is Not Enough (robotic.substack.com)
2 points by natolambert on June 21, 2021 | past
12.All machine learning becomes reinforcement learning (robotic.substack.com)
1 point by natolambert on June 14, 2021 | past
13.Debugging model-based reinforcement learning systems (natolambert.com)
2 points by natolambert on April 5, 2021 | past
14.Setting ourselves up for exploitation: RL in the wild (robotic.substack.com)
1 point by natolambert on March 19, 2021 | past
15.Decoupling AI from the latent variable of spoken languages (robotic.substack.com)
2 points by natolambert on Feb 12, 2021 | past
16.Covid didn’t give us personal robots, it gave us Woebot (robotic.substack.com)
1 point by natolambert on Feb 5, 2021 | past
17.Boston Dynamics: Studying Athletic Intelligence (robotic.substack.com)
1 point by natolambert on Jan 29, 2021 | past
18.Robotic Startups 2.0: Horizontal Modularity (democraticrobots.substack.com)
1 point by natolambert on Jan 22, 2021 | past
19.Social Networks and Degradation to the Public Square of Discourse (democraticrobots.substack.com)
2 points by natolambert on Jan 15, 2021 | past | 1 comment
20.Our idea of Free Will can bias AIs we create (democraticrobots.substack.com)
2 points by natolambert on Jan 8, 2021 | past
21.The Ubiquity and Future of Model-Based Reinforcement Learning (democraticrobots.substack.com)
2 points by natolambert on Jan 4, 2021 | past
22.Constructing Axes for (Legal) Reinforcement Learning Policy (democraticrobots.substack.com)
2 points by natolambert on Dec 2, 2020 | past
23.The Collingridge Dilemma and Current Policy on Robots (democraticrobots.substack.com)
2 points by natolambert on Dec 1, 2020 | past
24.EE's See the World in Models (democraticrobots.substack.com)
2 points by natolambert on Oct 2, 2020 | past
25.Automated: The levers tech companies pull to direct our lives (democraticrobots.substack.com)
1 point by natolambert on July 31, 2020 | past
26.Recommender systems are a game – a dangerous game (for us) (democraticrobots.substack.com)
1 point by natolambert on July 24, 2020 | past
27.View from online courses and digital degrees at UC Berkeley (democraticrobots.substack.com)
2 points by natolambert on July 10, 2020 | past
28.Democratizing Automation (democraticrobots.substack.com)
1 point by natolambert on July 3, 2020 | past
29.Drones, Swarms, and Storms of Drones (democraticrobots.substack.com)
1 point by natolambert on June 26, 2020 | past
30.“10 years of automation in 1 year” from Covid-19 (democraticrobots.substack.com)
1 point by natolambert on June 13, 2020 | past

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: