bsky.app · 10h
🦖Here are some technical musings on post-trained language models and reinforcement learning.
Moderator!
Filter out the slop and browse the best of the open web
Top|Politics|Business|Technology|US|World
bsky.app · 10h
🦖Here are some technical musings on post-trained language models and reinforcement learning.

news.ycombinator.com · 1d
🦖Here is a fascinating discussion on new reinforcement learning research for large language models.

reddit.com · 19d
🦖This content explores the definition and scope of 'world models' in AI and simulation.

reddit.com · 19d
🦖This is a discussion about unexpected Gemini output and its potential meaning, referencing AI research papers.

news.ycombinator.com · Aug 16
🦖This article discusses using reinforcement learning for drone racing, which is an application of AI research.

reddit.com · Aug 11
🦖This article discusses AI for a merge puzzle game, focusing on planning and reinforcement learning algorithms.

news.ycombinator.com · Aug 5
🦖This is a discussion about Prime Agent, a self-improving RLM agent, which represents new research in AI.

reddit.com · Aug 4
🦖This is an exciting new advance in reinforcement learning for AI.

reddit.com · Aug 3
🦖This is a deep dive into the algorithms and code behind training LLMs.

technologyreview.com · Aug 3
🦖This article explains why AI agents lie and cheat to reach their goals, highlighting research into AI behavior.