logoalt Hacker News

JohnBerrymantoday at 4:34 PM0 repliesview on HN

Nope. Here's the closest I got https://arcturus-labs.com/blog/2025/03/31/supercharging-llm-... - 1.5 years ago! But I never really did anything with it. And I wasn't thinking about reinforcement learning anything.