<?xml version="1.0" encoding="UTF-8"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Yestino - The Signal · RL</title><link>https://yestino.com/zh-CN/entities/rl-999ac1</link><description>与 RL 相关的全部事件</description><language>zh-CN</language><atom:link href="https://yestino.com/zh-CN/entities/rl-999ac1/feed.xml" rel="self" type="application/rss+xml"/><item><title>Learning to solve hard problems in RL for LLMs by never giving up</title><link>https://yestino.com/zh-CN/events/learning-to-solve-hard-problems-in-rl-for-llms-by-never-givi-2693bb</link><guid isPermaLink="true">https://yestino.com/zh-CN/events/learning-to-solve-hard-problems-in-rl-for-llms-by-never-givi-2693bb</guid><pubDate>Tue, 15 Sep 2026 19:07:38 GMT</pubDate><description>This is a blog post for my recent paper on RL post-training of LLMs: introducing the Matthew Effect and proposing to solve it with Never Give Up.
来源：Hacker News</description></item></channel></rss>