EN Submit a tool
Tip

RLM Paper Reveals: Harness Allows Models to 'Cheat' Generalization

Published: Source: X: swyx (@swyx)

ShareXFacebookTelegramWhatsApp

very notable trajectory comparison writeup here buried in the RLM paper from @a1zhang and @lateinter…

Read the original (opens in a new tab)

News stream data aggregated by AI HOT

Related newsLatest in this category