Is One Layer Enough? A Single Transformer Layer Matches Full-Parameter RL Train

(arxiv.org)

32 points | by tcp_handshaker 2 hours ago ago

7 comments