CARREGANDO O RADAR…
OneLA: Scaling Linear-Attention Decoding to Large Beams in Generative Recommendation | Radar arXiv · portela.dev