CARREGANDO O RADAR…
GRPO-QPS: Target-Preserving Reinforcement Learning for Quantum Posterior Sampling | Radar arXiv · portela.dev