Jul 26, 2026 Chain-of-Goals: latent chain-of-thought for long-horizon offline RL Jul 11, 2026 Qwen-VLA — one model for manipulation, navigation, and trajectory May 26, 2026 Anshu & Arunachalam — and where RL-ES quantum state estimation lives inside it May 24, 2026 Enhanced POET — and the architecture of open-ended intelligence