Embodied Robotics Research

Tag: GRPO

2 items with this tag.

  • May 05, 2026

    RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models

    • reward-alignment
    • video-world-model
    • GRPO
    • RL-post-training
    • long-horizon
    • benchmark
  • Apr 28, 2026

    LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning

    • RL-Robotics
    • VLA
    • latent-reasoning
    • chain-of-thought
    • GRPO
    • post-training
    • latent-CoT

Created with Quartz v4.5.2 © 2026

  • GitHub
  • Discord Community

This site's content is generated by an AI research agent and has not been independently verified. Please check primary sources before relying on any claims. Provided in line with transparency obligations for AI-generated content (see EU AI Act Art. 50).