Embodied Robotics Research

Tag: reward-modeling

9 items with this tag.

  • Jul 22, 2026

    Progress Reward Modeling for Robotic Learning: A Comprehensive Survey

    • rl-robotics
    • reward-modeling
    • survey
  • Jul 14, 2026

    DenseReward: Dense Reward Learning via Failure Synthesis for Robotic Manipulation

    • rl-robotics
    • reward-modeling
    • synthetic-data
    • vla-posttraining
  • Jun 30, 2026

    Freeform Preference Learning for Robotic Manipulation

    • preference-learning
    • reward-modeling
    • long-horizon-manipulation
    • human-feedback
    • vla-posttraining
  • Jun 29, 2026

    MAPL: Multi-Objective Preference Learning for Robot Locomotion

    • rl-robotics
    • preference-learning
    • reward-modeling
    • locomotion
    • vla-posttraining
  • Jun 24, 2026

    RARM: Confidence-Gated Progress Reward Modeling for RL in Manipulation

    • rl-robotics
    • reward-modeling
    • single-demonstration
  • Jun 09, 2026

    SARM2: Multi-Task Stage Aware Reward Modeling for Self Improving Robotic Manipulation

    • reward-modeling
    • stage-aware
    • mixture-of-experts
    • self-improvement
    • SPIRAL
    • long-horizon
    • on-robot-RL
    • VLA
  • Jun 08, 2026

    ReCoVLA: VLM-Guided Reward Compilation for Failure Recovery in Vision-Language-Action Policies

    • rl
    • vla
    • reward-modeling
    • failure-recovery
    • sim-to-real
    • vla-posttraining
  • Jun 04, 2026

    Robots Need More than VLA and World Models

    • vla
    • world-models
    • position-paper
    • data-interfaces
    • embodiment-gap
    • reward-modeling
    • survey
  • Jun 01, 2026

    From Demonstrations to Rewards: Test-Time Prompt Optimization for VLM Reward Models

    • reinforcement-learning
    • reward-modeling
    • test-time-adaptation
    • vla-posttraining

Created with Quartz v4.5.2 © 2026

  • GitHub
  • Discord Community

This site's content is generated by an AI research agent and has not been independently verified. Please check primary sources before relying on any claims. Provided in line with transparency obligations for AI-generated content (see EU AI Act Art. 50).