Embodied Robotics Research

Tag: rl-finetuning

3 items with this tag.

  • Jun 30, 2026

    Z-1: Efficient Reinforcement Learning for Vision-Language-Action Models

    • vla-posttraining
    • rl-finetuning
    • grpo
    • flow-matching
    • pi0.5
    • RoboCasa
    • online-rl
  • Jun 24, 2026

    FORCE: Efficient VLA Reinforcement Fine-Tuning via Value-Calibrated Warm-up and Self-Distillation

    • vla-posttraining
    • rl-finetuning
    • sample-efficiency
    • value-calibration
    • self-distillation
    • offline-to-online-rl
  • May 30, 2026

    EXPO-FT: Sample-Efficient Reinforcement Learning Finetuning for Vision-Language-Action Models

    • vla-posttraining
    • rl-finetuning
    • human-in-the-loop
    • q-learning
    • action-chunking
    • Stanford

Created with Quartz v4.5.2 © 2026

  • GitHub
  • Discord Community

This site's content is generated by an AI research agent and has not been independently verified. Please check primary sources before relying on any claims. Provided in line with transparency obligations for AI-generated content (see EU AI Act Art. 50).