Series
Simplevla Rl
The "Simplevla Rl" series has 11 parts — read them in order from part 1.
Whole-body VLA
SimpleVLA-RL (1): Overview & Key Ideas
Introducing SimpleVLA-RL — using simple binary 0/1 rewards with RL to improve VLA models by 430%, no reward engineering needed.

SimpleVLA-RL (2): Architecture & Algorithm
Deep-dive into OpenVLA-OFT backbone, GRPO optimizer, dynamic sampling, and the exploration mechanisms that let VLA models self-improve.

SimpleVLA-RL (3): Setup & Training
Step-by-step guide to setting up the environment, running SFT cold-start, and RL training for SimpleVLA-RL on LIBERO and RoboTwin.

SimpleVLA-RL (4): Results & Key Takeaways
Analyzing SimpleVLA-RL results: ablation studies, the pushcut phenomenon, real-world transfer, and 5 key lessons learned.
SimpleVLA-RL (5): Comparison with LeRobot
In-depth comparison of SimpleVLA-RL and LeRobot: RL approach, VLA models, sim vs real, data efficiency — two complementary frameworks.

SimpleVLA-RL (6): OpenArm — Training Roadmap
A detailed analysis of training the 7-DoF OpenArm robot for carton box grasping — comparing 3 paths: LeRobot native, SimpleVLA-RL style, and hybrid.
SimpleVLA-RL (7): Collecting Data for OpenArm
Step-by-step guide to setting up OpenArm, calibrating, teleoperating, and collecting 50 box-grasping episodes with LeRobot.

SimpleVLA-RL (8): Training & Deploying on OpenArm
Train SmolVLA, ACT, Pi0-FAST for OpenArm box grasping — from fine-tuning to real robot deployment and improvement with HIL-SERL.

SimpleVLA-RL (9): OpenArm Simulation & Data Collection
Set up OpenArm in Isaac Lab, collect demonstration data in simulation, and convert to SimpleVLA-RL training format.

SimpleVLA-RL (10): SFT & RL Training for OpenArm
Step-by-step guide to SFT fine-tuning and RL training with SimpleVLA-RL for OpenArm — from environment config to running GRPO.

SimpleVLA-RL (11): Sim-to-Real Transfer for OpenArm
Deploy a SimpleVLA-RL model from simulation to real OpenArm — camera setup, action mapping, and tips for reducing the sim-to-real gap.