Robotic Manipulation on OGBench cube-double online
100Success RateQC-FQL
Evaluation Results
| Method | Links | |
|---|---|---|
| QC-FQLPolicy classification=Q-chunking, Base method=FQL2025.07 | 100 | |
| RLPDTemporal Difference (TD) steps=1-step, Policy classification=From Scratch, Base method=RLPD2025.07 | 99 | |
| QCPolicy classification=Q-chunking, Base method=QC2025.07 | 98 | |
| RLPD-ACTemporal Difference (TD) steps=1-step, Policy classification=From Scratch, Base method=RLPD-AC2025.07 | 96 | |
| BFNTemporal Difference (TD) steps=1-step, Policy classification=Flow, Base method=BFN2025.07 | 79 | |
| QC-IFQLPolicy classification=Q-chunking, Base method=IFQL2025.07 | 78 | |
| FQL-nTemporal Difference (TD) steps=n-step, Policy classification=Flow, Base method=FQL2025.07 | 77 | |
| FQLTemporal Difference (TD) steps=1-step, Policy classification=Flow, Base method=FQL2025.07 | 76 | |
| SUPE-GTTemporal Difference (TD) steps=1-step, Policy classification=From Scratch, Base method=SUPE-GT2025.07 | 67 | |
| BFN-nTemporal Difference (TD) steps=n-step, Policy classification=Flow, Base method=BFN2025.07 | 65 | |
| IFQLTemporal Difference (TD) steps=1-step, Policy classification=Flow, Base method=IFQL2025.07 | 59 | |
| ReBRACTemporal Difference (TD) steps=1-step, Policy classification=Gaussian, Base method=ReBRAC2025.07 | 30 | |
| IFQL-nTemporal Difference (TD) steps=n-step, Policy classification=Flow, Base method=IFQL2025.07 | 29 | |
| IQLTemporal Difference (TD) steps=1-step, Policy classification=Gaussian, Base method=IQL2025.07 | 0 |