Dear Authors,
Thank you for open-sourcing DriveMA. The action-centric pretraining, meta-action SFT, and turn-level reinforcement learning designs presented in the paper have been very inspiring for my research.
I have recently been conducting some experiments on SFT reproduction. May I ask what the overall PDMS scores and component metrics of DriveMA-2B and DriveMA-4B are on NAVSIM after SFT but before reinforcement learning? These results would provide a valuable reference for my reproduction experiments.
Thank you for your time and excellent work. I look forward to your reply.
Dear Authors,
Thank you for open-sourcing DriveMA. The action-centric pretraining, meta-action SFT, and turn-level reinforcement learning designs presented in the paper have been very inspiring for my research.
I have recently been conducting some experiments on SFT reproduction. May I ask what the overall PDMS scores and component metrics of DriveMA-2B and DriveMA-4B are on NAVSIM after SFT but before reinforcement learning? These results would provide a valuable reference for my reproduction experiments.
Thank you for your time and excellent work. I look forward to your reply.