[NeurIPS 2025] AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning
-
Updated
May 29, 2026 - Python
[NeurIPS 2025] AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning
Official code for the paper, "Stop Summation: Min-Form Credit Assignment Is All Process Reward Model Needs for Reasoning"
A Foundation Model for Crystal Structure Generation and Prediction
PersonaMem-v2: Towards Personalized Intelligence via Learning Implicit User Personas and Agentic Memory
ARIA (Autonomous Research Intelligence Agent): a closed-loop research agent that retrieves methodology from arXiv, implements it faithfully, and validates results through a non-compensable statistical gate — or abstains.
Add a description, image, and links to the reinforcement-finetuning topic page so that developers can more easily learn about it.
To associate your repository with the reinforcement-finetuning topic, visit your repo's landing page and select "manage topics."