Fine-tuning Multi-modal LLMs with ART: Art-based Reinforcement Training Paper • 2606.11854 • Published Jun 10 • 3