ALoDLM: Adaptively Looped Diffusion Language Models Paper • 2610.04198 • Published 8 days ago • 77
Rethinking Latent Visual Reasoning: Grounding Latent Reasoning in Visual Evidence Paper • 2609.34563 • Published 13 days ago • 219
DeFormer: Integrating Transformers with Deformable Models for 3D Shape Abstraction from a Single Image Paper • 2309.12594 • Published Sep 22, 2023
Optimal Transport-Guided Source-Free Adaptation for Face Anti-Spoofing Paper • 2503.22984 • Published Mar 29, 2025
The Hidden Life of Tokens: Reducing Hallucination of Large Vision-Language Models via Visual Information Steering Paper • 2502.03628 • Published Feb 5, 2025 • 12
MLLM-as-a-Judge for Image Safety without Human Labeling Paper • 2501.00192 • Published Dec 31, 2024 • 31
Decoupling Vision and Language: Codebook Anchored Visual Adaptation Paper • 2602.19449 • Published Feb 23
T3D: Few-Step Diffusion Language Models via Trajectory Self-Distillation with Direct Discriminative Optimization Paper • 2602.12262 • Published Feb 12 • 9
Running on CPU Upgrade Featured 3.32k The Smol Training Playbook 📚 3.32k The secrets to building world-class LLMs
EPO: Entropy-regularized Policy Optimization for LLM Agents Reinforcement Learning Paper • 2509.22576 • Published Sep 26, 2025 • 137
Token-Efficient Long Video Understanding for Multimodal LLMs Paper • 2503.04130 • Published Mar 6, 2025 • 97
The Hidden Life of Tokens: Reducing Hallucination of Large Vision-Language Models via Visual Information Steering Paper • 2502.03628 • Published Feb 5, 2025 • 12