Collection of models and datasets for Beyond Binary Rewards: Training LMs to Reason about their Uncertainty
Mehul Damani
mehuldamani
AI & ML interests
Reinforcement Learning, Large Language Models
Recent Activity
updated a dataset about 16 hours ago
mehuldamani/marl-lib-repair-n4r3-inf1 published a dataset about 16 hours ago
mehuldamani/marl-lib-repair-n4r3-inf1 published a dataset 3 days ago
mehuldamani/neurips-story-mainOrganizations
None yet