metacognition Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs Paper • 2606.32032 • Published Jun 30 • 30
Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs Paper • 2606.32032 • Published Jun 30 • 30
Multidoc MDCure: A Scalable Pipeline for Multi-Document Instruction-Following Paper • 2410.23463 • Published Oct 30, 2024 • 3
MDCure: A Scalable Pipeline for Multi-Document Instruction-Following Paper • 2410.23463 • Published Oct 30, 2024 • 3
metacognition Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs Paper • 2606.32032 • Published Jun 30 • 30
Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs Paper • 2606.32032 • Published Jun 30 • 30
Multidoc MDCure: A Scalable Pipeline for Multi-Document Instruction-Following Paper • 2410.23463 • Published Oct 30, 2024 • 3
MDCure: A Scalable Pipeline for Multi-Document Instruction-Following Paper • 2410.23463 • Published Oct 30, 2024 • 3