secrett2633's blog

[논문리뷰] TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning

October 1, 2025

이 [arXiv]에 게시한 ‘TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning’ 논문에 대한 자세한 리뷰입니다.

October 1, 2025

이 [arXiv]에 게시한 ‘Thinking Sparks!: Emergent Attention Heads in Reasoning Models During Post Training’ 논문에 대한 자세한 리뷰입니다.

October 1, 2025

이 [arXiv]에 게시한 ‘The Dragon Hatchling: The Missing Link between the Transformer and Models of the Brain’ 논문에 대한 자세한 리뷰입니다.

October 1, 2025

Yao Shu이 [arXiv]에 게시한 ‘Test-Time Policy Adaptation for Enhanced Multi-Turn Interactions with LLMs’ 논문에 대한 자세한 리뷰입니다.

October 1, 2025

Anpei Chen이 [arXiv]에 게시한 ‘TTT3R: 3D Reconstruction as Test-Time Training’ 논문에 대한 자세한 리뷰입니다.