☆ Yσɠƚԋσʂ ☆ to technology@hexbear.netEnglish · 1 year agoDeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learningarxiv.orgexternal-linkmessage-square0linkfedilinkarrow-up113arrow-down10cross-posted to: technologymachinelearning@lemmy.ml
arrow-up113arrow-down1external-linkDeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learningarxiv.org☆ Yσɠƚԋσʂ ☆ to technology@hexbear.netEnglish · 1 year agomessage-square0linkfedilinkcross-posted to: technologymachinelearning@lemmy.ml