Find the people behind AI papers and projects.
/
Search public work, explore its connections, and check the sources. How coverage works
4,006 saved profiles
People connected to this work in our saved records. This may not be the complete contributor list.
Search results
1–15 of 15
- Chenggang ZhaoAffiliation: DeepSeek AIWork: DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
- Deli ChenAffiliation: DeepSeek AIWork: DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
- Jianzhong GuoAffiliation: Not recordedWork: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- Liang WenfengAffiliation: DeepSeekWork: DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
- Qiancheng WangAffiliation: DeepSeek-AIWork: DeepSeek-R1 incentivizes reasoning in language models through reinforcement learning
- Ruiqi GeAffiliation: DeepSeek-AIWork: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- Runxin XuAffiliation: DeepSeek AIWork: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- Ruyi ChenAffiliation: Not recordedWork: DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
- Tian PeiAffiliation: Not recordedWork: DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
- Xinyuan LiAffiliation: Not recordedWork: DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
- Yichao ZhangAffiliation: Not recordedWork: DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
- Yuan OuAffiliation: Not recordedWork: DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
- Yue GongAffiliation: Not recordedWork: DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
- Yujia HeAffiliation: Not recordedWork: DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
- Zhihong ShaoAffiliation: DeepSeekWork: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning