Find the people behind AI papers and projects.
/
Search public work, explore its connections, and check the sources. How coverage works
4,006 saved profiles
Matching saved profiles. Coverage and source-listed affiliations may be incomplete or historical.
Search results
1–80 of 451
- Adrià Puigdomènech BadiaAffiliation: Google DeepMindWork: Asynchronous Methods for Deep Reinforcement Learning
- Alex GravesAffiliation: InstaDeepWork: Playing Atari with Deep Reinforcement Learning
- Andrew G. BartoAffiliation: University of Massachusetts AmherstWork: Reinforcement Learning: An Introduction (second edition)
- Arthur GuezAffiliation: Google DeepMindWork: Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
- Ashwin GopinathAffiliation: Sentra (Dynamis Labs Inc.)Work: Reflexion: Language Agents with Verbal Reinforcement Learning
- Atsushi SaitoAffiliation: Not recordedWork: Curriculum Learning Based on Reward Sparseness for Deep Reinforcement Learning of Task Completion Dialogue Management
- Aviral KumarAffiliation: Carnegie Mellon University, School of Computer ScienceWork: Training Language Models to Self-Correct via Reinforcement Learning
- Barret ZophAffiliation: Google DeepMindWork: Neural Architecture Search with Reinforcement Learning
- Benjamin LipkinAffiliation: Stealth startup (name undisclosed)Work: Entropy-Preserving Reinforcement Learning
- Binhang YuanAffiliation: The Hong Kong University of Science and Technology (HKUST)Work: AReaL: A Large-Scale Asynchronous Reinforcement Learning System
- Daan WierstraAffiliation: Not recordedWork: Playing Atari with Deep Reinforcement Learning
- Daniel J. MankowitzAffiliation: EthosWork: Faster sorting algorithms discovered using deep reinforcement learning
- David SilverAffiliation: Ineffable IntelligenceWork: Human-level control through deep reinforcement learning
- Daya GuoAffiliation: Not recordedWork: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- Deli ChenAffiliation: DeepSeek AIWork: DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
- Edward BermanAffiliation: Harvard UniversityWork: Reflexion: Language Agents with Verbal Reinforcement Learning
- Felipe Petroski SuchAffiliation: Not recordedWork: Deep Neuroevolution: Genetic Algorithms Are a Competitive Alternative for Training Deep Neural Networks for Reinforcement Learning
- Gabriel SynnaeveAffiliation: MetaWork: Reinforcement Learning for Code Optimization
- Henrique Pondé de Oliveira PintoAffiliation: Not recordedWork: Dota 2 with Large Scale Deep Reinforcement Learning
- Henryk MichalewskiAffiliation: Google DeepMindWork: Olympiad-Level Formal Mathematical Reasoning with Reinforcement Learning
- Hila NogaAffiliation: Not recordedWork: Multi-turn Reinforcement Learning with Preference Human Feedback
- Igor BabuschkinAffiliation: River AIWork: Grandmaster level in StarCraft II using multi-agent reinforcement learning
- Jan LeikeAffiliation: AnthropicWork: Deep reinforcement learning from human preferences
- Jean-Baptiste LespiauAffiliation: Not recordedWork: OpenSpiel: A Framework for Reinforcement Learning in Games
- Jingchang ChenAffiliation: Not recordedWork: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- Junhyuk OhAffiliation: Ineffable IntelligenceWork: Discovering state-of-the-art reinforcement learning algorithms
- Kai HuAffiliation: Nex-AGIWork: MM-Doc-R1: Training Agents for Long Document Visual Question Answering through Multi-turn Reinforcement Learning
- Koray KavukcuogluAffiliation: Google DeepMindWork: Playing Atari with Deep Reinforcement Learning
- Leandro von WerraAffiliation: Hugging FaceWork: TRL: Transformers Reinforcement Learning
- Liang WenfengAffiliation: DeepSeekWork: DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
- Lovish MadaanAffiliation: RecursiveWork: The Art of Scaling Reinforcement Learning Compute for LLMs
- Manaal FaruquiAffiliation: MetaWork: AdvancedIF: Rubric-Based Benchmarking and Reinforcement Learning for Advancing LLM Instruction Following
- Manuel KroissAffiliation: Not recordedWork: Grandmaster level in StarCraft II using multi-agent reinforcement learning
- Marc LanctotAffiliation: Google DeepMindWork: Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
- Martin A. RiedmillerAffiliation: Not recordedWork: Playing Atari with Deep Reinforcement Learning
- Matthew LaiAffiliation: Google DeepMindWork: RoboBallet: Planning for Multi-Robot Reaching with Graph Neural Networks and Reinforcement Learning
- Michael (Misha) LaskinAffiliation: Reflection AIWork: CURL: Contrastive Unsupervised Representations for Reinforcement Learning
- Miljan MarticAffiliation: Not recordedWork: Deep reinforcement learning from human preferences
- Nabeel ShanAffiliation: Not recordedWork: Advanced LLM Adaptation: From Supervised Fine-Tuning to Reinforcement Learning for Dialogue Summarization
- Noah ShinnAffiliation: Instinct / Spear Street Technology, Inc.Work: Reflexion: Language Agents with Verbal Reinforcement Learning
- Nova DasSarmaAffiliation: AnthropicWork: Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
- Oriol VinyalsAffiliation: Discovery LoopWork: Grandmaster level in StarCraft II using multi-agent reinforcement learning
- Paul ChristianoAffiliation: Alignment Research Center (ARC)Work: Deep reinforcement learning from human preferences
- Peiyi WangAffiliation: Not recordedWork: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- Richard PowellAffiliation: Not recordedWork: Grandmaster level in StarCraft II using multi-agent reinforcement learning
- Richard S. SuttonAffiliation: University of AlbertaWork: Reinforcement Learning: An Introduction (second edition)
- Rishabh AgarwalAffiliation: Periodic LabsWork: Deep Reinforcement Learning at the Edge of the Statistical Precipice
- Runxin XuAffiliation: DeepSeek AIWork: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- Ruoyu ZhangAffiliation: DeepSeekWork: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- Scott JohnstonAffiliation: Not recordedWork: Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
- Tarun GogineniAffiliation: OpenAIWork: TorsionNet: A Reinforcement Learning Approach to Sequential Conformer Search
- Thomas HubertAffiliation: Google DeepMindWork: Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
- Timothy P. LillicrapAffiliation: Not recordedWork: Continuous control with deep reinforcement learning
- Tobias PohlenAffiliation: Not recordedWork: Grandmaster level in StarCraft II using multi-agent reinforcement learning
- Volodymyr MnihAffiliation: Not recordedWork: Playing Atari with Deep Reinforcement Learning
- Wenjun GaoAffiliation: Not recordedWork: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- Xinyi (Cindy) WangAffiliation: DatabricksWork: KARL: Knowledge Agents via Reinforcement Learning
- Yang YueAffiliation: Tsinghua UniversityWork: Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?
- Yiyuan LiuAffiliation: Not recordedWork: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- Yundi QianAffiliation: MetaWork: AdvancedIF: Rubric-Based Benchmarking and Reinforcement Learning for Advancing LLM Instruction Following
- Yury SulskyAffiliation: GoogleWork: Grandmaster level in StarCraft II using multi-agent reinforcement learning
- Yuxiang YouAffiliation: Not recordedWork: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- Zhibin GouAffiliation: DeepSeekWork: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- Abram L. FriesenAffiliation: Not recordedWork: Acme: A Research Framework for Distributed Reinforcement Learning
- Adrien EcoffetAffiliation: OpenAIWork: Reinforcement Learning Under Moral Uncertainty
- Afroz MohiuddinAffiliation: Not recordedWork: Model-Based Reinforcement Learning for Atari
- Aixin LiuAffiliation: Not recordedWork: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- Albin CassirerAffiliation: Not recordedWork: Acme: A Research Framework for Distributed Reinforcement Learning
- Alex GoldinAffiliation: Not recordedWork: Improving Multimodal Interactive Agents with Reinforcement Learning from Human Feedback
- Alex RayAffiliation: Pioneer Square LabsWork: Multi-Goal Reinforcement Learning: Challenging Robotics Environments and Request for Research
- Alexander NovikovAffiliation: Not recordedWork: Discovering faster matrix multiplication algorithms with reinforcement learning
- Alon AlbalakAffiliation: Lila SciencesWork: Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
- Amol MandhaneAffiliation: Not recordedWork: Olympiad-level formal mathematical reasoning with reinforcement learning
- Anca DraganAffiliation: Google DeepMindWork: Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback
- Andrea TupiniAffiliation: Microsoft ResearchWork: Multimodal Reinforcement Learning with Adaptive Verifier for AI Agents
- Andreas Kirkeby FidjelandAffiliation: Not recordedWork: Human-level control through deep reinforcement learning
- Andreas KöpfAffiliation: benetura GmbHWork: REASONING GYM: Reasoning Environments for Reinforcement Learning with Verifiable Rewards
- Andy JonesAffiliation: AnthropicWork: Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
- Angeliki LazaridouAffiliation: Google DeepMindWork: Social Influence as Intrinsic Motivation for Multi-Agent Deep Reinforcement Learning
- Angelos FilosAffiliation: Google DeepMindWork: In-context Reinforcement Learning with Algorithm Distillation