Skip to main content

> ML_LITERATURE_ATLAS_v1.0

Research Literature Atlas

253 qualified literature records from foundational statistical learning to frontier reasoning LLMs: verified DOIs, arXiv IDs, and original bilingual syntheses.

Showing 18 of 253 Qualified Records (Page 12 of 15)Verified Academic Citations
Literature Record
2022 · Annualbenchmark

TruthfulQA: Measuring How Models Mimic Human Falsehoods

The primary benchmark for evaluating hallucination and epistemic truthfulness in language models, proving that larger models are often less truthful unless explicitly aligned.

question answering
Literature Record
2023 · Advancesbenchmark

Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

The most influential empirical benchmark of the generative AI era (LMSYS Chatbot Arena), establishing human blind preference battles as the definitive benchmark for frontier LLMs.

text generationmodel evaluation
Literature Record
1992 · Machinefoundational

Q-learning

The seminal paper that introduced Q-learning, the foundational model-free off-policy reinforcement learning algorithm in computer science.

reinforcement learning
Literature Record
1992 · Machinefoundational

Simple Statistical Gradient-Following Algorithms for Connectionist Reinforcement Learning (REINFORCE)

Foundational paper creating policy gradient methods, forming the theoretical bedrock of modern deep RL (PPO, TRPO, DPO, RLHF).

reinforcement learning
Literature Record
2015 · Natureseminal-architecture

Human-level control through deep reinforcement learning (Nature DQN)

Historic Nature cover paper that inaugurated the field of Deep Reinforcement Learning, bridging deep convolutional networks and dynamic programming.

reinforcement learning
Literature Record
2016 · AAAIalgorithm

Deep Reinforcement Learning with Double Q-learning (Double DQN)

Landmark AAAI paper resolving value overestimation in deep Q-networks, improving stability and score performance across complex environments.

reinforcement learning
Literature Record
2015 · Internationalalgorithm

Trust Region Policy Optimization (TRPO)

The landmark ICML paper establishing Trust Region Policy Optimization (TRPO), solving destructive step-size collapse in continuous control robotics and locomotion.

reinforcement learning
Literature Record
2018 · Internationalalgorithm

Addressing Function Approximation Error in Actor-Critic Methods (TD3)

Foundational continuous control paper resolving fundamental overestimation bias in actor-critic architectures, establishing TD3 as a standard benchmark in robotic simulation.

reinforcement learningcontinuous control
Literature Record
2017 · Scienceseminal-architecture

Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm (AlphaZero)

Groundbreaking Science paper demonstrating that a single general-purpose reinforcement learning algorithm can achieve superhuman mastery across multiple complex strategic domains tabula rasa.

game playing
Literature Record
2023 · Robotics:algorithm

Diffusion Policy: Visuomotor Policy Learning via Action Diffusion

Landmark RSS robotics paper establishing Diffusion Policy, the leading paradigm for training dextrous robot manipulation from human demonstrations.

continuous controlrobotics
Literature Record
2016 · Internationalalgorithm

Prioritized Experience Replay (PER)

Foundational ICLR paper introducing Prioritized Experience Replay, standard across DQN and off-policy actor-critic architectures.

reinforcement learning
Literature Record
2016 · Internationalseminal-architecture

Dueling Network Architectures for Deep Reinforcement Learning (Dueling DQN)

Winner of ICML 2016 Best Paper, introducing Dueling DQN to learn which states are valuable without having to learn the effect of each action for each state.

reinforcement learning
Literature Record
2016 · Internationalalgorithm

Asynchronous Methods for Deep Reinforcement Learning (A3C)

Landmark DeepMind ICML paper establishing A3C and A2C, enabling high-performance deep reinforcement learning directly on multi-core standard CPUs.

reinforcement learning
Literature Record
2016 · Internationalalgorithm

Continuous control with deep reinforcement learning (DDPG)

The landmark ICLR paper establishing DDPG, enabling deep reinforcement learning to solve 20+ continuous physical control tasks in physics simulators.

reinforcement learningcontinuous control
Literature Record
2017 · Advancesalgorithm

Hindsight Experience Replay (HER)

Landmark OpenAI paper solving sparse-reward robotic manipulation, enabling robots to master robotic arm pushing, sliding, and pick-and-place without reward shaping.

reinforcement learningrobotics
Literature Record
2018 · AAAIseminal-architecture

Rainbow: Combining Improvements in Deep Reinforcement Learning

Influential DeepMind AAAI paper establishing Rainbow, providing comprehensive ablation studies on what drives sample efficiency and score frontiers in discrete deep RL.

reinforcement learning
Literature Record
2022 · Natureseminal-architecture

Outracing champion Gran Turismo drivers with deep reinforcement learning (GT Sophy)

Historic Nature cover paper presenting Gran Turismo Sophy, solving real-time continuous vehicle dynamics, tire friction physics, and high-speed tactical etiquette.

game playingcontinuous control
Literature Record
2022 · Robotics:seminal-architecture

RT-1: Robotics Transformer for Real-World Control at Scale

Landmark robotics foundation paper demonstrating that large-scale multitask Transformer models generalize robustly to new tasks, environments, and objects.

roboticscontinuous control
Showing 199–216 of 253 Literature Records