Skip to main content

> ML_LITERATURE_ATLAS_v1.0

Research Literature Atlas

253 qualified literature records from foundational statistical learning to frontier reasoning LLMs: verified DOIs, arXiv IDs, and original bilingual syntheses.

Showing 18 of 253 Qualified Records (Page 13 of 15)Verified Academic Citations
Literature Record
2023 · Conferenceseminal-architecture

RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

The defining Vision-Language-Action (VLA) paper demonstrating emergent semantic reasoning and zero-shot physical affordance understanding in real robots.

roboticsmultimodal
Literature Record
2001 · Evolutionaryfoundational

Completely Derandomized Self-Adaptation in Evolution Strategies (CMA-ES)

The landmark evolutionary computation paper establishing CMA-ES as the gold standard algorithm for difficult, rugged, non-linear black-box optimization.

black box optimizationhyperparameter tuning
Literature Record
2017 · IEEEseminal-architecture

Xception: Deep Learning with Depthwise Separable Convolutions

Landmark CVPR paper by François Chollet creating Xception, establishing depthwise separable convolutions as the standard building block for efficient mobile vision.

image classification
Literature Record
2017 · arXivseminal-architecture

MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications

The landmark Google paper introducing MobileNetV1, enabling real-time computer vision inference on smartphones, drones, and edge embedded devices.

image classification
Literature Record
2018 · IEEEseminal-architecture

MobileNetV2: Inverted Residuals and Linear Bottlenecks

Monumental CVPR paper establishing MobileNetV2, the ubiquitous industry standard for mobile and on-device computer vision backbones.

image classificationobject detection
Literature Record
2020 · Internationalalgorithm

A Simple Framework for Contrastive Learning of Visual Representations (SimCLR)

The landmark Google ICML paper establishing SimCLR, transforming self-supervised visual representation learning through contrastive InfoNCE optimization.

feature extractionimage classification
Literature Record
2020 · IEEEalgorithm

Momentum Contrast for Unsupervised Visual Representation Learning (MoCo)

Landmark Facebook CVPR paper introducing MoCo, enabling large negative sample dictionaries without requiring thousands of GPU memory allocations.

feature extractionimage classification
Literature Record
2021 · IEEEalgorithm

Emerging Properties in Self-Supervised Vision Transformers (DINO)

Landmark ICCV paper creating DINO, demonstrating that self-supervised ViT attention maps discover semantic object boundaries without any human annotation.

feature extractionimage segmentation
Literature Record
2023 · arXivseminal-architecture

DINOv2: Learning Robust Visual Features without Supervision

Monumental Meta paper establishing DINOv2, the foundational vision feature backbone for monocular depth estimation, dense segmentation, and image retrieval.

feature extractiondepth estimationimage segmentation
Literature Record
2022 · IEEEalgorithm

Masked Autoencoders Are Scalable Vision Learners (MAE)

Landmark CVPR paper creating Masked Autoencoders (MAE), establishing the BERT-style masked autoencoding paradigm as a dominant foundation for scalable computer vision.

representation learningimage classification
Literature Record
2021 · Natureseminal-architecture

Highly accurate protein structure prediction with AlphaFold (AlphaFold 2)

The historic Nature cover paper announcing AlphaFold 2, resolving the 50-year-old protein folding grand challenge in biology and earning the 2024 Nobel Prize in Chemistry.

structural biology
Literature Record
2021 · Internationalseminal-architecture

Zero-Shot Text-to-Image Generation (DALL-E)

The historic ICML paper creating DALL-E, proving that multimodal autoregressive modeling enables creative zero-shot text-to-image synthesis.

image generation
Literature Record
2022 · arXivseminal-architecture

Hierarchical Text-Conditional Image Generation with CLIP Latents (DALL-E 2 / unCLIP)

The landmark unCLIP / DALL-E 2 paper from OpenAI, establishing two-stage CLIP latent diffusion for photorealistic text-to-image generation and semantic image variations.

image generation
Literature Record
2022 · Advancesseminal-architecture

Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding (Imagen)

Landmark Google NeurIPS paper establishing Imagen and the DrawBench benchmark, proving the decisive role of large language models in text-to-image synthesis.

image generation
Literature Record
2024 · OpenAIseminal-architecture

Video generation models as world simulators (Sora)

The landmark OpenAI technical report presenting Sora, establishing spacetime patch diffusion transformers as scalable world simulators for continuous video generation.

video generationimage generation
Literature Record
2016 · arXivseminal-architecture

SqueezeNet: AlexNet-level accuracy with 50x fewer parameters and <0.5MB model size

Foundational edge ML paper establishing SqueezeNet, proving that architectural compression enables high-accuracy deep learning on low-power microcontrollers and FPGAs.

image classification
Literature Record
2018 · IEEEseminal-architecture

ShuffleNet: An Extremely Efficient Convolutional Neural Network for Mobile Devices

Landmark CVPR mobile vision paper establishing ShuffleNet, delivering state-of-the-art speed-accuracy trade-offs for embedded robotics and mobile phones.

image classification
Literature Record
2020 · IEEEseminal-architecture

Designing Network Design Spaces (RegNet)

Winner of CVPR 2020 Best Paper, introducing the RegNet design space methodology for scalable, hardware-friendly convolutional architectures.

image classification
Showing 217–234 of 253 Literature Records