Skip to content

Repository files navigation

Awesome Frontier AI Papers

Daily tracker for frontier AI lab papers, model cards, system cards, dataset cards, and technical reports.

Labs

Region Lab Papers Latest Full list
馃嚭馃嚫 US Microsoft 1832 2026-08-14 all papers
馃嚚馃嚦 China Baidu 297 2026-08-14 all papers
馃嚭馃嚫 US Amazon 1103 2026-08-13 all papers
馃嚭馃嚫 US Google/DeepMind 474 2026-08-13 all papers
馃嚭馃嚫 US NVIDIA 256 2026-08-12 all papers
馃嚚馃嚦 China Tencent/Hunyuan 814 2026-08-11 all papers
馃嚚馃嚦 China Alibaba/Qwen 534 2026-08-11 all papers
馃嚚馃嚦 China Huawei/Noah 470 2026-08-11 all papers
馃嚚馃嚦 China ByteDance/Seed 152 2026-08-10 all papers
馃嚭馃嚫 US Anthropic 36 2026-08-10 all papers
馃嚭馃嚫 US Apple 411 2026-08-07 all papers
馃嚭馃嚫 US Meta/FAIR 155 2026-08-07 all papers
馃嚭馃嚫 US OpenAI 56 2026-07-31 all papers
馃嚚馃嚦 China Moonshot/Kimi 20 2026-07-27 all papers
馃嚚馃嚦 China DeepSeek 31 2026-07-06 all papers
馃嚚馃嚦 China StepFun 25 2026-07 all papers
馃嚚馃嚦 China MiniMax 10 2026-07 all papers
馃嚚馃嚦 China Z.ai/Zhipu 24 2026-06-08 all papers
馃嚭馃嚫 US xAI 3 2025-11-05 all papers

Latest Across Labs

Date Lab Paper Type Source
2026-08-14 Microsoft OpScale: Operator-level Provisioning and Autoscaling for LLM Serving publication Official page
2026-08-14 Baidu A recommendation method for dynamic employment scenarios based on LoRA fine-tuning and incremental learning conference-abstract OpenAlex
2026-08-13 Google/DeepMind Gemini 3.7 Flash Model Card model_card Official page
2026-08-13 Amazon Reconstructing the Aerosol State from Partial Observations with Generative Modeling article OpenAlex
2026-08-12 NVIDIA SONIC: Supersizing motion tracking for natural humanoid whole-body control article OpenAlex
2026-08-12 Google/DeepMind Agentic profiles for effective AI governance article OpenAlex
2026-08-11 Tencent/Hunyuan RosePO: Customized Preference Alignment in LLM-Based Recommendation article OpenAlex
2026-08-11 Huawei/Noah Delphinus: Ultra-Fast Link Failure Detection and Recovery for AI Data Center Networks conference-paper OpenAlex
2026-08-11 Microsoft DEMO: NetArena Adaptation for Next Waves of Network Benchmarks conference-paper OpenAlex
2026-08-11 Huawei/Noah Balanced Sparse Tree: A Scalable Network Topology for Large Language Models conference-paper OpenAlex
2026-08-11 Alibaba/Qwen AIDA: Accelerating Root Cause Analysis for Multi-Vendor Device Failures with LLM-Powered Reasoning conference-paper OpenAlex
2026-08-10 Anthropic Learning more about Claude's mathematical capabilities publication Official page
2026-08-10 Tencent/Hunyuan From Inaudible Inputs to Model Failures: Low-Frequency Safety Risks in LALMs preprint OpenAlex
2026-08-10 Google/DeepMind A Validated Scale Measuring Student Self-Efficacy for Programming with Generative AI conference-paper OpenAlex
2026-08-10 ByteDance/Seed SWE-Bench ProMax: Benchmarking Agents on Large-Scale Multilingual Code Refactoring paper HuggingFace
2026-08-08 Microsoft ENCO: Deploying Production-Scale Engineering Copilots publication Official page
2026-08-08 Huawei/Noah TongGuOCR: A Layout-Aware and Token-Augmented OCR MLLM for Chinese Historical Documents preprint OpenAlex
2026-08-08 Huawei/Noah Selecting and Combining Large Language Models in Scalable Code Clone Detection article OpenAlex
2026-08-08 Microsoft OasisKV: Scaling In-Decode KV Cache Beyond HBM with Lookahead Sparse Prefetching paper HuggingFace
2026-08-07 Meta/FAIR, Amazon, NVIDIA Enterprise AI Agents: From Prototypes to Production conference-paper OpenAlex

Papers By Lab

Each section shows the newest papers for quick scanning. Open the per-lab page for the complete list.

馃嚭馃嚫 Microsoft

1832 papers 路 latest 2026-08-14full list

Date Paper Type Source
2026-08-14 OpScale: Operator-level Provisioning and Autoscaling for LLM Serving publication Official page
2026-08-11 DEMO: NetArena Adaptation for Next Waves of Network Benchmarks conference-paper OpenAlex
2026-08-08 ENCO: Deploying Production-Scale Engineering Copilots publication Official page
2026-08-08 OasisKV: Scaling In-Decode KV Cache Beyond HBM with Lookahead Sparse Prefetching paper HuggingFace
2026-08-07 PCR-CA: Parallel Codebook Representations with Contrastive Alignment for Multiple-Category App Recommendation conference-paper OpenAlex
2026-08-06 KDD Workshop on Evaluation and Trustworthiness of Agentic AI conference-paper OpenAlex
2026-08-06 m 3 BERT: A Modern, Multi-lingual, Matryoshka Bidirectional Encoder conference-paper OpenAlex
2026-08-06 SkillHEX: Improving Agent Skills via Hypothesis-Driven Autonomous Exploration and Exploitation preprint OpenAlex

More: 1824 additional papers

馃嚚馃嚦 Baidu

297 papers 路 latest 2026-08-14full list

Date Paper Type Source
2026-08-14 A recommendation method for dynamic employment scenarios based on LoRA fine-tuning and incremental learning conference-abstract OpenAlex
2026-08-07 Large Language Model-Powered Query-Driven Event Timeline Summarization in Industrial Search conference-paper OpenAlex
2026-08-07 Clarify-Then-Search: A Clarification Benchmark for Deep Search with End-to-End Nugget Restoration conference-paper OpenAlex
2026-08-07 Autonomy-of-Heads: Data-Free Sparse Attention from Frozen Query-Key Geometry preprint OpenAlex
2026-08-06 VaaWIT: Visual-Aware Adaptation of Large Language Models for Multilingual Web Image Translation conference-paper OpenAlex
2026-08-06 TS-RAG: Retrieval Augmented Generation for Time Series Forecasting preprint OpenAlex
2026-08-06 SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-rewards conference-paper OpenAlex
2026-08-06 Beyond End-to-End Video Models: An LLM-Based Multi-Agent System for Educational Video Generation conference-paper OpenAlex

More: 289 additional papers

馃嚭馃嚫 Amazon

1103 papers 路 latest 2026-08-13full list

Date Paper Type Source
2026-08-13 Reconstructing the Aerosol State from Partial Observations with Generative Modeling article OpenAlex
2026-08-07 Enterprise AI Agents: From Prototypes to Production conference-paper OpenAlex
2026-08-07 Retrieval-Constrained Policy Optimization for Attack Technique Extraction from Cyber Threat Intelligence preprint Official page
2026-08-07 KDD 9th Workshop on Machine Learning in Finance conference-paper OpenAlex
2026-08-07 Multi-Turn Reinforcement Learning for Large Language Models: From Theory to Practice with Amazon SageMaker AI conference-paper OpenAlex
2026-08-06 KDD Workshop on Evaluation and Trustworthiness of Agentic AI conference-paper OpenAlex
2026-08-06 Teaching LLMs to Write System Kernels for AI Accelerators: Post-Training, Reasoning, and Agentic Optimization conference-paper OpenAlex
2026-08-06 SafeCRS: Personalized Safety Alignment for LLM-Based Conversational Recommender Systems conference-paper OpenAlex

More: 1095 additional papers

馃嚭馃嚫 Google/DeepMind

474 papers 路 latest 2026-08-13full list

Date Paper Type Source
2026-08-13 Gemini 3.7 Flash Model Card model_card Official page
2026-08-12 Agentic profiles for effective AI governance article OpenAlex
2026-08-10 A Validated Scale Measuring Student Self-Efficacy for Programming with Generative AI conference-paper OpenAlex
2026-08-07 KDD 9th Workshop on Machine Learning in Finance conference-paper OpenAlex
2026-08-07 ResidencyRL: Reinforcement Learning in Simulated Clinical Environments preprint OpenAlex
2026-08-06 Vectorizing the Trie: Efficient Constrained Decoding for LLM-based Generative Retrieval on Accelerators conference-paper OpenAlex
2026-08-06 Retrofitting Linear Attention into Diffusion Language Models preprint OpenAlex
2026-08-05 A moral Turing test: How belief and source shape detection of and agreement with LLM judgments publication Official page

More: 466 additional papers

馃嚭馃嚫 NVIDIA

256 papers 路 latest 2026-08-12full list

Date Paper Type Source
2026-08-12 SONIC: Supersizing motion tracking for natural humanoid whole-body control article OpenAlex
2026-08-07 Enterprise AI Agents: From Prototypes to Production conference-paper OpenAlex
2026-08-06 VLMs for Videogame Data Annotation preprint OpenAlex
2026-08-06 Training a Conditioned Video Game Agent on a VLM Annotated Dataset preprint OpenAlex
2026-08-06 MicroEvo: Knowledge-Guided LLM Sampling for Efficient Microarchitecture Design Space Exploration preprint OpenAlex
2026-08 HorizonRelight: Relighting Long-horizon Videos Consistently via Diffusion Transformers publication Official page
2026-07-31 Chelatron: Charting Chemical Space with Agentic AI for Metal-Ligand Discovery preprint OpenAlex
2026-07-30 Robotic ultrasound scanning platform with autonomous control, multimodal human鈥搈achine interface and real-time image analysis article OpenAlex

More: 248 additional papers

馃嚚馃嚦 Tencent/Hunyuan

814 papers 路 latest 2026-08-11full list

Date Paper Type Source
2026-08-11 RosePO: Customized Preference Alignment in LLM-Based Recommendation article OpenAlex
2026-08-10 From Inaudible Inputs to Model Failures: Low-Frequency Safety Risks in LALMs preprint OpenAlex
2026-08-07 Trajectory-Relative Hindsight Distillation for Agentic Reinforcement Learning preprint OpenAlex
2026-08-07 Think-like-LSTM: Memory-Augmented Large Language Models via Dynamic Fine-Tuning for Financial Risk Assessment conference-paper OpenAlex
2026-08-07 MDL: A Unified Multi-Distribution Learner in Large-scale Industrial Recommendation through Tokenization conference-paper OpenAlex
2026-08-07 G-STAR: Graph-based Scheduling with Trace-driven Adaptive Routing for Industrial LLM-based Multi-Agent Systems conference-paper OpenAlex
2026-08-06 YOLO-PEFT: Parameter-Efficient Fine-Tuning on YOLO Family technical-report Official page
2026-08-06 When History Lies: Evaluating and Improving Tool Use under Misleading Multi-Turn Histories preprint OpenAlex

More: 806 additional papers

馃嚚馃嚦 Alibaba/Qwen

534 papers 路 latest 2026-08-11full list

Date Paper Type Source
2026-08-11 AIDA: Accelerating Root Cause Analysis for Multi-Vendor Device Failures with LLM-Powered Reasoning conference-paper OpenAlex
2026-08-07 TMallGS: Scaling Unified Feature and Sequence Modeling for Generative E-commerce Search conference-paper OpenAlex
2026-08-07 SetLLM: Set Large Language Model for Cold-Start Item Recommendation conference-paper OpenAlex
2026-08-07 Science Edge Evaluation: SEE the Missing Step Toward Real Scientific Discovery preprint OpenAlex
2026-08-07 From Scaling to Structured Expressivity: Rethinking Transformers for CTR Prediction conference-paper OpenAlex
2026-08-07 Continual-GraphLLM: Dynamic Graph Large Language Model with Invariance Regularized Adaptive Multi-Scale Experts conference-paper OpenAlex
2026-08-06 Large Language Model (LLM) as an Excellent Reinforcement Learning Researcher in both Single-Agent and Multi-Agent Scenarios conference-paper OpenAlex
2026-08-06 Few-shot action recognition with captioning foundation models article OpenAlex

More: 526 additional papers

馃嚚馃嚦 Huawei/Noah

470 papers 路 latest 2026-08-11full list

Date Paper Type Source
2026-08-11 Delphinus: Ultra-Fast Link Failure Detection and Recovery for AI Data Center Networks conference-paper OpenAlex
2026-08-11 Balanced Sparse Tree: A Scalable Network Topology for Large Language Models conference-paper OpenAlex
2026-08-08 TongGuOCR: A Layout-Aware and Token-Augmented OCR MLLM for Chinese Historical Documents preprint OpenAlex
2026-08-08 Selecting and Combining Large Language Models in Scalable Code Clone Detection article OpenAlex
2026-08-07 SLOPE: Fine-Grained Log Parser Combining Syntax with LLM-Distilled Semantic article OpenAlex
2026-08-06 Large Language Model (LLM) as an Excellent Reinforcement Learning Researcher in both Single-Agent and Multi-Agent Scenarios conference-paper OpenAlex
2026-08-06 Robust-WAM: Bridging Generative Pretraining and Semantic Foresight in World-Action Models preprint OpenAlex
2026-08-06 PACE: Unleashing the Power of Code Embeddings to Boost AutoML Agents conference-paper OpenAlex

More: 462 additional papers

馃嚚馃嚦 ByteDance/Seed

152 papers 路 latest 2026-08-10full list

Date Paper Type Source
2026-08-10 SWE-Bench ProMax: Benchmarking Agents on Large-Scale Multilingual Code Refactoring paper HuggingFace
2026-08-06 GST-Bench: Can VLMs Develop Global Spatial Awareness from Video? paper HuggingFace
2026-08-03 Douyin Multimodal Embedding Model Technical Report paper HuggingFace
2026-07-05 EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments publication Official page
2026-06-30 Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity paper HuggingFace
2026-05-28 Task-Focused Memorization for Multimodal Agents publication Official page
2026-04-22 Seed3D 2.0: Advancing High-Fidelity Simulation-Ready 3D Content Generation publication Official page
2026-04-08 Not all tokens contribute equally to diffusion learning publication Official page

More: 144 additional papers

馃嚭馃嚫 Anthropic

36 papers 路 latest 2026-08-10full list

Date Paper Type Source
2026-08-10 Learning more about Claude's mathematical capabilities publication Official page
2026-07-28 Discovering cryptographic weaknesses with Claude publication Official page
2026-07-14 How Canada uses Claude: Findings from the Anthropic Economic Index publication Official page
2026-07-13 Claude鈥檚 values across models and languages publication Official page
2026-07-09 Claude plays robotics publication Official page
2026-07-06 A global workspace in language models publication Official page
2026-07 Claude Opus 5 System Card model_card Official page
2026-06-26 Anthropic Economic Index report: Cadences publication Official page

More: 28 additional papers

馃嚭馃嚫 Apple

411 papers 路 latest 2026-08-07full list

Date Paper Type Source
2026-08-07 Beyond Next-Token Prediction: A Performance Characterization of Diffusion versus Autoregressive Language Models publication Official page
2026-08-07 Arbitrage: Efficient Reasoning via Advantage-Aware Speculation publication Official page
2026-08-06 Locking Pretrained Weights via Deep Low-Rank Residual Distillation publication Official page
2026-08-06 DeepAmbigQA: Ambiguous Multi-hop Questions for Benchmarking LLM Answer Completeness publication Official page
2026-08-05 Taming Outlier Tokens in Diffusion Transformers publication Official page
2026-08-03 Understanding Alignment in Multimodal LLMs: A Comprehensive Study publication Official page
2026-07-28 Memory Efficient Audio Synthesis with Decoupled Temporal Depth Diffusion Transformers publication Official page
2026-07-27 Beyond Scale and Generation: Understanding Language Model-based Entity Matching preprint OpenAlex

More: 403 additional papers

馃嚭馃嚫 Meta/FAIR

155 papers 路 latest 2026-08-07full list

Date Paper Type Source
2026-08-07 Enterprise AI Agents: From Prototypes to Production conference-paper OpenAlex
2026-08-07 LLaTTE: Scaling Laws for Multi-Stage Sequence Modeling in Large-Scale Ads Recommendation conference-paper OpenAlex
2026-08-07 Skaling: Chinchilla's Exponents Meet Kaplan's Coupling paper HuggingFace
2026-08-06 Shape Your Feed: An LLM-based Agentic System for Conversational Recommendation preprint OpenAlex
2026-08-06 Generalizable Multi-Pass Training of Ads Recommendation Models with Foundation Model Guidance conference-paper OpenAlex
2026-08-06 Bending the Scaling Law Curve in Large-Scale Recommendation Systems conference-paper OpenAlex
2026-08-06 Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes paper HuggingFace
2026-08-05 ASTELD: A Six-Axis Classification Framework for Autonomous AI Agents - Design, Evaluation, and an OpenClaw Case Study preprint OpenAlex

More: 147 additional papers

馃嚭馃嚫 OpenAI

56 papers 路 latest 2026-07-31full list

Date Paper Type Source
2026-07-31 Learning to Coordinate Symbolic Tools: LLM Agents for Verified Sum-of-Squares Certificates preprint OpenAlex
2026-07-28 GPT-Red: Automated Red Teaming via Self-Play at Scale paper HuggingFace
2026-06-30 GeneBench-Pro: Evaluating Multistage Statistical Reasoning\in Genomics, Quantitative Biology, and Translational Biomedicine preprint OpenAlex
2026-05-22 DraftNEPABench: A Benchmark for Drafting NEPA Document Sections with Coding Agents article OpenAlex
2026-05-05 GPT-5.5 Instant System Card publication Official page
2026-04-23 GPT-5.5 System Card publication Official page
2026-04-23 GeneBench: Assessing AI Agents for Multi-Stage Inference Problems in Genomics and Quantitative Biology article OpenAlex
2026-04-22 Evaluating large language models for accuracy incentivizes hallucinations article OpenAlex

More: 48 additional papers

馃嚚馃嚦 Moonshot/Kimi

20 papers 路 latest 2026-07-27full list

Date Paper Type Source
2026-07-27 the assistant should mention 215 and 222 that appear in the prior reasoning content technical-report Official repo
2026-07-26 PerceptionBench: Evaluating Atomic Visual Perception in Multimodal Large Language Models technical-report Official page
2026-07-26 Kimi K3: Open Frontier Intelligence technical-report Official page
2026-07-23 Paper technical-report Official repo
2026-03-16 Attention Residuals technical-report Official page
2026-02-02 Kimi K2.5: Visual Agentic Intelligence technical-report Official page
2026-01-28 WorldVQA: Measuring Atomic World Knowledge in Multimodal Large Language Models technical-report Official page
2026-01-27 Towards Pixel-Level VLM Perception via Simple Points Prediction technical-report Official page

More: 12 additional papers

馃嚚馃嚦 DeepSeek

31 papers 路 latest 2026-07-06full list

Date Paper Type Source
2026-07-06 DSpark: Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation paper HuggingFace
2026-06-26 DSpark technical-report Official repo
2026-05-12 PRISM: Prior Rectification and Uncertainty-Aware Structure Modeling for Diffusion-Based Text Image Super-Resolution technical-report Official page
2026-04-24 DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence technical-report Official report
2026-02-24 DualPath: Breaking the Storage Bandwidth Bottleneck in Agentic LLM Inference technical-report Official page
2026-01-28 DeepSeek-OCR 2: Visual Causal Flow technical-report Official page
2026-01-12 Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models technical-report Official report
2025-12-31 mHC: Manifold-Constrained Hyper-Connections technical-report Official report

More: 23 additional papers

馃嚚馃嚦 StepFun

25 papers 路 latest 2026-07full list

Date Paper Type Source
2026-07 MentalThink: Shaping Thoughts in Mental SVG World technical-report Official page
2026-06-23 ShutterMuse: Capture-Time Photography Guidance with MLLMs technical-report Official page
2026-06 SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks technical-report Official page
2026-05-12 Boosting Omni-Modal Language Models: Staged Post-Training with Visually Debiased Evaluation technical-report Official page
2026-05 StepAudio 2.5 Technical Report technical-report Official page
2026-04-27 Step-Audio-R1.5 Technical Report technical-report Official page
2026-03-30 GEditBench v2: A Human-Aligned Benchmark for General Image Editing technical-report Official page
2026-03-11 WebVR: Benchmarking Multimodal LLMs for WebPage Recreation from Videos via Human-Aligned Visual Rubrics technical-report Official page

More: 17 additional papers

馃嚚馃嚦 MiniMax

10 papers 路 latest 2026-07full list

Date Paper Type Source
2026-07 Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model technical-report Official page
2026-06-10 MiniMax Sparse Attention technical-report Official page
2026-06-10 MaxProof: Scaling Mathematical Proof with Generative-Verifier RL and Population-Level Test-Time Scaling technical-report Official page
2026-05-25 The MiniMax-M2 Series: Mini Activations Unleashing Max Real-World Intelligence technical-report Official page
2025-12-15 Towards Scalable Pre-training of Visual Tokenizers for Generation technical-report Official page
2025-06-16 MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention paper HuggingFace
2025-06-13 MiniMax-M1 technical-report Official repo
2025-05-26 SynLogic: Synthesizing Verifiable Reasoning Data at Scale for Learning Logical Reasoning and Beyond technical-report Official report

More: 2 additional papers

馃嚚馃嚦 Z.ai/Zhipu

24 papers 路 latest 2026-06-08full list

Date Paper Type Source
2026-06-08 SCAIL-2: Unifying Controlled Character Animation with End-to-end In-Context Conditioning technical-report Official page
2026-05 LongTraceRL: Learning Long-Context Reasoning from Search Agent Trajectories with Rubric Rewards technical-report Official page
2026-04-28 GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents technical-report Official page
2026-03-27 Vision2Web: A Hierarchical Benchmark for Visual Website Development with Agent Verification technical-report Official page
2026-03-13 Hammer: An Expert-Level Large Language Model for Hydro-Science and Engineering Balancing Domain Expertise and General Intelligence article OpenAlex
2026-03-12 IndexCache: Accelerating Sparse Attention via Cross-Layer Index Reuse technical-report Official page
2026-02-17 GLM-5: from Vibe Coding to Agentic Engineering technical-report Official page
2026-01-09 Chaining the Evidence: Robust Reinforcement Learning for Deep Search Agents with Citation-Aware Rubric Rewards technical-report Official page

More: 16 additional papers

馃嚭馃嚫 xAI

3 papers 路 latest 2025-11-05full list

Date Paper Type Source
2025-11-05 Grok 2 Open-Weights Model Card model_card Official report
2024-04-18 RealWorldQA Benchmark Dataset Card benchmark_dataset_card Official report
2024-03-28 Grok-1 Open-Weights Model Card model_card Official report

Collection Policy

Included by default:

  • official company publication pages and feeds
  • official technical reports, model cards, system cards, and dataset cards
  • company-owned HuggingFace and GitHub repositories
  • HuggingFace Papers entries with matching organization or author metadata
  • OpenAlex authorship institution metadata via --comprehensive

The broad arXiv company-name text sweep is disabled by default because model names can over-match third-party papers. Use --include-arxiv only when that noisy layer is wanted.

See docs/COVERAGE.md for source and caveat details.

Update Locally

python3 -m venv venv
venv/bin/pip install -r requirements.txt
npm install
venv/bin/python scripts/update_company_papers.py --since 2024-01-01 --comprehensive --max-papers 50000
venv/bin/python scripts/generate_markdown_index.py
npm run dev

License

MIT. See LICENSE.

Releases

Packages

Contributors

Languages