Foundations

Research Papers Explained

The latest AI papers broken down — attention, RLHF, diffusion, MoE and more

17,121
lessons
Skills in this topic
View full skill map →
Reading ML Papers
beginner
Navigate Intro/Method/Experiments/Conclusion sections efficiently
Paper Reproduction
intermediate
Re-implement a published architecture from the paper
Research Methods
advanced
Design ablation studies
All Reads (1,397) Articles (460)Blog Posts (114)Tutorials (20)Research Papers (795)News (8)
Reddit r/MachineLearning 📄 Research Papers Explained ⚡ AI Lesson 1w ago
AC comment and our reply disappeared on OpenReview [D]
Hi everyone, we noticed that the AC's comment, along with our reply, has disappeared, and we are wondering if anyone else has experienced the same thing. The co
Comparing embedding models with synthetic query probing [R]
Reddit r/MachineLearning 📄 Research Papers Explained 1w ago
Comparing embedding models with synthetic query probing [R]
Say you want to swap out your embedding m
MIT Technology Review 📄 Research Papers Explained 1w ago
These startups are chasing the next big thing in LLMs
MIT Technology Review’s What’s Next series looks across industries, trends, and technologies to give you a first look at the future. You can read the rest of th
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
EntropyMoE: Entropy-Aware Sparse Expert Routing for Tokenizer-Free LLMs
arXiv:2608.06398v1 Announce Type: new Abstract: Recent byte-level large language models (LLMs) have made tokenizer-free modeling increasingly competitive by gro
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
NxN E-valuation: Hypothesis Certification via a Conformal CRT Null
arXiv:2608.06621v1 Announce Type: new Abstract: We propose NxN E-valuation, a handy, e-value-based hypothesis-certification algorithm that lets a hypothesis be
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
MolBioKG: Grounding Out-of-Graph Molecules in Biomedical Knowledge Graphs via Multi-Resolution Structural Anchoring
arXiv:2608.06713v1 Announce Type: new Abstract: Biomedical knowledge graphs (KGs) accelerate drug discovery, but standard pipelines assume query molecules alrea
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
Capek 0.5: An Execution-Centric Vision-Language Model for Embodied Intelligence
arXiv:2608.06756v1 Announce Type: new Abstract: Vision-language models are increasingly serving as the reasoning core of embodied agents. Robot execution is inh
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
LiFTER: A Grounded Neuro-Symbolic Microscope for Continuous-Time Dynamic Graph Forecasting
arXiv:2608.06765v1 Announce Type: new Abstract: Continuous-time dynamic graph models predict future links by compressing past interactions into neural states. A
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
Surg-UniWorld: A Unified Surgical World Model with Multimodal Control Experts
arXiv:2608.06770v1 Announce Type: new Abstract: Controllable surgical world models can provide a generative foundation for surgical artificial intelligence and
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
CEDAR: Agent-Orchestrated Tree Search for Goal-Directed Optimization of Complex Systems
arXiv:2608.06871v1 Announce Type: new Abstract: Complex systems, core objects of study in artificial life, model diverse phenomena through nonlinear, feedback-d
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
SkillEval: Decomposing Agent Skill Quality into Interpretable Signals
arXiv:2608.06891v1 Announce Type: new Abstract: Agent skills provide reusable procedural knowledge that helps agents solve specialized tasks. As their use expan
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
From Points to Edges: Edge-Conditioned Spectral Operators for Physics-Sensitive PDE Learning
arXiv:2608.06894v1 Announce Type: new Abstract: Neural operators have become a central tool for solving partial differential equations (PDEs), with spectral ope
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
ReGraph: Learning to Generate Recipe Graphs from Food Images
arXiv:2608.06917v1 Announce Type: new Abstract: Recent Large Multimodal Models (LMMs) have achieved impressive performance in recipe generation from food images
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
TRIBE: Predicting Team Performance via Communication Behavior Ensembles
arXiv:2608.06926v1 Announce Type: new Abstract: Designing autonomous agents that effectively assist human teams hinges on understanding team dynamics, often wit
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
Science Edge Evaluation: SEE the Missing Step Toward Real Scientific Discovery
arXiv:2608.06931v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly involved in scientific discovery, yet it remains unclear whether t
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
Blind to the Pivotal Vote: Aggregate Independence Metrics Miss Where Verification Actually Helps
arXiv:2608.06940v1 Announce Type: new Abstract: LLM judge panels are a standard evaluation tool, but prior work reports highly correlated panel errors: nine jud
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
CAi Copilot: Reducing Operational Workload in Molecular Design through Intent-Driven Agentic Workflows
arXiv:2608.06961v1 Announce Type: new Abstract: Early-stage molecular design is an iterative process, not just a task of generating molecules. Researchers turn
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
Learning in Deep Networks under Dale's Constraint
arXiv:2608.06963v1 Announce Type: new Abstract: Biologically plausible learning models aim to explain how neural circuits can implement effective learning under
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
Finding Usable Weight Mechanisms with Tiled SVD
arXiv:2608.06969v1 Announce Type: new Abstract: The dominant approach to mechanistic interpretability trains proxy dictionaries such as sparse autoencoders and
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
FedLBW: A Loss-Based Weighting Strategy for Federated Learning on Non-IID Data in Wireless Networks
arXiv:2608.07007v1 Announce Type: new Abstract: Federated Learning (FL) enables collaborative machine learning (ML) across distributed clients while preserving
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
ZIPBrain: Can EEG Foundation Models Be Faster, Locally Deployable, but Accurate?
arXiv:2608.07033v1 Announce Type: new Abstract: This work investigates whether Electroencephalograph (EEG) foundation models (EFMs) can be made faster and local
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
Not All Problems Are Best Modeled as MILP: A DSL-Centric Framework for Flexible and Accurate Optimization Modeling
arXiv:2608.07040v1 Announce Type: new Abstract: Solving combinatorial optimization problems (COPs) requires not only efficient algorithms but also carefully cra
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
BONSAI: Evolvability-Guided Tree Search over Skills
arXiv:2608.07056v1 Announce Type: new Abstract: A skill is a naturallanguage document that steers a frozen agent whose weights cannot be updated so any capabili
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
PTQ4SNN: Membrane-Aware Post-Training Quantization for Spiking Neural Networks
arXiv:2608.07066v1 Announce Type: new Abstract: Spiking neural networks (SNNs) enable sparse and event-driven computation, but their low-bit deployment remains
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
DocMemo: Dynamic Evidence Discovery via Probabilistic Memory-Guided Retrieval for Multi-Modal Document Understanding
arXiv:2608.07067v1 Announce Type: new Abstract: Long-document understanding requires locating sparse and heterogeneous evidence across hundreds of pages, yet ex
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
Transformers Struggle to Use Their Emergent World Models: Revisiting the Tower of Hanoi, and the Illusion of Thinking
arXiv:2608.07077v1 Announce Type: new Abstract: The Tower of Hanoi is a simple planning puzzle that in prior work has proven challenging for large reasoning mod
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
MemWM: Memory-Augmented Text-Based World Model
arXiv:2608.07107v1 Announce Type: new Abstract: World models are increasingly used to support planning in agents by predicting how environment states evolve in
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
How Much, Then Where: Credit-Conserving Action-to-Token Allocation for Multi-Turn Agent Reinforcement Learning
arXiv:2608.07118v1 Announce Type: new Abstract: Credit assignment in multi-turn agent reinforcement learning operates at two levels: assigning trajectory-level
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
DiDPO: Diff-in-Diff Policy Optimization for Coding Agent Training
arXiv:2608.07147v1 Announce Type: new Abstract: Reinforcement learning with Verifiable Reward (RLVR) has emerged as a powerful paradigm for training coding agen
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
SetEasy: A Multi-Modal Classroom Engagement Assessment and Seating Optimization Framework
arXiv:2608.07188v1 Announce Type: new Abstract: SetEasy optimizes classroom engagement in fixed seating grids. It fuses multimodal sensing (wristband physiology
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
Authoring and Management of Transparent Research Integrity Assessments of Randomised Clinical Trial Publications Using LLM-assisted Tools and Provenance Knowledge Graphs
arXiv:2608.07202v1 Announce Type: new Abstract: Systematic reviews of Randomised Controlled Trials (RCTs) are routinely used as evidence for clinical care guide
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
Beyond the Black Box: Interpretable Models of Human Randomisation Failures
arXiv:2608.07220v1 Announce Type: new Abstract: Mixed strategy equilibrium predicts i.i.d play: past actions should not help predict future decisions. Human pla
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
From probability to causality in probabilistic logic programming
arXiv:2608.07230v1 Announce Type: new Abstract: Probabilistic logic programming is a formalism of statistical relational artificial intelligence that supports c
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
WNM-3D: A World Navigation Model with 3D Scene Conditioning for Closed-Loop VLN
arXiv:2608.07267v1 Announce Type: new Abstract: Recent vision-language navigation (VLN) systems increasingly adapt pretrained vision-language models (VLMs) into
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
Winning by Peeking: Unenforced Budgets and Test-Set Selection Inflate Short-Budget AutoML Comparisons
arXiv:2608.07303v1 Announce Type: new Abstract: Comparisons between AutoML systems at short time budgets -- tens of seconds rather than hours -- are common in t
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
QFCQT: A Chaotically Gated Quantformer Framework for Volatile Time-Series Forecasting
arXiv:2608.07363v1 Announce Type: new Abstract: Forecasting non-stationary time series remains difficult due to long-range dependencies, local volatility bursts
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
Curriculum as Code: An AI-Assisted Architecture for Instructional Design in STEM Education
arXiv:2608.07364v1 Announce Type: new Abstract: Contribution: This paper presents a six-phase AI-assisted instructional design architecture based on the Curricu
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
FinRank: An Evidence-Grounded Benchmark for Financial Question Answering and Retrieval over SEC Filings
arXiv:2608.07400v1 Announce Type: new Abstract: Financial question answering is typically evaluated by answer correctness, yet in SEC filings a plausible and ev
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks
arXiv:2608.07411v1 Announce Type: new Abstract: In the context of geodata, existing Large Language Models have often been studied in a homogeneous setting, whic
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
CoBa: Cost-Effective Test-Time Scaling via Compute-Balanced Routing
arXiv:2608.07424v1 Announce Type: new Abstract: Test-time scaling is often implemented by spending more compute along one axis: sampling more solutions, extendi
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
TEPA: Revoking Stale Memories for Conflict-Robust Language Agents
arXiv:2608.07429v1 Announce Type: new Abstract: Long-term memory enables language agents to reuse past facts, preferences, and task experience. Persistence also
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
Post-Grokking Collapse at the Representation-Readout Interface in Muon-Trained Transformers
arXiv:2608.07436v1 Announce Type: new Abstract: Under the standard split, Muon gets hidden matrices and AdamW embeddings/output head. Muon groks modular additio
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
PsychoAgent: An Affect-Sensitive Cognitive Architecture for Conflict-Aware Memory in LLM Agents
arXiv:2608.07438v1 Announce Type: new Abstract: Human-like cognition does not select past experience by topical similarity alone: affective significance and unr
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
Blast Radius
arXiv:2608.07440v1 Announce Type: new Abstract: Agentic coding faces growing problems of affordability and wasted tokens. We introduce Blast Radius, a predictiv
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
SkillProx: Self-Evolving Agent Skills via Proximal Textual Gradient Descent
arXiv:2608.07449v1 Announce Type: new Abstract: LLM agents increasingly adapt to recurring tasks by accumulating procedural knowledge in skills. These skills ar
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
Interaction Creates Dynamical AI Behavior Absent in Isolation
arXiv:2608.07457v1 Announce Type: new Abstract: What will happen when AI agents interact in daily life, e.g. when one AI starts bossing another around? We find
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
Multimodal Drivers' Emotion Recognition and Safety-Oriented Intervention for Intelligent Transportation Systems
arXiv:2608.06378v1 Announce Type: cross Abstract: Driver emotions can affect risk perception, decision-making, and vehicle control under complex road conditions
ArXiv cs.AI 📄 Research Papers Explained 📄 Paper 1w ago
Mobile Interaction for Assessing Fatigue, Sleep, and Activity in Neurodegenerative and Chronic Diseases
arXiv:2608.06380v1 Announce Type: cross Abstract: Fatigue, sleep, or disturbances in daily activities are common symptoms among patients with neurodegenerative