Foundations
Research Papers Explained
The latest AI papers broken down — attention, RLHF, diffusion, MoE and more
Skills in this topic
3 skills — Sign in to track your progress
Reddit r/MachineLearning
📄 Research Papers Explained
⚡ AI Lesson
1w ago
AC comment and our reply disappeared on OpenReview [D]
Hi everyone, we noticed that the AC's comment, along with our reply, has disappeared, and we are wondering if anyone else has experienced the same thing. The co
![Comparing embedding models with synthetic query probing [R]](https://preview.redd.it/eauhd4hdyiih1.png?width=140&height=47&auto=webp&s=7594a52bcc580426082f61ebb75cecded686b9a9)
Reddit r/MachineLearning
📄 Research Papers Explained
1w ago
Comparing embedding models with synthetic query probing [R]
Say you want to swap out your embedding m
MIT Technology Review
📄 Research Papers Explained
1w ago
These startups are chasing the next big thing in LLMs
MIT Technology Review’s What’s Next series looks across industries, trends, and technologies to give you a first look at the future. You can read the rest of th
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
EntropyMoE: Entropy-Aware Sparse Expert Routing for Tokenizer-Free LLMs
arXiv:2608.06398v1 Announce Type: new Abstract: Recent byte-level large language models (LLMs) have made tokenizer-free modeling increasingly competitive by gro
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
NxN E-valuation: Hypothesis Certification via a Conformal CRT Null
arXiv:2608.06621v1 Announce Type: new Abstract: We propose NxN E-valuation, a handy, e-value-based hypothesis-certification algorithm that lets a hypothesis be
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
MolBioKG: Grounding Out-of-Graph Molecules in Biomedical Knowledge Graphs via Multi-Resolution Structural Anchoring
arXiv:2608.06713v1 Announce Type: new Abstract: Biomedical knowledge graphs (KGs) accelerate drug discovery, but standard pipelines assume query molecules alrea
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
Capek 0.5: An Execution-Centric Vision-Language Model for Embodied Intelligence
arXiv:2608.06756v1 Announce Type: new Abstract: Vision-language models are increasingly serving as the reasoning core of embodied agents. Robot execution is inh
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
LiFTER: A Grounded Neuro-Symbolic Microscope for Continuous-Time Dynamic Graph Forecasting
arXiv:2608.06765v1 Announce Type: new Abstract: Continuous-time dynamic graph models predict future links by compressing past interactions into neural states. A
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
Surg-UniWorld: A Unified Surgical World Model with Multimodal Control Experts
arXiv:2608.06770v1 Announce Type: new Abstract: Controllable surgical world models can provide a generative foundation for surgical artificial intelligence and
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
CEDAR: Agent-Orchestrated Tree Search for Goal-Directed Optimization of Complex Systems
arXiv:2608.06871v1 Announce Type: new Abstract: Complex systems, core objects of study in artificial life, model diverse phenomena through nonlinear, feedback-d
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
SkillEval: Decomposing Agent Skill Quality into Interpretable Signals
arXiv:2608.06891v1 Announce Type: new Abstract: Agent skills provide reusable procedural knowledge that helps agents solve specialized tasks. As their use expan
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
From Points to Edges: Edge-Conditioned Spectral Operators for Physics-Sensitive PDE Learning
arXiv:2608.06894v1 Announce Type: new Abstract: Neural operators have become a central tool for solving partial differential equations (PDEs), with spectral ope
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
ReGraph: Learning to Generate Recipe Graphs from Food Images
arXiv:2608.06917v1 Announce Type: new Abstract: Recent Large Multimodal Models (LMMs) have achieved impressive performance in recipe generation from food images
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
TRIBE: Predicting Team Performance via Communication Behavior Ensembles
arXiv:2608.06926v1 Announce Type: new Abstract: Designing autonomous agents that effectively assist human teams hinges on understanding team dynamics, often wit
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
Science Edge Evaluation: SEE the Missing Step Toward Real Scientific Discovery
arXiv:2608.06931v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly involved in scientific discovery, yet it remains unclear whether t
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
Blind to the Pivotal Vote: Aggregate Independence Metrics Miss Where Verification Actually Helps
arXiv:2608.06940v1 Announce Type: new Abstract: LLM judge panels are a standard evaluation tool, but prior work reports highly correlated panel errors: nine jud
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
CAi Copilot: Reducing Operational Workload in Molecular Design through Intent-Driven Agentic Workflows
arXiv:2608.06961v1 Announce Type: new Abstract: Early-stage molecular design is an iterative process, not just a task of generating molecules. Researchers turn
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
Learning in Deep Networks under Dale's Constraint
arXiv:2608.06963v1 Announce Type: new Abstract: Biologically plausible learning models aim to explain how neural circuits can implement effective learning under
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
Finding Usable Weight Mechanisms with Tiled SVD
arXiv:2608.06969v1 Announce Type: new Abstract: The dominant approach to mechanistic interpretability trains proxy dictionaries such as sparse autoencoders and
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
FedLBW: A Loss-Based Weighting Strategy for Federated Learning on Non-IID Data in Wireless Networks
arXiv:2608.07007v1 Announce Type: new Abstract: Federated Learning (FL) enables collaborative machine learning (ML) across distributed clients while preserving
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
ZIPBrain: Can EEG Foundation Models Be Faster, Locally Deployable, but Accurate?
arXiv:2608.07033v1 Announce Type: new Abstract: This work investigates whether Electroencephalograph (EEG) foundation models (EFMs) can be made faster and local
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
Not All Problems Are Best Modeled as MILP: A DSL-Centric Framework for Flexible and Accurate Optimization Modeling
arXiv:2608.07040v1 Announce Type: new Abstract: Solving combinatorial optimization problems (COPs) requires not only efficient algorithms but also carefully cra
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
BONSAI: Evolvability-Guided Tree Search over Skills
arXiv:2608.07056v1 Announce Type: new Abstract: A skill is a naturallanguage document that steers a frozen agent whose weights cannot be updated so any capabili
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
PTQ4SNN: Membrane-Aware Post-Training Quantization for Spiking Neural Networks
arXiv:2608.07066v1 Announce Type: new Abstract: Spiking neural networks (SNNs) enable sparse and event-driven computation, but their low-bit deployment remains
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
DocMemo: Dynamic Evidence Discovery via Probabilistic Memory-Guided Retrieval for Multi-Modal Document Understanding
arXiv:2608.07067v1 Announce Type: new Abstract: Long-document understanding requires locating sparse and heterogeneous evidence across hundreds of pages, yet ex
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
Transformers Struggle to Use Their Emergent World Models: Revisiting the Tower of Hanoi, and the Illusion of Thinking
arXiv:2608.07077v1 Announce Type: new Abstract: The Tower of Hanoi is a simple planning puzzle that in prior work has proven challenging for large reasoning mod
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
MemWM: Memory-Augmented Text-Based World Model
arXiv:2608.07107v1 Announce Type: new Abstract: World models are increasingly used to support planning in agents by predicting how environment states evolve in
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
How Much, Then Where: Credit-Conserving Action-to-Token Allocation for Multi-Turn Agent Reinforcement Learning
arXiv:2608.07118v1 Announce Type: new Abstract: Credit assignment in multi-turn agent reinforcement learning operates at two levels: assigning trajectory-level
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
DiDPO: Diff-in-Diff Policy Optimization for Coding Agent Training
arXiv:2608.07147v1 Announce Type: new Abstract: Reinforcement learning with Verifiable Reward (RLVR) has emerged as a powerful paradigm for training coding agen
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
SetEasy: A Multi-Modal Classroom Engagement Assessment and Seating Optimization Framework
arXiv:2608.07188v1 Announce Type: new Abstract: SetEasy optimizes classroom engagement in fixed seating grids. It fuses multimodal sensing (wristband physiology
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
Authoring and Management of Transparent Research Integrity Assessments of Randomised Clinical Trial Publications Using LLM-assisted Tools and Provenance Knowledge Graphs
arXiv:2608.07202v1 Announce Type: new Abstract: Systematic reviews of Randomised Controlled Trials (RCTs) are routinely used as evidence for clinical care guide
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
Beyond the Black Box: Interpretable Models of Human Randomisation Failures
arXiv:2608.07220v1 Announce Type: new Abstract: Mixed strategy equilibrium predicts i.i.d play: past actions should not help predict future decisions. Human pla
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
From probability to causality in probabilistic logic programming
arXiv:2608.07230v1 Announce Type: new Abstract: Probabilistic logic programming is a formalism of statistical relational artificial intelligence that supports c
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
WNM-3D: A World Navigation Model with 3D Scene Conditioning for Closed-Loop VLN
arXiv:2608.07267v1 Announce Type: new Abstract: Recent vision-language navigation (VLN) systems increasingly adapt pretrained vision-language models (VLMs) into
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
Winning by Peeking: Unenforced Budgets and Test-Set Selection Inflate Short-Budget AutoML Comparisons
arXiv:2608.07303v1 Announce Type: new Abstract: Comparisons between AutoML systems at short time budgets -- tens of seconds rather than hours -- are common in t
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
QFCQT: A Chaotically Gated Quantformer Framework for Volatile Time-Series Forecasting
arXiv:2608.07363v1 Announce Type: new Abstract: Forecasting non-stationary time series remains difficult due to long-range dependencies, local volatility bursts
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
Curriculum as Code: An AI-Assisted Architecture for Instructional Design in STEM Education
arXiv:2608.07364v1 Announce Type: new Abstract: Contribution: This paper presents a six-phase AI-assisted instructional design architecture based on the Curricu
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
FinRank: An Evidence-Grounded Benchmark for Financial Question Answering and Retrieval over SEC Filings
arXiv:2608.07400v1 Announce Type: new Abstract: Financial question answering is typically evaluated by answer correctness, yet in SEC filings a plausible and ev
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks
arXiv:2608.07411v1 Announce Type: new Abstract: In the context of geodata, existing Large Language Models have often been studied in a homogeneous setting, whic
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
CoBa: Cost-Effective Test-Time Scaling via Compute-Balanced Routing
arXiv:2608.07424v1 Announce Type: new Abstract: Test-time scaling is often implemented by spending more compute along one axis: sampling more solutions, extendi
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
TEPA: Revoking Stale Memories for Conflict-Robust Language Agents
arXiv:2608.07429v1 Announce Type: new Abstract: Long-term memory enables language agents to reuse past facts, preferences, and task experience. Persistence also
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
Post-Grokking Collapse at the Representation-Readout Interface in Muon-Trained Transformers
arXiv:2608.07436v1 Announce Type: new Abstract: Under the standard split, Muon gets hidden matrices and AdamW embeddings/output head. Muon groks modular additio
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
PsychoAgent: An Affect-Sensitive Cognitive Architecture for Conflict-Aware Memory in LLM Agents
arXiv:2608.07438v1 Announce Type: new Abstract: Human-like cognition does not select past experience by topical similarity alone: affective significance and unr
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
Blast Radius
arXiv:2608.07440v1 Announce Type: new Abstract: Agentic coding faces growing problems of affordability and wasted tokens. We introduce Blast Radius, a predictiv
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
SkillProx: Self-Evolving Agent Skills via Proximal Textual Gradient Descent
arXiv:2608.07449v1 Announce Type: new Abstract: LLM agents increasingly adapt to recurring tasks by accumulating procedural knowledge in skills. These skills ar
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
Interaction Creates Dynamical AI Behavior Absent in Isolation
arXiv:2608.07457v1 Announce Type: new Abstract: What will happen when AI agents interact in daily life, e.g. when one AI starts bossing another around? We find
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
Multimodal Drivers' Emotion Recognition and Safety-Oriented Intervention for Intelligent Transportation Systems
arXiv:2608.06378v1 Announce Type: cross Abstract: Driver emotions can affect risk perception, decision-making, and vehicle control under complex road conditions
ArXiv cs.AI
📄 Research Papers Explained
📄 Paper
1w ago
Mobile Interaction for Assessing Fatigue, Sleep, and Activity in Neurodegenerative and Chronic Diseases
arXiv:2608.06380v1 Announce Type: cross Abstract: Fatigue, sleep, or disturbances in daily activities are common symptoms among patients with neurodegenerative
DeepCamp AI