✕ Clear all filters
530 articles
▶ Videos →

Articles

530 articles · Updated every 3 hours · View all reads

All Articles 189,754Blog Posts 171,753Tech Tutorials 50,885Research Papers 36,791News 23,178 ⚡ AI Lessons
Papers Explained 618: SFT Conflicts, RL Coexists
Medium · Deep Learning 📄 Research Papers Explained 1w ago
Papers Explained 618: SFT Conflicts, RL Coexists
The paper investigates the differences between Supervised Fine-Tuning (SFT) and Reinforcement Learning (RL) for enhancing multi-task… Continue reading on Medium
Papers Explained 618: SFT Conflicts, RL Coexists
Medium · NLP 📄 Research Papers Explained 1w ago
Papers Explained 618: SFT Conflicts, RL Coexists
The paper investigates the differences between Supervised Fine-Tuning (SFT) and Reinforcement Learning (RL) for enhancing multi-task… Continue reading on Medium
arXiv Is Not a Journal. It Should Stop Acting Like One.
Medium · Data Science 📄 Research Papers Explained 1w ago
arXiv Is Not a Journal. It Should Stop Acting Like One.
A preprint server has one lane: circulate research before peer review. arXiv is drifting out of it. Continue reading on Medium »
Mixture-of-Kittens: Inside Cursor’s Open-Source MoE Megakernel for NVL72s
Medium · LLM 📄 Research Papers Explained 1w ago
Mixture-of-Kittens: Inside Cursor’s Open-Source MoE Megakernel for NVL72s
How a five-person research team fused an entire distributed training bottleneck into a single deterministic kernel — and how you can… Continue reading on Medium
Reddit r/deeplearning 📄 Research Papers Explained 1w ago
A question about discrete representations of numerical data
I recently came across a topic which I found really interesting. [ https://arxiv.org/html/2507.00078v1](https://arxiv.org/html/2507.00078v1(The) ( () The Langua
llama: add Maple 20B-A1B ternary MoE architecture (CPU) by AlexGabbia · Pull Request #27000 · ggml-org/llama.cpp
Reddit r/LocalLLaMA 📄 Research Papers Explained 1w ago
llama: add Maple 20B-A1B ternary MoE architecture (CPU) by AlexGabbia · Pull Request #27000 · ggml-org/llama.cpp
<img src="https://external-preview.redd.it/DFTzyiS0CLWzQ0O1KOr8efed00gqhw-g9U7ZD4C1V78.png?width=640&crop=smart&auto=webp&s=6fd21fbeffa3530f3eab0e25
InfoQ AI/ML 📄 Research Papers Explained 2w ago
GitHub Copilot's Project HydraFusion Promises Frontier Level Performance Through Multi-Model Routing
GitHub's Project HydraFusion is a research preview for GitHub Copilot that enhances coding intelligence through runtime model orchestration. It dynamically asse
What Is RAG? A Practical Breakdown of Retrieval-Augmented Generation
Medium · NLP 📄 Research Papers Explained 3w ago
What Is RAG? A Practical Breakdown of Retrieval-Augmented Generation
In 2020, a team of Facebook AI researchers published a paper called “Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks.”… Continue reading on Med
What Is RAG? A Practical Breakdown of Retrieval-Augmented Generation
Medium · RAG 📄 Research Papers Explained 3w ago
What Is RAG? A Practical Breakdown of Retrieval-Augmented Generation
In 2020, a team of Facebook AI researchers published a paper called “Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks.”… Continue reading on Med
AI Can Prove Einstein Was Right. It Could Never Have Been Einstein.
Medium · Machine Learning 📄 Research Papers Explained 3w ago
AI Can Prove Einstein Was Right. It Could Never Have Been Einstein.
A Google DeepMind researcher just explained why, and the answer is more specific than you’d expect. Continue reading on Level Up Coding »
I Kept Adding AI Agents Until One Question Stopped Me
Medium · Python 📄 Research Papers Explained 3w ago
I Kept Adding AI Agents Until One Question Stopped Me
Building a research → write → review pipeline with LangGraph — and the mental model that saves you from over-engineering yours. Continue reading on Medium »
{INTRESTING PAPER BASED ON HBF}2607.10186] FlashAccel: Leveraging High-Bandwidth Flash (HBF) for High-Throughput LLM Inference
Reddit r/LocalLLaMA 📄 Research Papers Explained 4w ago
{INTRESTING PAPER BASED ON HBF}2607.10186] FlashAccel: Leveraging High-Bandwidth Flash (HBF) for High-Throughput LLM Inference
<img src="https://external-preview.redd.it/q3evP6JeDpAC2MdSQHWYxnCYTqbJkElIQsLFqVSdkss.png?width=640&crop=smart&auto=webp&s=de730fbf7ecace6df0036b21
Reddit r/deeplearning 📄 Research Papers Explained ⚡ AI Lesson 4w ago
Joining AI research
Hi, I want to join an ai research project. How can I find people to work with? I would like to publish a paper at the end. submitted by /u/ihateyou103 [link] [c
Reddit r/deeplearning 📄 Research Papers Explained 4w ago
[Request] arXiv endorsement for cs.AI - Published AI researcher (Graph Embeddings / NLP)
submitted by /u/GabrielCPond [link] [comments]
Medium · NLP 📄 Research Papers Explained 4w ago
The Math Behind Mamba: How State Space Models Remember the Past
How can history be projected into set of Orthogonal Polynomials Continue reading on Medium »
Medium · LLM 📄 Research Papers Explained 4w ago
The Math Behind Mamba: How State Space Models Remember the Past
How can history be projected into set of Orthogonal Polynomials Continue reading on Medium »
Reddit r/Entrepreneur 📄 Research Papers Explained 4w ago
Help me or Roast me on my Competitor Intelligence idea
I'm building a competitor intelligence tool (to remain nameless to avoid the whole self-promotion hassle). For those people that actively research and monitor t
Consciousness as a Three‑Level System: Quantum Core, Neural Architecture, and Metaconscious…
Medium · Machine Learning 📄 Research Papers Explained 1mo ago
Consciousness as a Three‑Level System: Quantum Core, Neural Architecture, and Metaconscious…
In consciousness research, two extremes usually dominate: either everything is explained by neurons, or it drifts into metaphysics. But… Continue reading on Med
Consciousness as a Three‑Level System: Quantum Core, Neural Architecture, and Metaconscious…
Medium · LLM 📄 Research Papers Explained 1mo ago
Consciousness as a Three‑Level System: Quantum Core, Neural Architecture, and Metaconscious…
In consciousness research, two extremes usually dominate: either everything is explained by neurons, or it drifts into metaphysics. But… Continue reading on Med
Reddit r/LocalLLaMA 📄 Research Papers Explained 1mo ago
N-gram vs Experts explained
Since Qwen's dropped the Qwen4Exp architecture bomb that focus on offloading parameters to n-gram instead of pure mixture of experts, I dug into this and learne