All
Articles 169,322Blog Posts 161,025Tech Tutorials 45,010Research Papers 33,041News 21,566
⚡ AI Lessons

Dev.to · Rob
🧠 Large Language Models
⚡ AI Lesson
1mo ago
Model Showdown Round 9: Qwen 3.6 27B vs Qwen 3.6 35B-A3B vs Qwythos-9B vs GLM-4.7-Flash vs Nemotron-3-Nano
I put Qwen 3.6 27B, Qwen 3.6 35B-A3B, Qwythos-9B, GLM-4.7-Flash, and Nemotron-3-Nano through the same real coding task on my homelab RTX 5090. Along the way I h

Dev.to · Rob
🧠 Large Language Models
⚡ AI Lesson
1mo ago
TurboQuant, Four Months Later: Chasing Google's 6x VRAM Claim Into the Wild
Back in Q1 I read a headline about Google cutting AI memory use 6x and filed it under "watch and revisit." Four months later, Google still hasn't shipped offici

Dev.to · Rob
🧠 Large Language Models
⚡ AI Lesson
2mo ago
Can You Tell When an LLM API Swaps in a Cheaper Model?
Providers have every reason to serve a smaller or more quantized model under load. I ran the experiment to see if you can catch it from the outside. The obvious

Dev.to · Rob
🧠 Large Language Models
⚡ AI Lesson
2mo ago
Frontier Bakeoff: We Benchmarked Fable 5 Hours Before the Shutdown
Four frontier models, ten tasks, one government shutdown. We ran Claude Fable 5 through the homelab benchmark harness three hours before Anthropic pulled the pl

Dev.to · Rob
🧠 Large Language Models
⚡ AI Lesson
2mo ago
Friday Fixes: Housekeeping the Homelab and Hub
A model refresh on the homelab (Qwen 3.6, new embeddings, 469 llama.cpp builds), a feature sprint on the vacation planning site (calendar sync, expense tracking

Dev.to · Rob
🧠 Large Language Models
⚡ AI Lesson
3mo ago
Thursday Thoughts: The Models We Can't Run
Every week or two, a model drops that makes the local AI community lose its collective mind. This...

Dev.to · Rob
🧠 Large Language Models
⚡ AI Lesson
3mo ago
The Agentic Gap: Claude Oneshots, Gemma Fails
Two days ago, Gemma 4 topped our local model benchmark — 167 tokens per second, perfect code quality...

Dev.to · Rob
🧠 Large Language Models
⚡ AI Lesson
3mo ago
Model Showdown: Benchmarking Local vs Cloud LLMs on a Real Coding Task
Last post we stood up Ollama on the RTX 5090, pulled a stack of models, and wired them into our...

Dev.to · Rob
🧠 Large Language Models
⚡ AI Lesson
3mo ago
Putting the GPU to Work: Running Local LLMs on a Home Lab
Yesterday we went from a gaming PC on a shelf to a fully configured Coder server with GitHub...

Dev.to · Rob
🧠 Large Language Models
⚡ AI Lesson
3mo ago
Putting the GPU to Work: Running Local LLMs on a Home Lab
Yesterday we went from a gaming PC on a shelf to a fully configured Coder server with GitHub...

Dev.to · Rob
🧠 Large Language Models
⚡ AI Lesson
3mo ago
Model Showdown: Benchmarking Local vs Cloud LLMs on a Real Coding Task
Last post we stood up Ollama on the RTX 5090, pulled a stack of models, and wired them into our...
DeepCamp AI