Blogs
Read the latest updates, guides, and thoughts from the team building AATMA.
Understanding Generative AI from the Ground Up with microgpt
Explore the algorithmic essence of large language models with this minimalist guide to training a GPT from scratch in 200 lines of pure Python.
When AI Gets Obsessed: The Eiffel Tower Llama Experiment
Researchers tweaked Llama's neurons to make it obsessed with the Eiffel Tower, generating April Fools pranks and pickup lines that always mention the famous landmark.
AI Agents Are Sending Angry Emails and Writing Hit Pieces Now
An AI agent sent six emails a minute, got banned from an open-source project, and wrote an angry blog post calling the maintainer a prejudiced gatekeeper.
Generative AI and AI Product Moats
A look at eight observations on generative AI and product moats, shared recently on the Cohere blog.
Moving To Substack
The author announces they are freezing their current blog and moving to Substack for a more convenient authoring experience.
Staj Günlükleri #2: Mühendislikte Proje Yönetimi, Agile/Scrum ve Kurumsal Süreçler
An intern shares insights from a week focused on project management, Agile/Scrum, and enterprise processes in an embedded systems company.
99.4% Accurate but Still Completely Useless?
Accuracy alone can be a misleading metric on imbalanced datasets. A model can be 99.4% correct and still catch zero fraud cases.
AI Migrated Legacy COBOL Programs to Java, Bugs Included
A new AI agentic method called the Locksmith Loop achieves up to 91.90% branch coverage when migrating COBOL to Java, but the remaining bugs are the real challenge.
Deep Neural Nets: 33 Years Ago and 33 Years From Now
Andrej Karpathy reproduces Yann LeCun's 1989 backpropagation paper and uses it as a case study on the nature of progress in deep learning over 33 years.
microgpt: 200 Lines of Pure Python That Trains a GPT
Andrej Karpathy's microgpt distills the entire GPT algorithm into 200 lines of pure Python with zero dependencies. It's a beautiful educational artifact.
Scaling Laws, Carefully
A deep dive into the empirical foundations of scaling laws, the famous Kaplan-Chinchilla disagreement, and why fitting these power laws is trickier than it looks.
Sam Altman, the Decel Debate, and Why Neither Frame Is Useful
Altman says we may need to 'harden around' new capability levels. His framing assumes one path, and speed is the only dial we actually have to set here.