Home
Back to home

Blog

Working notes on continual learning and electrode manufacturing.

I write about materials, machine learning, modelling, and automation, from visual explanations to technical working notes. New posts appear as ideas and experiments develop.

AI research · 10 min

How to (Not) Reinvent the Wheel With Jeff

Jev, Jeff, and typed decisions: shared representations, task-specific heads, calibrated probabilities, and a concrete policy for when a manufacturing model acts or requests review.

AI agents · 10 min

DisCo and the Case for Reusable Agent Skills

DisCo distills repositories and papers into reusable operating skills. Scope, ground, construct, and verify: a practical route from research knowledge to repeatable workflows.

Reading notes · 10 min

Reading ConvMem: Long-Context Reasoning as a Summary Tree

ConvMem reads a long document as a tree of question-conditioned summaries instead of a chain of memory updates, which cuts the path from a fact to the answer from linear to logarithmic. Notes on the method, results and cost, and seven ideas for a path shorter than logarithmic.

LLMs in science · 10 min

I Checked 2,294 arXiv Equations for Upside-Down Fractions

Copying rendered math from a chat window can put a fraction's denominator first. I reviewed every display equation in 60 arXiv papers and pattern-checked 200: 4.9% of display equations have a meaning-changing error, 2 of 118 errors are inverted fractions, and the rate did not change between 2021 and 2026.

AI agents · 7 min

I Ran a Game-Theory RL-Menu Agent in GLEE for 48 Hours (No LLM)

A weekend building a no-LLM game-theory bot for the GLEE economic games (bargaining, negotiation, persuasion), and what actually moved the percentile score: scaling the reward, measuring in the right seat, and closing before inflation eats the pot. Includes a playable version of the three games.

Continual learning · 11 min

Finding a Fair Evaluation Under Changing Representations

Certified retention is easy for frozen features and collapses once the representation moves. A bottom-up stripe benchmark that makes adaptation provably necessary, and CARG, the metric for the accuracy you lose when the retention certificate has to hold.

AI research · 9 min

The Student Becomes the Teacher

Three 2026 papers let a language model supply its own supervision, the external teacher, the ground-truth labels, and the privileged context, all replaced by the model itself. What it buys, and where it breaks.

AI research · 12 min

From Simulating Speech to Optimizing Reasoning

Three 2026 papers argue chain-of-thought text was never the reasoning. Reasoning is a latent direction, activated early, held in superposition, and compressible to a few tokens.

ML systems · 14 min

The Race to O(n)

Every subquadratic-attention family relaxes one constraint while quietly keeping another. Five families, the routing-absorption trap, and why the frontier keeps flinching back to dense attention.

Materials · 8 min

Lithium Dendrites Are Batteries' Dandelions

Ceramic solid electrolytes still fail by dendrite. A March 2026 model derives a closed-form critical current density from first principles, and writes down the same physics that lets a dandelion crack a driveway.

AI & infrastructure · 15 min

AI's Bottleneck Is Physical

AI becomes truly useful when a physical feedback loop (power, heat, latency, sensors, data lineage, safe actuation) can run reliably. Where the loop is closed and why capability is not deployability.

Reading notes · 8 min

Reading CausalPFN: Causal Effects as In-Context Prediction

A NeurIPS 2025 paper trains one transformer on simulated, identifiable causal worlds; new observational data goes in as context, treatment-effect posteriors come out. What holds it up, and where it breaks.