Astra’s “Reasoning Effort” Might Not Mean What You Think
A few benchmark curves reminded me of predictive coding. Now I have a hypothesis.
Deep technical analysis of AI systems and LLM safety.
A few benchmark curves reminded me of predictive coding. Now I have a hypothesis.
If curiosity begins with not knowing, what happens when answers become instant?
I swapped the activation functions inside Gemma. It tolerated far more than I expected.
Introducing EqPropMomentum: A physics-grounded optimizer for biologically plausible AI.
A hypernetwork that rewires GPT's value heads on every forward pass. And the answer is... not straightforward.
Automating the least glamorous, most expensive part of reinforcement learning: building the world.
Building a VLM from scratch using model stitching and LoRA — comparing CLIP's language alignment vs. I-JEPA's world modeling.
How a Single Function Call Gates Safety Alignment in Gemma, Qwen, and Other Open-Source LLMs
My thoughts on the frontier labs' gatekeeping, the rise of the app layer, and why we need open algorithms, not just open weights.
An experiment in making neural networks rewrite themselves for every single input. Inspired by the human brain.
Evaluating Local Feature Attribution and Decision Fidelity in Deep Vision Models via Perturbation-Based Explanations
New phase, new experiments.