I'm currently a research intern at MILeS, working on LLM reasoning. I work to make LLMs more capable and safer, even as they get smarter than the people steering them.
AI-safety interpretability work toward detecting hidden objectives and scheming, even when an LLM's outward behavior is held fixed.
Framework for visualizing, decomposing, and analyzing layer-wise latent trajectories during generation.
Modular agentic framework for automation, learning, and research.
Latent-space guidance system for Meta's Coconut model.