
ISM: Self-Improving Strategy Memory for Continual Mathematical Reasoning
A strategy memory that learns from successful and failed episodes to improve mathematical reasoning with a frozen language model. Evaluated on MATH-Hard and OlympiadBench.
Computer Science · UMBC
Mathematical reasoning. Agent memory. Reinforcement learning.
I am a PhD student in Computer Science at the University of Maryland, Baltimore County. I study how AI systems reason, learn from experience, and use memory, from mathematical problem solving with language models to reinforcement learning and robotics.

A strategy memory that learns from successful and failed episodes to improve mathematical reasoning with a frozen language model. Evaluated on MATH-Hard and OlympiadBench.
Schema-based reasoning, self-improving strategy memory, and learning from feedback in language models.
Deep RL, hierarchical agents, Sim2Real transfer, graph attention in R-GCN frameworks, sample efficiency.
YOLO-based object detection, autonomous drone navigation, real-world RL deployment.
An AI agent that fetches recent ArXiv research papers based on search — like Google Scholar but more user-friendly and readable.
ML model to detect whether an essay was written by a student or an LLM. Finished in the top 25% of the Kaggle leaderboard.
Python package fetching real-time International Space Station data. Over 60,000 downloads on PyPI.
Self-driving prototype using CNNs and OpenCV to predict steering angles from dash-cam image inputs.