Methodology-Driven AI Researcher.
Focusing on RLHF/RLVF, hallucination awareness, reward modeling, and Internalizing Consciousness in LLMs.
Undergraduate @ UW-Madison CS
Methodology-Driven AI Researcher
I am interested in reliable learning signals for large language models, especially through RLHF/RLVF, reward modeling, and Internalizing Consciousness in LLMs.
My recent work started from a question about hallucination and evaluation failure, and has gradually pushed me toward post-training as a practical entry point: how RLHF/RLVF pipelines handle real, high-variance human reasoning rather than only cleaner model-generated distributions.
At the same time, I remain deeply interested in hallucination awareness, internalized awareness, and the deeper question of how models come to internalize reliable reasoning rather than merely imitate it.