About

I am Daisuke Nohara, a first-year master’s student at the Institute of Science Tokyo (formerly Tokyo Institute of Technology), conducting research in the Rio Yokota Laboratory. My research focuses on reinforcement learning for improving reasoning in large language models, with particular interests in controlling reasoning behavior and understanding the optimization principles behind stable and efficient RL post-training.

I also enjoy building software systems as a hobby, from low-level Rust engines to self-hosted infrastructure.

Research Interests

My research centers on improving and analyzing reasoning in large language models through reinforcement learning. Specific topics I am currently exploring include:

Publications

First-Author

Co-Authored

Technical Writing

Experience

Education

Activities

Conference presentations and other research activities. View activities →

Hobbies