Shinnosuke Ono
Master's student, University of Tokyo
I am applying to PhD programs for Fall 2027. I hold the Toyota Riken Overseas PreDoctoral Fellowship, which funds two years of doctoral study abroad with up to three further years of support.
I am a second-year master’s student at the University of Tokyo, advised by Prof. Masashi Sugiyama and Prof. Takashi Ishida. Through December 2026 I am a visiting research student at Nanyang Technological University, hosted by Prof. Bo An. I am also a research assistant at the Matsuo-Iwasawa Lab and the lead engineer at EQUES.
personal
Outside of research, I love:
- running. Best half marathon: 1:43:38 (Yamanakako 2026).
- music. I also play (a bit of) the guitar.
- sports. I played soccer (was a 5′6″ goalkeeper). Also love watching various sports.
- my dear three cats.
research interests
My research goal is to develop reliable AI systems that humans can benefit from working with, and to help our society adapt to them. I am working on foundation models:
- Reward hacking in RLHF: True human preferences are hard to specify via rule-based rewards, so we need learned reward models (RMs). RMs are vulnerable to reward hacking, where language models exploit imperfect RMs during RL. EIML@ICML2026 quantifies this vulnerability via adversarial robustness and proposes a practical, lightweight algorithm to mitigate reward hacking.
- Application: IJCNLP-AACL2025 develops a Japanese pharmaceutical LLM via continual pre-training and model merging, and proposes three new evaluation benchmarks for this field. arXiv2025 introduces a large-scale hierarchical mobile manipulation dataset for VLA models.
- Societal perspective: How can we oversee the evolution of almost superhuman AI systems? Work coming soon!
I am always happy to chat about research. Reach out!
news
| Jul 07, 2026 | Awarded the Overseas PreDoctoral Fellowship of the Toyota Physical and Chemical Research Institute, funding doctoral study abroad from 2027 🎉 |
|---|---|
| Jul 01, 2026 | Started a six-month visiting research stay at Nanyang Technological University with Prof. Bo An, supported by the Heiwa Nakajima Foundation 🇸🇬 |
| May 19, 2026 | Our paper Mitigating Reward Hacking in RLHF via Advantage Sign Robustness was accepted to the EIML Workshop at ICML 2026 🎉 |
| Apr 09, 2026 | Presented a poster and chaired a session at the Statistical Safeguarding Workshop 2026, hosted by the University of Tokyo and the RIKEN AIP Imperfect Information Learning Team. |
| Apr 03, 2026 | New preprint: Mitigating Reward Hacking in RLHF via Advantage Sign Robustness. |
selected publications
- A Japanese Language Model and Three New Evaluation Benchmarks for Pharmaceutical NLPIn Proceedings of the 4th International Joint Conference on Natural Language Processing and the 14th Conference of the Asia-Pacific Chapter of the Association for Computational Linguistics (IJCNLP-AACL)equal contribution with I. Sukeda , Dec 2025