Shinnosuke Ono

Master's student, University of Tokyo

prof_pic.png

I am applying to PhD programs for Fall 2027. I hold the Toyota Riken Overseas PreDoctoral Fellowship, which funds two years of doctoral study abroad with up to three further years of support.

I am a second-year master’s student at the University of Tokyo, advised by Prof. Masashi Sugiyama and Prof. Takashi Ishida. Through December 2026 I am a visiting research student at Nanyang Technological University, hosted by Prof. Bo An. I am also a research assistant at the Matsuo-Iwasawa Lab and the lead engineer at EQUES.

personal

Outside of research, I love:

  • running. Best half marathon: 1:43:38 (Yamanakako 2026).
  • music. I also play (a bit of) the guitar.
  • sports. I played soccer (was a 5′6″ goalkeeper). Also love watching various sports.
  • my dear three cats.

research interests

My research goal is to develop reliable AI systems that humans can benefit from working with, and to help our society adapt to them. I am working on foundation models:

  • Reward hacking in RLHF: True human preferences are hard to specify via rule-based rewards, so we need learned reward models (RMs). RMs are vulnerable to reward hacking, where language models exploit imperfect RMs during RL. EIML@ICML2026 quantifies this vulnerability via adversarial robustness and proposes a practical, lightweight algorithm to mitigate reward hacking.
  • Application: IJCNLP-AACL2025 develops a Japanese pharmaceutical LLM via continual pre-training and model merging, and proposes three new evaluation benchmarks for this field. arXiv2025 introduces a large-scale hierarchical mobile manipulation dataset for VLA models.
  • Societal perspective: How can we oversee the evolution of almost superhuman AI systems? Work coming soon!

I am always happy to chat about research. Reach out!

news

Jul 07, 2026 Awarded the Overseas PreDoctoral Fellowship of the Toyota Physical and Chemical Research Institute, funding doctoral study abroad from 2027 🎉
Jul 01, 2026 Started a six-month visiting research stay at Nanyang Technological University with Prof. Bo An, supported by the Heiwa Nakajima Foundation 🇸🇬
May 19, 2026 Our paper Mitigating Reward Hacking in RLHF via Advantage Sign Robustness was accepted to the EIML Workshop at ICML 2026 🎉
Apr 09, 2026 Presented a poster and chaired a session at the Statistical Safeguarding Workshop 2026, hosted by the University of Tokyo and the RIKEN AIP Imperfect Information Learning Team.
Apr 03, 2026 New preprint: Mitigating Reward Hacking in RLHF via Advantage Sign Robustness.

selected publications

  1. Mitigating Reward Hacking in RLHF via Advantage Sign Robustness
    Shinnosuke Ono, Johannes Ackermann, Soichiro Nishimori, and 2 more authors
    In 2nd Workshop on Epistemic Intelligence in Machine Learning (EIML) at ICMLfull version under review , Apr 2026
  2. A Japanese Language Model and Three New Evaluation Benchmarks for Pharmaceutical NLP
    Shinnosuke Ono, Issey Sukeda, Takuro Fujii, and 2 more authors
    In Proceedings of the 4th International Joint Conference on Natural Language Processing and the 14th Conference of the Asia-Pacific Chapter of the Association for Computational Linguistics (IJCNLP-AACL)equal contribution with I. Sukeda , Dec 2025