publications

in reverse chronological order

2026

  1. Mitigating Reward Hacking in RLHF via Advantage Sign Robustness
    Shinnosuke Ono, Johannes Ackermann, Soichiro Nishimori, and 2 more authors
    In 2nd Workshop on Epistemic Intelligence in Machine Learning (EIML) at ICMLfull version under review , Apr 2026

2025

  1. A Japanese Language Model and Three New Evaluation Benchmarks for Pharmaceutical NLP
    Shinnosuke Ono, Issey Sukeda, Takuro Fujii, and 2 more authors
    In Proceedings of the 4th International Joint Conference on Natural Language Processing and the 14th Conference of the Asia-Pacific Chapter of the Association for Computational Linguistics (IJCNLP-AACL)equal contribution with I. Sukeda , Dec 2025
  2. preprint
    pub_airoa.png
    AIRoA MoMa Dataset: A Large-Scale Hierarchical Dataset for Mobile Manipulation
    Ryosuke Takanami, Petr Khrapchenkov, Shu Morikuni, and 32 more authors
    arXiv preprint arXiv:2509.25032, Sep 2025