About Me

I am a Ph.D. student in the Department of Technology Management for Innovation at the University of Tokyo, Japan, supervised by Prof. Yutaka Matsuo.

My research focuses on reinforcement learning for audio understanding and reward-guided alignment for audio generation. I am also interested in unified generation of speech and non-speech audio, as well as joint audio–video generation.

Publications

Experience

  1. –Present

    The University of Tokyo

    Ph.D. student in Technology Management for Innovation

  2. –

    ByteDance

    Speech Algorithm Engineer

  3. –

    The University of Tokyo

    Master’s in Electrical Engineering and Information Systems