Shinoda Laboratory · Institute of Science Tokyo

I’m a master’s student in Artificial Intelligence at the Institute of Science Tokyo. My research focuses on speech separation, neural audio codecs, and robust speech recognition.

Outside research, I enjoy photography and am a cinematography enthusiast. You can find my photos and videos in Photo.

Portrait of Phuong Dinh

Current Research

All publications

Improving Audio Codec-based Speech Separation By Stacking Residual Vector Quantization Layers

Nhu Minh Phuong Dinh, Roland Hartanto, Koichi Shinoda

Interspeech 2026 · Accepted

RVQ-Grid explores the structure of residual vector quantization layers for speech separation in neural audio codec representations.

Project

Education

  1. M.Sc. in Artificial Intelligence

    Institute of Science Tokyo

  2. B.Sc. in Information and Communication Technology

    University of Science and Technology of Hanoi

Experience

  1. Graduate Researcher

    Shinoda Laboratory

  2. Software Engineer

    LINE Technology Vietnam

  3. Software Engineer

    ABIVIN

Skills

Research
  • Speech separation
  • Neural audio codecs
  • Deep learning
  • PyTorch
Engineering
  • Python
  • Java
  • Spring Framework
  • Data processing & analytics
  • Docker
  • Kubernetes
  • CI/CD
Languages
  • 🇻🇳 Vietnamese · Native
  • 🇬🇧 English · Full working proficiency
  • 🇯🇵 Japanese · Basic
  • 🇫🇷 French · Basic

Awards

  1. LINE Engineering Culture winner

    LINE

  2. LINE Technology Vietnam Awards winner

    LINE Technology Vietnam

  3. Excellent Student Scholarship (Fully Funded)

    University of Science and Technology of Hanoi

Recently written

All writing
Loading writing