// Hello — I'm

Yuhang
Wu (Johann) · 吴昱杭

I'm interested in music technology.

Now / Study

MSc student in Sound & Music Computing at Universitat Pompeu Fabra (UPF), Barcelona.

Now / Work

R&D intern at BMAT Music Innovators.

Published at ISMIR 2024 · Findings of ACL 2024 · IEEE ICETCI 2024

The Route

// Haikou → Beijing → Barcelona

Haikou Beijing Barcelona
Haikou · origin Beijing · study Barcelona · basecamp
01 About

From Haikou, Hainan, now based in Barcelona. During undergrad at BUPT I started out in computer vision — deep-learning semantic segmentation of remote sensing images — which gradually led me to the question I now care about most: can machines genuinely understand music?

At M.A.P I led the benchmark team behind ChatMusician (Findings of ACL 2024), curating 1,000+ music-theory QA pairs later incorporated into the MMLU multimodal subset, and co-authored an ISMIR 2024 study of LLMs' capability of music understanding and generation.

Now an R&D intern at BMAT, actively exploring the intersection of academia and industry.

// Research interests
Music Information Retrieval LLMs for Music Audio Watermarking Generative Audio Computer Vision Signal Processing
Yuhang Wu
02 Publications
ISMIR 2024 San Francisco

Can LLMs "Reason" in Music? An Evaluation of LLMs' Capability of Music Understanding and Generation

Z. Zhou, Y. Wu, Z. Wu, X. Zhang, R. Yuan, Y. Ma, L. Wang, E. Benetos, W. Xue

Findings of ACL 2024 Bangkok

ChatMusician: Understanding and Generating Music Intrinsically with LLM

R. Yuan, H. Lin, Y. Wang, Z. Tian, S. Wu, T. Shen, G. Zhang, Y. Wu, et al.

IEEE ICETCI 2024 Changchun

Research on Fire Detection Based on the YOLOv9 Algorithm

L. Song, Y. Wu, W. Zhang

03 Experience
Mar 2026 – Present
Barcelona

BMAT Music Innovators

R&D Intern

Interning in R&D, actively exploring the intersection of academia and industry.

Aug 2024 – Jul 2025
Beijing

Chinese Academy of Sciences, Institute of Automation

Research Assistant · Audio Watermarking

Watermarking in TTS / music GenAI systems: reviewed the literature on shortcomings of existing watermarking systems, reproduced a classic baseline and constructed a complete workflow.

Aug 2023 – May 2024
Remote

Multimodal Art Projection (M.A.P)

Team Leader (Benchmark) → Team Member (LLM Music Evaluation)

Spearheaded construction of the MusicTheoryBenchmark — 1,000+ music theory QA pairs curated to rigorous academic standards, later incorporated into the MMLU multimodal subset. Then evaluated LLMs (GPT-4, Llama, Qwen, Gemma) on music understanding and generation over ~3,000 samples.

Nov 2023 – Mar 2024
Beijing

Beijing University of Posts and Telecommunications

Research Assistant · Sight-Singing Scoring

Deep-learning automatic scoring for sight-singing, with data-synthesis methods generating unlimited training samples to reduce dependence on fixed datasets.

Mar 2023 – Aug 2023
Beijing

Chinese Academy of Sciences, Aerospace Information Research Institute

Research Assistant · Computer Vision

Semantic segmentation for remote sensing: reproduced state-of-the-art methods, adapted network architectures, and validated across 10+ open and proprietary datasets.

04 Education & Skills

Education

2025 – 2027 (expected)

Universitat Pompeu Fabra — MSc in Sound & Music Computing

Barcelona. MIR, machine learning for music, generative algorithms, computational musicology.

2021 – 2025

Queen Mary University of London & BUPT — B.Eng. Internet of Things

Beijing · joint programme. First Class Honours · GPA 90/100 · top 8% (rank 15/186)

Skills

Programming

PythonC/C++JavaSQLLaTeX

ML & AI

PyTorchOpenMMLabLLMs / PromptingComputer VisionSemantic Segmentation

Audio & Music

MIRSignal ProcessingPure DataSonic VisualiserMuseScore

Tools

GitDockerLinuxJupyter

Languages

English · TOEFL 98Chinese · native