Yung-Hsu (Roy) Yang (楊詠旭)

I am a Ph.D. student at CVG ETH Zürich supervised by Prof. Marc Pollefeys and Prof. Hermann Blum. My research interests include scene understanding and 3D Object Detection and Tracking. Starting June 2026, I work as a student researcher at Google Zürich, hosted by Christina Tsalicoglou, Erik Sandström, and Federico Tombari.

I obtained my M.Sc. and B.Sc. degrees in Electrical Engineering department at National Tsing Hua University supervised by Prof. Min Sun. Previously, I worked with Dr. Samuel Rota Bulò and Dr. Peter Kontschieder in dense prediction tasks.

Email  /  Scholar  /  Github  /  Linkedin  /  X  /  Bluesky  /  Huggingface  /  CV

profile photo

Research

LeAD-M3D: Leveraging Asymmetric Distillation for Real-time Monocular 3D Detection
Johannes MeierJonathan MichelOussema DhaouadiYung-Hsu YangChristoph ReichZuria BauerStefan RothMarc PollefeysJacques KaiserDaniel Cremers
ECCV, 2026
arXiv / Code
DVPSFormer: Efficient Online Depth-aware Video Panoptic Segmentation for Autonomous Driving
Yung-Hsu YangLuigi PiccinelliSiyuan LiMattia SeguLei KeMartin DanelljanYuqian FuZuria BauerFisher YuHermann BlumMarc Pollefeys
arXiv, 2026
arXiv / Code
3D-MOOD: Lifting 2D to 3D for Monocular Open-Set Object Detection
Yung-Hsu YangLuigi PiccinelliMattia SeguSiyuan LiRui HuangYuqian FuMarc PollefeysHermann BlumZuria Bauer
ICCV, 2025
arXiv / Code / Demo
FlowR: Flowing from Sparse to Dense 3D Reconstructions
Tobias FischerSamuel Rota BulòYung-Hsu YangNikhil Varma KeethaLorenzo PorziNorman MüllerKatja SchwarzJonathon LuitenMarc PollefeysPeter Kontschieder
ICCV, 2025 (Highlight)
arXiv / Code
UniK3D: Universal Camera Monocular 3D Estimation
Luigi PiccinelliChristos SakaridisMattia SeguYung-Hsu YangSiyuan LiWim AbbeloosLuc Van Gool
CVPR, 2025
arXiv / Code / Demo
UniDepthV2: Universal Monocular Metric Depth Estimation Made Simpler
Luigi PiccinelliChristos SakaridisYung-Hsu YangMattia SeguSiyuan LiWim AbbeloosLuc Van Gool
TPAMI, 2025
arXiv / Code
Samba: Synchronized Set-of-Sequences Modeling for End-to-end Multiple Object Tracking
Mattia SeguLuigi PiccinelliSiyuan LiYung-Hsu YangBernt SchieleLuc Van Gool
ICLR, 2025 (Spotlight)
arXiv / Code
SLAck: Semantic, Location, and Appearance Aware Open-Vocabulary Tracking
Siyuan LiLei KeYung-Hsu YangLuigi PiccinelliMattia SeguMartin DanelljanLuc Van Gool
ECCV, 2024
arXiv / Code
UniDepth: Universal Monocular Metric Depth Estimation
Luigi PiccinelliYung-Hsu YangChristos SakaridisMattia SeguSiyuan LiLuc Van GoolFisher Yu
CVPR, 2024 (Highlight)
arXiv / Code / Video
CR3DT: Camera-RADAR Fusion for 3D Detection and Tracking
Nicolas Baumann*Michael Baumgartner*Edoardo Ghignone*Jonas Kühne*Tobias FischerYung-Hsu YangMarc PollefeysMichele Magno
IROS, 2024
arXiv / Code
Dense Prediction with Attentive Feature Aggregation
Yung-Hsu YangThomas E. HuangMin SunSamuel Rota BulòPeter KontschiederFisher Yu
WACV, 2023
arXiv / Code / Video
CC-3DT: Panoramic 3D Object Tracking via Cross-Camera Fusion
Tobias Fischer*Yung-Hsu Yang*Suryansh KumarMin SunFisher Yu
CoRL, 2022
arXiv / Code / Video
Monocular Quasi-Dense 3D Object Tracking
Hou-Ning HuYung-Hsu YangTobias FischerTrevor DarrellFisher YuMin Sun
TPAMI, 2022
arXiv / Code / Video

Project

Vis4D: A modular library for 4D scene understanding
Yung-Hsu Yang*Tobias Fischer*Thomas E. Huang*Tao SunRené ZurbrüggFisher Yu

Academic Services

  • Conference Reviewer
    • CVPR: 2023, 2024, 2025, 2026
    • ICCV: 2025
    • ECCV: 2024, 2026
    • NeurIPS: 2023, 2024, 2025
    • ICLR: 2024, 2025
    • ICML: 2025
    • AAAI: 2026
    • ACM MM: 2024
    • 3DV: 2025, 2026
    • WACV: 2026
    • BMCV: 2026
  • Journal Reviewer
    • TPAMI
    • RA-L
    • TIP
    • IJCV

Awesome website template