Yuheng Wu 吴宇恒

Headshot of Yuheng Wu

PhD Candidate

Department of Computer Science

University of Wisconsin-Madison

yuheng.wu [at] wisc.edu

I am a 5th year PhD candidate at the MadAbility Lab at the University of Wisconsin-Madison, supervised by Prof. Yuhang Zhao. My research interests include Human-Centered AI, Human-Computer Interaction, Intelligent Interactive System, Augmented Reality, and Accessibility.

I build context-aware multimodal systems across mobile devices and AR glasses that perceive and understand complex, dynamic real-world environments. I focus on accessibility, with applications extending to education and other everyday tasks.

Before UW-Madison, I received my B.S. in Computer Science from Peking University in 2022.

#Publications

Conference and Journal Papers

2026

  1. Two panels showing NavSight in use. Left: a person standing at a street corner, seen from behind, holding up a phone to scan the intersection ahead. Right: the phone screen, where the background is darkened and navigation-relevant objects are highlighted. A parked SUV, a pedestrian with a backpack, and a distant car are outlined in cyan, and a traffic signal pole is filled in yellow, so they stand out against the dimmed street scene.
    NavSight in the Wild: Understanding Real-World Use of a Mobile Augmented Reality Application for People with Low Vision in Outdoor Navigation
    Yuheng Wu, Kexin Zhang, Ben Kosa, Ru Wang, Sanbrita Mondal, and Yuhang Zhao
    IMWUT, 2026
    📱 Available on Apple App Store
  2. SceneGlance augmenting two cluttered scenes for low-vision users, shown side by side. Left, labeled OUTDOOR: a busy street where cyclists, bicycles, cars, pedestrians, and traffic and pedestrian signals are outlined in green to mark higher importance, while the sidewalk and other objects are outlined in blue. Right, labeled INDOOR: a kitchen counter where knives, cups, and a bottle carry yellow solid overlays for higher importance, while utensils, bowls, and other items are outlined in blue.
    What to Distinguish and How? Opportunities and Challenges of Augmenting Multiple, Cluttered Objects in Complex Scenes for People with Low Vision
    Yuheng Wu, Ruijia Chen, Jaewook Lee, Jia Li, Kexin Zhang, Meng Fong Lio, and 5 more authors
    ASSETS, 2026
  3. NaviNote enables blind and low vision (BLV) users to explore their surroundings via a five-stage pipeline after localizing their precise positions using Visual Positioning System (VPS) in a pre-scanned area on a smartphone: (1) User asks about the area using natural language; (2) User navigates to a specific location within the area by following the turn-by-turn instructions NaviNote provides; (3) User listens to nearby spatial annotations created by other BLV users; (4) User asks follow-up questions on spatial annotations and from other information sources; (5) User creates their own spatial annotations. Arrows in the figure indicate the internal flow of the pipeline. During the interaction, the user wears a vest that holds the smartphone with camera facing forward.
    NaviNote: Enabling In-situ Spatial Annotation Authoring to Support Exploration and Navigation for Blind and Low Vision People
    🏆 Honorable Mention | 📜 US Patent Pending
    Ruijia Chen*, Yuheng Wu*, Charlie Houseago, Filipe Gaspar, Filippo Aleotti, Dorian Gálvez-López, and 6 more authors
    * Equal contribution
    CHI, 2026
  4. Overview of AskNow, an interactive system powered by LLM, consisting of a Student Interface and an Instructor Interface. In a large-scale classroom: (1) AskNow listens to the instructor using streaming speech-to-text technology; (2) Students can ask questions through their own devices and get immediate feedback grounded in the ongoing lecture context; (3) The Instructor Interface displays common confusion areas in real time.
    AskNow: An LLM-powered Interactive System for Real-Time Question Answering in Large-Scale Classrooms
    Ziqi Liu, Yuankun Wang, Hui-Ru Ho, Yuheng Wu, Yuhang Zhao, and Bilge Mutlu
    CHI, 2026

2022

  1. Teaser image of the project: TreeVisual: Design and Evaluation of a Web-Based Visualization Tool for Teaching and Learning Tree Visualization
    TreeVisual: Design and Evaluation of a Web-Based Visualization Tool for Teaching and Learning Tree Visualization
    Brendan J. O'Handley, Yuheng Wu, Haobin Duan, and Chaoli Wang
    ASEE Annual Conference & Exposition, 2022

Doctoral Symposium

2026

  1. My dissertation comprises three research threads toward a single goal. The first, SceneGlance (ASSETS '26), addresses RQ1, how to augment cluttered scenes: it surfaces the design challenges of augmenting multiple objects in cluttered environments. The second, ongoing work on stair navigation, addresses RQ2, how to augment dynamic scenes: it senses stairways, obstacles, handrails, and people in real time, and explores augmentation designs across relevant objects and traversal stages. The third, NavSight (IMWUT '26), addresses RQ3, how AR is used in real-world settings: a deployable mobile AR app for outdoor navigation, studied in a seven-day diary study. All three threads motivate the final goal of scene-aware AR for complex, dynamic scenes, which infers what and how to augment from the scene and context and adapts augmentation automatically, without manual input.
    Scene-Aware Augmented Reality for People with Low Vision in Complex, Dynamic Environments
    Yuheng Wu
    UIST Adjunct, 2026

Workshop Papers

2026

  1. Preliminary Exploration on Intelligent Social Cues in Wearable Augmented Reality to Support Passing Encounters for People with Low Vision
    Ben Kosa, Kexin Zhang, Yuheng Wu, Sanbrita Mondal, and Yuhang Zhao
    Next Steps for Augmented Reality On-the-Move: Challenges & Opportunities, CHI Workshop, 2026

#Education

Logo of University of Wisconsin–Madison University of Wisconsin-Madison
Ph.D. in Computer Science
Advisor: Prof. Yuhang Zhao
Sep. 2022 - May 2027 (expected)
Logo of Peking University Peking University
B.S. in Computer Science, Turing Class (elite CS Program)
Aug. 2018 - Jul. 2022

#Professional Experience

Logo of Niantic Spatial Niantic Spatial
Research and Development Intern
Mentors: Dr. Jessica Van Brummelen and Prof. Gabriel Brostow
May. 2025 - Sep. 2025
London, United Kingdom
Logo of Microsoft Microsoft Research Asia
Research Intern, Data and Knowledge Intelligence (DKI)
Mentor: Dr. Yun Wang
Jan. 2022 - Jun. 2022
Beijing, China

#Service and Teaching

Reviewer

UIST'26, CHI'26, ICCV'25, CSCW'24, ISMAR'24

Workshops

Co-organizer, Workshop on Vision Foundation Models and Generative AI for Accessibility: Challenges and Opportunities, ICCV 2025

Teaching

  • CS320: Data Science Programming II
    Teaching Assistant, University of Wisconsin-Madison, Fall 2025, Spring 2026
  • CS220: Data Science Programming I
    Head Teaching Assistant, University of Wisconsin-Madison, Spring 2024
  • CS577: Introduction to Algorithms
    Teaching Assistant, University of Wisconsin-Madison, Spring 2023
  • CS220: Data Science Programming I
    Teaching Assistant, University of Wisconsin-Madison, Fall 2022, Fall 2023, Summer 2024

#Miscellaneous

  • Award & Honors: David G. Walsh Research Travel Awards, Fall 2025
  • AI / Large Language Models: Multi-Agent Orchestration, Vision-Language Models, Agentic LLM Systems, Multi-modal Processing, Real-Time Pipelines, Tool-use, Human-Centric Evaluation
  • AR / VR Development: Unity, MRTK (HoloLens), OpenXR, Meta XR, Niantic VPS/Lightship, ARKit, A-Frame, Sensor Fusion, Spatial Computing
  • Visual Computing: PyTorch, Core ML, OpenCV, Model Fine-Tuning, On-Device Inference, Client-Server Streaming, Real-time Recognition/Segmentation Pipelines
  • Programming Languages: Python, JavaScript, C#, C/C++, Swift, HTML/CSS, SQL
  • Other skills: MATLAB, R, ReactJS, VueJS, Flask, Git
  • Languages: Chinese Mandarin, Chinese Cantonese, English
    • Slowly learning Spanish 🇪🇸