Research

Our ongoing research topics:

Superhuman Perception with RF Signals

High-fidelity RF imaging, 3D reconstruction, non-line-of-sight scene understanding, and RF simulation.
PanoRadar
Enabling Visual Recognition at Radio Frequency
MobiCom 2024

TL;DR PanoRadar is a rotating single-chip mmWave radar that brings RF imaging resolution close to LiDAR while staying robust in smoke, fog, and darkness. Its combination of novel signal processing and machine learning enables, for the first time, visual recognition tasks such as semantic segmentation and object detection at radio frequency.

Best Demo Award ACM SRC Grand Finals Second Place
CartoRadar
RF-Based 3D SLAM Rivaling Vision Approaches
MobiCom 2025

TL;DR CartoRadar achieves RF-based 3D SLAM with accuracy rivaling vision-based approaches, overcoming the sparsity and noise of radar measurements with learning. It provides robust mapping and localization in conditions where optical sensors fail, such as smoke, fog, and darkness.

Best Artifact Award
HoloRadar
Non-Line-of-Sight 3D Reconstruction with Radar
NeurIPS 2025

TL;DR HoloRadar reconstructs 3D geometry beyond the line of sight by exploiting multi-bounce RF reflections: a first stage performs multi-return RF imaging, and a second stage carries out reflection-aware scene reconstruction. The result lets a single mmWave radar perceive around corners.

SurfRadar
Surface Characterization with mmWave Signals
MobiSys 2026

TL;DR SurfRadar recovers surface properties such as permittivity and roughness from high-resolution mmWave measurements, producing RF-based 3D material maps of indoor scenes. This physical understanding complements geometry for robotics and wireless applications.

mmWave Sensing and Communication

Millimeter-wave systems for motion capture, wearable sensing, and connectivity.
Mo3Cap
Motion Capture with Millimeter-Wave Tags
SenSys 2026

TL;DR Mo³Cap uses lightweight millimeter-wave tags on the body to enable motion capture without cameras. It extends RF sensing toward fine-grained, occlusion-resilient body tracking for robotics and interaction.

BlinkWise
Tracking Blink Dynamics and Mental States on Glasses
MobiSys 2025

TL;DR BlinkWise is a minimalist RF add-on for everyday glasses that tracks blink dynamics at millisecond resolution, with all processing running on an edge microcontroller for private, real-time use. Detailed blink dynamics unlock applications in drowsiness monitoring, workload assessment, and dry-eye disease management.

Acoustic & Audio Intelligence

Acoustic propagation & field modeling and generative audio.
SmartDJ
SmartDJ: Declarative Audio Editing with Audio Language Model
ICLR 2026

TL;DR SmartDJ turns high-level, declarative instructions into executable audio edits: an audio language model plans step-by-step operations, and editing models carry them out on stereo audio while preserving spatial cues. Users describe the result they want rather than the steps to get there.

AV-Twin
Building Audio-Visual Digital Twins with Smartphones
MobiSys 2026

TL;DR AV-Twin builds editable audio-visual digital twins of real spaces with just a commodity smartphone, combining mobile room impulse response capture, visual-assisted acoustic field modeling, and differentiable acoustic rendering. Edits to geometry, materials, and layout update both sound and visuals.

VERSA
Resounding Acoustic Fields with Reciprocity
NeurIPS 2025

TL;DR VERSA leverages the acoustic reciprocity principle to efficiently learn acoustic fields, enabling resounding: re-rendering how a scene sounds from new source positions. This makes spatial audio capture practical with far fewer measurements.

AVR
Acoustic Volume Rendering for Neural Impulse Response Fields
NeurIPS 2024

TL;DR AVR brings physically grounded volume rendering to spatial audio, learning neural impulse response fields that obey acoustic wave propagation. Together with the AcoustiX simulator, it synthesizes realistic, pose-consistent spatial sound at unseen positions.

Spotlight

Digital Health

Contactless health monitoring and computational biomarkers.
BlinkWise
Tracking Blink Dynamics and Mental States on Glasses
MobiSys 2025

TL;DR BlinkWise is a minimalist RF add-on for everyday glasses that tracks blink dynamics at millisecond resolution, with all processing running on an edge microcontroller for private, real-time use. Detailed blink dynamics unlock applications in drowsiness monitoring, workload assessment, and dry-eye disease management.

Fundus oculomics
Oculomics AI for Cardiovascular Risk Factors
Asia-Pacific Journal of Ophthalmology 2024

TL;DR A case study in fundus oculomics for HbA1c assessment, examining how retinal-imaging AI can screen for cardiovascular risk factors. The work discusses clinically relevant considerations for bringing such models into practice.

Medication self-administration sensing
Assessment of Medication Self-Administration Using Artificial Intelligence
Nature Medicine 2021

TL;DR A contactless wireless sensing system that detects when patients use insulin pens and inhalers and flags administration errors. Deployed unobtrusively in the home, it enables continuous monitoring of medication self-administration.

Immersive Media & Content Creation

Generative models for motion, audio, and world simulation.
WaveVerse
Scalable RF Simulation in Generative 4D Worlds
ICML 2026

TL;DR WaveVerse is a prompt-based framework that generates dynamic indoor 4D worlds and simulates phase-coherent RF signals within them via ray tracing. It provides scalable, realistic RF data for imaging and activity-recognition research when real measurements are scarce.

MoScale
Next-Scale Autoregressive Models for Text-to-Motion Generation
CVPR 2026

TL;DR MoScale generates human motion hierarchically from coarse to fine temporal scales with next-scale autoregressive models. Cross-scale and in-scale refinement improve text-to-motion quality, and the model generalizes zero-shot to motion editing and completion tasks.

SmartDJ
SmartDJ: Declarative Audio Editing with Audio Language Model
ICLR 2026

TL;DR SmartDJ turns high-level, declarative instructions into executable audio edits: an audio language model plans step-by-step operations, and editing models carry them out on stereo audio while preserving spatial cues. Users describe the result they want rather than the steps to get there.

AV-Twin
Building Audio-Visual Digital Twins with Smartphones
MobiSys 2026

TL;DR AV-Twin builds editable audio-visual digital twins of real spaces with just a commodity smartphone, combining mobile room impulse response capture, visual-assisted acoustic field modeling, and differentiable acoustic rendering. Edits to geometry, materials, and layout update both sound and visuals.