Jiayang Aurora Sun

Jiayang Sun (Aurora) 孙嘉阳

I am a Computer Science undergraduate at HKU and an exchange student at Princeton. My research goal is building reliable, learning-capable AI agents through rigorous evaluation, verification, and interaction. My research focuses on AI agents, AI evaluation, computer-use agents, AI for mathematics, and reinforcement learning.

At Princeton, I work with Professor Pramod Viswanath, mentored by Zerui Cheng. At HKU, I work with Professor Tao Yu in the XLang Lab.

I am currently looking for PhD opportunities beginning in Fall 2027.

Selected Research

AEGIS-D multi-agent debate pipeline

EMNLP 2026 ICML 2026 AI4Math Workshop

AEGIS-D

When Wrong Answers Look Right: Multi-Agent Debate for High-Precision Verification of Hard Competition Mathematics

Jiayang Sun, Jiawei Xu, Maxm Pan, Zerui Cheng, Pramod Viswanath

AEGIS-D uses five role-specialized agents to cross-check competition-math solutions through structured debate and positive-evidence voting. On 195 hard-to-verify variants, it raises precision from 55.1% to 92.5% and reduces false positives by 8.3×.

OSWorld 2.0 long-horizon computer workflow

Core Contributor XLang Lab

OSWorld 2.0

Benchmarking Computer-Use Agents on Long-Horizon Real-World Tasks

Mengqi Yuan*, Zilong Zhou*, Xinzhuang Xiong*, Weiming Wu, Jiayang Sun, Jiamin Song, Kaiqian Cui, Bowen Wang, Haoyuan Wu, Yitong Li, Dunjie Lu, Haikong Lu, Qi Zhen, Xinyuan Wang, Jiaqi Deng, Yuhao Yang, Cheng Chen, Boyuan Zheng, Alex Su, Xiao Yu, Hao Zou, Saaket Agashe, Xing Han Lu, Manpreet Kaur, Zhengyang Qi, Vincent Sunn Chen, Frederic Sala, Dayiheng Liu, Junyang Lin, Zhou Yu, Yu Su, Siva Reddy, Xin Eric Wang, Peng Qi, Tianbao Xie, Tao Yu

OSWorld 2.0 evaluates agents on 108 realistic, long-horizon, end-to-end workflows across web and desktop applications. Each task requires sustained execution, cross-application state tracking, hidden-state recovery, and reliable final-state verification.

VeRA executable-specification pipeline for verified math benchmark variants

ICML 2026 AI4Math Workshop

VeRA

Math Benchmarks as Executable Specifications

Zerui Cheng, Jiashuo Liu, Chunjie Wu, Jiayang Sun, Jianzhu Yao, Pramod Viswanath, Ge Zhang, Wenhao Huang

VeRA compiles static math benchmarks into executable specifications with templates, generators, and verifiers. VeRA-E probes familiarity, while VeRA-H generates fresh verified tasks to renew saturated benchmarks.

OpenFrontierCS benchmark platform showing open-ended computer science challenges

Community Autoresearch

OpenFrontierCS

Open-Ended Computer Science Challenges for AI Agents

OpenFrontierCS is a community challenge for improving executable solvers on eight open-ended computer-science problems. Deterministic hidden-test evaluators provide partial scores, while promoted solutions become shared baselines for further improvement.

Qualitative comparison of pretrained, baseline, and diversity-aware diffusion fine-tuning outputs

Research Project Reinforcement Learning

Diversity-Aware Fine-Tuning for Diffusion Models

DASW replaces DDPO's uniform timestep weighting with budget-preserving weights from feature-space spread and scheduler variance. Experiments show stronger reward-diversity trade-offs on aesthetic and prompt-image alignment tasks.

Education

The University of Hong Kong

B.Eng. in Computer Science

Selected coursework: artificial intelligence, machine learning, robotics, optimization, and numerical methods.

Princeton University

Undergraduate Exchange in Computer Science

Selected coursework: natural language processing, reinforcement learning, and deep learning.

Experience

Baidu, Inc.

Comate Team Intern

The University of Hong Kong

Teaching Assistant

Led weekly tutorials and supported students across three programming courses: ENGG1330, COMP2113, and COMP1117.

Honors

China Soong Ching Ling Foundation Scholarship

RMB 320,000 scholarship.

ACEHK Student Enrichment Award

Sole recipient selected across HKU.

HKU Dean's Honors List

Recognized for academic achievement at the University of Hong Kong.

Academic Service

Reviewer

ICML 2026 AI4Math Workshop.

Skills

English Proficiency

IELTS overall score of 8.0, including 8.0 in Speaking.

Math

Former Mathematics Olympiad competitor.

Athletics

Member of the HKU Rowing Team.

Miscellaneous

Outside the lab

Always planning the next trip ✈️, happiest in the water 🚣 🏊, and usually up for cooking, singing, dancing, or a good stand-up set.