Jacky Kwok

I am a third-year PhD student in Computer Science at Stanford University. My research focuses on developing scalable verifiers for agents (with Prof. Azalia Mirhoseini and Ion Stoica) and robots (with Prof. Marco Pavone and Chelsea Finn). Recent works include LLM-as-a-Verifier, RoboMonkey, and CoVer-VLA.

Before joining Stanford, I received my Bachelor's and Master's degrees in Computer Science at UC Berkeley, where I conducted research at the Sky Computing Lab, Berkeley AI Research (BAIR), and iCyPhy Center. My master's thesis focused on efficient and reliable systems for reinforcement learning and robotics under the guidance of Ion Stoica and Edward Ashford Lee.

Email  /  LinkedIn  /  X  /  Projects  /  Google Scholar

profile photo

Selected Publications

LLM-as-a-Verifier: A General-Purpose Verification Framework
Jacky Kwok, Shulu Li, Pranav Atreya, Yuejiang Liu, Yixing Jiang, Chelsea Finn, Marco Pavone, Ion Stoica, Azalia Mirhoseini
Submitted to Neural Information Processing Systems (NeurIPS 2026)
[pdf] [website] [code] [agent] [tweet]

A general-purpose verification framework that provides fine-grained feedback by scaling scoring granularity, repeated verification, and criteria decompositions. SOTA on Terminal-Bench, SWE-Bench, and other agentic benchmarks.

Scaling Verification Can Be More Effective Than Scaling Policy Learning for Vision-Language-Action Alignment
Jacky Kwok, Xilun Zhang, Mengdi Xu, Yuejiang Liu, Azalia Mirhoseini, Chelsea Finn, Marco Pavone
European Conference on Computer Vision (ECCV 2026)
CVPR 2026 Workshop on Scalable Robot Learning (Best Paper)
Invited Talk, NVIDIA RoboticsNVIDIA
[pdf] [code] [tweet]

Introduces a contrastive-based verifier for vision–language–action alignment and a hierarchical test-time verification pipeline that couples high-level prompt optimization with low-level action chunk selection for VLAs

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models
Jacky Kwok, Christopher Agia, Rohan Sinha, Matt Foutter, Shulu Li, Ion Stoica, Azalia Mirhoseini, Marco Pavone
Conference on Robot Learning (CoRL), 2025
Invited Talk, MetaMeta& AMDAMD
[pdf] [code] [tweet]

Proposes an action preference learning framework for VLAs and characterizes embodied test-time scaling laws

Towards Efficient and Deterministic Dataflow Systems for Machine Learning
Jacky Kwok, Edward A. Lee, Ion Stoica
UC Berkeley EECS Dissertations and Theses, 2024UC Berkeley
[pdf] [code]

Master's Thesis on efficient and reliable systems for reinforcement learning and robotics

HPRM: High-Performance Robotic Middleware for Intelligent Autonomous Systems
Jacky Kwok, Shulu Li, Marten Lohstroh, Edward A. Lee
International Conference on Robotics and Automation (ICRA), 2025
[pdf] [code]

A high-performance middleware for robots, achieving up to ~100x lower latency than ROS2

SkyPilot: An Intercloud Broker for Sky Computing
Zongheng Yang, Zhanghao Wu, Ion Stoica (Research Assistant: Jacky Kwok)
USENIX Symposium on Networked Systems Design and Implementation (NSDI), 2023
[pdf] [code]

A system to run, manage, and scale AI workloads on any AI infrastructure

Miscellanea

Teaching

Teaching Assistant, Principles of Robot Autonomy (CS137A)

Service

Reviewer: NeurIPS, CoLM, CoRL, CVPR, ICRA, RSS

Template