← Theme gallery Field Notes of Hwiwon Lee
Research notebook · No. 07 · 2026

Hwiwon
Lee

I build and evaluate AI agents for cybersecurity.

Computer Science PhD student at the University of Illinois Urbana-Champaign, advised by Lingming Zhang. My work sits between software security, systems, and AI.

Portrait of Hwiwon Lee
Hwiwon LeeUIUC · CS

Dispatches

Field log / 2026

SEC-bench Pro, our benchmark for long-horizon software security tasks, is now available on arXiv.

Recognized as a Microsoft Most Valuable Security Researcher for 2026 Q1.

Selected papers

Reading index / 04

SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks?

Hwiwon Lee, Jiawei Liu, Dongjun Kim, Wubing Xia, Ziqi Zhang, Chunqiu Steven Xia, Lingming Zhang

A benchmark of 344 validated vulnerabilities across browser engines and the Linux kernel, designed to measure realistic agent bug hunting.

PDF: SEC-bench Pro

Agentic Vulnerability Reasoning on COTS Binaries

Hwiwon Lee, Jongseong Kim, Lingming Zhang

An end-to-end study of autonomous vulnerability discovery and debugger-verified validation on commercial Windows binaries.

PDF: Agentic Vulnerability Reasoning on COTS Binaries

SEC-bench: Automated Benchmarking of LLM Agents on Real-World Software Security Tasks

Hwiwon Lee, Ziqi Zhang, Hanxiao Lu, Lingming Zhang

A fully automated framework for constructing and evaluating authentic proof-of-concept generation and vulnerability patching tasks.

PDF: SEC-bench

BENZENE: A Practical Root Cause Analysis System with an Under-Constrained State Mutation

Younggi Park, Hwiwon Lee, Jinho Jung, Kevin Koo, Huy Kang

A practical, automated crash-diagnosis system that uses under-constrained state mutation to rank root causes efficiently.

PDF: BENZENE