Software engineering AI systems

Junda He

Ph.D. Candidate Singapore Management University

I am a Ph.D. candidate in the SOAR group at Singapore Management University, advised by Prof. David Lo. Previously, I received my M.Sc. in Software Engineering and B.Sc. in Computer Science from University College London.

Background

Education

  • Ph.D., Computer Science
    Singapore Management University, Singapore
    School of Computing and Information Systems · Advisor: Prof. David Lo
  • M.Sc., Software Engineering (Distinction)
    University College London, United Kingdom
    Department of Computer Science
  • B.Sc., Computer Science (First Class Honours)
    University College London, United Kingdom
    Department of Computer Science

Honours & Awards

  • PhD Research Excellence Award
    Singapore Management University
  • SCIS Dean's List
    Singapore Management University
  • SMU Presidential Doctoral Fellowship
    Singapore Management University
  • ACM SIGSOFT CAPS Travel Fund
    ACM SIGSOFT
Experience, funding, talks & service

Experience

  • Research Engineer
    Singapore Management University
    AI for software engineering: LLM-based agents, benchmark construction, and evaluation methodology.
  • Research Assistant
    China Academy of Industrial Internet
  • Software Engineer
    iCarbonX

Grants & Funding

  • TitanCA: Vulnerability Detection using Large Code Models
    Government Technology Agency (GovTech), Singapore — total award over SGD $3.8M

Invited Talks

  • From Code to Courtroom: LLMs as the New Software Judges
    Workshop on the Next Generation of CodeLLMs and Their Applications, Korea University, Seoul
    Invited by Prof. Dongsun Kim
  • When LLMs Judge Software: Toward Trustworthy Automated Evaluation
    Lab seminar, Osaka University, Japan
    Invited by Prof. Raula Gaikovina Kula

Teaching

  • Course Development
    Singapore Management University, Singapore
    CS706: Software Mining and Analysis — contributed to course design discussions with Prof. David Lo and prepared lecture slides
  • Course Development
    North Carolina State University, USA
    CSC791 (graduate special topics): designed laboratory exercises and instructional materials for the course

Professional Service

  • Junior Program Committee, MSR 2026, 2024 (Technical Track)Shadow Program Committee, ICSE 2025 (Research Track)
  • AAAI 2027ACL 2025ICLR 2027UIST 2025
  • TSE (Distinguished Reviewer)TOSEMEMSETISTMachine LearningNeurocomputingCACMJSSAutomated Software Engineering
  • ICSE 2022–2024ASE 2022–2025FSE 2023–2024SANER 2023

Selected Software

  • BigCodeBench
    Code generation benchmark with diverse function calls and complex instructions (co-author, ICLR 2025). Public leaderboard of 160+ models; built into AI2 OLMES, NVIDIA NeMo-Skills and OpenCompass.
  • PTM4Tag
    Pre-trained-model tag recommendation for Stack Overflow posts.

Research

I develop methods to evaluate and test LLMs and AI agents, with a focus on their reliability in software engineering.

All research areas & related work

LLM-as-a-Judge & reliable automated evaluation

How LLMs are used to judge software artifacts — and what it would take for their verdicts to be trustworthy.

LLM-as-a-Judge for SE 2026

News

  1. 2026
  2. 2026

    Beyond Text Matching, our work on reference-free evaluation for binary reverse engineering, was accepted at ASE 2026.

  3. 2026

    Learning from the Test, our work on self-referential differential testing for deep RL agents, was accepted at ISSTA 2026.

  4. 2026

    LLM-as-a-Judge for Software Engineering, our literature review and research roadmap, is published in TOSEM.

  5. 2026

    Our community guidelines for empirical SE studies involving LLMs were accepted by EMSE.

  6. 2026

    Our systematic review of AI support for software architecture practice is published in TOSEM.

Show 11 more updatesShow recent updates only
  1. 2026
  2. 2026
  3. 2026
  4. 2025.12
  5. 2025.11

    Invited talk at Korea University, Seoul: From Code to Courtroom: LLMs as the New Software Judges.

  6. 2025.10

    Invited talk at Osaka University: When LLMs Judge Software: Toward Trustworthy Automated Evaluation.

  7. 2025.05

    LLM-Based Multi-Agent Systems for Software Engineering is published in TOSEM, as part of its 2030 roadmap special issue.

  8. 2025

    BigCodeBench was accepted at ICLR 2025 as an oral presentation.

  9. 2025
  10. 2025

    Received the SMU PhD Research Excellence Award and was named to the SCIS Dean’s List.

  11. 2024.11

    PTM4Tag+, our work on Stack Overflow tag recommendation with pre-trained models, is published online in EMSE (Volume 30, 2025).

Get in touch

Let’s connect.

For research conversations and collaborations.

jundahe.2022@phdcs.smu.edu.sg

School of Computing and Information Systems
Singapore Management University · Singapore

Selected Publications

Google Scholar
Show 23 more papersShow selected papers only