Software Engineering Empirical Research Radar

SE-Jury: An LLM-as-Ensemble-Judge Metric for Narrowing the Gap with Human Evaluation in SE.

Paper detail page in SEER Radar.

Authors

Xin Zhou 0014, Kisub Kim, Ting Zhang 0011, Martin Weyssow, Luís F. Gomes, Guang Yang 0019, Kui Liu 0001, Xin Xia 0001, David Lo 0001

Venue / Year

ASE 2025

Topics

AI / LLM for SE; Human / Empirical / Socio-technical

Abstract / Summary

Summary pending. This page currently provides paper metadata; an abstract or short summary will be added when available.

External Links

DOI / Publisher

PDF

Local PDF is not available on SEER Radar yet. When a public source is recorded, this page will add a local reading link with source attribution.