Skip to content
InfoResearchPeer-reviewedLLM-specific

LLMs Cannot Reliably Judge (Yet?): A Comprehensive Assessment on the Robustness of LLM-as-a-Judge

Published
Record updated
View JSON

Summary

RobustJudge is an automated, modular framework that systematically tests how robust LLM-as-a-Judge systems are across datasets, prompt templates, judge models, attacks and defenses. The study covers 15 attack methods and 8 defense strategies across 13 models, and finds that LLM-based judges stay susceptible under both pointwise and pairwise protocols. Robustness is highly sensitive to prompt-template and judge-model choice, and optimization-based attacks with long suffixes can substantially inflate scores from both PAI-Judge variants.