Skip to yearly menu bar Skip to main content


Poster
in
Workshop: 3rd AI for Math Workshop: Toward Self-Evolving Scientific Agents
Sat, Jul 11, 2026 • 4:00 PM – 5:00 PM KST

$\text{P}^4$Bench: Contamination-Proof, Publicly Verifiable, and Privacy-Preserving LLM Evaluation via Zero-Knowledge Proofs

Enhan Zhao ⋅ Yuanrui Zhang ⋅ Zhang Zhang ⋅ Wei Wu ⋅ Di He

Abstract

Chat is not available.