Skip to yearly menu bar Skip to main content


Poster
in
Workshop: 3rd AI for Math Workshop: Toward Self-Evolving Scientific Agents
Sat, Jul 11, 2026 • 12:00 AM – 1:00 AM PDT

$\text{P}^4$Bench: Contamination-Proof, Publicly Verifiable, and Privacy-Preserving LLM Evaluation via Zero-Knowledge Proofs

Enhan Zhao ⋅ Yuanrui Zhang ⋅ Zhang Zhang ⋅ Wei Wu ⋅ Di He

Abstract

Chat is not available.