Scholar
Florian E. Dorner
Google Scholar ID: aYHq31IAAAAJ
ETH Zürich, Max-Planck Institut für Intelligente Systeme
Follow
Google Scholar
↗
Citations & Impact
All-time
Citations
202
H-index
8
i10-index
8
Publications
13
Co-authors
9
list available
Publications
8 items
Cutting LLM Evaluation Costs with SySRs: A Bandit Algorithm that Provably Exploits Model Similarity
2026
Cited
0
Delay, Plateau, or Collapse: Evaluating the Impact of Systematic Verification Error on RLVR
arXiv.org · 2026
Cited
1
Stop Evaluating AI with Human Tests, Develop Principled, AI-specific Tests instead
2025
Cited
0
ROC-n-reroll: How verifier imperfection affects test-time scaling
2025
Cited
0
How Benchmark Prediction from Fewer Data Misses the Mark
2025
Cited
0
Limits to scalable evaluation at the frontier: LLM as Judge won't beat twice the data
arXiv.org · 2024
Cited
2
Training on the Test Task Confounds Evaluation and Emergence
arXiv.org · 2024
Cited
6
Incentivizing Honesty among Competitors in Collaborative Learning and Optimization
Neural Information Processing Systems · 2023
Cited
3
Co-authors
7 total
Moritz Hardt
Max Planck Institute for Intelligent Systems
Nikola Konstantinov
Tenure-track faculty, INSAIT, Sofia University
Martin Vechev
Full Professor of Computer Science, ETH Zurich; Scientific Director, INSAIT, Sofia University
Tom Sühr
Max Planck Institute for Intelligent Systems
Ross Gruetzemacher
Wichita State University
Elliott Ash
Associate Professor of Law, Economics, and Data Science
Vivian Y. Nastl
Max Planck Institute for Intelligent Systems, Tübingen, Germany and ETH Zürich