Topic hub
Tag: AI Benchmark
AI
32-Task Benchmark Shows AI Falls Short on Scientific Figures
32-Task Benchmark Shows AI Falls Short on Scientific Figures: SciDraw-Bench, a 32-task benchmark for scientific figure generation, finds…
AI
SciDraw-Bench Launches to Evaluate AI Scientific Figures
Introduced in a June 24, 2026 arXiv paper, SciDraw-Bench tests 32 tasks across 8 figure types and 10…
AI
CORE-Bench Extends Agent Benchmarks Past Accuracy Saturation
New arXiv research on CORE-Bench v1.1 finds saturated accuracy benchmarks deliver more insight when expanded to measure efficiency,…
The zBrandco Edition
Our pick of the week in AI, open source, gaming, crypto and the tools shaping them.
One email a week · unsubscribe anytime