AIO 20003 · Benchmark
The first open benchmark measured by V/E/S distribution
Instead of the binary question of whether an AI is safe, we measure which value hierarchy an AI answers with. Explore the distribution data from 8 frontier models × 366,120 responses directly in an interactive dashboard.
Model Profiles
AIO Framework evaluation results by foundation model
Click any card to open the model’s 9-section report in a modal. The rankings appendix loads on demand.