Loading published resultsβ¦
Select Datasets
Search & Filter
KPI
β
Models
β
Datasets
Last updated β
Model Comparison
Click β on rows to compare models
Leaderboard
Sorted by mean score across selected datasets
About this Leaderboard
This leaderboard presents published model evaluations on KETI's ethicality and veracity benchmark datasets.
Detailed K-Prism evaluation code and documentation: alsgur0720/K-Prism on GitHub.
K-Prism Dataset
Browse the questions, answers, and explanations, or download the annotations and source images from Hugging Face.
Model submissions and automatic evaluations are not available on this page. Results are updated when the maintainers publish a new evaluation.