Loading published results…
Select Datasets
Search & Filter
KPI
—
Models
—
Datasets
Last updated —
Model Comparison
Click ☆ on rows to compare models
Leaderboard
Sorted by mean score across selected datasets
About this Leaderboard
This leaderboard presents published model evaluations on KETI's ethicality and veracity benchmark datasets.
Detailed K-Prism evaluation code and documentation: alsgur0720/K-Prism on GitHub.
Model submissions and automatic evaluations are not available on this page. Results are updated when the maintainers publish a new evaluation.