Scenario viewer

807 metrics across 26 benchmarks

Autonomy-Preserving Information Provision (20)

Cog Bias (30)

Cognitive Offloading Asymmetry: Over-Scaffolding vs. Autonomy Preservation in LLM Responses (15)

Creativity Support Index (18)

Cultural Context Sensitivity in AI-Mediated Knowledge Production (20)

Emotional Dependence in Long-Term Use of Large Language Models (15)

Evaluation of LLM Chatbot Responses to Psychotic Prompts (13)

Financial (15)

Flourishing AI Benchmark (30)

Flourishing Considerations for AI (6)

Health Bench (20)

HumanAgencyBench (40)

HumaneBench (25)

KORA (130)

LLM Mental Health Refusal and Responsiveness Score (20)

Legal Bench (20)

Modulated Cognitive Autonomy Benchmark (MCAB) (20)

Name Bias (15)

Positive Alignment (8)

ProSocial AI Index (20)

Stanford (147)

The Human Creativity Benchmark (12)

The Spillunder Effect of AI: Cognitive Intrusion and Multitasking Degradation through AI-Assisted Task Planning (15)

User Bias (15)

Validation of Ethical and Responsible AI in Mental Health (VERA-MH) (20)

Weval (98)