โ† Values

Values evaluation profile

Agent-ValueBench Schwartz Basic Values (PVQ40)

14canonical models
10dimensions
0index weight

About this evaluation

Paper-reported per-model mean value adherence for the named inventory; no cross-axis or cross-source aggregate.

Descriptive value-adherence evidence from an agentic task-conflict benchmark. Higher is not universally better, and these profiles do not affect the Safety & Ethics Index.

Original source โ†— All values evaluations Download model results

Model results

ModelSelf-directionStimulationUniversalismConformitySecurityConfigurations
Claude Haiku 4.56.54.65.75.86.81
Claude Sonnet 4.67.34.97.35.581
DeepSeek V3.27.76.16.95.27.31
Gemini 3 Flash Preview7.15.675.36.51
Gemini 3.1 Pro Preview64.375.56.91
GLM 5.17.75.17.25.37.31
GPT-5.47.33.65.95.77.31
GPT-5.4 Mini5.84.666.86.51
Grok 4.206.74.75.456.81
Kimi K2.57.24.66.95.871
Llama 3.3 70B Instruct6.36.37.16.17.11
MiniMax M2.76.65.26.256.91
Qwen3 30B A3B5.45.25.84.85.81
Qwen3.5 397B A17B7.55.16.65.37.11

Dimensions

MeasureFamilyNative scale
Self-directionValues0 to 10
StimulationValues0 to 10
UniversalismValues0 to 10
ConformityValues0 to 10
SecurityValues0 to 10
TraditionValues0 to 10
AchievementValues0 to 10
HedonismValues0 to 10
BenevolenceValues0 to 10
PowerValues0 to 10