すべてのプロジェクト

jev-synthetic-survey

jjd-lab/jev-synthetic-survey

Jev vs GPT-4.1 as synthetic survey respondents on Twin-2K-500. How you ask mattered more than which model you used.

研究と評価Python
スター
1
フォーク
0

審査時の参照元

参照元を見る

トピック

jevbehavioral-economicsbenchmarkcalibrationdigital-twindigital-twinsllmllm-evaluation