jev-synthetic-survey
jjd-lab/jev-synthetic-survey
Jev vs GPT-4.1 as synthetic survey respondents on Twin-2K-500. How you ask mattered more than which model you used.
研究と評価Python
- スター
- 1
- フォーク
- 0
審査時の参照元
参照元を見るトピック
jevbehavioral-economicsbenchmarkcalibrationdigital-twindigital-twinsllmllm-evaluation