Artificial Intelligence · 01.09.2026, 13:17 UTC
Emulate or Estimate? The Divergent Strengths of Base and Post-Trained Language Models for Opinion Simulation
| Schweregrad | info |
|---|---|
| Kategorie | Artificial Intelligence |
| Quelle | arXiv cs.AI ↗ |
| Veröffentlicht | 01.09.2026 UTC |
Sicherheitsmeldung mit Schweregrad noch nicht bewertet. Technische Details im Tab „Originaltext“; empfohlene Schritte in der Checkliste.
arXiv:2608.03044v2 Announce Type: replace-cross Abstract: Large language models are increasingly used to simulate human opinions, but prior work reports conflicting results: some studies find promising alignment with human survey data, while others find persona collapse and weak demographic sensitivity. We propose that much of this conflict stems from conflating two distinct tasks. We call the first task emulation, in which models generate individual responses that aggregate into a population distribution. We call the second task estimation, in which models directly predict the population distribution. Evaluating six matched base and post-trained models on the Pew American Trends Panel, we find that base models are the stronger emulators: they produce response distributions closer to human ground truth and better preserve demographic structure. Post-trained models are generally the stronger estimators, producing more accurate distributional predictions when asked directly. We argue that model selection for human simulation should be guided by whether the task requires generating text or predicting distributions.
Maßnahmen
⬇ Als MarkdownVerwandte Beiträge
- info Exact Recovery Thresholds for Weighted Data Selection in Vector-Valued Linear Regression
- info Strong Drafts Need Compact Memories: Long-Context Speculative Decoding with Compressed KV Cache
- info Diffusion-Based Refinement for Kilometer-Scale Probabilistic Precipitation Nowcasting
- info Certified Safety Radii in Forecast-Error Space for Wasserstein Distributionally Robust Small Signal Stability-Constrained AC Optimal Power Flow via Lifted Spectrahedral Containment