Paper Detail

Distilling Aggregate Mobility Statistics into a Language Model Policy for Post-Event Crowd Simulation

Tatsuya Amano, Hirozumi Yamaguchi

arxiv Score 4.8

Published 2026-08-20 · First seen 2026-08-21

General AI

Abstract

Pedestrian simulators need a behaviour rule for every agent, but privacy usually limits the data for setting one to aggregate statistics, namely zone-level device counts and origin-to-destination (OD) flows, with no individual trajectories. Such aggregates under-determine individual behaviour, because many different sets of decisions reproduce the same counts. We fine-tune a language model crowd agent so that the simulated population matches the observed destination composition, the fraction of the departing crowd heading to each point of interest. We read this target from the OD flow and reweight the model's own destination distribution onto it by iterative proportional fitting. Because fine-tuning inflates the dominant destination class, we fit the low-rank adapter to trajectories resampled to a corrected training composition that reaches the target after this inflation. On mobile network counts from two baseball games the fine-tuned agent runs without inference-time correction, cutting the destination-share error by 25%, while the grid correlation remains similar across policies.

Workflow Status

Review status
pending
Role
unreviewed
Read priority
soon
Vote
Not set.
Saved
no
Collections
Not filed yet.
Next action
Not filled yet.

Reading Brief

No structured notes yet. Add `summary_sections`, `why_relevant`, `claim_impact`, or `next_action` in `papers.jsonl` to enrich this view.

Why It Surfaced

No ranking explanation is available yet.

Tags

No tags.

BibTeX

@article{amano2026distilling,
  title = {Distilling Aggregate Mobility Statistics into a Language Model Policy for Post-Event Crowd Simulation},
  author = {Tatsuya Amano and Hirozumi Yamaguchi},
  year = {2026},
  abstract = {Pedestrian simulators need a behaviour rule for every agent, but privacy usually limits the data for setting one to aggregate statistics, namely zone-level device counts and origin-to-destination (OD) flows, with no individual trajectories. Such aggregates under-determine individual behaviour, because many different sets of decisions reproduce the same counts. We fine-tune a language model crowd agent so that the simulated population matches the observed destination composition, the fraction of },
  url = {https://arxiv.org/abs/2608.19778},
  keywords = {cs.MA, cs.AI},
  eprint = {2608.19778},
  archiveprefix = {arXiv},
}

Metadata

{}