Spatial-Temporal Large Language Model for Traffic Prediction

This repository contains the code of ST-LLM "Spatial-Temporal Large Language Model for Traffic Prediction" paper

Abstract

Traffic prediction, an essential component for intelligent transportation systems, endeavours to use historical data to foresee future traffic features at specific locations. Although existing traffic prediction models often emphasize developing complex neural network structures, their accuracy has not improved. Recently, large language models have shown outstanding capabilities in time series analysis. Differing from existing models, LLMs progress mainly through parameter expansion and extensive pretraining while maintaining their fundamental structures. Motivated by these developments, we propose a Spatial-Temporal Large Language Model (ST-LLM) for traffic prediction. In the ST-LLM, we define timesteps at each location as tokens and design a spatial-temporal embedding to learn the spatial location and global temporal patterns of these tokens. Additionally, we integrate these embeddings by a fusion convolution to each token for a unified spatial-temporal representation. Furthermore, we innovate a partially frozen attention strategy to adapt the LLM to capture global spatial-temporal dependencies for traffic prediction. Comprehensive experiments on real traffic datasets offer evidence that ST-LLM is a powerful spatial-temporal learner that outperforms state-of-the-art models. Notably, the ST-LLM also exhibits robust performance in both few-shot and zero-shot prediction scenarios.

Dependencies

Python 3.11
PyTorch 2.1.2
cuda 11.5
torchvision 0.8.0

> conda env create -f env_ubuntu.yaml

Datasets

We provide preprocessed datasets, which you can access here.
If you need the original datasets, please refer to the ESG.

Training

CUDA_VISIBLE_DEVICES=0
nohup python train.py --data taxi_pick --device cuda:0  > your_log_name.log &

BibTeX

If you find our work useful in your research. Please consider giving a star ⭐ and citation 📚.

@inproceedings{liu2024spatial,
  title={Spatial-temporal large language model for traffic prediction},
  author={Liu, Chenxi and Yang, Sun and Xu, Qianxiong and Li, Zhishuai and Long, Cheng and Li, Ziyue and Zhao, Rui},
  booktitle={MDM},
  year={2024}
}

Acknowledgement

Our implementation adapts OFA as the code base and has extensively modified it for our purposes. are grateful to the authors for providing their implementations and related resources.

Name		Name	Last commit message	Last commit date
Latest commit History 35 Commits
LICENSE.txt		LICENSE.txt
README.md		README.md
ST-LLM.pdf		ST-LLM.pdf
env_ubuntu.yaml		env_ubuntu.yaml
model_GAT_GPT.py		model_GAT_GPT.py
model_GCN_GPT.py		model_GCN_GPT.py
model_ST_LLM.py		model_ST_LLM.py
pkl.py		pkl.py
ranger21.py		ranger21.py
test.py		test.py
train.py		train.py
train_gcn.py		train_gcn.py
util.py		util.py
weight.py		weight.py

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

Spatial-Temporal Large Language Model for Traffic Prediction

Abstract

Dependencies

Datasets

Training

BibTeX

Acknowledgement

About

Releases

Packages

Languages

License

weiwei5c5/ST-LLM

Folders and files

Latest commit

History

Repository files navigation

Spatial-Temporal Large Language Model for Traffic Prediction

Abstract

Dependencies

Datasets

Training

BibTeX

Acknowledgement

About

Resources

License

Stars

Watchers

Forks

Releases

Packages 0

Languages

Packages