ChaofanTao/Autoregressive-Models-in-Vision-Survey

★ 807⑂ 22

[TMLR 2025🔥] A survey for the autoregressive models in vision.

About ChaofanTao/Autoregressive-Models-in-Vision-Survey

ChaofanTao/Autoregressive-Models-in-Vision-Survey is an open-source project on GitHub, mainly written in several languages. [TMLR 2025🔥] A survey for the autoregressive models in vision. It currently holds 807 stars and 22 forks with 0 open issues, and was last pushed on an unknown date (repository created unknown).

Project Overview

AI Homed tracks it on the AI Image Projects board and on the AI AI Image Projects list.

GitHub Repository Details

Repository ChaofanTao/Autoregressive-Models-in-Vision-Survey · default branch - · size 0 KB · watchers 0 · source: GitHub REST API and repository README

README

[TMLR 2025] Awesome Autoregressive Models in Vision

If you like our project, please give us a star ⭐ on GitHub for the latest update.

Awesome arxiv TechBeat GitHub Repo stars

Autoregressive models have shown significant progress in generating high-quality content by modeling the dependencies sequentially. This repo is a curated list of papers about the latest advancements in autoregressive models in vision.

Paper: [[TMLR 2025🔥]](https://openreview.net/forum?id=1BqXkjNEGP) Autoregressive Models in Vision: A Survey | [[中文解读]](https://mp.weixin.qq.com/s/_O8W1qgvMZu37IKwgtskMA)
Authors: Jing Xiong1,†, Gongye Liu2,†, Lun Huang3, Chengyue Wu1, Taiqiang Wu1, Yao Mu1, Yuan Yao4, Hui Shen5, Zhongwei Wan5, Jinfa Huang4, Chaofan Tao1,‡, Shen Yan6, Huaxiu Yao7, Lingpeng Kong1, Hongxia Yang9, Mi Zhang5, Guillermo Sapiro8,10, Jiebo Luo4, Ping Luo1, Ngai Wong1
1The University of Hong Kong, 2Tsinghua University, 3Duke University, 4University of Rochester, 5The Ohio State University, 6Bytedance, 7The University of North Carolina at Chapel Hill, 8Apple, 9The Hong Kong Polytechnic University, 10Princeton University
Core Contributors, Corresponding Authors


💡 We also have other generative projects that may interest you ✨.

[Personalized Video Generation: Progress, Applications, and Challenges]()
Jinfa Huang, Shenghai Yuan, Kunyang Li, and Meng Cao etc.
github github

📑 Citation

Please consider citing 📑 our papers if our repository is helpful to your work. Thanks sincerely!
@misc{xiong2024autoregressive,
    title={Autoregressive Models in Vision: A Survey},
    author={Jing Xiong and Gongye Liu and Lun Huang and Chengyue Wu and Taiqiang Wu and Yao Mu and Yuan Yao and Hui Shen and Zhongwei Wan and Jinfa Huang and Chaofan Tao and Shen Yan and Huaxiu Yao and Lingpeng Kong and Hongxia Yang and Mi Zhang and Guillermo Sapiro and Jiebo Luo and Ping Luo and Ngai Wong},
    year={2024},
    eprint={2411.05902},
    archivePrefix={arXiv},
    primaryClass={cs.CV}
}

📣 Update News

[2025-11-01] ⏸️ After a year of rapid progress in autoregressive visual generation, two clear trends now define the field: unified multimodal models and autoregressive diffusion-forcing video generation. Our current repository categories no longer capture this evolving landscape, so we’re moving to maintenance mode and pausing proactive updates as of today. The repo remains available as a reference, and targeted PRs are welcome (additions, corrections, or reorganizations with new trends). Thanks for your support! 🙏

[2025-05-31] 🔥 Our survey has been revised in arXiv! The revised paper streamlines content and enhances discussions on:

[2025-03-11] 🔥 Our survey has been accepted by TMLR 2025!

[2024-11-11] We have released the survey: Autoregressive Models in Vision: A Survey.

[2024-10-13] We have initialed the repository.

⚡ Contributing

We welcome feedback, suggestions, and contributions that can help improve this survey and repository and make them valuable resources for the entire community. We will actively maintain this repository by incorporating new research as it emerges. If you have any suggestions about our taxonomy, please take a look at any missed papers, or update any preprint arXiv paper that has been accepted to some venue.

If you want to add your work or model to this list, please do not hesitate to email jhuang90@ur.rochester.edu or pull requests). Markdown format:

* [Name of Conference or Journal + Year] Paper Name. Paper Code

📖 Table of Contents

-----

Image Generation

Unconditional/Class-Conditioned Image Generation

##### Tokenizer ##### Autoregressive Modeling

GitHub Stars & Activity

807Stars
22Forks
0Open issues
-Language

GitHub Popularity

GitHub stars807
Forks22
Open issues0
Primary language-
License-
Stars gained today0
Created-
Last pushed-

Trending History

Trending statusnot on today's boards

Related AI Projects

1

opencv / opencv

C++★ 90,909⑂ 57,033
2

ultralytics / ultralytics

Python★ 61,829⑂ 11,787
3

ultralytics / yolov5

Python★ 58,057⑂ 17,469
4

roboflow / supervision

Python★ 50,960⑂ 4,848
5
6

Anil-matcha / Open-Generative-AI

JavaScript★ 28,946⑂ 5,267
7

lucidrains / vit-pytorch

Python★ 25,514⑂ 3,495
8

junyanz / pytorch-CycleGAN-and-pix2pix

Python★ 25,244⑂ 6,567

More AI Rankings