FoundationVision/var

Model

93

stars

10

commits

10

repos using this model

2

linked in READMEs

Dec 22, 2024

updated

README

Model Card for VAR (Visual AutoRegressive) Transformers 🔥

arXivdemo platform

VAR is a new visual generation framework that makes GPT-style models surpass diffusion models for the first time🚀, and exhibits clear power-law Scaling Laws📈 like large language models (LLMs).

VAR redefines the autoregressive learning on images as coarse-to-fine "next-scale prediction" or "next-resolution prediction", diverging from the standard raster-scan "next-token prediction".

This repo is used for hosting VAR's checkpoints.

For more details or tutorials see https://github.com/FoundationVision/VAR.

Contributors

co163

5 commits

keyu-tian

3 commits

JI
jiangyi.enjoy

1 commits

TI
tiankeyu

1 commits

FoundationVision/var

Model

93

stars

10

commits

10

repos using this model

2

linked in READMEs

Dec 22, 2024

updated

README

Model Card for VAR (Visual AutoRegressive) Transformers 🔥

arXivdemo platform

VAR is a new visual generation framework that makes GPT-style models surpass diffusion models for the first time🚀, and exhibits clear power-law Scaling Laws📈 like large language models (LLMs).

VAR redefines the autoregressive learning on images as coarse-to-fine "next-scale prediction" or "next-resolution prediction", diverging from the standard raster-scan "next-token prediction".

This repo is used for hosting VAR's checkpoints.

For more details or tutorials see https://github.com/FoundationVision/VAR.

Contributors

co163

5 commits

keyu-tian

3 commits

JI
jiangyi.enjoy

1 commits

TI
tiankeyu

1 commits