openai / imagegpt-medium

huggingface.co
Total runs: 386
24-hour runs: 0
7-day runs: 11
30-day runs: -1
Model's Last Updated: June 12 2023

Introduction of imagegpt-medium

Model Details of imagegpt-medium

ImageGPT (medium-sized model)

ImageGPT (iGPT) model pre-trained on ImageNet ILSVRC 2012 (14 million images, 21,843 classes) at resolution 32x32. It was introduced in the paper Generative Pretraining from Pixels by Chen et al. and first released in this repository . See also the official blog post .

Disclaimer: The team releasing ImageGPT did not write a model card for this model so this model card has been written by the Hugging Face team.

Model description

The ImageGPT (iGPT) is a transformer decoder model (GPT-like) pretrained on a large collection of images in a self-supervised fashion, namely ImageNet-21k, at a resolution of 32x32 pixels.

The goal for the model is simply to predict the next pixel value, given the previous ones.

By pre-training the model, it learns an inner representation of images that can then be used to:

  • extract features useful for downstream tasks: one can either use ImageGPT to produce fixed image features, in order to train a linear model (like a sklearn logistic regression model or SVM). This is also referred to as "linear probing".
  • perform (un)conditional image generation.
Intended uses & limitations

You can use the raw model for either feature extractor or (un) conditional image generation. See the model hub to all ImageGPT variants.

How to use

Here is how to use this model in PyTorch to perform unconditional image generation:

from transformers import ImageGPTImageProcessor, ImageGPTForCausalImageModeling
import torch
import matplotlib.pyplot as plt
import numpy as np

processor = ImageGPTImageProcessor.from_pretrained('openai/imagegpt-medium')
model = ImageGPTForCausalImageModeling.from_pretrained('openai/imagegpt-medium')

device = torch.device("cuda" if torch.cuda.is_available() else "cpu")
model.to(device)

# unconditional generation of 8 images
batch_size = 8
context = torch.full((batch_size, 1), model.config.vocab_size - 1) #initialize with SOS token
context = torch.tensor(context).to(device)
output = model.generate(pixel_values=context, max_length=model.config.n_positions + 1, temperature=1.0, do_sample=True, top_k=40)

clusters = processor.clusters
n_px = processor.size

samples = output[:,1:].cpu().detach().numpy()
samples_img = [np.reshape(np.rint(127.5 * (clusters[s] + 1.0)), [n_px, n_px, 3]).astype(np.uint8) for s in samples] # convert color cluster tokens back to pixels

f, axes = plt.subplots(1, batch_size, dpi=300)
for img, ax in zip(samples_img, axes):
   ax.axis('off')
   ax.imshow(img)
Training data

The ImageGPT model was pretrained on ImageNet-21k , a dataset consisting of 14 million images and 21k classes.

Training procedure
Preprocessing

Images are first resized/rescaled to the same resolution (32x32) and normalized across the RGB channels. Next, color-clustering is performed. This means that every pixel is turned into one of 512 possible cluster values. This way, one ends up with a sequence of 32x32 = 1024 pixel values, rather than 32x32x3 = 3072, which is prohibitively large for Transformer-based models.

Pretraining

Training details can be found in section 3.4 of v2 of the paper.

Evaluation results

For evaluation results on several image classification benchmarks, we refer to the original paper.

BibTeX entry and citation info
@InProceedings{pmlr-v119-chen20s,
  title = 	 {Generative Pretraining From Pixels},
  author =       {Chen, Mark and Radford, Alec and Child, Rewon and Wu, Jeffrey and Jun, Heewoo and Luan, David and Sutskever, Ilya},
  booktitle = 	 {Proceedings of the 37th International Conference on Machine Learning},
  pages = 	 {1691--1703},
  year = 	 {2020},
  editor = 	 {III, Hal Daumé and Singh, Aarti},
  volume = 	 {119},
  series = 	 {Proceedings of Machine Learning Research},
  month = 	 {13--18 Jul},
  publisher =    {PMLR},
  pdf = 	 {http://proceedings.mlr.press/v119/chen20s/chen20s.pdf},
  url = 	 {https://proceedings.mlr.press/v119/chen20s.html
}
@inproceedings{deng2009imagenet,
  title={Imagenet: A large-scale hierarchical image database},
  author={Deng, Jia and Dong, Wei and Socher, Richard and Li, Li-Jia and Li, Kai and Fei-Fei, Li},
  booktitle={2009 IEEE conference on computer vision and pattern recognition},
  pages={248--255},
  year={2009},
  organization={Ieee}
}

Runs of openai imagegpt-medium on huggingface.co

386
Total runs
0
24-hour runs
4
3-day runs
11
7-day runs
-1
30-day runs

More Information About imagegpt-medium huggingface.co Model

More imagegpt-medium license Visit here:

https://choosealicense.com/licenses/apache-2.0

imagegpt-medium huggingface.co

imagegpt-medium huggingface.co is an AI model on huggingface.co that provides imagegpt-medium's model effect (), which can be used instantly with this openai imagegpt-medium model. huggingface.co supports a free trial of the imagegpt-medium model, and also provides paid use of the imagegpt-medium. Support call imagegpt-medium model through api, including Node.js, Python, http.

imagegpt-medium huggingface.co Url

https://huggingface.co/openai/imagegpt-medium

openai imagegpt-medium online free

imagegpt-medium huggingface.co is an online trial and call api platform, which integrates imagegpt-medium's modeling effects, including api services, and provides a free online trial of imagegpt-medium, you can try imagegpt-medium online for free by clicking the link below.

openai imagegpt-medium online free url in huggingface.co:

https://huggingface.co/openai/imagegpt-medium

imagegpt-medium install

imagegpt-medium is an open source model from GitHub that offers a free installation service, and any user can find imagegpt-medium on GitHub to install. At the same time, huggingface.co provides the effect of imagegpt-medium install, users can directly use imagegpt-medium installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

imagegpt-medium install url in huggingface.co:

https://huggingface.co/openai/imagegpt-medium

Url of imagegpt-medium

imagegpt-medium huggingface.co Url

Provider of imagegpt-medium huggingface.co

openai
ORGANIZATIONS

Other API from openai

huggingface.co

Total runs: 6.7M
Run Growth: -1.2M
Growth Rate: -17.72%
Updated:August 27 2025
huggingface.co

Total runs: 5.3M
Run Growth: 739.3K
Growth Rate: 13.84%
Updated:August 27 2025
huggingface.co

Total runs: 2.0M
Run Growth: -176.2K
Growth Rate: -8.96%
Updated:February 29 2024
huggingface.co

Total runs: 1.7M
Run Growth: -1.4M
Growth Rate: -86.54%
Updated:February 29 2024
huggingface.co

Total runs: 908.4K
Run Growth: 272.1K
Growth Rate: 29.96%
Updated:February 29 2024
huggingface.co

Total runs: 800.2K
Run Growth: 268.3K
Growth Rate: 33.53%
Updated:February 29 2024
huggingface.co

Total runs: 252.6K
Run Growth: -210.0K
Growth Rate: -83.11%
Updated:April 23 2026
huggingface.co

Total runs: 212.7K
Run Growth: 125.1K
Growth Rate: 58.80%
Updated:January 23 2024
huggingface.co

Total runs: 140.4K
Run Growth: 87.9K
Growth Rate: 62.58%
Updated:January 23 2024
huggingface.co

Total runs: 24.8K
Run Growth: -59.3K
Growth Rate: -239.29%
Updated:February 29 2024
huggingface.co

Total runs: 2.9K
Run Growth: -232
Growth Rate: -7.87%
Updated:December 12 2023