By clicking "Accept", you agree to the storing of cookies on your device to enhance site navigation, analyze site usage, and assist in our marketing efforts. See our Privacy Policy for more information
Open Datasets
Movie Posters 100K ControlNet
Image

Movie Posters 100K ControlNet

A comprehensive visual dataset of annotated movie posters for training layout-guided generation models like ControlNet.

Download dataset
Size

10,000 high-resolution images + OCR annotations, approximately 42.5 GB, Parquet format

Licence

Apache 2.0

Description

Movie Posters 100K ControlNet is a corpus of 10,000 high-resolution movie posters, enriched with layout annotations and captions. Each image is associated with a title and one or more genres, forming a useful caption for content generation. The visual structures were extracted via PaddleOCR and can be used to train conditional models like ControlNet.

What is this dataset for?

  • Train image generation models from layouts (layout2image)
  • Use ControlNet to generate custom posters from a canvas
  • Analyze the graphic structure of posters for classification or automated design projects

Can it be enriched or improved?

Yes, this dataset can be enriched with additional metadata (movie summaries, dates, directors), standardization of OCR annotations, or by adding cultural or linguistic variations. It is also possible to cross it with other movie databases to extend its use.

🔎 In summary

Criterion Evaluation
🧩 Ease of use⭐⭐⭐⭐⭐ (Ready-to-use, well-structured)
🧼 Need for cleaning⭐⭐⭐⭐⭐ (Low – possibility to standardize captions or OCR)
🏷️ Annotation richness⭐⭐⭐⭐✩ (Extracted layout, useful captions)
📜 Commercial license✅ Yes (Apache 2.0)
👨‍💻 Beginner friendly⚠️ Accessible with some basics in diffusion/vision
🔁 Fine-tuning ready🎨 Ideal for ControlNet, Diffusers
🌍 Cultural diversity⚠️ Medium – depends on the initial poster corpus

🧠 Recommended for

  • Visual AI researchers
  • Generative artists
  • layout2image projects

🔧 Compatible tools

  • ControlNet
  • Stable Diffusion
  • PaddleOCR
  • Diffusers

💡 Tip

For precise control, use OCR layouts as an input mask in your build prompts.

Frequently Asked Questions

Does this dataset include the full texts of the movies?

No, it only includes poster images, genres and titles as captions, and their extracted layout.

Is it suitable for training a ControlNet model?

Yes, it is one of its main uses: to provide a visual base with conditioned layout to generate new posters.

Can we use this dataset to classify movie genres?

Yes, captions include genres, which makes it possible to train a visual or multimodal classifier according to graphic appearance.

Similar datasets

See more
Category

Lorem ipsum dolor sit amet, consectetur adipiscing elit. Suspendisse varius enim in eros elementum tristique.

Category

Lorem ipsum dolor sit amet, consectetur adipiscing elit. Suspendisse varius enim in eros elementum tristique.

Category

Lorem ipsum dolor sit amet, consectetur adipiscing elit. Suspendisse varius enim in eros elementum tristique.