Training images · Experimental dataset

Synthetic Webcam Scene Dataset

17,280 AI-generated webcam crops, organized into 12 scene categories. Original contact sheets, generation prompts and source-to-crop manifests included.

USD 99 One-time purchase · Images + source code

Checkout and digital delivery are handled by Payhip. The paid product includes both the image dataset and the source-code ZIP.

24 AI-generated webcam scenes, with two examples from each category
24 images from the free sample. View at full size.

What is included in the purchase?

17,280 RGB JPEG crops

Category folders under train/ and val/, ready for your own dataset tooling.

720 original PNG sheets

The generated contact sheets behind the crops, with 60 originals per category.

12 generation prompts

The active category prompts used to request scenes, camera geometry and lighting variation.

Provenance and integrity records

Crop/source manifests, SHA-256 checksums, extraction settings and a dataset card.

Source code included: a companion ZIP for grid extraction, training, ONNX export and a C++/OpenCV reference wrapper. Both ZIPs are delivered with the USD 99 product.

The originals and crops show the same underlying scenes. The 240-image sample is part of the full collection. Neither should be counted as additional independent training examples.

Look at the images before deciding

One example from each category is shown below. Open an image to inspect its original crop size. These are generated scenes, not photographs supplied by real exam participants.

The free sample ZIP contains 240 JPEGs: 20 per category, from 240 distinct source sheets. It includes a sample manifest and is approximately 5.08 MB. It is an inspection sample, not an independent benchmark or a prebuilt train/validation set.

Twelve categories, with explicit limits

Folder labelIntended visual contentCrops
book_visibleA book visible in the scene1,440
camera_occludedAn obstructed camera view1,446
extra_person_visibleAn additional person in the scene1,440
eyes_off_screenA subject depicted looking away1,440
hand_covering_mouthA hand covering the mouth1,440
mobile_phone_visibleA visible mobile phone1,440
mouth_open_speakingAn open mouth / speech-like pose1,440
multiple_facesMultiple visible faces1,428
no_faceNo visible face1,446
one_face_clearA clear single face / intended normal state1,440
paper_notes_visibleVisible paper notes1,440
partial_faceA partly visible or cropped face1,440
Total12 categories17,280

Labels come from the intended generation category. They are not exhaustive manual annotations of every object or event. Some concepts overlap; a still image cannot establish speech, a precise gaze target or misconduct.

Formats, dimensions and split

originals/<category>/<source>.png
images/train/<category>/<crop>.jpg
images/val/<category>/<crop>.jpg
metadata/crops.csv
metadata/sources.csv
metadata/dataset-audit.json
metadata/checksums.sha256
prompts/generation-prompts.md
DATASET_CARD.md

Generated entirely from text prompts

I generated the source sheets in Picsart Pro using the option labelled “GPT Image 2.5.” That is the service label I recorded, not an independently verified underlying model version. I supplied text prompts only and uploaded no real-person reference photographs.

The prompts requested 6 × 4 grids of webcam scenes. Five sheets had a different number of cells. Labels, anatomy and scene details can be imperfect; demographic balance, unique identities and absence of near-duplicate scenes have not been established.

The historical prompt book also has a blurry_or_compressed template. There are no images for that category in this dataset.

Where this collection can be useful

A concrete starting point for dataset-loader tests, training-pipeline prototypes, synthetic-to-real experiments, relabeling exercises and studying generator artifacts. The value is the existing collection, its organization and traceability; whether it improves your model is something to measure.

This is a self-service project snapshot. It does not include installation help, retraining, integration work, custom development, promised updates or a model-performance guarantee. Read the support policy.

Package availability

240-image sample

20 images per category, with a manifest and documentation. Get the free sample through Payhip; no source code is included.

Get the free sample on Payhip

Complete image dataset · USD 99

All crops, original sheets, 12 prompts, provenance records and the companion source-code ZIP. Read the dataset and source-code license terms.

Buy dataset + source · USD 99

The source-code ZIP is included in the USD 99 purchase as a companion download. The free preview contains only the 240-image sample and its documentation. There are no trained model files in the sample, source or full image archive.