Poplar: A Scalable Pipeline for Human-Centric Image Dataset Synthesis
Abstract
Recent image generators can synthesize convincing human-centric images, yet producing a useful collection remains different from producing a single successful image. A human-centric dataset must cover varied people and contexts, avoid implausible attribute combinations, preserve an everyday photographic character, and expose quality-control decisions at scale. We present Poplar, a reproducible Specify--Render--Inspect pipeline for human-centric image dataset synthesis. Specify samples structured attributes under commonsense constraints and verbalizes them as photography-oriented prompts. Render uses a realism-adapted image generator across composition-aware aspect ratios and retries obvious technical failures. Inspect applies a single structured vision--language review to each candidate, preserving the original prompt while rejecting intrinsic image defects or material prompt mismatches. Using Poplar, we construct Poplar-9K: 9,401 curated human-centric image--text pairs retained from 11,765 reviewed candidates (79.9\% acceptance). We release the dataset together with the pipeline, configurations, immutable generation prompts, and auditable inspection records as a compact resource for building customizable human-centric collections.
Community
nice job
wow
This is an automated message from the Librarian Bot. I found the following papers similar to this paper.
The following papers were recommended by the Semantic Scholar API
- CogCanvas: A Benchmark for Evaluating Multi-Subject Reference-Based Image Generation (2026)
- StructGen: Disambiguating Multi-Reference Image Generation via Structured Context Modeling (2026)
- DisciplineGen-1M: A Large-Scale Dataset for Multidisciplinary Visual Generation and Editing (2026)
- LCG: Long-Context Consistent Image Generation with Sparse Relational Attention (2026)
- DynEval: Holistic Evaluations of T2I Generative Models in the Wild (2026)
- SatEdit: Mask-Conditioned Image Editing via VLM-Guided Segment Annotation (2026)
- Object-Centric Dataset Resources for Constrained-Data Image Generation and Augmentation (2026)
Please give a thumbs up to this comment if you found it helpful!
If you want recommendations for any Paper on Hugging Face checkout this Space
You can directly ask Librarian Bot for paper recommendations by tagging it in a comment: @librarian-bot recommend
Get this paper in your agent:
hf papers read 2608.00440 Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash Models citing this paper 0
No model linking this paper
Datasets citing this paper 1
Spaces citing this paper 0
No Space linking this paper










