Use the uploaded model sheet only as a general reference for the human figure.
Create ONE full-body person as a reusable isolated human asset.
The goal is NOT to create a portrait or fashion photograph.
The goal is to generate a natural human figure that can be placed into an architectural visualization, city scene, exhibition, advertisement, or other environment.
CAMERA AND COMPOSITION:
- Fixed camera height and consistent perspective.
- Do not randomize the camera position, camera height, or perspective.
- 9:16 vertical composition.
- Always show the complete person from head to feet.
- Keep the entire body comfortably inside the frame.
- Do not crop the head, hands, or feet.
- Maintain a consistent full-body scale and photographic viewpoint across generations.
RANDOMIZE THE HUMAN:
Randomly vary the person's:
- everyday action or behavior
- body posture
- weight distribution
- arm and hand position
- leg position
- head direction
- gaze direction
- body orientation
- clothing and styling
The action should feel like an ordinary moment from everyday life:
walking, waiting, standing, looking around, checking a phone, carrying a bag, holding an object, gesturing, talking to someone, looking at a building, turning around, pausing while walking, or other natural everyday activities.
Do not use a fixed pose.
Do not use a standard fashion catalog pose.
Do not make the person pose for the camera.
Do not make every person look like a model.
ORIENTATION:
Do not prioritize showing the person's face.
Randomize the person's orientation naturally across the full range:
- front
- front three-quarter
- side
- rear three-quarter
- rear
Rear-facing figures are important and should occur regularly, but they must not dominate the results.
Side and three-quarter views should occur naturally and frequently.
Do not automatically turn the body or head toward the camera.
Do not automatically turn the person away from the camera either.
The face may be fully visible, partially visible, in profile, or completely unseen depending naturally on the person's orientation and activity.
A completely natural rear view is fully acceptable.
Prioritize spatial orientation and natural human behavior over face visibility.
CLOTHING:
Generate a different believable everyday outfit for each generation.
Use good contemporary fashion sense.
Coordinate the clothing naturally with the person's apparent age, activity, body type, and environment.
Avoid repetitive outfits.
Avoid generic fashion-catalog styling.
Avoid theatrical costumes, uniforms, or exaggerated runway fashion unless the generated context naturally calls for them.
The clothing should look like something a real person could naturally be wearing in an everyday urban environment.
HUMAN REALISM:
Maintain believable human anatomy and natural balance.
Allow subtle asymmetry, imperfect posture, relaxed gestures, natural body language, and small unexpected variations.
Do not exaggerate the pose.
Do not make the person acrobatic, crouching, kneeling, sitting, jumping, lying down, or performing an athletic action.
Prefer standing and walking situations because the resulting figure must remain broadly useful for architectural visualization.
BACKGROUND AND OUTPUT:
The background must be completely transparent.
Output as a PNG with true alpha transparency.
No white background.
No colored background.
No gradient.
No studio backdrop.
No floor.
No ground.
No cast shadow.
No environmental elements.
The person must be isolated cleanly from head to feet and ready for direct compositing into another image.
VARIETY:
Each generation should be meaningfully different from the previous generation.
Prioritize variety and discovery over identity preservation.
Do not repeatedly generate the same pose, same orientation, same clothing, or same body language.
The uploaded model sheet provides a general human reference, but exact facial identity preservation is NOT the primary goal.
The most important goal is to generate useful, believable, varied human figures for real-world image compositing.
Generate the image only.