# Original Style LoRA Dataset V0 Generation Guide

상태: `round_chibi_direction_reset_before_batch_1`  
목표: train 40장 + holdout 8장 = 총 48장  
현재 생성 범위: 12개 identity의 `neutral_front` 12장만

## 1. 48장 구성

각 캐릭터마다 다음 네 상태를 한 장씩 만든다.

| state | 목적 |
| --- | --- |
| `neutral_front` | identity와 공통 스타일 기준 원화 |
| `speaking_gesture` | 대화 장면의 표정·손짓 |
| `strong_action` | 전신 동작과 실루엣 변화 |
| `extreme_reaction` | 과장된 얼굴·몸 리액션 |

`train_*` 10명은 LoRA 학습에 사용하고 `holdout_*` 2명은 학습에서 완전히
제외한다. holdout은 학습 후 처음 보는 identity에도 스타일이 적용되는지 확인하는
정답 이미지다.

## 2. 반드시 지킬 생성 방식

1. 캐릭터마다 별도 ChatGPT 대화를 만든다. 서로 다른 캐릭터를 같은 대화에서
   연속 생성하지 않는다.
2. 첫 `neutral_front`를 만들 때 2~3등신을 통과한 새 round-chibi 참조만 첨부한다.
   기존의 키 큰 여성과 사람 비율 마스터는 첨부하지 않는다.
3. 참조 이미지에서는 공통 선, 비율 범위, 얼굴 문법, 평면 채색만 가져온다.
   참조 캐릭터의 얼굴, 헤어, 의상, 체형은 복제하지 않는다.
4. 한 이미지에는 한 캐릭터와 한 포즈만 넣는다. 캐릭터 시트, 4분할, 여러 시안,
   나란히 선 비교 이미지는 금지한다.
5. `neutral_front`가 승인된 뒤 같은 캐릭터 대화에서 그 neutral 원화를 첨부해
   나머지 세 상태를 한 장씩 만든다.
6. 모든 결과는 세로 3:4, 전신, 머리부터 신발까지 보이게 만든다. 권장 원본은
   1086x1448 이상의 PNG다.
7. 회색 검수 배경은 유지한다. 배경 제거와 투명 PNG는 승인 후 파생 자산으로 만든다.
8. 이미지 안에 글자, 로고, 워터마크, 소품, 장면 배경을 넣지 않는다.

## 3. Batch 1 Neutral Prompt

아래 프롬프트의 `[IDENTITY]`만 표의 고정 외형으로 바꾼다.

```text
The attached images are approved internal visual-style references.
Use them only for the shared line quality, proportion system, simple facial grammar,
clean flat colors, and restrained one-step cel shading.

Create exactly one new original character. Do not copy the face, hairstyle,
costume, body shape, or signature details of any reference character.

Use a round compact chibi body between 2 and 3 heads tall, targeting 2.6 heads.
The head must be large and round or soft-square. Keep the torso and limbs short.
The eyes, eyebrows, and mouth must be large and simple enough to carry comedy
without requiring a different full-body pose.

IDENTITY:
[IDENTITY]

Full-body front-facing neutral standing pose, eyes looking at the camera,
both arms relaxed down, closed mouth with a restrained slight smile.
Vertical 3:4 canvas, centered, head and both shoes fully visible,
plain light gray review background.

One character, one pose, one image only.
No side glance, no gesture, no prop, no text, no logo, no watermark,
no character sheet, no multiple views, no crop, no extra limbs or fingers,
no photorealism, no 3D render, no painterly texture, no detailed lighting.
```

## 4. Identity 목록

| character_id | split | `[IDENTITY]` 고정 외형 |
| --- | --- | --- |
| `train_01_braided_girl` | train | short child girl, twin braids, red overalls, cream t-shirt, teal sneakers |
| `train_02_curly_teen` | train | slim teenage boy, curly brown hair, green zip jacket, gray shorts, high-top shoes |
| `train_03_pixie_woman` | train | petite young adult woman, black pixie cut, magenta sweatshirt, navy skirt, white trainers |
| `train_04_glasses_man` | train | tall lean young adult man, round glasses, orange cardigan, khaki slacks, brown loafers |
| `train_05_apron_woman` | train | stocky middle-aged woman, wavy shoulder-length hair, mint apron, burgundy dress, flat shoes |
| `train_06_mustache_worker` | train | lean middle-aged man, small mustache, denim work shirt, tan work pants, black work shoes |
| `train_07_elder_bun` | train | short elderly woman, silver hair bun, teal tracksuit, white walking shoes |
| `train_08_elder_vest` | train | tall elderly man, bald crown with gray fringe, mustard check vest, brown trousers, dress shoes |
| `train_09_athletic_woman` | train | athletic adult woman, high ponytail, navy track jacket, coral track pants, running shoes |
| `train_10_heavy_man` | train | heavyset adult man, buzz cut, lavender sweatshirt, charcoal joggers, simple sneakers |
| `holdout_01_bowlcut_child` | holdout | small child boy, round face, dark bowl cut, sky-blue overalls, striped shirt, red canvas shoes |
| `holdout_02_curly_raincoat` | holdout | tall curvy adult woman, voluminous curly hair, yellow raincoat, dark leggings, purple ankle boots |

## 5. Batch 2-4 Identity Lock Prompt

Batch 1의 neutral이 승인된 뒤에만 사용한다. 이때 첫 번째 첨부 이미지는 해당
캐릭터의 승인 neutral, 두 번째 첨부 이미지는 승인 마스터 라인업으로 둔다.

```text
Image 1 is the exact identity source. Preserve the exact face shape, hairstyle,
body proportions, outfit shapes, outfit colors, shoes, and distinguishing features.
Image 2 is the shared internal visual-style reference.

Generate exactly one new full-body state of the same character.

STATE:
[STATE_INSTRUCTION]

Keep the same clean linework, flat colors, and restrained one-step cel shading.
Keep the exact approved 2-to-3-head-tall round compact proportions.
Vertical 3:4 canvas, centered, head and both shoes fully visible,
plain light gray review background.

One character, one pose, one image only.
No redesign, no costume change, no color change, no age change, no body-type change,
no prop, no text, no logo, no watermark, no character sheet, no multiple views,
no crop, no extra limbs or fingers, no photorealism or 3D render.
```

## 6. 캐릭터별 상태 동작

아래 문장을 `[STATE_INSTRUCTION]`에 넣는다. 동작과 감정을 분산해 모든 캐릭터가
같은 포즈를 취하지 않도록 한다.

| character_id | speaking_gesture | strong_action | extreme_reaction |
| --- | --- | --- | --- |
| `train_01_braided_girl` | friendly three-quarter explanation, left palm open | excited hop, one knee raised, both arms spread | sudden surprise, shoulders raised, hands beside cheeks, wide eyes |
| `train_02_curly_teen` | skeptical talking pose, right hand pointing sideways | fast running start, torso leaning forward, arms separated | hard laughter, torso bent slightly, one hand near belly |
| `train_03_pixie_woman` | one hand on hip, other palm turned upward while speaking | sharp side step with a broad arm sweep | angry stomp, clenched fists raised away from face |
| `train_04_glasses_man` | calm explanation, one hand lightly adjusting glasses, other hand open | startled recoil, one foot back, both arms extended | shocked disbelief, hands above shoulders, glasses unchanged |
| `train_05_apron_woman` | confident explanation, one open palm, other hand at waist | decisive forward step with one arm pushing forward | joyful loud laugh, both arms raised, face unobstructed |
| `train_06_mustache_worker` | firm talking pose with one raised index finger | urgent stop gesture while lunging forward | panic, body leaning back, both hands pulled away, mouth wide |
| `train_07_elder_bun` | gentle upright explanation with a small open-hand gesture | quick side step with arms balancing, stable footing | stern scolding reaction, elbows out, fists at waist |
| `train_08_elder_vest` | warm explanation with both hands loosely open | startled backward hop with clear limb separation | roaring laughter, one hand near belly, other arm open |
| `train_09_athletic_woman` | energetic talking pose, one palm open, other hand on hip | dynamic sprint pose with a strong forward line | victory yell, both arms high, wide confident expression |
| `train_10_heavy_man` | defensive explanation with both palms facing outward | frustrated stomp with a wide stance and forward lean | stunned shock, hands spread apart, jaw dropped |
| `holdout_01_bowlcut_child` | counting explanation with one hand, other arm relaxed | sideways leap with arms wide and both feet readable | tearful surprise without actual tears, hands away from face |
| `holdout_02_curly_raincoat` | doubtful explanation, one hand gesturing at shoulder height | dynamic side step, coat hem moving but outfit unchanged | exaggerated disbelief, both palms open, raised brows |

## 7. 파일명과 업로드

파일명은 다음 규칙을 쓴다.

```text
<character_id>__<state>.png
```

예시:

```text
train_01_braided_girl__neutral_front.png
train_01_braided_girl__speaking_gesture.png
train_01_braided_girl__strong_action.png
train_01_braided_girl__extreme_reaction.png
```

대시보드의 `Style LoRA dataset v0`에서 같은 `character_id`와 `state`를 선택해
한 장씩 올린다. 지금은 12개의 `neutral_front`만 올리고 중단한다. 12장을 먼저
검수해 identity 중복, 스타일 이탈, 손·얼굴 오류가 없는지 확인한 뒤 나머지 36장을
생성한다.

## 8. 즉시 재생성할 실패 조건

- 기존 승인 캐릭터와 얼굴, 헤어, 의상이 닮음
- 2등신 미만 또는 3등신 초과, 긴 팔다리, 사람 비율의 몸
- 얼굴이 작거나 각져서 표정이 360x640 화면에서 읽히지 않음
- 12명 중 둘이 같은 얼굴이나 체형으로 보임
- 전신이나 신발이 잘림
- 손가락, 팔, 다리, 얼굴 구조 오류
- 회색 배경 외 장면·소품·글자가 들어감
- 외곽선, 눈·눈썹·입 문법, 색면이 승인 스타일과 다름
- 상태를 바꿀 때 얼굴, 체형, 의상, 색상이 달라짐
- 한 이미지에 여러 인물이나 여러 시안이 들어감
