althjs/sig_q_50s_male_western
sigq50smalewestern
50대 백인 남성 · 영국 장년 신사 / 학자 / 사령관
한국 YouTube 콘텐츠 자동화 파이프라인 hejin 의 시그니처 캐릭터 라이브러리 LoRA.
핵심 정보
캐릭터 정보
외모 베이스: a 50s European/British man with mature distinguished presence, neat side-parted dark hair with significant grey, refined face with strong jawline and thoughtful eyes, well-groomed neat grey moustache (optional small grey beard), slightly heavier mature build with composed authoritative posture. 권장 의상 prior: Victorian formal suit / 1930s three-piece tweed / Edwardian frock coat / military officer's uniform / modern conservative suit. 권장 분위기: 영국 시대극의 장년 신사, 사령관, 50대 변호사/의사, 영국 시대극의 가장. 활용 작품 예: ABC 살인사건의 Sir Carmichael Clarke (조금 더 젊은 버전), 영국 시대극의 사장/사령관. dataset 작성 prompt 베이스: 'a 50s European man with neat dark hair greying at temples, refined strong jawline, thoughtful eyes, well-groomed grey moustache, dignified authoritative expression, painterly Studio Ghibli watercolor, soft warm interior light.'
학습 정보
- 데이터셋: 20 컷 (1024×1024) — 의상 / 표정 / 배경 다양화
- Caption 형식:
hejin_sig_q, <framing> of <identity>, <의상 1~2 단어>, <표정 1~2 단어>(60~90 chars sweet spot) - Caption 원칙: face 묘사 (눈색·머리 길이·피부 등) 는 caption 에 박지 않고 trigger word 만으로 face 호출 — face 학습 신호 dominant
- Base model: Flux 2 Klein 4B
- 학습 도구: ai-toolkit (Ostris) — rank 24 / alpha 16 / lr 1e-4 / steps 1500 / flowmatch scheduler
사용 예시 — WaveSpeed Cloud API
curl -X POST https://api.wavespeed.ai/api/v3/wavespeed-ai/flux-2-klein-4b/text-to-image-lora \
-H "Authorization: Bearer $WAVESPEED_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"prompt": "hejin_sig_q, close-up portrait of a Western man, modern casual shirt, calm smile, Studio Ghibli watercolor",
"size": "1024*1024",
"seed": 12345,
"loras": [{"path": "https://huggingface.co/althjs/sig_q_50s_male_western/resolve/main/sig_q_50s_male_western_klein4b_v1.safetensors", "scale": 0.8}]
}'라이센스
이 LoRA 는 base model Flux 2 Klein 4B 의 라이센스 (apache-2.0) 를 따른다.
Apache 2.0 — 상업 사용 / 수정 / 재배포 자유. royalty-free.
학습 데이터셋 미리보기
학습에 사용된 dataset 일부 (전체 20컷은 dataset/ 폴더 참조):
각 컷의 caption 파일 (dataset/NN.txt) 은 학습 시 사용된 짧은 형식 (60~90 chars) 으로 trigger + framing + identity + 의상 + 표정만 포함.
