인문학
사회과학
자연과학
공학
의약학
농수해양학
예술체육학
복합학
지원사업
학술연구/단체지원/교육 등 연구자 활동을 지속하도록 DBpia가 지원하고 있어요.
커뮤니티
연구자들이 자신의 연구와 전문성을 널리 알리고, 새로운 협력의 기회를 만들 수 있는 네트워킹 공간이에요.
초록·키워드
Recent advancements in text-to-image generation have led to the development of human-centric models that create face images based on text prompts and a given face image. Previous approaches have focused on preserving facial features by employing face-specific encoders to extract facial characteristics and train diffusion models. However, these face-specific encoders typically utilize only a single layer of the encoder, which limits their ability to preserve fine-grained facial details. In this paper, we propose a face- centric text-to-image generative model that enhances the preservation of detailed facial features through the integration of a multiscale feature extraction strategy. By extracting features at multiple layers, our method ensures that facial details are more effectively preserved, addressing the limitations of traditional face-specific encoders. The proposed method shows better performance in preserving semantic and delicate facial features compared to existing approaches, as evidenced by higher CLIP-I scores on the Celeb-HQ dataset.
본문·목차
인공지능 문자 인식 모델을 통해 추출된 텍스트로, 일부 오타나 오류가 포함될 수 있으나 지속적으로 개선 중입니다.
오류를 발견하셨다면 해당 부분을 드래그한 후 ' 를 통해 신고해주세요.
오류를 발견하셨다면 해당 부분을 드래그한 후 ' 를 통해 신고해주세요.
최근 본 자료 전체보기
UCI(KEPA) : I410-151-25-02-093761524