메뉴 건너뛰기
소속 기관 / 학교 인증
인증하면 논문, 학술자료 등을  무료로 열람할 수 있어요.
한국대학교, 누리자동차, 시립도서관 등 나의 기관을 확인해보세요
(국내 대학 90% 이상 구독 중)
고객센터 ENG
주제분류

논문 기본 정보

저자정보
(Soongsil University) (Soongsil University) (Archipin) (Soongsil university)
저널정보
한국멀티미디어학회 멀티미디어학회논문지 멀티미디어학회논문지 제28권 제8호
오류 신고하기

피인용 0

검색

    초록·키워드

    This paper presents a novel field-adaptive methodology for dense retrieval of structured documents, tackling the persistent semantic gap between natural language queries and field-based content organization. As structured document repositories proliferate in enterprise environments, traditional dense retrieval methods face challenges due to the heterogeneous composition of fields and uneven semantic density. Our approach introduces three key innovations. First, we employ fine-tuned language models with similarity filtering to generate high-fidelity training data, addressing the scarcity of reliable query-document pairs. Second, we implement query-length-based adaptive field weighting, dynamically adjusting the contribution of titles, descriptions, and metadata during bi-encoder contrastive training. Third, we design a two-stage hybrid ranking strategy that combines the efficiency of bi-encoders with the precision of cross-encoders through optimized score integration. Extensive experiments on the Crello dataset, comprising over 25,000 structured documents, demonstrate a 33.8% improvement in Mean Reciprocal Rank (MRR) compared to the baseline, while maintaining inference efficiency. These results establish a scalable and domain-independent solution for structured document retrieval, offering both theoretical contributions and practical feasibility for real-world deployment.

    최근 본 자료 전체보기

      UCI(KEPA) : I410-151-25-02-094042564