RISS 학술연구정보서비스

검색

인기 검색어

    다국어 입력

    http://chineseinput.net/에서 pinyin(병음)방식으로 중국어를 변환할 수 있습니다.

    변환된 중국어를 복사하여 사용하시면 됩니다.

    예시)
    • 中文 을 입력하시려면 zhongwen을 입력하시고 space를누르시면됩니다.
    • 北京 을 입력하시려면 beijing을 입력하시고 space를 누르시면 됩니다.
    닫기
    KCI등재

    문서 레이아웃 인식과 OCR을 결합한 학술 문서 목차 자동 생성 = Automatic Table-of-Contents Generation in Scholarly Documents via Joint Layout Analysis and OCR

    한글로보기

    https://www.riss.kr/link?id=A110180557

    • 0

      상세조회
    • 0

      다운로드
    서지정보 열기
    • 내보내기
    • 내책장담기
    • 공유하기
    • 오류접수

    부가정보

    국문 초록 (Abstract) kakao i 다국어 번역

    본 연구는 복잡한 구조의 학술 문서에서 목차를 자동으로 생성하기 위해, 이미지 기반 문서 레이아웃 분석과 OCR을 결합한 파이프라인을 제안한다. 먼저 DocLayout-YOLO 기반의 학술정보 구조화 인식 모델을 통해 섹션(장/절/항), 본문, 표, 그림, 수식, 페이지정보, 참고문헌정보, 참고문헌영역등 10개 구성 요소를 동시에 탐지하고, 탐지된 섹션 후보에 대하여 영역 단위 OCR을 수행한다. 이어 섹션 깊이 정보 정제 알고리즘을 적용하여문서별 표기 체계에 적응적으로 대응하고 섹션 레벨 정확도를 향상한다. 국내 과학기술 R&D 보고서를 기반으로 구축한 데이터셋으로 학습ㆍ평가를수행한 결과, 제안 시스템은 다양한 문서 형식과 스캔 PDF 환경에서도 End-to-End 목차 자동 생성이 가능함을 확인하였다.
    번역하기

    본 연구는 복잡한 구조의 학술 문서에서 목차를 자동으로 생성하기 위해, 이미지 기반 문서 레이아웃 분석과 OCR을 결합한 파이프라인을 제안한다. 먼저 DocLayout-YOLO 기반의 학술정보 구조화 ...

    본 연구는 복잡한 구조의 학술 문서에서 목차를 자동으로 생성하기 위해, 이미지 기반 문서 레이아웃 분석과 OCR을 결합한 파이프라인을 제안한다. 먼저 DocLayout-YOLO 기반의 학술정보 구조화 인식 모델을 통해 섹션(장/절/항), 본문, 표, 그림, 수식, 페이지정보, 참고문헌정보, 참고문헌영역등 10개 구성 요소를 동시에 탐지하고, 탐지된 섹션 후보에 대하여 영역 단위 OCR을 수행한다. 이어 섹션 깊이 정보 정제 알고리즘을 적용하여문서별 표기 체계에 적응적으로 대응하고 섹션 레벨 정확도를 향상한다. 국내 과학기술 R&D 보고서를 기반으로 구축한 데이터셋으로 학습ㆍ평가를수행한 결과, 제안 시스템은 다양한 문서 형식과 스캔 PDF 환경에서도 End-to-End 목차 자동 생성이 가능함을 확인하였다.

    더보기

    다국어 초록 (Multilingual Abstract) kakao i 다국어 번역

    This study presents a pipeline for automatic table-of-contents (TOC) generation in scholarly documents with complex structures bycombining image-based document layout analysis and OCR. A DocLayout-YOLO–based scholarly-information structuring model jointlydetects ten components—sections (chapter/section/subsection), body text, tables, figures, formulas, page markers, bibliography heading,and bibliography region—and then performs region-level OCR on detected section candidates. We further apply a Section-DepthRefinement algorithm to adapt to document-specific notation conventions and improve section-level accuracy. Trained and evaluatedon a dataset built from domestic science-and-technology R&D reports, the proposed system demonstrates reliable end-to-end TOCgeneration across diverse formats, including scanned PDFs.
    번역하기

    This study presents a pipeline for automatic table-of-contents (TOC) generation in scholarly documents with complex structures bycombining image-based document layout analysis and OCR. A DocLayout-YOLO–based scholarly-information structuring model j...

    This study presents a pipeline for automatic table-of-contents (TOC) generation in scholarly documents with complex structures bycombining image-based document layout analysis and OCR. A DocLayout-YOLO–based scholarly-information structuring model jointlydetects ten components—sections (chapter/section/subsection), body text, tables, figures, formulas, page markers, bibliography heading,and bibliography region—and then performs region-level OCR on detected section candidates. We further apply a Section-DepthRefinement algorithm to adapt to document-specific notation conventions and improve section-level accuracy. Trained and evaluatedon a dataset built from domestic science-and-technology R&D reports, the proposed system demonstrates reliable end-to-end TOCgeneration across diverse formats, including scanned PDFs.

    더보기

    분석정보

    View

    상세정보조회

    0

    Usage

    원문다운로드

    0

    대출신청

    0

    복사신청

    0

    EDDS신청

    0

    동일 주제 내 활용도 TOP

    더보기

    주제

    연도별 연구동향

    연도별 활용동향

    연관논문

    연구자 네트워크맵

    공동연구자 (7)

    유사연구자 (20) 활용도상위20명

    이 자료와 함께 이용한 RISS 자료

    나만을 위한 추천자료

    해외이동버튼