Search | Korea Science

Scene Text Detection with Length of Text (글자 수 정보를 이용한 이미지 내 글자 영역 검출 방법)

Yeong Woo Kim;Wonjun Kim
- Proceedings of the Korean Society of Broadcast Engineers Conference
- /
- 2022.11a
- /
- pp.177-179
- /
- 2022
딥러닝의 발전과 함께 합성곱 신경망 기반의 이미지 내 글자 영역 검출(Scene Text Detection) 방법들이 제안됐다. 그러나 이러한 방법들은 대부분 데이터셋이 제공하는 단어의 위치 정보만을 이용할 뿐 글자 영역이 갖는 고유한 정보인 글자 수는 활용하지 않는다. 따라서 본 논문에서는 글자 수 정보를 학습하여 효과적으로 이미지 내의 글자 영역을 검출하는 모듈을 제안한다. 제안하는 방법은 간단한 합성곱 신경망으로 구성된 이미지 내 글자 영역 검출 모델에 글자 수를 예측하는 모듈을 추가하여 학습을 진행하였다. 글자 영역 검출 성능 평가에 널리 사용되는 ICDAR 2015 데이터셋을 통해 기존 방법 대비 성능이 향상됨을 보였고, 글자 수 정보가 글자 영역을 감지하는 데 유효한 정보임을 확인했다.
PDF

Character-level Region Detection Using Attention Center (어텐션 중심을 이용한 글자 단위 영역 검출)

Kim, Jiin;Jeong, Chang-Sung
- Proceedings of the Korea Information Processing Society Conference
- /
- 2019.10a
- /
- pp.952-953
- /
- 2019
최근 딥러닝으로 진행되는 광학 문자 인식 분야는 대부분 단어 단위로 인식하는 것으로 글자 단위의 영역을 검출하는 데에는 적합하지 못하다. 본 연구는 각 글자의 영역을 검출하기 위해 기존의 딥러닝을 이용한 광학 문자 인식 절차인 단어 분리 과정과 단어 인식 과정을 유지하면서 어텐션 중심을 이용하여 각 글자의 영역을 보다 정확하게 검출하는 것을 목표로 한다. 제안하는 모델은 CRAFT 와 Attention Network 를 사용한 OCR 과정을 확장한 모델로 각 단어 문자열 결과물에 각 글자의 영역을 추가로 나타내게 되며 각 글자와 라벨 간의 IOU 평균은 0.671 로 나타났다.
https://doi.org/10.3745/PKIPS.y2019m10a.952 인용 PDF

A Lightweight Deep Learning Model for Text Detection in Fashion Design Sketch Images for Digital Transformation

Ju-Seok Shin;Hyun-Woo Kang
- Journal of the Korea Society of Computer and Information
- /
- v.28 no.10
- /
- pp.17-25
- /
- 2023
In this paper, we propose a lightweight deep learning architecture tailored for efficient text detection in fashion design sketch images. Given the increasing prominence of Digital Transformation in the fashion industry, there is a growing emphasis on harnessing digital tools for creating fashion design sketches. As digitization becomes more pervasive in the fashion design process, the initial stages of text detection and recognition take on pivotal roles. In this study, a lightweight network was designed by building upon existing text detection deep learning models, taking into consideration the unique characteristics of apparel design drawings. Additionally, a separately collected dataset of apparel design drawings was added to train the deep learning model. Experimental results underscore the superior performance of our proposed deep learning model, outperforming existing text detection models by approximately 20% when applied to fashion design sketch images. As a result, this paper is expected to contribute to the Digital Transformation in the field of clothing design by means of research on optimizing deep learning models and detecting specialized text information.
https://doi.org/10.9708/jksci.2023.28.10.017 인용 PDF HTML

Dynamic Synthesis of Pseudo 2D HMMs for Korean Characters in Key Character Recognition Tasks (키워드 인식을 위한 한글 Pseudo 2D HMM의 동적 합성 방법)

조범준
- The Journal of Korean Institute of Communications and Information Sciences
- /
- v.26 no.6B
- /
- pp.820-827
- /
- 2001
한글은 둘 또는 세 개의 자모가 사각형 영역 안에 적절히 배치된 구조로 되어 있다. 이와 같은 구성 방법에 따라 글자의 영상을 합성하고 이를 실시간에 Pseudo 2D HMM으로 변환하는 방법을 제안한다. 본 방법에 따라 실시간 합성된 모델과 추가의 필러(filler) 모델, 여백 모델을 문서 영상의 글자 영역에서 핵심어 검출에 적용하였다. 실험 결과 최소한의 설계 변수 조정으로도 오검출, 미검출률이 낮고 언어 모델 없이 숫자 89%, 한글 80%의 검출성능을 보였으며, 따라서 제안된 방법이 인쇄 문자 패턴의 실시간 모델링 및 키워드 검출에 효과가 있음을 보였다. 본 연구 결과는 내용 기반의 광학 문서 색인 등에 활용할 수 있다.
PDF

Learning-based Word Segmentation for Text Document Recognition (텍스트 문서 인식을 위한 학습 기반 단어 분할)

Lomaliza, Jean-Pierre;Moon, Kwang-Seok;Park, Hanhoon
- Proceedings of the Korean Society of Broadcast Engineers Conference
- /
- 2018.06a
- /
- pp.41-42
- /
- 2018
텍스트 문서 영상으로부터 단어를 검출하고, LLAH(locally likely arrangement hashing) 알고리즘을 이용하여 이웃 단어 사이의 기하 관계를 표현하는 특징 벡터를 계산한 후, 특징 벡터를 비교함으로써 텍스트 문서를 효과적으로 인식하거나 검색할 수 있다. 그러나, 이는 문서 내 각 단어가 정확하고 강건하게 검출된다는 전제를 필요로 한다. 본 논문에서는 텍스트 내 각 라인을 검출하고, 각 라인 내에서 단어 사이의 간격과 글자 사이의 간격을 깊은 신경망(deep neural network)을 이용하여 학습하고 분류함으로써, 보다 카메라와 텍스트 문서 사이의 거리나 방향이 동적으로 변하는 조건에서 각 단어를 강건하게 검출하는 방법을 제안한다. 모바일 환경에서 제안된 방법을 구현하였으며, 실험을 통해 단어 사이의 간격과 글자 사이의 간격을 92.5%의 정확도로 구별할 수 있으며, 이를 통해 동적인 환경에서 단어 검출의 강건성을 크게 개선할 수 있음을 확인하였다.
PDF

Pigments in the Letters of Hanging Boards of the Joseon Royal Court and Reproduction Experiments (조선왕실 현판 글자의 금색 안료와 재현 실험 연구)

LEE Hyeyoun;LEE Minhye;LEE Heeseung
- Korean Journal of Heritage: History & Science
- /
- v.56 no.3
- /
- pp.118-135
- /
- 2023
Hanging boards of the Joseon royal court are hung on buildings related to the royal family, such as palaces and Jongmyo Shrine, to show the hierarchy and character of the building. In addition, the manufacturing method and materials are recorded in the royal protocols of the Joseon Dynasty, so it is an important material for studying the manufacturing method and material changes at that time. However, the hanging boards were restored several times due to fire or war, and it is presumed that there is a change in the original form and material of the hanging boards. In particular, many hanging boards of the Joseon royal court were written with calligraphy by kings, so there are many forms consisting of gold letters on a black background. This study tried to analyze the pigments remaining in the letters of 44 of the Joseon royal hanging boards, which are presumed to be gold letters, and to find out the changes in the hanging board production method and materials by referring to the analysis results. The letters of the hanging boards studied were classified according to the current state of the gold pigment and the detected components. As a result of the analysis of character pigments, 24 embossing techniques and 5 intaglio techniques were mainly detected with gold (Au), but 15 embossing techniques were detected with brass (Cu, Zn). Only blue-green substances, not gold pigments, remain in some of the hanging boards in which brass components were detected. A reproduction experiment was conducted because the pigments of the brass component were not recorded in the literature and were not currently used as Dancheong pigments. In the reproduction experiment, it was difficult to confirm the application and use of brass pigments due to the limitations of materials, but it is judged that research on the timing and method of using brass pigments is needed in the future.
https://doi.org/10.22755/kjchs.2023.56.3.118 인용 PDF

Engraved Character Recognition of Automotive Airbag Part using Template Matching (템플릿 매칭을 이용한 자동차 에어백 부품의 각인 문자 인식)

Kim, Dong-Hyun;Koo, Bong-Geun;Lee, Hae-Yeoun
- Proceedings of the Korea Information Processing Society Conference
- /
- 2015.04a
- /
- pp.859-861
- /
- 2015
생산 기술이 발전함에 따라 제품의 생산량이 증가하고 컴퓨터 비전을 통한 제품의 양/불 판단 기술의 필요성이 증가하고 있다. 제품의 양/불 판단은 그 정확도가 중요하며, 동시에 빠른 검사를 위한 신속성이 요구된다. 기존 연구들에서 다양한 금속성 제품에 대한 양/불 판단과 각인된 글자에 대한 양/불 판단을 수행하는 연구가 지속되어 왔으나 자동차 에어백 부품 중 하나인 Upper Housing의 양/불을 판단하는 알고리즘은 부재하다. 본 논문에서는 Upper Housing에 대해 각인 문자의 양/불을 판정하는 알고리즘을 제안한다. 먼저 영상에서 기준점이 되는 원을 찾는 것부터 시작하여, 기준점을 기반으로 특정 각도로 회전시켜 미리 수집한 글자 이미지와의 템플릿 매칭을 통해 글자가 제대로 각인 되었는지를 판단한다. 실험에서는 에어백 부품에 대한 검사 장치에서 촬영한 동영상에 대하여 제안한 알고리즘을 적용하였고, 그 결과 높은 정확도로 글자를 검출할 수 있음을 확인하였다.
https://doi.org/10.3745/PKIPS.y2015m04a.859 인용 PDF

Text line extraction based on filtering and peak detection (필터링 및 피크검출을 이용한 텍스트 추출)

Jin, Bora;Cho, Nam-Ik
- Proceedings of the Korean Society of Broadcast Engineers Conference
- /
- 2013.11a
- /
- pp.41-42
- /
- 2013
본 논문에서는 문서 영상 처리의 중요한 전처리 과정인 텍스트 라인 추출을 위하여 가우시안 필터링 및 피크 검출을 이용하는 방법을 제안한다. 이는 문서 영상 내의 글자 영역의 픽셀 강도와 텍스트 라인 사이의 간격에 해당하는 강도의 차이로 인해 문서 영상의 각 열마다 높은 피크와 낮은 피크가 번갈아 가며 나타나는 것에 기반으로, 제안하는 알고리즘은 필터 스케일 추정, 필터량 및 피크 검출, 라인 성분 그룹화의 세 단계로 구성된다. 필터 스케일 추정 단계에서는 여러 초기 값으로 필터링하여 피크 차이 간의 히스토그램을 만듦으로써 글자 크기를 대략적으로 예축하며, 필터링 및 피크 검출 단계에서 앞서 예측된 스케일의 가우시안 필터를 이용하여 필터링 한 후, 각각의 열마다 피크를 검출한다. 마지막으로 라인 성분 그룹화를 통하여 검출된 피크를 서로 연결하여 하나의 텍스트 라인을 구성하는 성분들로 그룹화시켜 텍스트 라인을 추출한다. 실험 결과를 통하여, 제안하는 알고리즘은 이진화 과정을 거치지 않음으로써 균일하지 못한 조명환경 등으로 이진화 성능이 좋지 못할 경우에도 텍스트 라인을 추출할 수 있으며, 텍스트 라인 간격이 인정하지 않고 휘어진 라인을 포함하는 경우에도 적용할 수 있음을 확인 할 수 있다.
PDF

Remote Drawing Technology Based on Motion Trajectories Analysis (움직임 궤적 분석 기반의 원거리 판서 기술)

Leem, Seung-min;Jeong, Hyeon-seok;Kim, Sung-young
- The Journal of Korea Institute of Information, Electronics, and Communication Technology
- /
- v.9 no.2
- /
- pp.229-236
- /
- 2016
In this paper, we suggest new technology that can draw characters at a long distance by tracking a hand and analysing the trajectories of hand positions. It's difficult to recognize the shape of a character without discriminating effective strokes from all drawing strokes. We detect end points from input trajectories of a syllable with camera system and localize strokes by using detected end points. Then we classify the patterns of the extracted strokes into eight classes and finally into two categories of stroke that is part of syllable and not. We only draw the strokes that are parts of syllable and can display a character. We can get 88.3% in classification accuracy of stroke patterns and 91.1% in stroke type classification.
https://doi.org/10.17661/jkiiect.2016.9.2.229 인용 PDF KSCI

Readability Enhancement Algorithm for Patterned Retarder based Stereoscopic 3D display (Patterned Retarder 방식 입체 디스플레이에서의 가독성 향상 기법)

Lee, Hui Jung;Song, Byung Cheol
- Journal of the Institute of Electronics and Information Engineers
- /
- v.50 no.5
- /
- pp.175-182
- /
- 2013
This paper proposes a readability enhancement filter for Patterned Retarder (PR) display. In general, when some texts in stereoscopic images are shown on PR display, their readability tends to be lowered. In order to overcome this problem, we present a readability enhancement algorithm which consists of readability filtering stage and post-processing stage for specific characters. First, each input stereo image is divided into an odd line image and an even line image. Then, they are independently up-scaled vertically by using Lanczos filter. Next, two up-scaled line images are averaged considering vertical phase difference. In post-processing stage, two specific characters which are normally difficult to read on PR display are detected, and they are filtered for additional readability enhancement. Here, this additional filtering is based on a specific brightness adjustment, and is applied only for two characters. The experiment results show that the proposed method achieves significant improvement in terms of readability in comparison with the previous scheme.
https://doi.org/10.5573/ieek.2013.50.5.175 인용 PDF KSCI

Search Result 40, Processing Time 0.027 seconds

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

Detail Search

Image Search (β)