Search | Korea Science

Feature based Text Watermarking in Digital Binary Image (이진 문서 영상에서의 특징 기반 텍스트 워터마킹)

공영민;추현곤;최종욱;김희율
- Proceedings of the IEEK Conference
- /
- 2002.06d
- /
- pp.359-362
- /
- 2002
In this paper, we propose a new feature-based text watermarking for the binary text image. The structure of specific characters from preprocessed text image are modified to embed watermark. Watermark message are embedded and detected by the following method; Hole line disconnect using the connectivity of the character containing a hole, Center line shift using the hole area and Differential encoding using difference of flippable score points. Experimental results show that the proposed method is robust to rotation and scaling distortion.
PDF

A Comparative Study on OCR using Super-Resolution for Small Fonts

Cho, Wooyeong;Kwon, Juwon;Kwon, Soonchu;Yoo, Jisang
- International journal of advanced smart convergence
- /
- v.8 no.3
- /
- pp.95-101
- /
- 2019
Recently, there have been many issues related to text recognition using Tesseract. One of these issues is that the text recognition accuracy is significantly lower for smaller fonts. Tesseract extracts text by creating an outline with direction in the image. By searching the Tesseract database, template matching with characters with similar feature points is used to select the character with the lowest error. Because of the poor text extraction, the recognition accuracy is lowerd. In this paper, we compared text recognition accuracy after applying various super-resolution methods to smaller text images and experimented with how the recognition accuracy varies for various image size. In order to recognize small Korean text images, we have used super-resolution algorithms based on deep learning models such as SRCNN, ESRCNN, DSRCNN, and DCSCN. The dataset for training and testing consisted of Korean-based scanned images. The images was resized from 0.5 times to 0.8 times with 12pt font size. The experiment was performed on x0.5 resized images, and the experimental result showed that DCSCN super-resolution is the most efficient method to reduce precision error rate by 7.8%, and reduce the recall error rate by 8.4%. The experimental results have demonstrated that the accuracy of text recognition for smaller Korean fonts can be improved by adding super-resolution methods to the OCR preprocessing module.
https://doi.org/10.7236/IJASC.2019.8.3.95 인용 PDF KSCI

Qualitative Study on Group Decision Making with Synchronous Text Communication Medium (동시적 텍스트 기반 매체를 이용한 집단의사결정에 관한 질적 연구)

Park Sanghyuk;Cho Namjae
- Journal of Information Technology Applications and Management
- /
- v.11 no.4
- /
- pp.1-23
- /
- 2004
This study identifies communication patterns of groups using synchronous text communication medium for their group decision-making, and examines how these patterns are associated with creative solutions to problems. Our research suggests that certain communication behavior of groups, when appropriately organized, can be of help in enhancing creative production of outcomes. A qualitative study was conducted on communication patterns based on an analysis of text-based electronic conversation protocols. Specifically this research tried to overcome existing studies on electronic groups by focusing on interactive process of communication among participants. The major study conclusion; are: (1) The production of creative outcome may depend on the process or sequence of discussion among group members with synchronous text communication medium. That is, proper interactive responses and appropriate control of the discussion process are essential to obtain a high level of performance. (2) It is importantto make discuss rules based on meta-cognitive and interactive protocols in the early stage. Explicit rules relating to internal group processes as well as communication medium use are even more important to groups with electronic communication medium than face-to-face groups.
PDF

Citation-based Article Summarization using a Combination of Lexical Text Similarities: Evaluation with Computational Linguistics Literature Summarization Datasets

Kang, In-Su
- Journal of the Korea Society of Computer and Information
- /
- v.24 no.7
- /
- pp.31-37
- /
- 2019
Citation-based article summarization is to create a shortened text for an academic article, reflecting the content of citing sentences which contain other's thoughts about the target article to be summarized. To deal with the problem, this study introduces an extractive summarization method based on calculating a linear combination of various sentence salience scores, which represent the degrees to which a candidate sentence reflects the content of author's abstract text, reader's citing text, and the target article to be summarized. In the current study, salience scores are obtained by computing surface-level textual similarities. Experiments using CL-SciSumm datasets show that the proposed method parallels or outperforms the previous approaches in ROUGE evaluations against SciSumm-2017 human summaries and SciSumm-2016/2017 community summaries.
https://doi.org/10.9708/jksci.2019.24.07.031 인용 PDF KSCI HTML

Usability Evaluation of Text-based Search and Visual Search of a Multidisciplinary Library Database (상용 학술데이터베이스의 텍스트 기반 검색과 비주얼검색의 사용성에 관한 연구)

Kim, Jong-Ae
- Journal of the Korean Society for information Management
- /
- v.26 no.3
- /
- pp.111-129
- /
- 2009
This study examined the usability of text-based search and visual search of a large multidisciplinary library database to provide an empirical analysis of the acceptability of visual systems in the information retrieval environment. It also examined if there are differences in the usability assessment based on experimental order. The results indicated that the text-based search supported users' search behaviors more efficiently than the visual search. Also the text-based search was rated higher than the visual search in terms of user perceptions of four usability factors.
https://doi.org/10.3743/KOSIM.2009.26.3.111 인용 PDF

Mobile Phone Camera Based Scene Text Detection Using Edge and Color Quantization (에지 및 컬러 양자화를 이용한 모바일 폰 카메라 기반장면 텍스트 검출)

Park, Jong-Cheon;Lee, Keun-Wang
- Journal of the Korea Academia-Industrial cooperation Society
- /
- v.11 no.3
- /
- pp.847-852
- /
- 2010
Text in natural images has a various and important feature of image. Therefore, to detect text and extraction of text, recognizing it is a studied as an important research area. Lately, many applications of various fields is being developed based on mobile phone camera technology. Detecting edge component form gray-scale image and detect an boundary of text regions by local standard deviation and get an connected components using Euclidean distance of RGB color space. Labeling the detected edges and connected component and get bounding boxes each regions. Candidate of text achieved with heuristic rule of text. Detected candidate text regions was merged for generation for one candidate text region, then text region detected with verifying candidate text region using ectilarity characterization of adjacency and ectilarity between candidate text regions. Experctental results, We improved text region detection rate using completentary of edge and color connected component.
https://doi.org/10.5762/KAIS.2010.11.3.847 인용 PDF KSCI

Control of Duration Model Parameters in HMM-based Korean Speech Synthesis (HMM 기반의 한국어 음성합성에서 지속시간 모델 파라미터 제어)

Kim, Il-Hwan;Bae, Keun-Sung
- Speech Sciences
- /
- v.15 no.4
- /
- pp.97-105
- /
- 2008
Nowadays an HMM-based text-to-speech system (HTS) has been very widely studied because it needs less memory and low computation complexity and is suitable for embedded systems in comparison with a corpus-based unit concatenation text-to-speech one. It also has the advantage that voice characteristics and the speaking rate of the synthetic speech can be converted easily by modifying HMM parameters appropriately. We implemented an HMM-based Korean text-to-speech system using a small size Korean speech DB and proposes a method to increase the naturalness of the synthetic speech by controlling duration model parameters in the HMM-based Korean text-to speech system. We performed a paired comparison test to verify that theses techniques are effective. The test result with the preference scores of 73.8% has shown the improvement of the naturalness of the synthetic speech through controlling the duration model parameters.
PDF

Representation of Texts into String Vectors for Text Categorization

Jo, Tae-Ho
- Journal of Computing Science and Engineering
- /
- v.4 no.2
- /
- pp.110-127
- /
- 2010
In this study, we propose a method for encoding documents into string vectors, instead of numerical vectors. A traditional approach to text categorization usually requires encoding documents into numerical vectors. The usual method of encoding documents therefore causes two main problems: huge dimensionality and sparse distribution. In this study, we modify or create machine learning-based approaches to text categorization, where string vectors are received as input vectors, instead of numerical vectors. As a result, we can improve text categorization performance by avoiding these two problems.
https://doi.org/10.5626/JCSE.2010.4.2.110 인용 PDF

Improved Spam Filter via Handling of Text Embedded Image E-mail

Youn, Seongwook;Cho, Hyun-Chong
- Journal of Electrical Engineering and Technology
- /
- v.10 no.1
- /
- pp.401-407
- /
- 2015
The increase of image spam, a kind of spam in which the text message is embedded into attached image to defeat spam filtering technique, is a major problem of the current e-mail system. For nearly a decade, content based filtering using text classification or machine learning has been a major trend of anti-spam filtering system. Recently, spammers try to defeat anti-spam filter by many techniques. Text embedding into attached image is one of them. We proposed an ontology spam filters. However, the proposed system handles only text e-mail and the percentage of attached images is increasing sharply. The contribution of the paper is that we add image e-mail handling capability into the anti-spam filtering system keeping the advantages of the previous text based spam e-mail filtering system. Also, the proposed system gives a low false negative value, which means that user's valuable e-mail is rarely regarded as a spam e-mail.
https://doi.org/10.5370/JEET.2015.10.1.401 인용 PDF KSCI KPUBS HTML

Stroke Width-Based Contrast Feature for Document Image Binarization

Van, Le Thi Khue;Lee, Gueesang
- Journal of Information Processing Systems
- /
- v.10 no.1
- /
- pp.55-68
- /
- 2014
Automatic segmentation of foreground text from the background in degraded document images is very much essential for the smooth reading of the document content and recognition tasks by machine. In this paper, we present a novel approach to the binarization of degraded document images. The proposed method uses a new local contrast feature extracted based on the stroke width of text. First, a pre-processing method is carried out for noise removal. Text boundary detection is then performed on the image constructed from the contrast feature. Then local estimation follows to extract text from the background. Finally, a refinement procedure is applied to the binarized image as a post-processing step to improve the quality of the final results. Experiments and comparisons of extracting text from degraded handwriting and machine-printed document image against some well-known binarization algorithms demonstrate the effectiveness of the proposed method.
https://doi.org/10.3745/JIPS.2014.10.1.055 인용 PDF KSCI KPUBS HTML

Search Result 3,954, Processing Time 0.028 seconds

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

Detail Search

Image Search (β)