<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE ArticleSet PUBLIC "-//NLM//DTD PubMed 2.7//EN" "https://dtd.nlm.nih.gov/ncbi/pubmed/in/PubMed.dtd">
<ArticleSet>
<Article>
<Journal>
<PublisherName>OICC Press</PublisherName>
<JournalTitle>Majlesi Journal of Electrical Engineering</JournalTitle>
<Issn>2345-3796</Issn>
<Volume>3</Volume>
<Issue>3</Issue>
<PubDate PubStatus="epublish">
<Year>2024</Year>
<Month>02</Month>
<Day>28</Day>
</PubDate>
</Journal>
<ArticleTitle>A New Segmentation Method for Persian/Arabic OCR Based on Baseline Processing</ArticleTitle>
<VernacularTitle></VernacularTitle>
<FirstPage></FirstPage>
<LastPage></LastPage>
<ELocationID EIdType="doi">10.1234/mjee.v3i3.155</ELocationID>
<Language>EN</Language>
<AuthorList>
<Author>
<FirstName>Mahboubeh</FirstName>
<LastName>Shamsi</LastName>
<Affiliation>Islamic Azad University, Bardsir Branch</Affiliation>
<Identifier Source="ORCID"></Identifier>
</Author>
<Author>
<FirstName>Reza</FirstName>
<LastName>Rasouli</LastName>
<Affiliation>Azad Islamic University Bardsir</Affiliation>
<Identifier Source="ORCID"></Identifier>
</Author>
<Author>
<FirstName>Soudeh</FirstName>
<LastName>Shadravan</LastName>
<Affiliation>Unknown</Affiliation>
<Identifier Source="ORCID"></Identifier>
</Author>
</AuthorList>
<PublicationType>Journal Article</PublicationType>
<History>
<PubDate PubStatus="received">
<Year>2024</Year>
<Month>02</Month>
<Day>28</Day>
</PubDate>
</History>
<Abstract>One of the most important stages in Character Recognition Systems is âSegmentationâ, because any mistake will affect to all other tasks, especially to character recognition. This operation is more complex in Persian/Arabic writing than other Latin writing like English, and there has been an ongoing research on it. Other algorithms, that has been used as base as proposed algorithm, show 85% accuracy. In this paper, a new improved method has been presented by analyzing the visual features of the Persian/Arabic language. The proposed algorithm is able to segment existing fonts up to 98.5% accuracy or even 100% on some cases. The remaining error could be refined by applying a good character recognition technique and a precise vocabulary.</Abstract>
<ObjectList>
<Object Type="keyword">
<Param Name="value">Azad Islamic University Bardsir</Param>
</Object>
<Object Type="keyword">
<Param Name="value">fa</Param>
</Object>
<Object Type="keyword">
<Param Name="value">Image processing</Param>
</Object>
<Object Type="keyword">
<Param Name="value">Persian OCR</Param>
</Object>
<Object Type="keyword">
<Param Name="value">Segmentation. recognition</Param>
</Object>
<Object Type="keyword">
<Param Name="value">smoothing</Param>
</Object>
</ObjectList>
</Article>
</ArticleSet>