<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE ArticleSet PUBLIC "-//NLM//DTD PubMed 2.7//EN" "https://dtd.nlm.nih.gov/ncbi/pubmed/in/PubMed.dtd">
<ArticleSet>
<Article>
<Journal>
				<PublisherName>West Asia Organization for Cancer Prevention (WAOCP), APOCP's West Asia Chapter.</PublisherName>
				<JournalTitle>Asian Pacific Journal of Cancer Prevention</JournalTitle>
				<Issn>1513-7368</Issn>
				<Volume>18</Volume>
				<Issue>5</Issue>
				<PubDate PubStatus="epublish">
					<Year>2017</Year>
					<Month>05</Month>
					<Day>01</Day>
				</PubDate>
			</Journal>
<ArticleTitle>Modified Bat Algorithm for Feature Selection with the Wisconsin Diagnosis Breast Cancer (WDBC) Dataset</ArticleTitle>
<VernacularTitle></VernacularTitle>
			<FirstPage>1257</FirstPage>
			<LastPage>1264</LastPage>
			<ELocationID EIdType="pii">46478</ELocationID>
			
<ELocationID EIdType="doi">10.22034/APJCP.2017.18.5.1257</ELocationID>
			
			<Language>EN</Language>
<AuthorList>
<Author>
					<FirstName>Suganthi </FirstName>
					<LastName>Jeyasingh</LastName>
<Affiliation>Department of Computer Science and Engineering, Raja College of Engineering and Technology, Madurai, Tamilnadu, India	.</Affiliation>

</Author>
<Author>
					<FirstName>Malathy </FirstName>
					<LastName>Veluchamy</LastName>
<Affiliation>Department of Electrical and
Electronics Engineering, Anna University Regional Centre, Madurai, Tamilnadu, India.</Affiliation>

</Author>
</AuthorList>
				<PublicationType>Journal Article</PublicationType>
			<History>
				<PubDate PubStatus="received">
					<Year>2016</Year>
					<Month>11</Month>
					<Day>25</Day>
				</PubDate>
			</History>
		<Abstract> &lt;br /&gt; &lt;span style=&quot;font-size: small;&quot;&gt;Early diagnosis of breast cancer is essential to save lives of patients. Usually, medical datasets include a large variety of data that can lead to confusion during diagnosis. The Knowledge Discovery on Database (KDD) process helps to &lt;/span&gt;&lt;span style=&quot;font-family: Times New Roman,Times New Roman; font-size: small;&quot;&gt;&lt;span style=&quot;font-family: Times New Roman,Times New Roman; font-size: small;&quot;&gt;improve efficiency. It requires elimination of inappropriate and repeated data from the dataset before final diagnosis. &lt;/span&gt;&lt;/span&gt;&lt;span style=&quot;font-size: small;&quot;&gt;This can be done using any of the feature selection algorithms available in data mining. Feature selection is considered &lt;/span&gt;&lt;span style=&quot;font-family: Times New Roman,Times New Roman; font-size: small;&quot;&gt;&lt;span style=&quot;font-family: Times New Roman,Times New Roman; font-size: small;&quot;&gt;as a vital step to increase the classification accuracy. This paper proposes a Modified Bat Algorithm (MBA) for feature selection to eliminate irrelevant features from an original dataset. The Bat algorithm was modified using simple random &lt;/span&gt;&lt;/span&gt;&lt;span style=&quot;font-size: small;&quot;&gt;sampling to select the random instances from the dataset. Ranking was with the global best features to recognize the &lt;/span&gt;&lt;span style=&quot;font-family: Times New Roman,Times New Roman; font-size: small;&quot;&gt;&lt;span style=&quot;font-family: Times New Roman,Times New Roman; font-size: small;&quot;&gt;predominant features available in the dataset. The selected features are used to train a Random Forest (RF) classification algorithm. The MBA feature selection algorithm enhanced the classification accuracy of RF in identifying the occurrence &lt;/span&gt;&lt;/span&gt;&lt;span style=&quot;font-size: small;&quot;&gt;of breast cancer. The Wisconsin Diagnosis Breast Cancer Dataset (WDBC) was used for estimating the performance analysis of the proposed MBA feature selection algorithm. The proposed algorithm achieved better performance in &lt;/span&gt;&lt;span style=&quot;font-family: Times New Roman,Times New Roman; font-size: small;&quot;&gt;&lt;span style=&quot;font-family: Times New Roman,Times New Roman; font-size: small;&quot;&gt;terms of Kappa statistic, Mathew’s Correlation Coefficient, Precision, F-measure, Recall, Mean Absolute Error (MAE), &lt;/span&gt;&lt;/span&gt;&lt;span style=&quot;font-size: small;&quot;&gt;Root Mean Square Error (RMSE), Relative Absolute Error (RAE) and Root Relative Squared Error (RRSE). &lt;/span&gt;</Abstract>
		<ObjectList>
			<Object Type="keyword">
			<Param Name="value">breast cancer</Param>
			</Object>
			<Object Type="keyword">
			<Param Name="value">Wisconsin Diagnosis Breast Cancer (WDBC) dataset</Param>
			</Object>
			<Object Type="keyword">
			<Param Name="value">Modified Bat algorithm</Param>
			</Object>
		</ObjectList>
<ArchiveCopySource DocType="pdf">https://journal.waocp.org/article_46478_cfcbe23bed573f17d12d29860eedbdba.pdf</ArchiveCopySource>
</Article>
</ArticleSet>
