<?xml version="1.0" encoding="UTF-8"?>
<doi_batch version="5.3.1" xmlns="http://www.crossref.org/schema/5.3.1" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xmlns:jats="http://www.ncbi.nlm.nih.gov/JATS1" xmlns:ai="http://www.crossref.org/AccessIndicators.xsd" xsi:schemaLocation="http://www.crossref.org/schema/5.3.1 http://www.crossref.org/schema/deposit/crossref5.3.1.xsd">
 <head>
  <doi_batch_id>aspg-3-3041-1791417438</doi_batch_id>
  <timestamp>20261007235718</timestamp>
  <depositor>
   <depositor_name>American Scientific Publishing Group</depositor_name>
   <email_address>admin@americaspg.com</email_address>
  </depositor>
  <registrant>American Scientific Publishing Group</registrant>
 </head>
 <body>
  <journal>
   <journal_metadata language="en">
    <full_title>Fusion: Practice and Applications</full_title>
    <abbrev_title>FPA</abbrev_title>
    <issn media_type="print">2770-0070</issn>
    <issn media_type="electronic">2692-4048</issn>
   </journal_metadata>
   <journal_issue>
    <publication_date media_type="online">
     <year>2024</year>
    </publication_date>
    <journal_volume>
     <volume>16</volume>
    </journal_volume>
    <issue>2</issue>
   </journal_issue>
   <journal_article publication_type="full_text">
    <titles>
     <title>Speaker Identification in Crowd Speech Audio using Convolutional Neural Networks</title>
    </titles>
    <contributors>
     <person_name sequence="first" contributor_role="author">
      <given_name>Ghadeer Qasim</given_name>
      <surname>Ali</surname>
      <affiliations>
       <institution>
        <institution_name>Computer Science Department, College of Science, University of Baghdad</institution_name>
       </institution>
      </affiliations>
     </person_name>
     <person_name sequence="additional" contributor_role="author">
      <given_name>Husam Ali</given_name>
      <surname>Abdulmohsin</surname>
      <affiliations>
       <institution>
        <institution_name>Computer Science Department, College of Science, University of Baghdad</institution_name>
       </institution>
      </affiliations>
     </person_name>
    </contributors>
    <jats:abstract>
     <jats:p>Crowd speaker identification is the most advanced technology in the field of audio identification and personal user experience which researchers have extensively focused on, but still, science hasn’t been able to achieve high results in crowed identification. This work aims to design and implement a novel crowd speech identification method that can identify speakers in a multi speaker environment, (two, three, four and five speakers). This work will be implemented through two phases. The training phase is the Convolutional Neural Network (CNN) training and testing phase. Through this phase, the training will be implemented on data generated via the Combinatorial Cartesian Product approach. This approach uses two primary processes, the Computation of the Cartesian product process and combinatorial selection process. The second phase is the prediction phase. The aim of this phase is to check the CNN trained in the first phase, through testing it on new crowed audios than the data that the CNN was trained on in the first phase, these new crowded audios exist in the Ghadeer-Speech-Crowd-Corpus (GSCC) dataset, which is a new database designed through this work. Compared to the state-of-the-art speaker identification in multi speaker environment approaches, the results are impressive, with a recognition rate of 99.5% for audio with three speakers, 98.5% for music with four speakers, and 96.4% for audio with five speakers.</jats:p>
    </jats:abstract>
    <publication_date media_type="online">
     <year>2024</year>
    </publication_date>
    <pages>
     <first_page>118</first_page>
     <last_page>125</last_page>
    </pages>
    <publisher_item>
     <item_number item_number_type="article-number">3041</item_number>
    </publisher_item>
    <ai:program name="AccessIndicators">
     <ai:license_ref applies_to="vor">https://creativecommons.org/licenses/by/4.0/</ai:license_ref>
    </ai:program>
    <doi_data>
     <doi>10.54216/FPA.160208</doi>
     <resource>https://www.americaspg.com/journal/3/article/3041</resource>
     <collection property="text-mining">
      <item>
       <resource mime_type="application/pdf">https://www.americaspg.com/storage/91736184760.pdf</resource>
      </item>
     </collection>
    </doi_data>
   </journal_article>
  </journal>
 </body>
</doi_batch>
