<?xml version='1.0' encoding='utf-8'?>
<eprints xmlns='http://eprints.org/ep2/data/2.0'>
  <eprint id='https://researchdata.gla.ac.uk/id/eprint/2109'>
    <eprintid>2109</eprintid>
    <rev_number>35</rev_number>
    <documents>
      <document id='https://researchdata.gla.ac.uk/id/document/7923'>
        <docid>7923</docid>
        <rev_number>2</rev_number>
        <files>
          <file id='https://researchdata.gla.ac.uk/id/file/48249'>
            <fileid>48249</fileid>
            <datasetid>document</datasetid>
            <objectid>7923</objectid>
            <filename>Readme_2109.docx</filename>
            <mime_type>application/vnd.openxmlformats-officedocument.wordprocessingml.document</mime_type>
            <hash>daf16ae338a6f40b351c28110d62a4a5</hash>
            <hash_type>MD5</hash_type>
            <filesize>1414270</filesize>
            <mtime>2026-06-29 08:18:28</mtime>
            <url>https://researchdata.gla.ac.uk/2109/1/Readme_2109.docx</url>
          </file>
        </files>
        <eprintid>2109</eprintid>
        <pos>1</pos>
        <placement>1</placement>
        <mime_type>application/vnd.openxmlformats-officedocument.wordprocessingml.document</mime_type>
        <format>Text</format>
        <language>en</language>
        <security>public</security>
        <license>cc_by_4</license>
        <main>Readme_2109.docx</main>
        <content>readme</content>
      </document>
      <document id='https://researchdata.gla.ac.uk/id/document/7924'>
        <docid>7924</docid>
        <rev_number>1</rev_number>
        <files>
          <file id='https://researchdata.gla.ac.uk/id/file/48252'>
            <fileid>48252</fileid>
            <datasetid>document</datasetid>
            <objectid>7924</objectid>
            <filename>indexcodes.txt</filename>
            <mime_type>text/plain</mime_type>
            <hash>a05ea1e8f0c2b59e47f647bcfae10da4</hash>
            <hash_type>MD5</hash_type>
            <filesize>3555</filesize>
            <mtime>2026-06-29 08:18:42</mtime>
            <url>https://researchdata.gla.ac.uk/2109/2/indexcodes.txt</url>
          </file>
        </files>
        <eprintid>2109</eprintid>
        <pos>2</pos>
        <placement>2</placement>
        <mime_type>text/plain</mime_type>
        <format>other</format>
        <formatdesc>Generate index codes conversion from Text to indexcodes</formatdesc>
        <language>en</language>
        <security>public</security>
        <main>indexcodes.txt</main>
        <relation>
          <item>
            <type>http://eprints.org/relation/isVersionOf</type>
            <uri>https://researchdata.gla.ac.uk/id/document/7923</uri>
          </item>
          <item>
            <type>http://eprints.org/relation/isVolatileVersionOf</type>
            <uri>https://researchdata.gla.ac.uk/id/document/7923</uri>
          </item>
          <item>
            <type>http://eprints.org/relation/isIndexCodesVersionOf</type>
            <uri>https://researchdata.gla.ac.uk/id/document/7923</uri>
          </item>
        </relation>
      </document>
    </documents>
    <eprint_status>archive</eprint_status>
    <userid>32966</userid>
    <dir>disk0/00/00/21/09</dir>
    <datestamp>2026-06-29 08:28:20</datestamp>
    <lastmod>2026-06-29 10:36:03</lastmod>
    <status_changed>2026-06-29 08:28:20</status_changed>
    <type>data_collection</type>
    <metadata_visibility>show</metadata_visibility>
    <creators>
      <item>
        <name>
          <family>Ghadban</family>
          <given>Nour</given>
        </name>
        <enlightenid>68735</enlightenid>
      </item>
      <item>
        <name>
          <family>Hameed</family>
          <given>Hira</given>
        </name>
        <enlightenid>64581</enlightenid>
      </item>
      <item>
        <name>
          <family>Abbasi</family>
          <given>Qammer H.</given>
        </name>
        <enlightenid>39488</enlightenid>
        <orcid>0000-0002-7097-9969</orcid>
      </item>
      <item>
        <name>
          <family>Imran</family>
          <given>Muhammad</given>
        </name>
        <enlightenid>37001</enlightenid>
        <orcid>0000-0003-4743-9136</orcid>
      </item>
      <item>
        <name>
          <family>Cooper</family>
          <given>Jon</given>
        </name>
        <enlightenid>7063</enlightenid>
        <orcid>0000-0002-2358-1050</orcid>
      </item>
      <item>
        <name>
          <family>Le Kernec</family>
          <given>Julien</given>
        </name>
        <enlightenid>35927</enlightenid>
        <orcid>0000-0003-2124-6803</orcid>
      </item>
    </creators>
    <uniqueid>glaresearchdata:2025-12-05-2109</uniqueid>
    <title>Synergising Multi-Modal Data: Unveiling Audio-Visual and Radar Insights</title>
    <ispublished>pub</ispublished>
    <divisions>
      <item>30300000</item>
      <item>30303000</item>
      <item>30305000</item>
      <item>30304000</item>
    </divisions>
    <note>The data file associated with this dataset is too large for standard download (~17GB). Please use the request data button to arrange alternative access. This dataset is licensed with a Creative Commons Attribution CC-BY 4.0 license.</note>
    <abstract>The study uses a multimodal dataset comprising synchronized audio, video, and radar recordings collected to evaluate the effectiveness of audio–visual–radar (AVR) fusion in speech-recognition tasks. The dataset contains 800 labeled samples, produced by four fluent but non-native English speakers aged 28–40, each repeating ten phonetically challenging English words twenty times under controlled recording conditions.

Audio was recorded at 44.1 kHz/16-bit, video at 1080p/30 fps, and radar data were acquired using the XeThru X4M03 IR-UWB sensor positioned approximately 2 meters from the speaker to capture speech-related micro-Doppler articulatory signatures. The three modalities were automatically segmented and aligned using pretrained machine-learning models to ensure consistent synchronization across streams.

The purpose of this dataset is to serve as a resource for studying how radar sensing can enhance speech-recognition robustness under noise, occlusion, and other adverse conditions.

All audio–visual recordings contain identifiable human data and will therefore be shared only under controlled access to protect participant privacy.</abstract>
    <date>2025-12-05</date>
    <date_type>published</date_type>
    <publisher>University of Glasgow</publisher>
    <id_number>10.5525/gla.researchdata.2109</id_number>
    <data_type>
      <item>Audio</item>
      <item>Text</item>
      <item>Video</item>
    </data_type>
    <copyright_holders>
      <item>University of Glasgow</item>
    </copyright_holders>
    <funding>
      <item>
        <project_code>307829</project_code>
        <project_name>Quantum-Inspired Imaging for Remote Monitoring of Health &amp; Disease in Community Healthcare</project_name>
        <investigator_name>Jonathan Cooper</investigator_name>
        <funder_name>Engineering and Physical Sciences Research Council (EPSRC)</funder_name>
        <funder_code>EP/T021020/1</funder_code>
        <investigator_dept>ENG - Biomedical Engineering</investigator_dept>
      </item>
    </funding>
    <pending>FALSE</pending>
    <language>English</language>
    <collection_date>
      <date_from>2022-11-29</date_from>
      <date_to>2022-12-19</date_to>
    </collection_date>
    <related_resources>
      <item>
        <url>https://doi.org/10.5525/gla.researchdata.2111</url>
        <type>code</type>
      </item>
    </related_resources>
    <retention_date>2035-12-05</retention_date>
    <retention_action>R</retention_action>
    <dataset_origin>
      <item>file_share</item>
    </dataset_origin>
    <ingest_data>
      <item>
      </item>
      <item>
      </item>
    </ingest_data>
    <archive_data>
      <item>
      </item>
      <item>
      </item>
    </archive_data>
    <ethics_consent_required>TRUE</ethics_consent_required>
    <ethics_application_number>300230076</ethics_application_number>
    <ethics_committee>cose</ethics_committee>
    <request_copy>TRUE</request_copy>
    <repo_link>
      <item>
        <title>Artificial Intelligence and Radar-Enhanced Audio-Visual Speech Recognition for the Next Generation of Communication Technologies</title>
        <link>https://eprints.gla.ac.uk/id/eprint/352488</link>
      </item>
      <item>
        <title>Advancing Speech Recognition for the Hearing Impaired: Key Insights from the Language and Technology Workshop</title>
        <link>https://eprints.gla.ac.uk/id/eprint/348868</link>
      </item>
      <item>
        <title>Unlocking the Future: AI and Radar-Enhanced Audio-Visual Speech Recognition</title>
        <link>https://eprints.gla.ac.uk/id/eprint/345447</link>
      </item>
      <item>
        <title>Noise Reduction in Audio Visual Radar Speech Recognition System</title>
        <link>https://eprints.gla.ac.uk/id/eprint/341571</link>
      </item>
      <item>
        <title>Radar-enhanced multimodal speech recognition under occlusion and noise: methodology and experimental analysis</title>
        <link>https://eprints.gla.ac.uk/id/eprint/386960</link>
      </item>
    </repo_link>
  </eprint>
</eprints>
