<?xml version='1.0' encoding='utf-8'?>
<?xml-stylesheet type="text/xsl" href="/v2/static/oai2.xsl"?>
<OAI-PMH xmlns="http://www.openarchives.org/OAI/2.0/" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/ http://www.openarchives.org/OAI/2.0/OAI-PMH.xsd">
  <responseDate>2026-10-06T07:27:55Z</responseDate>
  <request identifier="oai:figshare.com:article/34018134" metadataPrefix="oai_dc" verb="GetRecord">https://api.figshare.com/v2/oai</request>
  <GetRecord>
    <record>
      <header>
        <identifier>oai:figshare.com:article/34018134</identifier>
        <datestamp>2026-09-28T18:12:22Z</datestamp>
        <setSpec>category_26407</setSpec>
        <setSpec>item_type_12</setSpec>
        <setSpec>month_year_09_2026</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance"  xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xmlns:dc="http://purl.org/dc/elements/1.1/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:title>Same Bits, Different Verdict: The Decoder in Speech Deepfake Detection</dc:title>
          <dc:creator>GANG SHI (25132929)</dc:creator>
          <dc:subject>Signal processing</dc:subject>
          <dc:subject>speech deepfake detection</dc:subject>
          <dc:subject>audio deepfake detection</dc:subject>
          <dc:subject>anti-spoofing</dc:subject>
          <dc:subject>speech coding</dc:subject>
          <dc:subject>Opus</dc:subject>
          <dc:subject>decoder-side processing</dc:subject>
          <dc:subject>receiver-side rendering</dc:subject>
          <dc:subject>threshold calibration</dc:subject>
          <dc:subject>domain shift</dc:subject>
          <dc:subject>cover-source mismatch</dc:subject>
          <dc:description>&lt;p dir="ltr"&gt;Speech deepfake detectors consume decoded waveforms, but for coded speech the bitstream is what was transmitted; which rendering of it the detector sees belongs to the receiver. Decoding a bit-identical Opus bitstream with libopus 1.6.1’s stronger optional decoder-side post-filter enabled rather than disabled shifts class-mean detector scores toward the synthetic class in 20 of 20 detector–class–corpus cells, with no packet loss and no change to a transmitted byte. At a threshold transferred from the unenhanced rendering at a 1% false-alarm target, false alarms rise in all eight cells; one public detector goes from 1.00% to 3.80% while its equal-error rate moves only from 0.30% to 0.40%. The shift is not a monotone rescaling that recalibration absorbs: items cross the threshold in both directions, and equal-error rate moves in every cell. On held-out data a stale threshold multiplies false alarms by 1.7 to 3.8. Refitting recovers the target rate, but no corpus, protocol or file header records that the configuration changed.&lt;/p&gt;&lt;p dir="ltr"&gt;&lt;br&gt;This work has been submitted to the IEEE for possible publication. Copyright may be transferred without notice, after which this version may no longer be accessible.&lt;/p&gt;</dc:description>
          <dc:date>2026-09-28T18:12:22Z</dc:date>
          <dc:type>Text</dc:type>
          <dc:type>Preprint</dc:type>
          <dc:identifier>10.6084/m9.figshare.34018134.v1</dc:identifier>
          <dc:relation>https://figshare.com/articles/preprint/Same_Bits_Different_Verdict_The_Decoder_in_Speech_Deepfake_Detection/34018134</dc:relation>
          <dc:rights>CC BY 4.0</dc:rights>
        </oai_dc:dc>
      </metadata>
    </record>
  </GetRecord>
</OAI-PMH>
