Data · dataset · 2026
Same Bits, Different Verdict: The Decoder in Speech Deepfake Detection
Listed in ZivaHub
<p dir="ltr">Speech deepfake detectors consume decoded waveforms, but for coded speech the bitstream is what was transmitted; which rendering of it the detector sees belongs to the receiver.
Description
Decoding a bit-identical Opus bitstream with libopus 1.6.1’s stronger optional decoder-side post-filter enabled rather than disabled shifts class-mean detector scores toward the synthetic class in 20 of 20 detector–class–corpus cells, with no packet loss and no change to a transmitted byte.
At a threshold transferred from the unenhanced rendering at a 1% false-alarm target, false alarms rise in all eight cells; one public detector goes from 1.00% to 3.80% while its equal-error rate moves only from 0.30% to 0.40%. The shift is not a monotone rescaling that recalibration absorbs: items cross the threshold in both directions, and equal-error rate moves in every cell. On held-out data a stale threshold multiplies false alarms by 1.7 to 3.8.
Read the rest (1 more)
Refitting recovers the target rate, but no corpus, protocol or file header records that the configuration changed.</p><p dir="ltr"><br>This work has been submitted to the IEEE for possible publication. Copyright may be transferred without notice, after which this version may no longer be accessible.</p>
Links
Where it is published
- DOI doi.org/10.6084/m9.figshare.34018134.v1 ↗
DOI / persistent id · from zivahub uct ac za
Catalogue records · 1
- OAI-PMH record api.figshare.com/v2/oai?verb=GetRecord&metadataPrefix=oai_dc&identifier=oai%3Af… ↗
metadata API · from zivahub uct ac za
Topics
- From keywords
- Earth & Environmental Science · Signal processing
- Inferred from text
- Audio 65%
Provenance · 1 source records, 8 field assertions
| Source | Key | Last seen | Raw |
|---|---|---|---|
| ZivaHub | oai:figshare.com:article/34018134 | 5 d ago | JSON v1 |
| Field | Assertion | Extractor | Evidence |
|---|---|---|---|
| access_level | source · zivahub uct ac za | connector:zivahub_uct_ac_za@1.0.0 | |
| concepts[field].anzsrc:field:400607 | mapping · zivahub uct ac za | vocabulary-mapper@1.0.0 | keywords['Signal processing'] |
| concepts[field].local:field:earth-environmental | mapping · zivahub uct ac za | connector:zivahub_uct_ac_za@1.0.0 | |
| concepts[modality].local:modality:audio | enrichment · zivahub uct ac za | keyword-concept-rules@1.0.0 | title+description (65%) |
| description | source · zivahub uct ac za | connector:zivahub_uct_ac_za@1.0.0 | /metadata/dc/description |
| license | source · zivahub uct ac za | connector:zivahub_uct_ac_za@1.0.0 | /metadata/dc/rights |
| publication_date | source · zivahub uct ac za | connector:zivahub_uct_ac_za@1.0.0 | |
| title | source · zivahub uct ac za | connector:zivahub_uct_ac_za@1.0.0 | /metadata/dc/title |