AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (1.7 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Publishing Language: Chinese

Comparative Study of AI-Synthesized Speech and Natural Speech

Fangling LIAO1Manqing CHEN1Shengxiang CHEN1Yuhang GUO1Yingcang YANG1( )Fan MU2
Department of Forensic Science and Technology, Guizhou Police College, Guiyang 550005, China
Criminal Investigation Division, Guiyang Municipal Public Security Bureau, Guiyang 550000, China
Show Author Information

Abstract

Objective

With the rapid development of artificial intelligence (AI)-synthesized speech technology, its detectability in forensic appraisal has become a key issue. This study systematically analyzes the differential features between AI-synthesized speech and natural speech through a two-dimensional comparative study of auditory perception and acoustic quantification, thereby providing an effective reference for the identification, prevention, inspection, and appraisal of synthesized speech in judicial practice.

Methods

In the auditory test, the Likert-type Scale was used to rate the consistency between natural speech and synthesized speech. Acoustic tests were conducted by extracting feature parameters such as fundamental frequency, formants, sound intensity, and duration using the Praat speech analysis software. Combined with SPSS 27 statistical analysis software, a paired-sample t-test was conducted to quantify the differences between natural speech and synthesized speech.

Results

Compared with natural speech, AI-synthesized speech exhibited poorer performance in terms of auditory features such as monosyllabic integrity, retroflex features, stress, speech rate, and fluency. Statistical analysis of the acoustic testing showed that there were significant differences in fundamental frequency and formants, while sound intensity and duration showed no significant differences.

Conclusion

The combined application of “human ear preliminary screening” and acoustic quantification two-dimensional testing techniques in forensic appraisal can effectively distinguish AI-synthesized speech from natural speech, providing technical support for the inspection and appraisal of AI-synthesized speech.

CLC number: DF794.1 Document code: A Article ID: 1671-2072-(2026)2-0038-08

References

【1】
【1】
 
 
Chinese Journal of Forensic Sciences
Pages 38-45

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
LIAO F, CHEN M, CHEN S, et al. Comparative Study of AI-Synthesized Speech and Natural Speech. Chinese Journal of Forensic Sciences, 2026, 2026(2): 38-45. https://doi.org/10.3969/j.issn.1671-2072.2026.02.005

160

Views

3

Downloads

0

Crossref

Received: 24 April 2025
Published: 15 March 2026
© 2026 Editorial Office of Chinese Journal of Forensic Sciences