<?xml version="1.1" encoding="utf-8"?>
<article xsi:noNamespaceSchemaLocation="http://jats.nlm.nih.gov/publishing/1.1/xsd/JATS-journalpublishing1-mathml3.xsd" dtd-version="1.1" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance"><front><journal-meta><journal-id journal-id-type="publisher-id">JERA</journal-id><journal-title-group><journal-title>Journal of Electronic Research and Application</journal-title></journal-title-group><issn>2208-3502</issn><eissn>2208-3510</eissn><publisher><publisher-name>Bio-Byword Scientific Publishing Pty. Ltd.</publisher-name></publisher></journal-meta><article-meta><article-id pub-id-type="doi">10.26689/jera.v10i5.15278</article-id><article-categories><subj-group subj-group-type="heading"><subject>Article</subject></subj-group></article-categories><title>Development and Application of an AI Popular Science Digital Human System Based on Local Private Large Model Technology</title><url>https://artdesignp.com/journal/JERA/10/5/10.26689/jera.v10i5.15278</url><author>WangYongqiang</author><pub-date pub-type="publication-year"><year>2026</year></pub-date><volume>10</volume><issue>5</issue><history><date date-type="pub"><published-time>2026-06-29</published-time></date></history><abstract>With the continuous development of artificial intelligence technology, digital human technology offers extensive applications in science popularization education. This paper designs an AI popular science digital human system using local private large model technology, which integrates key technologies including speech recognition, natural language processing, speech synthesis, and digital human driving to enable intelligent interactive Q&amp;amp;A with users. The system adopts a locally deployed architecture, fine-tuned based on the Qwen large language model, and combines SenseVoice speech recognition, CosyVoice speech synthesis, and the LiveTalking digital human driving engine to build a complete popular science interaction process. The system has been put into practical use in scenarios such as science and technology festivals in primary and secondary schools and science and technology exhibition halls, which effectively improves the fun and interactivity of science popularization education and provides a new solution for cultivating scientific literacy among teenagers.</abstract><keywords/></article-meta></front><body/><back><ref-list><ref id="B1" content-type="article"><label>1</label><element-citation publication-type="journal"><p>National Informatization Planning Working Group, 2021, 14th Five-Year National Informatization Plan.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B2" content-type="article"><label>2</label><element-citation publication-type="journal"><p>Brown T, Mann B, Ryder N, et al., 2020, Language Models are Few-Shot Learners, Advances in Neural Information Processing Systems, 1877–1901.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B3" content-type="article"><label>3</label><element-citation publication-type="journal"><p>Alibaba Cloud, 2023, Qwen Large Language Model Technical Report.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B4" content-type="article"><label>4</label><element-citation publication-type="journal"><p>Alibaba Group, 2024, CosyVoice Speech Synthesis Technical Documentation.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B5" content-type="article"><label>5</label><element-citation publication-type="journal"><p>Alibaba Group, 2024, SenseVoice Speech Recognition White Paper.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B6" content-type="article"><label>6</label><element-citation publication-type="journal"><p>OpenAI, 2023, GPT-4 technical report, arXiv preprint arXiv:2303.08774.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B7" content-type="article"><label>7</label><element-citation publication-type="journal"><p>Kwon W, Li Z, Zhuang S, et al., 2023, Efficient Memory Management for Large Language Model Serving with Pagedattention, SOSP, 611–626.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B8" content-type="article"><label>8</label><element-citation publication-type="journal"><p>LiveTalking Open Source Community, 2024, LiveTalking Digital Human Driving Engine Documentation.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B9" content-type="article"><label>9</label><element-citation publication-type="journal"><p>Prajwal K, Mukhopadhyay R, Namboodiri V, et al., 2020, A Lip Sync Expert is all you Need for Speech to Lip Generation in the Wild, ACM MM, 484–492.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B10" content-type="article"><label>10</label><element-citation publication-type="journal"><p>Yang F, Zhang W, 2024, Research on Deep Learning-Based Voice-Driven Digital Human Technology. Journal of Computer Research and Development, 61(2): 312–325.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B11" content-type="article"><label>11</label><element-citation publication-type="journal"><p>China Association for Science and Technology, 2021, National Scientific Literacy Action Plan (2021–2035).</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B12" content-type="article"><label>12</label><element-citation publication-type="journal"><p>Ministry of Education, 2018, Education Informatization 2.0 Action Plan.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B13" content-type="article"><label>13</label><element-citation publication-type="journal"><p>Liu Y, Zhao W, 2023, Analysis on Development Status and Trends of Digital Human Technology. Journal of Computer-Aided Design &amp; Computer Graphics, 35(8): 1156–1168.</p><pub-id pub-id-type="doi"/></element-citation></ref></ref-list></back></article>
