<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Transcription on 中文AI术语词典</title><link>https://terms-en.ai-term-hub.com/zh/tags/transcription/</link><description>Recent content in Transcription on 中文AI术语词典</description><generator>Hugo</generator><language>zh-cn</language><lastBuildDate>Sat, 18 Jul 2026 11:44:45 +0000</lastBuildDate><atom:link href="https://terms-en.ai-term-hub.com/zh/tags/transcription/index.xml" rel="self" type="application/rss+xml"/><item><title>说话人日志</title><link>https://terms-en.ai-term-hub.com/zh/terms/speaker_diarization/</link><pubDate>Sat, 18 Jul 2026 11:34:45 +0000</pubDate><guid>https://terms-en.ai-term-hub.com/zh/terms/speaker_diarization/</guid><description>&lt;h2 id="definition">Definition&lt;/h2>
&lt;p>说话人日志是将音频流根据说话人身份划分为同质片段的任务。它结合了说话人切换检测和说话人聚类技术，旨在为音频中的每一段语音打上说话人身份的标签，实现“谁说了什么”的完整记录。&lt;/p>
&lt;h3 id="summary">Summary&lt;/h3>
&lt;p>确定音频录音中“谁在何时说话”的过程。&lt;/p>
&lt;h2 id="key-concepts">Key Concepts&lt;/h2>
&lt;ul>
&lt;li>说话人聚类&lt;/li>
&lt;li>身份标注&lt;/li>
&lt;li>谁说了什么&lt;/li>
&lt;li>音频分段&lt;/li>
&lt;/ul>
&lt;h2 id="use-cases">Use Cases&lt;/h2>
&lt;ul>
&lt;li>自动生成会议纪要&lt;/li>
&lt;li>访谈转录&lt;/li>
&lt;li>广播媒体分析&lt;/li>
&lt;/ul>
&lt;h2 id="related-terms">Related Terms&lt;/h2>
&lt;ul>
&lt;li>&lt;a href="https://terms-en.ai-term-hub.com/en/terms/speaker_change_detection-%E8%AF%B4%E8%AF%9D%E4%BA%BA%E5%88%87%E6%8D%A2%E6%A3%80%E6%B5%8B/">speaker_change_detection (说话人切换检测)&lt;/a>&lt;/li>
&lt;li>&lt;a href="https://terms-en.ai-term-hub.com/en/terms/speech_to_text-%E8%AF%AD%E9%9F%B3%E8%BD%AC%E6%96%87%E6%9C%AC/">speech_to_text (语音转文本)&lt;/a>&lt;/li>
&lt;li>&lt;a href="https://terms-en.ai-term-hub.com/en/terms/voice_printing-%E5%A3%B0%E7%BA%B9%E8%AF%86%E5%88%AB/">voice_printing (声纹识别)&lt;/a>&lt;/li>
&lt;li>&lt;a href="https://terms-en.ai-term-hub.com/en/terms/audio_analysis-%E9%9F%B3%E9%A2%91%E5%88%86%E6%9E%90/">audio_analysis (音频分析)&lt;/a>&lt;/li>
&lt;/ul></description></item><item><title>自动语音识别</title><link>https://terms-en.ai-term-hub.com/zh/terms/automatic_speech_recognition/</link><pubDate>Sat, 18 Jul 2026 11:08:15 +0000</pubDate><guid>https://terms-en.ai-term-hub.com/zh/terms/automatic_speech_recognition/</guid><description>&lt;h2 id="definition">Definition&lt;/h2>
&lt;p>自动语音识别（ASR），也称为语音转文字，是语音处理的一个子领域，它利用人工智能将音频信号转录为书面文本。现代ASR系统能够处理各种口音、背景噪音和连续语音，广泛应用于人机交互领域。&lt;/p>
&lt;h3 id="summary">Summary&lt;/h3>
&lt;p>一种利用深度学习模型将口语转换为文本的技术。&lt;/p>
&lt;h2 id="key-concepts">Key Concepts&lt;/h2>
&lt;ul>
&lt;li>声学建模&lt;/li>
&lt;li>语言建模&lt;/li>
&lt;li>深度学习&lt;/li>
&lt;li>转录&lt;/li>
&lt;/ul>
&lt;h2 id="use-cases">Use Cases&lt;/h2>
&lt;ul>
&lt;li>语音助手（如Siri、Alexa）&lt;/li>
&lt;li>视频实时字幕生成&lt;/li>
&lt;li>专业人士的听写软件&lt;/li>
&lt;/ul>
&lt;h2 id="related-terms">Related Terms&lt;/h2>
&lt;ul>
&lt;li>&lt;a href="https://terms-en.ai-term-hub.com/en/terms/natural_language_processing-%E8%87%AA%E7%84%B6%E8%AF%AD%E8%A8%80%E5%A4%84%E7%90%86/">natural_language_processing (自然语言处理)&lt;/a>&lt;/li>
&lt;li>&lt;a href="https://terms-en.ai-term-hub.com/en/terms/speaker_identification-%E8%AF%B4%E8%AF%9D%E4%BA%BA%E8%AF%86%E5%88%AB/">speaker_identification (说话人识别)&lt;/a>&lt;/li>
&lt;li>&lt;a href="https://terms-en.ai-term-hub.com/en/terms/audio_processing-%E9%9F%B3%E9%A2%91%E5%A4%84%E7%90%86/">audio_processing (音频处理)&lt;/a>&lt;/li>
&lt;li>&lt;a href="https://terms-en.ai-term-hub.com/en/terms/deep_learning-%E6%B7%B1%E5%BA%A6%E5%AD%A6%E4%B9%A0/">deep_learning (深度学习)&lt;/a>&lt;/li>
&lt;/ul></description></item></channel></rss>