<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
	<channel>
		<title>GTCRN on 编程心语</title>
		<link>https://www.ithome.me/tags/gtcrn/</link>
		<description>Recent content in GTCRN on 编程心语</description>
		<generator>Hugo</generator>
		<language>zh-CN</language>
		
		
		
		
			<lastBuildDate>Mon, 05 Oct 2026 03:07:55 +0800</lastBuildDate>
		
			<atom:link href="https://www.ithome.me/tags/gtcrn/index.xml" rel="self" type="application/rss+xml" />
			<item>
				<title>端侧语音前处理实战：Silero VAD 断句 &#43; GTCRN 降噪，把嘈杂语音洗干净</title>
				<link>https://www.ithome.me/post/2026/10/05/on-device-speech-vad-gtcrn/</link>
				<pubDate>Mon, 05 Oct 2026 08:00:00 +0800</pubDate>
				<guid>https://www.ithome.me/post/2026/10/05/on-device-speech-vad-gtcrn/</guid>
				<description>&lt;p&gt;做端侧语音应用，很多人的第一步就是上 ASR，结果识别率怎么调都上不去。问题往往不在识别模型，而在送进去的音频：地铁里的风噪、办公室的人声、麦克风底噪，都会让识别率断崖式下跌。更糟的是整段录音里大量是静音，直接喂给 ASR 既浪费算力，又容易触发幻觉，静音段被模型硬生生编出文字。&lt;/p&gt;</description>
			</item>
	</channel>
</rss>
