Clear.
设备端语音增强,让你的录音获得播客录音棚那种温暖、近距离拾音的声音,无需上传。
演示
录音棚音质,没有云端账单。
Adobe Podcast、Dolby 和 Auphonic 都按分钟计费,而且都要求把音频交到它们的服务器上。Clear 在 iPhone 16 Pro 上处理一段 5 分钟素材只需 1 秒,全程在设备上完成。
送进去的是一段笔记本麦克风录的声音:房间没做过声学处理,窗外还有车流。返回的声音像是贴着麦克风录的:中低频温暖,人声在混音中靠前,背后没有房间,响度也已经调到 Spotify、Apple Podcasts 和 YouTube 要求的标准。
Clear 提供两个变体:clear-studio 和 clear-natural。Studio 是默认变体,会清理到接近全静。Natural 会保留房间声、呼吸声和唇齿细节,适合做过声学处理的房间和有意为之的旁白。
5 分钟音频,1 秒处理完。
在 iPhone 16 Pro 上 302 倍实时,在 MacBook Pro M5 上 345 倍:对一段 60 秒素材完成增强、母带处理和重新编码,三次取最好,全部在设备上完成。
clear-studio,对一段 60 秒素材完成增强、母带处理和重新编码,三次取最好
| 设备 | 实时倍率 |
|---|---|
| iPhone 16 Pro | 302x |
| MacBook Pro (M5) | 345x |
在实际发布的 Core ML 构建上的内部测量:一段 5 分钟的单声道素材走完包含增强、母带处理和重新编码的完整流程,iOS 27.0。模型卡给出的数据是:一段 60 秒素材在 iPhone 16 Pro 上 302 倍、在 MacBook Pro(M5)上 345 倍,三次取最好。Android 和浏览器上的速度尚未测量,我们也还没有公布与云端工具对比的音质评分。
播客与视频母带处理.
除掉的不只是底噪,还有整个房间
大多数降噪工具处理到底噪就停了。Clear 连房间反射也一并去掉,在空卧室里录的一条素材,回来时听着像贴着麦克风录的,而不是闷在盒子里。
不用上传,没有按分钟计费
把录音棚级清理做成一项功能,对用户不必按量计费。增强在设备上运行,因此不会有一张随音频处理时长不断增长的云端音频账单。
把糟糕房间里的录音救回来
语音备忘和通话录音在进入转写或归档之前先清理一遍,在厨房里接的电话,转写出来也干干净净。
在手机上修好一段一小时的录音
写入磁盘时,文件路径会分块流式处理,因此一段一小时的录音和十秒的片段占用同样的内存。增强后的文件就落在原文件旁边。
灵感
用 Clear 构建的点子。把提示词复制到你的编码 agent 里就能开始。
Add a one-tap 'clean audio' button that removes noise and room echo.
Add a 'clean audio' action that runs Desert Ant's Clear on-device to remove background noise and room reverb from a recording, with presets for podcast, video, and voiceover loudness. Nothing uploaded. Build it with the Desert Ant SDK. Clear (Swift, Kotlin, JavaScript / TypeScript). Install and API: https://desertant.com/docs/clear/. Install: SwiftPM desert-ant-core; Maven ai.desertant:clear:3.1.0; npm i @desert-ant-labs/clear @litertjs/core # browser. SDK source: https://github.com/Desert-Ant-Labs/desert-ant-core. Machine-readable catalog of every model and SDK: https://desertant.com/llms.txt.
Master a podcast episode to platform loudness, on the device.
Build a podcast export step that runs Desert Ant's Clear to enhance and master the episode to Apple Podcasts / Spotify / YouTube loudness locally, so it's ready to publish without a cloud mastering service. Build it with the Desert Ant SDK. Clear (Swift, Kotlin, JavaScript / TypeScript). Install and API: https://desertant.com/docs/clear/. Install: SwiftPM desert-ant-core; Maven ai.desertant:clear:3.1.0; npm i @desert-ant-labs/clear @litertjs/core # browser. SDK source: https://github.com/Desert-Ant-Labs/desert-ant-core. Machine-readable catalog of every model and SDK: https://desertant.com/llms.txt.
Rescue a recording made in a noisy, untreated room.
Build a 'rescue this recording' feature: run Desert Ant's Clear on-device to make a take recorded in a kitchen or hotel room sound close-miked, warm, and free of background noise. Build it with the Desert Ant SDK. Clear (Swift, Kotlin, JavaScript / TypeScript). Install and API: https://desertant.com/docs/clear/. Install: SwiftPM desert-ant-core; Maven ai.desertant:clear:3.1.0; npm i @desert-ant-labs/clear @litertjs/core # browser. SDK source: https://github.com/Desert-Ant-Labs/desert-ant-core. Machine-readable catalog of every model and SDK: https://desertant.com/llms.txt.
Clean the audio on a phone video in one pass.
In a video app, add a step that runs Desert Ant's Clear over the audio track of a phone video to remove wind, hiss, and echo before sharing. On-device, no per-minute cost. Build it with the Desert Ant SDK. Clear (Swift, Kotlin, JavaScript / TypeScript). Install and API: https://desertant.com/docs/clear/. Install: SwiftPM desert-ant-core; Maven ai.desertant:clear:3.1.0; npm i @desert-ant-labs/clear @litertjs/core # browser. SDK source: https://github.com/Desert-Ant-Labs/desert-ant-core. Machine-readable catalog of every model and SDK: https://desertant.com/llms.txt.
Record, clean, transcribe, and clip: a whole creator pipeline, offline.
Build a creator pipeline: Desert Ant's Clear cleans the recording, Voz transcribes it, and Clips pulls the shorts, all on the device so a full episode never touches a server. Build it with the Desert Ant SDK. Clear (Swift, Kotlin, JavaScript / TypeScript). Install and API: https://desertant.com/docs/clear/. Voz (Swift). Install and API: https://desertant.com/docs/voz/. Clips (Swift). Install and API: https://desertant.com/docs/clips/. SDK source: https://github.com/Desert-Ant-Labs/desert-ant-core. Machine-readable catalog of every model and SDK: https://desertant.com/llms.txt.
Build a voice-notes app that cleans the audio and transcribes it offline.
Build a voice-notes app. Run Desert Ant's Clear on the recording for studio-clean audio, then Voz to transcribe it with word timestamps, entirely on-device so a note taken in a noisy room still reads well and never leaves the phone. Build it with the Desert Ant SDK. Clear (Swift, Kotlin, JavaScript / TypeScript). Install and API: https://desertant.com/docs/clear/. Voz (Swift). Install and API: https://desertant.com/docs/voz/. SDK source: https://github.com/Desert-Ant-Labs/desert-ant-core. Machine-readable catalog of every model and SDK: https://desertant.com/llms.txt.
A clean, filler-free transcript from a raw, noisy recording.
Build a transcript pipeline: Desert Ant's Clear cleans the audio, Voz transcribes it, and Uhm removes the fillers, so a messy recording becomes a clean transcript, on-device. Build it with the Desert Ant SDK. Clear (Swift, Kotlin, JavaScript / TypeScript). Install and API: https://desertant.com/docs/clear/. Voz (Swift). Install and API: https://desertant.com/docs/voz/. Uhm (Swift). Install and API: https://desertant.com/docs/uhm/. SDK source: https://github.com/Desert-Ant-Labs/desert-ant-core. Machine-readable catalog of every model and SDK: https://desertant.com/llms.txt.
A Shortcut that cleans up any audio file's sound.
Build an App Intent for Apple Shortcuts that runs Desert Ant's Clear over an audio file to remove noise and room echo and master it to a chosen loudness, on-device. Build it with the Desert Ant SDK. Clear (Swift, Kotlin, JavaScript / TypeScript). Install and API: https://desertant.com/docs/clear/. Install: SwiftPM desert-ant-core; Maven ai.desertant:clear:3.1.0; npm i @desert-ant-labs/clear @litertjs/core # browser. SDK source: https://github.com/Desert-Ant-Labs/desert-ant-core. Machine-readable catalog of every model and SDK: https://desertant.com/llms.txt.
模型能做什么
- 针对饱满、在场、贴近麦克风的播客声音训练:中低频靠前,人声不薄也不远。
- 由微调过的 DeepFilterNet 3 模型完成降噪和去混响。没做声学处理的卧室或酒店房间,出来的效果更接近处理过的录音棚,而且不会再添回混响。
- 齿音安全,无处理痕迹:S、T、F 上不会出现刺耳的尖峰,也没有抽吸感或音乐噪声。呼吸声和爆破音完整保留。
- 响度预设:Apple Podcasts、Spotify、YouTube、EBU R128(LUFS / EBU R128 归一化)。输出为 48 kHz。
- 可解码 AVFoundation 能读取的任何音频文件,输出扩展名决定编码:
.wav为 16-bit PCM,.m4a、.mp4或.aac为 AAC,.caf或.aiff为 PCM。没有文件系统时,enhance(bytes:)返回 WAV 字节。 - 两个变体:
clear-studio用于安静的录音棚清理,clear-natural保留房间氛围声。权重在首次使用时从 Hub 获取并缓存。两者都提供Strength和Mastering控制项,可按喜好调节。
快速上手
只需几行代码,即可为你的 iOS or macOS, Android or web 应用加上 语音增强。 Clear 文档.
// Swift Package Manager
.package(url: "https://github.com/Desert-Ant-Labs/desert-ant-core", from: "3.1.0")
// target dependency
.product(name: "Clear", package: "desert-ant-core")
import Clear
let clear = Clear()
let result = try await clear.enhance(path: "in.wav", to: "out.wav")
print(result.realtimeFactor, result.measuredLUFS ?? 0)
Add Clear from Desert Ant Labs to this Swift project (iOS, macOS). What it does: Clear:面向语音录音的设备端录音棚音质. SDK: Swift (iOS, macOS) Repo: https://github.com/Desert-Ant-Labs/desert-ant-core#readme // Swift Package Manager .package(url: "https://github.com/Desert-Ant-Labs/desert-ant-core", from: "3.1.0") // target dependency .product(name: "Clear", package: "desert-ant-core") Reference: - Model page: https://desertant.com/models/clear/ - Full catalog and other models: https://desertant.com/llms.txt Add the SDK, then follow its README for the exact API and current version. Do not invent API names or method signatures; confirm them against the README.
// build.gradle.kts (Maven Central)
implementation("ai.desertant:clear:3.1.0")
import ai.desertant.clear.Clear
Clear(context).use { clear ->
val result = clear.enhance(samples, 48_000.0) // 48kHz mono out
result.measuredTruePeakDbfs
}
Add Clear from Desert Ant Labs to this Kotlin project (Android).
What it does: Clear:面向语音录音的设备端录音棚音质.
SDK:
Kotlin (Android)
Repo: https://github.com/Desert-Ant-Labs/desert-ant-core#readme
// build.gradle.kts (Maven Central)
implementation("ai.desertant:clear:3.1.0")
Reference:
- Model page: https://desertant.com/models/clear/
- Full catalog and other models: https://desertant.com/llms.txt
Add the SDK, then follow its README for the exact API and current version. Do not invent API names or method signatures; confirm them against the README.
npm i @desert-ant-labs/clear @litertjs/core # browser
npm i @desert-ant-labs/clear # Node
import { Clear } from "@desert-ant-labs/clear";
const clear = await Clear.load();
const result = await clear.enhance(samples, 48_000); // Float32Array in, 48kHz out
await clear.enhance(samples, 48_000, { targetLUFS: "spotify" });
Add Clear from Desert Ant Labs to this JavaScript / TypeScript project (Web, Node.js). What it does: Clear:面向语音录音的设备端录音棚音质. SDK: JavaScript / TypeScript (Web, Node.js) Repo: https://github.com/Desert-Ant-Labs/desert-ant-core#readme npm i @desert-ant-labs/clear @litertjs/core # browser npm i @desert-ant-labs/clear # Node Reference: - Model page: https://desertant.com/models/clear/ - Full catalog and other models: https://desertant.com/llms.txt Add the SDK, then follow its README for the exact API and current version. Do not invent API names or method signatures; confirm them against the README.
规格
- 速度
- 在 iPhone 16 Pro 上 302 倍实时
- 模型
- 微调过的 DeepFilterNet 3
- 设备端体积
- 每个变体 9.0 MB Core ML,24 MB ONNX
- 平台
- iOS 18+、macOS 15+、tvOS 18+、visionOS 2+(Core ML);Android 7+(LiteRT);浏览器(WebAssembly 和 LiteRT.js);Node
Clear 是为人声打造的。音乐、音效和其他非人声会被当作噪声压下去,所以请把 Clear 用在人声轨上,而不是已经混好的成品上。Clear 也不是声源分离工具:两个人抢着说话,出来仍然是叠在一起的。clear-studio 清理得很彻底,连呼吸声和房间氛围声都会被剥掉,房间做过声学处理时请改用 clear-natural。
常见问题
Clear 是什么?
设备端语音增强,让你的录音获得播客录音棚那种温暖、近距离拾音的声音,无需上传。
Clear 在设备上运行吗?
是的。Clear 在设备上运行,不调用任何服务器,因此数据始终留在用户手中。
Clear 支持哪些平台?
Clear 以面向 Swift, Kotlin, JavaScript / TypeScript 的原生设备端 SDK 形式提供。
Clear 的价格是多少?
每个模型的每个 SDK 均可免费支持最多 10 万台月活跃设备。每位用户调用模型的次数不设上限。 如需定制授权,请联系我们。
Clear 的准确度和速度如何?
在 iPhone 16 Pro 上 302 倍实时,在 MacBook Pro M5 上 345 倍:对一段 60 秒素材完成增强、母带处理和重新编码,三次取最好,全部在设备上完成。