Skip to content

Tag

#Audio Processing

9 posts and news items use this tag.

GitHub TechAI & Tools

LosslessCut Tested: A 5-Minute, 600 MB Video Cut Almost Instantly

LosslessCut copies media streams through FFmpeg without re-encoding. I tested it with a roughly 5-minute, 600 MB video and documented the free download, I/O shortcuts, and keyframe limitation.

5 min readEasy
LOCAL
macOSWindowsLinux
GPU: None
GitHub TechAI & Tools

Lively Wallpaper Review: Free Open-Source Windows Live Wallpapers with Audio Sync and Auto-Pause

My hands-on look at Lively Wallpaper, a free and open-source Windows live-wallpaper app. It accepts videos, GIFs, web pages, and URLs, supports audio visualizers, and can pause wallpapers automatically for full-screen apps.

7 min readEasy
LOCAL
Windows
GPU: Optional
Local AIAI & Tools

Meetily Hands-On: A Local AI Tool for Meeting Transcription, Transcripts, and Traditional Chinese Summaries

Meetily is an open-source, privacy-first AI meeting assistant that can record, transcribe, and generate summaries locally. I tested v0.4.0 and organized the installation flow, Chinese transcription model setup, Traditional Chinese summaries, and the situations where it makes sense to use.

6 min readMedium
LOCAL
macOSWindowsLinux
GPU: Optional
Local AIAI & Tools

Hands-on with Handy Offline Voice Input: Paired with Breeze ASR 25, a Chinese-English Mixed Input Setup Built for Taiwanese Users

Handy is a free, open-source desktop voice input tool that runs fully offline. In this article, I test installation and configuration on macOS/Windows/Linux, paired with the Breeze ASR 25 model optimized for everyday Chinese-English mixed speech in Taiwan, to build a privacy-focused voice input experience.

5 min readEasy
LOCAL
macOSWindowsLinux
GPU: None
AI AgentAI & Tools

Gemini 3.5 Live Translate Hands-On: New Two-Way Real-Time Voice Translation and Live API Developer Guide

Google has released the new Gemini 3.5 Live Translate model, supporting low-latency two-way real-time voice translation across more than 70 languages. This article walks through a hands-on test of the Google AI Studio web experience, explains how it works, and provides Live API development examples and WebSocket connection setup.

7 min readMedium
ONLINE
WebAPI
GPU: None
Local AITools

Vibe Review: A Cross-Platform Offline AI Speech-to-Text Tool for Beginners — Clean, Intuitive, One-Click

Dread the Whisper command line or find other tools too complicated to set up? Vibe is a minimal, hassle-free offline speech-to-subtitle tool. The feature set is simple, but the interface is clean and easy to pick up — a solid choice for less technical users.

4 min readEasy
LOCAL
macOSWindowsLinux
GPU: None
Local AIAI & Tools

Hands-on with OmniVoice Studio, a Local AI Video Dubbing Tool, plus a macOS Installation Pitfall Guide

I recently tested OmniVoice Studio, an open-source alternative to ElevenLabs + HeyGen. It supports 646 languages, runs automatic local video dubbing, and even works on a Mac mini. This article covers my hands-on notes, how to bypass macOS quarantine, and a voice comparison between Traditional Chinese and Simplified Chinese input.

10 min readHard
LOCAL
macOSWindows
GPU: Apple Silicon / NVIDIA 8GB+
Local AIOpen Source

VideoLingo Local AI Video Subtitle Translation & Chinese Dubbing Deployment Guide

I tested VideoLingo, from raw video to Chinese subtitles and Chinese-dubbed video, all automated. This post covers features, actual results, and my recommended model settings.

9 min readHard
LOCAL
WindowsLinux
GPU: 6GB+ VRAM
Local AIAI & Tools

Installing Voicebox: A Local AI Voice Studio Guide

A developer-oriented guide to Voicebox — from macOS/Windows installation to voice cloning, and how to make your AI agent speak via MCP.

8 min readMedium
LOCAL
WindowsLinux
GPU: 6GB+ VRAM