local-transcription

Author	SHA1	Message	Date
Josh Knapp	a5556c475d	Fix uv index configuration: Use PyTorch CUDA as additional index - Changed from 'default' to named additional index - Added tool.uv.sources to specify torch comes from pytorch-cu121 index - Other packages (fastapi, uvicorn, etc.) still come from PyPI - Fixes: 'fastapi was not found in the package registry' error How it works: - PyPI remains the default index for most packages - torch package explicitly uses pytorch-cu121 index - Best of both worlds: CUDA PyTorch + all other packages from PyPI	2025-12-26 12:13:40 -08:00
Josh Knapp	0bcd8e8d21	Configure uv to always use PyTorch CUDA index Changes: - Set PyTorch CUDA index (cu121) as default for all builds - CUDA builds support both GPU and CPU (auto-fallback) - Fixes uv run reinstalling CPU-only PyTorch - Updated dependency-groups syntax (fixes deprecation warning) Benefits: - Simpler build process - no CPU vs CUDA distinction needed - uv sync and uv run now get CUDA-enabled PyTorch automatically - Builds work on systems with or without NVIDIA GPUs - Fixes issue where uv run check_cuda.py was getting CPU version Index: https://download.pytorch.org/whl/cu121 (PyTorch 2.5.1+cu121)	2025-12-26 12:08:42 -08:00
Josh Knapp	d51b24e2e5	Move FastAPI and uvicorn to main dependencies - Web server is always-running (not optional) for OBS integration - Users no longer need to manually install fastapi and uvicorn - Previously required: uv pip install "fastapi[standard]" uvicorn - Now auto-installed with: uv sync Fixes: Missing FastAPI/uvicorn dependencies on fresh Windows installs	2025-12-26 11:57:50 -08:00
Josh Knapp	472233aec4	Initial commit: Local Transcription App v1.0 Phase 1 Complete - Standalone Desktop Application Features: - Real-time speech-to-text with Whisper (faster-whisper) - PySide6 desktop GUI with settings dialog - Web server for OBS browser source integration - Audio capture with automatic sample rate detection and resampling - Noise suppression with Voice Activity Detection (VAD) - Configurable display settings (font, timestamps, fade duration) - Settings apply without restart (with automatic model reloading) - Auto-fade for web display transcriptions - CPU/GPU support with automatic device detection - Standalone executable builds (PyInstaller) - CUDA build support (works on systems without CUDA hardware) Components: - Audio capture with sounddevice - Noise reduction with noisereduce + webrtcvad - Transcription with faster-whisper - GUI with PySide6 - Web server with FastAPI + WebSocket - Configuration system with YAML Build System: - Standard builds (CPU-only): build.sh / build.bat - CUDA builds (universal): build-cuda.sh / build-cuda.bat - Comprehensive BUILD.md documentation - Cross-platform support (Linux, Windows) Documentation: - README.md with project overview and quick start - BUILD.md with detailed build instructions - NEXT_STEPS.md with future enhancement roadmap - INSTALL.md with setup instructions 🤖 Generated with [Claude Code](https://claude.com/claude-code) Co-Authored-By: Claude <noreply@anthropic.com>	2025-12-25 18:48:23 -08:00

4 Commits