v1.0.1 Feature Update and Polish

Full Changelog: [New Features] - Added Native Translation Mode: - Whisper model now fully supports Translating any language to English - Added 'task' and 'language' parameters to Transcriber core - Dual Hotkey Support: - Added separate Global Hotkeys for Transcribe (default F8) and Translate (default F10) - Both hotkeys are fully customizable in Settings - Engine dynamically switches modes based on which key is pressed [UI/UX Improvements] - Settings Window: - Widened Hotkey Input fields (240px) to accommodate long combinations - Added Pretty-Printing for hotkey sequences (e.g. 'ctrl+f9' display as 'Ctrl + F9') - Replaced Country Code dropdown with Full Language Names (99+ languages) - Made Language Dropdown scrollable (max height 300px) to prevent screen overflow - Removed redundant 'Task' selector (replaced by dedicated hotkeys) - System Tray: - Tooltip now displays both Transcribe and Translate hotkeys - Tooltip hotkeys are formatted readably [Core & Performance] - Bootstrapper: - Implemented Smart Incremental Sync - Now checks filesize and content hash before copying files - Drastically reduces startup time for subsequent runs - Preserves user settings.json during updates - Backend: - Fixed HotkeyManager to support dynamic configuration keys - Fixed Language Lock: selecting a language now correctly forces the model to use it - Refactored bridge/main connection for language list handling
2026-01-24 18:29:10 +02:00
parent f184eb0037
commit 4b84a27a67
11 changed files with 342 additions and 72 deletions
--- a/src/core/transcriber.py
+++ b/src/core/transcriber.py
@@ -74,11 +74,11 @@ class WhisperTranscriber:
            logging.error(f"Failed to load model: {e}")
            self.model = None

-    def transcribe(self, audio_data, is_file: bool = False) -> str:
+    def transcribe(self, audio_data, is_file: bool = False, task: Optional[str] = None) -> str:
        """
        Transcribe audio data.
        """
-        logging.info(f"Starting transcription... (is_file={is_file})")
+        logging.info(f"Starting transcription... (is_file={is_file}, task={task})")
        
        # Ensure model is loaded
        if not self.model:
@@ -91,6 +91,10 @@ class WhisperTranscriber:
            beam_size = int(self.config.get("beam_size"))
            best_of = int(self.config.get("best_of"))
            vad = False if is_file else self.config.get("vad_filter")
+            language = self.config.get("language")
+            
+            # Use task override if provided, otherwise config
+            final_task = task if task else self.config.get("task")
            
            # Transcribe
            segments, info = self.model.transcribe(
@@ -98,6 +102,8 @@ class WhisperTranscriber:
                beam_size=beam_size,
                best_of=best_of,
                vad_filter=vad,
+                task=final_task,
+                language=language if language != "auto" else None,
                vad_parameters=dict(min_silence_duration_ms=500),
                condition_on_previous_text=self.config.get("condition_on_previous_text"),
                without_timestamps=True