@mei2
First and foremost thank you for the insane work on version 1.9.3, this is trully impressive.
I have some question about the setting in order to make sure I checked the right boxes.
1.
pass1_pipeline : Is transformer worth it or I should keep fidelity as previous version?
2.
Speech detection pass1_builtin_vad: 3.1 and
pass1_speech_segmenter: silero-v4.0
, are they the correct choice?
3.
Pass1_scene_detector: Silero or I should keep it on automatic?
4.
pass1_speech_enhancer: ffmpeg-dsp as I like the two-pass but should I tick any other boxes other than
pass1_ffmpeg_loudnorm?
5. For the second pass, is using
Qwen the right choice?
Once again your work is trully appraciate and rest assure that I will contribute by supporting the project with the ''coffee''
Edit: I tested Transformer and I got that message: whisperjav - WARNING - Translation mode was requested but output appears to be in Japanese (15/18 chars are Japanese). This may indicate HuggingFace translation is not working as expected. The subtitles were in Japanese despite the fact that subtitle language was English (whisper direct). If you know what I did wrong let me know, testing fidelity now.
Edit 2 (fidelity/agressive): whisperjav - ERROR - Run exits with status 1: at least one file is in a failing state (failed). /usr/lib/python3.13/multiprocessing/resource_tracker.py:479: UserWarning: resource_tracker: There appear to be 2 leaked semaphore objects to clean up at shutdown: {'/loky-3973-eqp7chec', '/loky-12621-ip_mpn2i'} warnings.warn( FAIL Transcription failed. An exception has occurred.
Got subtitle but only from second pass and they were in japanese even with english subtitle language selected.
Edit 3: Testing Command: whisperjav /content/drive/MyDrive/WhisperJAV --output-dir /content/drive/MyDrive/WhisperJAV --mode fidelity --sensitivity aggressive --subs-language direct-to-english with only one pass this time,
More to follow tomorrow, ran out of memory..
-Besh