- Add new transcript file `generated_transcripts/transcript_TWV4JQtGxTA.txt` containing 946 lines of content
- Update .gitignore to ignore transcript files in root directory with leading slash pattern
- Added 'has', 'acquire', and 'without' to COMMON_ENGLISH_WORDS dictionary
- Modified is_english detection threshold from 0.05 to 0.04 matches ratio
- Added debug logging to show matches count during English detection
Move COMMON_ENGLISH_WORDS set from youtube_transcript_downloader.py to language_codes.py as a shared constant. This improves code organization and reusability, allowing other modules to reference the same word list for English language detection.
Move COMMON_ENGLISH_WORDS set from youtube_transcript_downloader.py to language_codes.py as a shared constant. This improves code organization and reusability, allowing other modules to reference the same word list for English language detection
- Introduce `language_codes.py` with complete YouTube language codes and fallback lists
- Refactor `youtube_transcript_downloader.py` to use centralized language lists
- Replace interactive prompt with automatic English translation
- Implement sequential fallback: English → common languages → all supported languages
Tested video v75iZrtI6hw and researched APIs for live stream transcripts. The get_transcript() method doesn't exist in youtube-transcript-api library. Should use YouTubeTranscriptApi().fetch() instead.