โ† Logs

social_media_monitor_2026-06-24.log

2026-06-24 20:30:02,600 - Yiddish24Scraper - INFO - File downloader initialized with 55 terms, prefix 'Y2', and file type organization enabled
2026-06-24 20:30:02,600 - Yiddish24Scraper - INFO - Yiddish24 scraper initialized: data_id='latest', page_limit=100, total_pages=10, debug_html=True, include_terms=55, exclude_terms=0
2026-06-24 20:30:02,603 - AhBlickLiveScraper - INFO - AhBlickLive scraper initialized: urls=1, max_items=100, include_terms=49, exclude_terms=0
2026-06-24 20:30:02,607 - TorahAnytimeScraper - INFO - File downloader initialized with 53 terms, prefix 'TA', and file type organization enabled
2026-06-24 20:30:02,608 - TorahAnytimeScraper - INFO - TorahAnytime scraper initialized: limit=400, offset=0, project_id=1, include_terms=53, exclude_terms=10
2026-06-24 20:30:02,610 - BatorahScraper - INFO - File downloader initialized with 19 terms, prefix 'BT', and file type organization enabled
2026-06-24 20:30:02,612 - BatorahScraper - INFO - Batorah scraper initialized: api_urls=2, page_size=100, include_terms=19, exclude_terms=1
2026-06-24 20:30:02,612 - KolHalashonScraper - INFO - Kol Halashon scraper initialized: base_url=https://www2.kolhalashon.com, max_shiurim=50, fetch_all_speakers=True, headless=False, exclude_filters=1, include_filters=0
2026-06-24 20:30:02,612 - __main__ - INFO - Running 3 scrapers concurrently...
2026-06-24 20:30:02,612 - GitHubScraper - INFO - Searching GitHub for terms: ['paperclip', '"Hermes Agent"', 'multica']
2026-06-24 20:30:02,706 - LinkedInFeedScraper - INFO - Fetching LinkedIn feed (25 posts)
2026-06-24 20:30:03,035 - LinkedInFeedScraper - INFO - Auto-grabbed fresh LinkedIn cookies from running Chrome (fallback).
2026-06-24 20:30:13,033 - LinkedInFeedScraper - INFO - Successfully sorted fetched LinkedIn posts chronologically (most recent first)
2026-06-24 20:30:13,051 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Joel Goldstein - The one thing missing from every strategy you have ever followed.
2026-06-24 20:30:13,070 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Sol Jacobs - Job #2131: Property Compliance & Grant Applications Coordinator | Brooklyn, NY
Our client, a compan
2026-06-24 20:30:13,089 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Yossi Wachtel - Words are important. They shape how we think about things. And there's a word my whole industry is i
2026-06-24 20:30:13,155 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Pickleton - The Pickleton Classic is proud to welcome Theย Wellstone Groupย as a Team Sponsor for their 2nd consec
2026-06-24 20:30:13,171 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Robert Rahmanian - WE'RE HIRING: New Development Leasing Agents
REAL New York has 3000+ new development units active an
2026-06-24 20:30:13,183 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Yoav Nordmann - What a day at MLcon in #Munich
For me it's all about opening your mind to new ideas! That's what I
2026-06-24 20:30:13,203 - LinkedInFeedScraper - INFO - New LinkedIn feed post: StandWithUs - Ben & Jerry's Israel has launched its most Israeli flavor yet - "Milk & Honey," a tribute to the sou
2026-06-24 20:30:13,220 - LinkedInFeedScraper - INFO - LinkedIn Feed scrape complete. Found 7 new posts.
2026-06-24 20:30:14,424 - httpx - INFO - HTTP Request: GET https://api.github.com/users/abeperl/received_events?per_page=100&page=1 "HTTP/1.1 200 OK"
2026-06-24 20:30:15,222 - httpx - INFO - HTTP Request: GET https://api.github.com/search/repositories?q=paperclip+in%3Aname%2Cdescription%2Creadme+stars%3A%3E150&sort=updated&order=desc&per_page=20 "HTTP/1.1 200 OK"
2026-06-24 20:30:15,224 - GitHubScraper - INFO - Found 241 GitHub repositories for term: paperclip
2026-06-24 20:30:16,586 - httpx - INFO - HTTP Request: GET https://api.github.com/search/repositories?q=%22Hermes+Agent%22+in%3Aname%2Cdescription%2Creadme+stars%3A%3E150&sort=updated&order=desc&per_page=20 "HTTP/1.1 200 OK"
2026-06-24 20:30:16,590 - GitHubScraper - INFO - Found 395 GitHub repositories for term: "Hermes Agent"
2026-06-24 20:30:16,727 - GitHubScraper - INFO - Added new GitHub repository: conductor
2026-06-24 20:30:16,762 - httpx - INFO - HTTP Request: GET https://api.github.com/users/abeperl/received_events?per_page=100&page=2 "HTTP/1.1 200 OK"
2026-06-24 20:30:17,788 - httpx - INFO - HTTP Request: GET https://api.github.com/search/repositories?q=multica+in%3Aname%2Cdescription%2Creadme+stars%3A%3E150&sort=updated&order=desc&per_page=20 "HTTP/1.1 200 OK"
2026-06-24 20:30:17,792 - GitHubScraper - INFO - Found 54 GitHub repositories for term: multica
2026-06-24 20:30:17,942 - GitHubScraper - INFO - Total GitHub results found: 1
2026-06-24 20:30:19,019 - httpx - INFO - HTTP Request: GET https://api.github.com/users/abeperl/received_events?per_page=100&page=3 "HTTP/1.1 200 OK"
2026-06-24 20:30:19,950 - GitHubScraper - INFO - GitHub feed: 298 fetched across pages, 0 new events for abeperl
2026-06-24 20:30:19,952 - __main__ - INFO - Scraper github completed: 1 results
2026-06-24 20:30:19,952 - __main__ - INFO - Scraper github_feed completed: 0 results
2026-06-24 20:30:19,952 - __main__ - INFO - Scraper linkedin_feed completed: 7 results
2026-06-24 20:30:19,952 - social_media_monitor.sheets.webinar_sheets_writer - INFO - Checking 7 LinkedIn feed posts for webinar/course mentions...
2026-06-24 20:30:21,120 - social_media_monitor.sheets.webinar_sheets_writer - INFO - Loaded 23 existing URLs from sheet for dedup
2026-06-24 20:30:21,121 - social_media_monitor.sheets.webinar_sheets_writer - INFO - Webinar/course match: Pickleton - keyword 'class' - https://www.linkedin.com/feed/update/urn:li:activity:7475643573170900992
2026-06-24 20:30:21,121 - social_media_monitor.sheets.webinar_sheets_writer - INFO - Webinar/course match: Yoav Nordmann - keyword 'lecture' - https://www.linkedin.com/feed/update/urn:li:activity:7475631690418413569
2026-06-24 20:30:22,121 - social_media_monitor.sheets.webinar_sheets_writer - INFO - Appended 2 rows to Google Sheet 'Webinar Recording DMs'
2026-06-24 20:30:22,121 - __main__ - INFO - Webinar/Course: appended 2 posts to Google Sheet
2026-06-24 20:30:22,121 - __main__ - CRITICAL - Critical error in main: name 'scrapers' is not defined
Traceback (most recent call last):
File "/home/openclaw/Social_Media_Monitor/social_media_monitor.py", line 219, in main
_check_linkedin_feed_failure('linkedin_feed' in scrapers, len(results.get('linkedin_feed_results', [])))
^^^^^^^^
NameError: name 'scrapers' is not defined
2026-06-24 20:30:22,123 - social_media_monitor.database.connection - INFO - Closing all database connections
2026-06-24 20:30:22,194 - social_media_monitor.database.connection - INFO - Closed 5 database connections
2026-06-24 20:30:22,194 - __main__ - INFO - Closed all database connections
2026-06-24 20:30:22,196 - social_media_monitor.database.connection - INFO - Closing all database connections
2026-06-24 20:30:22,196 - social_media_monitor.database.connection - INFO - Closed 0 database connections
2026-06-24 21:00:01,838 - social_media_monitor.utils.logging_config - INFO - Posts log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_posts_2026-06-24.log
2026-06-24 21:00:01,838 - social_media_monitor.utils.logging_config - INFO - Downloads log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_downloads_2026-06-24.log
2026-06-24 21:00:01,838 - social_media_monitor.utils.logging_config - INFO - Logging configured with level INFO, log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_2026-06-24.log
2026-06-24 21:00:01,838 - __main__ - INFO - === Social Media Monitor v2.0 (Async) Started ===
2026-06-24 21:00:01,863 - social_media_monitor.config.settings - INFO - Configuration loaded from /home/openclaw/Social_Media_Monitor/config.yaml
2026-06-24 21:00:01,863 - __main__ - INFO - Configuration loaded and validated successfully
2026-06-24 21:00:01,865 - social_media_monitor.database.connection - INFO - Database connection pool initialized with 5 connections
2026-06-24 21:00:01,866 - social_media_monitor.database.connection - INFO - add_speaker_filters column already exists
2026-06-24 21:00:01,911 - social_media_monitor.database.connection - INFO - Database setup completed successfully
2026-06-24 21:00:01,911 - __main__ - INFO - Database initialized at: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/social_media_monitor.db
2026-06-24 21:00:01,916 - social_media_monitor.scrapers.reddit_scraper - INFO - Reddit API initialized successfully
2026-06-24 21:00:01,918 - RSSFeedScraper - INFO - RSS scraper initialized: feeds=2, html_urls=0 (0 with custom selectors), include_terms=0, exclude_terms=0
2026-06-24 21:00:01,918 - __main__ - INFO - RSS scraper initialized and enabled
2026-06-24 21:00:01,924 - Yiddish24Scraper - INFO - File downloader initialized with 55 terms, prefix 'Y2', and file type organization enabled
2026-06-24 21:00:01,924 - Yiddish24Scraper - INFO - Yiddish24 scraper initialized: data_id='latest', page_limit=100, total_pages=10, debug_html=True, include_terms=55, exclude_terms=0
2026-06-24 21:00:01,924 - __main__ - INFO - Yiddish24 scraper initialized and enabled
2026-06-24 21:00:01,930 - AhBlickLiveScraper - INFO - AhBlickLive scraper initialized: urls=1, max_items=100, include_terms=49, exclude_terms=0
2026-06-24 21:00:01,941 - TorahAnytimeScraper - INFO - File downloader initialized with 53 terms, prefix 'TA', and file type organization enabled
2026-06-24 21:00:01,941 - TorahAnytimeScraper - INFO - TorahAnytime scraper initialized: limit=400, offset=0, project_id=1, include_terms=53, exclude_terms=10
2026-06-24 21:00:01,945 - BatorahScraper - INFO - File downloader initialized with 19 terms, prefix 'BT', and file type organization enabled
2026-06-24 21:00:01,945 - BatorahScraper - INFO - Batorah scraper initialized: api_urls=2, page_size=100, include_terms=19, exclude_terms=1
2026-06-24 21:00:01,946 - KolHalashonScraper - INFO - Kol Halashon scraper initialized: base_url=https://www2.kolhalashon.com, max_shiurim=50, fetch_all_speakers=True, headless=False, exclude_filters=1, include_filters=0
2026-06-24 21:00:01,946 - __main__ - INFO - Ivelt scraper initialized and enabled
2026-06-24 21:00:01,946 - __main__ - INFO - Running 3 scrapers concurrently...
2026-06-24 21:00:01,946 - Yiddish24Scraper - INFO - Starting Yiddish24 scraping (max 10 pages, limit 100 items)
2026-06-24 21:00:02,032 - Yiddish24Scraper - INFO - Fetching Yiddish24 page 1 with data_id='latest', page_limit=100
2026-06-24 21:00:02,036 - IveltScraper - INFO - Loaded 181 include, 0 exclude filters from DB topic_filters
2026-06-24 21:00:02,037 - IveltScraper - INFO - Fetching ivelt active topics from https://www.ivelt.com/forum/search.php?search_id=active_topics...
2026-06-24 21:00:02,037 - RSSFeedScraper - INFO - Fetching RSS feed: https://www.theyeshivaworld.com/feed
2026-06-24 21:00:02,277 - RSSFeedScraper - INFO - Fetching RSS feed: https://www.thegatewaypundit.com/feed/
2026-06-24 21:00:02,660 - RSSFeedScraper - INFO - Total RSS/HTML results: 3
2026-06-24 21:00:09,010 - httpx - INFO - HTTP Request: POST https://www.yiddish24.com/ajax/cat_pagination.php "HTTP/1.1 200 OK"
2026-06-24 21:00:09,230 - Yiddish24Scraper - INFO - Received response with 494089 characters
2026-06-24 21:00:09,233 - Yiddish24Scraper - INFO - Extracted HTML from JSON result field (310687 chars)
2026-06-24 21:00:09,236 - Yiddish24Scraper - INFO - Saved HTML debug file to: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/debug/yiddish24_page1_response.html
2026-06-24 21:00:09,322 - Yiddish24Scraper - INFO - Found 100 potential article elements on page 1
2026-06-24 21:00:09,363 - Yiddish24Scraper - INFO - Page 1 extraction: 100 successful, 0 failed out of 100 elements
2026-06-24 21:00:09,363 - Yiddish24Scraper - INFO - Parsed 100 articles from HTML on page 1
2026-06-24 21:00:09,374 - Yiddish24Scraper - INFO - Term 'ืื™ื ืกืคื™ืจืืฆื™ืข -' matched in DESCRIPTION: 'ืื™ื ืกืคื™ืจืืฆื™ืข - ื‘ืขื ื˜ืฉ ืžื™ืจ ืื•ื™ืš...'
2026-06-24 21:00:09,374 - Yiddish24Scraper - INFO - Downloading 1 media file(s) for article: ื”ื›ืœ ื‘ื›ืœ
2026-06-24 21:00:09,875 - social_media_monitor.utils.file_downloader - INFO - Starting download from: https://cloudfront.yiddish24.com/____4___260_260624030729.mp3
2026-06-24 21:00:10,080 - social_media_monitor.utils.file_downloader - INFO - Successfully downloaded: Y2_____4___260_260624030729.mp3 (8.6 MB)
2026-06-24 21:00:10,081 - social_media_monitor.utils.file_downloader - INFO - Organized audio file to: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/downloads/Audio (audio/mpeg)
2026-06-24 21:00:10,089 - social_media_monitor.utils.media_tagger - INFO - Tagged Y2_____4___260_260624030729.mp3: title, artist, album, comment, source_url, genre, publisher
2026-06-24 21:00:10,089 - Yiddish24Scraper - INFO - Renamed downloaded file to include match term: Y2_____4___260_260624030729_ืื™ื ืกืคื™ืจืืฆื™ืข_-.mp3
2026-06-24 21:00:10,843 - Yiddish24Scraper - INFO - Page 1 summary: 5 new out of 100 total articles
2026-06-24 21:00:10,844 - Yiddish24Scraper - INFO - Page 1: 5/100 new items
2026-06-24 21:00:10,844 - Yiddish24Scraper - INFO - Low new item ratio (5.0%), stopping pagination
2026-06-24 21:00:10,844 - Yiddish24Scraper - INFO - Total Yiddish24 results found: 5
2026-06-24 21:00:13,852 - IveltScraper - INFO - Found 50 active topics on ivelt
2026-06-24 21:00:20,082 - IveltScraper - INFO - Filtering post containing exclude term: '*ื‘ื™ื˜ืข ืžืื›ื˜ ื–ื™ื›ืขืจ ื‘ืœื•ื™ื– ืฆื• ื‘ืืจื™ื›ื˜ืŸ ืงืจืื ื˜ืข ื ื™ื™ืขืก*'
2026-06-24 21:00:27,089 - IveltScraper - INFO - Filtering post containing exclude term: '*ื‘ื™ื˜ืข ืžืื›ื˜ ื–ื™ื›ืขืจ ื‘ืœื•ื™ื– ืฆื• ื‘ืืจื™ื›ื˜ืŸ ืงืจืื ื˜ืข ื ื™ื™ืขืก*'
2026-06-24 21:00:56,565 - IveltScraper - INFO - ivelt scrape complete. Checked 6 topics.
2026-06-24 21:00:56,565 - IveltScraper - INFO - All forums scrape complete. Found 32 new posts.
2026-06-24 21:00:56,566 - __main__ - INFO - Scraper rss completed: 3 results
2026-06-24 21:00:56,566 - __main__ - INFO - Scraper yiddish24 completed: 5 results
2026-06-24 21:00:56,566 - __main__ - INFO - Scraper ivelt completed: 32 results
2026-06-24 21:00:56,567 - __main__ - CRITICAL - Critical error in main: name 'scrapers' is not defined
Traceback (most recent call last):
File "/home/openclaw/Social_Media_Monitor/social_media_monitor.py", line 219, in main
_check_linkedin_feed_failure('linkedin_feed' in scrapers, len(results.get('linkedin_feed_results', [])))
^^^^^^^^
NameError: name 'scrapers' is not defined
2026-06-24 21:00:56,570 - social_media_monitor.database.connection - INFO - Closing all database connections
2026-06-24 21:00:56,671 - social_media_monitor.database.connection - INFO - Closed 5 database connections
2026-06-24 21:00:56,671 - __main__ - INFO - Closed all database connections
2026-06-24 21:00:56,673 - social_media_monitor.database.connection - INFO - Closing all database connections
2026-06-24 21:00:56,673 - social_media_monitor.database.connection - INFO - Closed 0 database connections
2026-06-24 21:30:02,177 - social_media_monitor.utils.logging_config - INFO - Posts log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_posts_2026-06-24.log
2026-06-24 21:30:02,177 - social_media_monitor.utils.logging_config - INFO - Downloads log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_downloads_2026-06-24.log
2026-06-24 21:30:02,178 - social_media_monitor.utils.logging_config - INFO - Logging configured with level INFO, log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_2026-06-24.log
2026-06-24 21:30:02,178 - __main__ - INFO - === Social Media Monitor v2.0 (Async) Started ===
2026-06-24 21:30:02,203 - social_media_monitor.config.settings - INFO - Configuration loaded from /home/openclaw/Social_Media_Monitor/config.yaml
2026-06-24 21:30:02,203 - social_media_monitor.config.settings - WARNING - Twitter monitoring enabled but credentials validation not implemented
2026-06-24 21:30:02,203 - __main__ - INFO - Configuration loaded and validated successfully
2026-06-24 21:30:02,206 - social_media_monitor.database.connection - INFO - Database connection pool initialized with 5 connections
2026-06-24 21:30:02,206 - social_media_monitor.database.connection - INFO - add_speaker_filters column already exists
2026-06-24 21:30:02,225 - social_media_monitor.database.connection - INFO - Database setup completed successfully
2026-06-24 21:30:02,226 - __main__ - INFO - Database initialized at: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/social_media_monitor.db
2026-06-24 21:30:02,274 - social_media_monitor.scrapers.reddit_scraper - INFO - Reddit API initialized successfully
2026-06-24 21:30:02,274 - __main__ - INFO - GitHub scraper initialized and enabled
2026-06-24 21:30:02,274 - __main__ - INFO - GitHub feed scraper initialized and enabled
2026-06-24 21:30:02,276 - RSSFeedScraper - INFO - RSS scraper initialized: feeds=2, html_urls=0 (0 with custom selectors), include_terms=0, exclude_terms=0
2026-06-24 21:30:02,276 - __main__ - INFO - LinkedIn Feed scraper initialized and enabled
2026-06-24 21:30:02,284 - Yiddish24Scraper - INFO - File downloader initialized with 55 terms, prefix 'Y2', and file type organization enabled
2026-06-24 21:30:02,284 - Yiddish24Scraper - INFO - Yiddish24 scraper initialized: data_id='latest', page_limit=100, total_pages=10, debug_html=True, include_terms=55, exclude_terms=0
2026-06-24 21:30:02,290 - AhBlickLiveScraper - INFO - AhBlickLive scraper initialized: urls=1, max_items=100, include_terms=49, exclude_terms=0
2026-06-24 21:30:02,300 - TorahAnytimeScraper - INFO - File downloader initialized with 53 terms, prefix 'TA', and file type organization enabled
2026-06-24 21:30:02,300 - TorahAnytimeScraper - INFO - TorahAnytime scraper initialized: limit=400, offset=0, project_id=1, include_terms=53, exclude_terms=10
2026-06-24 21:30:02,303 - BatorahScraper - INFO - File downloader initialized with 19 terms, prefix 'BT', and file type organization enabled
2026-06-24 21:30:02,303 - BatorahScraper - INFO - Batorah scraper initialized: api_urls=2, page_size=100, include_terms=19, exclude_terms=1
2026-06-24 21:30:02,303 - KolHalashonScraper - INFO - Kol Halashon scraper initialized: base_url=https://www2.kolhalashon.com, max_shiurim=50, fetch_all_speakers=True, headless=False, exclude_filters=1, include_filters=0
2026-06-24 21:30:02,303 - __main__ - INFO - Running 3 scrapers concurrently...
2026-06-24 21:30:02,303 - GitHubScraper - INFO - Searching GitHub for terms: ['paperclip', '"Hermes Agent"', 'multica']
2026-06-24 21:30:02,404 - LinkedInFeedScraper - INFO - Fetching LinkedIn feed (25 posts)
2026-06-24 21:30:02,730 - LinkedInFeedScraper - INFO - Auto-grabbed fresh LinkedIn cookies from running Chrome (fallback).
2026-06-24 21:30:09,476 - LinkedInFeedScraper - INFO - Successfully sorted fetched LinkedIn posts chronologically (most recent first)
2026-06-24 21:30:09,488 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Nussi Einhorn - Just finished a UX audit for "Simcha Set", built by Raphael Lasry for the amazing organization Eim L
2026-06-24 21:30:09,496 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Aaron Rubin - Yosef Haas and I road-tripped from Dallas to Austin today to visit the wonderful team at Manifest.ec
2026-06-24 21:30:09,509 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Microsoft 365 & SharePoint & Teams & Power Platform & Viva & Digital Transformation & Copilot & AI - #QuestionForGroup
Why do I get a reminder that there are new posts but LinkedIn doesn't show them?
2026-06-24 21:30:09,521 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Abraham Jacobowitz - I witnessed a owner of a multi-million-dollar company call a meeting and start reading from document
2026-06-24 21:30:09,532 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Elya Zieg - It was around 2 PM.
Iโ€™d already been sitting in the office for a few hours, making calls, sending em
2026-06-24 21:30:09,546 - LinkedInFeedScraper - INFO - New LinkedIn feed post: AG DESIGNS - Elegant from every angle.
Warm wood tones and subtle reveals woven into the molding create depth an
2026-06-24 21:30:09,556 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Dameeko Knight, LNHA - As of this Friday, I am officially open to work and actively searching for a role as a Nursing Home
2026-06-24 21:30:09,571 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Lucila T. - Literally every designer's workflow after Figma Config 2026 ๐Ÿ˜‚
2026-06-24 21:30:09,580 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Isaac Fried - I came across this article and thought it was worth sharing.
It discusses how rent concessions cont
2026-06-24 21:30:09,595 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Simon Klein, CPA - Have you heard of the Kwong case?
Did you pay IRS penalties during 2019-2022?
Many taxpayers paid
2026-06-24 21:30:09,631 - LinkedInFeedScraper - INFO - LinkedIn Feed scrape complete. Found 10 new posts.
2026-06-24 21:30:10,740 - httpx - INFO - HTTP Request: GET https://api.github.com/users/abeperl/received_events?per_page=100&page=1 "HTTP/1.1 200 OK"
2026-06-24 21:30:11,443 - httpx - INFO - HTTP Request: GET https://api.github.com/search/repositories?q=paperclip+in%3Aname%2Cdescription%2Creadme+stars%3A%3E150&sort=updated&order=desc&per_page=20 "HTTP/1.1 200 OK"
2026-06-24 21:30:11,446 - GitHubScraper - INFO - Found 241 GitHub repositories for term: paperclip
2026-06-24 21:30:12,611 - httpx - INFO - HTTP Request: GET https://api.github.com/search/repositories?q=%22Hermes+Agent%22+in%3Aname%2Cdescription%2Creadme+stars%3A%3E150&sort=updated&order=desc&per_page=20 "HTTP/1.1 200 OK"
2026-06-24 21:30:12,615 - GitHubScraper - INFO - Found 395 GitHub repositories for term: "Hermes Agent"
2026-06-24 21:30:12,821 - httpx - INFO - HTTP Request: GET https://api.github.com/users/abeperl/received_events?per_page=100&page=2 "HTTP/1.1 200 OK"
2026-06-24 21:30:13,687 - httpx - INFO - HTTP Request: GET https://api.github.com/search/repositories?q=multica+in%3Aname%2Cdescription%2Creadme+stars%3A%3E150&sort=updated&order=desc&per_page=20 "HTTP/1.1 200 OK"
2026-06-24 21:30:13,778 - GitHubScraper - INFO - Found 54 GitHub repositories for term: multica
2026-06-24 21:30:13,913 - GitHubScraper - INFO - Total GitHub results found: 0
2026-06-24 21:30:15,024 - httpx - INFO - HTTP Request: GET https://api.github.com/users/abeperl/received_events?per_page=100&page=3 "HTTP/1.1 200 OK"
2026-06-24 21:30:15,955 - GitHubScraper - INFO - GitHub feed: 298 fetched across pages, 4 new events for abeperl
2026-06-24 21:30:15,956 - __main__ - INFO - Scraper github completed: 0 results
2026-06-24 21:30:15,956 - __main__ - INFO - Scraper github_feed completed: 4 results
2026-06-24 21:30:15,956 - __main__ - INFO - Scraper linkedin_feed completed: 10 results
2026-06-24 21:30:15,957 - social_media_monitor.sheets.webinar_sheets_writer - INFO - Checking 10 LinkedIn feed posts for webinar/course mentions...
2026-06-24 21:30:17,517 - social_media_monitor.sheets.webinar_sheets_writer - INFO - Loaded 25 existing URLs from sheet for dedup
2026-06-24 21:30:17,517 - social_media_monitor.sheets.webinar_sheets_writer - INFO - No webinar/course mentions found in LinkedIn feed posts
2026-06-24 21:30:17,518 - __main__ - CRITICAL - Critical error in main: name 'scrapers' is not defined
Traceback (most recent call last):
File "/home/openclaw/Social_Media_Monitor/social_media_monitor.py", line 219, in main
_check_linkedin_feed_failure('linkedin_feed' in scrapers, len(results.get('linkedin_feed_results', [])))
^^^^^^^^
NameError: name 'scrapers' is not defined
2026-06-24 21:30:17,519 - social_media_monitor.database.connection - INFO - Closing all database connections
2026-06-24 21:30:17,670 - social_media_monitor.database.connection - INFO - Closed 5 database connections
2026-06-24 21:30:17,670 - __main__ - INFO - Closed all database connections
2026-06-24 21:30:17,672 - social_media_monitor.database.connection - INFO - Closing all database connections
2026-06-24 21:30:17,672 - social_media_monitor.database.connection - INFO - Closed 0 database connections
2026-06-24 22:00:02,243 - social_media_monitor.utils.logging_config - INFO - Posts log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_posts_2026-06-24.log
2026-06-24 22:00:02,243 - social_media_monitor.utils.logging_config - INFO - Downloads log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_downloads_2026-06-24.log
2026-06-24 22:00:02,243 - social_media_monitor.utils.logging_config - INFO - Logging configured with level INFO, log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_2026-06-24.log
2026-06-24 22:00:02,243 - __main__ - INFO - === Social Media Monitor v2.0 (Async) Started ===
2026-06-24 22:00:02,270 - social_media_monitor.config.settings - INFO - Configuration loaded from /home/openclaw/Social_Media_Monitor/config.yaml
2026-06-24 22:00:02,270 - __main__ - INFO - Configuration loaded and validated successfully
2026-06-24 22:00:02,272 - social_media_monitor.database.connection - INFO - Database connection pool initialized with 5 connections
2026-06-24 22:00:02,273 - social_media_monitor.database.connection - INFO - add_speaker_filters column already exists
2026-06-24 22:00:02,287 - social_media_monitor.database.connection - INFO - Database setup completed successfully
2026-06-24 22:00:02,287 - __main__ - INFO - Database initialized at: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/social_media_monitor.db
2026-06-24 22:00:02,290 - social_media_monitor.scrapers.reddit_scraper - INFO - Reddit API initialized successfully
2026-06-24 22:00:02,290 - RSSFeedScraper - INFO - RSS scraper initialized: feeds=2, html_urls=0 (0 with custom selectors), include_terms=0, exclude_terms=0
2026-06-24 22:00:02,290 - __main__ - INFO - RSS scraper initialized and enabled
2026-06-24 22:00:02,294 - Yiddish24Scraper - INFO - File downloader initialized with 55 terms, prefix 'Y2', and file type organization enabled
2026-06-24 22:00:02,294 - Yiddish24Scraper - INFO - Yiddish24 scraper initialized: data_id='latest', page_limit=100, total_pages=10, debug_html=True, include_terms=55, exclude_terms=0
2026-06-24 22:00:02,294 - __main__ - INFO - Yiddish24 scraper initialized and enabled
2026-06-24 22:00:02,297 - AhBlickLiveScraper - INFO - AhBlickLive scraper initialized: urls=1, max_items=100, include_terms=49, exclude_terms=0
2026-06-24 22:00:02,304 - TorahAnytimeScraper - INFO - File downloader initialized with 53 terms, prefix 'TA', and file type organization enabled
2026-06-24 22:00:02,304 - TorahAnytimeScraper - INFO - TorahAnytime scraper initialized: limit=400, offset=0, project_id=1, include_terms=53, exclude_terms=10
2026-06-24 22:00:02,307 - BatorahScraper - INFO - File downloader initialized with 19 terms, prefix 'BT', and file type organization enabled
2026-06-24 22:00:02,307 - BatorahScraper - INFO - Batorah scraper initialized: api_urls=2, page_size=100, include_terms=19, exclude_terms=1
2026-06-24 22:00:02,307 - KolHalashonScraper - INFO - Kol Halashon scraper initialized: base_url=https://www2.kolhalashon.com, max_shiurim=50, fetch_all_speakers=True, headless=False, exclude_filters=1, include_filters=0
2026-06-24 22:00:02,307 - __main__ - INFO - Ivelt scraper initialized and enabled
2026-06-24 22:00:02,307 - __main__ - INFO - Running 3 scrapers concurrently...
2026-06-24 22:00:02,307 - Yiddish24Scraper - INFO - Starting Yiddish24 scraping (max 10 pages, limit 100 items)
2026-06-24 22:00:02,394 - Yiddish24Scraper - INFO - Fetching Yiddish24 page 1 with data_id='latest', page_limit=100
2026-06-24 22:00:02,398 - IveltScraper - INFO - Loaded 181 include, 0 exclude filters from DB topic_filters
2026-06-24 22:00:02,398 - IveltScraper - INFO - Fetching ivelt active topics from https://www.ivelt.com/forum/search.php?search_id=active_topics...
2026-06-24 22:00:02,399 - RSSFeedScraper - INFO - Fetching RSS feed: https://www.theyeshivaworld.com/feed
2026-06-24 22:00:02,646 - RSSFeedScraper - INFO - Fetching RSS feed: https://www.thegatewaypundit.com/feed/
2026-06-24 22:00:03,007 - RSSFeedScraper - INFO - Total RSS/HTML results: 3
2026-06-24 22:00:06,290 - httpx - INFO - HTTP Request: POST https://www.yiddish24.com/ajax/cat_pagination.php "HTTP/1.1 200 OK"
2026-06-24 22:00:06,502 - Yiddish24Scraper - INFO - Received response with 505489 characters
2026-06-24 22:00:06,504 - Yiddish24Scraper - INFO - Extracted HTML from JSON result field (315046 chars)
2026-06-24 22:00:06,506 - Yiddish24Scraper - INFO - Saved HTML debug file to: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/debug/yiddish24_page1_response.html
2026-06-24 22:00:06,583 - Yiddish24Scraper - INFO - Found 100 potential article elements on page 1
2026-06-24 22:00:06,624 - Yiddish24Scraper - INFO - Page 1 extraction: 100 successful, 0 failed out of 100 elements
2026-06-24 22:00:06,624 - Yiddish24Scraper - INFO - Parsed 100 articles from HTML on page 1
2026-06-24 22:00:07,393 - Yiddish24Scraper - INFO - Page 1 summary: 5 new out of 100 total articles
2026-06-24 22:00:07,394 - Yiddish24Scraper - INFO - Page 1: 5/100 new items
2026-06-24 22:00:07,394 - Yiddish24Scraper - INFO - Low new item ratio (5.0%), stopping pagination
2026-06-24 22:00:07,394 - Yiddish24Scraper - INFO - Total Yiddish24 results found: 5
2026-06-24 22:00:14,110 - IveltScraper - INFO - Found 50 active topics on ivelt
2026-06-24 22:01:03,197 - IveltScraper - INFO - ivelt scrape complete. Checked 6 topics.
2026-06-24 22:01:03,197 - IveltScraper - INFO - All forums scrape complete. Found 25 new posts.
2026-06-24 22:01:03,198 - __main__ - INFO - Scraper rss completed: 3 results
2026-06-24 22:01:03,198 - __main__ - INFO - Scraper yiddish24 completed: 5 results
2026-06-24 22:01:03,198 - __main__ - INFO - Scraper ivelt completed: 25 results
2026-06-24 22:01:03,198 - __main__ - CRITICAL - Critical error in main: name 'scrapers' is not defined
Traceback (most recent call last):
File "/home/openclaw/Social_Media_Monitor/social_media_monitor.py", line 219, in main
_check_linkedin_feed_failure('linkedin_feed' in scrapers, len(results.get('linkedin_feed_results', [])))
^^^^^^^^
NameError: name 'scrapers' is not defined
2026-06-24 22:01:03,200 - social_media_monitor.database.connection - INFO - Closing all database connections
2026-06-24 22:01:03,261 - social_media_monitor.database.connection - INFO - Closed 5 database connections
2026-06-24 22:01:03,261 - __main__ - INFO - Closed all database connections
2026-06-24 22:01:03,263 - social_media_monitor.database.connection - INFO - Closing all database connections
2026-06-24 22:01:03,263 - social_media_monitor.database.connection - INFO - Closed 0 database connections
2026-06-24 22:30:01,725 - social_media_monitor.utils.logging_config - INFO - Posts log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_posts_2026-06-24.log
2026-06-24 22:30:01,725 - social_media_monitor.utils.logging_config - INFO - Downloads log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_downloads_2026-06-24.log
2026-06-24 22:30:01,725 - social_media_monitor.utils.logging_config - INFO - Logging configured with level INFO, log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_2026-06-24.log
2026-06-24 22:30:01,725 - __main__ - INFO - === Social Media Monitor v2.0 (Async) Started ===
2026-06-24 22:30:01,753 - social_media_monitor.config.settings - INFO - Configuration loaded from /home/openclaw/Social_Media_Monitor/config.yaml
2026-06-24 22:30:01,753 - social_media_monitor.config.settings - WARNING - Twitter monitoring enabled but credentials validation not implemented
2026-06-24 22:30:01,753 - __main__ - INFO - Configuration loaded and validated successfully
2026-06-24 22:30:01,755 - social_media_monitor.database.connection - INFO - Database connection pool initialized with 5 connections
2026-06-24 22:30:01,756 - social_media_monitor.database.connection - INFO - add_speaker_filters column already exists
2026-06-24 22:30:01,769 - social_media_monitor.database.connection - INFO - Database setup completed successfully
2026-06-24 22:30:01,769 - __main__ - INFO - Database initialized at: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/social_media_monitor.db
2026-06-24 22:30:01,771 - social_media_monitor.scrapers.reddit_scraper - INFO - Reddit API initialized successfully
2026-06-24 22:30:01,771 - __main__ - INFO - GitHub scraper initialized and enabled
2026-06-24 22:30:01,771 - __main__ - INFO - GitHub feed scraper initialized and enabled
2026-06-24 22:30:01,772 - RSSFeedScraper - INFO - RSS scraper initialized: feeds=2, html_urls=0 (0 with custom selectors), include_terms=0, exclude_terms=0
2026-06-24 22:30:01,772 - __main__ - INFO - LinkedIn Feed scraper initialized and enabled
2026-06-24 22:30:01,775 - Yiddish24Scraper - INFO - File downloader initialized with 55 terms, prefix 'Y2', and file type organization enabled
2026-06-24 22:30:01,775 - Yiddish24Scraper - INFO - Yiddish24 scraper initialized: data_id='latest', page_limit=100, total_pages=10, debug_html=True, include_terms=55, exclude_terms=0
2026-06-24 22:30:01,777 - AhBlickLiveScraper - INFO - AhBlickLive scraper initialized: urls=1, max_items=100, include_terms=49, exclude_terms=0
2026-06-24 22:30:01,782 - TorahAnytimeScraper - INFO - File downloader initialized with 53 terms, prefix 'TA', and file type organization enabled
2026-06-24 22:30:01,782 - TorahAnytimeScraper - INFO - TorahAnytime scraper initialized: limit=400, offset=0, project_id=1, include_terms=53, exclude_terms=10
2026-06-24 22:30:01,784 - BatorahScraper - INFO - File downloader initialized with 19 terms, prefix 'BT', and file type organization enabled
2026-06-24 22:30:01,784 - BatorahScraper - INFO - Batorah scraper initialized: api_urls=2, page_size=100, include_terms=19, exclude_terms=1
2026-06-24 22:30:01,784 - KolHalashonScraper - INFO - Kol Halashon scraper initialized: base_url=https://www2.kolhalashon.com, max_shiurim=50, fetch_all_speakers=True, headless=False, exclude_filters=1, include_filters=0
2026-06-24 22:30:01,784 - __main__ - INFO - Running 3 scrapers concurrently...
2026-06-24 22:30:01,784 - GitHubScraper - INFO - Searching GitHub for terms: ['paperclip', '"Hermes Agent"', 'multica']
2026-06-24 22:30:01,880 - LinkedInFeedScraper - INFO - Fetching LinkedIn feed (25 posts)
2026-06-24 22:30:02,119 - LinkedInFeedScraper - INFO - Auto-grabbed fresh LinkedIn cookies from running Chrome (fallback).
2026-06-24 22:30:11,114 - LinkedInFeedScraper - INFO - Successfully sorted fetched LinkedIn posts chronologically (most recent first)
2026-06-24 22:30:11,123 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Moishe Goldstein - At 22, Iโ€™m often one of the youngest people in the room.
Sometimes by 10, 20 or even 30 years.
Yet
2026-06-24 22:30:11,131 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Michael Simons - Is your supply chain struggling after containers arrive at the Port of LA/Long Beach? ๐Ÿšข
The gap bet
2026-06-24 22:30:11,140 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Yisroel Wahl - Bari Azman continues to rock the show as he brings OrahVision Inc across the country.
I am tremendo
2026-06-24 22:30:11,148 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Hershy Goldstein - Running a business means learning things you never planned to learn, making decisions fast, and figu
2026-06-24 22:30:11,155 - LinkedInFeedScraper - INFO - New LinkedIn feed post: .Naftuly (Tuli) Kraus - 1 week left to go till we make Aliyah.
Can't believe I am even writing this sentence.
With hashems
2026-06-24 22:30:11,163 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Shmuel Pearlman - I saw a little kid drop his ice cream.
Complete meltdown.
Tears.
Screaming.
The whole thing.
Th
2026-06-24 22:30:11,172 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Yishai Lieser - Team work!!
Swiss Madisonยฎ
2026-06-24 22:30:11,180 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Mordy Friedman - Fantastic experience at #SAAGNY Promotions East NYC! Lots of conversations about custom branded 3D p
2026-06-24 22:30:11,242 - LinkedInFeedScraper - INFO - LinkedIn Feed scrape complete. Found 8 new posts.
2026-06-24 22:30:12,307 - httpx - INFO - HTTP Request: GET https://api.github.com/users/abeperl/received_events?per_page=100&page=1 "HTTP/1.1 200 OK"
2026-06-24 22:30:12,344 - httpx - INFO - HTTP Request: GET https://api.github.com/search/repositories?q=paperclip+in%3Aname%2Cdescription%2Creadme+stars%3A%3E150&sort=updated&order=desc&per_page=20 "HTTP/1.1 200 OK"
2026-06-24 22:30:12,370 - GitHubScraper - INFO - Found 241 GitHub repositories for term: paperclip
2026-06-24 22:30:13,532 - httpx - INFO - HTTP Request: GET https://api.github.com/search/repositories?q=%22Hermes+Agent%22+in%3Aname%2Cdescription%2Creadme+stars%3A%3E150&sort=updated&order=desc&per_page=20 "HTTP/1.1 200 OK"
2026-06-24 22:30:13,537 - GitHubScraper - INFO - Found 395 GitHub repositories for term: "Hermes Agent"
2026-06-24 22:30:14,431 - httpx - INFO - HTTP Request: GET https://api.github.com/search/repositories?q=multica+in%3Aname%2Cdescription%2Creadme+stars%3A%3E150&sort=updated&order=desc&per_page=20 "HTTP/1.1 200 OK"
2026-06-24 22:30:14,434 - GitHubScraper - INFO - Found 54 GitHub repositories for term: multica
2026-06-24 22:30:14,568 - GitHubScraper - INFO - Total GitHub results found: 0
2026-06-24 22:30:14,616 - httpx - INFO - HTTP Request: GET https://api.github.com/users/abeperl/received_events?per_page=100&page=2 "HTTP/1.1 200 OK"
2026-06-24 22:30:16,614 - httpx - INFO - HTTP Request: GET https://api.github.com/users/abeperl/received_events?per_page=100&page=3 "HTTP/1.1 200 OK"
2026-06-24 22:30:17,462 - GitHubScraper - INFO - GitHub feed: 298 fetched across pages, 1 new events for abeperl
2026-06-24 22:30:17,463 - __main__ - INFO - Scraper github completed: 0 results
2026-06-24 22:30:17,463 - __main__ - INFO - Scraper github_feed completed: 1 results
2026-06-24 22:30:17,463 - __main__ - INFO - Scraper linkedin_feed completed: 8 results
2026-06-24 22:30:17,463 - social_media_monitor.sheets.webinar_sheets_writer - INFO - Checking 8 LinkedIn feed posts for webinar/course mentions...
2026-06-24 22:30:18,421 - social_media_monitor.sheets.webinar_sheets_writer - INFO - Loaded 25 existing URLs from sheet for dedup
2026-06-24 22:30:18,422 - social_media_monitor.sheets.webinar_sheets_writer - INFO - No webinar/course mentions found in LinkedIn feed posts
2026-06-24 22:30:18,422 - __main__ - CRITICAL - Critical error in main: name 'scrapers' is not defined
Traceback (most recent call last):
File "/home/openclaw/Social_Media_Monitor/social_media_monitor.py", line 219, in main
_check_linkedin_feed_failure('linkedin_feed' in scrapers, len(results.get('linkedin_feed_results', [])))
^^^^^^^^
NameError: name 'scrapers' is not defined
2026-06-24 22:30:18,423 - social_media_monitor.database.connection - INFO - Closing all database connections
2026-06-24 22:30:18,485 - social_media_monitor.database.connection - INFO - Closed 5 database connections
2026-06-24 22:30:18,485 - __main__ - INFO - Closed all database connections
2026-06-24 22:30:18,487 - social_media_monitor.database.connection - INFO - Closing all database connections
2026-06-24 22:30:18,487 - social_media_monitor.database.connection - INFO - Closed 0 database connections
2026-06-24 23:00:02,029 - social_media_monitor.utils.logging_config - INFO - Posts log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_posts_2026-06-24.log
2026-06-24 23:00:02,029 - social_media_monitor.utils.logging_config - INFO - Downloads log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_downloads_2026-06-24.log
2026-06-24 23:00:02,029 - social_media_monitor.utils.logging_config - INFO - Logging configured with level INFO, log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_2026-06-24.log
2026-06-24 23:00:02,029 - __main__ - INFO - === Social Media Monitor v2.0 (Async) Started ===
2026-06-24 23:00:02,056 - social_media_monitor.config.settings - INFO - Configuration loaded from /home/openclaw/Social_Media_Monitor/config.yaml
2026-06-24 23:00:02,056 - __main__ - INFO - Configuration loaded and validated successfully
2026-06-24 23:00:02,058 - social_media_monitor.database.connection - INFO - Database connection pool initialized with 5 connections
2026-06-24 23:00:02,059 - social_media_monitor.database.connection - INFO - add_speaker_filters column already exists
2026-06-24 23:00:02,134 - social_media_monitor.database.connection - INFO - Database setup completed successfully
2026-06-24 23:00:02,134 - __main__ - INFO - Database initialized at: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/social_media_monitor.db
2026-06-24 23:00:02,191 - social_media_monitor.scrapers.reddit_scraper - INFO - Reddit API initialized successfully
2026-06-24 23:00:02,192 - RSSFeedScraper - INFO - RSS scraper initialized: feeds=2, html_urls=0 (0 with custom selectors), include_terms=0, exclude_terms=0
2026-06-24 23:00:02,193 - __main__ - INFO - RSS scraper initialized and enabled
2026-06-24 23:00:02,200 - Yiddish24Scraper - INFO - File downloader initialized with 55 terms, prefix 'Y2', and file type organization enabled
2026-06-24 23:00:02,200 - Yiddish24Scraper - INFO - Yiddish24 scraper initialized: data_id='latest', page_limit=100, total_pages=10, debug_html=True, include_terms=55, exclude_terms=0
2026-06-24 23:00:02,200 - __main__ - INFO - Yiddish24 scraper initialized and enabled
2026-06-24 23:00:02,202 - AhBlickLiveScraper - INFO - AhBlickLive scraper initialized: urls=1, max_items=100, include_terms=49, exclude_terms=0
2026-06-24 23:00:02,208 - TorahAnytimeScraper - INFO - File downloader initialized with 53 terms, prefix 'TA', and file type organization enabled
2026-06-24 23:00:02,208 - TorahAnytimeScraper - INFO - TorahAnytime scraper initialized: limit=400, offset=0, project_id=1, include_terms=53, exclude_terms=10
2026-06-24 23:00:02,210 - BatorahScraper - INFO - File downloader initialized with 19 terms, prefix 'BT', and file type organization enabled
2026-06-24 23:00:02,210 - BatorahScraper - INFO - Batorah scraper initialized: api_urls=2, page_size=100, include_terms=19, exclude_terms=1
2026-06-24 23:00:02,211 - KolHalashonScraper - INFO - Kol Halashon scraper initialized: base_url=https://www2.kolhalashon.com, max_shiurim=50, fetch_all_speakers=True, headless=False, exclude_filters=1, include_filters=0
2026-06-24 23:00:02,211 - __main__ - INFO - Ivelt scraper initialized and enabled
2026-06-24 23:00:02,211 - __main__ - INFO - Running 3 scrapers concurrently...
2026-06-24 23:00:02,211 - Yiddish24Scraper - INFO - Starting Yiddish24 scraping (max 10 pages, limit 100 items)
2026-06-24 23:00:02,333 - Yiddish24Scraper - INFO - Fetching Yiddish24 page 1 with data_id='latest', page_limit=100
2026-06-24 23:00:02,337 - IveltScraper - INFO - Loaded 181 include, 0 exclude filters from DB topic_filters
2026-06-24 23:00:02,337 - IveltScraper - INFO - Fetching ivelt active topics from https://www.ivelt.com/forum/search.php?search_id=active_topics...
2026-06-24 23:00:02,339 - RSSFeedScraper - INFO - Fetching RSS feed: https://www.theyeshivaworld.com/feed
2026-06-24 23:00:02,611 - RSSFeedScraper - INFO - Fetching RSS feed: https://www.thegatewaypundit.com/feed/
2026-06-24 23:00:02,975 - RSSFeedScraper - INFO - Total RSS/HTML results: 3
2026-06-24 23:00:06,268 - httpx - INFO - HTTP Request: POST https://www.yiddish24.com/ajax/cat_pagination.php "HTTP/1.1 200 OK"
2026-06-24 23:00:06,491 - Yiddish24Scraper - INFO - Received response with 490300 characters
2026-06-24 23:00:06,493 - Yiddish24Scraper - INFO - Extracted HTML from JSON result field (310012 chars)
2026-06-24 23:00:06,495 - Yiddish24Scraper - INFO - Saved HTML debug file to: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/debug/yiddish24_page1_response.html
2026-06-24 23:00:06,598 - Yiddish24Scraper - INFO - Found 100 potential article elements on page 1
2026-06-24 23:00:06,641 - Yiddish24Scraper - INFO - Page 1 extraction: 100 successful, 0 failed out of 100 elements
2026-06-24 23:00:06,641 - Yiddish24Scraper - INFO - Parsed 100 articles from HTML on page 1
2026-06-24 23:00:06,725 - Yiddish24Scraper - INFO - Term 'ืฉืžืขืœืฆืขืจ' matched in DESCRIPTION: 'ืจ' ืœื™ืคื ืฉืžืขืœืฆืขืจ'
2026-06-24 23:00:06,725 - Yiddish24Scraper - INFO - Downloading 1 media file(s) for article: ื”ื›ืœ ื‘ื›ืœ
2026-06-24 23:00:07,225 - social_media_monitor.utils.file_downloader - INFO - Starting download from: https://cloudfront.yiddish24.com/BADCHUNES_lipa_260_260624060845.mp3
2026-06-24 23:00:07,461 - social_media_monitor.utils.file_downloader - INFO - Successfully downloaded: Y2_BADCHUNES_lipa_260_260624060845.mp3 (38.2 MB)
2026-06-24 23:00:07,462 - social_media_monitor.utils.file_downloader - INFO - Organized audio file to: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/downloads/Audio (audio/mpeg)
2026-06-24 23:00:07,479 - social_media_monitor.utils.media_tagger - INFO - Tagged Y2_BADCHUNES_lipa_260_260624060845.mp3: title, artist, album, comment, source_url, genre, publisher
2026-06-24 23:00:07,479 - Yiddish24Scraper - INFO - Renamed downloaded file to include match term: Y2_BADCHUNES_lipa_260_260624060845_ืฉืžืขืœืฆืขืจ.mp3
2026-06-24 23:00:08,638 - Yiddish24Scraper - INFO - Page 1 summary: 7 new out of 100 total articles
2026-06-24 23:00:08,639 - Yiddish24Scraper - INFO - Page 1: 7/100 new items
2026-06-24 23:00:08,639 - Yiddish24Scraper - INFO - Low new item ratio (7.0%), stopping pagination
2026-06-24 23:00:08,639 - Yiddish24Scraper - INFO - Total Yiddish24 results found: 7
2026-06-24 23:00:14,243 - IveltScraper - INFO - Found 50 active topics on ivelt
2026-06-24 23:00:17,218 - IveltScraper - INFO - Filtering post containing exclude term: '*ื‘ื™ื˜ืข ืžืื›ื˜ ื–ื™ื›ืขืจ ื‘ืœื•ื™ื– ืฆื• ื‘ืืจื™ื›ื˜ืŸ ืงืจืื ื˜ืข ื ื™ื™ืขืก*'
2026-06-24 23:00:36,720 - IveltScraper - INFO - ivelt scrape complete. Checked 4 topics.
2026-06-24 23:00:36,721 - IveltScraper - INFO - All forums scrape complete. Found 16 new posts.
2026-06-24 23:00:36,721 - __main__ - INFO - Scraper rss completed: 3 results
2026-06-24 23:00:36,721 - __main__ - INFO - Scraper yiddish24 completed: 7 results
2026-06-24 23:00:36,721 - __main__ - INFO - Scraper ivelt completed: 16 results
2026-06-24 23:00:36,721 - __main__ - CRITICAL - Critical error in main: name 'scrapers' is not defined
Traceback (most recent call last):
File "/home/openclaw/Social_Media_Monitor/social_media_monitor.py", line 219, in main
_check_linkedin_feed_failure('linkedin_feed' in scrapers, len(results.get('linkedin_feed_results', [])))
^^^^^^^^
NameError: name 'scrapers' is not defined
2026-06-24 23:00:36,723 - social_media_monitor.database.connection - INFO - Closing all database connections
2026-06-24 23:00:36,768 - social_media_monitor.database.connection - INFO - Closed 5 database connections
2026-06-24 23:00:36,768 - __main__ - INFO - Closed all database connections
2026-06-24 23:00:36,770 - social_media_monitor.database.connection - INFO - Closing all database connections
2026-06-24 23:00:36,770 - social_media_monitor.database.connection - INFO - Closed 0 database connections
2026-06-24 23:30:02,326 - social_media_monitor.utils.logging_config - INFO - Posts log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_posts_2026-06-24.log
2026-06-24 23:30:02,326 - social_media_monitor.utils.logging_config - INFO - Downloads log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_downloads_2026-06-24.log
2026-06-24 23:30:02,326 - social_media_monitor.utils.logging_config - INFO - Logging configured with level INFO, log file: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/logs/social_media_monitor_2026-06-24.log
2026-06-24 23:30:02,326 - __main__ - INFO - === Social Media Monitor v2.0 (Async) Started ===
2026-06-24 23:30:02,353 - social_media_monitor.config.settings - INFO - Configuration loaded from /home/openclaw/Social_Media_Monitor/config.yaml
2026-06-24 23:30:02,354 - social_media_monitor.config.settings - WARNING - Twitter monitoring enabled but credentials validation not implemented
2026-06-24 23:30:02,354 - __main__ - INFO - Configuration loaded and validated successfully
2026-06-24 23:30:02,357 - social_media_monitor.database.connection - INFO - Database connection pool initialized with 5 connections
2026-06-24 23:30:02,357 - social_media_monitor.database.connection - INFO - add_speaker_filters column already exists
2026-06-24 23:30:02,380 - social_media_monitor.database.connection - INFO - Database setup completed successfully
2026-06-24 23:30:02,380 - __main__ - INFO - Database initialized at: /mnt/data/openclaw-shared/home/shared/Social_Media_Monitor/social_media_monitor.db
2026-06-24 23:30:02,383 - social_media_monitor.scrapers.reddit_scraper - INFO - Reddit API initialized successfully
2026-06-24 23:30:02,383 - __main__ - INFO - GitHub scraper initialized and enabled
2026-06-24 23:30:02,383 - __main__ - INFO - GitHub feed scraper initialized and enabled
2026-06-24 23:30:02,383 - RSSFeedScraper - INFO - RSS scraper initialized: feeds=2, html_urls=0 (0 with custom selectors), include_terms=0, exclude_terms=0
2026-06-24 23:30:02,384 - __main__ - INFO - LinkedIn Feed scraper initialized and enabled
2026-06-24 23:30:02,386 - Yiddish24Scraper - INFO - File downloader initialized with 55 terms, prefix 'Y2', and file type organization enabled
2026-06-24 23:30:02,388 - Yiddish24Scraper - INFO - Yiddish24 scraper initialized: data_id='latest', page_limit=100, total_pages=10, debug_html=True, include_terms=55, exclude_terms=0
2026-06-24 23:30:02,391 - AhBlickLiveScraper - INFO - AhBlickLive scraper initialized: urls=1, max_items=100, include_terms=49, exclude_terms=0
2026-06-24 23:30:02,396 - TorahAnytimeScraper - INFO - File downloader initialized with 53 terms, prefix 'TA', and file type organization enabled
2026-06-24 23:30:02,396 - TorahAnytimeScraper - INFO - TorahAnytime scraper initialized: limit=400, offset=0, project_id=1, include_terms=53, exclude_terms=10
2026-06-24 23:30:02,398 - BatorahScraper - INFO - File downloader initialized with 19 terms, prefix 'BT', and file type organization enabled
2026-06-24 23:30:02,398 - BatorahScraper - INFO - Batorah scraper initialized: api_urls=2, page_size=100, include_terms=19, exclude_terms=1
2026-06-24 23:30:02,398 - KolHalashonScraper - INFO - Kol Halashon scraper initialized: base_url=https://www2.kolhalashon.com, max_shiurim=50, fetch_all_speakers=True, headless=False, exclude_filters=1, include_filters=0
2026-06-24 23:30:02,398 - __main__ - INFO - Running 3 scrapers concurrently...
2026-06-24 23:30:02,398 - GitHubScraper - INFO - Searching GitHub for terms: ['paperclip', '"Hermes Agent"', 'multica']
2026-06-24 23:30:02,497 - LinkedInFeedScraper - INFO - Fetching LinkedIn feed (25 posts)
2026-06-24 23:30:02,812 - LinkedInFeedScraper - INFO - Auto-grabbed fresh LinkedIn cookies from running Chrome (fallback).
2026-06-24 23:30:12,772 - LinkedInFeedScraper - INFO - Successfully sorted fetched LinkedIn posts chronologically (most recent first)
2026-06-24 23:30:12,781 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Yaakov Weiner - Need a CFO, Controller, or VP of Finance?
Weโ€™ve got em.
Our pipeline is filled with elite, experie
2026-06-24 23:30:12,795 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Jewish Professionals - Rabbi Shlomo Ezagui presents a thought-provoking exploration of a 1953 Chassidic discourse by the Re
2026-06-24 23:30:12,861 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Driscoll's Locksmith - Closing in on 200 โญ๏ธโญ๏ธโญ๏ธโญ๏ธโญ๏ธ reviews
Itโ€™s been a journey!
These are all actual customer reviews,
2026-06-24 23:30:12,873 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Mike Rosenberg - "Good to see you - btw I love your LinkedIn posts" - some MFer who has been following me for 3 years
2026-06-24 23:30:12,885 - LinkedInFeedScraper - INFO - New LinkedIn feed post: Jinjing Liang - Weโ€™re trending on GitHub!!
In < 100 days, weโ€™ve grown from 0 to 6.7k stars ๐ŸŒŸ
0 paid ads, no Hack
2026-06-24 23:30:12,897 - LinkedInFeedScraper - INFO - LinkedIn Feed scrape complete. Found 5 new posts.
2026-06-24 23:30:14,092 - httpx - INFO - HTTP Request: GET https://api.github.com/search/repositories?q=paperclip+in%3Aname%2Cdescription%2Creadme+stars%3A%3E150&sort=updated&order=desc&per_page=20 "HTTP/1.1 200 OK"
2026-06-24 23:30:14,117 - GitHubScraper - INFO - Found 241 GitHub repositories for term: paperclip
2026-06-24 23:30:14,256 - httpx - INFO - HTTP Request: GET https://api.github.com/users/abeperl/received_events?per_page=100&page=1 "HTTP/1.1 200 OK"
2026-06-24 23:30:15,217 - httpx - INFO - HTTP Request: GET https://api.github.com/search/repositories?q=%22Hermes+Agent%22+in%3Aname%2Cdescription%2Creadme+stars%3A%3E150&sort=updated&order=desc&per_page=20 "HTTP/1.1 200 OK"
2026-06-24 23:30:15,222 - GitHubScraper - INFO - Found 395 GitHub repositories for term: "Hermes Agent"
2026-06-24 23:30:16,166 - httpx - INFO - HTTP Request: GET https://api.github.com/users/abeperl/received_events?per_page=100&page=2 "HTTP/1.1 200 OK"
2026-06-24 23:30:17,131 - httpx - INFO - HTTP Request: GET https://api.github.com/search/repositories?q=multica+in%3Aname%2Cdescription%2Creadme+stars%3A%3E150&sort=updated&order=desc&per_page=20 "HTTP/1.1 200 OK"
2026-06-24 23:30:17,136 - GitHubScraper - INFO - Found 54 GitHub repositories for term: multica
2026-06-24 23:30:17,278 - GitHubScraper - INFO - Total GitHub results found: 0
2026-06-24 23:30:18,148 - httpx - INFO - HTTP Request: GET https://api.github.com/users/abeperl/received_events?per_page=100&page=3 "HTTP/1.1 200 OK"
2026-06-24 23:30:19,332 - GitHubScraper - INFO - GitHub feed: 298 fetched across pages, 0 new events for abeperl
2026-06-24 23:30:19,333 - __main__ - INFO - Scraper github completed: 0 results
2026-06-24 23:30:19,333 - __main__ - INFO - Scraper github_feed completed: 0 results
2026-06-24 23:30:19,333 - __main__ - INFO - Scraper linkedin_feed completed: 5 results
2026-06-24 23:30:19,333 - social_media_monitor.sheets.webinar_sheets_writer - INFO - Checking 5 LinkedIn feed posts for webinar/course mentions...
2026-06-24 23:30:21,405 - social_media_monitor.sheets.webinar_sheets_writer - INFO - Loaded 25 existing URLs from sheet for dedup
2026-06-24 23:30:21,405 - social_media_monitor.sheets.webinar_sheets_writer - INFO - Webinar/course match: Jewish Professionals - keyword 'course' - https://www.linkedin.com/feed/update/urn:li:activity:7475685282277801984
2026-06-24 23:30:21,405 - social_media_monitor.sheets.webinar_sheets_writer - INFO - Webinar/course match: Driscoll's Locksmith - keyword 'course' - https://www.linkedin.com/feed/update/urn:li:activity:7475683678015389696
2026-06-24 23:30:22,417 - social_media_monitor.sheets.webinar_sheets_writer - INFO - Appended 2 rows to Google Sheet 'Webinar Recording DMs'
2026-06-24 23:30:22,417 - __main__ - INFO - Webinar/Course: appended 2 posts to Google Sheet
2026-06-24 23:30:22,417 - __main__ - CRITICAL - Critical error in main: name 'scrapers' is not defined
Traceback (most recent call last):
File "/home/openclaw/Social_Media_Monitor/social_media_monitor.py", line 219, in main
_check_linkedin_feed_failure('linkedin_feed' in scrapers, len(results.get('linkedin_feed_results', [])))
^^^^^^^^
NameError: name 'scrapers' is not defined
2026-06-24 23:30:22,419 - social_media_monitor.database.connection - INFO - Closing all database connections
2026-06-24 23:30:22,482 - social_media_monitor.database.connection - INFO - Closed 5 database connections
2026-06-24 23:30:22,482 - __main__ - INFO - Closed all database connections
2026-06-24 23:30:22,484 - social_media_monitor.database.connection - INFO - Closing all database connections
2026-06-24 23:30:22,485 - social_media_monitor.database.connection - INFO - Closed 0 database connections