{"repo":"ThomsenDrake/ainews-source-extractor","free":true,"listed":false,"github":"https://github.com/ThomsenDrake/ainews-source-extractor","clone":"git clone https://github.com/ThomsenDrake/ainews-source-extractor.git","description":"Extracts all URLs from the most recent AI News issue (from news.smol.ai) and prepares them for seamless import into Google's NotebookLM","language":"Python","stars":10,"topics":["automation","markdown","notebooklm","python","web-scraping"],"license":"MIT","category":"scrapers-browser-automation","readme_excerpt":"AI News Source Extractor Description AI News Link Scraper extracts all URLs from the most recent AI News issue (from news.smol.ai) and prepares them for seamless import into Google's NotebookLM. It organizes sources into a dedicated folder, separates non-social URLs into a sources.txt , and generates individual markdown files for quoted tweet content. Features Folder Generation: Creates a timestamped folder for each issue’s sources. sources.txt: Lists all URLs from the issue, excluding twitter.com , x.com , and discord.com . Tweet Markdown: Saves the full text of each quoted tweet as a separate markdown file. WebSync Ready: sources.txt can be pasted directly into the WebSync for NotebookLM Chrome extension to auto-import into NotebookLM. Installation Usage Simply run the main scraper: This will: 1. Generate a folder named with the current date for the latest AI News issue. 2. Create sources.txt inside that folder, containing all non-social URLs. 3. Produce individual .md files for each tweet quoted in the issue. Roadmap Improve URL-filtering logic to separate twitter.com , x.com , and discord.com links. Build discord scraper.py to fetch and save referenced Discord messages as markdown. Parameterize the output folder path and issue source URL for greater flexibility. Contributing Contributions welcome! Fork, branch, and submit a pull request.","default_branch":null,"files":null,"tree":[],"storefront":"/r/ThomsenDrake","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/ThomsenDrake/ainews-source-extractor/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}