<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0">
  <channel>
    <title>extractfeed</title>
    <link>https://extractfeed.io/</link>
    <description>A rolling, agent-readable changefeed for web scraping and data extraction news.</description>
    <item>
      <title>Firecrawl releases official ChatGPT plugin</title>
      <link>https://extractfeed.io/story/firecrawl-releases-official-chatgpt-plugin-595c490/</link>
      <guid>https://extractfeed.io/story/firecrawl-releases-official-chatgpt-plugin-595c490/</guid>
      <pubDate>Sun, 06 Sep 2026 15:05:26 GMT</pubDate>
      <description>Firecrawl has launched an official ChatGPT plugin, enabling users to extract web data directly through conversational AI.</description>
    </item>
    <item>
      <title>Firecrawl introduces Anydoc and PDF Inspector for document extraction</title>
      <link>https://extractfeed.io/story/firecrawl-introduces-anydoc-and-pdf-inspector-for-document-e-3cefcd2/</link>
      <guid>https://extractfeed.io/story/firecrawl-introduces-anydoc-and-pdf-inspector-for-document-e-3cefcd2/</guid>
      <pubDate>Sun, 06 Sep 2026 15:05:21 GMT</pubDate>
      <description>Firecrawl announced two new tools, Anydoc and PDF Inspector, aimed at improving document and PDF data extraction.</description>
    </item>
    <item>
      <title>Firecrawl introduces Agentic OCR for structured data extraction from images</title>
      <link>https://extractfeed.io/story/firecrawl-introduces-agentic-ocr-for-structured-data-extract-eae1acb/</link>
      <guid>https://extractfeed.io/story/firecrawl-introduces-agentic-ocr-for-structured-data-extract-eae1acb/</guid>
      <pubDate>Sun, 06 Sep 2026 15:05:07 GMT</pubDate>
      <description>Firecrawl has released a new OCR feature that uses AI agents to extract structured data from images and documents.</description>
    </item>
    <item>
      <title>Firecrawl launches Research Index for life sciences</title>
      <link>https://extractfeed.io/story/firecrawl-launches-research-index-for-life-sciences-1f24346/</link>
      <guid>https://extractfeed.io/story/firecrawl-launches-research-index-for-life-sciences-1f24346/</guid>
      <pubDate>Sun, 06 Sep 2026 15:05:01 GMT</pubDate>
      <description>Firecrawl announced the launch of a Research Index tailored for the life sciences domain.</description>
    </item>
    <item>
      <title>Firecrawl releases Convex component for web scraping integration</title>
      <link>https://extractfeed.io/story/firecrawl-releases-convex-component-for-web-scraping-integra-86b90c0/</link>
      <guid>https://extractfeed.io/story/firecrawl-releases-convex-component-for-web-scraping-integra-86b90c0/</guid>
      <pubDate>Sun, 06 Sep 2026 15:04:56 GMT</pubDate>
      <description>Firecrawl announced a new Convex component that allows developers to integrate web scraping capabilities directly into their Convex applications.</description>
    </item>
    <item>
      <title>Firecrawl announces integration with Eden AI</title>
      <link>https://extractfeed.io/story/firecrawl-announces-integration-with-eden-ai-33f5214/</link>
      <guid>https://extractfeed.io/story/firecrawl-announces-integration-with-eden-ai-33f5214/</guid>
      <pubDate>Sun, 06 Sep 2026 15:04:54 GMT</pubDate>
      <description>Firecrawl has published a blog post announcing an integration with Eden AI.</description>
    </item>
    <item>
      <title>Firecrawl publishes case study on 11x's use of its web scraping API for prospect research</title>
      <link>https://extractfeed.io/story/firecrawl-publishes-case-study-on-11x-s-use-of-its-web-scrap-9f9b4c2/</link>
      <guid>https://extractfeed.io/story/firecrawl-publishes-case-study-on-11x-s-use-of-its-web-scrap-9f9b4c2/</guid>
      <pubDate>Sun, 06 Sep 2026 15:04:32 GMT</pubDate>
      <description>Firecrawl's blog details how sales automation startup 11x uses its web scraping API to automate prospect research.</description>
    </item>
    <item>
      <title>Firecrawl launches Developer Index for web scraping performance metrics</title>
      <link>https://extractfeed.io/story/firecrawl-launches-developer-index-for-web-scraping-performa-5aa37fc/</link>
      <guid>https://extractfeed.io/story/firecrawl-launches-developer-index-for-web-scraping-performa-5aa37fc/</guid>
      <pubDate>Sun, 06 Sep 2026 15:04:29 GMT</pubDate>
      <description>Firecrawl announced the launch of a Developer Index to provide benchmarks and performance data for web scraping tools.</description>
    </item>
    <item>
      <title>Browserless publishes enterprise Docker deployment guide for self-hosting</title>
      <link>https://extractfeed.io/story/browserless-publishes-enterprise-docker-deployment-guide-for-cb637d4/</link>
      <guid>https://extractfeed.io/story/browserless-publishes-enterprise-docker-deployment-guide-for-cb637d4/</guid>
      <pubDate>Sun, 06 Sep 2026 15:04:22 GMT</pubDate>
      <description>Browserless released a guide detailing how to deploy its browser automation service in an enterprise Docker environment.</description>
    </item>
    <item>
      <title>Browserless introduces Skill Bucket for agent-based browser automation</title>
      <link>https://extractfeed.io/story/browserless-introduces-skill-bucket-for-agent-based-browser-aabceab/</link>
      <guid>https://extractfeed.io/story/browserless-introduces-skill-bucket-for-agent-based-browser-aabceab/</guid>
      <pubDate>Sun, 06 Sep 2026 14:43:28 GMT</pubDate>
      <description>Browserless announces a new Skill Bucket feature that allows AI agents to access and use predefined browser automation skills.</description>
    </item>
    <item>
      <title>Browserless blog post examines the pitfalls of persisting browser profiles</title>
      <link>https://extractfeed.io/story/browserless-blog-post-examines-the-pitfalls-of-persisting-br-219e98a/</link>
      <guid>https://extractfeed.io/story/browserless-blog-post-examines-the-pitfalls-of-persisting-br-219e98a/</guid>
      <pubDate>Sun, 06 Sep 2026 14:43:25 GMT</pubDate>
      <description>Browserless published a blog post discussing the challenges and drawbacks of persisting browser profiles in automated environments.</description>
    </item>
    <item>
      <title>Browserless publishes guide on browser infrastructure for computer use agents</title>
      <link>https://extractfeed.io/story/browserless-publishes-guide-on-browser-infrastructure-for-co-2a7a8a2/</link>
      <guid>https://extractfeed.io/story/browserless-publishes-guide-on-browser-infrastructure-for-co-2a7a8a2/</guid>
      <pubDate>Sun, 06 Sep 2026 14:43:20 GMT</pubDate>
      <description>Browserless has published a blog post discussing the infrastructure requirements for running browser-based agents that interact with web interfaces on behalf of users.</description>
    </item>
    <item>
      <title>Browserless publishes guide on bypassing Datadome anti-bot protection</title>
      <link>https://extractfeed.io/story/browserless-publishes-guide-on-bypassing-datadome-anti-bot-p-eac61a5/</link>
      <guid>https://extractfeed.io/story/browserless-publishes-guide-on-bypassing-datadome-anti-bot-p-eac61a5/</guid>
      <pubDate>Sun, 06 Sep 2026 14:43:06 GMT</pubDate>
      <description>Browserless released a blog post detailing techniques to circumvent Datadome's anti-bot system.</description>
    </item>
    <item>
      <title>Browserless introduces the Browser Automation Protocol</title>
      <link>https://extractfeed.io/story/browserless-introduces-the-browser-automation-protocol-6bba0fd/</link>
      <guid>https://extractfeed.io/story/browserless-introduces-the-browser-automation-protocol-6bba0fd/</guid>
      <pubDate>Sun, 06 Sep 2026 14:42:49 GMT</pubDate>
      <description>Browserless has announced a new protocol for browser automation.</description>
    </item>
    <item>
      <title>Browserless launches Agent for AI-driven browser automation</title>
      <link>https://extractfeed.io/story/browserless-launches-agent-for-ai-driven-browser-automation-ccdffa4/</link>
      <guid>https://extractfeed.io/story/browserless-launches-agent-for-ai-driven-browser-automation-ccdffa4/</guid>
      <pubDate>Sun, 06 Sep 2026 14:42:44 GMT</pubDate>
      <description>Browserless has introduced a new product called Browserless Agent, designed to enable AI agents to control browser sessions.</description>
    </item>
    <item>
      <title>Bright Data Blog discusses robot training data from public web video</title>
      <link>https://extractfeed.io/story/bright-data-blog-discusses-robot-training-data-from-public-w-9da3ca9/</link>
      <guid>https://extractfeed.io/story/bright-data-blog-discusses-robot-training-data-from-public-w-9da3ca9/</guid>
      <pubDate>Sun, 06 Sep 2026 14:42:39 GMT</pubDate>
      <description>Bright Data's blog explores the use of publicly available web video as a source of training data for robotic AI systems.</description>
    </item>
    <item>
      <title>Bright Data integrates with Paperclip for AI-driven data extraction</title>
      <link>https://extractfeed.io/story/bright-data-integrates-with-paperclip-for-ai-driven-data-ext-f0a45a6/</link>
      <guid>https://extractfeed.io/story/bright-data-integrates-with-paperclip-for-ai-driven-data-ext-f0a45a6/</guid>
      <pubDate>Sun, 06 Sep 2026 14:42:37 GMT</pubDate>
      <description>Bright Data announces a partnership or integration with Paperclip, an AI tool, as detailed in a blog post on their site.</description>
    </item>
    <item>
      <title>Bright Data Defends Its Network Against Misuse Claims</title>
      <link>https://extractfeed.io/story/bright-data-defends-its-network-against-misuse-claims-310a2e7/</link>
      <guid>https://extractfeed.io/story/bright-data-defends-its-network-against-misuse-claims-310a2e7/</guid>
      <pubDate>Sun, 06 Sep 2026 14:42:32 GMT</pubDate>
      <description>Bright Data publishes a blog post arguing that its proxy network cannot be used in the ways critics allege.</description>
    </item>
    <item>
      <title>Bright Data explains Context as a Service for AI data retrieval</title>
      <link>https://extractfeed.io/story/bright-data-explains-context-as-a-service-for-ai-data-retrie-10366cc/</link>
      <guid>https://extractfeed.io/story/bright-data-explains-context-as-a-service-for-ai-data-retrie-10366cc/</guid>
      <pubDate>Sun, 06 Sep 2026 14:42:26 GMT</pubDate>
      <description>Bright Data published a blog post defining Context as a Service, a concept where external context is provided to AI models via structured data feeds.</description>
    </item>
    <item>
      <title>Bright Data compares its Cursor integration with default coding agent</title>
      <link>https://extractfeed.io/story/bright-data-compares-its-cursor-integration-with-default-cod-3f18756/</link>
      <guid>https://extractfeed.io/story/bright-data-compares-its-cursor-integration-with-default-cod-3f18756/</guid>
      <pubDate>Sun, 06 Sep 2026 14:42:11 GMT</pubDate>
      <description>Bright Data published a blog post comparing its Cursor integration against the default coding agent for web data tasks.</description>
    </item>
    <item>
      <title>Playwright 1.63.0 adds test locks for safe concurrent access to shared resources</title>
      <link>https://extractfeed.io/story/playwright-1-63-0-adds-test-locks-for-safe-concurrent-access-2738b90/</link>
      <guid>https://extractfeed.io/story/playwright-1-63-0-adds-test-locks-for-safe-concurrent-access-2738b90/</guid>
      <pubDate>Sat, 05 Sep 2026 14:56:32 GMT</pubDate>
      <description>Playwright 1.63.0 introduces named test locks that prevent concurrent execution of tests sharing the same lock name across files, workers, and projects.</description>
    </item>
    <item>
      <title>Apify outlines six methods for downloading Instagram images in 2026</title>
      <link>https://extractfeed.io/story/apify-outlines-six-methods-for-downloading-instagram-images-451ee85/</link>
      <guid>https://extractfeed.io/story/apify-outlines-six-methods-for-downloading-instagram-images-451ee85/</guid>
      <pubDate>Sat, 05 Sep 2026 14:56:49 GMT</pubDate>
      <description>Apify published a guide covering six techniques for downloading Instagram images, ranging from a simple browser trick to a two-Actor pipeline for full profile extraction.</description>
    </item>
    <item>
      <title>SerpApi publishes guide for scraping Zillow listings</title>
      <link>https://extractfeed.io/story/serpapi-publishes-guide-for-scraping-zillow-listings-9d6de1a/</link>
      <guid>https://extractfeed.io/story/serpapi-publishes-guide-for-scraping-zillow-listings-9d6de1a/</guid>
      <pubDate>Sat, 05 Sep 2026 14:56:53 GMT</pubDate>
      <description>SerpApi released a tutorial showing how to scrape Zillow real estate listings using its API with multiple programming languages.</description>
    </item>
    <item>
      <title>ScrapingBee compares eight top Python web scraping tools for 2026</title>
      <link>https://extractfeed.io/story/scrapingbee-compares-eight-top-python-web-scraping-tools-for-2b8ce4b/</link>
      <guid>https://extractfeed.io/story/scrapingbee-compares-eight-top-python-web-scraping-tools-for-2b8ce4b/</guid>
      <pubDate>Sat, 05 Sep 2026 14:56:58 GMT</pubDate>
      <description>ScrapingBee published a guide comparing eight Python web scraping tools, covering parsers, browser automation, crawling, proxies, and full-stack scraping services.</description>
    </item>
    <item>
      <title>Puppeteer 25.10.0 adds video-stream screen recording and Firefox 155.0 support</title>
      <link>https://extractfeed.io/story/puppeteer-25-10-0-adds-video-stream-screen-recording-and-fir-9409190/</link>
      <guid>https://extractfeed.io/story/puppeteer-25-10-0-adds-video-stream-screen-recording-and-fir-9409190/</guid>
      <pubDate>Sat, 05 Sep 2026 14:57:08 GMT</pubDate>
      <description>Puppeteer released version 25.10.0 of puppeteer-core, introducing a video-stream-based screen recording feature via page.record() and rolling to Firefox 155.0.</description>
    </item>
    <item>
      <title>Stagehand Python adds WebMCP tool support inside iframes</title>
      <link>https://extractfeed.io/story/stagehand-python-adds-webmcp-tool-support-inside-iframes-d633a05/</link>
      <guid>https://extractfeed.io/story/stagehand-python-adds-webmcp-tool-support-inside-iframes-d633a05/</guid>
      <pubDate>Sat, 05 Sep 2026 14:57:13 GMT</pubDate>
      <description>Stagehand Python's latest dev release enables discovery and invocation of WebMCP tools registered inside iframes by routing calls through the main CDP session and all adopted OOPIF sessions.</description>
    </item>
    <item>
      <title>Apify streamlines Actor creation with one-click Git repo setup</title>
      <link>https://extractfeed.io/story/apify-streamlines-actor-creation-with-one-click-git-repo-set-7f456bc/</link>
      <guid>https://extractfeed.io/story/apify-streamlines-actor-creation-with-one-click-git-repo-set-7f456bc/</guid>
      <pubDate>Sat, 05 Sep 2026 14:57:15 GMT</pubDate>
      <description>Apify now lets users select a Git provider during Actor setup, automatically creating a repository, pushing template code, and configuring builds on every push.</description>
    </item>
    <item>
      <title>ScrapingBee explains CAPTCHA solvers and when to avoid them</title>
      <link>https://extractfeed.io/story/scrapingbee-explains-captcha-solvers-and-when-to-avoid-them-80921e5/</link>
      <guid>https://extractfeed.io/story/scrapingbee-explains-captcha-solvers-and-when-to-avoid-them-80921e5/</guid>
      <pubDate>Sat, 05 Sep 2026 14:57:28 GMT</pubDate>
      <description>ScrapingBee published an article defining CAPTCHA solvers, how they automate challenge responses, and when developers should skip them in scraping workflows.</description>
    </item>
    <item>
      <title>ScrapingBee explains how CodeWhale gives AI agents live web access</title>
      <link>https://extractfeed.io/story/scrapingbee-explains-how-codewhale-gives-ai-agents-live-web-b2bf90a/</link>
      <guid>https://extractfeed.io/story/scrapingbee-explains-how-codewhale-gives-ai-agents-live-web-b2bf90a/</guid>
      <pubDate>Sat, 05 Sep 2026 14:57:24 GMT</pubDate>
      <description>ScrapingBee publishes an article detailing how CodeWhale enables AI agents to access live web data by integrating web scraping tools, MCP servers, or framework tools.</description>
    </item>
    <item>
      <title>ScrapingBee ranks top proxies for Amazon scraping in 2026</title>
      <link>https://extractfeed.io/story/scrapingbee-ranks-top-proxies-for-amazon-scraping-in-2026-a0e35f9/</link>
      <guid>https://extractfeed.io/story/scrapingbee-ranks-top-proxies-for-amazon-scraping-in-2026-a0e35f9/</guid>
      <pubDate>Sat, 05 Sep 2026 14:57:18 GMT</pubDate>
      <description>ScrapingBee published a guide evaluating residential, ISP, and mobile proxies for Amazon scraping, noting that Amazon's anti-bot updates can render previously effective proxies obsolete.</description>
    </item>
    <item>
      <title>Stagehand SDK exposes Browserbase Search and Fetch across TypeScript, Python, and Go</title>
      <link>https://extractfeed.io/story/stagehand-sdk-exposes-browserbase-search-and-fetch-across-ty-f686c14/</link>
      <guid>https://extractfeed.io/story/stagehand-sdk-exposes-browserbase-search-and-fetch-across-ty-f686c14/</guid>
      <pubDate>Sat, 05 Sep 2026 14:57:31 GMT</pubDate>
      <description>Stagehand SDK version 4.1.0a0.dev1494 adds Browserbase Search and Fetch APIs to its TypeScript and Python facades, with equivalent Go support via the existing HTTP transport.</description>
    </item>
    <item>
      <title>Zyte tests Claude Fable 5.1 and GLM-5.3-Flash in a live extraction benchmark</title>
      <link>https://extractfeed.io/story/zyte-tests-claude-fable-5-1-and-glm-5-3-flash-in-a-live-extr-ed217db/</link>
      <guid>https://extractfeed.io/story/zyte-tests-claude-fable-5-1-and-glm-5-3-flash-in-a-live-extr-ed217db/</guid>
      <pubDate>Sat, 05 Sep 2026 14:57:37 GMT</pubDate>
      <description>Zyte published a personal benchmark comparing Claude Fable 5.1 and GLM-5.3-Flash on real extraction tasks, revealing that the GLM model matched a model the author had previously encountered.</description>
    </item>
    <item>
      <title>Zyte analysis finds 75% of top sites use robots.txt, but few name specific crawlers</title>
      <link>https://extractfeed.io/story/zyte-analysis-finds-75-of-top-sites-use-robots-txt-but-few-n-e5e97be/</link>
      <guid>https://extractfeed.io/story/zyte-analysis-finds-75-of-top-sites-use-robots-txt-but-few-n-e5e97be/</guid>
      <pubDate>Sat, 05 Sep 2026 14:57:46 GMT</pubDate>
      <description>Zyte published a study showing that three-quarters of the world's top websites publish a robots.txt file, yet most do not name individual crawlers.</description>
    </item>
    <item>
      <title>ScrapingBee publishes guide to AI agent web scraping with wigolo and MCP</title>
      <link>https://extractfeed.io/story/scrapingbee-publishes-guide-to-ai-agent-web-scraping-with-wi-42dda92/</link>
      <guid>https://extractfeed.io/story/scrapingbee-publishes-guide-to-ai-agent-web-scraping-with-wi-42dda92/</guid>
      <pubDate>Sat, 05 Sep 2026 14:57:54 GMT</pubDate>
      <description>ScrapingBee released a guide covering how to equip AI agents with web scraping capabilities using wigolo and the Model Context Protocol.</description>
    </item>
    <item>
      <title>ScrapingBee publishes guide on scraping website text for LLM training</title>
      <link>https://extractfeed.io/story/scrapingbee-publishes-guide-on-scraping-website-text-for-llm-dca44ea/</link>
      <guid>https://extractfeed.io/story/scrapingbee-publishes-guide-on-scraping-website-text-for-llm-dca44ea/</guid>
      <pubDate>Sat, 05 Sep 2026 14:57:51 GMT</pubDate>
      <description>ScrapingBee released a tutorial covering how to extract all text from a website for use in LLM training pipelines.</description>
    </item>
    <item>
      <title>Zyte adds CDP support for browser automation on its infrastructure</title>
      <link>https://extractfeed.io/story/zyte-adds-cdp-support-for-browser-automation-on-its-infrastr-b858bca/</link>
      <guid>https://extractfeed.io/story/zyte-adds-cdp-support-for-browser-automation-on-its-infrastr-b858bca/</guid>
      <pubDate>Sat, 05 Sep 2026 14:58:00 GMT</pubDate>
      <description>Zyte has introduced Chrome DevTools Protocol (CDP) support, allowing users to run browser automation scripts on Zyte's managed infrastructure.</description>
    </item>
    <item>
      <title>Apify launches Apartments.com scraper for rental market analysis</title>
      <link>https://extractfeed.io/story/apify-launches-apartments-com-scraper-for-rental-market-anal-4ea2e9e/</link>
      <guid>https://extractfeed.io/story/apify-launches-apartments-com-scraper-for-rental-market-anal-4ea2e9e/</guid>
      <pubDate>Sat, 05 Sep 2026 14:58:02 GMT</pubDate>
      <description>Apify released a tool to scrape Apartments.com listings at scale and integrate them with ChatGPT for interactive market reports.</description>
    </item>
    <item>
      <title>SerpApi countersues Reddit over API access restrictions</title>
      <link>https://extractfeed.io/story/serpapi-countersues-reddit-over-api-access-restrictions-dfab48d/</link>
      <guid>https://extractfeed.io/story/serpapi-countersues-reddit-over-api-access-restrictions-dfab48d/</guid>
      <pubDate>Sat, 05 Sep 2026 14:58:05 GMT</pubDate>
      <description>SerpApi has filed counterclaims against Reddit in response to Reddit's lawsuit, alleging broken promises of an open internet and unfair API pricing.</description>
    </item>
    <item>
      <title>Scrapfly reviews nine Scrapy extensions and middlewares for 2026</title>
      <link>https://extractfeed.io/story/scrapfly-reviews-nine-scrapy-extensions-and-middlewares-for-bd62997/</link>
      <guid>https://extractfeed.io/story/scrapfly-reviews-nine-scrapy-extensions-and-middlewares-for-bd62997/</guid>
      <pubDate>Sat, 05 Sep 2026 14:58:11 GMT</pubDate>
      <description>Scrapfly tested nine Scrapy extensions and middlewares against Scrapy 2.18, covering rendering, TLS fingerprints, proxies, extraction, shared queues, and deployment.</description>
    </item>
    <item>
      <title>Scrapfly publishes 2026 diagnostic guide for blocked Scrapy spiders</title>
      <link>https://extractfeed.io/story/scrapfly-publishes-2026-diagnostic-guide-for-blocked-scrapy-6ae174b/</link>
      <guid>https://extractfeed.io/story/scrapfly-publishes-2026-diagnostic-guide-for-blocked-scrapy-6ae174b/</guid>
      <pubDate>Sat, 05 Sep 2026 14:58:14 GMT</pubDate>
      <description>Scrapfly released a walkthrough covering nine mechanisms that can block a Scrapy spider, from IP reputation to CAPTCHA, with evidence and mitigation steps for each.</description>
    </item>
    <item>
      <title>Scrapingdog launches MCP Server for AI-powered web scraping</title>
      <link>https://extractfeed.io/story/scrapingdog-launches-mcp-server-for-ai-powered-web-scraping-89825de/</link>
      <guid>https://extractfeed.io/story/scrapingdog-launches-mcp-server-for-ai-powered-web-scraping-89825de/</guid>
      <pubDate>Sat, 05 Sep 2026 14:58:17 GMT</pubDate>
      <description>Scrapingdog has introduced an MCP Server that integrates its web scraping and data extraction capabilities directly into AI-powered applications.</description>
    </item>
    <item>
      <title>SerpApi publishes roundup of best web scraping tools for 2026</title>
      <link>https://extractfeed.io/story/serpapi-publishes-roundup-of-best-web-scraping-tools-for-202-f437a05/</link>
      <guid>https://extractfeed.io/story/serpapi-publishes-roundup-of-best-web-scraping-tools-for-202-f437a05/</guid>
      <pubDate>Sat, 05 Sep 2026 14:58:25 GMT</pubDate>
      <description>SerpApi has published a guide covering popular web scraping tools, from open-source frameworks like Scrapy and Crawlee to commercial scraping platforms and search APIs.</description>
    </item>
    <item>
      <title>Builder spotlight: Goldmine automated outreach and won on Apify</title>
      <link>https://extractfeed.io/story/builder-spotlight-goldmine-automated-outreach-and-won-on-api-12278f3/</link>
      <guid>https://extractfeed.io/story/builder-spotlight-goldmine-automated-outreach-and-won-on-api-12278f3/</guid>
      <pubDate>Sat, 05 Sep 2026 14:58:29 GMT</pubDate>
      <description>Apify profiles a developer who built a LinkedIn scraper for his team and later became a top-rated Apify developer, winning an EMEA prize in the Apify $1 Million Challenge.</description>
    </item>
    <item>
      <title>Zenrows publishes practical guide on web data for LLM fine-tuning</title>
      <link>https://extractfeed.io/story/zenrows-publishes-practical-guide-on-web-data-for-llm-fine-t-3e46095/</link>
      <guid>https://extractfeed.io/story/zenrows-publishes-practical-guide-on-web-data-for-llm-fine-t-3e46095/</guid>
      <pubDate>Sat, 05 Sep 2026 14:58:40 GMT</pubDate>
      <description>Zenrows outlines how to source clean web data for LLM fine-tuning, warning that a single bad extraction can measurably degrade a small seed set.</description>
    </item>
    <item>
      <title>Zenrows blog walks through scraping 2026 FIFA World Cup data across three access patterns</title>
      <link>https://extractfeed.io/story/zenrows-blog-walks-through-scraping-2026-fifa-world-cup-data-31c6573/</link>
      <guid>https://extractfeed.io/story/zenrows-blog-walks-through-scraping-2026-fifa-world-cup-data-31c6573/</guid>
      <pubDate>Sat, 05 Sep 2026 14:58:35 GMT</pubDate>
      <description>Zenrows published a tutorial demonstrating how to scrape 2026 FIFA World Cup data from a JSON endpoint behind JavaScript, an API with per-session JWT, and server-rendered HTML using Python and its own scraping API.</description>
    </item>
    <item>
      <title>Crawlbase shows how to build a web scraping pipeline with Zapier using async callbacks</title>
      <link>https://extractfeed.io/story/crawlbase-shows-how-to-build-a-web-scraping-pipeline-with-za-402ff8c/</link>
      <guid>https://extractfeed.io/story/crawlbase-shows-how-to-build-a-web-scraping-pipeline-with-za-402ff8c/</guid>
      <pubDate>Sat, 05 Sep 2026 14:58:32 GMT</pubDate>
      <description>Crawlbase published a guide explaining how to split web retrieval from Zapier's execution window by dispatching async crawls with a callback and catching the finished page in a second Zap.</description>
    </item>
    <item>
      <title>Google introduces /goto redirect URLs in Search, SerpApi works on resolution</title>
      <link>https://extractfeed.io/story/google-introduces-goto-redirect-urls-in-search-serpapi-works-872fc85/</link>
      <guid>https://extractfeed.io/story/google-introduces-goto-redirect-urls-in-search-serpapi-works-872fc85/</guid>
      <pubDate>Sat, 05 Sep 2026 14:58:42 GMT</pubDate>
      <description>Google is rolling out new /goto redirect URLs across Search, and SerpApi is actively resolving affected links as the implementation evolves.</description>
    </item>
    <item>
      <title>Stagehand Python 4.0.3a0.dev1483 adds stagehand_facade tool surface for eval benchmarking</title>
      <link>https://extractfeed.io/story/stagehand-python-4-0-3a0-dev1483-adds-stagehand-facade-tool-0c6ec60/</link>
      <guid>https://extractfeed.io/story/stagehand-python-4-0-3a0-dev1483-adds-stagehand-facade-tool-0c6ec60/</guid>
      <pubDate>Sat, 05 Sep 2026 14:58:45 GMT</pubDate>
      <description>A dev release of Stagehand Python introduces a facade tool surface that allows evals to benchmark the exact byte-identical interface shipped to Claude Code, Codex, and Pi integrations.</description>
    </item>
    <item>
      <title>SerpApi publishes guide on scraping Walmart product reviews with its dedicated API</title>
      <link>https://extractfeed.io/story/serpapi-publishes-guide-on-scraping-walmart-product-reviews-b3a3b9a/</link>
      <guid>https://extractfeed.io/story/serpapi-publishes-guide-on-scraping-walmart-product-reviews-b3a3b9a/</guid>
      <pubDate>Sat, 05 Sep 2026 14:58:48 GMT</pubDate>
      <description>SerpApi released a tutorial showing how to use its Walmart Product Reviews API to extract ratings, review text, feedback counts, and reviewer details in structured JSON and export to CSV.</description>
    </item>
    <item>
      <title>SerpApi explains HTTP 429 errors and rate-limit best practices</title>
      <link>https://extractfeed.io/story/serpapi-explains-http-429-errors-and-rate-limit-best-practic-284d4e5/</link>
      <guid>https://extractfeed.io/story/serpapi-explains-http-429-errors-and-rate-limit-best-practic-284d4e5/</guid>
      <pubDate>Sat, 05 Sep 2026 14:58:54 GMT</pubDate>
      <description>SerpApi published a guide covering the causes of HTTP 429 errors, how to read Retry-After headers, and strategies for retrying requests without escalating blocking.</description>
    </item>
    <item>
      <title>Zenrows integrates with smolagents to give AI agents production-grade web access</title>
      <link>https://extractfeed.io/story/zenrows-integrates-with-smolagents-to-give-ai-agents-product-bbb5156/</link>
      <guid>https://extractfeed.io/story/zenrows-integrates-with-smolagents-to-give-ai-agents-product-bbb5156/</guid>
      <pubDate>Sat, 05 Sep 2026 14:59:10 GMT</pubDate>
      <description>A tutorial shows how to swap smolagents' plain-requests VisitWebpageTool for a Zenrows-backed tool that bypasses bot checks.</description>
    </item>
    <item>
      <title>Zenrows MCP brings live scraping to Cursor's AI editor</title>
      <link>https://extractfeed.io/story/zenrows-mcp-brings-live-scraping-to-cursor-s-ai-editor-c22b25e/</link>
      <guid>https://extractfeed.io/story/zenrows-mcp-brings-live-scraping-to-cursor-s-ai-editor-c22b25e/</guid>
      <pubDate>Sat, 05 Sep 2026 14:59:07 GMT</pubDate>
      <description>Zenrows released an MCP integration that lets Cursor users scrape JavaScript-rendered and bot-protected sites directly from the editor.</description>
    </item>
    <item>
      <title>ScrapingBee publishes guide on MCP servers for web scraping, emphasizing control over data</title>
      <link>https://extractfeed.io/story/scrapingbee-publishes-guide-on-mcp-servers-for-web-scraping-114f3a0/</link>
      <guid>https://extractfeed.io/story/scrapingbee-publishes-guide-on-mcp-servers-for-web-scraping-114f3a0/</guid>
      <pubDate>Sat, 05 Sep 2026 14:59:04 GMT</pubDate>
      <description>ScrapingBee explains how a Model Context Protocol (MCP) server can improve scraping agent performance by keeping context lean and avoiding page bloat.</description>
    </item>
    <item>
      <title>Crawlbase introduces Web Bot Auth as an identity layer for bots, coinciding with pay-per-crawl pricing</title>
      <link>https://extractfeed.io/story/crawlbase-introduces-web-bot-auth-as-an-identity-layer-for-b-5dad3c5/</link>
      <guid>https://extractfeed.io/story/crawlbase-introduces-web-bot-auth-as-an-identity-layer-for-b-5dad3c5/</guid>
      <pubDate>Sat, 05 Sep 2026 14:59:01 GMT</pubDate>
      <description>Crawlbase ships Web Bot Auth, a specification that treats bot identity as a verifiable signature rather than a self-declared claim, on the same day it launches pay-per-crawl pricing.</description>
    </item>
    <item>
      <title>Scrapfly publishes 2026 guide to e-commerce scraping tools</title>
      <link>https://extractfeed.io/story/scrapfly-publishes-2026-guide-to-e-commerce-scraping-tools-d28e260/</link>
      <guid>https://extractfeed.io/story/scrapfly-publishes-2026-guide-to-e-commerce-scraping-tools-d28e260/</guid>
      <pubDate>Sat, 05 Sep 2026 14:59:13 GMT</pubDate>
      <description>Scrapfly's blog post surveys nine e-commerce scraping tools for developers, covering retrieval, extraction, browser automation, and discovery layers.</description>
    </item>
    <item>
      <title>Zyte argues EU AI scraping guidelines rely on outdated robots.txt standard</title>
      <link>https://extractfeed.io/story/zyte-argues-eu-ai-scraping-guidelines-rely-on-outdated-robot-62ff000/</link>
      <guid>https://extractfeed.io/story/zyte-argues-eu-ai-scraping-guidelines-rely-on-outdated-robot-62ff000/</guid>
      <pubDate>Sat, 05 Sep 2026 15:10:31 GMT</pubDate>
      <description>Zyte publishes a blog post arguing that Europe's new generative AI scraping guidelines, which lean on the robots.txt protocol, will harm users and entrench monopolies.</description>
    </item>
    <item>
      <title>Zyte pitches WebFetch as a drop-in replacement for coding agents' built-in fetch tool</title>
      <link>https://extractfeed.io/story/zyte-pitches-webfetch-as-a-drop-in-replacement-for-coding-ag-09ecfc0/</link>
      <guid>https://extractfeed.io/story/zyte-pitches-webfetch-as-a-drop-in-replacement-for-coding-ag-09ecfc0/</guid>
      <pubDate>Sat, 05 Sep 2026 15:10:34 GMT</pubDate>
      <description>Zyte published a blog post arguing that its WebFetch CLI tool outperforms the default webfetch tool in coding agents for research and coding workflows.</description>
    </item>
    <item>
      <title>Scrapfly ranks six open-source YouTube scrapers by job, flags two failures</title>
      <link>https://extractfeed.io/story/scrapfly-ranks-six-open-source-youtube-scrapers-by-job-flags-9da6470/</link>
      <guid>https://extractfeed.io/story/scrapfly-ranks-six-open-source-youtube-scrapers-by-job-flags-9da6470/</guid>
      <pubDate>Sat, 05 Sep 2026 15:10:37 GMT</pubDate>
      <description>Scrapfly published a blog post comparing six open-source YouTube scrapers, including a GitHub snapshot from August 11, 2026, and noting two projects that failed in their tests.</description>
    </item>
    <item>
      <title>Scrapfly publishes guide to scraping Target.com via Redsky API and bypassing PerimeterX</title>
      <link>https://extractfeed.io/story/scrapfly-publishes-guide-to-scraping-target-com-via-redsky-a-8cad6c2/</link>
      <guid>https://extractfeed.io/story/scrapfly-publishes-guide-to-scraping-target-com-via-redsky-a-8cad6c2/</guid>
      <pubDate>Sat, 05 Sep 2026 15:10:44 GMT</pubDate>
      <description>Scrapfly released a blog post detailing how to extract product and pricing data from Target.com using the internal Redsky API while handling store-keyed prices and PerimeterX anti-bot defenses.</description>
    </item>
    <item>
      <title>Zyte report reveals retailers as second most aggressive sector in blocking AI crawlers</title>
      <link>https://extractfeed.io/story/zyte-report-reveals-retailers-as-second-most-aggressive-sect-65d0ea4/</link>
      <guid>https://extractfeed.io/story/zyte-report-reveals-retailers-as-second-most-aggressive-sect-65d0ea4/</guid>
      <pubDate>Sat, 05 Sep 2026 15:10:49 GMT</pubDate>
      <description>Zyte published a blog post detailing how major retail marketplaces deploy anti-bot technology to block AI crawlers and automated data extraction.</description>
    </item>
    <item>
      <title>Crawlbase publishes technical guide on scaling headless browser fleets to 10,000 concurrent sessions</title>
      <link>https://extractfeed.io/story/crawlbase-publishes-technical-guide-on-scaling-headless-brow-32f02ba/</link>
      <guid>https://extractfeed.io/story/crawlbase-publishes-technical-guide-on-scaling-headless-brow-32f02ba/</guid>
      <pubDate>Sun, 06 Sep 2026 14:40:43 GMT</pubDate>
      <description>Crawlbase's blog post details the architecture and capacity planning required to run 10,000 concurrent Playwright sessions across roughly 100 nodes.</description>
    </item>
    <item>
      <title>Zyte research argues web scraping faces pricing barriers, not outright blocking</title>
      <link>https://extractfeed.io/story/zyte-research-argues-web-scraping-faces-pricing-barriers-not-00b8b4e/</link>
      <guid>https://extractfeed.io/story/zyte-research-argues-web-scraping-faces-pricing-barriers-not-00b8b4e/</guid>
      <pubDate>Sat, 05 Sep 2026 15:10:52 GMT</pubDate>
      <description>Zyte's State of Web Access report, discussed in an interview with the researcher, finds that new economic barriers are making web scraping more difficult rather than technical blocks.</description>
    </item>
    <item>
      <title>Zyte report finds fashion websites among the most heavily defended against scraping</title>
      <link>https://extractfeed.io/story/zyte-report-finds-fashion-websites-among-the-most-heavily-de-9cb9bf8/</link>
      <guid>https://extractfeed.io/story/zyte-report-finds-fashion-websites-among-the-most-heavily-de-9cb9bf8/</guid>
      <pubDate>Sat, 05 Sep 2026 15:10:56 GMT</pubDate>
      <description>A Zyte blog post examines the anti-bot and blocking measures used by fashion e-commerce sites, finding they are some of the most aggressively protected on the web.</description>
    </item>
    <item>
      <title>Scrapfly publishes tutorial on scraping Skyscanner flight prices with Python</title>
      <link>https://extractfeed.io/story/scrapfly-publishes-tutorial-on-scraping-skyscanner-flight-pr-6f59aa0/</link>
      <guid>https://extractfeed.io/story/scrapfly-publishes-tutorial-on-scraping-skyscanner-flight-pr-6f59aa0/</guid>
      <pubDate>Sat, 05 Sep 2026 15:10:59 GMT</pubDate>
      <description>Scrapfly released a blog post showing how to extract flight data from Skyscanner by constructing deep-link URLs and capturing itinerary JSON from the rendered page.</description>
    </item>
    <item>
      <title>Scrapfly publishes guide on scraping Airbnb listings and prices</title>
      <link>https://extractfeed.io/story/scrapfly-publishes-guide-on-scraping-airbnb-listings-and-pri-057da85/</link>
      <guid>https://extractfeed.io/story/scrapfly-publishes-guide-on-scraping-airbnb-listings-and-pri-057da85/</guid>
      <pubDate>Sat, 05 Sep 2026 15:11:01 GMT</pubDate>
      <description>Scrapfly released a blog post detailing how to scrape Airbnb search results, listing details, prices, reviews, and availability using Python and its own scraping platform.</description>
    </item>
    <item>
      <title>Scrapfly ranks five open-source LinkedIn scrapers on GitHub by auth model and ban risk</title>
      <link>https://extractfeed.io/story/scrapfly-ranks-five-open-source-linkedin-scrapers-on-github-348f842/</link>
      <guid>https://extractfeed.io/story/scrapfly-ranks-five-open-source-linkedin-scrapers-on-github-348f842/</guid>
      <pubDate>Sat, 05 Sep 2026 15:11:05 GMT</pubDate>
      <description>Scrapfly published a blog post evaluating five open-source LinkedIn scraping repositories on GitHub, ranking them by authentication approach, maintenance status, and real-world blocking risk as of August 2026.</description>
    </item>
    <item>
      <title>Zyte blog post examines how AI and web scraping turn scattered personal data into security risks</title>
      <link>https://extractfeed.io/story/zyte-blog-post-examines-how-ai-and-web-scraping-turn-scatter-b0c02d8/</link>
      <guid>https://extractfeed.io/story/zyte-blog-post-examines-how-ai-and-web-scraping-turn-scatter-b0c02d8/</guid>
      <pubDate>Sat, 05 Sep 2026 15:11:08 GMT</pubDate>
      <description>Domagoj Marić explores the intersection of AI, web scraping, and OSINT to show how fragmented personal data is assembled into profiles, scams, and security threats at Extract Summit.</description>
    </item>
    <item>
      <title>Scrapfly publishes guide on scraping Lowe's product data and bypassing Akamai</title>
      <link>https://extractfeed.io/story/scrapfly-publishes-guide-on-scraping-lowe-s-product-data-and-11a87d7/</link>
      <guid>https://extractfeed.io/story/scrapfly-publishes-guide-on-scraping-lowe-s-product-data-and-11a87d7/</guid>
      <pubDate>Sat, 05 Sep 2026 15:11:11 GMT</pubDate>
      <description>Scrapfly released a blog post detailing how to scrape Lowe's product, price, search, and store location data using embedded page state and their maintained Python scraper.</description>
    </item>
    <item>
      <title>Scrapfly Publishes Guide to Scraping DigiKey Data Past Cloudflare</title>
      <link>https://extractfeed.io/story/scrapfly-publishes-guide-to-scraping-digikey-data-past-cloud-1945102/</link>
      <guid>https://extractfeed.io/story/scrapfly-publishes-guide-to-scraping-digikey-data-past-cloud-1945102/</guid>
      <pubDate>Sat, 05 Sep 2026 15:11:18 GMT</pubDate>
      <description>Scrapfly's blog post details how to scrape DigiKey pricing, stock, and parametric specs while navigating its Cloudflare challenge, and compares this approach to using the official API v4.</description>
    </item>
    <item>
      <title>Zyte Audit Reveals Industry-Specific Bot Access Policies</title>
      <link>https://extractfeed.io/story/zyte-audit-reveals-industry-specific-bot-access-policies-9815786/</link>
      <guid>https://extractfeed.io/story/zyte-audit-reveals-industry-specific-bot-access-policies-9815786/</guid>
      <pubDate>Sat, 05 Sep 2026 15:11:22 GMT</pubDate>
      <description>Zyte published a large-scale audit of web access controls showing how different industries enforce different policies toward bots.</description>
    </item>
    <item>
      <title>Crawlbase details the infrastructure behind 8,000 CAPTCHAs per second</title>
      <link>https://extractfeed.io/story/crawlbase-details-the-infrastructure-behind-8-000-captchas-p-5421d0e/</link>
      <guid>https://extractfeed.io/story/crawlbase-details-the-infrastructure-behind-8-000-captchas-p-5421d0e/</guid>
      <pubDate>Sun, 06 Sep 2026 14:40:46 GMT</pubDate>
      <description>Crawlbase published a blog post explaining the throughput math, Go-based control plane, and scaling challenges required to solve 8,000 CAPTCHAs per second.</description>
    </item>
    <item>
      <title>Scrapfly ranks six open-source Instagram scrapers with notes on auth and ban risk</title>
      <link>https://extractfeed.io/story/scrapfly-ranks-six-open-source-instagram-scrapers-with-notes-97474c6/</link>
      <guid>https://extractfeed.io/story/scrapfly-ranks-six-open-source-instagram-scrapers-with-notes-97474c6/</guid>
      <pubDate>Sat, 05 Sep 2026 15:11:25 GMT</pubDate>
      <description>Scrapfly published a comparison of six open-source Instagram scrapers for 2026, covering auth models, ban risk, and maintenance status.</description>
    </item>
    <item>
      <title>Zyte blog profiles case study of $70 AI-coded app replacing $5,000 platform</title>
      <link>https://extractfeed.io/story/zyte-blog-profiles-case-study-of-70-ai-coded-app-replacing-5-5f8c76d/</link>
      <guid>https://extractfeed.io/story/zyte-blog-profiles-case-study-of-70-ai-coded-app-replacing-5-5f8c76d/</guid>
      <pubDate>Sat, 05 Sep 2026 15:11:29 GMT</pubDate>
      <description>Zyte published a blog post detailing how developer Fran Muñoz used AI coding and specification-driven development to build a production app that replaced a costly platform.</description>
    </item>
    <item>
      <title>Scrapfly compares five MCP servers for web scraping and browser automation</title>
      <link>https://extractfeed.io/story/scrapfly-compares-five-mcp-servers-for-web-scraping-and-brow-6c2fb2b/</link>
      <guid>https://extractfeed.io/story/scrapfly-compares-five-mcp-servers-for-web-scraping-and-brow-6c2fb2b/</guid>
      <pubDate>Sun, 06 Sep 2026 14:40:49 GMT</pubDate>
      <description>Scrapfly published a blog post comparing five MCP servers by their capabilities in protected-site scraping, browser control, debugging, static fetching, and cross-browser automation.</description>
    </item>
    <item>
      <title>Scrapfly compares six modern command-line tools as alternatives to cURL and Wget</title>
      <link>https://extractfeed.io/story/scrapfly-compares-six-modern-command-line-tools-as-alternati-9dd3c68/</link>
      <guid>https://extractfeed.io/story/scrapfly-compares-six-modern-command-line-tools-as-alternati-9dd3c68/</guid>
      <pubDate>Sun, 06 Sep 2026 14:40:52 GMT</pubDate>
      <description>A blog post from Scrapfly evaluates HTTPie, aria2, and other tools that address specific limitations of cURL and Wget, including a managed fetch tool for blocked requests.</description>
    </item>
    <item>
      <title>Scrapfly compares 8 Python HTTP clients for web scraping in 2026</title>
      <link>https://extractfeed.io/story/scrapfly-compares-8-python-http-clients-for-web-scraping-in-fd86041/</link>
      <guid>https://extractfeed.io/story/scrapfly-compares-8-python-http-clients-for-web-scraping-in-fd86041/</guid>
      <pubDate>Sun, 06 Sep 2026 14:40:55 GMT</pubDate>
      <description>Scrapfly published a blog post evaluating eight Python HTTP clients on async support, HTTP/2, HTTP/3, TLS impersonation, and maintenance, with runnable examples.</description>
    </item>
    <item>
      <title>Scrapfly ranks 7 lead scraping tools for 2026</title>
      <link>https://extractfeed.io/story/scrapfly-ranks-7-lead-scraping-tools-for-2026-644e664/</link>
      <guid>https://extractfeed.io/story/scrapfly-ranks-7-lead-scraping-tools-for-2026-644e664/</guid>
      <pubDate>Sun, 06 Sep 2026 14:40:57 GMT</pubDate>
      <description>Scrapfly published a ranked list of seven lead scraping tools covering no-code extensions and production APIs, with honest assessments of each tool's limitations.</description>
    </item>
    <item>
      <title>Study finds only 34.5% of free proxies work, thousands tamper with traffic</title>
      <link>https://extractfeed.io/story/study-finds-only-34-5-of-free-proxies-work-thousands-tamper-a6a450f/</link>
      <guid>https://extractfeed.io/story/study-finds-only-34-5-of-free-proxies-work-thousands-tamper-a6a450f/</guid>
      <pubDate>Sun, 06 Sep 2026 14:41:00 GMT</pubDate>
      <description>Crawlbase reports on two peer-reviewed studies that tested 640,600 free proxies, finding that just over a third were functional and many altered traffic.</description>
    </item>
    <item>
      <title>Scrapfly publishes diagnostic guide for browser fingerprint testing tools</title>
      <link>https://extractfeed.io/story/scrapfly-publishes-diagnostic-guide-for-browser-fingerprint-c9a9db6/</link>
      <guid>https://extractfeed.io/story/scrapfly-publishes-diagnostic-guide-for-browser-fingerprint-c9a9db6/</guid>
      <pubDate>Sun, 06 Sep 2026 14:41:03 GMT</pubDate>
      <description>Scrapfly released a layer-by-layer guide covering fingerprint and bot detection tools, explaining what detectable results mean and how to fix each leak.</description>
    </item>
    <item>
      <title>Vacation rental intelligence platform scales to 1 billion monthly crawl requests with Crawlbase Enterprise Crawler</title>
      <link>https://extractfeed.io/story/vacation-rental-intelligence-platform-scales-to-1-billion-mo-aae458f/</link>
      <guid>https://extractfeed.io/story/vacation-rental-intelligence-platform-scales-to-1-billion-mo-aae458f/</guid>
      <pubDate>Sun, 06 Sep 2026 14:41:09 GMT</pubDate>
      <description>Crawlbase published a case study detailing how a vacation rental intelligence platform processed 5.52 billion requests in six months with 99.96% success using its Enterprise Crawler.</description>
    </item>
    <item>
      <title>Chrome's new navigator.cpuPerformance API opens a fresh fingerprinting vector</title>
      <link>https://extractfeed.io/story/chrome-s-new-navigator-cpuperformance-api-opens-a-fresh-fing-9faf1c6/</link>
      <guid>https://extractfeed.io/story/chrome-s-new-navigator-cpuperformance-api-opens-a-fresh-fing-9faf1c6/</guid>
      <pubDate>Sun, 06 Sep 2026 14:41:12 GMT</pubDate>
      <description>Zyte reports that Chrome 152 will expose a navigator.cpuPerformance property, giving sites a new way to fingerprint browsers.</description>
    </item>
    <item>
      <title>Scrapfly publishes tutorial on scraping Google Jobs with Python</title>
      <link>https://extractfeed.io/story/scrapfly-publishes-tutorial-on-scraping-google-jobs-with-pyt-e84a57a/</link>
      <guid>https://extractfeed.io/story/scrapfly-publishes-tutorial-on-scraping-google-jobs-with-pyt-e84a57a/</guid>
      <pubDate>Sun, 06 Sep 2026 14:41:15 GMT</pubDate>
      <description>Scrapfly released a blog post showing how to scrape Google Jobs listings using Python and its own scraping platform.</description>
    </item>
    <item>
      <title>Zyte publishes largest ever audit of web access control mechanisms</title>
      <link>https://extractfeed.io/story/zyte-publishes-largest-ever-audit-of-web-access-control-mech-1445548/</link>
      <guid>https://extractfeed.io/story/zyte-publishes-largest-ever-audit-of-web-access-control-mech-1445548/</guid>
      <pubDate>Sun, 06 Sep 2026 14:41:18 GMT</pubDate>
      <description>Zyte released a comprehensive audit of how websites regulate programmatic visits, revealing the current state of web access barriers.</description>
    </item>
    <item>
      <title>Zyte publishes tutorial on building custom fetch tools for AI agents with Claude Agent SDK</title>
      <link>https://extractfeed.io/story/zyte-publishes-tutorial-on-building-custom-fetch-tools-for-a-b10db61/</link>
      <guid>https://extractfeed.io/story/zyte-publishes-tutorial-on-building-custom-fetch-tools-for-a-b10db61/</guid>
      <pubDate>Sun, 06 Sep 2026 14:41:20 GMT</pubDate>
      <description>Zyte released a tutorial showing how to create a custom fetch tool using the Claude Agent SDK to help AI agents extract structured data from the web.</description>
    </item>
    <item>
      <title>Zyte releases scrapy-spidey-sense, a preflight CLI for Scrapy projects</title>
      <link>https://extractfeed.io/story/zyte-releases-scrapy-spidey-sense-a-preflight-cli-for-scrapy-c54cf08/</link>
      <guid>https://extractfeed.io/story/zyte-releases-scrapy-spidey-sense-a-preflight-cli-for-scrapy-c54cf08/</guid>
      <pubDate>Sun, 06 Sep 2026 14:41:23 GMT</pubDate>
      <description>Zyte has open-sourced a command-line tool that performs static analysis on Scrapy projects before a crawl begins, scoring production-readiness and linking findings to fixes.</description>
    </item>
    <item>
      <title>Scrapfly publishes guide on scraping Google Play app reviews and metadata with Python</title>
      <link>https://extractfeed.io/story/scrapfly-publishes-guide-on-scraping-google-play-app-reviews-539f8b6/</link>
      <guid>https://extractfeed.io/story/scrapfly-publishes-guide-on-scraping-google-play-app-reviews-539f8b6/</guid>
      <pubDate>Sun, 06 Sep 2026 14:41:26 GMT</pubDate>
      <description>Scrapfly released a tutorial showing how to extract full Google Play app reviews, ratings, and metadata using Python, bypassing the typical few-hundred-review limit of free libraries.</description>
    </item>
    <item>
      <title>Scrapfly ranks four open-source proxy scrapers still viable in 2026</title>
      <link>https://extractfeed.io/story/scrapfly-ranks-four-open-source-proxy-scrapers-still-viable-6737496/</link>
      <guid>https://extractfeed.io/story/scrapfly-ranks-four-open-source-proxy-scrapers-still-viable-6737496/</guid>
      <pubDate>Sun, 06 Sep 2026 14:41:29 GMT</pubDate>
      <description>A blog post from Scrapfly filters the crowded open-source proxy tool landscape down to four actively maintained scrapers and checkers worth using this year.</description>
    </item>
    <item>
      <title>Scrapfly ranks 7 AI browser agents for production scraping in 2026</title>
      <link>https://extractfeed.io/story/scrapfly-ranks-7-ai-browser-agents-for-production-scraping-i-19a0fe3/</link>
      <guid>https://extractfeed.io/story/scrapfly-ranks-7-ai-browser-agents-for-production-scraping-i-19a0fe3/</guid>
      <pubDate>Sun, 06 Sep 2026 14:41:32 GMT</pubDate>
      <description>Scrapfly published a ranked guide to the best AI browser agents for automation and scraping, evaluating them on production stability and anti-blocking capability rather than demo performance.</description>
    </item>
    <item>
      <title>Scrapfly publishes guide on scraping Marriott hotel data through Akamai defenses</title>
      <link>https://extractfeed.io/story/scrapfly-publishes-guide-on-scraping-marriott-hotel-data-thr-d57755b/</link>
      <guid>https://extractfeed.io/story/scrapfly-publishes-guide-on-scraping-marriott-hotel-data-thr-d57755b/</guid>
      <pubDate>Sun, 06 Sep 2026 14:41:35 GMT</pubDate>
      <description>Scrapfly released a tutorial covering how to extract Marriott hotel prices and availability using Python, including bypassing Akamai bot protection.</description>
    </item>
    <item>
      <title>Zyte blog post explores rendering JavaScript pages with Playwright and Scrapy</title>
      <link>https://extractfeed.io/story/zyte-blog-post-explores-rendering-javascript-pages-with-play-0b34b8f/</link>
      <guid>https://extractfeed.io/story/zyte-blog-post-explores-rendering-javascript-pages-with-play-0b34b8f/</guid>
      <pubDate>Sun, 06 Sep 2026 14:41:37 GMT</pubDate>
      <description>Zyte published a guide on using Playwright to render dynamic content within a Scrapy workflow.</description>
    </item>
    <item>
      <title>Crawlbase argues AI agent failures are infrastructure failures, not code problems</title>
      <link>https://extractfeed.io/story/crawlbase-argues-ai-agent-failures-are-infrastructure-failur-cb2c217/</link>
      <guid>https://extractfeed.io/story/crawlbase-argues-ai-agent-failures-are-infrastructure-failur-cb2c217/</guid>
      <pubDate>Sun, 06 Sep 2026 14:41:49 GMT</pubDate>
      <description>Crawlbase publishes a blog post claiming that most AI agent failures stem from infrastructure issues like Markdown normalization, retrieval circuit breakers, and storage-backed memory.</description>
    </item>
    <item>
      <title>Scrapfly publishes guide to scraping Kayak flight data with its SDK</title>
      <link>https://extractfeed.io/story/scrapfly-publishes-guide-to-scraping-kayak-flight-data-with-831ad88/</link>
      <guid>https://extractfeed.io/story/scrapfly-publishes-guide-to-scraping-kayak-flight-data-with-831ad88/</guid>
      <pubDate>Sun, 06 Sep 2026 14:41:52 GMT</pubDate>
      <description>Scrapfly released a blog post walking through the process of scraping Kayak flight search results using its own SDK, covering JavaScript rendering and parsing internal poll JSON.</description>
    </item>
    <item>
      <title>Zyte launches 'Modern Scrapy for experienced developers' tutorial series</title>
      <link>https://extractfeed.io/story/zyte-launches-modern-scrapy-for-experienced-developers-tutor-e3a5408/</link>
      <guid>https://extractfeed.io/story/zyte-launches-modern-scrapy-for-experienced-developers-tutor-e3a5408/</guid>
      <pubDate>Sun, 06 Sep 2026 14:41:54 GMT</pubDate>
      <description>Zyte published the first part of a new blog series aimed at experienced developers building production-ready Scrapy projects.</description>
    </item>
    <item>
      <title>Scrapfly publishes guide on scraping RS-Online for product data</title>
      <link>https://extractfeed.io/story/scrapfly-publishes-guide-on-scraping-rs-online-for-product-d-7f2f49b/</link>
      <guid>https://extractfeed.io/story/scrapfly-publishes-guide-on-scraping-rs-online-for-product-d-7f2f49b/</guid>
      <pubDate>Sun, 06 Sep 2026 14:41:57 GMT</pubDate>
      <description>Scrapfly released a tutorial on extracting pricing, stock, specifications, and datasheet links from RS-Online's North American listings and product pages.</description>
    </item>
    <item>
      <title>Scrapfly compares Browser Use and Playwright for web scraping</title>
      <link>https://extractfeed.io/story/scrapfly-compares-browser-use-and-playwright-for-web-scrapin-0b85da8/</link>
      <guid>https://extractfeed.io/story/scrapfly-compares-browser-use-and-playwright-for-web-scrapin-0b85da8/</guid>
      <pubDate>Sun, 06 Sep 2026 14:42:00 GMT</pubDate>
      <description>Scrapfly published a blog post comparing Browser Use and Playwright, covering architectural differences, speed and cost tradeoffs, silent failure risks, and a hybrid approach for production scraping.</description>
    </item>
    <item>
      <title>Scrapfly publishes guide on bypassing AWS WAF Bot Control for web scraping</title>
      <link>https://extractfeed.io/story/scrapfly-publishes-guide-on-bypassing-aws-waf-bot-control-fo-bac190f/</link>
      <guid>https://extractfeed.io/story/scrapfly-publishes-guide-on-bypassing-aws-waf-bot-control-fo-bac190f/</guid>
      <pubDate>Sun, 06 Sep 2026 14:42:03 GMT</pubDate>
      <description>Scrapfly released a blog post detailing how AWS WAF Bot Control detects scrapers across five layers and how to bypass it using their Scrapfly ASP product.</description>
    </item>
    <item>
      <title>Scrapfly rounds up top open-source Facebook Marketplace scrapers on GitHub for 2026</title>
      <link>https://extractfeed.io/story/scrapfly-rounds-up-top-open-source-facebook-marketplace-scra-25a8e58/</link>
      <guid>https://extractfeed.io/story/scrapfly-rounds-up-top-open-source-facebook-marketplace-scra-25a8e58/</guid>
      <pubDate>Sun, 06 Sep 2026 14:42:06 GMT</pubDate>
      <description>Scrapfly published a blog post listing the five best open-source Facebook Marketplace scrapers on GitHub as of 2026, along with repos to avoid.</description>
    </item>
  </channel>
</rss>
