
Closed
Posted
Paid on delivery
I need an automated scraper that gathers data from several news sites in near real-time. The tool should loop through a list of URLs I will provide, respect each site’s [login to view URL] where possible, and export the captured information to CSV or JSON so I can feed it straight into my analysis pipeline. I’ll share the exact fields during kickoff, but the scraper must be flexible enough to handle common article elements—headline, body text, author byline, publication date, and source URL—and easy to extend if I add more outlets later. Time is critical. Delivery within 24–48 hours is preferred, so please lean on a proven stack such as Python with Scrapy/BeautifulSoup, Node with Cheerio, or any robust alternative you already master. The script should: • Rotate user agents and accept a proxy list to avoid blocks • Log failed requests for easy reruns • Be clearly commented and organized so I can update selectors myself Deliverables 1. Executable script or notebook with all dependencies noted 2. Sample output file containing at least 50 successfully scraped articles 3. Short README explaining setup, configuration, and how to add new sites The project is complete once I can run the scraper locally and reproduce your sample output without errors.
Project ID: 40584845
41 proposals
Remote project
Active 2 days ago
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs
41 freelancers are bidding on average $47 AUD for this job

I can develop a flexible near real-time news scraper that monitors your list of news sources and exports structured data to CSV or JSON. The solution can be built in Python using Scrapy/BeautifulSoup (or another preferred stack) and designed to be easily extended as you add new outlets. Features will include: • Extraction of common article fields such as headline, article body, author, publication date, and source URL • Configurable user-agent rotation and proxy support • Failed-request logging and retry capability for easy reruns • Clean, well-commented code with site-specific selectors separated for simple maintenance Deliverables will include the executable scraper, sample output containing at least 50 articles, and a concise README covering installation, configuration, scheduling, and adding new sources. The goal is to provide a reliable scraper that you can run locally and reproduce consistently without manual intervention.
$150 AUD in 3 days
7.3
7.3

Hey there Glane here, I can build a robust Python-based news scraping solution using Scrapy, BeautifulSoup,/Selenium, Requests, and pandas to collect articles from multiple news websites in near real time. The scraper will support rotating user agents, proxy integration, failed request logging, and export data to CSV or JSON with a modular, well-commented codebase that is easy to extend for new sources. I'll also provide a sample dataset with at least 50 scraped articles, along with a concise README covering setup, configuration, and adding new websites.
$50 AUD in 1 day
6.2
6.2

Hello, Check my previous project https://www.freelancer.com/projects/web-scraping/Extract-Text-Data-from-URLs/details I can build this scraper within 24–48 hours using Python (Scrapy/BeautifulSoup) with a modular architecture that's easy to extend for additional news sites. Collect headlines, article content, author, publication date, and source URL Export data to CSV and JSON Support proxy rotation and random user agents Log failed requests for easy retries Thanks Mamun
$30 AUD in 7 days
5.6
5.6

As an experienced and highly-rated freelancer in the field of data extraction, processing, and web scraping, I am confident that I possess the right skill set to efficiently and accurately complete your project. Not only have I worked with Python and Scrapy—proven stacks for this task—I have also extensively used BeautifulSoup and its colossal range of data collection capacities. Using these tools, I can organize, format, and clean your scraped data according to your specifications as well as generate descriptive reports to help you paint a clearer picture. Moreover, for a project of this nature where time is of the essence, you need someone who understands the value of meticulousness combined with quick thinking. I pride myself on both speed and accuracy; my work is always concise and error-free. Additionally, I bring in an attention to detail that ensures all aspects are covered—including extensive documentation—making it easier for you to maintain the scraper beyond our engagement. Lastly, my all-around experience in various tasks related to data management—including Excel proficiency for easy CSV exporting—shows my adaptability and how quickly I can understand diverse requirements. Rest assured that choosing me (or rather an 'extended' representation of me!) will be opting for diligence, reliability, and a solution-oriented mindset that never compromises on quality or efficiency. So let's get started!
$30 AUD in 1 day
4.6
4.6

Hello, I am an experienced Python Web Scraping and Data Extraction specialist, and I have already completed multiple similar scraping projects involving large-scale data extraction, automated scrapers, data cleaning, and CSV/Excel/JSON output. I can build this scraper using Python with Scrapy and BeautifulSoup, with flexible selectors for headline, article body, author, publication date, and source URL. I can also implement user-agent rotation, proxy support, failed-request logging, and an easy-to-extend structure for adding new news sites. You will receive the complete executable script, required dependencies, 50+ sample scraped articles, and a clear README for running and modifying the scraper locally. I can prioritize this project and deliver within 2 hours after receiving the target URLs and exact required fields. Ready to start immediately. Best regards, Rajaoul
$100 AUD in 1 day
4.7
4.7

Hi, I can build a Python-based news scraper that monitors multiple news sites and exports structured data to CSV or JSON. The scraper will support rotating user agents, optional proxy lists, failed request logging, and a clean, modular structure so new websites can be added easily. I will provide: Python source code with all dependencies Sample output with 50+ scraped articles README with setup and instructions for adding new news sites Well-commented code for easy maintenance The implementation approach will be based on the target websites. For standard HTML sites I'll use Scrapy/BeautifulSoup, and if any site relies heavily on JavaScript I'll use an appropriate browser automation solution where needed. I'm available to start immediately and can deliver within your requested timeline.
$30 AUD in 7 days
4.0
4.0

Greetings! I am a proficient web scraper, with extensive experience scraping data from several online sources and website using Python along with tools such as BeautifulSoup and Scrapy. I can help you making a script to extract the information you need for each article on the news website. I'll be able to return the script in less than 24 hours so you can verify it and test it on the website. Feel free to reach out, so we can start right away! Best regards, Sebastián H.
$40 AUD in 1 day
3.2
3.2

Hi, I will build a robust, modular Python scraper (Scrapy or BeautifulSoup/Requests) that extracts headlines, body text, bylines, dates, and URLs, delivering it in clean CSV/JSON within 24 hours. The script will be built for high resilience, featuring automatic user-agent rotation, a proxy list integration, and detailed error logging to handle blocks and trace failed requests smoothly. You will receive the executable code, a sample 50-article dataset, and a clear README with structured CSS selectors so you can easily add new URLs yourself.
$30 AUD in 3 days
4.0
4.0

Hello, I hope you are doing well. I have carefully checked your requirements and understand that you need a reliable news data extraction system that can collect structured article information from multiple sources in near real-time and deliver clean data for your analysis workflow. I will build the scraper using Python with Scrapy and BeautifulSoup, creating a modular extraction pipeline where each news source can have configurable selectors and parsing rules. I will implement request handling with user-agent rotation, proxy support, error logging, retry management, and structured CSV/JSON output so the collected data remains consistent and easy to process. The system will also include clear configuration files and documentation, allowing new websites and fields to be added without rebuilding the scraper. I will validate the scraper with sample article exports, optimize the extraction flow for reliability, and provide a ready-to-run package with setup instructions. Once you share the target websites and required fields, I can start immediately and deliver the working scraper within the required timeframe. Best Regards,
$50 AUD in 2 days
3.2
3.2

Hi, this is Joshua from Davis. I build practical automation tools and data workflows. You need a fast scraper for multiple news sites that is easy to extend and simple to run locally. A key risk here is blocking and inconsistent article layouts, so the scraper should be modular from the start. I would build this in Python with Scrapy for crawling and BeautifulSoup or lxml for parsing. I would add rotating user agents, optional proxy support, retry logging, and clean CSV and JSON exports. I would keep selectors in a config file so new outlets can be added without rewriting the core script. I can communicate in real time in your time zone and share a simple demo or part of the project within 12 hours of starting. Q1: Do you want the scraper to fetch full article pages, or only index pages with article links? Q2: Will you provide the initial URL list and any proxy credentials at kickoff? Q3: Should the output keep one unified schema even when some sites miss fields like author or date? Best, Joshua
$15 AUD in 1 day
0.0
0.0

Hey , Good morning! I’ve carefully checked your requirements and really interested in this job. I’m full stack node.js developer working at large-scale apps as a lead developer with U.S. and European teams. I’m offering best quality and highest performance at lowest price. I can complete your project on time and your will experience great satisfaction with me. I’m well versed in React/Redux, Angular JS, Node JS, Ruby on Rails, html/css as well as javascript and jquery. I have rich experienced in Python, Data Processing, BeautifulSoup, Scrapy, Web Scraping and Data Extraction. For more information about me, please refer to my portfolios. I’m ready to discuss your project and start immediately. Looking forward to hearing you back and discussing all details.. Best Regards
$35 AUD in 3 days
0.0
0.0

Hi there, Fast and reliable news collection is valuable because delayed or inconsistent data can weaken your analysis pipeline. I can build a flexible Python scraper with Scrapy and BeautifulSoup, CSV or JSON export, retries, logs, proxy support, rate limits, and easy site selectors. I’ve built similar multi source data scrapers with automated reruns and clean structured output. How many news sites are included initially? Thanks Srdan
$30 AUD in 1 day
0.0
0.0

With extensive experience in web development and a deep understanding of the ever-evolving world of data processing, I believe I am the perfect fit for your news site data scraping project. My skills extend through Python with scrapy, BeautifulSoup, Node with Cheerio, and other alternatives that can deliver near real-time results whilst respecting each site's robots.txt. Time is critical for you, which is why I consistently maintain a logical flow in my coding process, leaving no room for errors or delays. As you've mentioned, the tool should scrape common article elements such as headline, body text, author byline, publication date, and source URL. My proficiency in handling even complex scraping tasks assures you that not just 50 but a whole lot more successfully scraped articles within a short time-frame is definitely doable. Furthermore, I'm committed to providing you with comprehensible documentation on setup and configuration so that even the extension of new outlets later would be a breezy task for you. Remember, I'm not just doing this task - I'm ensuring my creation empowers your analysis pipeline to function efficiently and systematically for years to come. Choose me if you’re looking for an adept professional who loves rising up to challenges and delivering quality solutions on time. Let's turn your data scarcity into streamlined analysis!
$10 AUD in 1 day
4.1
4.1

Ethan here, From South Africa. I've read through your project and I'm definitely interested in assisting you. I see you need an automated scraper to gather data from various news sites in near real-time. My approach would be to use Python with Scrapy for its flexibility, ensuring we can easily extend it for new outlets later. I’d set it up to rotate user agents and accept a proxy list to minimize the risk of blocks, while also logging any failed requests for easy troubleshooting. What you really need is a reliable data source that feeds directly into your analysis pipeline, not just a script. I’m confident I can deliver this within your 24–48 hour timeline. Kind regards, Ethan
$21 AUD in 8 days
0.0
0.0

Hello, I can develop a flexible, well-structured news scraper in Python using Scrapy/BeautifulSoup that collects article data in near real-time and exports it to CSV or JSON for seamless integration with your analysis pipeline. The scraper will support user-agent rotation, proxy lists, request logging, configurable selectors, and clean documentation so adding new news sources is straightforward. I can deliver a production-ready solution within your 24–48 hour timeframe, including sample output, a detailed README, and organized, maintainable code. Best Regards Muhammad Shariq
$30 AUD in 7 days
0.0
0.0

Hi, The challenge isn't scraping news websites—it's building a scraper that remains reliable as websites change and new sources are added. That's what immediately caught my attention in your project. I like your focus on flexibility, maintainability, and clean data rather than a one-time extraction script. Instead of hardcoding individual websites, I'd structure the scraper as a modular pipeline where new sources, selectors, and output fields can be added with minimal effort while maintaining consistent logging, retry handling, and export formats. The result should be a dependable collection system that continuously feeds your analysis pipeline without becoming difficult to maintain as your list of news sources grows. One question: Do you expect the scraper to evolve into a continuously running monitoring service, or will it be executed on demand at scheduled intervals? I'd love to discuss your vision.
$30 AUD in 7 days
0.0
0.0

Hi, I have strong experience building high-performance web scrapers using Python (Scrapy, BeautifulSoup) and Node.js, including proxy rotation, user-agent rotation, retry logic, structured logging, and CSV/JSON exports. I can deliver a clean, well-documented scraper that is easy to extend with new news sources and meets your 24–48 hour timeline. Please share the target news sites and required fields, and I'll start immediately. Kindly initiate the chat to discuss the project in detail and bring your idea to life.
$30 AUD in 2 days
0.0
0.0

Noticing your emphasis on near real-time data gathering, it’s clear you require a highly efficient and adaptable scraper. The core challenge is to create an automated tool that respects site guidelines while delivering structured data swiftly. I specialize in building scrapers that are not only robust but also flexible enough to adapt to evolving requirements. In previous work, I delivered a similar scraper that successfully managed multiple news sites, capturing critical elements like headlines and publication dates, while ensuring ease of extension for future outlets. The approach I take includes using Python with Scrapy, enabling features like user agent rotation and detailed logging for failed requests. To kick off effectively, I’d appreciate knowing the specific number of URLs you plan to start with. This will help in scoping the project accurately. Chat Soon, Warm Regards Mthoko
$15 AUD in 7 days
0.0
0.0

Hi! I've previously built a news aggregation system that continuously scraped articles from multiple news sources, extracted structured content, and generated summaries in near real-time. I'm experienced with Python, BeautifulSoup/Scrapy, robust scraping practices, logging, and building maintainable, well-documented code. I can deliver a flexible scraper with proxy/user-agent support, CSV/JSON export, and clear documentation within your 24–48 hour timeline. I'd be happy to discuss the target news sites and get started immediately.
$35 AUD in 7 days
0.0
0.0

Hello, I can deliver a flexible and well-documented news scraping solution within your 24-48 hour timeframe. I have experience building Python and Node.js scrapers for content aggregation, monitoring, and analytics pipelines using Scrapy, BeautifulSoup, Requests, and Cheerio. For this project, I will provide: • A configurable scraper that processes your list of news URLs and extracts fields such as headline, article body, author, publication date, and source URL. • CSV and JSON export support for direct integration with your analysis workflow. • Modular site-specific selectors so new sources can be added easily. • Request failure logging and retry support for efficient reruns. • Clear comments and documentation to simplify future maintenance. Where appropriate and permitted by site policies, I can also implement: • User-agent rotation • Proxy list support • Rate limiting and request throttling to improve reliability Deliverables: ✔ Executable scraper with dependency list ✔ Sample output with scraped articles ✔ README covering setup, configuration, and adding new websites I understand the urgency of this project and can prioritize rapid delivery while keeping the code clean and maintainable. Best regards, Techavinya
$30 AUD in 1 day
0.0
0.0

Muridke, Pakistan
Payment method verified
Member since Oct 14, 2022
$8-15 AUD / hour
$10-30 AUD
$10-60 AUD
$8-15 AUD / hour
$8-15 AUD / hour
₹600-1500 INR
₹1500-12500 INR
$25-50 USD / hour
$250-750 USD
$250-750 USD
$2-8 USD / hour
₹12500-37500 INR
$30-250 USD
₹750-1250 INR / hour
$2000-6000 HKD
₹1500-12500 INR
$10-50 AUD
$15-25 USD / hour
$30-250 USD
₹750-1250 INR / hour
₹1500-12500 INR
$250-750 USD
$250-750 USD
₹100-400 INR / hour
$30-250 USD