
In Progress
Posted
Paid on delivery
I need a reliable scraper built (Python, Scrapy / BeautifulSoup or similar) to pull data for roughly 1,000 UK companies directly from the publicly available Companies House register. For every company I want four fields captured: 1. Company name 2. Registered address 3. Full director name(s) 4. (Optional extra columns for multiple directors per company are fine) Please deliver the finished dataset as a clean, UTF-8 encoded CSV file with consistent column headers. A quick sample of 10 records up front will let me confirm the structure before you run the full job. Acceptance criteria • 1,000 distinct companies returned (no duplicates) • All mandatory fields populated; blanks only where the source truly has no data • CSV opens without formatting errors in Excel/Google Sheets • Script or notebook supplied so I can re-run it later (well-commented, with any required libraries listed in a README) Let me know your timeframe and any rate limits or captchas you anticipate so we can plan around them.
Project ID: 40612593
101 proposals
Remote project
Active 6 days ago
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs
101 freelancers are bidding on average £392 GBP for this job

Hi — Elias here from Miami. I see you're looking to scrape data for around 1,000 UK companies. The goal here is to efficiently gather reliable information while ensuring the scraper can handle potential changes on target sites. What usually matters most here is the scraper's ability to adapt to varying website structures and handle potential CAPTCHAs or rate limits. The tricky part is usually managing data quality and ensuring the scraper runs smoothly over time without frequent maintenance. My approach would involve using Scrapy for its robust handling of complex sites and BeautifulSoup for parsing. I’d structure the scraper to include error handling and logging to monitor performance and identify issues early. This way, we ensure stability and maintainability. I've developed similar scrapers for various industries, focusing on both efficiency and compliance with scraping best practices, so I understand the nuances involved. A few questions to better understand the scope: Q1 – Are there specific data fields you need from each company? Q2 – What are your expectations regarding the frequency of data updates? Q3 – Do you have any specific websites in mind that we should target or avoid? Happy to go through the details and suggest the best technical approach. Looking forward to hearing from you.
£500 GBP in 3 days
7.8
7.8

Hi, you're looking to get clear details of UK companies, especially names and addresses, in a neat way. I think automating with a scraper can speed things up and avoid manual work. I can build a Python script using Scrapy that pulls company info from Companies House, including multiple directors if needed. Have you thought about how often you'll need updates—weekly or monthly? Let's discuss how to plan this and create something big together. Regards, Nick. I’ve worked on projects like CarCRMDemo and TicketingHIPAAMedicalRecords that involved data extraction and cleanup, so I understand your needs.
£250 GBP in 3 days
8.0
8.0

Hello there, I will build a Python scraper targeting the Companies House register to extract company name, registered address, and full director name(s) for 1,000 distinct UK companies, delivered as a clean UTF-8 CSV. You will also get the commented script with a README listing all dependencies so you can re-run it anytime. I will send a 10 record sample within the first day so you can confirm column structure before the full run. Questions: 1) Are you targeting a specific SIC code or region, or is any 1,000 active companies acceptable? 2) Do you have a Companies House API key already, or should I register one under your account? Share your filtering criteria and I will have the sample CSV ready within 24 hours. Looking forward to potentially working together. Thanks, Kamran
£278 GBP in 13 days
7.3
7.3

With expertise in web scraping, I can extract data from the UK Companies House register for 1,000 companies efficiently and accurately. Using Python tools like Scrapy or BeautifulSoup, I will retrieve company names, addresses, and director information. I will deliver a well-structured CSV file with the extracted data fields, validated with a sample of 10 records for your review. The script will be documented clearly, along with a README file listing all required libraries. I am prepared to handle challenges such as rate limits and captchas with robust error handling mechanisms to ensure timely delivery of the complete dataset. My commitment includes providing 1,000 unique and well-populated company records in a clean CSV format, compatible with Excel/Google Sheets. I look forward to establishing a long-term partnership to cater to your data extraction needs effectively. Let's collaborate to exceed your expectations and create valuable data insights together.
£675 GBP in 5 days
7.1
7.1

Hi, To build a reliable scraper for UK companies, I can create a Python script using Scrapy or BeautifulSoup. I have experience in web scraping and data extraction. I have worked on similar projects where I pulled data from various sources and delivered it in clean formats. I will start by creating a sample of 10 records to confirm the structure with you. After your approval, I will scrape the data for 1,000 companies, ensuring no duplicates and all fields filled as required. I will also provide a well-commented script for you to re-run later. Regarding rate limits or captchas, I will implement strategies to handle them effectively. If you need to see my previous work, please check my portfolio: https://www.freelancer.com/u/techplusintl What is your preferred method for receiving the sample data? I invite you to chat further about this project.
£250 GBP in 2 days
7.0
7.0

Hello, Hope you are doing well, i am expert in web scraping , i can scrape your targeted data as per your given requirements, please come on chat so we can discuss all requirements in details, thank you Regards Gaurav Garg
£500 GBP in 7 days
6.6
6.6

Hi there, I understand you need a reliable Python scraper to extract company information for approximately 1,000 UK companies from the public Companies House register, delivering clean, accurate data along with a reusable script. I am confident I can build a robust scraping solution that captures the required fields while producing a well-structured CSV ready for Excel or Google Sheets. My approach will be to develop a Python scraper using Scrapy or BeautifulSoup, depending on the most efficient approach, to collect company names, registered addresses, and director details while handling pagination, validation, and duplicate prevention. The script will clean and normalize the extracted data, support multiple directors with separate columns where required, and generate a UTF-8 encoded CSV with consistent headers. I'll first provide a 10-record sample for your approval before processing the complete dataset, and the final solution will include well-commented source code, a README with setup instructions, required libraries, and guidance for rerunning the scraper in the future. Could you clarify whether you already have a list of the 1,000 target companies, or should the scraper select companies based on specific industries or search criteria? I'm ready to start immediately. Warm Regards, Aneesa.
£250 GBP in 2 days
6.6
6.6

Hi, I can build a reliable Python scraper for Companies House that returns 1,000 distinct UK companies with clean Company name, registered address, and director name fields in a UTF-8 CSV. I’ll structure the Data Collection so multiple directors are handled consistently, provide a 10-record sample first, then run the full export once you approve the format. The Web Scraping script/notebook will be commented and include a README with libraries, run steps, and any rate-limit notes. Do you already have a target list of company numbers, or should I select 1,000 distinct companies from the public register? Should director names be kept in separate columns or combined into one field per company? Confirm whether you want active companies only, or should dissolved companies be included if Companies House provides complete director and address data? Best, Ahtesham Ahmed
£555 GBP in 4 days
7.0
7.0

Hi, I reviewed the request to scrape 1,000 UK companies from the public Companies House register, collecting company name, registered address, and full director names into a clean UTF-8 CSV. I’ll build a reliable Python Scrapy data scraping pipeline, using Software Architecture patterns to keep the crawl logic, parsing, and deduplication clear. I’ll also handle optional multiple directors by expanding columns consistently, and I’ll export a structured CSV that opens cleanly in Excel/Google Sheets. I’ll deliver a small sample of 10 records first, then the full run with no duplicates and well-commented code plus a README. Let’s discuss here now.
£250 GBP in 30 days
6.5
6.5

Hello, I understand you need a reliable scraper to pull structured company data from Companies House without getting blocked or missing fields. I’ve built similar scrapers before, like one that collected thousands of business listings from public EU directories, where we had to handle inconsistent HTML and rate limits gracefully. I’d approach this using Scrapy with rotating proxies and delays because it keeps the scraping efficient while avoiding IP bans. The biggest improvement will come from consistent selectors and fallback logic for missing director names. I’ll test the script locally with a small batch first, then run the full job while logging any skipped entries so you can verify the output before the full run. The goal is a clean dataset you can use immediately and re-run if needed, without manual cleanup. I can start right now. Thanks, Denis
£400 GBP in 3 days
6.0
6.0

Hello, I'll develop a reliable Python scraper to extract company names, registered addresses, and director details directly from the Companies House register. The scraper will include robust parsing, duplicate detection, retry handling, and structured data validation to ensure the final dataset is complete, consistent, and easy to regenerate. Before running the full extraction, I'll provide a 10-record sample for your approval. Once confirmed, the script will collect 1,000 unique companies, export the results as a clean UTF-8 CSV with consistent headers, and support multiple directors through additional columns while gracefully handling records with missing source data. The final delivery will include the complete Python source code, a well-documented README with setup instructions and dependencies, and testing to verify CSV compatibility with Excel and Google Sheets, ensuring the scraper can be rerun whenever needed. I can START the WORK NOW and I can FINISH ASAP Relevant Work: https://www.freelancer.in/projects/beautifulsoup/Maritime-Job-Board-Scraping https://www.freelancer.in/projects/beautifulsoup/Python-Meetup-Events-Scraper Kind regards, Gowtham
£450 GBP in 1 day
5.8
5.8

Hi there, I hope this message finds you well. I am excited about the opportunity to assist you with your project to scrape 1,000 UK company details from the Companies House register. With extensive experience in Python, web scraping, and data mining, I am confident in delivering a reliable and efficient solution tailored to your needs. I understand the importance of accuracy and consistency in data collection. My approach will involve using Python with libraries like Scrapy and BeautifulSoup to extract the required fields: company name, registered address, and full director names. I'll ensure the data is cleaned and formatted as a UTF-8 encoded CSV file, easily accessible in Excel or Google Sheets. To validate the structure, I will provide a sample of 10 records for your approval before proceeding with the entire dataset. Additionally, I will deliver a well-documented script that you can re-run at your convenience. I'll include a README file listing any necessary libraries and provide guidance on handling potential rate limits or captchas to ensure seamless data extraction. Thank you for considering my proposal. I look forward to the possibility of working together and contributing to the success of your project. Best Regards, Efanntyo
£500 GBP in 10 days
5.9
5.9

Hello!, I am a US-based senior software engineer(frontend, backend, ecommerce, etc) with 15 years of experience in Python, web scraping, software architecture, data mining, Scrapy, BeautifulSoup, and reliable data pipelines. I read your project carefully, and the goal is clear: build a dependable scraper to collect around 1,000 UK company records accurately and in a clean, usable format. This kind of work needs someone who pays attention to the small details, handles edge cases, and delivers results that actually work. My approach would be: 1. Review the target source and confirm fields, pagination, and any anti-bot behavior 2. Build a robust scraper with retries, throttling, and structured output 3. Validate, dedupe, and deliver the data in CSV, JSON, or your preferred format 4. If needed, package it so you can rerun it later with minimal effort Could you please clarify the following questions to help me better understand the project? 1. What exact company fields do you need collected? 2. Are there specific websites or directories I should scrape from, or should I identify the sources? 3. Do you need a one-time scrape of 1,000 records, or an ongoing scraper you can reuse? I’ve built similar Python scraping and data collection tools for business directories, lead gen, and research workflows. A few relevant examples: local business extraction tools, directory crawlers, and structured data pipelines for SaaS and e-commerce teams.
£520 GBP in 3 days
5.7
5.7

Hi. To build this, I’ll use Python with Scrapy for reliable collection and BeautifulSoup for any edge-page parsing, then normalize the Companies House data into a clean UTF-8 CSV. I’ll first deliver a 10-row sample so you can verify headers, address formatting, and director handling before I run the full 1,000-company pass. I’ll also add deduping, retry logic, and rate-limit handling so the final dataset stays consistent and import-ready in Excel or Google Sheets. As a Senior Backend Engineer, I have mastered Python, Scrapy, BeautifulSoup, CSV/data cleaning, and resilient web extraction workflows and have strong experience in public-record scraping and reusable automation scripts. I’ll include a well-commented script or notebook plus a short README with dependencies and run steps. I am sure I can deliver high-quality results within 7 days. Let’s get in touch and discuss more. Thanks.
£420 GBP in 7 days
5.2
5.2

Companies House actually offers a free public API alongside the website, so pulling data through that route rather than raw HTML scraping would give you cleaner, more reliable results with far fewer rate limit or CAPTCHA concerns, especially at 1,000 records. I understand exactly what you need, company name, registered address, and full director names for 1,000 distinct UK companies, delivered as a clean UTF-8 CSV with consistent headers, plus the reusable script so you can rerun it later. Do you have a specific list or filter criteria for which 1,000 companies, such as by SIC code, region, or incorporation date, or should selection be left open as long as the count and fields are met? I will start by registering for Companies House API access and building the extraction script to pull company details and officer or director records for each entry, structuring multiple directors into additional columns as you suggested. Once I confirm the field mapping is accurate I will send you a 10-record sample to validate structure, then scale up to the full 1,000 distinct companies, checking for duplicates and blank mandatory fields before final export. I have solid experience with Python data extraction from public UK registries, so I anticipate minimal rate limiting through the official API, and I will deliver the well-commented script with a README so you can rerun the job independently going forward.
£300 GBP in 3 days
5.2
5.2

Hi, I’m a senior developer with 20+ years of experience, Top Rated on Freelancer.com with 1,500+ completed projects and a 5-star record, and I’ve built similar scrapers before, pulling public business data at scale. I’ll write a Python scraper using Scrapy or BeautifulSoup to fetch the 1,000 UK companies from Companies House, extract the four fields you need, and store the results in a clean UTF-8 CSV with consistent headers. The script will handle pagination, respect rate limits to avoid captchas, and include error handling for missing fields or failed requests. I’ll provide a sample of 10 records up front, then deliver the full dataset once validated. The code will be well-commented and include a README with dependency lists so you can re-run it later. If the API or structure changes, I’ll adjust the scraper accordingly. I can start immediately and deliver within 48 hours.
£300 GBP in 3 days
5.3
5.3

Hi, I will deliver a Scrapy spider and a clean UTF8 CSV of 1,000 distinct Companies House records within 3 working days, and I will send a 10 record sample first so you can confirm the structure. I built a Scrapy spider that extracted 3,200 Companies House records for a compliance client last year. I will run the scrape at 1 request per second with retries and proxy rotation to avoid blocks; no captchas are anticipated on the public register but I will use a headless browser fallback if one appears. Do you have a list of company numbers to target or should I discover companies via Companies House search? Happy to jump on a quick chat. Ali Zain
£500 GBP in 7 days
4.8
4.8

I can help with this, We will build a Python scraper (Scrapy or BeautifulSoup) to extract company name, registered address, and full director name(s) for 1,000 UK companies from the Companies House register. A clean, UTF-8 CSV with consistent headers will be the final output. Companies House offers a free REST API with 600 requests per five minutes. We will use that instead of raw HTML scraping. It is more reliable, avoids captcha issues, and handles pagination cleanly. We will also deduplicate by company number before export. A couple of quick things to confirm: 1) Do you have a specific list of 1,000 companies, or should we pull by SIC code, region, or incorporation date? 2) Do you need any additional fields (incorporation date, company status, SIC codes)? The number quoted here is a starting estimate. Looking forward to your response. Best regards, Faizan
£277 GBP in 13 days
4.6
4.6

Hello there. I hope you are donig well. I have successfully completed similar web scraping projects, utilizing tools like Python with Scrapy and BeautifulSoup to extract structured data from various sources. My experience ensures that I understand the nuances of scraping and data handling, allowing me to deliver accurate and reliable datasets. I understand you need to scrape data for 1,000 distinct UK companies from the Companies House register. I will implement a robust scraping solution that captures all required fields while adhering to best practices to avoid rate limits and captchas, ensuring a smooth and efficient process. I will deliver a high-quality, UTF-8 encoded CSV file with consistent headers, along with a well-commented script for future use. My approach prioritizes data integrity and usability, guaranteeing the final product meets your expectations and can be easily manipulated in Excel or Google Sheets. Please feel free to reach out to me. I look forward to working with you. Best regards, Billy Bryan
£450 GBP in 3 days
4.3
4.3

Hi, I got that you are looking for a reliable scraper to extract data from the Companies House register for 1,000 UK companies, capturing company name, registered address, and full director name(s). This is what I can help you with, let's chat. My approach is to use Python with Scrapy to ensure efficient and accurate data extraction. By leveraging advanced web scraping techniques, I will guarantee that all 1,000 distinct companies are captured without duplicates. The final dataset will be delivered in a clean, UTF-8 encoded CSV file with consistent column headers, ensuring seamless integration with Excel or Google Sheets. Additionally, I will provide a well-commented script or notebook for future reusability, meeting your acceptance criteria effectively. As final deliverables, you will receive a comprehensive CSV file containing the requested company details, along with a detailed script for your convenience. One thing I'd like to confirm before we start: Do you have any specific preferences for the format of the director names? Looking forward to discussing this project further. Regards, Imran
£250 GBP in 1 day
4.4
4.4

City of London, United Kingdom
Payment method verified
Member since Jul 23, 2026
£20-250 GBP
$15-20 USD
$10-30 CAD
$18.5 USD / hour
₹1500-12500 INR
$46.5 USD / hour
₹750-1250 INR / hour
€250-750 EUR
₹600-1500 INR
min $50 USD / hour
₹1500-12500 INR
$2-8 USD / hour
$3000-5000 USD
$10-150 USD
$25-50 USD / hour
$30-250 USD
₹1500-12500 INR
$250-750 AUD
$30-250 USD
$2-8 USD / hour
$10-30 USD