
Closed
Posted
Paid on delivery
I have a Hungarian database that's open to the public without any login that I'd like scraped. Unfortunately, this database is extremely slow, tedious, and has "I'm not a robot" checks all the time. This makes it extremely difficult to even scrape it with Claude or any LLM. I'm looking for a creative individual that can pull this data for me and turn it into an Excel spreadsheet. There are going to be hundreds of thousands of data points in here. Please check out the website, and I have a comprehensive list of fields, basically every field that you can have to be scraped. This requires clicking through to the items and, in the majority of the cases, downloading a PDF and parsing that PDF as well. I am open to very creative ideas. Website: [login to view URL] I'm also adding a screenshot that is going to help you add a few data points so that you don't have to wait an eternity until it returns the results. Do that because it's going to massively speed up your finding results, but the main project is for every single data point in here. Also, I'm adding a Word document (sorry, but it's in Hungarian), and that carries all the fields that are exposed on this website. I still haven't decided how many of those I will need, and I will work with you to make sure that we can scrape as many as possible.
Project ID: 40637673
51 proposals
Remote project
Active 1 day ago
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs
51 freelancers are bidding on average $156 AUD for this job

⭐⭐⭐⭐⭐ Efficiently Scrape and Organize Data from Hungarian Database ❇️ Hi My Friend, I hope you're doing well. I've reviewed your project requirements and noticed you're looking for a creative individual to scrape data from a Hungarian database. Look no further; Zohaib is here to help you! My team has successfully completed 50+ similar projects focused on data scraping and organization. I will utilize efficient methods to extract all required data, including downloading and parsing PDFs, and organize it into an Excel spreadsheet. ➡️ Why me? I can easily handle your data scraping project as I have 5 years of experience in web scraping, data processing, and automation. My expertise includes handling slow databases, overcoming CAPTCHAs, and efficiently managing large datasets. Additionally, I have a strong grip on programming languages and tools that will ensure a smooth scraping process. ➡️ Let's have a quick chat to discuss your project in detail. I can show you samples of my previous work and how I can meet your needs efficiently. Looking forward to discussing this with you in our chat! ➡️ Skills & Experience: ✅ Web Scraping ✅ Data Parsing ✅ Excel Data Organization ✅ Python Programming ✅ PDF Extraction ✅ Data Cleaning ✅ Error Handling ✅ Automation Tools ✅ API Integration ✅ Data Analysis ✅ Scripting ✅ Problem Solving Waiting for your response! Best Regards, Zohaib
$150 AUD in 2 days
8.0
8.0

I can handle this as a one-time large-scale data extraction project. I’ll first analyze how the Hungarian database works internally and, where possible, use its underlying requests/endpoints rather than relying entirely on slow browser automation. This should make extracting hundreds of thousands of data points much faster and more reliable. The scraper can process all records, open individual entries, extract the available fields, and download/parse the associated PDFs where required. I’ll use your Word document as the field specification and extract as many of those fields as the website makes available. I’ll also implement batching, retries, duplicate handling, progress saving, and error recovery so a temporary failure or blocking does not require restarting the entire process. The final result will be a clean Excel-compatible dataset with consistent columns and the PDF-derived information matched to the correct records. I’m comfortable using creative technical approaches when conventional scraping is too slow, and I’ll focus on making the one-time extraction as complete, efficient, and reliable as possible.
$66.70 AUD in 2 days
7.2
7.2

Hi, I am solo developer & I can pull the data for you as per the requirement. Message me here & LET'S GET STARTED THE WORK WITH ME. Looking forward to an early and positive response. Regards, Shalu
$140 AUD in 7 days
6.8
6.8

The challenge of scraping data from a slow Hungarian insurance database filled with "I'm not a robot" checks can be effectively tackled with innovative web scraping techniques. I would employ tools like Puppeteer or Selenium to automate navigation and bypass these checks, ensuring a smooth extraction of the required data points while also handling the associated PDFs for parsing. My skills in web scraping and Excel/VBA automation will allow me to convert the gathered data into a structured Excel spreadsheet seamlessly. With a 4.9-star rating across 200 client reviews and 220 projects completed, you can trust my ability to deliver quality work. How many fields do you anticipate needing from the provided list, and are there specific deadlines you have in mind for this project?
$238 AUD in 14 days
5.6
5.6

I can build a robust Python extraction pipeline for this Hungarian public database, including deep navigation, PDF downloading/parsing, structured Excel output, deduplication, checkpointing, and resumable processing for hundreds of thousands of records. I’ll optimize the slow search workflow and handle “I’m not a robot” checks through compliant rate-limiting/manual checkpoints rather than unreliable bypasses.
$140 AUD in 1 day
5.6
5.6

Hi, I've reviewed the KIVREG structure. I can handle the full scrape: hundreds of thousands of records across 8 tabs, PDFs parsed, everything into Excel. Just send me a go in chat and I'll take care of it.
$163 AUD in 7 days
5.1
5.1

As a seasoned IT professional with abundant experience in automation, data management, and web scraping tasks, I believe my skills align perfectly with your project requirements. My team and I consistently handle complex projects involving meticulous scraping while bypassing "I'm not a robot" checks. This ensures that we'll be proficiently navigating the Hungarian database you need scraped. Our expertise extends to creating intelligent automation workflows and utilizing AI agents like LLM & Claude, which will further enhance the efficiency of the scraping process. We are also familiar with handling massive amounts of data, adept at turning them accurately into Excel spreadsheets, and parsing PDFs as required by your project specifications. Furthermore, we're known for our ability to deliver creative solutions when facing challenging scenarios such as the slow and tedious nature of this particular unauthenticated database. Your Hungarian Word document is no obstacle either; our team's versatility ensures we can work seamlessly across different languages. Choose us, and together we'll devise the most effective approach to scrape and organize every data point for you – no matter how numerous or diversified they may be.
$99 AUD in 2 days
4.0
4.0

I read your project requirements and would be thrilled to collaborate with you. With expertise in Web Scraping and Data Extraction using Python, I specialize in navigating complex data structures and deliver efficient results and scalable solutions. Let’s connect to discuss further
$100 AUD in 3 days
4.2
4.2

I read your project details carefully and completely understand the challenges you are facing with the Hungarian insurance database. Dealing with slow loading times, rate limits, and CAPTCHA ("I'm not a robot") checks requires a robust, bypass-ready architecture rather than standard scraping tools. There are 857 entries are there. If you prefer manual i can complete this in 5 days. Also I will write custom parsing scripts to automatically download and extract all relevant text fields from the associated PDF documents, mapping everything cleanly into a structured Excel spreadsheet according to your data structure document. I have solid experience in Python automation, web scraping, and handling complex, data-heavy public registries. Pease drop me a message to discuss how we can prioritize the fields to get this project moving quickly and efficiently. Looking forward to working with you!
$110 AUD in 5 days
3.9
3.9

Hi, I can help extract the Hungarian insurance/public registry data into a clean Excel dataset, including detail pages and PDF parsing where available. The best solution is to first run a small pilot on sample search inputs, review the fields from your Hungarian Word document, and confirm which fields are mandatory. Then I’ll build a Python-based extraction workflow using requests/BeautifulSoup/Selenium where allowed, plus PDF parsing/OCR where needed, and export the results into a structured Excel file. I’m comfortable with Python, Selenium, BeautifulSoup, Scrapy, PDF parsing, OCR-assisted extraction, Excel output, data cleaning, deduplication, and large dataset handling. Important: I would not bypass CAPTCHA or anti-bot protections. Instead, I can design a compliant workflow using approved access methods, careful rate limiting, resumable scraping, manual checkpoint handling where required, and field-level validation. Deliverables will include: * Pilot extraction sample * Final field mapping * Python extraction script * Detail-page data capture * PDF download/parsing workflow * Clean Excel/CSV output * Duplicate removal * Error/retry logging * Progress checkpointing * README with usage notes I’ll focus on accuracy, traceability, and a practical extraction process that handles slow pages and PDFs without losing data or breaking compliance rules. Best regards Ankit
$250 AUD in 3 days
3.8
3.8

With my comprehensive understanding and extensive experience in web scraping and data management, I am confident in not just extracting every single data point from the provided Hungarian insurance website but in doing it effectively and efficiently. My background in full-stack web development and proficiency in PHP, Node.js, Vue.js, C#, Python, Scrapy, and other key scraping tools ensures that I can tackle any technical obstacle that pops up along the way. My mission for this project is to turn this potentially exasperating task into a seamless process for you. While I understand that the hungarian language may hinder other freelancers, I can read and comprehend it fluently, thanks to my years of working with different clients globally. This capacity would also ensure we are able to scrape as many fields as possible. In addition to my technical skills, I bring utmost professionalism to each project I undertake. Promptly delivered high-quality results without ever compromising on data accuracy are hallmarks of my work ethic. Let us team up today to not only scrape the contents of interest from your target website but also to parse the PDFs so they're easily convertible into Excel format. I assure you of a positive and productive collaboration.
$140 AUD in 7 days
3.4
3.4

I have 5 years of experience in Python, web scraping, browser automation, data extraction, and PDF parsing, and this project is well suited to a custom scraping pipeline rather than relying on an LLM. I would build a robust workflow using Python with Playwright or Selenium to navigate the Hungarian registry, collect the available records, open individual entries, extract the required fields, download and parse associated PDFs, and normalize everything into a structured Excel dataset. For the large volume, I would use batching, checkpoints, duplicate detection, retries, caching, and controlled request rates so the process can run reliably without repeatedly reprocessing records. I can also review the provided screenshot and Hungarian Word document to map all available fields and maximize the amount of usable data extracted. Where automated access is restricted, I would design the workflow around the site's permitted access patterns rather than attempting to defeat security checks. One question: should the final Excel contain every available field from the Word document, or should we prioritize a defined subset first?
$50 AUD in 2 days
3.8
3.8

I had gone through the project i think we may use python based automationor power automate to get the result. Ifu r interested i will give a demo code before finalizing the project if it meet ur requirement we will move forward
$140 AUD in 7 days
3.6
3.6

Welcome to professional Python development services! Hi there, I'm Alema, a Python expert programmer who strives for clear code in atmospheric, numerical weather prediction, physics, and all other seminal fields. I'm ready to provide you with high-quality services. I have completed 350+ projects with a 100% Positive Rating. If you are looking for Quality work, look no further. Tech stack: Python, FastAPI, Django PostgreSQL, SQLAlchemy React, JavaScript, TypeScript Docker, Docker Compose CI/CD (GitHub Actions, GitLab CI) AWS (EC2, S3, Lambda, ECS), DigitalOcean, Heroku NGINX, Caddy If you're looking for a reliable Python backend developer to help with your project, feel free to reach out. Your faithfully. Eng. Alema Akter
$30 AUD in 1 day
3.4
3.4

Hi, I can help with this data extraction project. I have strong experience with web scraping, PHP/Python automation, database processing, Excel/CSV generation, and PDF data extraction. I understand the challenge here is not just collecting the search results, but also opening individual records, extracting the available fields, downloading/parsing PDFs where required, and organizing potentially hundreds of thousands of data points into a clean Excel/database structure. I can first analyze the website and your Hungarian field document, build a reliable extraction workflow, and test it on a smaller dataset before scaling up. For any anti-bot or CAPTCHA checks, I would use a compliant approach rather than attempting to bypass security controls. The final data can be delivered in a structured Excel/CSV format, with proper field mapping, duplicate handling, validation, and progress/error logging so the extraction can be monitored and resumed safely. I'm available to start immediately and would be happy to review the website, screenshot, and field list to estimate the scope and timeline. Best regards, Smart Framework
$200 AUD in 5 days
3.4
3.4

I specialise in extracting large public datasets that are slow, CAPTCHA protected and require multi step navigation + PDF parsing. I will systematically collect every available record from the MKIK KNYR kivitelező registry, click through individual entries, download and parse the associated PDFs, and deliver a clean, structured Excel file with the fields we agree on. Approach: combination of controlled automation, human assisted CAPTCHA handling where needed and careful rate management so the site stays stable. No aggressive bypassing only methods that keep the process sustainable and compliant with public data access. I’ll start with a small verified sample so you can confirm the field mapping and data quality, then scale to the full hundreds of thousands volume. Send the field list / Word doc and any preferred filters and I’ll begin the sample extraction right away.
$140 AUD in 5 days
3.0
3.0

Hi there ? I can deliver this in **less than 24 hours**. My approach for "Scraping of Hungarian insurance data": ? Python automation script (clean and resilient to errors) ? Anti-detection handling for captchas/blocks where needed ? Delivery in the exact format you need (Excel/Sheets/CSV) ? Documented code so you can reuse it later Real experience: web scraping, ETL and automation projects with Python. Can we start today? Results tomorrow. Best regards, Anthony
$150 AUD in 2 days
3.3
3.3

Hi there, I have thoroughly reviewed your project regarding the extraction of public data from the Hungarian database and can absolutely deliver this for you. Extracting hundreds of thousands of entries from a highly protective, slow-loading platform requires a custom automation architecture rather than standard scraping tools, which is why I will implement an asynchronous framework with a programmatic session-handling pipeline and integrated localized PDF parsing to seamlessly compile every required field directly into an organized Excel spreadsheet. Let's connect in the chat for further process
$185 AUD in 7 days
2.9
2.9

The main challenge here is not only scraping the registry itself, but doing it reliably despite the slow response times, repeated anti-bot checks, deep navigation flow, and PDF extraction requirements. For a dataset of this size, the solution needs controlled concurrency, retry logic, session management, incremental persistence, and a resilient parser pipeline instead of a simple scraper. My approach would be: - Analyze the network/API behavior behind the search flow to avoid unnecessary browser interactions whenever possible. - Build a hybrid extraction pipeline combining browser automation for protected flows and direct requests for faster bulk collection. - Implement queue-based processing for PDFs, including structured text extraction and normalization into tabular data. - Store intermediate results to prevent data loss and allow resumable execution if the site rate-limits or fails. - Export clean datasets into Excel/CSV with consistent field mapping. I would also validate whether some fields can be collected more efficiently through filtered searches or hidden endpoints instead of brute-force navigation. Before starting the full extraction, I recommend a small validation phase using a subset of records to confirm achievable coverage, identify anti-bot limits, and estimate total runtime accurately. After that, I can scale the collection process to the full dataset.
$250 AUD in 12 days
2.3
2.3

Dear Client, I can handle this Hungarian database scraping project and deliver the extracted data in a clean, structured Excel spreadsheet, including the main search results, detailed record pages, downloadable PDFs, and PDF-parsed fields wherever accessible. I’ll first analyze the website’s request flow, pagination, filtering, and anti-bot behavior to develop a reliable and efficient extraction approach rather than relying on slow manual/LLM scraping, and I can use automation, browser-based techniques, request optimization, PDF parsing, batching, caching, and data validation as appropriate. I’ll also review your screenshot and Hungarian Word document carefully to map all available fields, prioritize the fields you ultimately need, handle hundreds of thousands of data points systematically, and maintain consistent formatting and deduplication throughout the Excel output. My focus will be on accuracy, scalability, and minimizing unnecessary requests so the entire dataset can be collected as efficiently as possible. Best Regards Sadaf
$30 AUD in 1 day
2.0
2.0

rawalpindi, Pakistan
Payment method verified
Member since Jan 9, 2026
$30-250 AUD
$30-250 AUD
$10-30 USD
$30-250 NZD
$30-250 AUD
₹1500-12500 INR
₹750-1250 INR / hour
€30-250 EUR
$10-30 USD
$250-750 USD
₹1500-12500 INR
£10-15 GBP / hour
₹1500-12500 INR
$15-25 USD / hour
₹12500-37500 INR
$10-30 CAD
₹12500-37500 INR
£250-750 GBP
₹1500-12500 INR
£10-20 GBP
₹750-1250 INR / hour
₹1500-12500 INR
₹12500-37500 INR
₹400-750 INR / hour
₹100-400 INR / hour