
Closed
Posted
Paid on delivery
We are looking for an experienced Python engineer with web scraping / structured data acquisition / source integration experience to extend and generalize an existing multi-source funding opportunity acquisition system. This is NOT a greenfield crawler project. We already have an existing Python implementation with: source acquisition runner; write-gated importer; downstream promotion of validated discoveries; source registry and routing configuration; normalized candidate contracts; stable source/candidate identities; SHA-256 content hashing; deduplication and idempotency concepts; event classification and triage; automated tests; several working deterministic source adapters; an existing n8n environment. Your task is to reuse, consolidate, extend and generalize the current implementation, not rebuild it from scratch. A detailed Technical Pricing Pack is attached. Please review it before bidding. The attachment contains the full architecture context, source corpus, acceptance requirements, sample contracts, test expectations and Definition of Done. Main Goal Build a maintainable Multi-Source Opportunity Acquisition Framework v0.1 that can: regularly monitor our full approved source corpus; detect genuinely new funding opportunities; distinguish new calls from results, amendments, information pages, archives and already known opportunities; normalize source-specific data into the existing acquisition contract; preserve source identity, provenance and raw-content identity; deduplicate repeated discoveries; detect changes to previously known opportunities; preserve relevant PDF/DOCX attachment links; prepare valid opportunities for the existing downstream system; allow most future sources to be added through configuration rather than new custom Python code; expose a stable operational entry point suitable for n8n. Mandatory High-Value Source Areas The v0.1 scope explicitly includes: 1. [login to view URL] / [login to view URL] We already have qualification work and a previously validated discovery path. [login to view URL] is treated as a secondary aggregator/discovery source. Expected principle: [login to view URL] discovery -> official-source resolution where possible -> normalized candidate It must remain part of the acquisition system. 2. [login to view URL] Witkac is an important primary/public application platform. Existing repository work includes: known public contest index; public discovery evidence; incremental monitoring specification; polite-fetch rules; resolver/implementation specifications; URL/contest identity handling. The selected freelancer should reuse this existing work rather than research Witkac from zero. Known access limitations must be handled safely and explicitly. 3. Corporate foundations / private grant operators We already have a dedicated source corpus. The selected freelancer will receive a 46-row 2026-qualified/verified corporate-foundation dataset containing fields such as: official domain; grants/program page; news page; regulations; results; application platform; RSS/newsletter; evidence and status. A wider 84-row corpus also exists as supporting context. The objective is not to create 46 separate scrapers. The freelancer should determine which sources can use: generic HTML; RSS; list/detail patterns; common foundation/private-operator adapters; Witkac/application-platform resolution; configuration-only onboarding; and which genuinely require custom adapters. Every row in the verified 46-source dataset must be accounted for in the final coverage matrix. Existing Source Universe Existing work also covers or researches sources such as: NIW; Warsaw ETO; BIP families / BIP Logonet; Fundusze Europejskie; EEA Grants; ARiMR / [login to view URL]; PARP; BIPLO; Grantona; [login to view URL]; Atlas Dotacji; generic HTML sources; RSS/sitemap sources. At project start we will provide a frozen source acceptance corpus. We are deliberately not limiting this project to 10 or 20 URLs. Every source in the frozen corpus must receive an explicit final status, for example: SUPPORTED_EXISTING_ADAPTER SUPPORTED_GENERIC_CONFIG SUPPORTED_NEW_CUSTOM_ADAPTER BLOCKED_EXTERNAL_LIMITATION OUT_OF_SCOPE_BY_EXPLICIT_AGREEMENT Sources must not simply be skipped. Generalization Where technically reasonable, we want reusable source classes instead of one scraper per website. Expected classes may include: RSS / Atom; RSS + detail page; HTML list -> detail; generic configurable HTML; BIP patterns; registry/tenant-driven sources; official programme/call pages; JSON/API; Witkac/application-platform style sources; corporate foundation grant/news/program pages; PDF/DOCX attachment discovery. The desired future workflow is: add source to registry -> choose source class -> configure -> validate -> test -> enable instead of writing another scraper. Classification & Deduplication A newly fetched URL is not automatically a new opportunity. The system must distinguish: new real call -> INGEST_READY; ambiguous possible call -> REVIEW_REQUIRED; results -> NOT_A_NEW_CALL; information/archive page -> NOT_A_NEW_CALL; known unchanged call -> idempotent no-op; known changed call -> update/change signal; aggregator article -> discovery + official-source resolution. Deduplication must reuse or extend existing stable IDs, canonical URLs, hashes and source identities. It must also handle: repeated scheduled runs; URL variants; Witkac identity/fragment variants; aggregator -> official-source relationships; changed content; the same opportunity discovered from multiple listing paths. Safety & Reliability Expected safeguards include: request timeouts; response-size limits; allowed-domain validation; safe redirect handling; bounded concurrency; request budgets; sensible retry/backoff; no uncontrolled crawling; partial source failures must not crash the entire batch; anti-bot/403/JS-only limitations must be visible; no secrets in logs; no new paid infrastructure without prior approval. Repeated execution must be idempotent and must not create duplicate records. n8n Integration The wider system already uses n8n. We need a small stable integration point: n8n trigger -> acquisition runner -> normalize/classify/deduplicate -> existing persistence -> run summary This project does not include redesigning the entire n8n architecture. Testing & Documentation Expected deliverables include: regression tests for existing adapters; adapter contract tests; registry validation; generic source-class tests; [login to view URL] tests; Witkac resolver/identity/monitor tests; corporate-foundation onboarding tests; idempotency/dedup/change tests; source-failure tests; attachment-discovery tests; proof that several new sources can be added through configuration only. Documentation should include: final architecture; supported source classes; source registry documentation; adapter contract; official-source resolution rules; how to add a source without changing Python; how to add a custom adapter; failure/recovery notes; final coverage matrix for the complete source corpus; separate coverage matrix for all 46 verified corporate-foundation sources; [login to view URL] and Witkac support/limitations. Out of Scope This project does not include: organization/client intake; organization profiling; full semantic extraction of grant requirements from PDFs; matching organizations to opportunities; rewriting the downstream decision engine; application writing; new frontend development; database redesign; rebuilding the acquisition system from scratch. Before Bidding Please read the attached Technical Pricing Pack carefully. It contains significantly more technical detail than this listing. After reviewing it, please confirm: whether the full scope is clear; how you would generalize the existing implementation without rebuilding working adapters; how you would approach [login to view URL]; how you would reuse the existing Witkac work; how you would cover the 46-source corporate-foundation dataset without creating 46 custom scrapers; how you would handle deduplication, change detection and idempotency; your realistic delivery time; estimated hours; whether your quoted price covers the complete agreed v0.1 scope; anything described in the listing or attachment that is not included in your price. We want to agree the complete scope and price before development starts
Project ID: 40653264
51 proposals
Remote project
Active 5 days ago
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs
51 freelancers are bidding on average $48 USD for this job

You need to extend an existing Python-based funding opportunity acquisition system by generalizing source adapters, improving reliability, and expanding coverage without rebuilding the working architecture. I specialize in Python backend systems, data processing pipelines, web scraping, and structured integrations where clean contracts and reliable automation are critical. At Marin Software, I worked on Python-based serverless systems, AI/data pipelines, and production workflows where event processing, validation, and stable data handling were essential. I can work with your existing source registry, acquisition runner, normalized contracts, hashing, deduplication, and n8n workflow instead of replacing them. My approach would be to build reusable source classes for RSS, HTML, BIP-style pages, application platforms, and foundation sources, while adding custom adapters only where necessary. I can implement classification rules, change detection, idempotent execution, attachment discovery, tests, documentation, and final source coverage matrices. Can you share the Technical Pricing Pack and current repository structure so I can review the existing adapters before estimating the final effort?
$200 USD in 3 days
4.4
4.4

Hi there, The moment I read "Extend & Generalize an Existing Multi-Source Funding Opportunity Acquisition System (Python)", I knew it was a strong match for exactly what we do best. I love turning a solid brief like yours into a polished result you’ll be proud to put your name on. From your brief I can see this involves content, article, writing — all areas we handle in-house. We specialise in Python, Data Processing, Web Scraping, Software Architecture, which lines up directly with what you need. How we'd approach it: - Research your topic, audience and tone of voice - Draft well-structured, original copy - Revise against your feedback - Deliver SEO-ready, polished content I also noted you don’t need uncontrolled crawling — that’s clear, and I’ll keep the work clean and strictly on-brief. Delivering at the scale of 256 content is no problem for us — we're set up for volume without dropping quality. Expect smooth progress updates throughout, full respect for your specifications, and zero surprises along the way. I can start right away and keep you updated at every step — let’s make this a great one. Best regards, FreeLancers360 Let’s connect in chat and get started — message me anytime and I’ll reply right away!
$10 USD in 5 days
4.1
4.1

Generalizing this across new funding sources without breaking the existing n8n pipeline is the real challenge. I'd build a config-driven adapter layer so each new source is a set of rules, not new code, keeping the scraper and API stable. Can start today, working version in 4 days. These numbers are based on the post as written, we'll refine them after a quick scope call. I can build a free demo of the adapter pattern on one source first.
$30 USD in 4 days
4.0
4.0

Hi , You need an expert in Data Mining, Software Architecture, API Development, Data Processing, n8n, Web Scraping, Python and Data Integration, and I have a tailor-made solution ready for you. Your project brief instantly reminded me of a recent client who faced similar challenges, and I know exactly how to execute this flawlessly for your specific needs. To ensure we hit the ground running, I have three quick questions: Are there any additional technical details or constraints not mentioned in the brief? What is the primary hurdle currently blocking your progress on this? What is your strict timeline for completion? Why trust me with your project? The Record: 250+ Projects. 6+ Years. 100+ consecutive 5-star reviews. The Standard: Zero misses. I don’t just finish the job; I guarantee flawless execution. The Availability: Full-time freelancer, online 9 AM - 9 PM EST. My biggest "heavy-hitter" projects are kept off my public portfolio to protect client confidentiality. Click 'CHAT', and I’ll immediately send over relevant, private samples so you can see the standard of my work firsthand. Best regards, Muhammad Arsalan
$10 USD in 3 days
2.1
2.1

Hi there, With a passion for automation and AI, the ideal candidate to extend and generalize your existing multi-source funding opportunity acquisition system is me, Jazib Siraj. As an AI Agent and Automation Developer at Jumpnest, I have extensive experience in creating AI agents, RAG pipelines, and automation systems that can truly eliminate manual work. What sets my team apart is our focus on scoping and fixing any problem before writing a line of code. This means you won't be paying to automate a process that needs tweaking first. In regards to the specific requirements of your project, I have strong hands-on skills with Python, web scraping, and structured data acquisition which have been part of my core responsibilities throughout my professional journey. Moreover, I'm adept at source integration which matches one of the core aspects you require. Ultimately, what I offer is a comprehensive solution based on years of practical experience and a deep understanding of your project's goals. Let's have a free 30-minute scope call to discuss how my skill set can streamline your existing multi-source funding opportunity acquisition system.
$30 USD in 2 days
2.0
2.0

As an AI and Cloud Data Engineering specialist, I bring a unique blend of skills that are ideally suited for your project. I have extensive experience in not only web scraping and structured data acquisition, but I also excel at building maintainable frameworks capable of monitoring multiple sources and identifying new opportunities - exactly what you need! Moreover, I've been successful in distinguishing between different types of data to filter out information you may not need. Hence, with deep domain expertise in finance, healthcare and insurance sectors involved within these sources like Fundusze Europejskie, EEA Grants, BIP families/BIP Logonet, BIPLO, Grantona, Warsaw ETO etc. I bring wide ranges of knowledge to your project. My technical proficiency includes Python, SQL, TensorFlow and PyTorch to improve efficiency at every stage of the process. Being well-versed with AWS and Azure cloud infrastructure; I can efficiently manage your existing system and consolidate all the disparate elements into an easy-to-use and scalable interface.
$40 USD in 12 days
2.6
2.6

Hey , I just went through the project description, and I see you are looking for someone experienced in n8n, Data Integration, Software Architecture, API Development, Python, Data Mining, Data Processing and Web Scraping. It instantly reminded me of a client who faced similar challenges, and I knew I had a tailor-made solution for it. Please review my profile to confirm that I have great experience working with these tech stacks. While I have few questions: • Is there anything else you’d like to add to the project details? • What’s the top hurdle you’re facing with this project? • What is the timeline to get this done? Why Choose Me? 250+ Projects. 5 Years. Zero Misses. My reputation is built on a single metric: Flawless Execution. While others promise quality, my last 100+ consecutive 5-star reviews prove it. I don’t just finish the job; I set the standard. The portfolio here is just the tip of the iceberg. To respect client confidentiality, my recent heavy-hitters aren't public, but I can share them 1-on-1. Regards, Ali .
$10 USD in 6 days
2.2
2.2

GIVE ME 30 SECONDS TO SHOW YOU WHY I'M THE RIGHT FIT FOR THIS PROJECT. I successfully extended a similar Python-based funding acquisition system, where I integrated multiple data sources and enhanced deduplication mechanisms. This resulted in a 40% increase in valid funding opportunities detected. I have over 5 years of experience in Python engineering, specializing in web scraping and data integration. My expertise aligns perfectly with your requirements for reusing and generalizing existing components. I understand your goal to create a robust framework that efficiently monitors sources and detects new opportunities. I would achieve this by leveraging existing workflows and ensuring seamless integration with n8n. My focus will be on clear communication, meticulous execution, and ensuring long-term success through a maintainable architecture. I am confident I can deliver exceptional results that exceed your expectations. The difference between an average result and an exceptional one is usually decided before the work even begins. Regards Patrick
$15 USD in 7 days
1.0
1.0

Hello, With the comprehensive background I bring in Python and web scraping, I am fully equipped to tackle your project to extend and generalize the funding opportunity acquisition system. My experience goes beyond simply building from scratch; instead, I specialize in reusing, consolidating, extending, and generalizing existing implementations. That makes me perfect for this task as it even enables me to proactively prevent any unnecessary code duplication and ensure a maintainable framework.
$25 USD in 3 days
0.0
0.0

Hi there, I am a Full Stack Software Engineer with extensive experience in Python and data integration. My background in building scalable systems and automation solutions positions me well to successfully extend your multi-source funding opportunity acquisition system. This project is crucial for efficiently monitoring and acquiring funding opportunities. I propose to enhance the existing implementation by generalizing reusable source classes and integrating the systems through configuration options. My focus will be on ensuring deduplication, change detection, and maintaining source integrity, which are essential for delivering a robust solution. Please send a message so we can discuss the details further. Looking forward to working with you. Thank you, Andre
$12 USD in 10 days
0.0
0.0

With my extensive experience as a Full-Stack Python Developer, I'm confident that I can rise to the challenge of extending and generalizing your existing multi-source funding opportunity acquisition system. I have a proven track record in web scraping, structured data acquisition, and source integration - exactly the skills your project demands. I have comprehensive knowledge of working with varied data formats and handling diverse APIs, enabling me to effectively reuse, consolidate, and extend your current implementation according to your specific requirements. Moreover, my expertise in implementing automated testing regimes ensures high-quality additions to your system without introducing unnecessary risks. You can rely on me to preserve the vital aspects of your existing system like source identity, provenance, and raw-content identity while incorporating new sources seamlessly. Additionally, my proficiency in API development and data processing empowers me to not only make your acquisition system compatible with n8n but also expand its source capacity through configuration rather than requiring new custom Python code. Lastly, my dedication towards scalability and clean coding aligns perfectly with your project's goals. By leveraging my skills in optimization and cleaning up codebases for improved performance, I'll deliver a maintainable Multi-Source Opportunity Acquisition Framework v0.1 that will serve you well into the future.
$20 USD in 7 days
0.0
0.0

Hey — saw your post about extending a multi-source funding opportunity acquisition system. Keeping data acquisition reliable across varied sources is often the toughest part. Are you aiming to generalize the scraper for any funding site, or focused on a fixed set of sources? I’ve built Python scrapers and data pipelines that handle multiple changing formats smoothly. Send over your current code or specs and I’ll take a look to see how to best extend it.
$20 USD in 7 days
0.0
0.0

Hi — I am Anas from Tempe, AZ. It seems your primary goal is to enhance and broaden the capabilities of your existing multi-source funding opportunity acquisition system. This means ensuring the system can efficiently gather, process, and integrate data from diverse sources while maintaining reliability and performance. A key consideration here is how to structure the data acquisition process to ensure scalability and maintainability. By implementing a modular architecture, we can facilitate future enhancements without the risk of significant disruptions. This approach will also allow for easier integration with new data sources as they arise. To move forward, I would focus on optimizing the current scraping logic to handle structured data more effectively and ensure that API integrations are robust. Additionally, I can assist in developing documentation and a handover guide to ensure smooth transitions during updates. Could you clarify which existing functionalities you want to preserve, and are there specific sources you need prioritized for scraping? Looking forward to hearing from you.
$25 USD in 1 day
0.0
0.0

As an accomplished AI engineer with a penchant for tackling complex problems, I believe I bring a unique set of skills to your extended funding acquisition system project. With extensive experience in data mining, processing, and software architecture combined with my fluency in Python and proficiency in web scraping, I am well-versed in the exact areas this project requires. My knowledge of API development will be beneficial in creating a maintainable Multi-Source Opportunity Acquisition Framework v0.1 that can handle your large source corpus and identify genuinely new funding opportunities. Additionally, I've developed robust solutions for various organisations like yours which involved monitoring large data sets, detecting changes, normalizing unstructured data seamlessly and preserving content identity - all essential functionalities mentioned in your project description. Lastly, my communication skills are something that sets me apart. My ability to absorb complex requirements and apply them to real-world solutions while keeping the client informed at every step has earned me high praise from my former clients. I am excited by the challenge ahead and confident that together we can orchestrate a cost-effective solution without compromising on quality for your multi-source funding system. Look no further - let's get started on this important venture today!
$100 USD in 7 days
0.0
0.0

Poland
Member since Feb 24, 2010
$10-30 USD
$10-30 USD
min $10 USD
$10-30 USD
$10-20 USD
₹12500-37500 INR
₹600-1500 INR
$250-750 AUD
₹1500-12500 INR
₹600-1500 INR
$30-250 USD
$250-750 USD
₹750-1250 INR / hour
₹12500-37500 INR
$25-50 USD / hour
₹12500-37500 INR
₹1500-12500 INR
₹750-1250 INR / hour
$30-250 USD
$250-750 USD
$150-200 USD
₹12500-37500 INR
$15-25 USD / hour
₹1800-2800 INR / hour
$8-15 USD / hour