Filter

My recent searches
Filter by:
Budget
to
to
to
Type
Skills
Languages
    Job State
    2,149 beautifulsoup jobs found

    I need the structured data—every table or list—pulled from a publicly available HTML page and delivered in a neatly formatted Excel workbook. Please build a small, repeatable script so I can refresh the file whenever the page updates. Your solution can use any reliable stack you prefer (Python + BeautifulSoup/Requests, Selenium, Playwright, or similar); just keep it lightweight and well-commented. When you hand it over, I expect: • the finished .xlsx file with clean headers and no broken characters • the source code with a brief read-me showing how to run it and change the URL if the page moves No login or API keys are required—the page is completely public—so the focus is on accurate parsing and tidy Excel output.

    $113 Average bid
    $113 Avg Bid
    167 bids

    I need a robust web-scraping solution that can reliably pull 100,000 text-based records from a set of e-commerce sites. The data I’m after is purely textual—product titles, prices, descriptions, categories and any other publicly visible details that help build a clean catalogue. No images are required. Here’s how I envision the job: • You create an automated scraper (Python, Scrapy, BeautifulSoup, Selenium or a comparable stack) that navigates the target stores, handles pagination, variations and search filters, and respects reasonable crawl rates while bypassing blocks or CAPTCHAs when they appear. • The script should be reusable: site list, category paths and output format must be easy for me to tweak later. • Final deliverables: the fully co...

    $20 Average bid
    $20 Avg Bid
    53 bids

    ...user account. Server-side Excel Export (.xlsx): Auto-generated downloadable files with date/time tracking. Automated Webhooks: Webhook support for external CRM pushes. Deduplication: Hash-based checks (user_id + unique record identifiers) to prevent duplicate record storage.| Required Tech Stack Backend Framework: Python (Django or FastAPI) Scraping & Crawling: Scrapy, Playwright, Selenium, BeautifulSoup Anti-Bot Countermeasures: Rotating residential proxy pools, CAPTCHA bypass, stealth headless browsers Task Queue & Scheduler: Celery + Redis + Celery Beat Database: PostgreSQL (indexed for fast deduplication lookups) Frontend: React.js / with Tailwind CSS or Bootstrap How to Apply Please submit your proposal with: Relevant Portfolio: 2–3 examples of complex we...

    $190 Average bid
    $190 Avg Bid
    185 bids

    I need to automate the entire journey of pulling financial data from the web, saving it into a clean .csv file, and then funneling that data into Excel where I can explore it through a purpose-built dashboard with interactive filters. Here is what I have in mind: • A reusable script (Python with requests/BeautifulSoup, Selenium, or a similar stack) that scrapes up-to-date financial data. We can decide together which public sources make the most sense once we talk through feasibility and any site restrictions. • The script should output a single, well-structured .csv file ready for import. • An Excel workbook that automatically refreshes from that .csv and presents the data in a clear dashboard: charts, key metrics, and slicers or drop-down filters so I can drill...

    $104 Average bid
    $104 Avg Bid
    47 bids

    ...sources and deliver them in a clean, structured format I can quickly import into my database. That means gathering every public-facing datapoint a customer would expect to see—name, pack size, detailed description, brand / manufacturer, current selling price, any listed MRP or discount, and live stock availability—without missing a single SKU that appears on the target sites. You may use Python (BeautifulSoup, Scrapy, Selenium, or similar), Node, or another proven stack as long as the extraction is accurate, duplicates are handled, and the run can be repeated on demand. Please return: • A CSV or Excel file containing the full dataset • The well-commented script or notebook so I can re-run the crawl whenever needed I will validate by spot-checking prod...

    $71 Average bid
    $71 Avg Bid
    65 bids

    I need a reliable scraper built (Python, Scrapy / BeautifulSoup or similar) to pull data for roughly 1,000 UK companies directly from the publicly available Companies House register. For every company I want four fields captured: 1. Company name 2. Registered address 3. Full director name(s) 4. (Optional extra columns for multiple directors per company are fine) Please deliver the finished dataset as a clean, UTF-8 encoded CSV file with consistent column headers. A quick sample of 10 records up front will let me confirm the structure before you run the full job. Acceptance criteria • 1,000 distinct companies returned (no duplicates) • All mandatory fields populated; blanks only where the source truly has no data • CSV opens without formatting errors in Excel/Google...

    $405 Average bid
    $405 Avg Bid
    1 bids

    Looking for an Automation & AI Expert to build a Facebook Scraper, AI Parser, and Moderation Dashboard Project. I have build a free hyper-local mobil...Dashboard (Web App): Build a simple, ultra-fast backend dashboard for me. It should display the scraped image, original text, and the AI-extracted data (date, time, tags) pre-filled into input fields. I need to be able to review it, make minor edits if needed, and click a single "Approve & Publish" button to push it live to the app database in under 2 seconds. Ideal Freelancer Skills: Python (BeautifulSoup, Scrapy, or API integration) Open AI API / LLM text processing Backend dashboard development (e.g., Node.js, Django, or a low-code framework like Retool for speed)Experience with database mapping/geocoding (Pos...

    $489 Average bid
    $489 Avg Bid
    133 bids

    ...orchestration: coordinate scraping and file ingestion, then trigger downstream jobs. Extra capabilities • Email notifications at each major step, with templated messages. • Detailed logging to both console and rotating log files. • Secure data storage in a relational database so results are queryable later. I prefer developers who are comfortable with popular scraping libraries (Requests, BeautifulSoup, Selenium), file-handling packages (pandas, PyPDF2, openpyxl), and orchestration tools such as Celery or Prefect. Experience building RESTful APIs or adding AI/ML modules for future smart features is a plus. Proposed delivery structure - Milestone 1: Project skeleton, environment setup, CI/CD pipeline - Milestone 2: Web scraping module with tests - Mil...

    $1071 Average bid
    $1071 Avg Bid
    63 bids

    ... This data is used strictly for educational and statistical analysis. I am looking for a fixed-price offer. Bids that change or increase after the initial application will be immediately rejected. Technical Scope & Deliverables • The Engine: A self-contained Python script utilizing standard industry libraries (e.g., Scrapy, Selenium, BeautifulSoup, or requests). • The Output: One clean, fully deduplicated CSV file containing 25,000 to 30,000 unique rows. • Data Fields: o Mandatory: Establishment Name, Full Address, Region/City, Category (Restaurant/Activity/Cruise). o Optional (To be finalized together based on site policy constraints): Image URLs, Menu Prices (if available), Email, Website URL, Telephone, and Google Maps metadata. •

    $152 Average bid
    $152 Avg Bid
    87 bids

    I’m looking for a Python-based scraping solution that can harvest 65,000 unique records covering hotels and apartments located anywhere in Greece. The data will be used strictly for educational and statistical analysis. Scope • Build a self-contained Python script (ideally using requests, BeautifulSoup, Selenium, Scrapy, or similar) that visits agreed-upon sources—such as travel aggregators, official hotel pages, or review platforms—to pull the information. • Save every record to a single, well-structured CSV file. • Core fields I need are the establishment’s name and full address area/ city. I’m also interested in current pricing and availability. Images, facilities, and types of rooms and info pinned on the google map if those can...

    $249 Average bid
    $249 Avg Bid
    59 bids

    I need a reliable scraping solution that pulls every publicly-available customer e-mail address from the following sportsbooks: bet9ja, BetWinner, Sunbet, Nairabet, Melbet and Bet King. The final list must be delivered as a clean TXT file—one address per line, de-duplicated and free of formatting noise. Please build whatever combination of Python, Selenium, BeautifulSoup, Scrapy or in-house tooling you prefer as long as the method is repeatable. I will ask for a short proof-of-concept run (100–200 addresses) before we move to a full scrape so I can verify data quality and coverage. Acceptance criteria • All six brands above crawled end-to-end • TXT file delivered, UTF-8 encoded, no duplicates • Brief note describing the steps or script so the proce...

    $52 Average bid
    $52 Avg Bid
    23 bids

    ...should run on a schedule without manual intervention, gracefully handle captchas or page-layout changes, and drop the results into a clean, structured file or database table I can query—CSV or JSON is fine. Accuracy and consistency matter more than sheer speed; if a product page fails on one run, the script should log the error and retry. I’m comfortable working with Python, so Scrapy, Requests/BeautifulSoup, or Selenium are all acceptable. If you prefer another stack that achieves the same reliability, let me know. Please include in your bid: • A brief outline of your approach to avoiding Amazon’s anti-bot measures • Sample output for one ASIN so I can confirm the field structure • Estimated turnaround to get the first successful daily run ...

    $208 Average bid
    $208 Avg Bid
    54 bids

    ...Craigs list, CareerBuilder, or niche sites as long as they host downloadable resumes or profile data. • Industries: no narrow filter—I’m interested in technology, healthcare, finance, hospitality, construction, and any “various jobs” you come across. • Geography: restrict to the greater Las Vegas area (use location filters or keywords). Technical notes A Python solution using Scrapy, BeautifulSoup, Selenium, or a comparable stack is fine if it avoids detection and respects rate limits. Handle CAPTCHAs or login requirements gracefully. Please build in simple configuration so I can adjust city or keyword filters later without touching code. Deliverables 1. Clean CSV file and an identical Excel workbook containing all scraped records. 2....

    $130 Average bid
    $130 Avg Bid
    115 bids

    I need a fully-functioning lead-generation and sales operating system tailored to my granite and stone business and restricted to prospects located in Pune ci...crawl. 3. Export to a clean CSV or direct sync with the CRM I specify. 4. Simple dashboard where I can trigger a new crawl, view statistics and download the latest file. Acceptance criteria • Minimum 95 % valid email and phone coverage on the final list. • All leads mapped to Pune; no out-of-city records. • Source code, documentation and hand-off session included. Python, Selenium, BeautifulSoup, Scrapy or similar tools are fine; feel free to recommend alternatives if they boost reliability. The sooner the first usable dataset is in my hands, the better—I’m ready to move ahead as soon a...

    $13 Average bid
    $13 Avg Bid
    15 bids

    Project D...access) to fetch challan data from the PSCA website Bypasses or avoids Google reCAPTCHA v2 ("I'm not a robot") without using paid CAPTCHA-solving services (like 2Captcha or AntiCaptcha) Parses and displays the result on my own styled HTML template Creates a backend PHP-based API endpoint to return JSON response for submitted values Preferred Technologies: PHP (Core PHP / Laravel optional) BeautifulSoup, Playwright, or Selenium (optional in Python if PHP can't do it) Simple HTML DOM Parser or similar PHP scraping library No CAPTCHA solving APIs – strictly no paid services like 2Captcha Key Features: Frontend Form: Two input fields: Vehicle Number (e.g. LEA-19-8218) CNIC or Chassis Number (e.g. 3440301829785) Styled HTML form page O...

    $143 Average bid
    $143 Avg Bid
    56 bids

    I have a specific website from which I need all available email addresses together with the associated names ...from which I need all available email addresses together with the associated names and physical addresses extracted in a single, clean pass. This is a one-time scrape, not an ongoing collection, so accuracy and completeness on the first delivery are critical. Please return the results in a single CSV file, clearly labelled columns for Email, Name, and Address. If you automate with Python, BeautifulSoup, Scrapy, Selenium, or a similar toolset, that’s great—just make sure the final file opens correctly in any spreadsheet editor. Deliverable • CSV containing every accessible email, name, and address from the target site, with no duplicates and valid f...

    $1856 Average bid
    $1856 Avg Bid
    136 bids

    We are building a robust, production-grade automated fulfillment and pricing engine for an store. We have already designed the system architecture, business logic, and error handling rules. We need an experienced backend developer to implement the code quickly and cleanly. DO NOT APPLY if you only know standard web scraping (Selenium/BeautifulSoup) or basic scripts. We need a developer who understands asynchronous Python architecture, state machines, API authentication flows, and queue-based order handling. Core System Requirements & Scope Authentication & Session Management: Implement OAuth2/Cognito-based token acquisition (/api/client-credentials & /api/authentication/seller/token). Manage persistent seller sessions with dynamic token refreshing, automated retries...

    $188 Average bid
    $188 Avg Bid
    91 bids

    I already have flat-image mock-ups of every page, so the visual side is locked in; I simply need those screens turned into a clean, responsive informational website. Beyond the front-end build, the key requirement is a small data-p...call that I can trigger manually or via a cron job is fine. Please: • Recreate each page exactly as shown in my images (HTML5/CSS3/JS or a framework you prefer). • Build a simple admin or script where I can run “update” and watch the catalogue refresh. • Keep code well-commented so I can tweak selectors or timing later. If you’re comfortable with typical scraping stacks—Python (BeautifulSoup, Scrapy), Node, or PHP cURL—and can hand over the source along with basic deployment instructions, this ...

    $94 Average bid
    $94 Avg Bid
    86 bids

    ...costs I can define through a settings panel • A simple sales-trend module so I can see how each item has moved over the past weeks or months and spot consistent winners I’m happy with a cron job, Windows Task Scheduler or any lightweight scheduler you prefer, as long as the update cadence is reliable and I can trigger a manual refresh when needed. Preferred stack is flexible—Python with BeautifulSoup/Scrapy, Node with Puppeteer or any robust combination you’re comfortable maintaining. What matters is clean, well-annotated code, clear setup instructions and resilience to minor site layout changes. Deliverables I expect: – Executable (desktop or headless script) plus source code – Modules that pull data from each of the three market...

    $469 Average bid
    $469 Avg Bid
    310 bids

    ...photos so I can revisit the brochure months after the original ad has vanished. I will pay a set up and small monthly management fee of the process Data format isn’t fixed—CSV, Excel, JSON or anything similarly straightforward all work as long as I can open it later. What matters is that the archive is tidy and searchable. Key deliverables • A dependable scraper (Python, Scrapy/Selenium/BeautifulSoup—your call) scheduled to run weekly without manual input. • De-duping logic based on listing ID or URL so the same property is never saved twice. • For each listing, a folder containing the pictures plus a metadata file holding price and description. • Initial back-fill of current live listings so the archive starts complete. • ...

    $154 Average bid
    $154 Avg Bid
    173 bids

    ...where available. For every listing it finds, the script must capture: • Job title and company name • Job description and requirements • Location and salary range Please structure the CSV like format so each of those data points sits in its own clearly labeled column. I expect well-commented code that runs from the command line (Python 3.12+), uses common libraries such as requests/BeautifulSoup or Selenium if dynamic content makes that necessary, and includes a quick README explaining how to install any dependencies and run a sample search. Acceptance will be based on running the script with a sample keyword; if the resulting CSV faithfully mirrors what I see on the first several pages of Indeed and the columns match the fields above, the job is done. T...

    $93 Average bid
    $93 Avg Bid
    212 bids

    I need a clean, up-to-date extract of specific listings from an accommodation website. The focus is on Apartments, Cottages, and Bed & Breakfast properties, and for each listing I wa...property types covered (apartments, cottages, B&Bs) • Accurate capture of property name, full address, geo-coordinates if shown, amenities summary, and contact phone number • No duplicates, no missing rows, no truncated text • Scraper respects and rate limits or uses rotating proxies where needed • Clear, well-commented code supplied so I can rerun the scrape later (Python, BeautifulSoup/Scrapy/Selenium—use what suits you) If anything on the site requires special handling—JavaScript-rendered content, pagination, or CAPTCHA—please note it up front...

    $30 Average bid
    $30 Avg Bid
    31 bids

    ...community name • Full address • Total area in acres (or square feet if acres are missing) • Number of houses or flats in the project • Key amenities mentioned by the developer or seller • Cost per house / flat (latest listed price or price range) Deliverables 1. A CSV or Excel file with at least 1,000 complete rows that meet the above criteria. 2. The scraping script (Python + BeautifulSoup/Scrapy, Node.js with Cheerio/Puppeteer, or a comparable stack) well-commented so I can rerun it later. 3. Brief run instructions and any environment requirements. Acceptance criteria • No duplicate projects. • All mandatory fields populated for 95 %+ of rows. • Only Bengaluru addresses; anything outside the city will be reject...

    $71 Average bid
    $71 Avg Bid
    35 bids

    ...notice, etc.) captured consistently across all sites. If additional obvious fields appear, include them as well. Output format Everything must arrive in a single, well-structured Excel spreadsheet, one row per tender with clear column headings. Please normalise dates and strip duplicates so the file is immediately usable for analysis. Technical approach Feel free to employ Python, Scrapy, BeautifulSoup, Selenium or any combination that handles pagination, dynamic content or CAPTCHA challenges. The key is repeatability: I’d like to rerun the script later, so provide clean, documented code. Deliverables • Final Excel workbook containing all scraped tender records • Runnable script or notebook with brief setup instructions • I will want the script t...

    $83 Average bid
    $83 Avg Bid
    54 bids

    ​We are looking for an exp...must strictly check existing product SKUs to prevent duplicate entries or overwrite errors. Only new products should be added or existing ones updated safely. ​Bulk Import: Format the final data into a clean CSV file and successfully import it into WooCommerce without crashing the server (batch imports recommended). ​Requirements for Applicants: ​Proven experience in Web Scraping (Python, BeautifulSoup, Scrapy, Apify, or similar tools). ​Strong background in WooCommerce data management, CSV mapping, and bulk product imports. ​Please provide examples of similar data extraction or e-commerce import projects you have completed. ​Note: We prefer to start with a milestone test of 20-30 products to ensure accuracy before proceeding with the full catalog.

    $32 Average bid
    $32 Avg Bid
    82 bids

    ...date (if available), and the extracted description. A simple SMTP configuration file is enough—no need for a full mail server build. • Scheduling: It should be runnable via cron / Task Scheduler; a 10-15 minute crawl cycle is fine as long as no single run overlaps the next. Deliverables 1. Clean, well-commented source code. 2. A short README describing dependencies (e.g., requests, BeautifulSoup, Selenium if needed for JavaScript sites) and setup steps. 3. Example configuration files for the site list and SMTP settings. 4. A brief video or screenshot walkthrough proving the alert fires on a test posting. I will consider the task complete once I can plug in three live company URLs, run the script unattended for a day, and consistently receive email wheneve...

    $87 Average bid
    $87 Avg Bid
    83 bids

    I am looking exclusively for a local UAE-based developer; the total budget for this project is $10–$15, and it i...missing, the plug-in quietly launches a web-scraper (Python) to grab supplier specs and fills the gaps before updating Airtable. Acceptance criteria 1. Signed-installer (.msi or .exe) that adds a ribbon tab in Inventor 2024+ with the above workflow. 2. Python source, neatly documented, including the computer-vision model (TensorFlow or PyTorch—your choice) and the web-scraping routines (BeautifulSoup/Scrapy). 3. REST hooks or direct connectors for Airtable and Power BI proven to work with my API keys. 4. Short video demo plus a written deployment guide. If something in the toolchain needs tweaking, I’m open to suggestions—as long as Pyt...

    $35 Average bid
    $35 Avg Bid
    18 bids

    I need a reliable scraper that automatically captures price, reviews, and the full product description from both Amazon and Flipkart, then serves that data through a simple API I can call each day. The flow I ...description) • Setup guide plus one test run proving the data refresh Acceptance criteria 1. Hitting the API after the daily run returns up-to-date fields for at least ten test products. 2. No hard-coded credentials or paths; environment variables fully supported. 3. Documentation covers installation, scaling the crawler count, and updating the SKU list. Any stack is welcome—Python (Scrapy, BeautifulSoup, Selenium), Node (Puppeteer, Cheerio), or another proven toolset—so long as it meets the requirements above and can be deployed on a standard V...

    $85 Average bid
    $85 Avg Bid
    76 bids

    ...number revealed when the WhatsApp icon is tapped The final deliverable is a single, well-structured Excel file where every row represents one company and each field is placed in its own column. Data integrity is crucial, so duplicates, blank rows, and broken characters should be removed or flagged for review. Because there are thousands of entries, I expect an automated approach (Python, BeautifulSoup, Selenium, or similar mobile-app scraping tools) combined with manual spot-checks to ensure accuracy. Please be sure your method respects any rate limits or anti-scraping measures built into the app. Acceptance will be based on: 1. Full coverage: at least 10,000 unique companies returned. 2. Correct mapping of columns (name, address, contact details, WhatsApp number). 3. Ze...

    $150 Average bid
    $150 Avg Bid
    172 bids

    I need a reliable scraper built (Python, Scrapy / BeautifulSoup or similar) to pull data for roughly 1,000 UK companies directly from the publicly available Companies House register. For every company I want four fields captured: 1. Company name 2. Registered address 3. Full director name(s) 4. (Optional extra columns for multiple directors per company are fine) Please deliver the finished dataset as a clean, UTF-8 encoded CSV file with consistent column headers. A quick sample of 10 records up front will let me confirm the structure before you run the full job. Acceptance criteria • 1,000 distinct companies returned (no duplicates) • All mandatory fields populated; blanks only where the source truly has no data • CSV opens without formatting errors in ...

    $526 Average bid
    $526 Avg Bid
    101 bids

    ...a day; no manual trigger should be required after the first configuration. If it discovers a posting that was not present on the previous run, it should save the new record (CSV or SQLite is fine) and raise a simple desktop notification or send an email—whichever is easier for you to wire up quickly. I am comfortable with common web-automation stacks such as Python + Selenium/Playwright, BeautifulSoup, Scrapy, or even a lightweight C# solution, so feel free to choose the toolset you can ship fastest. Just package everything into a single installer or portable EXE; I do not want my end users touching the command line. Acceptance criteria: • Daily scheduled crawl without Windows Task Scheduler hacks — the app should handle timing internally. • Accura...

    $12 Average bid
    $12 Avg Bid
    25 bids

    I have a running list of rest...current offers, and customer-facing ratings. In short, if the information is displayed on Talabat, I want it in the file. A few guardrails: • Accuracy is critical; I will spot-check random entries against the live site. • No duplicates—each restaurant should appear exactly once. • Please return the data in CSV or XLSX along with the original scrape script (Python-based preferred, using BeautifulSoup, Scrapy, or Selenium—whatever you find most reliable against Talabat’s layout). • Script must be reusable so I can refresh the dataset later without starting from scratch. If you have prior experience scraping dynamic food-delivery platforms and can deliver clean, well-documented code plus the compiled datase...

    $27 Average bid
    $27 Avg Bid
    59 bids

    ...stock • the current selling price A simple CSV or Google-Sheet style output is perfect, but a lightweight dashboard is welcome too if it doesn’t slow development. Real-time data is not essential; an automatic refresh every 15–30 minutes is fine. You are free to pull the information through official APIs if they exist or by clean, well-commented scraping routines (Python with Requests/BeautifulSoup/Selenium, or a comparable stack). Whatever approach you pick, I need: • the full source code • a quick setup guide that lets me run it on Windows or a small VPS • one sample run demonstrating correct availability and price capture for at least three different pincodes If any site blocks too many requests, build in basic throttling or rota...

    $19 Average bid
    $19 Avg Bid
    11 bids

    ...to the scraped numbers and keeps that leaderboard up to date. Please scrape all three data groups—player statistics, game results, and player profiles—then funnel them into a database. I’m comfortable using whatever engine makes the most sense (MySQL, PostgreSQL, or even SQLite); I simply need you to explain the trade-offs and set it up for me. Deliverables • A robust scraper (Python + BeautifulSoup/Scrapy/Selenium—use what’s best for the target site) • A well-structured relational database with the imported data • Clear documentation on the schema and how to refresh the data • A ranking algorithm/script that calculates player standings from the live database • A short write-up comparing the database options you c...

    $138 Average bid
    $138 Avg Bid
    145 bids

    I need data scraped from the Google and relevant sources. The final result should be a clean, well-str...600 INR. Work Description: Need data of all the doctors in India. Need data of all the hospitals, clinics, and medical facilities in India. Key points I care about: • Accuracy: only valid, fully-formatted phone numbers. • Efficiency: the scrape should run unattended and respect reasonable rate limits so IPs don’t get blocked. • Reusability: deliver the source code (Python, Scrapy/BeautifulSoup/Selenium—use what you prefer) plus a brief README so I can rerun the extraction later or point it at new URLs. Let me know how you plan to tackle CAPTCHAs or dynamically loaded pages, and provide an estimated turnaround time along with a short sample ...

    $20 Average bid
    $20 Avg Bid
    17 bids

    ...upload product URLs, set a target or percentage-off value, and the system checks those pages frequently enough to feel “real time” without triggering the sites’ anti-bot measures. The moment a drop is detected, an email—my preferred notification channel—should land in my inbox with the new price, the old price and a direct link back to the product. I’m flexible on the tech stack, but Python (BeautifulSoup, Selenium, Scrapy), JavaScript (Puppeteer) or any other proven web-scraping framework that can handle dynamic content, rotating proxies and captchas is fine with me. A small database or flat-file store that remembers each product’s history will be useful for trend graphs later, though that’s a nice-to-have rather than a blocker....

    $222 Average bid
    $222 Avg Bid
    63 bids

    ...performance and data accuracy - Provide documentation and setup instructions Required Data Fields (examples): - Full name - Job title - Company name - Location - Industry - Profile URL - Company information - Publicly available contact information (if available) - Other custom fields based on requirements Technical Requirements: - Strong experience with: Python scraping frameworks (Scrapy, BeautifulSoup) Selenium / Playwright automation Browser automation Data processing pipelines APIs and integrations Proxy/session management Database storage Preferred Experience: - Previous LinkedIn scraping projects - Experience building lead generation tools - Experience handling large datasets - Knowledge of scraping reliability and maintenance Deliverables: - Working scraper/tool - Sour...

    $165 Average bid
    $165 Avg Bid
    95 bids

    ...login are not in scope for now, so authentication flows can be ignored. The problems I see: • Some pages return only partial HTML, losing key sections that appear after JavaScript executes. • PDF links are discovered but not downloaded consistently. • Text files get downloaded, yet their contents arrive empty or garbled. I would like you to: 1. Review the current codebase (requests/BeautifulSoup for static parts, a lightweight Selenium fallback for dynamic ones) and pinpoint where extraction breaks. 2. Patch the logic so it reliably gathers the three asset types above, saving them to the folder structure already defined. 3. Provide a short report or inline comments explaining the fix so I can maintain it later. A successful hand-off means running the ...

    $144 Average bid
    $144 Avg Bid
    161 bids

    ...every single record I expect four fields—name, phone number, crash report, and the driver’s insurance details—so please do not send partial data. You may pull the information from any combination of public or private/insurance databases; I care more about speed and accuracy than the exact source, provided it is lawful and verifiable. You are free to automate the process with Python, Selenium, BeautifulSoup, or any other reliable scraping stack as long as the output meets the quality bar. Deliverables • A clean, de-duplicated CSV or Excel file containing the four required fields for each lead • A brief note on the data sources and the method you used so I can replicate or audit if needed • Verification step completed (random sample cross-ch...

    $48 / hr Average bid
    $48 / hr Avg Bid
    20 bids

    I need the full, clean dataset of every ...unique ID) • School Name (विद्यालय का नाम) • School Category (Primary / Upper Primary / Secondary / Higher Secondary) • Management Type (Government) No extra fields are required, and you are free to name the CSV however you like. The crucial acceptance criteria are 100 % accurate UDISE codes, zero duplicates, and full district/block coverage. Feel free to choose whichever scraping stack—BeautifulSoup, Selenium, Scrapy, or your own blend—gets the job done quickly and reliably; I care only about correctness and the 24-hour turnaround. Payment is a fixed ₹500 released on successful verification of the dataset. If you’re confident you can meet the accuracy and time demands, I’m ready to award immed...

    $10 Average bid
    $10 Avg Bid
    25 bids

    I need a robust, well-structured Python script that can automatically scrape data from a target site (details shared in chat once an NDA is accepted). The script should be able to navigate through multiple pages, handle dynamic content or AJAX if present, and output clean, structured data to CSV or JSON. Please build it with widely supported libraries such as Requests/BeautifulSoup for static pages or Selenium/Scrapy if JavaScript rendering is required; I’m open to your recommendation as long as the final code is clear and fully commented. A lightweight virtual-env setup and a short README explaining how to run the script are important so I can reproduce the results on my own machine. Deliverables: • Complete, error-free Python script • or Pipfile • READM...

    $378 Average bid
    NDA
    $378 Avg Bid
    28 bids

    ...— I need to audit any row. - Reference numbers (watches) and model names (bags) in raw format, never truncated — they are the matching keys. - Deliver original currency, never pre-converted. - 10-year coverage as continuous as sources allow — **flag any gaps**. - Structured output only (CSV / Excel / JSON). No PDFs or screenshots. ## Skills I'm looking for - Python web scraping (requests / BeautifulSoup / Playwright or similar), handling JS-rendered pages and embedded JSON. - Structured data cleaning and de-duplication. - Bonus: familiarity with Supabase / PostgreSQL, and with luxury / auction data. ## Before you start — confirm back to me 1. Which sources you'll use per category, and expected row count / coverage per source. 2. Whether any sou...

    $194 Average bid
    $194 Avg Bid
    132 bids

    ...shown in the listing • Seller information (username and any public contact details) • Event details (event name, venue, city, date and any section/row/seat notes) The workflow is simple: identify ticket listings across the entire U.S. market on eBay and Facebook Marketplace, extract the data points above, and deliver them in a single CSV or JSON file each run. A lightweight Python script (BeautifulSoup, Selenium, Scrapy or a comparable solution) that I can schedule on my own server would be ideal, but I’m open to alternative stacks if they meet Facebook’s and eBay’s current anti-bot measures. Acceptance criteria 1. Script runs without manual intervention and finishes a full marketplace sweep in a reasonable time. 2. Output file contains zero ...

    $24 Average bid
    $24 Avg Bid
    32 bids

    ...before it goes live and quickly filter or sort the queue (by score, subreddit, date, etc.). After approval, publishing back to Reddit is required; auto-posting to blog comment sections or other sources can wait until a later phase. Brand awareness is the driving metric, so analytics are “nice to have” but not mandatory for this first milestone. Tech stack is your call—Python with PRAW, BeautifulSoup, and an OpenAI or similar API would work fine—but keep it modular so more sources and scoring signals (sentiment, performance analytics) can plug in later. Deliverables • Source code with clear setup instructions • Deployed MVP accessible via a secure URL • Brief hand-off doc outlining how to extend sources or scoring rules Acceptance...

    $453 Average bid
    $453 Avg Bid
    260 bids

    ...performance; I’ll supply the URLs as soon as we start. The scraper must extract for every athlete: full name, weight class, team or club, most recent match outcome, cumulative win-loss record, points scored, and ranking movement over time. I want the data normalised into a single CSV and a companion JSON feed so it can drop straight into my analytics pipeline. Python is my usual stack, so Scrapy, BeautifulSoup, or a light Selenium layer for the occasional dynamic page all work. Please build in polite rate limiting, user-agent rotation, and a quick retry strategy so the job runs cleanly without stressing the sites. Deliverables • Well-documented source code with setup instructions • One-click script or scheduled task that updates the dataset automatically ...

    $435 Average bid
    $435 Avg Bid
    197 bids

    ...model, with clean, clearly-labelled columns. A standard tabular layout is fine; no special template is required beyond what you see in the small example I have already shared. I wrote an example and explanation and have the excel example sheet ready The final deliverable is: 1. An .xlsx file containing the full dataset, ordered by model name. 2. The script or method you used (Python + BeautifulSoup/Scrapy, Power Query, etc.) so I can rerun it if the site updates. I’ll review by doing a random spot-check against the live site; every sampled figure must match what is currently displayed. If anything is missing or mis-aligned in Excel, I’ll return it for correction. Please include a short note on your chosen scraping approach and an estimated turnaround time ...

    $131 Average bid
    $131 Avg Bid
    46 bids

    ...it straight into my analysis pipeline. I’ll share the exact fields during kickoff, but the scraper must be flexible enough to handle common article elements—headline, body text, author byline, publication date, and source URL—and easy to extend if I add more outlets later. Time is critical. Delivery within 24–48 hours is preferred, so please lean on a proven stack such as Python with Scrapy/BeautifulSoup, Node with Cheerio, or any robust alternative you already master. The script should: • Rotate user agents and accept a proxy list to avoid blocks • Log failed requests for easy reruns • Be clearly commented and organized so I can update selectors myself Deliverables 1. Executable script or notebook with all dependencies noted 2. S...

    $32 Average bid
    $32 Avg Bid
    41 bids

    Website Content Scraping Required I need content to be extracted from a website and organized in a structured format. The task includes scraping text, images (if required), and other relevant information while maintaining accuracy and proper formatting. Requirements...maintaining accuracy and proper formatting. Requirements: - Extract content from the specified website. - Preserve headings, paragraphs, and content structure. - Organize the extracted data in Excel, CSV, or Word (as required). - Ensure the data is clean, complete, and free from duplicates. - Deliver the project within the agreed timeline. Experience with web scraping tools (such as Python, BeautifulSoup, Scrapy, Selenium, or similar) is preferred. Please mention your approach, estimated timeline, and cost in your...

    $212 Average bid
    $212 Avg Bid
    76 bids

    ...Announcement, Specialty, Experience, Experience, Book on, summary, badges and designations, (Clinic schedule, Fee, Clinic Name, Clinic Location,) of all clinics they have. Education, Med School, Residency, Fellowship Training, Certifications, online clinic hours and availability, Online clinic fee, Affiliations You may harvest the information with the tooling of your choice—Python (BeautifulSoup, Scrapy, Selenium), R, or another reliable stack—so long as the final file imports seamlessly. I am flexible on the exact format (Excel, CSV, or database dump); let me know what works best for your workflow and I’ll confirm before we start. Accuracy is critical. I will verify that every profile on the site is represented and that each required field is populat...

    $125 Average bid
    $125 Avg Bid
    227 bids