Your guide to getting data entry done for your business
Data entry is an important task, but choosing the wrong solution can seriously harm your company's productivity.
Data Extraction is the process of extracting data from a variety of sources for further analysis. A Data Extractor is someone who helps businesses and organizations gain insight from their data and create descriptive and predictive models. They specialize in finding patterns and relationships that guide decisions and uncover meaningful information. Through carefully crafted queries and processes, our Data Extractors can transform raw data into a useful format that can be used for reporting, analytics, machine learning and more.
Here's some projects that our expert Data Extractors made real:
When you partner with an experienced team of Freelancer's Data Extractors you can access valuable insights from your data that can guide decisions, uncover opportunities and create predictive models with new data sources. Our experts can help you unlock deeper insights with advanced filtering methods and complex coding. Explore the full range of possibilities with our talented community of professionals, capable of delivering comprehensive solutions tailored to your needs.
Ready to launch your very own project on Freelancer.com? We invite you to try us out and hire our experienced Data Extractors to make your design goals a reality. Let their creativity, skill, and proficiency bring something special to your project!
จาก 138,803 รีวิว ลูกค้าให้คะแนน Data Extractors 4.9 จาก 5 ดาวData Extraction is the process of extracting data from a variety of sources for further analysis. A Data Extractor is someone who helps businesses and organizations gain insight from their data and create descriptive and predictive models. They specialize in finding patterns and relationships that guide decisions and uncover meaningful information. Through carefully crafted queries and processes, our Data Extractors can transform raw data into a useful format that can be used for reporting, analytics, machine learning and more.
Here's some projects that our expert Data Extractors made real:
When you partner with an experienced team of Freelancer's Data Extractors you can access valuable insights from your data that can guide decisions, uncover opportunities and create predictive models with new data sources. Our experts can help you unlock deeper insights with advanced filtering methods and complex coding. Explore the full range of possibilities with our talented community of professionals, capable of delivering comprehensive solutions tailored to your needs.
Ready to launch your very own project on Freelancer.com? We invite you to try us out and hire our experienced Data Extractors to make your design goals a reality. Let their creativity, skill, and proficiency bring something special to your project!
จาก 138,803 รีวิว ลูกค้าให้คะแนน Data Extractors 4.9 จาก 5 ดาวI have a batch of PDF-based clinical notes and I want an automated routine—ideally powered by Anthropic’s Claude or a comparable large-language-model pipeline—that will (1) confirm each file is indeed a clinical note and (2) pull out two key data groups: • Patient information (full name, DOB, medical record number, and any other standard demographics present) • Every physician referenced in the note, including those listed in the “Cc” section or mentioned elsewhere in the narrative Source files may vary in layout, so the parser has to cope with scanned text (OCR may be required), mixed fonts, and occasional handwritten annotations. I can supply a small, representative sample for calibration and a larger set once the script is stable. Please retu...
I receive customer POs in a fixed-layout PDF that mixes plain text with structured tables. I need a small utility that will: • Parse each new PDF automatically, capturing every product line and the surrounding transaction details (quantities, prices, dates, PO number, etc.). • Add a configurable margin—3 % for now—to the unit price before the data is forwarded. • Push the final dataset straight into ERPNext through its REST API (or other methods may be discussed) so an outbound PO to our chosen vendor is created without manual intervention. I am comfortable running a Python script on a small Linux server and can provide ERPNext API keys plus sample PDFs. Clean, well-commented code, along with a brief read-me and test run that shows the data inside ERPNext, will...
This task involves typings into the application I will supply, ensuring every character matches the source exactly. Source files will be shared through email attachments or a cloud-storage link—whichever proves most convenient once we begin. You will get important the images perfectly and esuring the settings are fully accureate Please provide code :DATAGUTU if you read this. No computer programmers. NO data scientists. NO ai gurus or SEO mareketings. No ENGLISH no JOB
I need access to the CAN bus communication of a 2010 Toyota Prius, specifically to diagnose errors in the battery management system (BMS). This is crucial for troubleshooting and ensuring the hybrid battery operates efficiently. Key requirements include: - Access to CAN bus data - Ability to extract and interpret error codes from the Voltage Sensor or monitor the sensor to ensure it is operating correctly Ideal skills and experience: - Experience with automotive CAN bus systems - Knowledge of hybrid vehicle battery management systems - Proficiency in reading and analyzing CAN bus data Please provide a detailed approach and timeline with your bids.
I need data entry support to pull text from two documents — one is Excel spreadsheet and other is Word file —and transfer that information into a single PDF form that I will provide. Each form record must be populated only after you confirm the text matches across both sources, so a keen eye for consistency is essential. You’ll be working with text data and some numbers but no Math is involved. Accuracy outweighs speed, but I do expect steady progress and a short daily update so I can track completion. Deliverable • Fully completed forms with text verified against both document sources, ready for my final review. If you’re comfortable navigating Excel, Word and PDF viewers and can keep everything confidential, let’s get started right away.
My Zomato and Swiggy accounts hold years of purchase data that I can no longer view through the normal dashboards. The goal is strictly to retrieve that order history—nothing more. Credentials are still active, yet the platforms’ standard export tools fail or deliver incomplete results, so I need an ethical-hacking–oriented approach that respects the terms of service while uncovering every past transaction. Here’s what I need from you: • A complete, chronologically ordered list of every order ever placed on both Zomato and Swiggy, delivered in CSV or JSON. • A short, plain-language summary of the technical steps you took, so I can reproduce the process if the issue reappears. You’re free to use browser-based scraping, authenticated API calls, or ...
More details: What format are the source files in? Both PDF and Images What do you need the data typed into? Both MS Word and Excel What kind of data is being typed from the PDFs/Images into Excel? Text data
I have full-device backups created on Android 13 with Samsung Smart Switch as well as the matching Google Takeout archive. I need a small utility—or a clear, repeatable script—that digs into those encrypted/compressed files, extracts three very specific datasets, and spits them out in a clean, usable format. The targets are browser history, app history (install/uninstall and usage metadata), and system or per-app settings. All other content such as photos, contacts, or messages can be ignored. The device is my own and can load up using USB debugging mode to access if necessary You are free to choose the stack you prefer; Python with sqlite3/pandas, a lightweight Kotlin/Java CLI, or even a cross-platform Node.js solution are perfectly fine as long as they run on Windows or Lin...
I have a set of scanned images that combine narrative text with financial figures. Your task is to transcribe every element into a clean, well-structured Excel workbook so the numbers remain fully usable for analysis while the surrounding descriptions are preserved word-for-word. The numeric portion is strictly financial data—amounts, dates, account names, subtotals, and totals—so absolute accuracy and correct placement in the sheet are critical. The text portions (headings, notes, explanations) must appear in the corresponding rows or columns exactly as they do in the source images. Please work directly in Microsoft Excel (no CSV), maintain the original order of items, and keep all currency symbols, commas, and decimal points intact. I will supply the scanned files as high-r...
I need a repeatable way to pull ticket-listing data from both eBay and Facebook Marketplace, covering concert, sporting and theatre events anywhere in the United States. Each week the scraper should cycle through the two platforms, capture every new or updated listing, and return a clean dataset that includes: • Price shown in the listing • Seller information (username and any public contact details) • Event details (event name, venue, city, date and any section/row/seat notes) The workflow is simple: identify ticket listings across the entire U.S. market on eBay and Facebook Marketplace, extract the data points above, and deliver them in a single CSV or JSON file each run. A lightweight Python script (BeautifulSoup, Selenium, Scrapy or a comparable solution) that I ca...
I need a Python developer to create an automation script and also a website running specifically for data extraction. The script should efficiently gather and organize data from specified sources. Ideal skills and experience: - Proficiency in Python - Experience with automation and scripting - Knowledge of data extraction techniques and tools - Ability to handle various data formats (e.g., JSON, CSV, XML) - Attention to detail and problem-solving skills Please share relevant past work and experience.
I’m building a production-grade large-language model for Kurdish, covering both Sorani and Badini, and I need an experienced fine-tuner to guide the technical core of the project. The first milestone is our data pipeline. The dataset structure is critical: I want clearly separated CPT and SFT splits, stored as clean JSONL, with solid deduplication and quality filters baked in rather than patched on later. You’ll help me design that pipeline end-to-end instead of just running someone else’s script. Next comes training. We are using QLoRA on top of Hugging Face Transformers with PEFT and, ideally, Unsloth for speed-ups. I need hands-on advice across the entire stack—model configuration, training optimisation and hyper-parameter tuning—so my in-house engineer c...
I want a concise, verifiable pull-out of one Indian company’s latest financial statements straight from the MCA portal. The scope is limited to the Balance Sheet, Income Statement, and Cash-Flow Statement, plus any notes in the filings that reveal subsidiaries or overseas branches. I am not looking for director or registration details, only the financial picture. You may work directly on the MCA site, through XBRL downloads, or with any parsing tools you prefer, as long as the final data is clean and traceable back to the original AOC-4 or MGT-7 PDFs. Please return: • A spreadsheet (Excel or Google Sheets) with each statement on a separate tab, figures properly formatted and labeled. • A short PDF memo summarising key numbers and listing every subsidiary / foreign branc...
## Objective Develop an automated workflow that monitors Microsoft Outlook for criminal discovery notification emails from `noreply@`, downloads the linked discovery documents, identifies the correct Clio Manage matter using the court case number in the email subject, and uploads the documents into the appropriate matter without manual intervention. ## Current Process 1. Receive an email from `noreply@`. 2. Open the email. 3. Click one or more links to download discovery documents. 4. Save the downloaded files. 5. Search for the appropriate matter in Clio Manage. 6. Upload each document into the correct matter. 7. Repeat this process for every discovery email received. This process is currently manual, repetitive, and time-consuming. ## Desired Automated Workflow * Monitor Outlook fo...
Preciso de um script que navegue pelo Ifood Brasil e extraia, para cada restaurante encontrado, o nome, o endereço completo (incluindo cidade e coordenadas se disponíveis) e o cardápio integral com preços e descrições. Esses são os únicos campos obrigatórios; avaliações e outras métricas podem ficar de fora por agora. Aceito receber o resultado em qualquer formato — CSV, banco de dados ou planilha Excel — portanto escolha o meio mais prático dentro da sua stack. O importante é que eu consiga abrir o arquivo ou conectar-me ao repositório e já ver os dados prontos para análise. O scraper deve: • percorrer múltiplas cidades brasileiras; &bull...
I need the text content pulled from a set of PDF files and transferred into a clean, editable format—CSV or Excel works best for me, but I’m open to your recommended structure if it preserves every character exactly as it appears. The work is strictly text-based; there are no tables, images, or graphics to worry about. Here’s what I expect: • Open each PDF and extract all text accurately, including headings, paragraphs, and any footnotes. • Keep the original order and formatting cues (e.g., section titles, bullet points) so the final file is easy to read and reference. • Flag any unreadable or ambiguous sections so I can double-check them quickly. • Deliver the compiled file along with a brief summary of the tools or methods you used for transpa...
I have a set of PDFs that contain structured tables, and I need those tables reproduced accurately in classic XLS format—not XLSX or CSV. Your task is straightforward: extract every table from each PDF and place it into an Excel spreadsheet while preserving the original row and column layout, numeric precision, and any header formatting that appears in the source. Efficiency matters, but faithfulness to the source matters more; merged cells, multi-line headers, or footnotes that sit inside the tables should all make the trip intact. If you use tools like Adobe Acrobat Pro, Able2Extract, Tabula, or a custom Python script with pdfplumber and openpyxl, that’s fine—just let me know so I can replicate the process later if needed. Deliverable: one clean .xls file per PDF ...
I have a collection of PDFs and scanned images that need to be converted into clean, well-structured files in both Excel and Word. Every piece of information must be lifted exactly as it appears, without spelling or formatting slips, and then grouped by clear categories so the final sheets are easy to navigate. Here’s what the job involves: • Carefully read each source file, transpose the contents into Excel tables and mirrored Word documents, and double-check that figures, punctuation, and spacing match the originals. • Keep the new files organised by category names I’ll supply up front, using separate tabs or headings where appropriate. • Perform your own quality review before handing the work back so I don’t have to hunt for errors. I’ll share th...
โปรดลงทะเบียน หรือเข้าสู่ระบบ เพื่อดูรายละเอียด
I have a list of subreddits and need a one-off extraction of three specific data layers from each of them: • post titles with the body text • every comment thread • publicly available user info tied to those posts and comments (username, karma, join date where the API allows) •any details relative to any contact information for individuals that work in SBA EIDL service department Please return everything in a clean, well-structured CSV—one file per subreddit or a single consolidated file with clear subreddit, post, and comment identifiers. A reusable script is important to me. Python with PRAW, Pushshift, Scrapy, or any Reddit-API friendly tooling is fine as long as you document the install steps and rate-limit handling so I can rerun it later without...
I need a reliable web-scraping solution that pulls up-to-date market statistics focused on individual wrestlers. The target sites are public sports portals and federations that list match results, rankings, and season performance; I’ll supply the URLs as soon as we start. The scraper must extract for every athlete: full name, weight class, team or club, most recent match outcome, cumulative win-loss record, points scored, and ranking movement over time. I want the data normalised into a single CSV and a companion JSON feed so it can drop straight into my analytics pipeline. Python is my usual stack, so Scrapy, BeautifulSoup, or a light Selenium layer for the occasional dynamic page all work. Please build in polite rate limiting, user-agent rotation, and a quick retry strategy so th...
Θέλω να αυτοματοποιήσω πλήρως τη διαδικασία αντιγραφής προϊόντων από ένα υπάρχον eshop τρίτου σε δικό μου κατάστημα. Στόχος μου είναι να δίνω απλώς το URL του ξενόγλωσσ&omic...
I accidentally wiped several important WhatsApp conversations from my Android phone, and since then WhatsApp has run its scheduled cloud backups to Google Drive multiple times. The deletions happened over a month ago, so the current backup no longer contains the chats I need. I want to know whether those earlier, overwritten messages can still be extracted—either from an older Google Drive snapshot, residual local *.crypt12* files, or any other forensic avenue—and, if so, to have the complete chat threads restored or at least exported in a readable format (TXT, PDF, or HTML). You should already be comfortable working with Android data recovery, ADB, WhatsApp key database decryption, and Google Drive backup parsing. If root access or a temporary unlock bootloader is required, ...
I have a set of records sitting in an online database that I need transferred into a clean, well-structured Excel spreadsheet. The data is straightforward data entry—mostly standard fields that include both words and numbers—so accuracy is everything. You’ll log into the database with the credentials I provide, copy each record exactly as displayed, and place it in the corresponding columns of the Excel file I’ll share. No formatting tricks are required beyond keeping column headers consistent; I simply need the information moved over without errors or omissions. Please be ready to start right away and return the completed spreadsheet within 24 hours. If you spot inconsistencies while copying, flag them in a separate tab so I can review. High attention to deta...
I have a collection of handwritten records that mix text descriptions with numerical figures, and I need every line faithfully transcribed straight into my database. Accuracy is critical—names, dates, codes, quantities, and any notes must move over exactly as they appear in the originals. You will receive high-resolution scans of the documents. Your task is to read each page, interpret the handwriting, and enter the information directly into the database tables I provide (MySQL via phpMyAdmin, though I can export a CSV template if that is easier for you to edit offline). Key points: • Mixed content: both words and numbers on almost every line • Source is entirely handwritten, so a good eye for varying scripts is essential • I will run spot-checks; any field l...
converting of pdf tabular data into indexable editable searchable excel format. data @ 1317 records.
I have a batch of survey responses that must be transferred quickly and accurately into a structured Google Sheet. The raw data is currently in mixed formats—mainly PDF exports and a few Word documents—and I need each response entered into the appropriate columns I’ll share. Scope of work • Create or update my provided Google Sheet template without altering the existing header structure. • Enter every survey response exactly as written, preserving capitalization, punctuation, and any special characters. • Double-check for typos or misplaced fields before marking the row complete. • Flag any illegible or ambiguous entries in a separate “Notes” column so I can review them later. Acceptance criteria • 100% of supplied records ar...
I need help pulling strictly numerical data from a handful of online databases and entering it into a spreadsheet template I will supply. Everything you collect will come directly from those databases—no PDFs, scans, or web-scraping from regular sites—so the task is about careful extraction, accurate transcription, and a quick cross-check rather than guessing or reformatting text. You will log in, locate the specified tables or dashboards, copy the required figures (row by row), and paste them into the correct columns of my Excel/Google Sheets file. A few fields may require you to apply the database’s built-in filters or basic calculations before recording the result, but there is no heavy analysis involved. Deliverable • A fully completed spreadsheet with every ...
I receive a recurring Excel/CSV file of 8,000 records as an email attachment that holds customer data. 1. Create a separate record locator for each distinct address. Some record locators will have only one customer attached, others will have more. 2. Create 5 files with data drawn from each record locator.
I have a batch of printed documents that have already been converted to high-resolution scanned images. Your task is to extract every word from those scans and enter it verbatim into a structured text file (Word or plain-text) while preserving paragraph breaks and basic formatting. No numerical tables or mixed data—this is purely text transcription, so accuracy in spelling, punctuation, and line order matters most. You will receive a shared folder containing the scans. Feel free to use any reliable OCR tool—ABBYY FineReader, Adobe Acrobat, or Google Vision—but please proof-read the output and correct mistakes before submission. Manual typing is equally acceptable if that gives cleaner results; the final file just needs to be 100 % faithful to the source pages. Deliverab...
I have an old FreeBSD installation that boots; I copied the entire disk image onto a USB drive and now need the document files extracted. The priority is .doc and .pdf documents that were kept inside a handful of specific directories (I’ll send their full paths once we start). The job is simple in scope but requires solid FreeBSD / UFS-2 recovery know-how. You will need to: • Mount or image-clone the USB safely (read-only first). • Locate the requested directories and pull every .doc and .pdf you can find. • Verify file integrity—open the recovered documents or run hash checks so I know they are intact. • Package the results in a tidy archive and hand it back to me, along with a brief log of what you did (commands, tools, and any issues encountere...
I need 2,000 lines of mixed text-and-numeric information lifted from a batch of scanned PDFs and placed into a single Excel spreadsheet. Every character must match the source—accuracy is the top priority because the client will perform a full audit before any payment is cleared. Schedule • All 2,000 rows must be handed over by 19 July, 6:00 PM (IST). • Payment of ₹1,700 will be released on 3 August after the review confirms the file is error-free. You will receive the PDFs and a ready-made column template as soon as you accept, so you can begin straight away. Deliverable • One .xlsx file containing every record in the original order with zero transcription errors. Only reach out if you are completely confident you can finish on time with perfect precision.
I need help collecting data from clinical letters. The data includes: - Patient information - Diagnosis details - Treatment plans The extracted data should be delivered in a spreadsheet format. The clinical letters will be provided as digital text files. Ideal skills and experience: - Experience with data extraction and processing - Proficiency in spreadsheet software (e.g., Excel, Google Sheets) - Attention to detail and accuracy
Hi, I'm Ayush. I've built similar training materials before – mostly for small teams who need to get better at Excel without sitting through boring theory. I can put together a complete package for your 3-hour workshop. Here's what I'll deliver: PowerPoint deck – Clean, visual, and easy to follow. Covers core formulas (IF, VLOOKUP/XLOOKUP, SUMIFS, TEXT functions) and pivot tables. Each slide has a clear takeaway and a prompt for the hands-on exercise. Demo files and datasets – Realistic IT data. Think log files with timestamps and error codes, ticket metrics (open/closed by priority, SLA breaches), system inventory lists with hardware specs. Everything your team can manipulate during the session. Facilitator notes – Step-by-step. Timings for e...
I need 2,000 lines of mixed text-and-numeric information lifted from a batch of scanned PDFs and placed into a single Excel spreadsheet. Every character must match the source—accuracy is the top priority because the client will perform a full audit before any payment is cleared. Schedule • All 2,000 rows must be handed over by 19 July, 6:00 PM (IST). • Payment of ₹1,700 will be released on 3 August after the review confirms the file is error-free. You will receive the PDFs and a ready-made column template as soon as you accept, so you can begin straight away. Deliverable • One .xlsx file containing every record in the original order with zero transcription errors. Only reach out if you are completely confident you can finish on time with perfect precision.
I need an automated scraper that gathers data from several news sites in near real-time. The tool should loop through a list of URLs I will provide, respect each site’s where possible, and export the captured information to CSV or JSON so I can feed it straight into my analysis pipeline. I’ll share the exact fields during kickoff, but the scraper must be flexible enough to handle common article elements—headline, body text, author byline, publication date, and source URL—and easy to extend if I add more outlets later. Time is critical. Delivery within 24–48 hours is preferred, so please lean on a proven stack such as Python with Scrapy/BeautifulSoup, Node with Cheerio, or any robust alternative you already master. The script should: • Rotate user agent...
I have a PDF made up entirely of tables that I need transferred into an editable Excel workbook. Accuracy is critical—I want every row, column, and cell value reproduced exactly where it appears in the original file, with the same column order, headings, and basic layout preserved. No extra styling is required beyond what already exists; the goal is a faithful, spreadsheet-ready replica that I can work with immediately. Please use whichever extraction method you’re most comfortable with—manual entry, automated OCR, or a hybrid approach—as long as the final Excel file mirrors the PDF tables perfectly and is free from typos or shifted data. If you spot any inconsistencies in the source, flag them so I can clarify. Deliverable: a single .xlsx file reflecting the PDF ...
Facebook Group Outreach Automation Specialist Needed We're looking for someone to build an automated (or semi-automated) system that can: - Scrape relevant Facebook Groups — identify and extract members from groups where our target audience (creative agencies, UGC creators, video editors, content studios) is active. - Automate outreach — send personalized initial messages to identified members/leads at scale. - Manage replies automatically — a system to detect, categorize, and respond to inbound replies (e.g., interested / not interested / questions), with handoff to a human for qualified leads. - Report on a safe daily volume — tell us realistically how many groups/members can be scraped and how many messages can be sent per day/account without triggering Fa...
WHEN YOU BID, PLEASE BID FOR THE ENTIRE PROJECT - NOT PER PAGE!! URGENT DATA CONVERSION PROJECT: 24-HOUR DEADLINE WHO IS SUITABLE FOR THIS JOB (CRITICAL CRITERIA): Expert in Advanced Data Cleaning: You must have extensive experience using regular expressions (Regex), text-parsing tools, or advanced scripting to split messy addresses. If you plan to type this out manually row-by-row, do not apply. You will not finish in time. Flawless Attention to Detail: You must know how to cleanly isolate data fields without leaving broken lines, trailing commas, or raw code formatting artifacts inside cells. High Availability: You must be able to start immediately upon hiring and dedicate the necessary hours to deliver the complete file on time. Proven Track Record: You must show past proof of man...
I have a collection of PDFs that contain plain text fields I need transferred into an Excel spreadsheet. Accuracy is critical—the information will feed directly into a reporting tool, so even minor typos or mis-aligned cells will cause problems downstream. You will receive: • A folder of the source PDF documents. • A pre-formatted Excel file with clearly labeled columns that match the fields in each PDF. Your task is straightforward: open each PDF, copy the required text exactly as it appears, and paste it into the corresponding cells of the spreadsheet. No re-typing is needed if you can copy/paste cleanly, but every entry must be double-checked so the final sheet mirrors the PDFs character for character. Acceptance criteria • 100 % of PDFs processed and their...
I have between one and ten PDF files that each contain tables sharing the exact same column order and layout. I need every row and column transferred into Excel with 100 % accuracy—no missing cells, no merged-cell surprises, and no typos. Please give me back either a single workbook with one sheet per source file or individual workbooks that mirror the originals; I am flexible as long as the structure is preserved. The finished spreadsheets must be ready for filtering, sorting, and basic formulas, so numeric fields should remain numeric and text fields text. Timing is tight: I will supply the PDFs as soon as we start, and I need the completed Excel files returned within three calendar days. To keep us both confident in the result, I will spot-check totals and random rows once y...
I have a sizeable spreadsheet that needs a thorough clean-up in Microsoft Excel. The raw file contains duplicates, inconsistent formats, stray blanks and a few obvious typos. I’m looking for someone who can methodically work through the sheet—using whatever mix of formulas, Power Query, or manual checks you prefer—to deliver a polished, analysis-ready workbook. Accuracy is more important than speed, so the timeline is flexible. When you respond, please highlight your experience handling similar Excel data-cleaning jobs; examples of past successes will help me gauge fit. Final deliverable: the cleaned Excel file plus a short note summarising the main issues you found and the fixes you applied.
I have a batch of printed documents that have already been converted to high-resolution scanned images. Your task is to extract every word from those scans and enter it verbatim into a structured text file (Word or plain-text) while preserving paragraph breaks and basic formatting. No numerical tables or mixed data—this is purely text transcription, so accuracy in spelling, punctuation, and line order matters most. You will receive a shared folder containing the scans. Feel free to use any reliable OCR tool—ABBYY FineReader, Adobe Acrobat, or Google Vision—but please proof-read the output and correct mistakes before submission. Manual typing is equally acceptable if that gives cleaner results; the final file just needs to be 100 % faithful to the source pages. Deliverab...
I have roughly 5 000 PDFs that are a mix of Bills of Lading and Commercial Invoices. I need a reliable script—Python preferred, but any language is fine—that can open each file, read the key parties on the document, and aggregate everything into a single Excel workbook. The script must capture: • Shipper details • Receiver details • Broker details Accuracy matters more than speed; some files are machine-readable, others are scanned, so you may have to blend text parsing with OCR (think , PyPDF2, Camelot, Tesseract, or any stack you trust). The output should be a clean .xlsx file with one row per shipment and clearly labeled columns for each data point. Please send a brief but detailed proposal that explains: – The libraries or tools you will use ...
I already have a Python-based Scrapy project and now need dedicated spiders that can harvest data from the mobile apps Rabbitmart (Egypt), Voo (Egypt) and Oscar (Egypt). The crawlers should pull every publicly available endpoint (or reverse-engineered one) required to deliver: product details, user reviews, pricing information, the underlying product IDs, any active promotions or discounts and—where the apps expose it—current stock levels. All captured information must be written to clean, well-structured JSON files so that I can feed it straight into my existing pipeline. I’m still unsure whether the apps will demand login or token-based authentication; therefore, please build the spider with the flexibility to plug in credentials, headers or session cookies if we disc...
I need a clean, ready-to-use Excel file drawn directly from that covers every Indian IT company incorporated or that filed documents during May, June and July 2026. The sheet has to hold at least 10,000 unique records and be organised one row per company so I can filter and work quickly. For each record I must see: • Company name • City, full postal address, state and stated business type • Corporate phone number and generic company email ID • Director’s full name plus a working director-level email ID Keep the data strictly real and current to the period specified; no fabricated or recycled contacts. I’ve chosen “By company” as the organising principle, so be sure all subsidiary details line up clearly under each firm. The final hand-o...
Website Content Scraping Required I need content to be extracted from a website and organized in a structured format. The task includes scraping text, images (if required), and other relevant information while maintaining accuracy and proper formatting. Requirements: - Extract content from the specified website. - Preserve headings, paragraphs, and content structure. - Organize the extracted data in Excel, CSV, or Word (as required). - Ensure the data is clean, complete, and free from duplicates. - Deliver the project within the agreed timeline. Experience with web scraping tools (such as Python, BeautifulSoup, Scrapy, Selenium, or similar) is preferred. Please mention your approach, estimated timeline, and cost in your proposal.
I need a DXL scripting expert for data migration tasks in IBM DOORS Classic. Ideal Skills and Experience: - Proficiency in DXL scripting - Experience with IBM DOORS Classic data modelling - Data extraction and Data migration skills, Data mapping - Familiarity with CSV, XML, or JSON formats Please provide examples of similar work done.
I have a collection of PDFs, scanned documents, and a few legacy Excel files that all need to end up in one clean, well-structured spreadsheet. The content is mixed—text fields, numbers, the occasional ID code—so each cell must be copied or transcribed exactly as shown, then double-checked for accuracy. You will pull everything into both formats I use: an .xlsx file (for offline backup) and a synced Google Sheets version that mirrors it. Before delivery, please run your own error scan for typos, mis-alignments, or misplaced decimals and keep all information strictly confidential. When you respond, include a quick sample or link to past work that proves you have handled similar multi-source data entry projects. I am less interested in a long proposal and more interested in s...
I have a collection of PDFs, scanned documents, and a few legacy Excel files that all need to end up in one clean, well-structured spreadsheet. The content is mixed—text fields, numbers, the occasional ID code—so each cell must be copied or transcribed exactly as shown, then double-checked for accuracy. You will pull everything into both formats I use: an .xlsx file (for offline backup) and a synced Google Sheets version that mirrors it. Before delivery, please run your own error scan for typos, mis-alignments, or misplaced decimals and keep all information strictly confidential. When you respond, include a quick sample or link to past work that proves you have handled similar multi-source data entry projects. I am less interested in a long proposal and more interested in s...
Data entry is an important task, but choosing the wrong solution can seriously harm your company's productivity.
Learn how to hire and collaborate with a freelance Typeform Specialist to create impactful forms for your business.
A complete guide to finding, hiring, and working with a skilled freelance typist for your typing projects.