Data & Analytics

Data Cleaning & Preparation

Messy spreadsheets and exports de-duplicated, standardised and merged into one reliable dataset, with a reusable cleaning script.

Custom quote after a short consultation · Timeline agreed in your quote

Overview

Data cleaning fixes the problems that make a spreadsheet or export unreliable: duplicate records, inconsistent spellings, mixed date formats, blank or misplaced values and columns that do not line up between files. Clean data is the foundation for any dashboard, analysis or CRM import.

RF ETS works through each file systematically. We profile the data first to find the issues, agree the cleaning rules with you, then apply them using Python or Power Query so every change is repeatable rather than made by hand. Where several files describe the same records, we match and merge them into one table.

The Advanced package covers up to ten files and up to 250,000 rows, with merging, de-duplication, standardising, a data-quality report showing what was found and changed, and a reusable cleaning script so the same rules can be applied to next month's export. Typing data from scanned or handwritten documents is not included.

Who it’s for

  • Teams preparing a CRM, ERP or inventory import from several old spreadsheets
  • Analysts who spend more time fixing data than analysing it
  • Businesses merging customer or product lists after a system change
  • Anyone receiving the same messy export every month

Problems it solves

  • Remove duplicate and conflicting records before an import or campaign
  • Combine several files into one consistent table
  • Know exactly what was wrong with the data and what was changed
  • Repeat the same cleaning on future exports without starting again

What RF ETS delivers

  • Cleaned dataset in your preferred spreadsheet or CSV format
  • Merged and de-duplicated master file (Professional and Advanced)
  • Data-quality report listing issues found and fixes applied
  • Reusable cleaning script in Python or Power Query (Advanced)
  • Short notes on the cleaning rules used

How we work

  1. Profile the data

    We review each file to find duplicates, format problems, gaps and mismatched columns.

  2. Agree cleaning rules

    We confirm how to treat duplicates, standard formats and merge keys with you.

  3. Clean and merge

    We apply the rules with Python or Power Query and merge files into one dataset.

  4. Report and hand over

    You receive the clean data, the quality report and, on Advanced, the reusable script.

What we need from you

  • The files to be cleaned, in spreadsheet, CSV or similar digital format
  • The target format or system the clean data is going into
  • Any business rules, such as which record wins when duplicates conflict
  • A contact to confirm ambiguous cases
Pricing

How pricing works

This work varies too much for fixed packages. We scope it with you, then send a written quote with price, timeline and deliverables before any work starts.

  1. Tell us what you need

    Use the quote form or book a consultation and describe your goal, constraints and deadline.

  2. We scope it with you

    We review your material and agree deliverables, assumptions and what is out of scope.

  3. Written quote

    You receive a fixed price or milestone plan, timeline and revision terms before any work starts.

Not included

  • Data entry from scanned or handwritten documents
  • Third-party costs: domains, hosting, paid plugins and themes, software licences, API or AI-model usage fees, data-provider credits
  • Work outside the written scope agreed before work starts (handled as an add-on or a custom quote)
  • Ongoing support after the delivery and launch-support window unless a support package is bought
Custom solution

Need something different?

Tell us what you need, and we’ll prepare a solution and pricing based on your requirements.

FAQ

Data Cleaning & Preparation: common questions

How is data cleaning priced?

This service uses fixed packages that you can book, based on the number of files and rows. If your data is slightly over a tier limit, extra rows can be added as an add-on. See the packages above.

Can you type up data from scanned or handwritten documents?

No. This service works with data that is already digital, such as spreadsheets, CSV files and system exports. Manual data entry from scans, photos or handwriting is excluded.

Will you delete records we might need?

No record is removed without an agreed rule. Duplicates and dropped rows are listed in the data-quality report (Professional and Advanced), so you can see what changed and ask for anything to be restored.

What if we find more issues after delivery?

Revisions within the agreed files and rules are included as listed for your package. If new files or new cleaning rules are needed, we handle them as an add-on or a short quote.

Can we reuse the cleaning on future exports?

Yes, on the Advanced package. You receive a reusable script in Python or Power Query that applies the same rules to new files with the same structure, along with notes on running it.