Extracting Unique, Duplicate and Missing Items using Formulas [spreadcheats]

Share

Facebook
Twitter
LinkedIn

Often I wish Microsoft had spent the effort and time on a data genie (and a set of powerful formulas) that can automate common data cleanup tasks like extracting duplicates, makings lists unique, find missing items, remove spaces etc. Alas, instead they have provided features like clippy which are intrusive to say the least.

So as part of our second installment of spreadcheats we will learn how to tackle few of the most common data processing tasks:

Getting Unique Items from a List of Cells

There are 3 simple ways to do this:

  1. Using Advanced Data Filter
  2. Using countif() and auto filter
  3. Using formulas as described here

Assuming you have data as shown in the picture aside (and wishing you will have customers like those):

  • First add a column to the left of the list. Here we will use formulas to fill numbers based on the uniqueness of the cell next to it.
  • Essentially our formula should generate numbers in increasing order as long as the corresponding item is unique and not increase the number otherwise.
  • So the formula for order column can be like this: =IF(COUNTIF(list-upto-that-point, current element)=1,previous-order+1, previous-order)
    See the example below:

    remember, the first cell order is 1.
  • See how we are using both absolute and relative references to fetch the counts.
  • Now add another column to the right of the list, here we will fetch unique items.
  • We will use vlookup() to fetch each of the 12 unique items. The formula goes like this:
    =VLOOKUP(running number,$B$4:$C$22,2,FALSE)
    You can wrap the vlookup() with if() formula to avoid seeing #value errors.

That is all. Using this method you can extract unique items froma list.

Eliminating Doubles from a List

There are 2 ways in which you can find and remove duplicates(doubles) in excel lists with ease:

  1. Using countif() and then auto-filter
  2. Using formulas

The process for finding duplicates using formulas is same as that of finding unique items.

Instead of writing COUNTIF(list-upto-that-point, current element)=1, we now write COUNTIF(list-upto-that-point, current element)=2. Also the first element’s count should be changed to zero.

Once done the list should look like what you see on the side.

Finding Missing Items by comparing one list with another:

Even though this might seem like a different challenge, it is infact same as the above techniques. You need to use countif() to compare first list’s elements with second list. How? that is your home work.

Download and see these formulas in action:

Still having some doubts? Download the excel tutorial – unique & duplicate items and learn by poking around.

Facebook
Twitter
LinkedIn

Share this tip with your colleagues

Excel and Power BI tips - Chandoo.org Newsletter

Get FREE Excel + Power BI Tips

Simple, fun and useful emails, once per week.

Learn & be awesome.

Welcome to Chandoo.org

Thank you so much for visiting. My aim is to make you awesome in Excel & Power BI. I do this by sharing videos, tips, examples and downloads on this website. There are more than 1,000 pages with all things Excel, Power BI, Dashboards & VBA here. Go ahead and spend few minutes to be AWESOME.

Read my storyFREE Excel tips book

Overall I learned a lot and I thought you did a great job of explaining how to do things. This will definitely elevate my reporting in the future.
Rebekah S
Reporting Analyst
Excel formula list - 100+ examples and howto guide for you

From simple to complex, there is a formula for every occasion. Check out the list now.

Calendars, invoices, trackers and much more. All free, fun and fantastic.

Advanced Pivot Table tricks

Power Query, Data model, DAX, Filters, Slicers, Conditional formats and beautiful charts. It's all here.

Still on fence about Power BI? In this getting started guide, learn what is Power BI, how to get it and how to create your first report from scratch.

13 Responses to “Data Validation using an Unsorted column with Duplicate Entries as a Source List”

  1. Vipul says:

    Pivot Table will involve manual intervention; hence I prefer to use the 'countif remove duplicate trick' along with 'text sorting formula trick; then using the offset with len to name the final range for validation.

  2. Rich says:

    if using the pivot table, set the sort to Ascending, so the list in the validation cell comes back alphabetically.

  3. Kieranz says:

    Hui: Brillant neat idea.
    Vipul: I am intrigued by what you are saying. Please is it possible to show us how it can be done, because as u said Hui's method requires user intervention.
    Thks to PHD and all
    K

  4. sam says:

    Table names dont work directly inside Data validation.
    You will have to define a name and point it to the table name and then use the name inside validation
    Eg MyClient : Refers to :=Table1[Client]
    And then in the list validation say = MyClient

  5. Vipul says:

    Kieranz,
    Pls download the sample here http://cid-e98339d969073094.skydrive.live.com/self.aspx/.Public/data-validation-unsorted-list-example.xls
    Off course there are many other ways of doing the same and integrating the formulae in multiple columns into one.

  6. Vipul says:

    Pls refer to column FGHI in that file. Cell G4 is where my validation is.

  7. Kieranz says:

    Vipul:
    Many thks, will study it latter.
    Rgds
    K

  8. [...] to chandoo for the idea of getting unique list using Pivot tables.  What we do is that create a pivot table [...]

  9. Playercharlie says:

    @Vipul:

    Thanks, that was awesome! 🙂

  10. Vipul says:

    @Playercharlie Happy to hear that 🙂

  11. Enrique says:

    Great contribution, Hui. Solved a problem of many years!

  12. FARIS says:

    Thanks to you, A LOT

  13. Mohamed says:

    Hi Hui,
    Greeting
    hope you are doing well.
    I'm interested to send you a private vba excel file which i need to show detail of pivot in new workbook instead of showing in same workbook as new sheet.

    Please contact me on muhammed.ye@gmail.com

    Best Regards

Leave a Reply