Extracting Unique, Duplicate and Missing Items using Formulas [spreadcheats]

Share

Facebook
Twitter
LinkedIn

Often I wish Microsoft had spent the effort and time on a data genie (and a set of powerful formulas) that can automate common data cleanup tasks like extracting duplicates, makings lists unique, find missing items, remove spaces etc. Alas, instead they have provided features like clippy which are intrusive to say the least.

So as part of our second installment of spreadcheats we will learn how to tackle few of the most common data processing tasks:

Getting Unique Items from a List of Cells

There are 3 simple ways to do this:

  1. Using Advanced Data Filter
  2. Using countif() and auto filter
  3. Using formulas as described here

Assuming you have data as shown in the picture aside (and wishing you will have customers like those):

  • First add a column to the left of the list. Here we will use formulas to fill numbers based on the uniqueness of the cell next to it.
  • Essentially our formula should generate numbers in increasing order as long as the corresponding item is unique and not increase the number otherwise.
  • So the formula for order column can be like this: =IF(COUNTIF(list-upto-that-point, current element)=1,previous-order+1, previous-order)
    See the example below:

    remember, the first cell order is 1.
  • See how we are using both absolute and relative references to fetch the counts.
  • Now add another column to the right of the list, here we will fetch unique items.
  • We will use vlookup() to fetch each of the 12 unique items. The formula goes like this:
    =VLOOKUP(running number,$B$4:$C$22,2,FALSE)
    You can wrap the vlookup() with if() formula to avoid seeing #value errors.

That is all. Using this method you can extract unique items froma list.

Eliminating Doubles from a List

There are 2 ways in which you can find and remove duplicates(doubles) in excel lists with ease:

  1. Using countif() and then auto-filter
  2. Using formulas

The process for finding duplicates using formulas is same as that of finding unique items.

Instead of writing COUNTIF(list-upto-that-point, current element)=1, we now write COUNTIF(list-upto-that-point, current element)=2. Also the first element’s count should be changed to zero.

Once done the list should look like what you see on the side.

Finding Missing Items by comparing one list with another:

Even though this might seem like a different challenge, it is infact same as the above techniques. You need to use countif() to compare first list’s elements with second list. How? that is your home work.

Download and see these formulas in action:

Still having some doubts? Download the excel tutorial – unique & duplicate items and learn by poking around.

Facebook
Twitter
LinkedIn

Share this tip with your colleagues

Excel and Power BI tips - Chandoo.org Newsletter

Get FREE Excel + Power BI Tips

Simple, fun and useful emails, once per week.

Learn & be awesome.

Welcome to Chandoo.org

Thank you so much for visiting. My aim is to make you awesome in Excel & Power BI. I do this by sharing videos, tips, examples and downloads on this website. There are more than 1,000 pages with all things Excel, Power BI, Dashboards & VBA here. Go ahead and spend few minutes to be AWESOME.

Read my storyFREE Excel tips book

Overall I learned a lot and I thought you did a great job of explaining how to do things. This will definitely elevate my reporting in the future.
Rebekah S
Reporting Analyst
Excel formula list - 100+ examples and howto guide for you

From simple to complex, there is a formula for every occasion. Check out the list now.

Calendars, invoices, trackers and much more. All free, fun and fantastic.

Advanced Pivot Table tricks

Power Query, Data model, DAX, Filters, Slicers, Conditional formats and beautiful charts. It's all here.

Still on fence about Power BI? In this getting started guide, learn what is Power BI, how to get it and how to create your first report from scratch.

18 Responses to “Best Charts to Compare Actual Values with Targets – What is your take?”

  1. Andy Cotgreave says:

    Great post. I can't vote, though, because the answer I want to put down is "it depends". As with all visualisations, you've got to take into account your audience, your purpose, technical skills, where it will be viewed, etc.

  2. Jon Peltier says:

    I'm with Andy: It depends. Some I would use, some I might use, some I won't touch with a barge pole.
     
    Naturally I have comments 🙂
     
    The dial gauge, though familiar, is less easy to read than a linear type of chart (thermometer or bullet). It's really no better than the traffic lights, because all it can really tell you is which category the point falls in: red, yellow, or green.
     
    By the same token, pie charts are so familiar, people don't know they can't read them. Remember how long it takes kids to learn to read an analog clock?
     
    Bullet charts don't show trends.
     
    With any of the charts that have a filled component and a marker or ine component, it makes more sense to use the filled component (area/ column) for target, and the lines or markers for actual.

  3. [...] Best Charts to Compare Actual values with Targets (or Budgets … [...]

  4. Tony Rose says:

    I voted for #6 even though I agree with the other comments that it depends.

    The majority of the votes are for the #2, thermometer chart. I still have yet to understand what happens when you are above plan/goal, which was brought up in yesterday's post.

    Also, I agree with Jon in that it would be better to flip the series and make the filled part the target or goal and the line or marker the actual.

    I am also a fan of using text when appropriate if the data is among other metrics in a type of dashboard. Calling it out by saying actual and % achievement is a good option.

  5. Another "it depends" vote. Are you just looking at one or are you comparing a number of targets with actuals? You didn't include a text box. The problem with sentences is that they can get lost in a page of gray text. A text box can call attention to the numbers and line them up effectively.

    I'm with Jon: "Some I would use, some I might use, some I won’t touch with a barge pole" and I'm surprised that some of your readers voted for the last group.

  6. Bob Gannon says:

    Jon says:
    With any of the charts that have a filled component and a marker or line component, it makes more sense to use the filled component (area/ column) for target, and the lines or markers for actual.
    Why does this make more sense? I like 6 the way it is, although I would use a heavy dash for the plan/target marker.

  7. "It depends" is also my take. What I usually try to drill into my clients dashboard design is the fu ndamental difference between spot results (am I on target for this month) and long term trends.. I always try to create 3 different set of graphs to represent real perormance:
    - spot results vs objectives
    - cumulative results vs objectives
    - long-term trend (moving average) mostly) to see where we're going

  8. [...] Best Charts to Compare Actual Values with Targets – What is your take? (tags: excel charts) [...]

  9. Jamie Regan says:

    Jon says:
    With any of the charts that have a filled component and a marker or line component, it makes more sense to use the filled component (area/ column) for target, and the lines or markers for actual.
    Why does this make more sense? I like 6 the way it is, although I would use a heavy dash for the plan/target marker.

    I totally agree, Bob. I would normally favour a line for the target and a column for the actual, you can see quite easily then which columns break through the line, then.

  10. [...] best charts to compare actual values with targets — den Status mal anders zeigen, z. B. als Tacho [...]

  11. zzz says:

    Thermometer charts: "Not appropriate when actual values exceed targets" - this is easily solved by making the "mercury" portion a different color from the border, then you can clearly see where the expected range ends and the actual values keep going.

  12. Godsbod says:

    People seem to knock gauges quite a bit in dashboarding, but trying to show comparison of realtime data between operating sites and targets for each site can easily be done with a bank of gauges that have the optimal operating points at 12 o'clock.

    The human eye is great at pattern stripping, and any deviation of a gauge from the expected 12 position will quickly register with an operator and attract his attention. Using a colour background, or meter edge, will also indicate the sensitivity of a particular site.

  13. […] work laptop I have a favorites folder just dedicated to Excel charts.  Its got things like “Best Charts to Compare Actuals vs Targets” and “Best charts to show progress“. I love me some charts […]

  14. Albert says:

    I am wondering how will the plotting work, for some of the targets which may have been achieved before time. E.g. for the month of Jul the target was 226 and the actual was 219. So the chart will show a deficit in meeting the target by 7 points but what if this 7 may have been completed earlier in month of June. So ideally it not a deficit.

Leave a Reply