Using an Array Formula to Find and Count the Maximum Text Occurrences in a Range

Share

Facebook
Twitter
LinkedIn

A week ago Tarun asked a question on the Chandoo.org Forums.

“I have got multiple names in each row and would like to have what name is repeated maximum number of times and how many times?

Eg. Ram, Amita, Obama, Ram, Willi, Ram, Amita, Chandoo, Ram, Willi

Ans: Ram (4 times)”

(The list and answers are edited)

Chandoo responded with a neat Array Formula:

=INDEX(B2:K2,MATCH(MAX(COUNTIF(B2:K2,B2:K2)), COUNTIF(B2:K2,B2:K2),0))  &

” (“&MAX(COUNTIF(B2:K2,B2:K2))&” times)”

Lets take a look inside this and see how it works

 

THE EXAMINATION

The formula has two parts separated by a &

=INDEX(B2:K2,MATCH(MAX(COUNTIF(B2:K2,B2:K2)), COUNTIF(B2:K2,B2:K2),0))

and

&

and

” (“&MAX(COUNTIF(B2:K2,B2:K2))&” times)”

Each part is separate and can be used independently, the & character simply joins the two parts together to make a single string which answers Tarun’s question, Ram (4 times).

Now, lets look at each part.

You can follow along with this forensic examination by downloading the Sample Data File.

 

=INDEX(B2:K2,MATCH(MAX(COUNTIF(B2:K2,B2:K2)), COUNTIF(B2:K2,B2:K2),0))

This is a single Index Function with 2 components, being:

a Range B2:K2 and

a Count  MATCH(MAX(COUNTIF(B2:K2,B2:K2)), COUNTIF(B2:K2,B2:K2),0)

Typically an Index Function uses 3 components

=Index(Array, Row Number,[Column Number])

In this example the Range is a single Row, B2:K2

And so using the Counter in the Row spot has the effect of counting down the first Column and then continuing at the top of the second Column etc

So the formula used:

=INDEX(B2:K2,MATCH(MAX(COUNTIF(B2:K2,B2:K2)), COUNTIF(B2:K2,B2:K2),0))

Is equivalent to:

=INDEX(B2:K2,1,MATCH(MAX(COUNTIF(B2:K2,B2:K2)), COUNTIF(B2:K2,B2:K2),0))

 

Now lets jump ahead to the COUNTIF(B2:K2,B2:K2) bit

If you copy =COUNTIF(B2:K2,B2:K2) to a cell, Press F2 and then evaluate the Formula using F9

You will see that it returns an array. The array is highlighted by the squiggly brackets {  } ‘s

={4,2,1,4,2,4,2,1,4,2}

This is the heart of the solution.

What this is showing us is that for each position in the range B2:K2, the count of how many times that cells value occurs in the range B2:K2

So the formula

=INDEX(B2:K2,MATCH(MAX(COUNTIF(B2:K2,B2:K2)), COUNTIF(B2:K2,B2:K2),0))

Is equivalent to

=INDEX(B2:K2,MATCH(MAX({4,2,1,4,2,4,2,1,4,2}), {4,2,1,4,2,4,2,1,4,2},0))

Looking at the MAX({4,2,1,4,2,4,2,1,4,2}) part, this simplifies to 4, the Maximum value of the array (Remember this line, we’ll come back to it later).

So our simplified formula is now: =INDEX(B2:K2,MATCH(4, {4,2,1,4,2,4,2,1,4,2},0))

Now looking at the MATCH(4, {4,2,1,4,2,4,2,1,4,2},0) part of the equation

You can see that Match is looking for the value 4, in the array {4,2,1,4,2,4,2,1,4,2}, which is the First value , Position 1, the 0 requesting that an exact match is found.

So that MATCH(4, {4,2,1,4,2,4,2,1,4,2},0) is equivalent to 1

So our equation =INDEX(B2:K2,MATCH(4, {4,2,1,4,2,4,2,1,4,2},0))

Is now simplified even more to =INDEX(B2:K2, 1)

Index will then look in B2:K2 and will return the first cell or “Ram” in this example.

 

& “(” & MAX(COUNTIF(B2:K2,B2:K2)) & ” times)”

The second part of the equation is responsible for counting the number of Times Ram occurs and displaying it with some text.

& “(” & MAX(COUNTIF(B2:K2,B2:K2)) & ” times)”

The parts displayed in Red above add the text ( and times) to the Count

Remember the section MAX(COUNTIF(B2:K2,B2:K2)) which was explained above and evaluates to 4 in this case

So the & “(” & MAX(COUNTIF(B2:K2,B2:K2)) & ” times)”

Part evaluates to: ( 4 times)

With the initial & adding it to the text of the first part Ram for the final result – Ram ( 4 times)

 

LEARN MORE ABOUT ARRAY FORMULAS

You can learn more about Array Formulas at the following links:

http://www.cpearson.com/excel/ArrayFormulas.aspx

http://www.databison.com/index.php/excel-array-formulas-excel-array-formula-syntax-array-constants/

http://office.microsoft.com/en-us/excel-help/introducing-array-formulas-in-excel-HA001087290.aspx

 

Chandoo.org has several articles on Array Formulas

http://chandoo.org/wp/tag/array-formulas/

 

FORENSIC FORMULAS

Would you like to see more “Forensic” examination of complex formulas ?

Let us know in the comments below and it may become a regular section at Chandoo.org.

 

Facebook
Twitter
LinkedIn

Share this tip with your colleagues

Excel and Power BI tips - Chandoo.org Newsletter

Get FREE Excel + Power BI Tips

Simple, fun and useful emails, once per week.

Learn & be awesome.

Welcome to Chandoo.org

Thank you so much for visiting. My aim is to make you awesome in Excel & Power BI. I do this by sharing videos, tips, examples and downloads on this website. There are more than 1,000 pages with all things Excel, Power BI, Dashboards & VBA here. Go ahead and spend few minutes to be AWESOME.

Read my storyFREE Excel tips book

Overall I learned a lot and I thought you did a great job of explaining how to do things. This will definitely elevate my reporting in the future.
Rebekah S
Reporting Analyst
Excel formula list - 100+ examples and howto guide for you

From simple to complex, there is a formula for every occasion. Check out the list now.

Calendars, invoices, trackers and much more. All free, fun and fantastic.

Advanced Pivot Table tricks

Power Query, Data model, DAX, Filters, Slicers, Conditional formats and beautiful charts. It's all here.

Still on fence about Power BI? In this getting started guide, learn what is Power BI, how to get it and how to create your first report from scratch.

33 Responses to “Show Months & Years in Charts without Cluttering”

  1. eladberko says:

    Very CooOOOoool 🙂

  2. JP says:

    Would it work if I merely change the display format for the dates, or do they actually need to be retyped in that format (Nov, Dec, etc)?

    ps- it's only about 34 donuts per month, or slightly more than 1 per day. Yum!

  3. Jon Peltier says:

    To make it work automatically when you create a chart, delete the labels above the Year and Month columns, but keep the label above the Y data (Donuts). The blank cells tell Excel that the first row and first two columns (indicated by the blanks) are special, so it uses the first row for series names an the first two columns for X axis labels.
     
    This is better than the other kind of donut chart, but you'll soon be carrying a big donut around your midsection.

  4. Erin Smith says:

    First off, thank you Chandoo for being respectful and taking out the "Jesus" comment. Not that I'd threaten to kill you, or start world-wide riots, or make you go into hiding if you didn't (as OTHERS would; wink, wink, nudge, nudge)... I just really appreciate your respectulness and consideration; so thank you. I was meaning to write you about it, but when I came to your site you'd already made the edit... so again, thank you!

    Secondly, I wanna say I think there's an easier way to do what you are demonstrating. I've got a pivot chart with months of data and all I had to do was right-click the x axis and then select "format axis", under "Axis Options" there's a check-box that says "Multi-level Category Labels". The chart I was able to do this on was a pivotchart however so maybe it wouldn't be that easy for a non-pivotchart.

    Anyway, love the site. Keep up the good work. Thanks also for being so open about your success, it's very encouraging and motivating.

    God (aka Jesus) Bless. 🙂

  5. Terry Dukes says:

    Hi Chandoo - great site! Another option to save space is to simply rotate the orientation of the text by 90 degrees, so the dates read vertical rather than horizontal. However, I like the elegance of your solution also.

  6. Kien Leong says:

    Hey Chandoo -- Great tip. Only yesterday I was working through some strange behaviour with formatting dates in PivotCharts. Seems the axes never want to cooperate. This is a neat and elegant solution I hadn't thought of using. May need to abandon pivotcharts to use formulas like that, but if we use dynamic named ranges, no big sacrifice.

    BTW, whatever did you do to get your site blocked in China? Never heard of regime change by a grass-root spreadsheet movement. Maybe your ISP is hosting some problem sites. Chandoo.org is certainly worth it for me to fire up the VPN, but I'm sure you would lose a lot of other visitors from the middle kingdom.

  7. Kapil says:

    Chandoo ... pls help.. the link is blocked over here... pls can you put the regular link... 🙂

  8. Chandoo says:

    @JP... Excel Axis formatting is linked to cell formatting by default. So you can just have the dates which are formatted to look like months (mmm).

    @Erin: It was not my intention to mock anyone's faith or religion. I just used the word as it is quite common. I decided to remove it as I got 2 emails from readers requesting for the same.

    Also, the pivot charts take pivot table groupings by default, so you need not do any of the above while making charts from pivot tables.

    @Kein: I am not sure why Chinese authorities decided to block my site. I wish they would actually look at the content instead of blocking sites based on simple text matching rules.

    @Kapil: The file is mirrored here: http://chandoo.org/img/d/date-axis-months-years-trick.xls

  9. Prateek says:

    Cool, really cool...

  10. SS says:

    Nice one Chandoo,

    Also would like to mention abt useful method while creating dynamic charts.

    In any chart where in the months keep on adding - instead of changing the range for the chart every time we add a month, we can actually format the months as dates (probably 1st of every month) still keep the format as "mmm" AND while selecting the data, we can select a huge rows (date column) once and for all, and the chart adjusts automatically with the data that we entered. So next month when I enter Dec's data, I need not change the source data of the chart, however it automatically adjusts.

    Hope I made sense.!

    Regards,
    SS

  11. Tom says:

    Thanks, Chandoo! This is a great tip - one that I will definitely put to use. I typically have an axis with mmm yy format, aligned vertically, but this will definitely look a bit cleaner (except in cases where the chart is too small for the axis labels to be displayed horizontally, even without the mmm yy on one line). Thanks again!

    Tom

  12. Josh says:

    Chandoo,
    Thank you for the posts you are very diligent not to mention very helpful. I would like to know how to get the separation lines on the axis? For example your candy sales chart has longer lines separating east and west how do you format that?

    Thanks for being very awesome!
    -Josh

  13. Alvaro says:

    Hi Chandoo, we can look the formulas because there is a message:"Unsupported features".
    Could you send a diferent Link ?
    Thanks.

  14. Matt says:

    @SS But what if you've got formulas in the data block (i.e where you would enter static data for the month of december)? My chart now shows #N/A #N/A in the axis with no data for all future dates.

    Chandoo, I've got a dynamic range set up showing #N/A errors for future dates. The MMM-DD date format format in row works fine, but when I use YYYY and MMM in two rows, the axis shows #N/A #N/A for all future dates with no data. How would you go about keeping those future months hidden?

  15. Jon Peltier says:

    Matt -

    In order for the axis to automatically extend to the dates within the range and ignore #N/A at the end, you need a date-scale axis, and for this you need to use one column with the complete date, not two columns with year and month.

    If you want to use two columns, you need to generate Names in the worksheet which define ranges only as long as the number of months. I have a review of dynamic chart approaches in http://peltiertech.com/WordPress/dynamic-chart-review/ and a whole category on my blog at http://peltiertech.com/WordPress/category/dynamic-charts/. Chandoo also has examples of his own on this site.

  16. Ethan says:

    How do you make a dynamic chart out of this?
    I can't get the axis labels range right.
    I tried something like this:
    =OFFSET(REPORT!$H$10:$I$10;0;0;COUNTA(REPORT!$H$10:$I$100);1)

    Any idea?

  17. Jon Peltier says:

    Ethan -
     
    Your offset formula defines a range 1 row in size, but the technique here requires 2 rows. Your definition should end with
     
    ;2)
     
    instead of
     
    ;1)

  18. Ethan says:

    Thanks Jon,
    Got it working now

  19. Neal says:

    Great! Now, is there any way to do this directly in Powerpoint? I don't like having linked excel files, so I create the graphs right inside Powerpoint, any way to do this there? I tried and was unsuccessful.

    Thanks.

  20. Joe says:

    Cool tip Chandoo......thanks

  21. [...] extract year and month from dates to avoid a mess in our stock chart. Chandoo has a great post: Show Months & Years in Charts without ClutteringIn cell B2:=YEAR(D2)In cell B3:=IF(YEAR(D3)=YEAR(D2), "", YEAR(D3))Cell C2:=IF(TEXT(D2, [...]

  22. Bilal says:

    Hi there,
    I have got a data ranging for 3 years. I want to show a chart which shows Jan of 2011, 2012 and 2013 together side by side; then Feb11, Feb12 and Feb13 side by side, then Mar11, Mar12 and Mar13, and so on until December.
    Please help. Thanks.

  23. Down With This Sort Of Thing says:

    Hi there

    Very good solution this. I have another question on it, though. How do you format the X-axis with monthly gaps (ie, with labels "Jan 2012", "Apr", "Jul", "Oct", "Jan 2013", "Mar", etc), when you're dealing with a data series with weekly or daily data points? The Axis Options dialogue box doesn't appear to offer "Date axis" as an option under the "Axis Type" section.

    I've managed to do it in one case with weekly data by setting the interval between tick marks at 13 -- the approximate number of weeks in a quarter -- to get 3-month intervals. But this wouldn't work if I wanted to show 1-month intervals, or had a more detailed daily data series to work with.

  24. Herro says:

    Any luck getting the dates to work on a scatter graph? I'm only getting numbers. Works fine on line graphs though.

  25. Apoorve says:

    How can we do the vice versa? i.e. on the x-axis showing year on the level 1, and months on level 2.
    I wanted to build these kind of axis labels for 5 years, with year on top and months at the bottom, but it should form in such a way that the seperating lines should seperate the entire data set only at December of each year, and no lines in between any month.

  26. Carlos says:

    Like!!
    Three times already today I have used this website and saved a ton of work time in researching excel tricks.

    Suggestion: Why not have a "like" or "this article was useful to me" button. That way you can see what is most useful by your users and maybe generate more content based on those "likes".

    Just saying. Thanks again and you're doing a great job!

  27. Haj says:

    Thanks for the tip. However, I couldn't download your file. The link is broken.

  28. JeteMc says:

    Thank You for taking the time to post this tip. I hope that you have a blessed day.

  29. Tom says:

    The link does not work properly and I'm not sure how to actually get the graph to display like this, its frustrating me a tonne. I cant work out what to google either to find an answer elsewhere! 🙁

  30. parag says:

    Is this possible with waterfall chart. Data hereunder -

    Years Abbrevation Amt

    2020 BEG 2,006
    REV 1,950
    EMP 1,058
    DM (3,244)
    OOE 1,078
    OPMT 182
    AB (638)
    END 2,392
    2021 REV 8,534
    EMP 67
    DM (2,142)
    OOE (3,120)
    OPMT 510
    AB 1,008
    END 7,249

Leave a Reply