The Grammy Bump Chart in Excel

Share

Facebook
Twitter
LinkedIn

The folks at Washington Post made an interesting chart to understand whether winning a Grammy award makes any difference to album sales. Go ahead and browse it if you have not already seen it. Go, I will wait.

Are you impressed?

I really liked this chart. This is what I liked about the chart,

  • It tells a story. [why charts should tell a story]
  • It is an ego chart. We would all instantly search for our favorite artists and learn about how Grammy award changed their album sales.
  • It is a simple chart. No clutter, no gaudy colors, just a bunch of lines and the story is out there.
  • It lets you play. You can hover your most over an artist to see their sales before and after the award, and how much % bump they got.

In fact, I liked the chart so much that I wanted to make it in Excel.

Here is what I came up with:

The Grammy Bump Chart Replica in Excel - Demo

How does the chart work?

1. Data for the chart: The Washington Post guys did not give any details about the source of data. So I manually typed the data myself by looking at their chart. It took a few minutes. But totally worth it. I put the data in 5 columns – Year, Artist, Album, Before and After sales.

2. The chart: The chart is an XY Scatter plot. I took numbers from 0 to 37 (there is a total of 19 years of data – from 1992 to 2010. Each year has 2 data points – before and after). For even numbers I used the before sales and odd numbers I used the after sales. For this I wrote simple INDEX formula with a bit of MOD(). Again, nothing too fancy.

3. Getting the gaps in the chart:

This is the tricky part. By default, if you have 4 points (0,98000),(1,135000), (2,155000), (3, 427000) in the XY Scatter plot, Excel will draw a line connecting all 4. But we want to have a gap between first 2 points and second 2 points. How?!?

Thankfully, there is a simple workaround. You can insert blank rows between 2nd and 3rd row of your data and instantly you will see a gap in the chart. Repeat the same for remaining 18 points.

Before and after adding blank rows - scatter plot

4. Year Selection & Highlighting

This is done using conditional formatting & Worksheet_SelectionChange Event Macro. First, I wrote a simple macro that would change the named range valSelectedYear to the selected year. The code is very simple. You can examine it in the download file.

Then, I used the valSelectedYear to drive the conditional formatting that would fill blue color across the column. As you can guess, the chart is transparent (ie no fill color for both chart area and plot area). So whatever color the cells beneath the chart have, they will show up in the chart too.

5. Creating the Dynamic Legend:

Here I have used picture links to fetch the artist image dynamically. (well, I was too lazy to download the actual images of Nora Jones and U2 etc. So I just used clip art).

Then, I used text boxes to make the dynamic legend, same as the technique demonstrated in smart chart legends & excel product catalog articles.

6. Formatting and aligning everything:

Once the basic setup is ready, I just moved and re-arranged the chart, legend box etc. so that everything looks right.

Grammy bump chart replica in Excel

Download the Chart Workbook:

Click here to download the workbook in Excel 2007 format.

(click here to download the file in Excel 2003 format. I have not tested this, but it should work alright)

Mirror location for the files.

Please note that you must enable macros to select years.

Recommended Reading to make charts like these

How would you have made this chart?

I liked the original chart design and interactivity provided by Washington Post people. So I closely mimicked the same my Excel chart. But you may want to visualize the same data in a different way. So go ahead and download the workbook. It has data (hidden in columns A thru G). Play with it and make your own chart. Post them in comments.

I would love to see how you would have visualized the same information. Especially this type of data has a lot of relevance in business situations, so it would be fun to see your views and learn from each other. Go ahead and chip in.

Thanks to Washington post for the chart. Hat tip to Flowing Data for the link.

Facebook
Twitter
LinkedIn

Share this tip with your colleagues

Excel and Power BI tips - Chandoo.org Newsletter

Get FREE Excel + Power BI Tips

Simple, fun and useful emails, once per week.

Learn & be awesome.

Welcome to Chandoo.org

Thank you so much for visiting. My aim is to make you awesome in Excel & Power BI. I do this by sharing videos, tips, examples and downloads on this website. There are more than 1,000 pages with all things Excel, Power BI, Dashboards & VBA here. Go ahead and spend few minutes to be AWESOME.

Read my storyFREE Excel tips book

Overall I learned a lot and I thought you did a great job of explaining how to do things. This will definitely elevate my reporting in the future.
Rebekah S
Reporting Analyst
Excel formula list - 100+ examples and howto guide for you

From simple to complex, there is a formula for every occasion. Check out the list now.

Calendars, invoices, trackers and much more. All free, fun and fantastic.

Advanced Pivot Table tricks

Power Query, Data model, DAX, Filters, Slicers, Conditional formats and beautiful charts. It's all here.

Still on fence about Power BI? In this getting started guide, learn what is Power BI, how to get it and how to create your first report from scratch.

19 Responses to “How to Distribute Players Between Teams – Evenly”

  1. Roshan Thayyil says:

    An excellent solution, especially for large data sets.

    Another solution without using solver would be to assign the player with the highest score to Team 1, the 2nd to team 2, 3rd to team 3, 4th to team 3, 5th to team 2, 6th to team 1, 7th to team 1 and it continues. This method would end up with a Std Dev of 0.001247219. This works best with a distribution with lower Std Dev for the dataset.

    Full Disclosure: this is not my idea, remember reading something a few years ago. Think it may have been Ozgrid

    • Roshan Thayyil says:

      thinking back I now remember why I read about it. About 10 years back I had to distribute around 300 team members into 25-30 odd teams. Used this method based on their performance scores. I used the method I described to do this and the distribution was pretty fair.

      Solver would have saved me a ton of time though 🙂

  2. I think the issue with you first Solver approach was that you took the absolute value of the sum of team deviations (which should always be zero except for rounding) instead of the sum of the absolute values (which is a reasonable measure of how unbalanced the teams are).

  3. Here's another simple algorithm you could use: you start from the top (with players sorted from high to low), and at each step allocate the next player to whichever team has the smallest total so far. You can implement it dynamically with some formulas so it will update automatically when the data changes.

    If the scores were more widely distributed (so that this might end up with not all teams the same size), you could add a constraint to only pick among the teams which currently have fewest players at each step, or just stop adding to any team when it hits its quota.

    When I tried it on the sample, I got the three teams below, with a STDEV of 0.000942809 (i.e. about half of what Solver got to).

    Team 1: John, Hugo, Tom, Josh, Eric, Zane, Charles, Andrew
    Team 2: Barry, Michael, Kenny, Joe, Xavier, Patrick, Oliver, William
    Team 3: Henry, Steven, Ben, Frank, Kyle, Edward, Cameron, Lachlan

    Thanks for sharing!

    • Ishaan says:

      Hi,
      I was looking at all the solutions and this is closest to what I intended to do. I am dividing a bunch of players into 3 soccer teams. Players availability is also a factor while deciding the teams.
      So the steps the excel needs to do is as follows:
      1) In availability column if "yes" go to next
      2) Equally divide 'Goalkeepers', 'Strikers', 'Defenders' basis their quality
      So the end result gives each 3 teams a balance of players playing at different positions.
      Can this be done on Google spreadsheet with only availability as an input from the user and rest calculates by itself.
      Sorry for asking such a pointed question, but I have been struggling to find a solution for it for sometime now!

      • Robin says:

        Hi Ishaan,

        I am working on a similar problem at the moment, so I am wondering if you ever found a solution and if you are willing to share what you did.

  4. Konrad says:

    Hi everyone, this is a variation of the famous Knapsack Problem https://en.wikipedia.org/wiki/Knapsack_problem.

    I had to use a VBA implementation recently as part of a problem, where we ar trying to allocate teams of an organization into different locations (we are a large company with many different team). The goal was to optimally allocate teams to individual buildings without putting too many teams into one building and not splitting teams apart.
    As we had around 400 teams of different sizes, solver couldn't handle it anymore. Luckily there is a Knapsack algorithm implementation in VBA readily available on the internet :).

    I also went with a heuristic approach first!

  5. Joe Egan says:

    An interesting mathematical solution but what if Eric and Xavier can't stand each other or Patrick is best friends with Steven - the real life problems that effect "even" teams.

    • Hui... says:

      @Joe

      You can add more criteria like
      If Eric and Xavier can't stand each other
      =OR(AND(E15=1,E16=1),AND(F15=1,F16=1),AND(G15=1,G16=1))
      It must be False

      If Patrick is best friends with Steven
      =OR(AND(E5=1,E17=1),AND(F5=1,F17=1),AND(G5=1,G17=1))
      It must be True

      Note that the 2 formulas above are exactly the same
      except for the ranges
      One must be True = Friends
      One must be False = Not Friends

  6. Gustavo Sousa says:

    Nice post Hui!

    I download your workbook and just try to change in options the Precision Restriction from 10E-6 to 10-8 and the Convergence from 10E-4 to 10E-10. The process take almost the same time, but the results was great.

    The standard deviation I got was 0,000471.

    Team 1: John, Tom, Kenny, Frank, Eric, Xavier, Edward, Zane
    Team 2: Steven, Hugo, Ben, Joe, Josh, Oliver, Cameron, William
    Team 3: Barry, Henry, Michael, Kyle, Patrick, Charles, Andrew, Lachlan

  7. Charlie says:

    Great application of Solver! Thanks for the link!

  8. Chuck says:

    Great explanation. Well done... However, I tried with 6 teams of 4 players and solver never did finish.

  9. Akbar says:

    How about vba code for the same data set.
    I have 3 column A B C wherein A has text and B has number Wherein C is blank. And in C1 been the header C2 where I want the name to come evenly distributed the number which is in Column B.
    My Lastcolumn is 1000.

  10. HRMFT says:

    Sorry if I'm being slow here, but how is 'Team Score' calculated? I've gone through the explanation several times but it seems to just appear.

    • Hui... says:

      @Hrmft

      This process uses the Solver Excel addin

      Solver is effectively taking the model and trying different solutions until it gets a solution that meets all the criteria
      Then solver puts the solution into the cell and moves to the next cell

      So yes it appears to "just appear"

  11. Caroline says:

    Hi ! Thank you so much ! Works great 🙂

  12. Jim Cruse says:

    I cannot get the fourth Equation to work in my excel spreadsheet
    You have =($E$2:$G$25=0)+($E$2:$G$25=1)=1 as a SUMIF solution, I have, =($F$2:$H$13=0)+($F$2:$H$13=1)=1 as my solution but it does not work. The only thing I changed is the ranges. Any suggestions?
    Thank you.
    Jim

  13. Jim Cruse says:

    I cannot get the fourth Equation of TURE or FALSE statements to work in my excel spreadsheet You have =($E$2:$G$25=0)+($E$2:$G$25=1)=1 as a SUMIF solution, I have, =($F$2:$H$13=0)+($F$2:$H$13=1)=1 as my solution but it does not work. The only thing I changed is the ranges. Any suggestions?
    Sorry I left some of it out in the previous question,
    Thank you. Jim

Leave a Reply