Looking up when data won’t play nice – few more alternatives

Share

Facebook
Twitter
LinkedIn

Recently, we discussed about the case of unwieldy data and how we lookup what we want using formulas like SUMIFS. Today, let us learn few more ways to solve the same problem.

First, a re-cap of the problem:

Here is a data-set:

2D Lookup problem - Example dataset

The problem – build a lookup formula

And the problem. Oh, simple. Write a lookup formula to find how many customer walk-ins we have on any given day.

In the previous article, we discussed how to use SUMIFS to solve this problem. There were several amazing & awesome solutions shared by our readers in the comments section too.

Suitable structure spawns simple solutions

Poorly structured is the 2nd biggest problem of analysts. The first one is not enough coffee. That is why there is a dictum in the data analytics world.

Structure is everything

So, we can easily solve our lookup problem, if our data were to magically re-arranged in 2 column fashion – Data & Value.

Transforming data to solve problem easily - Example

This transformation can be done in 2 ways:

Option #1: Transforming Data – Using Formulas

We can use data fetching formulas like OFFSET or INDEX to re-arrange data in 2 columns.

Assuming,

  • Our 2D data is in a named range data,
  • There are running numbers starting with 0 in the cell J5

We can use below formula to fetch first column:

=IFERROR(INDEX(data,2*(INT(J5/7))+1,MOD(J5,7)+1),"")

for the second column, below formula works:

=IFERROR(INDEX(data,2*(INT(J5/7)+1),MOD(J5,7)+1),"")

How does this formula work?

I will explain the formula for first column. Deciphering 2nd column formula is your homework.

Here is the formula again: =IFERROR(INDEX(data,2*(INT(J5/7))+1,MOD(J5,7)+1),"")

Before understanding the formula, let’s take a minute to examine the structure of  our raw data.

  • Odd rows contain dates
  • Even rows contain values
  • There are 7 columns in total
  • So to get the first date, we need to go to row 1 (first odd number), column 1
  • To get the first value, we need to go to row 2 (first even number), column 1
  • But to get 8th date, we need to go to row 3(2nd odd number), column 1
  • So on

Let’s go from inside out.

  • 2*(INT(J5/7))+1 portion: This gives row number (ie odd number). J5 refers to running number and its value is 0. So we get 2*(INT(0/7))+1 = 1
    • This will be 3 when J5 becomes J12 (ie 8th date)
  • MOD(J5,7)+1 portion: This gives column number. It will result in values 1 thru 7 in a cyclical fashion. Thanks to MOD.
  • INDEX(data, ..., ...) portion: Now that we have both row & column numbers, INDEX formula kicks in and gets the corresponding date.
  • IFERROR(INDEX(...),"") portion: This is to help in case we ran out of all dates & values in our INDEX formula. Read about IFERROR here.

Once you have the formulas for first date & value, simply drag them to get rest of the values.

Option #2: Transforming data – Using VBA

VBA Macros are perfect for scenarios like this. Usually transformation is something you need to do every-time you import data from external systems. So simply write a macro that can do this automatically.

Assuming our data is in the range data and the first cell of our extraction range is startHere, you can use below macro:


Sub rearrangeData()
    'takes the values in DATA named range and rearranges them
    'from the named cell startHere

    Dim cell As Range, i As Long, j As Long, evenRow As Boolean, firstRow As Long
    
    i = 0
    j = 0
    firstRow = Range("data").Cells(1).Row
    
    For Each cell In Range("data")
        evenRow = (cell.Row - firstRow + 1) Mod 2 = 0
        If evenRow Then
            Range("startHere").Offset(j, 1).Value = cell.Value
            j = j + 1
        Else
            Range("startHere").Offset(i, 0).Value = cell.Value
            i = i + 1
        End If        
    Next cell
End Sub

How does this macro work?
Before jumping in to the lines of code and demystifying the logic, Let’s understand what we need to do:

  1. For each cell of data,
    1. If it is in odd row, put the cell data in Date column at end
    2. Else, put the cell data in Value column at end
  2. Repeat

This is what our code is trying to do.

Let’s examine the For Each loop, as this is the most critical part of our macro.

  • For each cell in the range data
  • We check if we are in evenRow using simple arithmetic on row numbers
  • If we are in evenRow then
    • We put the cell value in row j (number of values so far), column 2
    • We increment j
  • Else
    • We put the cell value in row i (number of dates so far), column 1
    • We increment i
  • Close the IF condition
  • We check for next cell in the data range

Advantages of Transformation over SUMIFS approach

Both options for transforming data have few advantages:

  • They work with any type of data (unlike SUMIFS, which works only for numeric lookups and has few other issues)
  • Once data is restructured, you can do other types of analysis like creating pivot tables, adding extra calculated columns etc. easily.

Download Example Workbook

Click here to download example workbook that shows original SUMIFS solution, both options for transforming data & few other formulas. Play with it to learn more. Check out the code by pressing ALT+F11.

How would you transform data?

My favorite techniques for transforming data are – VBA, formulas, Power Query, pivot tables & SQL. Depending on the situation, time availability, where my data is, I choose one of these options to scrub my data.

What about you? How do you clean up / scrub data like this? Please share you thoughts & tips with us in comments.

Instructions for washing your dirty data

If your work involves scrubbing dirty data, check out below tutorials too:


Facebook
Twitter
LinkedIn

Share this tip with your colleagues

Excel and Power BI tips - Chandoo.org Newsletter

Get FREE Excel + Power BI Tips

Simple, fun and useful emails, once per week.

Learn & be awesome.

Welcome to Chandoo.org

Thank you so much for visiting. My aim is to make you awesome in Excel & Power BI. I do this by sharing videos, tips, examples and downloads on this website. There are more than 1,000 pages with all things Excel, Power BI, Dashboards & VBA here. Go ahead and spend few minutes to be AWESOME.

Read my storyFREE Excel tips book

Overall I learned a lot and I thought you did a great job of explaining how to do things. This will definitely elevate my reporting in the future.
Rebekah S
Reporting Analyst
Excel formula list - 100+ examples and howto guide for you

From simple to complex, there is a formula for every occasion. Check out the list now.

Calendars, invoices, trackers and much more. All free, fun and fantastic.

Advanced Pivot Table tricks

Power Query, Data model, DAX, Filters, Slicers, Conditional formats and beautiful charts. It's all here.

Still on fence about Power BI? In this getting started guide, learn what is Power BI, how to get it and how to create your first report from scratch.

27 Responses to “Sum of Values Between 2 Dates [Excel Formulas]”

  1. dexter says:

    I would apply a filter and use function subtotal, with option 9. This way you can see multiple views based on the filter.

  2. Michael Azer says:

    hey Chandoo, the solutions you proposed are very efficient, but if I wanted to be fancy I would do it this way .. the references are as your example workbook.
    =SUM(INDIRECT("C"&(MATCH(F5,B5:B95)+4)):INDIRECT("C"&(MATCH(F6,B5:B95)+4)))

  3. Luke M says:

    I like things simple:
    =SUMIF(B5:B95,">="&F5,C5:C95)-SUMIF(B5:B95,">"&F6,C5:C95)

  4. Matt S says:

    use something like: =SUM(OFFSET(B1,0,0,DATEDIF(A1,D1,"d")))
    and have D1 be the date that I want to sum to.

  5. Tom J says:

    In Excel 2003 (and earlier) I'd use an array formula to calculate either with nested if statements (as shown here) or with AND.

    {=SUM(IF(B5:B95>F5,IF(B5:B95<F6,C5:C95,0),0))}

    Note that I truly made this for BETWEEN the dates, not including the dates

  6. Andrew says:

    I turned the data set into a table named Dailies.
    I named the two limits StartDate and EndDate.

    And used an array formula:

    {=SUM((Dailies[Date]>=StartDate)*(Dailies[Date]<=EndDate)*Dailies[Sales])}

  7. Frank Linssen says:

    If I would still be using the old Excel I would do it as follows:

    SUMIF($B$5:$B$95,"<="&H6,$C$5:$C$95)-SUMIF($B$5:$B$95,"<"&H5,$C$5:$C$95)

    Works as simple as it is.

    Regards

  8. ikkeman says:

    =sum(index(c:c,match(startdate,c:c,1)+1):index(c:c,match(enddate,c:c,1))

  9. ikkeman says:

    =sum(index(c:c,match(startdate,b:b,1)+1):index(c:c,match(enddate,b:b,1))

  10. ram says:

    Great examples and thanks to Chandoo. You have simplified my work.

  11. Rony says:

    Hi! great tips I have found in your page, have you seen this
    http://runakay.blogspot.com/2011/10/searching-in-multiple-excel-tabs.html

  12. [...] I'm not sure I understand your question fully, but have a look at this: Sum of Values Between 2 Dates [Excel Formulas] | Chandoo.org - Learn Microsoft Excel Online [...]

  13. Amanda says:

    Thank you! Thank you! Thank you!

  14. abdalurhman says:

    =SUMIF(A2:A11;">="&B13;B2:B11)-SUMIF(A2:A11;"<"&A11;B2:B11)

  15. Eliza says:

    awesome... thank yoo Chandoo!

  16. dockhem says:

    which is most efficient and fast, if all are efficient ?

  17. jmassiah says:

    Thank you for this formula, I've just spent ages trying to find something to work on my data, I knew it would be possible! Don't care if others think there are easier/other ways to do it, you explained it so I understood it and could apply it to what I was doing so I'm happy!

  18. Nagaraju says:

    The above said example is awesome for calculating values between dates,

    can you pls let know how to calculate sale values if we have 10 sales boys for
    ex: 1,rama
    2,krishna
    3,ashwin
    4,naga
    5,suresh

    how much rama sale value between 1/jan/2015 to 10/jun/15
    how much krishna sale value between 10/jan/2015 to 15/july/2015
    i think you understood can you pls let me know the formula for how to calculate the sale between diffrent sale man sale value from master data file

    Thanks,
    Nagaraju

  19. Viv says:

    Hi

    I have a list of people's names in column A, I have a list of dates in column B which records the dates they have been off sick, in column C I have either 1 if it is a full sick day or 0.5 if it is a half day.

    What I would like to do is to add up the number of dates a specific person has been off within two dates.

    For example, I want to look at my list of names and to find Joe Bloggs (column A), then add up all his sick days (column C). The start date will be in cell E1 and the end date will be in F1.

    If this possible using SUMIFS?

    List of names are in range A2:A100

    List of dates in B2:B100

    List of sick days (either 0.5 or 1 in C2:C100

    The start date is in cell E2

    The end date is in cell F2

    Your help would be greatly appreciated.

    • Loknathan says:

      Yes, with the help of SUMIFS you can have the solution.
      Note: you need have an extra col. D2 where you will input Name of the person.
      =SUMIFS(C2:C100,A2:A100,D2,C2:C100,">="&E2,C2:C100,"<"&F2)

      Col. A Col. B Col. C Col.D Col. E Col. F
      Name Date Sales
      ABC 28-Jun-11 1 MNO 28-Jun-11 25-Sep-11
      XYZ 29-Jun-11 0.5
      MNO 30-Jun-11 1
      PQR 1-Jul-11 1

      • Loknathan says:

        Typo ERROR / Correction in formula:
        Yes, with the help of SUMIFS you can have the solution.
        Note: you need have an extra col. D2 where you will input Name of the person.
        =SUMIFS(C2:C100,A2:A100,D2,B2:B100,">="&E2,B2:B100,"<"&F2)

  20. Viv says:

    Hi

    I have a list of people's names in column A, I have a list of dates in column B which records the dates they have been off sick, in column C I have either 1 if it is a full sick day or 0.5 if it is a half day.

    What I would like to do is to add up the number of dates a specific person has been off within two dates.

    For example, I want to look at my list of names and to find Joe Bloggs (column A), then add up all his sick days (column C). The start date will be in cell E1 and the end date will be in F1.

    If this possible using SUMIFS?

    List of names are in range A2:A100

    List of dates in B2:B100

    List of sick days (either 0.5 or 1 in C2:C100

    The start date is in cell E2

    The end date is in cell F2

    Your help would be greatly appreciated.

    Viv

  21. AC says:

    Thanks for this - it solved the problem that I was having. However can someone please explain to me why the "" needs to be around >= and <= as well as why we need to add & in order for the formula to work? Thanks in advance!

  22. Ufoo says:

    This formula works perfectly as well. Any ideas?: =SUM(INDEX(C5:C95,MATCH(H5,B5:B95,1)):INDEX(C5:C95,MATCH(H6,B5:B95,1)))

  23. Ufoo says:

    ikkeman had posted the same thing.

  24. murray says:

    I am trying to sum total a range of cells between date ranges ie column n has $ amounts column d has the transaction dates ie 1/3/2015 or 25/3/2015 or 25/4/2015 column b has the text saying drp or distribution - reinv

    In another cell I am trying to sum or total (in column n) with the value of a range of different dates (column d) that contain different text (column b) ie cell n48 is 50, n65 is 85, n165 is 36

    with the dates ie cell d48 is 1/3/2015, d65 is 25/3/2015 and d165 is 25/4/2015

    with different text that says drp or distribution - reinv ie cell b48 is drp, b65 is distribution - reinv, b165 is drp

    If I wanted to sum the amounts between 1/3/2015 to 31/3/2015 with drp then the total would be 50. Also if I wanted to sum the amounts between 1/4/2015 to 30/4/2015 with drp the sum total would be 36 If I wanted to sum the amounts between 1/3/2015 to 31/3/2015 with drp and distribution - reinv the sum would be 115

    What would the formula be for these different questions

    hope you can help, it has been driving me nuts and cant work it out

Leave a Reply