Learn Statistics & Probability using MS Excel

Share

Facebook
Twitter
LinkedIn

One of the most dreaded courses during my under-graduation is Probability, Statistics & Queuing Theory. We called it PSQT. I struggled to understand the significance and concept of this course as I could barely concentrate in the class. We had a professor, who is probably a genius, but the moment he started the class, I would magically fall in to one of my after-noon naps. When I woke up, we are either in the middle of an elaborate t-test or going thru intricacies of a Markovian queue.

This was all 11 years ago. Later in life, I have embraced the world of probability & statistics. I still fear queues. May be I will get there one day. 😉

A good understanding of statistics & probability theory is necessary if you want to model complex real-life problems using Excel or similar tools. Naturally, Excel has several functions, features & supported add-ins to help you in this area.

Today, I want to share some of this with you. This article is broken down in to 3 parts.

  1. Learning Statistics & Probability using Excel
  2. Downloadable Excel Workbooks to understand
  3. Full blown models & simulations in Excel

#1 – Learning Statistics & Probability Concepts using Excel

Using Excel RAND functions

Excel has several powerful functions (formulas) to generate random numbers, random data. You can combine these functions to generate data that has certain parameters – like a give mean, standard deviation or follows a certain type of distribution.

Go thru Using Excel’s Random Functions for a detailed overview these techniques.

Simulating Dice Throws in Excel

One of the fundamental ways to learn about Probability is to look at dice throws. A dice has 6 faces and on each throw, any of the 6 faces turning up is equally likely. So, we say, each face has 1/6th probability of showing up. If you want to simulate this in Excel, you can use the formula RANDBETWEEN like this, =RANDBETWEEN(1,6). On each run, this formula would throw up a random number between 1 & 6 (including both).

For more, Simulating Dice Throws in Excel

Shuffling a List of Values in Excel

Understanding permutations and combinations is essential when it comes to modeling many real-world problems. Using Excel’s RAND, VLOOKUP and SMALL formulas we can generate a random permutation of a given list of values (in other words – we can shuffle the list).

Shuffling a list of values in Excel

To learn this read, Shuffling a list of values in Excel

Generate Frequency Distribution from Data

Often, when you are analyzing data, you need to understand how the data is distributed. Again, Excel has just the right function for this sort of thing. FREQUENCY(). In this simple tutorial, learn how to use Excel’s FREQUENCY formula to generate frequency distribution of given data.

Calculating Statistical Frequency Distribution in Excel

 

Read Frequency Distributions in Excel

Trend Analysis & Forecasting using Excel

One of the most common applications of statistics is trend analysis & forecasting. Again, Excel shines with a lot of powerful formulas, built-in features and charting tools to help you understand the data & predict future based on that.

Since this is a big topic, we have covered it in 3 parts –

Part 1- Introduction to Trend Analysis & Forecasting: In this, we will learn what is trend analysis & forecasting. We will see manual forecasting technique in Excel. We will use Excel charts to depict our analysis and results.

Trend Analysis & Forecasting using Manual Forecasting Technique in Excel

 

Part 2 – Trend Analysis & Forecasting using Excel Functions

In this second part, we learn about Excel’s functions like LINEST, TREND, FORECAST, SLOPE, INTERCEPT, LOGEST and GROWTH. These powerful formulas can process lots of data and extract the trend information dynamically.

 

Trend Analysis and Forecasting using Excel's functions & charts

 

Part 3 – Trend Analysis & Forecasting using Charts & Macros

In the final part, we talk about how to use Excel chart’s trend analysis & forecasting features to estimate the trend & predict future values based on the data.

Trend Analysis & Forecasting using Excel Charts & VBA

 

We also learn how to use Macros (VBA) to augment Excel chart’s trend-lines with useful information.

Visualizing Distribution of data with Box Plots

Box plots are an excellent way to understand the distribution of data. Unfortunately, there is no direct option to make a box plot from given data in Excel. That is where, this tutorial comes handy.

Box plots in Excel - How to

Learn how to create box plots in Excel.

more on Box plots.

 

#2 Downloadable Excel Workbooks

Learn Basic Statistics & Gaussian Distribution using this Excel Workbook

Glen, one of our long time readers shared this file with me. It lets you perform statistical analysis, quality control analysis, visualize Gaussian distribution based on the data you enter.

Click here to download the workbook.

Gaussian Distribution of Data in Excel

Thanks Glen.

More Downloadable Workbooks

Almost all of the links in this page will take you to detailed articles on Chandoo.org, where you can also find downloadable workbook with examples. So just click thru and learn. 🙂

#3 Full blown models & simulations in Excel

A full blown model lets you learn various statistical concepts, Excel features and how to bring them all together to mimic a real-life situation.

Simulating Deal or No Deal game in Excel

In this simulation of Deal or No Deal, a popular television game, we use basic probability, permutations and Excel formulas features. You will learn how to assign random values to the suit-cases, how to use circular references, how to calculate the banker’s offer.

Simulation of Deal or No Deal game in Excel

 

Simulation of Deal or No Deal game in Excel

Generating Housie / Bingo Tickets in Excel

Housie (Bingo) is a popular recreational game where the tickets contain 15 numbers between 1 to 90, arranged in 10 columns (3×10 grid). First column has numbers between 1 to 9, second column has 10 to 19 so on..

Generating a bingo ticket in Excel is a nice exercise in statistics, permutations and Excel formulas.

Generating Housie / Bingo tickets using Excel

Learn from Bingo / Housie Tickets in Excel

Data Tables & Monte Carlo Simulations in Excel

Excel has powerful features to let us do complex simulations of real world situations. One such feature is called as data table.

The Data Table allows a set of what if questions to be posed and answered simply, and is useful in sensitivity analysis, variance analysis and even Monte Carlo (Stochastic) analysis of real life model within Excel.

The case of Blue Sky Mining Company

To help you learn about data tables, Monte Carlo simulations, we have put together a fictional mining company – Blue Sky co. and analyzed its performance under various assumptions & simulations.

Data Tables & Monte-Carlo Simulations using Excel

To learn about this, visit Data Tables & Monte-Carlo Simulations page.

Modeling & Scheduling a FIFO (First In First Out) Queue in Excel

FIFO queues are very common in life. You can see them at Airports, coffee shops, Apple stores; Except at Airports it is FIFOUYC (FIFO Unless You are Crew).

In this article, we model & schedule a FIFO queue using Excel.

More Full Blown Models & Simulations in Excel

For more examples, check out these links.

Do you use Statistical Concepts for your work?

As a small business owner, a good portion of my work involves statistical analysis, forecasting and simulation. I run estimates for our website traffic, revenues. I run statistical tests (split tests etc.) to optimize our sales pages, website. I estimate when my kids wake up from their nap (based on past experience) and plan my work accordingly. Thankfully, for the last part, I do not use Excel 😀

What about you? Do you use statistical concepts for your work? What are the things you use and how does Excel help you in that? What are your favorite formulas, features and tips? Please share using comments.

Special thanks to Hui & Glen

Many thanks to Hui, our resident Excel ninja for writing many of the articles on statistics, simulation, forecasting & trend analysis.

Special thanks to Glen for sharing the analyze-this file with us.

Say thanks to them if you enjoyed this.

Facebook
Twitter
LinkedIn

Share this tip with your colleagues

Excel and Power BI tips - Chandoo.org Newsletter

Get FREE Excel + Power BI Tips

Simple, fun and useful emails, once per week.

Learn & be awesome.

Welcome to Chandoo.org

Thank you so much for visiting. My aim is to make you awesome in Excel & Power BI. I do this by sharing videos, tips, examples and downloads on this website. There are more than 1,000 pages with all things Excel, Power BI, Dashboards & VBA here. Go ahead and spend few minutes to be AWESOME.

Read my storyFREE Excel tips book

Overall I learned a lot and I thought you did a great job of explaining how to do things. This will definitely elevate my reporting in the future.
Rebekah S
Reporting Analyst
Excel formula list - 100+ examples and howto guide for you

From simple to complex, there is a formula for every occasion. Check out the list now.

Calendars, invoices, trackers and much more. All free, fun and fantastic.

Advanced Pivot Table tricks

Power Query, Data model, DAX, Filters, Slicers, Conditional formats and beautiful charts. It's all here.

Still on fence about Power BI? In this getting started guide, learn what is Power BI, how to get it and how to create your first report from scratch.

20 Responses to “Untrimmable Spaces – Excel Formula”

  1. MF says:

    Hi Chandoo,
    First of all, HAPPY NEW YEAR!!! Wish you and your family another fruitful year ahead.

    To answer your question: Power Query is the best way to trim. 🙂

    Btw, if Power Query is not available, then formula would absolutely do... but did you forget to mention also Char 32?

    One more question: Is the trailing minus meant to be a negative number? Maybe only the sender knows... 🙂

    Cheers,

  2. Duncan Williamson says:

    I know these spaces can be a real pain but these days I advise Excel users to learn and use Flash Fill and that will learn what to do pretty quickly.

  3. David Hager says:

    Highlight range to be cleaned. Then, in Replace, hold down the Alt key and type 0160. Replace with nothing.

  4. Steve Jones says:

    I accomplished this by writing a macro to go through all the possible unprintable characters. Looped through the range.

  5. Ramnath D says:

    I use a different method here. First, I will copy the data from Excel and paste it in a notepad. In Notepad, I will do a Find Blanks (Space " ") and Replace (Empty) with nothing.

    Then you can copy the data from Notepad and paste it back to Excel which will be a perfect number as you desire.

    But Thanks for the formula. Its probably the 2nd out of 8 tricks as Chandoo mentioned. Waiting for the rest among 8 from other users 🙂

  6. Andrew says:

    I don't understand the x's. Why weren't they removed in the formula? Or are they part of some sort of numeric formatting that I'm not familiar with? I saw how you handled the non-breaking spaces and the dashes, but am confused about what role the x's played in all this.

    Thanks!

    • NARAYAN says:

      Hi Andrew ,

      The xs have been used solely to demarcate the actual data text ; thus , without the x in place at the end of text , as in :

      x 4,124,500.00 x

      it would be impossible to know that there are unwanted trailing characters , in this case , after the last 0.

      These xs are not part of the original data text , nor are they used in the formulae ; they are put in only so that readers can visualize the individual items of data as they are in practice. Think of them as imaginary delimiters.

      • Andrew Patceg says:

        Oh, that makes sense! Thank you for the explanation. I had a feeling it was something along those lines.

  7. Mucio says:

    You can type this character using the Keys Alt+0160.
    Very useful to replace this Character using Find and Select resource.

  8. Neva says:

    For many years, my jobs have included ETL tasks and I built this macro to help long, long ago. I tweak it every now and again. Many co-workers, past and present, have it wired to a button on their toolbar.

    Sub Clean_and_Trim()
    'CAUTION: Strips leading zeroes -- do not use on zipcodes, etc.

    If Application.Calculation = xlCalculationAutomatic Then
    Application.Calculation = xlCalculationManual
    Revert = 1
    ElseIf Application.Calculation = xlCalculationManual Then
    Revert = 0
    End If

    For Each Cell In Selection
    For x = Len(Cell.Value) To 1 Step -1
    If Asc(Mid(Cell.Value, x, 1)) = 160 Then
    Cell.Replace What:=Chr(160), Replacement:=" ", LookAt:=xlPart, MatchCase:=True
    End If
    If Asc(Mid(Cell.Value, x, 1)) = 32 Then
    Cell.Replace What:=Chr(32), Replacement:=" ", LookAt:=xlPart, MatchCase:=True
    End If
    Next x
    If Cell.Value "" Then
    Cell.Value = Application.Clean(Application.Trim(Cell.Value))
    End If
    Next

    If Revert = 1 Then
    Application.Calculation = xlCalculationAutomatic
    ElseIf Revert = 0 Then
    Application.Calculation = xlCalculationManual
    End If

    End Sub

  9. Brigitte Calahate says:

    This is awesome! What if you have several characters you need to have removed? What would be the easiest way as I can imagine there are several ways.?

    # - 35
    $ - 36
    - 62
    / - 47
    , - 44
    . - 46
    " - 34
    : - 58

  10. Roby says:

    This is typical case of a Fitbit data export to Csv file. Each number has CHAR160 as thousand separator.. how smart Fitbit, thank you 😉

    By the way, i prefer to copy the character, and use find and replace.

  11. Suhas Shetty says:

    Sometimes it happens if you copy a table from outlook and paste it in excel. When you apply formula on those cells you will get error. What i use to do is
    copy one character that looks like space,
    select the entire range,
    go to Find and replace,
    Paste the copied character in Find option
    Leave the replace option unfilled..
    click on replace all..

    All the errors shall be converted in to proper values..

    Process looks lengthier.. but it is one of the simplest method

  12. Gerry says:

    If Clean, Trim, and Substitute, or Find and Replace does not complete the job, I usually enter a value of 1 in an empty cell. Copy the Value of 1, Highlight the range of text numbers, and Paste Special, Values, Multiply. This site is great!

  13. king faisal says:

    You can use Dose for Excel Add-In that can quickly clean huge data with one click besides more than +100 new functions and features to add to your Excel to save time and effort.

    https://www.zbrainsoft.com

  14. R.Ranjit says:

    Hi,
    I have a problem in excel. The sheet attached herewith.

    TABLE CONFIG 2/6
    A B C D E F G H
    1 WEIGHT1 43,599 WEIGH2 62500 WEIGHT3 77000 WEIGHT4 66,500
    2 DEDUCTION1 15,000 DEDUCTION1 15,000 TEMP 0 DEDUCTION2 11,005
    3 RESULT 58,599 RESULT-1 77,500 RESULT-2 77,000 RESULT-3 77,505
    4 RESULT SUBSTRACT 0 0 0
    5 REQUIRED VALUE 77,500 77,000 77,505

    Note: 1- RESULT (58599) IS TO BE DEDUCTION EITHER FROM D4 OR F4 OR H4 WHICHEVER IS MOST
    LEAST CELL AMONG RESULT-1 OR RESULT-2 OR RESULT 3.
    2-HENCE, RESULT VALUE $B$3 IS TO BE PRESENTED ON CELL EITHER D4 OR F4 OR H4 WHICHER IS
    MOST LEAST VALUE
    3-FORMULA =IF(E8<H8,$B$9,IF(E8<J8,$B$9,IF(H8<J8,$B$9,IF(H8<E8,$B$9,IF(J8<H8,$B$9))))))
    CREATED ON CELL D4,F4 & H4 DID NOT WORK.
    PLS FOR YOUR HELP.
    THANK YOU

Leave a Reply