Every dataset you will ever touch is built from numbers, but not all numbers behave the same way. Some count discrete things like customers or clicks. Others measure continuous quantities like temperature or revenue growth. A few even help engineers and data scientists model signals that oscillate over time. To use numbers correctly in analysis, you first need to know which “family” a number belongs to. This post walks through the complete hierarchy of number sets, from the simplest counting numbers to the more abstract complex numbers, and shows how each set builds on the one before it.

Table of Contents

Natural, whole and integer numbers: the building blocks

Every number system starts with the numbers we use to count. These form the base layer, and everything else in mathematics is constructed on top of them.

Natural numbers (N)

Natural numbers are the counting numbers: 1, 2, 3, 4, and so on, continuing forever. They do not include zero, fractions, decimals, or negative values. When you count the number of rows in a spreadsheet or the number of respondents in a survey, you are using natural numbers.

Whole numbers (W)

Whole numbers include every natural number plus zero: 0, 1, 2, 3, and so forth. The addition of zero might look trivial, but it matters. Zero represents the idea of “nothing,” which natural numbers alone cannot express. Interestingly, several modern textbooks and mathematical bodies now treat zero as part of the natural numbers themselves, since it behaves consistently under addition and is essential in fields like computer science, where counting often starts at zero. Because of this, you will sometimes see natural numbers defined as starting from 0 rather than 1, depending on the convention a particular course or country follows.

Integers (Z)

Integers extend whole numbers by adding negative values: …, -3, -2, -1, 0, 1, 2, 3, …. The symbol Z comes from the German word “Zahlen,” meaning numbers. Integers let you represent a loss as easily as a gain, which is why every profit-and-loss statement, temperature scale, or bank balance depends on them.

These three sets are nested inside one another. Every natural number is a whole number, and every whole number is an integer, but the reverse is not true. Zero is a whole number but not a natural number under the classical definition, and -5 is an integer but neither a whole number nor a natural number.

Rational numbers and irrational numbers: understanding real numbers

Integers alone cannot represent everything you encounter in data analysis. Averages, proportions, and ratios often fall between whole numbers, which is where the next two sets come in.

Rational numbers (Q)

A rational number is any number that can be written as a fraction p/q, where p and q are integers and q is not zero. This definition, formalised in classical arithmetic, means all integers are automatically rational too, since any integer n can be written as n/1. Rational numbers include every fraction, along with all decimals that either terminate or repeat in a pattern. For example, 0.75 is rational because it equals 3/4, and 0.333… is rational because it equals 1/3. This is why the symbol used for this set is Q, referring to the word “quotient.”

Irrational numbers

Some numbers refuse to be written as a simple fraction, no matter how hard you try. These are irrational numbers. An irrational number is any real number that cannot be expressed as the ratio of two integers, and its decimal expansion goes on forever without settling into a repeating pattern. The square root of 2 is the classic example. It first appeared as a problem in geometry, when mathematicians tried to calculate the diagonal of a square with sides of length one and found that no fraction could describe it exactly. Other well-known irrational numbers include pi (used constantly in circular and periodic calculations) and Euler’s number e (central to exponential growth models used in statistics and finance).

Real numbers (R)

Put rational and irrational numbers together, and you get the real numbers, denoted R. Real numbers include every positive and negative integer, every fraction built from those integers, and every irrational number, making up the complete, unbroken number line you likely first encountered in school. What makes the real number line special is a property called completeness: there are no “gaps” on it. Every point on the line corresponds to a real number, and every real number corresponds to a point. Rational numbers alone do not have this property, since there are infinitely many “holes” on the rational line where irrational values like the square root of 2 should sit.

In practical terms, almost every measurement you record in a dataset (heights, weights, prices, percentages, sensor readings) is a real number. Whether that real number happens to be rational or irrational rarely matters for everyday analysis, but understanding the difference matters when you study number theory, computational limits, or the behaviour of algorithms that round decimals.

Complex numbers: beyond the real number system

Real numbers cover almost everything you will measure, but they cannot solve every equation. Consider xยฒ + 1 = 0. Solving for x means finding a number whose square is -1. No real number satisfies this, because squaring any real number, positive or negative, always produces a non-negative result.

The imaginary unit

To solve this class of problem, mathematicians introduced a new number called i, defined so that iยฒ = -1. Any real multiple of this imaginary unit is called an imaginary number. This was not an easy idea for mathematicians to accept at first. The concept emerged gradually during the 16th century, when Italian mathematicians working on cubic equations found that intermediate steps in their formulas required the square roots of negative numbers, even when the final answers turned out to be perfectly ordinary. It took a few more centuries before the idea was placed on solid theoretical footing and connected to geometry, which finally gave imaginary numbers a legitimate place in mathematics rather than being dismissed as a curious trick.

Why complex numbers matter

A complex number combines a real part and an imaginary part, written as x + iy, where x and y are real numbers. The set of all such numbers is denoted C. Complex numbers are not just an abstract exercise. They are essential in electrical engineering for analysing alternating current, in signal processing for handling waveforms, and in quantum mechanics for describing particle behaviour. The complex number system also contains ordinary real numbers as a special case, where the imaginary part is simply zero. This means 7 can be written as 7 + 0i, making it a complex number as well as a real one.

Solving what real numbers cannot

The real reason complex numbers exist is to guarantee that every polynomial equation has a solution. This idea, known as the fundamental theorem of algebra, states that any polynomial equation of degree n has exactly n solutions once you allow complex numbers. Real numbers alone cannot make this promise, since equations like xยฒ + 1 = 0 have no real roots at all.

Containment relations in number sets

Now that you have met every set, it helps to see them as one continuous chain, each one nested inside the next:

โˆ… โŠ‚ N โŠ‚ W โŠ‚ Z โŠ‚ Q โŠ‚ R โŠ‚ C

Here is what each step in this chain represents:

โˆ… โŠ‚ N: The empty set is contained in every set, including the natural numbers, simply by the definition of subset relationships in set theory.

N โŠ‚ W: Every counting number is a whole number, since whole numbers are just natural numbers with zero added.

W โŠ‚ Z: Every whole number is an integer, since integers are whole numbers extended to include negatives.

Z โŠ‚ Q: Every integer is rational, because any integer n can be expressed as the fraction n/1.

Q โŠ‚ R: Every rational number is real, since real numbers include both rationals and irrationals together.

R โŠ‚ C: Every real number is complex, because it can be written as x + 0i.

Each containment is strict, meaning the larger set always has elements the smaller one lacks. Zero sits in W but not N. Negative five sits in Z but not W. One-half sits in Q but not Z. The square root of two sits in R but not Q. And i itself sits in C but not R. This chain is not arbitrary. Each extension exists to solve a problem the previous set could not handle: whole numbers gave us zero, integers gave us negative values, rational numbers gave us division and fractions, real numbers gave us continuous measurement, and complex numbers gave us complete solutions to polynomial equations.

For anyone working with data, this hierarchy is more than a classification exercise. It determines what kind of arithmetic and statistical operations make sense on a variable, how a programming language stores and rounds a value, and why certain equations in modelling and simulation require complex numbers to solve completely. A solid grip on where a number “lives” in this chain makes every later concept in quantitative reasoning easier to follow.

What do you think?

What do you think? If zero was only added to the natural numbers later, why do you think some modern definitions now include it from the start? And can you think of a real-world dataset where the difference between a rational and an irrational value would actually change your analysis?

How useful was this post?

Click on a star to rate it!

Average rating 0 / 5. Vote count: 0

No votes so far! Be the first to rate this post.

We are sorry that this post was not useful for you!

Let us improve this post!

Tell us how we can improve this post?

References
  1. https://www.britannica.com/science/rational-number
  2. https://www.britannica.com/science/irrational-number
  3. https://www.britannica.com/science/real-number
  4. https://www.britannica.com/science/imaginary-number
  5. https://www.sciencefocus.com/science/a-brief-introduction-to-imaginary-numbers
  6. https://www.britannica.com/science/number-mathematics

Comments

Leave a Reply

Your email address will not be published. Required fields are marked *

Data Analysis

1 Mathematical Concept

  1. Set Theory
  2. Number Sets (with Standard Notations)
  3. Set Operations
  4. Relation and Functions
  5. Logic
  6. Proof Techniques

2 Statistical Concepts

  1. Some Elementary Concepts
  2. Descriptive Statistics
  3. Quantitative Data – Percentages and Measures of Central Tendency
  4. Quantitative Data – Measures of Dispersion
  5. Quantitative Data – Measures of Position

3 Introduction to Statistical Software

  1. Need of Statistical Software
  2. Data Handling
  3. Use of Formula and Functions
  4. Making Charts
  5. Activating Data Analysis Tab

4 Data Collection- Methods and Sources

  1. Methods of Data Collection
  2. Planning and Organisation of Census and Surveys
  3. Errors in Data or Data Collection
  4. Cost of the Enquiry
  5. Census or Survey?
  6. Sources of Secondary Data

5 Tools of Data Collection

  1. Quantitative and Qualitative Research
  2. Questionnaire
  3. Schedule
  4. Interview
  5. Participant Observation
  6. Non-participant Observation
  7. Focused Interview
  8. Oral Histories
  9. Case Study Method
  10. Group Discussion
  11. Focus Group Discussion
  12. Narratives

6 Data Presentation

  1. Classification of Data
  2. Simple Array
  3. Discrete Frequency Distribution
  4. Grouped Frequency Distribution
  5. Types of Grouped Frequency Distribution
  6. How to Use Spreadsheet Software for Frequency Distribution?
  7. Tabulation of Data
  8. Diagrammatic Presentation of Data
  9. Graphical Representation of Data

7 Univariate Data Analysis

  1. Exploratory Data Analysis
  2. Inferential Statistics: Basic Concepts and Significance of Measures of Central Tendency and Dispersions in Decision Making
  3. Inferential Statistics: Point Estimation and Setting up Confidence Intervals for Population Parameters

8 Bivariate Data Analysis

  1. Scatter Plots and Correlation
  2. Concept of Correlation
  3. Correlation Coefficient
  4. Test of Significance for the Correlation Coefficient
  5. Correlation and Causation
  6. Line of Best Fit
  7. Regression Lines Equation
  8. Regression Coefficients
  9. Predictability of Regression Equations
  10. Coefficient of Determination
  11. Standard Error of Estimate: Concept and Estimation
  12. Prediction Interval
  13. Testing the Difference between Two Means: Using the z-test and t-test
  14. Testing the Difference between Proportions Using z-test
  15. Testing the Difference between Two Variances: F-Test
  16. Analysis of Variances

9 Multivariate Data Analysis

  1. What is Multivariate Analysis?
  2. Classification of Multivariate Techniques
  3. Principal Components and Common Factor Analysis
  4. Multiple Regression
  5. Multiple Discriminant Analysis (MDA) and Logistic Regression
  6. Canonical Correlation Analysis
  7. Multivariate Analysis of Variance (MANOVA)
  8. Conjoint Analysis
  9. Cluster Analysis
  10. Perceptual Mapping
  11. Correspondence Analysis
  12. Structural Equation Modeling (SEM)
  13. Guidelines for Multivariate Techniques and Interpretation
  14. A Structured Approach to Multivariate Model Building

10 Construction of Composite Index in Social Sciences

  1. Composite Index: the Concept
  2. Steps in Constructing Composite Index
  3. Dealing with Missing Values and Outliers
  4. Simple Ranking Method
  5. Indices Method
  6. Mean Standardisation Method
  7. Range Equalisation Method
  8. Physical Quality of Life Index (PQLI)
  9. Human Development Index (HDI)
  10. Gender Development Index (GDI)
  11. Merits and Limitations of Composite Index

11 Analysis of Qualitative Data

  1. Qualitative Research
  2. Qualitative vs. Quantitative Research
  3. Qualitative Data: Research Methods
  4. Qualitative Data and Techniques
  5. Qualitative Data Collection Methods
  6. Qualitative Data Analysis: Approaches and Techniques
  7. Qualitative Data Analysis: Procedure and Computer Softwares