When you’re staring at a fresh dataset in SPSS, the urge to jump straight into regressions, t-tests, and fancy inferential statistics is real. But before you chase those complex outputs, there’s a humbler step you absolutely cannot skip: looking at each variable on its own. This is what researchers call univariate analysis, and it’s the foundation on which every meaningful statistical conclusion rests. Let’s walk through how to perform it in SPSS, what each statistic actually tells you, and how to interpret the output like a trained researcher.

Table of Contents

What univariate analysis really means

Univariate analysis is exactly what it sounds like: the analysis of one variable at a time. There is no cause-and-effect relationship being tested here, because only a single variable is involved. The goal is simple but powerful – describe the data. You want to know what the typical value looks like, how spread out the observations are, and whether the distribution has any unusual shape.

Before running any inferential test, researchers use univariate analysis to check each variable for accuracy, identify typos or errors, count how many cases exist, and examine the values the variable contains. Skipping this step is one of the fastest ways to end up with a result that looks impressive but is quietly built on flawed data.

Why it matters for your research

Think of univariate analysis as the diagnostic X-ray of your dataset. Examining frequency distributions can reveal unexpected categories, miscoded values, or implausible outliers that need to be addressed before the analysis moves forward. If someone’s age is recorded as 998 because that was the code for a missing value, your mean age could be wildly off. Catching such issues early saves you from reporting nonsense later.

Knowing your variable’s level of measurement

Before clicking a single menu in SPSS, identify what type of variable you’re dealing with. This decision determines which statistics are meaningful and which are garbage. The four standard levels are nominal (categories with no order, like gender or religion), ordinal (ordered categories, like education level), interval, and ratio (both of which are continuous/scale variables, like age or income).

Here’s the catch: SPSS will calculate statistics even if the measure of central tendency and dispersion are not appropriate for the variable type. If you ask SPSS to calculate the mean of a nominal variable coded as 1 for male and 2 for female, it will dutifully return 1.4 – a number that means absolutely nothing. The software doesn’t know better. You have to.

The command path in SPSS

The primary route for univariate analysis in SPSS is through the Frequencies command. Navigate to Analyze โ†’ Descriptive Statistics โ†’ Frequencies. Once the dialogue box opens, you’ll see your list of variables on the left and an empty “Variable(s)” box on the right. Move the variables you want to analyse into this box using the blue arrow.

For categorical variables (nominal or ordinal), keep the Display frequency tables checkbox ticked. For continuous variables, you may want to uncheck it to avoid producing a giant table with hundreds of rows – one for every unique value. Next, click the Statistics button to choose which measures you want SPSS to compute.

Choosing the right statistics for the job

A concise workflow, adapted from standard guidance, looks like this: select Mode for nominal variables; Mode and Median for ordinal variables; and Mean, Median, Mode, Standard Deviation, Range, Skewness, and Kurtosis for continuous variables. You can also request percentiles and quartiles if needed. Then click Continue, and back in the main dialogue, click Charts to add a bar chart or pie chart for categorical data, or a histogram for continuous data. Click OK to run it.

Understanding the frequency table

Once SPSS produces the output, the first thing you’ll see for categorical variables is the frequency table. It contains four important columns: Frequency (the raw count of respondents in each category), Percent (the percentage of the total sample), Valid Percent (percentage after excluding missing cases), and Cumulative Percent (a running total).

The Valid Percent column is usually the one you report, because it accounts for respondents who skipped the question. A large gap between Percent and Valid Percent tells you that missing data is substantial – itself a finding worth flagging in your write-up.

Measures of central tendency

Central tendency statistics give you a single value that represents the “middle” of your data. The three classic measures are the mean, median, and mode – each appropriate in different situations.

Mean

The mean is the arithmetic average – the sum of all values divided by the number of observations. It includes every value in the dataset as part of its calculation, and it is the only measure where the sum of deviations of each value from the mean is always zero. That’s mathematically elegant, but it comes with a cost: the mean is highly sensitive to outliers. A single billionaire in a neighbourhood income dataset will drag the mean upward and paint a misleading picture of the “typical” resident.

Median

The median is the middle value when data is arranged in order. It’s the go-to measure when your distribution is skewed or contains outliers. The median is appropriate to use with ordinal variables, and with interval variables that have a skewed distribution. Household income, wealth, housing prices – these are all classic examples where the median tells a truer story than the mean.

Mode

The mode is simply the most frequently occurring value. It’s the only measure of central tendency you can meaningfully use with nominal data. If you ask respondents about their preferred mode of transport and “metro” shows up most often, that’s your mode. Note that a dataset can be bimodal (two modes) or even multimodal, which itself reveals something about the underlying population.

Measures of dispersion

Central tendency alone is never enough. Two datasets can have the same mean but look completely different – one tightly clustered, the other wildly spread. That’s why you also need measures of dispersion to describe how far individual values fall from the centre.

Range

The range is the simplest measure: the difference between the maximum and minimum values. It’s easy to calculate but highly sensitive to extreme values. Still, it gives you a quick sense of how wide the spread is.

Variance and standard deviation

Variance quantifies the average squared deviation from the mean. Because it’s in squared units, interpreting it directly is awkward – if your variable is measured in rupees, variance is in “rupees squared.” That’s why we usually report the standard deviation, which is simply the square root of variance and is expressed in the same units as the original variable. A small standard deviation means observations cluster tightly around the mean; a large one means they’re widely scattered.

When the distribution is roughly normal, the mean and standard deviation are the preferred summary. If the distribution is skewed or has notable outliers, report the median and interquartile range (IQR) instead, and briefly mention the skew or outliers.

Shape of the distribution: skewness and kurtosis

Beyond central tendency and dispersion, SPSS lets you examine the shape of a distribution through two statistics: skewness and kurtosis. These are often overlooked by beginners but are vital when you plan to run parametric tests that assume normality.

Skewness

Skewness measures the symmetry of a distribution. A value of zero indicates perfect symmetry. A positive value means the distribution has a long tail on the right (right-skewed – think of income data), while a negative value means a long tail on the left (left-skewed – think of exam scores when most students do well). As a rule of thumb, skewness values between -2 and +2 are generally considered acceptable for assuming a normal univariate distribution.

Kurtosis

Kurtosis describes the “tailedness” of a distribution – whether the data produces more or fewer extreme outliers compared to a normal distribution. If a distribution has kurtosis greater than zero, it tends to produce more outliers than the normal distribution. A negative kurtosis indicates lighter tails and a flatter shape. SPSS reports excess kurtosis, meaning zero represents a perfectly normal distribution.

When both skewness and kurtosis fall within acceptable ranges, you can generally proceed with parametric tests. When they don’t, consider data transformations (like taking a log) or switching to non-parametric alternatives.

Interpreting the output – an example

Suppose you’re analysing the age variable in a public policy survey of 500 respondents. SPSS returns: mean = 38.4, median = 37, mode = 35, standard deviation = 11.2, range = 62, skewness = 0.42, kurtosis = -0.18. What does this tell you?

The mean and median are close, suggesting a reasonably symmetric distribution. Skewness at 0.42 confirms a mild rightward lean – a small group of older respondents pulls the mean slightly above the median. Kurtosis near zero suggests tails close to normal. The standard deviation of 11.2 indicates respondents typically fall within about 11 years above or below the average age. Based on this, age is likely suitable for parametric analyses without transformation.

Common pitfalls to avoid

Beginners often fall into a few traps. First, calculating a mean for a categorical variable just because SPSS allowed it. Second, reporting raw frequencies without percentages, which makes comparison across studies difficult. Third, forgetting to define missing value codes – if “99” represents “refused to answer” and you don’t tell SPSS, it treats 99 as a real age. Finally, treating univariate output as the end of the analysis rather than the beginning. These statistics set up everything that comes next, including your choice of inferential tests.

What do you think? When analysing a dataset from a citizen satisfaction survey with hundreds of variables, how would you prioritise which variables to examine first through univariate analysis? And when skewness and kurtosis suggest non-normality, would you prefer transforming the data or switching to non-parametric tests – and why?

How useful was this post?

Click on a star to rate it!

Average rating 0 / 5. Vote count: 0

No votes so far! Be the first to rate this post.

We are sorry that this post was not useful for you!

Let us improve this post!

Tell us how we can improve this post?

References
  1. https://manifold.open.umn.edu/read/chapter-5-data-management-and-cleaning-in-spss
  2. https://sociology.institute/research-methodologies-methods/univariate-analysis-spss-guide/
  3. https://subjectguides.sunyempire.edu/c.php?g=659059&p=4626896
  4. https://statistics.laerd.com/statistical-guides/measures-central-tendency-mean-mode-median.php
  5. https://www.betterevaluation.org/methods-approaches/methods/measures-central-tendency
  6. https://spssservices.com/univariate-analysis-spss-guide/
  7. https://imaging.mrc-cbu.cam.ac.uk/statswiki/FAQ/Simon
  8. https://www.statology.org/skewness-kurtosis-in-spss/

Comments

Leave a Reply

Your email address will not be published. Required fields are marked *

Research Methodologies

1 Logic of Inquiry in Social Research

  1. A Science of Society
  2. Comteโ€™s Ideas on the Nature of Sociology
  3. Observation in Social Sciences
  4. Logical Understanding of Social Reality

2 Empirical Approach

  1. Empirical Approach
  2. Rules of Data Collection
  3. Cultural Relativism
  4. Problems Encountered in Data Collection
  5. Difference between Common Sense and Science
  6. What is Ethical?
  7. What is Normal?
  8. Understanding the Data Collected
  9. Managing Diversities in Social Research
  10. Problematising the Object of Study

3 Diverse Logic of Theory Building

  1. Concern with Theory in Sociology
  2. Concepts: Basic Elements of Theories
  3. Why Do We Need Theory?
  4. Hypothesis, Description and Experimentation
  5. Controlled Experiment
  6. Designing an Experiment
  7. How to Test a Hypothesis
  8. Common Methods of Testing a Hypothesis
  9. Sensitivity to Alternative Explanations
  10. Rival Hypothesis Construction

4 Theoretical Analysis

  1. Premises of Evolutionary and Functional Theories
  2. Critique of Evolutionary and Functional Theories
  3. Turning away from Functionalism
  4. What after Functionalism
  5. Post-modernism
  6. Trends other than Post-modernism

5 Issues of Epistemology

  1. Some Major Concerns of Epistemology
  2. Rationalism
  3. Empiricism
  4. Idealism
  5. Phenomenology: Bracketing Experience

6 Philosophy of Social Science

  1. Foundations of Science
  2. Science, Modernity and Sociology
  3. Rethinking Science
  4. Crisis in Foundation

7 Positivism and its Critique

  1. Heroic Science and Origin of Positivism
  2. Early Positivism
  3. Consolidation of Positivism
  4. Critiques of Positivism

8 Hermeneutics

  1. Methodological Disputes in the Social Sciences
  2. Tracing the History of Hermeneutics
  3. Hermeneutics and Sociology
  4. Philosophical Hermeneutics
  5. The Hermeneutics of Suspicion
  6. Phenomenology and Hermeneutics

9 Comparative Method

  1. Relationship with Common Sense; Interrogating Ideological Location
  2. The Historical Context
  3. Elements of the Comparative Approach

10 Feminist Approach

  1. Relationship with Common Sense; Interrogating Ideological Location
  2. The Historical Context
  3. Features of the Feminist Method
  4. Feminist Methods adopt the Reflexive Stance
  5. Feminist Discourse in India

11 Participatory Method

  1. Relationship with Common Sense; Interrogating Ideological Location
  2. The Historical Context
  3. Delineation of Key Features

12 Types of Research

  1. What is Research?
  2. Types of Research

13 Methods of Research

  1. Centrality of Research Methods in Social Sciences
  2. Interface between Methodology and Methods
  3. Elements of Research Methodology
  4. Types of Data Used in Social Research
  5. Research Methods

14 Elements of Research Design

  1. Structuring the Research Process
  2. Defining Your Research Problem
  3. Choice of Field Site(s)
  4. Consideration of Time and Resources
  5. Reviewing Secondary Material
  6. Hypothesis
  7. Theoretical Orientation
  8. Universe and Unit of Study
  9. Pilot Study
  10. Sampling
  11. Data Collection
  12. Analysis and Report Writing

15 Sampling Methods and Estimation of Sample Size

  1. Sampling
  2. Classification of Sampling Methods
  3. Sample Size
  4. Probability Sampling
  5. Non-Probability Sampling

16 Measures of Central Tendency

  1. Mean
  2. Median
  3. Mode
  4. Relationship between Mean, Mode and Median
  5. Choosing a Measure of Central Tendency

17 Measures of Dispersion and Variability

  1. The Range
  2. The Variance
  3. The Standard Deviation
  4. Coefficient of Variation
  5. Measures of Dispersion and Variability

18 Statistical Inference- Tests of Hypothesis

  1. Statistical Inference
  2. Steps in Hypothesis Testing
  3. Types of Errors in Hypothesis Testing
  4. Tests of Significance: Chi-Square Test
  5. Tests of Significance: Student’s t Test

19 Correlation and Regression

  1. Correlation
  2. Method of Calculating Correlation of Ungrouped Data
  3. Method of Calculating Correlation of Grouped Data
  4. Regression

20 Survey Method

  1. Rationale of Survey Research Method
  2. History of Survey Research
  3. Defining Survey Research
  4. Sampling and Survey Techniques
  5. Operationalising Survey Research Tools
  6. Advantages and Weaknesses of Survey Methods

21 Survey Design

  1. Preliminary Considerations
  2. Stages / Phases in Survey Research
  3. Formulation of Research Question
  4. Survey Research Designs
  5. Sampling Design

22 Survey Instrumentation

  1. Techniques/Instruments for Data Collection
  2. Questionnaire Construction
  3. Issues in Designing a Survey Instrument

23 Survey Execution and Data Analysis

  1. Problems and Issues in Executing Survey Research
  2. Data Analysis
  3. Ethical Issues in Survey Research

24 Field Research – I

  1. History of Field Research
  2. Ethnography
  3. Theme Selection
  4. Designing Research
  5. Gaining Entry in the Field
  6. Key Informants
  7. Participant Observation

25 Field Research – II

  1. Genealogy
  2. Interview, its Types and Process
  3. Feminist and Postmodernist Perspectives on Interviewing
  4. Narrative Analysis
  5. Interpretation

26 Reliability, Validity and Triangulation

  1. Concepts of Reliability and Validity
  2. Three types of “Reliability”
  3. Working towards Reliability
  4. Procedural Validity
  5. Field Research as a Validity Check

27 Qualitative Data Formatting and Processing

  1. Qualitative Data Processing and Analysis
  2. Description
  3. Classification
  4. Making Connections
  5. Theoretical Coding

28 Writing up Qualitative Data

  1. Problems of Writing Up
  2. Grasp and Then Render
  3. Writing Down and “Writing Up”
  4. Write Early
  5. Writing Styles

29 Using Internet and Word Processor

  1. What is Internet and How Does it Work?
  2. Internet Services
  3. Searching on the Web: Search Engines
  4. Accessing and Using Online Information
  5. Uses of E-mail Services in Research

30 Using SPSS for Data Analysis Contents

  1. Starting and exiting SPSS
  2. Creating a data file
  3. Univariate analysis
  4. Bivariate analysis
  5. Multivariate analysis

31 Using SPSS in Report Writing

  1. Why to Use SPSS
  2. Charts
  3. Working with SPSS Output
  4. Copying SPSS output to MS Word Document
  5. Conclusion

32 Tabulation and Graphic Presentation- Case Studies

  1. Structure for Presentation of Research Findings
  2. Data Presentation: Editing, Coding and Transcribing
  3. Case Studies
  4. Qualitative Data Analysis and Presentation through Computer Software
  5. Types of ICT used for Research

33 Guidelines to Research Project Assignment

  1. Overview of Research Methodologies and Methods (MSO 002)
  2. Research Project Objectives
  3. Preparation for Research Project
  4. Stages of the Research Project
  5. Supervision During the Research Project