play icon for videos

Qualitative Data: What It Is, Types & Real Examples

Qualitative data is the non-numerical evidence from words, images, and observation. Definition, four types, six characteristics, and eight worked examples.

Updated
July 30, 2026
360 feedback training evaluation
Use Case

What is qualitative data?

Qualitative data is non-numerical evidence that describes qualities, experiences, and meaning — interview transcripts, open-ended survey answers, observation notes, documents, images, and recordings. It answers why and how something happened, where quantitative data answers how many and how much. It is the half of the evidence that explains the other half.

Qualitative data is rarely scarce. Most organizations hold far more of it than they use: recordings nobody transcribed, open-ended answers nobody coded, notes filed after a session and never opened. The problem is not collection — it is that the evidence arrives in fragments, in different tools, unattached to the person who gave it, so it never becomes something you can count on.

Key takeaways

  • Qualitative data is non-numerical evidence — transcripts, open-ended answers, notes, documents, images — that explains why a number moved.
  • Four types cover most of it: textual, observational, audio-visual, and documentary. Each needs a different capture method but the same discipline afterward.
  • Sopact calls the real failure the Fragment Problem: qualitative evidence scatters across tools and is never attached to the participant who gave it, so it can be quoted but not counted.
  • Unread qualitative data is not evidence. Fifty transcripts nobody coded contribute nothing to a finding, however rich they are.
  • Attach it to a person, then read it on arrival. One participant record plus a fixed codebook turns qualitative data from anecdote into analyzable evidence.

The Fragment Problem.

The reason qualitative data underdelivers is structural, not analytical. The interview lives in a transcription tool, the open-ended answers in a survey platform, the case notes in a case system, and the demographics in a CRM. Each fragment is meaningful on its own and none of them is connected to the others, so a theme can be quoted in a report but never counted across a cohort or broken out by who said it.

Sopact calls that the Fragment Problem: qualitative evidence scattered across tools and detached from the participant who produced it. The fix is a persistent participant ID from first contact and a codebook fixed before collection, so every fragment lands on one record and is themed as it arrives. The seven methods that produce this evidence are on the qualitative data collection methods page, and the questions that elicit it on qualitative questions.

Once the fragments are joined, qualitative data stops being the part of the report that gets skimmed. A theme carries a count, a segment, and a verbatim line, which is what lets it sit beside a number rather than beneath it.

The four types of qualitative data.

Qualitative data falls into four types: textual (interviews, open-ended answers, case notes), observational (field notes, behavior records), audio-visual (recordings, photographs, video), and documentary (reports, policies, existing records). The type decides how you capture it; it does not change what has to happen next.

Whatever the type, the same two things determine whether it becomes evidence: is it attached to a person, and is it read against a consistent codebook. The table pairs each type with an example and what it takes to make it analyzable. Running it at scale in a survey is covered on qualitative survey.

Where qualitative data becomes evidence.

Qualitative data becomes evidence at the moment it is attached to a participant and themed against a fixed codebook — not when it is collected, and not when someone finally reads it. The stage below shows the same workflow run the usual way and run as a loop.

The comparison that matters is not manual versus automated; it is whether the reading keeps pace with the collection. Choosing the software that does this is on qualitative data analysis software, and the difference from numeric evidence on qualitative vs quantitative.

Stage 1
From interview to evidence
where qualitative data usually breaks
TodayInterviews recorded · Transcripts dropped in a shared folder · Coded by hand months later
⚠ By analysis time half the transcripts are unread and the codebook has drifted from the first ten.
The Loop on this stage with Sopact
1
Collect — clean at the source
InterviewsOpen-ended surveyDocumentsField notes
→ every source lands on one persistent ID
2
On arrival — read automatically
Intelligent Cell
Each open response is themed against your locked codebook the moment it lands, with sentiment and the verbatim line kept.
Intelligent Row
Every participant resolves to one row — their rating, their reason, their demographics — across every wave.
3
Ask & act — the Assistant
“Which themes explain the confidence drop at the Oakland site, and who said them?”
→ Answer with cited verbatims in minutes instead of scheduling a coding sprint.

Four types, one requirement.

Qualitative data comes in four types — textual, observational, audio-visual, and documentary — and each needs the same two things to become evidence: a person attached to it, and a codebook applied to it. Read the last column.

Types of qualitative data
TypeExamplesWhat makes it analyzable
TextualInterviews, open-ended answers, case notesThemed against a fixed codebook on arrival
ObservationalField notes, behavior recordsA protocol, so two observers record comparably
Audio-visualRecordings, photos, videoTranscribed and bound to the participant
DocumentaryReports, policies, existing recordsCoded against a rubric, with the source cited

Read the last column and the types converge on one requirement: attached to a person, read against a consistent rule. That is what turns four kinds of raw material into evidence, and its absence is the Fragment Problem.

Qualitative data read at year-end is a folder. The Loop makes it evidence.

Fifty transcripts collected in March and read in November explain a cohort that has already left. Reading each response as it lands means the theme is available while the program can still act on it, and the codebook holds instead of drifting. That is the premise of the Loop, Sopact's method for continuous impact intelligence: collect clean at the source, analyze the moment data arrives, improve while you can still act.

The Loop is also what makes a qualitative claim defensible. Every theme count traces back to the response it came from, so a percentage resolves to the people who said it. That standard has its own chapter in traceability and transparency.

One method, three moves that never stop

1 · CollectClean at the source; every fragment on one participant record.
2 · AnalyzeOn arrival; themed against a codebook fixed before collection.
3 · ImproveIn time to act; the theme lands while the cohort is still here.

Then the cycle runs again, a little sharper each wave. Read the method: the Loop methodology →

Turn your qualitative data into evidence this week

The fastest way to solve the Fragment Problem is to code one batch against a fixed codebook. Each prompt below pastes into Sopact Sense's Assistant, or reasons through with your team; the arrow above each links the Academy walkthrough that shows the expected output and the tips.

Academy walkthrough → Fix the codebook before you code

Draft a codebook for this qualitative data from our framework and a sample: [PASTE FRAMEWORK + 10-15 RESPONSES]. For each code give a short name, a one-line definition, an include-when rule, an exclude-when rule, and an example quote. Keep it to 6-10 codes and flag overlaps where two codes would catch the same sentence.

Academy walkthrough → Theme it on arrival

Theme this batch against the codebook, one row per respondent: [PASTE CODEBOOK + RESPONSES with respondent_id]. Return respondent_id, assigned theme(s), sentiment, and the percentage distribution across the batch. Keep the codebook fixed; only add NEW_THEME if more than 5% fit nothing.

Academy walkthrough → Attach every fragment to a person

Review how we collect qualitative data today: [PASTE SOURCES AND TOOLS]. For each source, say whether a response can be traced to a specific participant and to their quantitative record. Flag every fragment that arrives detached, and give the fix that would attach it at collection rather than afterward.

Academy walkthrough → Give the theme a segment

Using this themed dataset with demographics on each record: [PASTE], show the theme distribution by [SITE / GENDER / COHORT], report where a theme appears in one subgroup but not another, and cite the strongest verbatim line for each difference.

Learn the how-to in the Academy

Each walkthrough is short and practical: what to do, the prompt to run, the output to expect, and the tips that keep it reliable.

Watch: multi-model collection — interviews, PDFs, and surveys — read on one participant record.

Frequently asked questions

What is qualitative data?

Qualitative data is non-numerical evidence describing qualities, experiences, and meaning — interview transcripts, open-ended survey answers, observation notes, documents, images, and recordings. It answers why and how, where quantitative data answers how many. In Sopact's framing, it only becomes evidence once it escapes the Fragment Problem: attached to a participant and read against a fixed codebook.

What are the types of qualitative data?

Four types cover most of it: textual (interviews, open-ended answers, case notes), observational (field notes, behavior records), audio-visual (recordings, photos, video), and documentary (reports, policies, existing records). The type decides how you capture it; all four still need to be attached to a person and coded consistently to be analyzable.

What are examples of qualitative data?

Examples include an interview transcript about why a participant left a program, an open-ended survey answer explaining a low confidence rating, a facilitator's field notes on group dynamics, a photograph documenting a site condition, and a policy document coded against a rubric. Sopact keeps each example bound to the participant it came from so it can be counted, not just quoted.

What is the difference between qualitative and quantitative data?

Quantitative data is numerical — counts, ratings, measurements — and answers how many and how much; qualitative data is non-numerical and answers why and how. Most real evidence needs both, ideally captured together so a rating and its reason sit on one record. The fuller comparison is on the qualitative vs quantitative page.

How do you analyze qualitative data?

Define a codebook before collection, theme each response against it as the response arrives, keep sentiment and the verbatim line, and disaggregate the themes by segment. Hand-coding does not scale past a few dozen responses, and an unconstrained model re-guesses each run. Sopact themes on arrival against a locked codebook, which keeps the analysis reproducible.

How do you collect qualitative data?

Through seven established methods — interviews, focus groups, open-ended surveys, document analysis, observation, case studies, and ethnography — covered in depth on the qualitative data collection methods page. The decision that matters more than the method is whether each response binds to a participant at collection, which is what Sopact's persistent Contact ID does.

Why does most qualitative data go unused?

Because it arrives in fragments across different tools, detached from the person who gave it, so reading it is a separate project nobody has time for. Fifty transcripts nobody coded contribute nothing to a finding. Sopact calls this the Fragment Problem and solves it by binding every fragment to one participant record and theming it as it lands.

Is qualitative data reliable?

It is as reliable as the method applied to it. A fixed codebook applied the same way to the same data returns the same distribution twice, which is what makes a qualitative finding defensible; ad-hoc reading does not. Sopact fixes the codebook and keeps every theme traceable to its source sentence, so a qualitative claim can be checked rather than trusted.

Next: see the seven methods on the qualitative data collection methods page, or compare the tools on the qualitative data analysis software page.