Sequences

Find recurring combinations of adjacent words

Sequences counts two-, three-, or four-word combinations in the Active selection. It reveals repeated phrasing while preserving sentence, speaker-turn, and text boundaries, and it lets you open any sequence as an exact phrase in Concordance. This is especially useful for spotting potential collocations in your data.

15 min
The basic unit

A sequence is an uninterrupted window of adjacent words

Sequence

Two, three, or four consecutive word tokens in their original order. No intervening word can be skipped.

Occurrence

One appearance of that complete combination. Repeated appearances contribute to its Frequency.

Unique sequence

One distinct combination represented by one result-table row, regardless of how often it occurs.

Choose the scope with Active selection

Sequences uses the shared Active selection. Every count and normalized rate refers only to that scope.

Entire corpus

Counts sequences throughout every imported corpus text.

Selected text files

Counts only the selected interviews or documents.

Selected speakers

Counts sequences inside turns attributed to the selected speaker codes across their texts.

Text groups

Counts the combined members of one or more custom text groups.

Speaker groups

Counts turns attributed to members of one or more custom speaker groups.

Choose 2, 3, or 4 words and understand the boundaries

Source:well you know what
well youyou knowknow what

Illustration · Adjacent windows overlap. In a two-word analysis, the middle words can participate in more than one sequence.

2 words

Also called bigrams. This is the default and usually produces the broadest result set.

3 words

Also called trigrams. These capture more specific recurring phrasing.

4 words

Four-word sequences are more specific and often less frequent.

Join at apostrophes

Uses the shared project rule. With it off, we’re contributes we and re; with it on, it contributes one word. Changing the rule requires Update Sequences.

Where a sequence must stop

  • A sequence never crosses a sentence boundary marked by a period, question mark, or exclamation mark.
  • A sequence never joins the end of one speaker’s turn to the beginning of another speaker’s turn.
  • A sequence never crosses from one corpus text into another.
  • A visual or imported line break alone does not stop a sequence when the same sentence and speaker turn continue.
  • Speaker codes and parenthetical comments are excluded before sequences are counted.
  • Letter case is combined. The displayed sequence uses normalized lowercase forms.

Generate sequences step by step

1
Open Sequences

Opening the workspace does not begin counting automatically.

2
Confirm Active selection

Verify the texts, speakers, or groups that answer your research question.

3
Choose the length

Select 2 words, 3 words, or 4 words.

4
Confirm apostrophe handling

Decide whether forms containing apostrophes count as one word or separate tokens.

5
Choose Generate Sequences

Progress, percentage, and elapsed time remain visible. The button becomes Cancel while counting.

6
Read the completion status

Exempla reports total occurrences and unique sequences above the completed table.

Read the three result columns

Filter sequences…ContainsMinimum: 2Time: 0.61 seconds
Sequence length2 words3 words4 wordsJoin at apostrophes48,392 occurrences · 12,614 unique sequences
SequenceFrequencyPer 1,000
you know1,28426.53
in the1,04121.51
I think81616.86
we were60412.48

Illustration

1Sequence

The complete adjacent word combination in normalized form.

2Frequency

The exact number of occurrences in the generated scope.

3Per 1,000

Frequency ÷ all sequence occurrences of the selected length × 1,000.

Use sequence frequencies as a hypothesis list

A high-frequency sequence is a useful lead, not automatic proof of a collocation. Use it to identify candidates, then validate them in context.

Start with recurrence

Use frequency to prioritise phrases, then confirm whether recurrence is stable across texts, turns, and speakers.

Validate in Concordance

Open candidates as exact phrases and review context before interpreting a sequence as a meaningful pattern.

Compare neighboring forms

Look for near-regular variants (for example, one extra function word) to avoid over-reading one frozen form.

Filter and sort completed results

Contains

Keeps sequences containing the entered characters anywhere. Because this is a text containment filter, a short fragment can also match inside a word.

Begins with

Keeps sequences beginning with the entered complete word or words. Filtering you know can return you know that, but not you knowledge.

Minimum

Keeps sequences whose raw Frequency is at least the entered value. The default is 2, which hides combinations occurring only once.

Show everything counted

Clear Filter sequences and set Minimum to 1. The cached results reappear without generation.

  • Filters are case-insensitive and do not change the completion totals.
  • Changing Filter, mode, or Minimum updates the displayed table immediately.
  • Export follows the currently displayed rows after filtering.
  • Changing sequence length, Active selection, or apostrophe handling requires Update Sequences.

Sort from the headings

Sequence

First click A–Z; second click Z–A.

Frequency

First click highest to lowest; second click lowest to highest.

Per 1,000

First click highest to lowest. Its order matches Frequency within one completed analysis.

Complete table

Sorting applies to all displayed results, not only rows currently visible on screen.

Double-click a sequence to find its occurrences

Double-click any row. Exempla opens Concordance and runs an Exact phrase search for that sequence in the current Active selection.

Double-click a row

Choose the recurring combination you want to inspect.

Exact phrase search

Concordance finds the adjacent words in the same order.

Read the evidence

Review speaker, context, source text, and transcript line.

Export sequences as CSV

Three columns

The CSV contains Sequence, Frequency, and Per 1,000.

Current filters

Only displayed rows are exported. Clear the filter and use Minimum 1 for the complete generated table.

Meaningful filename

Exempla proposes a date-first filename containing the project name and Sequences.

Large exports

At least 5,000 rows use a responsive progress window with Cancel. Cancelling leaves no incomplete CSV.

Check the rules in order

When sequence results are unexpected

No sequences were found

Confirm that the scope contains enough adjacent spoken words for the selected length.

The button says Update Sequences

Scope, length, or apostrophe handling changed. Choose Update Sequences to recount.

A known phrase is missing

Check whether it crosses a sentence, speaker turn, or text boundary, or contains a parenthetical comment.

Only repeated sequences appear

Minimum defaults to 2. Set it to 1 to include sequences occurring once.

The filter returns too many rows

Contains can match inside words. Choose Begins with for a complete initial word or phrase.

The table has no visible rows

Clear Filter sequences and set Minimum to 1.

Concordance has a different count

Confirm the same Active selection and apostrophe rule, and turn off Include comments in Concordance.

Export has fewer rows

The CSV follows the filtered table. Reset Filter and Minimum before exporting everything.