LW IT Solutions
« Blog Overview /Digital Analytics / Three Spellings, One Row
This post in other languages:

Three Spellings, One Row

Three Spellings, One Row
Contents
  1. What Source group pulls together
  2. Why the channel report was already right
  3. What stays fragmented
  4. Where the new dimension can get in the way
  5. The filter from the same day, working the other way round
  6. What can be checked

One platform, five rows. facebook, fb, m.facebook.com, Meta-facebook and facebook.com are five distinct values of the source dimension, so five rows in the report – and none of them is wrong. That is simply how the referral arrived.

On 11 June 2026 Google put a dimension up against this. Source group pulls the usual spellings of a platform into a single value, with no change to the tag and no setting in the admin section. The benefit is real, but it sits somewhere other than the announcement suggests: the channel report was already correct.

Flow diagram with three columns: on the left five raw values of the source dimension with their session counts, in the middle a single collecting node holding the value Facebook, on the right two channels Organic Social and Paid Social; the ribbons merge on the left and split again on the right, with two of them crossing
The ribbons merge in the middle and separate again on the right. The group bundles by platform, the channel splits by medium – and the fact that two ribbons cross while doing so is the whole point.

What Source group pulls together

Google’s note describes it as a new dimension concept that consolidates source values for common online platforms, naming Facebook, Instagram and TikTok as examples. Beyond those, the dimension recognises Pinterest, Amazon, YouTube, Google Search and Google Maps, along with ChatGPT and Perplexity. Google maintains the mapping table. No custom list can be supplied, and no unwanted mapping can be switched off.

The dimension is formed at query time, not at collection time. That is the reason it applies retroactively across the whole period: what remains stored is m.facebook.com, and the group only comes into being on reading. A year-over-year comparison therefore has no break anywhere in it, which sets Source group apart from almost every other change ever made to a mapping.

Why the channel report was already right

The obvious expectation is that five source rows also mean five scrambled channels. That was never the case. Channel grouping does not read source literally but through a source category maintained by Google. Both facebook and fb carry SOURCE_CATEGORY_SOCIAL, and whether that becomes Organic Social or Paid Social is decided by the medium alone.

So even before June there were two Google-maintained naming lists stacked on top of each other, and since then there are three: the source category for the channel, the source group for the display, and the collected raw values underneath. The three are not reconciled with one another, because they answer different questions. Treating them as one thing means looking for the fault in the wrong place later on.

What stays fragmented

Google’s list knows platforms, not in-house spellings. Everything originating in a company’s own campaigns remains exactly as inconsistent as it was created.

source values from an ordinary property

  facebook          fb                              -> Facebook
  m.facebook.com    Meta-facebook   facebook.com    -> Facebook
  chatgpt.com       chat.openai.com                 -> ChatGPT

  newsletter        Newsletter      nl              -> three rows
  email             e-mail          E-Mail          -> three rows
  partner-a.com     Partner-A       partnera        -> three rows

  Google maintains the upper half. The lower half is created
  in-house, and there Source group changes nothing.

The remedy is the same as before: a fixed spelling for utm_source and utm_medium, written down in a table and looked up whenever a campaign is created. A dimension that tidies up other people’s platform names does not take that work away.

Where the new dimension can get in the way

A saved exploration filtering on source exactly matching facebook still filters on the raw value and leaves the other four spellings outside. Two reports covering the same period can now show different row counts without either being broken – one reads the group, the other the raw value.

The BigQuery export continues to write the collected values. Analysis done there still has the five rows and has to reproduce the grouping by hand. Wherever interface reports and export queries are used side by side, this creates a new place where two numbers legitimately disagree.

The rollout is gradual. A dimension present in one property and not yet in another is not a fault at the moment, it is the current state.

The filter from the same day, working the other way round

A second change appeared on 11 June that reads well beside the first: hostname filters, a new kind of data filter in the admin section that excludes events based on their hostname. They are meant against hits from unfamiliar domains and against figures from a staging environment.

Both changes concern data cleanliness, but they act in opposite directions in time. Source group applies retroactively to everything already collected. A data filter acts only forward and leaves the past untouched. Switching both on the same day produces a tidied source list across the entire period and a data set that is filtered differently from precisely that day onward.

What can be checked

Three steps are enough to establish the local situation. First, a look at a report with the Source group dimension alongside: if it is offered, the rollout has arrived. Second, a comparison of the session total per group with the total of the matching raw values – if they disagree, a spelling is present that Google does not recognise. Third, a pass through the saved explorations looking for filters that point at raw values.

What remains afterwards is the old task in new light. Source group tidies the part of the list nobody in the building was responsible for anyway. The part somebody is responsible for stands exactly where it stood.

Lukas Wojcik

Lukas Wojcik

Systems architect and technology enthusiast specializing in scalable tracking solutions, GMP Stack (GA4 & GTM), and robust backend architectures. Advocate for clean code and privacy-first design.

Get in Touch

Briefly describe your project or inquiry for a tailored response. This site is protected by reCAPTCHA.

ALL ARTICLES & CATEGORIES

CCTV

Follow this category by RSS

Data Privacy

Follow this category by RSS

Digital Analytics

Follow this category by RSS

Digital Marketing

Follow this category by RSS

IT & Networks

Follow this category by RSS

Raspberry PI

Follow this category by RSS

Smart Home

Follow this category by RSS

Web Development

Follow this category by RSS

Wordpress Hacks

Follow this category by RSS