You Can’t Match What You Can’t Understand: Why Name & Address Parsing Is the Bedrock of Data Processing

By |2026-08-24T11:08:55+00:00August 24th, 2026|

There is a part of the data processing world that rarely gets the attention it deserves. It isn't glamorous. It doesn't generate exciting dashboards. It probably won't feature prominently in an AI strategy presentation. But get it wrong and virtually everything that follows can go wrong. Name and address parsing. At The Software Bureau, we believe accurate parsing is the bedrock of batch data processing. Whether the objective is address matching, deduplication, sorting, indexing, suppression, enrichment or the creation of a Single Customer View (SCV), it all starts with understanding exactly what is contained within a name and address record. And that is considerably more complicated than it sounds. What exactly is parsing? At its simplest, parsing is the process [...]

The History of Royal Mail’s Postcode Address File (PAF): A Timeline of Addressing Evolution

By |2026-06-10T09:27:23+00:00June 10th, 2026|

The Royal Mail’s Postcode Address File (PAF) is one of the most important yet often overlooked datasets in the UK. It underpins everything from mail delivery to e-commerce checkouts, credit checks, and emergency response systems. But PAF didn’t appear overnight. It evolved over decades alongside changes in addressing, technology, and society. This post explores the history and evolution of PAF, from the early days of postal districts to the digital infrastructure we rely on today. Before PAF: The Foundations of Modern Addressing (Pre-1970) 1850s–1910s: Postal districts and sorting efficiency Major cities such as London began using postal districts (e.g. “EC” or “SW”) in the mid-19th century. These were introduced to improve sorting efficiency as urban populations grew rapidly. However, outside [...]

How confident are you in the accuracy of your contact data?

By |2026-06-08T13:14:52+00:00June 8th, 2026|

Data hygiene is not just about tidy databases. It is about knowing when your contacts have moved home or sadly passed away, and ensuring your communications reflect that reality. Failing to keep this information up to date can lead to wasted budget, poor campaign performance and reputational risk. That is why we have created our FREE Complete Guide to Data Hygiene. It focuses on helping organisations identify movers and the deceased, so you can maintain accurate, respectful and effective outreach. Inside the guide, you will learn how to: - Identify customers who have moved home - Remove records linked to deceased individuals - Reduce wasted mail and improve campaign efficiency - Protect your brand by communicating responsibly - Keep your [...]

A New Era for Data Matching in the UK: Powerful, Proven and Priced for Reality

By |2026-05-20T10:18:05+00:00May 20th, 2026|

For organisations running large CRM, ERP and customer data platforms, accurate data matching has never been more critical, or more expensive. Licence fees on long-established tools are rising, innovation has stalled, and support is drifting away from UK data teams. At The Software Bureau, we believe there is a better way. And we are building it. The Market Has Changed but Pricing Has Not Kept Pace with Value Many organisations rely on long-established data matching tools that have been embedded into their operations for years. However, as ownership structures evolve and software portfolios are consolidated into larger global organisations, customers are often faced with: Significant increases in ongoing licence costs Reduced flexibility in how solutions are deployed and used Support [...]

Why Name and Address Parsing Is the Foundation of Every Data Quality Success

By |2026-05-11T14:47:09+00:00May 11th, 2026|

What is the single most important capability behind every effective name and address solution? Accurate parsing. Before data can be enhanced, cleansed, screened, sorted or reformatted, it must first be understood. Name and address parsing is the process that makes that understanding possible, transforming raw text into structured, reliable components that software can act on with confidence. At The Software Bureau, parsing is not a feature. It is the foundation. Parsing: the invisible engine of data quality Every contact record starts life as unstructured data. Names arrive in countless formats. Addresses vary by country, convention, abbreviation and free‑form input. Titles, suffixes, building names, sub‑premises and delivery points are often mixed together in ways that defy simple rules. Without accurate parsing, [...]

Eliminating Mojibake: How Our New SwiftCore Translation Fix Improves Data Integrity

By |2026-04-22T10:13:43+00:00April 22nd, 2026|

Text corruption has long been one of the most frustrating obstacles in data processing. Anyone who has worked with large, diverse data sources will have encountered the odd tangle of characters that appear in place of clean text. This problem, known as Mojibake, arises when character encoding is misinterpreted. It looks like a small nuisance on the surface, but in practice it can disrupt analytics, weaken matching, and undermine entire workflows. To address this, we have introduced a translation fix within our SwiftCore processing engine. It is designed to prevent Mojibake at source, repair corrupt text when encountered, and improve the overall integrity of every dataset that passes through the platform. What causes Mojibake in the first place? Mojibake is [...]

Why Returned Mail Can No Longer Be Ignored: An Industry Wide Call to Action

By |2026-04-13T08:55:14+00:00April 13th, 2026|

Returned Mail has long been an inconvenient truth in the UK postal ecosystem. Despite its scale, cost and operational drag, it remains largely unmeasured and therefore unmanaged. My recent LinkedIn post highlighted a simple but troubling reality: Royal Mail continues to double handle significant volumes of Return to Sender mail without recording volume, root cause or resulting waste. This is not solely a Royal Mail issue. It represents a wider industry failure to create effective feedback loops that improve address quality, reduce waste and lower costs for everyone involved. It is time to move this conversation beyond anecdote and towards collective action. The Silent Cost of Returned Mail RTS mail remains one of the least transparent operational processes within Royal [...]

Elevate Your Data Hygiene Expertise with Free One to One Training

By |2026-03-16T16:59:49+00:00March 16th, 2026|

Data hygiene has never been more important. Whether you are an intermediary delivering data services for clients or a brand owner managing large volumes of customer information, maintaining clean, accurate and compliant data is fundamental to performance and trust. To support everyone involved in promoting best practice, we are offering free one to one data hygiene training. This is available to all, whether you are already a client or entirely new to us. Why Data Hygiene Matters For intermediaries such as data bureaux and mail producers, data hygiene services are a proven driver of incremental revenue. They improve the quality of the datasets you process, reduce wasted output, increase client satisfaction and often open up entirely new service lines. Brand [...]

Data Cleansing Cadence: Why “Set and Forget” Is Costing You More Than You Think

By |2026-01-21T13:45:32+00:00January 21st, 2026|

I joined a great webinar today hosted by Paragon and the DMA, where one of the topics covered by Hannah Stapleford really stood out to me: Data Hygiene Cadence. I have to admit, I initially had to look up the word cadence. In simple terms, it means a regular and repeated pattern of activity. Once I had done that, it struck me just how perfectly the term describes where brands need to be when it comes to managing data quality within their customer data environments. Hannah was absolutely spot on, and she shared a couple of statistics that make this topic impossible to ignore: Around 10 percent of the UK population moves home each year Around 1 percent of the [...]

Why Processing Customer Data Against Change of Address Files is Crucial for Business Success

By |2026-01-14T12:45:08+00:00January 14th, 2026|

Every year, UK businesses waste £1 billion on mistargeted mailings: sending communications to people who have moved or even passed away. Beyond the financial cost, this erodes brand reputation and risks GDPR non-compliance. With 3.5 million households moving annually and 548,000 deaths each year, customer data decays at an alarming rate. This is why processing your data against Change of Address (COA) datasets is not just best practice; it is essential. What Are COA Data Sets? Change of Address files, such as Royal Mail’s National Change of Address (NCOA) and Experian’s Absolute Contacts, identify when a person has moved and provide their new address. These datasets allow businesses to: Suppress goneaways (people who have moved and not informed you). Update [...]

Go to Top