Defined Icon

REPORT · JULY 2026

The "free" address database from OpenStreetMap

porównanie usług standaryzacji i geokodowania RANKING comparison standardization geocoding tools

8.4M

address records in OSM (after deduplication)

−410K

fewer addresses than in the Algolytics Technologies database

45%

of OSM records have had no update since 2019

~2.05M

records nationwide that need correction, completion or a location fix

156,900 PLN

total cost of preparing and maintaining the data over 5 years

"Free data" is not the same as a production address database. Value is determined by completeness, consistency,
timeliness and stable address identification - and the decision to use OSM is worth making not on the price of downloading the data,
but on the total cost of bringing it up to production quality and maintaining it over time.

RECOMMENDATION FROM THE REPORT

RESULTS OVERVIEW

Free access to data does not mean there are no costs

  • Completeness. More than 2 million records to complete or correct

OSM contains 8.39 million addresses - 410,000 fewer than the Algolytics database. 7% (~617,000) have no counterpart, and 17.5% (~1.4 million) differ in at least one field.

  • Correctness. Agreement drops once TERYT is taken into account

Full field agreement is 82.5%, but with TERYT identifiers only 36.2%. Postal codes are the weakest - wrong on average in ~13% of addresses (nearly 32% in the Lubuskie Voivodeship).

  • Timeliness. The data is refreshed by volunteers - selectively

As many as 45% of OSM records have not changed since 2019. The coordinates stay accurate (median ~3 m; errors over 50 m are < 1%) because they come from the EMUiA register.

  • Heterogeneity. A national average is not enough

Quality is not evenly distributed: in extreme cases the same metrics vary between municipalities from 0% to nearly 100%. Good quality in one region is no guarantee of it nationwide.

OSM field completness

See what a "free" OpenStreetMap database really costs

Download the full report as a PDF: methodology, detailed regional statistics, an analysis of timeliness and positional accuracy, and a full cost calculation (TCO, 5 years).

porównanie usług standaryzacji i geokodowania
HOW WE TESTED IT

A methodology built on a production reference database

STEP 1

8.4 million addresses from the OSM project

We downloaded every object carrying "addr" tags (as of 5 June 2026) - 3.2 million points and 5.4 million polygons. After mapping the tags to the Polish address model, standardising the ambiguous addr:place, removing records without a number and deduplicating, a set of 8,392,914 records remained.

STEP 2

The Algolytics reference database

We compared OSM with the Algolytics database (as of 15 April 2026, 8.8 million active addresses), developed over more than 10 years on the basis of EMUiA, TERYT and the PNA postal-code directory - with stable identifiers and quarterly updates. In production we use it to standardise more than 600 million addresses a year at over 99% accuracy.

STEP 3

Three quality dimensions → TCO

We compared the two datasets on completeness, correctness, timeliness and positional accuracy - nationally, by voivodeship and at municipality level. Finally, we calculated the cost of bringing OSM up to production quality over a 5-year horizon.

WHO IT'S FOR

This report is for you if you're considering basing your processes on free address data

Logistics

Accurate address location, fewer failed deliveries and complaints, better route and service-area planning.

Finance and telecoms

Stable address and TERYT identifiers: the foundation for reliably matching and enriching data in the warehouse, used in scoring and anti-fraud.

E-commerce and CRM

A consistent customer base with no duplicates or typos, a reliable postal code and geographic segmentation.

Data architects and CTOs

Hard numbers (completeness, TCO) for the "build a database from OSM vs. buy a ready-made one" decision.

Meet Algolytics Technologies

Algolytics Technologies is a Polish technology company specialising in artificial intelligence, machine learning and advanced data analytics. We deliver a comprehensive Data Science Platform focused on automating business processes and supporting data-driven decision-making.

In Location Intelligence we develop the Data Quality service - our own tool for address standardisation and geocoding, available for Poland and four other Central European countries. More than 50 of the largest financial institutions, telecoms, e-commerce and logistics companies trust us.