NewYour coding agent can read the release notes before it upgrades.Set up the MCP server →
PyPI · #4100 most downloaded on PyPI
A record linkage toolkit for linking and deduplication
Last release 3 years ago
no release in 18 months
Release timing varies
gaps range from 2 weeks to 2.4 years
Most releases are documented
notes for 17 of 23 stable releases
Nothing withdrawn
no release was ever pulled
10 years old
23 releases · first in 2016
A new release of recordlinkage after a long time (too long, I'm sorry). This release bumps the minor version to 0.16. This version supports pandas 2 a
A new release of recordlinkage after a long time (too long, I'm sorry). This release bumps the minor version to 0.16. This version supports pandas 2 and pandas 1. It doesn't contain any structural changes or improvements to the API.
Full Changelog: https://github.com/J535D165/recordlinkage/compare/v0.15...v0.16
One column per quarter.
A new release of recordlinkage after a long time (too long, I'm sorry). This release bumps the minor version to 0.16. This version supports pandas 2 and pandas 1. It doesn't contain any structural changes or improvements to the API.
Full Changelog: v0.15...v0.16
Remove deprecated recordlinkage classes
Special thanks to Tomasz Waleń @twalen and other contributors for their work on this release.
Various updates in relation to deprecation warnings in third-party libraries such as sklearn, pandas and networkx.
.labels by .codes for pandas.MultiIndex objects for newer versions of pandas (>0.24). (#103)Fix distribution problem.
Fix distribution problem.
resolve conflict with threshold and missing value
Nothing published for this version
Minor installation improvement. Exclude unwanted files
Fix installation issue. Submodule 'preprocessing' was not added to the source distribution.
The submodule 'standardise' is renamed. The new name is 'preprocessing'. The submodule 'standardise' will get deprecated in a next version.
Note: In the next release, the Pairs class will get removed. Migrate now.
print statement in the geo compare algorithm removed.
A new compare API. The new Compare class no longer takes the datasets and pairs as arguments. The actual computation is now performed when calling .co
.compute(PAIRS, DF1, DF2). The documentation is updated as well, but
still needs improvement.A new index API. The new index API is no longer a single class (recordlinkage.Pairs(...)) with all the functionality in it. The new API is based on Te
recordlinkage.Pairs(...)) with all the functionality in it. The new API
is based on Tensorflow and FEBRL. With the new structure, it easier to
parallise the record linkage process. In future releases, this will be
implemented natively. See the reference page for more information and migrating. <http://recordlinkage.readthedocs.io/en/latest/ref-index.html>_binary_comparisons is renamed. The new name of the function
is binary_vectors. Documentation added to RTD.Issues solved with rendering docs on ReadTheDocs. Still not clear what is going on with the autodoc_mock_imports in the sphinx conf.py file. Maybe a b
autodoc_mock_imports in the sphinx conf.py file. Maybe
a bug in sphinx.Add additional arguments to the function that downloads and loads the krebsregister data. The argument missing_values is used to fill missing values.
missing_values is used to fill missing
values. Default: nothing is done. The argument shuffle is used to
shuffle the records. Default is True.Compare class and its algorithms. Making use
of nose-parameterized module.max_number_of_pairs to get the maximum number of pairs.low_memory for compare class.binary_comparisons in the recordlinkage.datasets.random module.tox.ini to test packaging and installation of package.Compare
module. Especially label handling is improved.Nothing published for this version
Nothing published for this version
Nothing published for this version
This version includes the following updates:
This version includes the following updates:
__sub__ is no longer used for computing the difference of Index objects. It is now replaced by ``INDEX.difference(OTHER_INDEX).clean function.Nothing published for this version
Fixes a serious bug with deduplication (thanks to https://github.com/dserban).
Nothing published for this version
This version contains a lot of changes to the API. Hopefully, there are no large API changes needed for now.
This version contains a lot of changes to the API. Hopefully, there are no large API changes needed for now.
numerical is now named numeric and fuzzy is now named string.Update the parameters of the Logistic Regression Classifier manually. In literature, this is often denoted as the _deterministic record linkage_.
Your coding agent can read these notes before it upgrades. Set up the MCP server →