diff --git a/PKG-INFO b/PKG-INFO
index 83e34df0..fab895dd 100644
--- a/PKG-INFO
+++ b/PKG-INFO
@@ -1,250 +1,250 @@
 Metadata-Version: 2.1
 Name: swh.storage
-Version: 0.39.0
+Version: 0.40.0
 Summary: Software Heritage storage manager
 Home-page: https://forge.softwareheritage.org/diffusion/DSTO/
 Author: Software Heritage developers
 Author-email: swh-devel@inria.fr
 License: UNKNOWN
 Project-URL: Bug Reports, https://forge.softwareheritage.org/maniphest
 Project-URL: Funding, https://www.softwareheritage.org/donate
 Project-URL: Source, https://forge.softwareheritage.org/source/swh-storage
 Project-URL: Documentation, https://docs.softwareheritage.org/devel/swh-storage/
 Platform: UNKNOWN
 Classifier: Programming Language :: Python :: 3
 Classifier: Intended Audience :: Developers
 Classifier: License :: OSI Approved :: GNU General Public License v3 (GPLv3)
 Classifier: Operating System :: OS Independent
 Classifier: Development Status :: 5 - Production/Stable
 Requires-Python: >=3.7
 Description-Content-Type: text/markdown
 Provides-Extra: testing
 Provides-Extra: journal
 License-File: LICENSE
 License-File: AUTHORS
 
 swh-storage
 ===========
 
 Abstraction layer over the archive, allowing to access all stored source code
 artifacts as well as their metadata.
 
 See the
 [documentation](https://docs.softwareheritage.org/devel/swh-storage/index.html)
 for more details.
 
 ## Quick start
 
 ### Dependencies
 
 Python tests for this module include tests that cannot be run without a local
 Postgresql database, so you need the Postgresql server executable on your
 machine (no need to have a running Postgresql server). They also expect a
 cassandra server.
 
 #### Debian-like host
 
 ```
 $ sudo apt install libpq-dev postgresql-11 cassandra
 ```
 
 #### Non Debian-like host
 
 The tests expects the path to `cassandra` to either be unspecified, it is then
 looked up at `/usr/sbin/cassandra`, either specified through the environment
 variable `SWH_CASSANDRA_BIN`.
 
 Optionally, you can avoid running the cassandra tests.
 
 ```
 (swh) :~/swh-storage$ tox -- -m 'not cassandra'
 ```
 
 ### Installation
 
 It is strongly recommended to use a virtualenv. In the following, we
 consider you work in a virtualenv named `swh`. See the
 [developer setup guide](https://docs.softwareheritage.org/devel/developer-setup.html#developer-setup)
 for a more details on how to setup a working environment.
 
 
 You can install the package directly from
 [pypi](https://pypi.org/p/swh.storage):
 
 ```
 (swh) :~$ pip install swh.storage
 [...]
 ```
 
 Or from sources:
 
 ```
 (swh) :~$ git clone https://forge.softwareheritage.org/source/swh-storage.git
 [...]
 (swh) :~$ cd swh-storage
 (swh) :~/swh-storage$ pip install .
 [...]
 ```
 
 Then you can check it's properly installed:
 ```
 (swh) :~$ swh storage --help
 Usage: swh storage [OPTIONS] COMMAND [ARGS]...
 
   Software Heritage Storage tools.
 
 Options:
   -h, --help  Show this message and exit.
 
 Commands:
   rpc-serve  Software Heritage Storage RPC server.
 ```
 
 
 ## Tests
 
 The best way of running Python tests for this module is to use
 [tox](https://tox.readthedocs.io/).
 
 ```
 (swh) :~$ pip install tox
 ```
 
 ### tox
 
 From the sources directory, simply use tox:
 
 ```
 (swh) :~/swh-storage$ tox
 [...]
 ========= 315 passed, 6 skipped, 15 warnings in 40.86 seconds ==========
 _______________________________ summary ________________________________
   flake8: commands succeeded
   py3: commands succeeded
   congratulations :)
 ```
 
 Note: it is possible to set the `JAVA_HOME` environment variable to specify the
 version of the JVM to be used by Cassandra. For example, at the time of writing
 this, Cassandra does not support java 14, so one may want to use for example
 java 11:
 
 ```
 (swh) :~/swh-storage$ export JAVA_HOME=/usr/lib/jvm/java-14-openjdk-amd64/bin/java
 (swh) :~/swh-storage$ tox
 [...]
 ```
 
 ## Development
 
 The storage server can be locally started. It requires a configuration file and
 a running Postgresql database.
 
 ### Sample configuration
 
 A typical configuration `storage.yml` file is:
 
 ```
 storage:
   cls: postgresql
   db: "dbname=softwareheritage-dev user=<user> password=<pwd>"
   objstorage:
     cls: pathslicing
     root: /tmp/swh-storage/
     slicing: 0:2/2:4/4:6
 ```
 
 which means, this uses:
 
 - a local storage instance whose db connection is to
   `softwareheritage-dev` local instance,
 
 - the objstorage uses a local objstorage instance whose:
 
   - `root` path is /tmp/swh-storage,
 
   - slicing scheme is `0:2/2:4/4:6`. This means that the identifier of
     the content (sha1) which will be stored on disk at first level
     with the first 2 hex characters, the second level with the next 2
     hex characters and the third level with the next 2 hex
     characters. And finally the complete hash file holding the raw
     content. For example: 00062f8bd330715c4f819373653d97b3cd34394c
     will be stored at 00/06/2f/00062f8bd330715c4f819373653d97b3cd34394c
 
 Note that the `root` path should exist on disk before starting the server.
 
 
 ### Starting the storage server
 
 If the python package has been properly installed (e.g. in a virtual env), you
 should be able to use the command:
 
 ```
 (swh) :~/swh-storage$ swh storage rpc-serve storage.yml
 ```
 
 This runs a local swh-storage api at 5002 port.
 
 ```
 (swh) :~/swh-storage$ curl http://127.0.0.1:5002
 <html>
 <head><title>Software Heritage storage server</title></head>
 <body>
 <p>You have reached the
 <a href="https://www.softwareheritage.org/">Software Heritage</a>
 storage server.<br />
 See its
 <a href="https://docs.softwareheritage.org/devel/swh-storage/">documentation
 and API</a> for more information</p>
 ```
 
 ### And then what?
 
 In your upper layer
 ([loader-git](https://forge.softwareheritage.org/source/swh-loader-git/),
 [loader-svn](https://forge.softwareheritage.org/source/swh-loader-svn/),
 etc...), you can define a remote storage with this snippet of yaml
 configuration.
 
 ```
 storage:
   cls: remote
   url: http://localhost:5002/
 ```
 
 You could directly define a postgresql storage with the following snippet:
 
 ```
 storage:
   cls: postgresql
   db: service=swh-dev
   objstorage:
     cls: pathslicing
     root: /home/storage/swh-storage/
     slicing: 0:2/2:4/4:6
 ```
 
 ## Cassandra
 
 As an alternative to PostgreSQL, swh-storage can use Cassandra as a database backend.
 It can be used like this:
 
 ```
 storage:
   cls: cassandra
   hosts:
     - localhost
   objstorage:
     cls: pathslicing
     root: /home/storage/swh-storage/
     slicing: 0:2/2:4/4:6
 ```
 
 The Cassandra swh-storage implementation supports both Cassandra >= 4.0-alpha2
 and ScyllaDB >= 4.4 (and possibly earlier versions, but this is untested).
 
 While the main code supports both transparently, running tests
 or configuring the schema requires specific code when using ScyllaDB,
 enabled by setting the `SWH_USE_SCYLLADB=1` environment variable.
 
 
diff --git a/debian/changelog b/debian/changelog
index bd2a0dfe..1ac3ca6f 100644
--- a/debian/changelog
+++ b/debian/changelog
@@ -1,2754 +1,2765 @@
-swh-storage (0.39.0-1~swh1~bpo10+1) buster-swh; urgency=medium
+swh-storage (0.40.0-2~swh1) unstable-swh; urgency=medium
 
-  * Rebuild for buster-swh
+  * Update dependencies in d/control.
 
- -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 29 Oct 2021 09:23:51 +0000
+ -- David Douard <david.douard@sdfa3.org>  Tue, 16 Nov 2021 14:37:20 +0100
+
+swh-storage (0.40.0-1~swh1) unstable-swh; urgency=medium
+
+  * New upstream release 0.40.0     - (tagged by David Douard
+    <david.douard@sdfa3.org> on 2021-11-16 10:11:25 +0100)
+  * Upstream changes:     - v0.40.0     - Add support for a redis-based
+    reporting for invalid mirrorred objects     - Remove now useless
+    fixers in storage/fixer.py     - Add a new --type option to 'swh
+    strorage repay'     - Update extrinsic metadata specs
+
+ -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 16 Nov 2021 09:45:39 +0000
 
 swh-storage (0.39.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.39.0     - (tagged by Antoine Lambert
     <anlambert@softwareheritage.org> on 2021-10-28 17:09:40 +0200)
   * Upstream changes:     - version 0.39.0
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 29 Oct 2021 09:17:36 +0000
 
 swh-storage (0.38.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.38.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2021-10-11 13:28:27
     +0200)
   * Upstream changes:     - v0.38.0     - buffer: add some debug logging
     for number of objects sent     - buffer: add a threshold for the
     estimated size of revision and release batches     - buffer: add a
     threshold for the number of revision parents in one batch     -
     buffer: add a threshold for the number of directory entries in one
     batch     - filter: add filtering for release_add     - filter: do
     not call the underlying functions if there's nothing to add     -
     buffer: Ensure that we don't send data from empty buffers
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 11 Oct 2021 14:58:23 +0000
 
 swh-storage (0.37.1-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.37.1     - (tagged by David Douard
     <david.douard@sdfa3.org> on 2021-09-29 11:31:33 +0200)
   * Upstream changes:     - v0.37.1
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 29 Sep 2021 10:12:31 +0000
 
 swh-storage (0.37.0-1~swh2) unstable-swh; urgency=medium
 
   * Bump new release
 
  -- Antoine R. Dumont (@ardumont) <ardumont@softwareheritage.org>  Thu, 16 Sep 2021 09:51:51 +0200
 
 swh-storage (0.37.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.37.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2021-09-15 18:44:19
     +0200)
   * Upstream changes:     - v0.37.0     - Allow filtering extids per
     extid_version/extid_type when reading     -
     migrate_extrinsic_metadata: Fix edge cases (missing f-
     stringification, remaining pypi     - issues, ...)     - cassandra:
     Make directory_ls fetch contents in batch instead of one-by-one     -
     cassandra: Add option to select (hopefully) more efficient batch
     insertion algos     - cassandra: Remove stat_counters.     -
     cassandra: generate statsd metrics on method calls     -
     content_get: Fetch rows concurrently     -
     directory_entry_add_batch: Remove the temporary prepared statement
     entirely     - directory_entry_add_batch: Reduce churn of prepared
     statements     - postgresql: Fix a column order mismatch between the
     query and object builder     - Add counting storage proxy
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 16 Sep 2021 06:42:23 +0000
 
 swh-storage (0.36.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.36.0     - (tagged by Vincent SELLIER
     <vincent.sellier@softwareheritage.org> on 2021-08-24 16:51:07 +0200)
   * Upstream changes:     - v0.36.0     - changelog:     - Add cvs as
     supported revision_type     - Add test for origin_visit_get_latest
     in presence of mismatched id and date orders     - cassandra: Bump
     next_visit_id when origin_visit_add is called by a replayer     -
     cassandra: Make content_missing query in batches     - backfill: add
     extra where clause to use the right index for extid requests
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 24 Aug 2021 15:01:32 +0000
 
 swh-storage (0.35.1-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.35.1     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2021-08-20 11:53:33 +0200)
   * Upstream changes:     - v0.35.1     - * cassandra: Fix crash when
     using _missing() functions with more than 100 ids with ScyllaDB.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 20 Aug 2021 10:01:16 +0000
 
 swh-storage (0.35.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.35.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2021-07-28 10:36:09
     +0200)
   * Upstream changes:     - v0.35.0     - Implement storage of the
     ExtID.extid_version field
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 28 Jul 2021 08:43:06 +0000
 
 swh-storage (0.34.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.34.0     - (tagged by Vincent SELLIER
     <vincent.sellier@softwareheritage.org> on 2021-07-07 18:22:00 +0200)
   * Upstream changes:     - v0.34.0     - cassandra: allow to configure
     the consistency level
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 07 Jul 2021 16:58:42 +0000
 
 swh-storage (0.33.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.33.0     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2021-07-05 16:48:16 +0200)
   * Upstream changes:     - v0.33.0     - * Add endpoint
     raw_extrinsic_metadata_get_authorities
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 05 Jul 2021 15:00:12 +0000
 
 swh-storage (0.32.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.32.0     - (tagged by Vincent SELLIER
     <vincent.sellier@softwareheritage.org> on 2021-06-28 15:35:44 +0200)
   * Upstream changes:     - v0.32.0     - * cassandra: Add support for
     non-ASCII origin 'URLs'.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 28 Jun 2021 16:20:21 +0000
 
 swh-storage (0.31.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.31.0     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2021-06-25 11:17:50 +0200)
   * Upstream changes:     - v0.31.0     - * cassandra: Add partial
     support for ScyllaDB     - * mypy: Fix errors with release >= v0.900
     (but breaks older mypy versions)     - * Add endpoints to access
     REMD by id
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 25 Jun 2021 09:26:09 +0000
 
 swh-storage (0.30.1-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.30.1     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2021-05-21 10:09:02
     +0200)
   * Upstream changes:     - v0.30.1     - Finalize the config "local"
     deprecation in favor of "postgresql"     - tests: Make test
     parameters order deterministic, so they don't crash pytest-xdist
     - test_cassandra: Improve error when the process is started but not
     listening
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 21 May 2021 08:22:33 +0000
 
 swh-storage (0.30.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.30.0     - (tagged by David Douard
     <david.douard@sdfa3.org> on 2021-05-18 16:34:25 +0200)
   * Upstream changes:     - v0.30.0
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 18 May 2021 14:45:21 +0000
 
 swh-storage (0.29.1-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.29.1     - (tagged by Nicolas Dandrimont
     <nicolas@dandrimont.eu> on 2021-05-14 18:31:52 +0200)
   * Upstream changes:     - Release swh.storage 0.29.1     - Add missing
     db migration
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 14 May 2021 16:59:42 +0000
 
 swh-storage (0.29.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.29.0     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2021-05-11 15:04:58 +0200)
   * Upstream changes:     - v0.29.0     - * Make the
     TenaciousProxyStorage retry when a single object add fails     - *
     Move all proxy storages in swh/storage/proxies/     - * Deprecate
     the "local" storage cls in favor of "postgresql"     - * cassandra:
     Add tests checking directory_add and snapshot_add are atomic.     -
     * Add endpoint directory_get_entries, to quickly list a directory's
     entries     - * content_get: Add support for queries by sha1_git
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 11 May 2021 13:12:42 +0000
 
 swh-storage (0.28.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.28.0     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2021-05-06 15:52:03 +0200)
   * Upstream changes:     - v0.28.0     - * Normalize all
     Storage.xxx_add() methods to return a summary     - * cassandra: Add
     'check_missing' option, to allow updating objects     - * cassandra:
     Add a test of a 'complex' migration, with a PK update     - * Add a
     new TenaciousProxyStorage     - * Make postgresql's origin_add not
     raise an error in case of conflict     - * Stop storing
     authority/fetcher metadata.     - * tenacious: Document potential
     issues about objects being dropped     - * Use swh.core 0.14
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 06 May 2021 14:06:51 +0000
 
 swh-storage (0.27.4-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.27.4     - (tagged by Antoine Lambert
     <antoine.lambert@inria.fr> on 2021-04-29 14:38:49 +0200)
   * Upstream changes:     - version 0.27.4
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 29 Apr 2021 13:04:46 +0000
 
 swh-storage (0.27.3-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.27.3     - (tagged by Antoine Lambert
     <antoine.lambert@inria.fr> on 2021-04-09 14:59:36 +0200)
   * Upstream changes:     - version 0.27.3
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 09 Apr 2021 13:06:58 +0000
 
 swh-storage (0.27.2-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.27.2     - (tagged by David Douard
     <david.douard@sdfa3.org> on 2021-04-07 15:06:41 +0200)
   * Upstream changes:     - v0.27.2
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 08 Apr 2021 08:05:43 +0000
 
 swh-storage (0.27.1-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.27.1     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2021-03-30 17:47:03 +0200)
   * Upstream changes:     - v0.27.1     - * buffer: Add support for
     'extid'
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 30 Mar 2021 15:59:01 +0000
 
 swh-storage (0.27.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.27.0     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2021-03-29 14:33:24 +0200)
   * Upstream changes:     - v0.27.0     - * origin_visit_status_add: Fix
     inconsistent/incorrect errors when type is None and visit is
     missing.     - * extid: remove unicity on (extid_type, extid) and
     (target_type, target)
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 29 Mar 2021 12:44:14 +0000
 
 swh-storage (0.26.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.26.0     - (tagged by Nicolas Dandrimont
     <nicolas@dandrimont.eu> on 2021-03-22 14:44:35 +0100)
   * Upstream changes:     - Release swh.storage v0.26.0     - Move
     raw_extrinsic_metadata deduplication to use a new id column.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 22 Mar 2021 21:53:39 +0000
 
 swh-storage (0.25.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.25.0     - (tagged by Antoine Lambert
     <antoine.lambert@inria.fr> on 2021-03-18 13:55:10 +0100)
   * Upstream changes:     - version 0.25.0
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 18 Mar 2021 13:02:02 +0000
 
 swh-storage (0.24.1-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.24.1     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2021-03-04 23:32:36 +0100)
   * Upstream changes:     - v0.24.1     - * tests: Drop hypothesis < 6
     requirement     - * Remove the remaining references to the
     deprecated SWHID class     - * postgresql: Ensure a minimum limit
     for the snapshot branches query
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 04 Mar 2021 22:39:03 +0000
 
 swh-storage (0.24.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.24.0     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2021-03-02 10:00:23 +0100)
   * Upstream changes:     - v0.24.0     - * storage_tests: recompute ids
     when evolving RawExtrinsicMetadata objects.     - *
     RawExtrinsicMetadata: update to use the API in swh-model 1.0.0
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 02 Mar 2021 09:11:15 +0000
 
 swh-storage (0.23.2-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.23.2     - (tagged by Antoine Lambert
     <antoine.lambert@inria.fr> on 2021-02-19 11:47:03 +0100)
   * Upstream changes:     - version 0.23.2
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 19 Feb 2021 10:58:50 +0000
 
 swh-storage (0.23.1-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.23.1     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2021-02-16 17:19:00
     +0100)
   * Upstream changes:     - v0.23.1     - Switch anonymized replayer
     test to use pytest parametrization
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 16 Feb 2021 16:28:25 +0000
 
 swh-storage (0.23.0-1~swh2) unstable-swh; urgency=medium
 
   * Fix dependency issue
 
  -- Antoine R. Dumont (@ardumont) <ardumont@softwareheritage.org>  Tue, 16 Feb 2021 14:34:57 +0100
 
 swh-storage (0.23.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.23.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2021-02-15 15:20:21
     +0100)
   * Upstream changes:     - v0.23.0     - storage: Refactor
     OriginVisitStatus instantiation     - db: Unify sql joins on
     origin_visit_status using "USING"     - storage.postgresql: Use
     origin_visit_status.type value as source     - test_replay: Fix hang
     since confluent-kafka 1.6 release     - postgresql: Fix dbversion()
     to return the max version instead of a random one.     - buffer:
     ensure objects are flushed in topological order     - Return an
     accurate summary from buffer's flush() method     - buffer: add
     support for snapshots     - buffer: add type annotations for tests
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 15 Feb 2021 14:39:04 +0000
 
 swh-storage (0.22.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.22.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2021-02-03 12:09:29
     +0100)
   * Upstream changes:     - v0.22.0     - storage: Make
     origin_get_latest_visit_status return OriginVisitStatus     -
     storage: Change origin_visit_status_get_random interface to return
     visit_status     - Write introduction to swh-storage
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 03 Feb 2021 11:15:27 +0000
 
 swh-storage (0.21.1-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.21.1     - (tagged by Vincent SELLIER
     <vincent.sellier@softwareheritage.org> on 2021-01-28 14:11:26 +0100)
   * Upstream changes:     - v0.21.1     - * Correctly return
     origin_visit_status.type value everywhere
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 28 Jan 2021 13:19:24 +0000
 
 swh-storage (0.21.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.21.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2021-01-20 15:42:40
     +0100)
   * Upstream changes:     - v0.21.0     - db: Allow new status values
     not_found, failed to OriginVisitStatus
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 20 Jan 2021 14:52:20 +0000
 
 swh-storage (0.20.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.20.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2021-01-20 10:24:00
     +0100)
   * Upstream changes:     - v0.20.0     - storage: Add persistence of
     the field OriginVisitStatus.type     - backfiller: Add type to the
     origin_visit_status topic     - tests: Make test_content_add_race
     fail for the right reason.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 20 Jan 2021 09:29:54 +0000
 
 swh-storage (0.19.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.19.0     - (tagged by Vincent SELLIER
     <vincent.sellier@softwareheritage.org> on 2021-01-14 11:09:17 +0100)
   * Upstream changes:     - v0.19.0     - * 2021-01-12 Adapt cassandra
     storage to ignore the new OriginVisitStatus.type field     - * 2021-
     01-08 Allow to use the JAVA_HOME environment for cassandra tests
     - * 2021-01-13 Enforce hypothesis <6 to prevent test breakage     -
     * 2021-01-08 Make the CREATE_TABLES_QUERIES in cassandra/schema.py
     an explicit list     - * 2020-12-18 Add a cli section in the doc
     - * 2020-11-24 storage.backfill: Allow cli run for
     origin_visit_status as well     - * 2020-11-24 conftest: Reference
     swh.core.db.pytest_plugin
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 14 Jan 2021 10:18:31 +0000
 
 swh-storage (0.18.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.18.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-11-23 14:46:41
     +0100)
   * Upstream changes:     - v0.18.0     - requirements-test.txt: Drop no
     longer needed pytest-postgresql requirement     - backfill: Reverse
     flawed logic in SnapshotBranch generation     -
     migrate_extrinsic_metadata: don't crash when deb revisions aren't
     referenced by any snapshot
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 23 Nov 2020 13:52:32 +0000
 
 swh-storage (0.17.2-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.17.2     - (tagged by Nicolas Dandrimont
     <nicolas@dandrimont.eu> on 2020-11-13 11:56:37 +0100)
   * Upstream changes:     - Release swh.storage 0.17.2     - Future-
     proof get_journal_writer by setting the value_sanitizer argument
     - migrate_extrinsic_metadata improvements     - backfill: only flush
     on every batch
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 13 Nov 2020 11:05:35 +0000
 
 swh-storage (0.17.1-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.17.1     - (tagged by Antoine Lambert
     <antoine.lambert@inria.fr> on 2020-11-05 13:50:35 +0100)
   * Upstream changes:     - version 0.17.1
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 05 Nov 2020 12:56:53 +0000
 
 swh-storage (0.17.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.17.0     - (tagged by Nicolas Dandrimont
     <nicolas@dandrimont.eu> on 2020-11-03 18:09:53 +0100)
   * Upstream changes:     - Release swh.storage v0.17.0     - Migrate
     all raw extrinsic metadata attributes from id to target     - Add an
     `algos` function to resolve branch aliases     - Prepare updates to
     make swh.journal more generic     - Improve api server
     initialization     - Various updates to the
     migrate_extrinsic_metadata script, notably writing     - most
     metadata on directories instead of revisions
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 03 Nov 2020 17:20:45 +0000
 
 swh-storage (0.16.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.16.0     - (tagged by Nicolas Dandrimont
     <nicolas@dandrimont.eu> on 2020-10-09 18:23:24 +0200)
   * Upstream changes:     - Release swh.storage v0.16.0     - Updates to
     the intrinsic metadata migration script     - Various improvements
     to the buffer storage     - Update swh storage backfill to use
     common configuration keys
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 09 Oct 2020 16:33:11 +0000
 
 swh-storage (0.15.3-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.15.3     - (tagged by Nicolas Dandrimont
     <nicolas@dandrimont.eu> on 2020-09-24 20:14:39 +0200)
   * Upstream changes:     - Release swh.storage v0.15.3     - hopefully
     fix the documentation build
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 24 Sep 2020 18:24:14 +0000
 
 swh-storage (0.15.2-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.15.2     - (tagged by Nicolas Dandrimont
     <nicolas@dandrimont.eu> on 2020-09-24 19:22:11 +0200)
   * Upstream changes:     - Release swh.storage v0.15.2     - no change
     rebuild to clean up jenkins fsckup accumulating old files.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 24 Sep 2020 17:28:22 +0000
 
 swh-storage (0.15.1-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.15.1     - (tagged by Nicolas Dandrimont
     <nicolas@dandrimont.eu> on 2020-09-24 18:34:54 +0200)
   * Upstream changes:     - Release swh.storage v0.15.1     - Restore
     buffer proxy behavior with default arguments
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 24 Sep 2020 16:44:22 +0000
 
 swh-storage (0.15.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.15.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-09-24 16:54:07
     +0200)
   * Upstream changes:     - v0.15.0     - Support different database
     flavors in the SQL scripts     - Add the SQL commands used to set up
     the logical replication publication     - Output a warning when the
     version of the database is different than expected     - Improve
     code quality and doc in BufferedProxyStorage     - Adapt cli
     declaration entrypoint to swh.core 0.3     - Add warning about
     skipped_content (sneaking into the 'content' topics)     - graph-
     replayer: fix to prevent wrong warning     - pre-commit: Add isort
     hook and reorder imports with isort     - pytest_plugin: Change
     dbname to storage to avoid clash in tests     - pytest_plugin: Use
     psql to load SQL files instead of connecting with psycopg2
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 24 Sep 2020 15:03:58 +0000
 
 swh-storage (0.14.3-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.14.3     - (tagged by David Douard
     <david.douard@sdfa3.org> on 2020-09-17 16:58:59 +0200)
   * Upstream changes:     - v0.14.3
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 17 Sep 2020 16:53:56 +0000
 
 swh-storage (0.14.2-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.14.2     - (tagged by David Douard
     <david.douard@sdfa3.org> on 2020-09-11 15:31:22 +0200)
   * Upstream changes:     - v0.14.2
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 11 Sep 2020 13:37:11 +0000
 
 swh-storage (0.14.1-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.14.1     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-09-04 15:43:51
     +0200)
   * Upstream changes:     - v0.14.1     - algos.diff: Add missed
     revision_get conversion
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 04 Sep 2020 13:52:17 +0000
 
 swh-storage (0.14.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.14.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-09-04 12:23:52
     +0200)
   * Upstream changes:     - v0.14.0     - Refactor revision_get storage
     API to return Revision objects     - cassandra: Discard Content
     ctime field in content_get_partition
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 04 Sep 2020 10:59:54 +0000
 
 swh-storage (0.13.3-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.13.3     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-09-01 14:34:57
     +0200)
   * Upstream changes:     - v0.13.3     - storage*: release_get(...) ->
     List[Optional[Release]]     - Make StorageInterface a Protocol.     -
     Add a validating storage proxy, to check ids before insertion.     -
     Add a --check-config option for cli commands     - Remove the
     deprecated config-path option from `swh storage rpc-serve` command
     - Add support for a new "check_config" config option in
     get_storage()     - Check for db version mismatch in
     PgStorage.check_config()     - Add a check_dbversion() method to the
     Db class     - Fix pytest_plugin's database janitor: do not truncate
     the dbversion table     - algos.snapshot: Add
     visits_and_snapshots_get_from_revision     - storage/interface:
     Remove deprecated diff endpoints     - storage_tests: Remove
     duplicated postgresql-specific tests.     - Move postgresql-related
     files to swh/storage/postgresql/
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 01 Sep 2020 12:40:29 +0000
 
 swh-storage (0.13.2-1~swh2) unstable-swh; urgency=medium
 
   * Add mypy-extensions to build-dependencies
 
  -- Nicolas Dandrimont <olasd@debian.org>  Fri, 21 Aug 2020 12:17:05 +0200
 
 swh-storage (0.13.2-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.13.2     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2020-08-20 08:59:39 +0200)
   * Upstream changes:     - v0.13.2     - * pg: Fix crash in
     snapshot_get when the snapshot does not exist.     - * cassandra:
     fix signatures     - * in_memory: rewrite as a backend for the
     cassandra storage     - * remove endpoint
     snapshot_get_by_origin_visit.     - * pg: rewrite converters to work
     with model objects
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 20 Aug 2020 07:18:50 +0000
 
 swh-storage (0.13.1-1~swh3) unstable-swh; urgency=medium
 
   * Update dependencies
 
  -- Antoine R. Dumont (@ardumont) <ardumont@softwareheritage.org>  Fri, 07 Aug 2020 21:17:01 +0000
 
 swh-storage (0.13.1-1~swh2) unstable-swh; urgency=medium
 
   * Update dependencies
 
  -- Antoine R. Dumont (@ardumont) <ardumont@softwareheritage.org>  Fri, 07 Aug 2020 21:02:01 +0000
 
 swh-storage (0.13.1-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.13.1     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2020-08-07 18:14:32 +0200)
   * Upstream changes:     - v0.13.1     - * Make snapshot_get_branches
     return a TypedDict containing SnapshotBranch objects.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 07 Aug 2020 16:23:01 +0000
 
 swh-storage (0.13.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.13.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-08-07 12:38:47
     +0200)
   * Upstream changes:     - v0.13.0     - storage*: Rename and type
     content_get(List[Sha1]) -> List[Optional[Content]]     - storage*:
     Rename content_get_data(Sha1) -> Optional[bytes]     - Simplify as
     Content.ctime None is popped out of a to_dict call in recent model
     - cassandra.storage: Use next token for pagination instead of
     computing it
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 07 Aug 2020 10:49:28 +0000
 
 swh-storage (0.12.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.12.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-08-06 08:50:17
     +0200)
   * Upstream changes:     - v0.12.0     - Type storage endpoints     -
     Drop content_get_range endpoint in favor of content_get_partition
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 06 Aug 2020 06:55:26 +0000
 
 swh-storage (0.11.10-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.11.10     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-08-04 14:10:21
     +0200)
   * Upstream changes:     - v0.11.10     - tests: Improve coverage on
     directory_ls endpoints     - storage*: Type content_find(...) ->
     List[Content]     - storage*: Type
     {cnt,dir,rev,rel,snp}_get_random(...) -> Sha1Git
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 04 Aug 2020 12:15:21 +0000
 
 swh-storage (0.11.9-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.11.9     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-08-03 11:55:10
     +0200)
   * Upstream changes:     - v0.11.9     - storage*: Drop origin-get-
     range in favor of origin-list     - storage*: Do not allow unknown
     visit status in origin_visit*_get_latest     - storage*: Add type
     annotation to origin_count     - Reuse swh.core stream_results
     function
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 03 Aug 2020 10:02:56 +0000
 
 swh-storage (0.11.8-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.11.8     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2020-07-31 14:57:09 +0200)
   * Upstream changes:     - v0.11.8     - * test_replay: update for
     swh.journal 0.4.1.     - * Add support for metadata-related object
     types to the backfiller and replayer.     - * pg: Rewrite
     _origin_query to force the query planner to filter on URLs before
     filtering on visits.     - * Make raw_extrinsic_metadata_get return
     PagedResult instead of Dict.     - * Rename argument 'object_type'
     of raw_extrinsic_metadata_get to 'type'.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 31 Jul 2020 13:17:40 +0000
 
 swh-storage (0.11.6-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.11.6     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-07-30 16:20:48
     +0200)
   * Upstream changes:     - v0.11.6     - storage*: Adapt
     origin_list(...) -> PagedResult[Origin]     - algos.snapshot: Open
     snapshot_id_get_from_revision     - storage*: add
     origin_visit_status_get(...) -> PagedResult[OriginVisitStatus]     -
     Add type annotations on get_storage.     - buffer: Pass lists to
     backend functions, not iterables.     - storage*: Simplify next-page-
     token computation     - filter: Fix types passed to the proxied
     storage.     - Fix upcoming type warning with swh.core > v0.1.2.
     - Make API endpoints take Lists instead of Iterables as arguments
     - storage*: use an enum to explicit the order in origin_visit_get
     - storage*: origin_visit_get(...) -> PagedResult[OriginVisit]     -
     Write metadata + metadata authorities/fetchers to the journal.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 30 Jul 2020 14:29:10 +0000
 
 swh-storage (0.11.5-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.11.5     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2020-07-28 09:55:34 +0200)
   * Upstream changes:     - v0.11.5     - in_memory: fix tie-breaking
     when two visits have the same date.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 28 Jul 2020 08:10:21 +0000
 
 swh-storage (0.11.4-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.11.4     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-07-27 16:08:42
     +0200)
   * Upstream changes:     - v0.11.4     - Rename object_metadata to
     raw_extrinsic_metadata     - metadata_{authority,fetcher}_add: Fix
     crash when the iterable argument is empty     - storage*:
     origin_visit_get_by -> Optional[OriginVisit]     - storage*:
     origin_visit_find_by_date -> Optional[OriginVisit]     - storage*:
     type origin_visit_get_latest endpoint result     - algos.origin:
     Simplify origin_get_latest_visit_status function
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 27 Jul 2020 14:16:18 +0000
 
 swh-storage (0.11.3-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.11.3     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-07-27 08:01:03
     +0200)
   * Upstream changes:     - v0.11.3     - storage*:
     origin_get(Iterable[str]) -> Iterable[Optional[Origin]]     -
     storage*.origin_visit_get_random: Read model objects
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 27 Jul 2020 06:08:55 +0000
 
 swh-storage (0.11.2-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.11.2     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-07-23 12:09:51
     +0200)
   * Upstream changes:     - v0.11.2     - pgstorage: Drop unnecessary
     indirection from reading origin_visit     - pytest-plugin: Make
     sample_data return data model objects     - tests: Use only model
     objects for testing     - Drop validate storage proxy
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 23 Jul 2020 10:18:15 +0000
 
 swh-storage (0.11.1-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.11.1     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2020-07-20 13:01:20 +0200)
   * Upstream changes:     - v0.11.1     - * Use model objects in tests
     - * Rename 'deposit' authority type to 'deposit_client'.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 20 Jul 2020 11:14:39 +0000
 
 swh-storage (0.11.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.11.0     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2020-07-20 11:01:10 +0200)
   * Upstream changes:     - v0.11.0     - * Make metadata-related
     endpoints consistent with other endpoints by using Iterables of swh-
     model objects instead of a dict.     - * Update tests to use model
     objects
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 20 Jul 2020 09:12:25 +0000
 
 swh-storage (0.10.6-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.10.6     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-07-16 15:31:19
     +0200)
   * Upstream changes:     - v0.10.6     - pytest_plugin: Ensure fixture
     instantiates correctly
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 16 Jul 2020 13:36:34 +0000
 
 swh-storage (0.10.5-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.10.5     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-07-16 14:24:50
     +0200)
   * Upstream changes:     - v0.10.5     - pytest_plugin: Do not expose
     the validate proxy storage     - pytest-plugin: Expose a
     sample_data_model fixture     - tests: Start using model objects and
     drop validate proxy when possible
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 16 Jul 2020 12:34:44 +0000
 
 swh-storage (0.10.4-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.10.4     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-07-16 11:25:25
     +0200)
   * Upstream changes:     - v0.10.4     - pytest_plugin: Avoid fixture
     client to declare optional dependency     - Allow cassandra binary
     path to be configured through env variable     - 158: Make schema
     and migration converge so the migration works
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 16 Jul 2020 09:37:24 +0000
 
 swh-storage (0.10.3-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.10.3     - (tagged by Antoine Lambert
     <antoine.lambert@inria.fr> on 2020-07-10 16:26:27 +0200)
   * Upstream changes:     - version 0.10.3
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 10 Jul 2020 14:40:28 +0000
 
 swh-storage (0.10.2-1~swh2) unstable-swh; urgency=medium
 
   * Fix debian rules to avoid double pytest-plugin loading clash
 
  -- Antoine R. Dumont (@ardumont) <ardumont@softwareheritage.org>  Fri, 10 Jul 2020 09:21:14 +0200
 
 swh-storage (0.10.2-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.10.2     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-07-10 08:30:37
     +0200)
   * Upstream changes:     - v0.10.2     - tests: Do no expose the pytest-
     plugin through setuptools entry     - Convert ImmutableDict to dict
     before passing it to json.dumps     - docs: Rework dia -> pdf
     pipeline for inkscape 1.0
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 10 Jul 2020 06:52:42 +0000
 
 swh-storage (0.10.1-1~swh2) unstable-swh; urgency=medium
 
   * Update runtime dependencies
 
  -- Antoine R. Dumont (@ardumont) <ardumont@softwareheritage.org>  Wed, 08 Jul 2020 14:56:01 +0200
 
 swh-storage (0.10.1-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.10.1     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-07-08 14:32:52
     +0200)
   * Upstream changes:     - v0.10.1     - extract-pytest-fixture Move
     shareable fixtures out of conftest into a dedicated pytest plugin
     - Migrate from vcversioner to setuptools-scm
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 08 Jul 2020 12:39:15 +0000
 
 swh-storage (0.10.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.10.0     - (tagged by David Douard
     <david.douard@sdfa3.org> on 2020-07-08 09:20:49 +0200)
   * Upstream changes:     - v0.10.0
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 08 Jul 2020 10:11:09 +0000
 
 swh-storage (0.9.3-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.9.3     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-07-06 09:55:56
     +0200)
   * Upstream changes:     - v0.9.3     - storage: Send metrics from the
     origin_add endpoint
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 06 Jul 2020 08:06:13 +0000
 
 swh-storage (0.9.2-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.9.2     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-07-03 18:48:39
     +0200)
   * Upstream changes:     - v0.9.2     - pg-storage: Add missing cur
     parameter passing
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 03 Jul 2020 16:54:13 +0000
 
 swh-storage (0.9.1-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.9.1     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-07-03 16:50:45
     +0200)
   * Upstream changes:     - v0.9.1     - storage.db: Drop
     db.origin_visit_upsert behavior
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 03 Jul 2020 15:00:32 +0000
 
 swh-storage (0.9.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.9.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-07-01 09:53:34
     +0200)
   * Upstream changes:     - v0.9.0     - storage*: Drop intermediary
     conversion step into OriginVisit     - pg: use 'on conflict do
     nothing' strategy for duplicate metadata rows.     - Make the code
     location of metadata endpoints consistent across backends.     - Add
     content_metadata_{add,get}.     - Add context columns to
     object_metadata table and object_metadata_{add,get}.     -
     Generalize origin_metadata to allow support for other object types
     in the future.     - Work around the segmentation faults caused by
     pytest-coverage + multiprocessing.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 01 Jul 2020 08:02:08 +0000
 
 swh-storage (0.8.1-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.8.1     - (tagged by David Douard
     <david.douard@sdfa3.org> on 2020-06-30 10:08:21 +0200)
   * Upstream changes:     - v0.8.1
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 30 Jun 2020 08:36:45 +0000
 
 swh-storage (0.8.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.8.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-06-29 09:33:12
     +0200)
   * Upstream changes:     - v0.8.0     - Iterate over paginated visits
     in batches to retrieve latest visit/snapshot     - storage*: Open
     order parameter to origin-visit-get endpoint     -
     tests/replayer/storage*: Drop obsolete origin visit fields     -
     Relax checks on journal writes regarding origin-visit*     -
     replayer: Fix isoformat datetime string for origin-visit     -
     Deprecate the origin_add_one() endpoint     - test_storage: Add
     missing tests on origin_visit_get method
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 29 Jun 2020 07:44:00 +0000
 
 swh-storage (0.7.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.7.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-06-22 15:42:25
     +0200)
   * Upstream changes:     - v0.7.0     - test_origin: Rename
     appropriately tests     - algos: Improve origin visit get latest
     visit status algorithm     - test_snapshot: Do not use
     origin_visit_add returned result     - algos.snapshot: Fix edge case
     when snapshot is not resolved     - Ensure ids are correct in tests'
     storage_data     - Fix tests' storage_data revisions     - SQL:
     replace the hash(url) index by a unique btree(url) on the origin
     table     - Make sure the pagination in swh_snapshot_get_by_id uses
     the proper indexes
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 22 Jun 2020 14:09:33 +0000
 
 swh-storage (0.6.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.6.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-06-19 11:29:42
     +0200)
   * Upstream changes:     - v0.6.0     - Move deprecated endpoint
     snapshot_get_latest from api endpoint to algos     - algos.origin:
     Open origin-get-latest-visit-status function     - storage*: Allow
     origin-visit-get-latest to filter on type     - test_origin: Align
     storage initialization within tests
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 19 Jun 2020 12:45:32 +0000
 
 swh-storage (0.5.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.5.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-06-17 16:03:15
     +0200)
   * Upstream changes:     - v0.5.0     - test_storage: Fix flakiness in
     round to milliseconds test util method     - storage*: Add origin-
     visit-status-get-latest endpoint     - Fix/update the backfiller
     - validate: accept model objects as well as dicts on all add
     endpoints     - cql: Fix blackified strings     - storage: Add
     missing cur parameter     - Fix db_to_author() converter to return
     None is all fields are None
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 17 Jun 2020 14:19:37 +0000
 
 swh-storage (0.4.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.4.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-06-16 09:50:25
     +0200)
   * Upstream changes:     - v0.4.0     - ardumont/master storage*: Drop
     leftover code     - storage*: Drop origin_visit_upsert endpoint     -
     storage*: Remove origin-visit-update endpoint     - replay: Replay
     origin-visit and origin-visit-status     - in_memory: Make origin-
     visit-status-add respect "on conflict ignore" policy     -
     test_storage: Add journal behavior coverage for origin-visit-*add
     - Start migrating the validate proxy toward using BaseModel objects
     - storage*: Do not write twice origin-visit-status in journal
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 16 Jun 2020 07:58:23 +0000
 
 swh-storage (0.3.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.3.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-06-12 09:08:23
     +0200)
   * Upstream changes:     - v0.3.0     - origin-visit-add storage*:
     Align origin-visit-add to take iterable of OriginVisit objects
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 12 Jun 2020 07:22:03 +0000
 
 swh-storage (0.2.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.2.0     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-06-10 11:51:30
     +0200)
   * Upstream changes:     - v0.2.0     - origin-visit-upsert: Write
     visit status objects to the journal     - origin-visit-update: Write
     visit status objects to the journal     - origin-visit-add: Write
     visit status to the journal     - Add pagination to
     origin_metadata_get.     - Deduplicate origin-metadata when they
     have the same authority + discovery_date + fetcher.     - Open
     `origin_visit_status_add` endpoint to add origin visit statuses     -
     Add a replayer test for anonymized journal topics     - Small
     refactoring of the InMemoryStorage to make it more consistent
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 10 Jun 2020 10:02:45 +0000
 
 swh-storage (0.1.1-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.1.1     - (tagged by Nicolas Dandrimont
     <nicolas@dandrimont.eu> on 2020-06-04 16:49:22 +0200)
   * Upstream changes:     - Release swh.storage v0.1.1     - Work around
     tests hanging during Debian build
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 04 Jun 2020 14:56:54 +0000
 
 swh-storage (0.1.0-2~swh1) unstable-swh; urgency=medium
 
   * Update dependencies.
 
  -- David Douard <david.douard@sdfa3.org>  Thu, 04 Jun 2020 13:40:52 +0200
 
 swh-storage (0.1.0-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.1.0     - (tagged by David Douard
     <david.douard@sdfa3.org> on 2020-06-04 12:08:46 +0200)
   * Upstream changes:     - v0.1.0
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 04 Jun 2020 10:28:43 +0000
 
 swh-storage (0.0.193-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.193     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-05-28 14:28:54
     +0200)
   * Upstream changes:     - v0.0.193     - pg: Write origin visit
     updates & status, read from origin_visit_status     - Make
     content.blake2s256 not null.     - Remove unused SQL functions.     -
     README: Update necessary dependencies for test purposes     - Add a
     pre-commit hook to check there are version bumps in
     sql/upgrades/*.sql     - Add missing dbversion bump in 150.sql.     -
     Add artifact metadata to the extrinsic metadata storage
     specification.     - Add not null constraints to
     metadata_authority/origin_metadata     - Realign schema with latest
     149 migration script
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 28 May 2020 12:37:58 +0000
 
 swh-storage (0.0.192-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.192     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2020-05-19 18:42:00 +0200)
   * Upstream changes:     - v0.0.192     - * origin_metadata_add: Reject
     non-bytes types for 'metadata'.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 19 May 2020 16:54:00 +0000
 
 swh-storage (0.0.191-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.191     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2020-05-19 13:43:35 +0200)
   * Upstream changes:     - v0.0.191     - * Implement the new extrinsic
     metadata specification/vocabulary.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 19 May 2020 11:52:00 +0000
 
 swh-storage (0.0.190-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.190     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-05-18 14:10:39
     +0200)
   * Upstream changes:     - v0.0.190     - storage: metadata_provider:
     Ensure idempotency when creating provider     - journal: add a
     skipped_content topic dedicated to SkippedContent objects     - Add
     missing return annotations on JournalWriter methods     - Improve a
     bit the exception message of JournalWriter.content_update     -
     Refactor the JournalWriter class to normalize its methods     -
     tests: fix test_replay; do only use aware datetime objects     -
     test_kafka_writer: Add missing object type skipped_content
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 18 May 2020 12:18:09 +0000
 
 swh-storage (0.0.189-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.189     - (tagged by Antoine R. Dumont
     (@ardumont) <ardumont@softwareheritage.org> on 2020-04-30 14:50:54
     +0200)
   * Upstream changes:     - v0.0.189     - pg: Write both origin visit
     updates & status, read from origin_visit     - pg-storage: Add new
     created state     - setup.py: add documentation link     - metadata
     spec: Fix title hierarchy     - tests: Use aware datetimes instead
     of naive ones.     - cassandra: Adapt internal implementations to
     use origin visit status     - in_memory: Adapt internal
     implementations to use origin visit status
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 30 Apr 2020 12:58:57 +0000
 
 swh-storage (0.0.188-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.188     - (tagged by David Douard
     <david.douard@sdfa3.org> on 2020-04-28 13:44:20 +0200)
   * Upstream changes:     - v0.0.188
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 28 Apr 2020 11:52:08 +0000
 
 swh-storage (0.0.187-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.187     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2020-04-14 18:13:08
     +0200)
   * Upstream changes:     - v0.0.187     - storage.interface: Actually
     define the remote flush operation
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 14 Apr 2020 16:23:41 +0000
 
 swh-storage (0.0.186-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.186     - (tagged by Nicolas Dandrimont
     <nicolas@dandrimont.eu> on 2020-04-14 17:09:22 +0200)
   * Upstream changes:     - Release swh.storage v0.0.186     - Drop
     backwards-compatibility code with swh.journal < 0.0.30
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 14 Apr 2020 15:20:57 +0000
 
 swh-storage (0.0.185-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.185     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2020-04-14 14:15:32
     +0200)
   * Upstream changes:     - v0.0.185     - storage.filter: Remove
     internal state     - test: update storage tests to (future)
     swh.journal 0.0.30
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 14 Apr 2020 12:22:06 +0000
 
 swh-storage (0.0.184-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.184     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2020-04-10 16:07:32
     +0200)
   * Upstream changes:     - v0.0.184     - storage*: Add flush endpoints
     to storage implems (backend, proxy)     - test_retry: Add missing
     skipped_content_add tests
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 10 Apr 2020 14:14:20 +0000
 
 swh-storage (0.0.183-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.183     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2020-04-09 12:35:53
     +0200)
   * Upstream changes:     - v0.0.183     - proxy storage: Add a
     clear_buffers endpoint     - buffer proxy storage: Filter out
     duplicate objects prior to storage write     - storage: Prevent
     erroneous HashCollisions by using the same ctime for all rows.     -
     Enable black     - origin_visit_update: ensure it raises a
     StorageArgumentException     - Adapt cassandra backend to validating
     model types     - tests: many refactoring improvements     - tests:
     Shut down cassandra connection before closing the fixture down     -
     Add more type annotations
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 09 Apr 2020 10:46:29 +0000
 
 swh-storage (0.0.182-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.182     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2020-03-27 07:02:13
     +0100)
   * Upstream changes:     - v0.0.182     - storage*: Update
     origin_visit_update to make status parameter mandatory     - test:
     Adapt origin validation test according to latest model changes     -
     Respec discovery_date as a Python datetime instead of an ISO string.
     - origin_visit_add: Add missing db/cur argument to call to
     origin_get.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 27 Mar 2020 06:13:17 +0000
 
 swh-storage (0.0.181-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.181     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2020-03-25 09:50:49
     +0100)
   * Upstream changes:     - v0.0.181     - storage*: Hex encode content
     hashes in HashCollision exception     - Add format of discovery_date
     in the metadata specification.     - Store the value of
     token(partition_key) in skipped_content_by_* table, instead of three
     hashes.     - Store the value of token(partition_key) in
     content_by_* table, instead of three hashes.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 25 Mar 2020 09:03:43 +0000
 
 swh-storage (0.0.180-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.180     - (tagged by Nicolas Dandrimont
     <nicolas@dandrimont.eu> on 2020-03-18 18:24:41 +0100)
   * Upstream changes:     - Release swh.storage v0.0.180     - Stop
     counting origin additions multiple times in statsd
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 18 Mar 2020 17:45:36 +0000
 
 swh-storage (0.0.179-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.179     - (tagged by Nicolas Dandrimont
     <nicolas@dandrimont.eu> on 2020-03-18 16:05:13 +0100)
   * Upstream changes:     - Release swh.storage v0.0.179.     - fix
     requirements-swh.txt to use proper version restriction     - reduce
     the transaction load for content writes and reads
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 18 Mar 2020 15:50:50 +0000
 
 swh-storage (0.0.178-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.178     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2020-03-16 12:51:28
     +0100)
   * Upstream changes:     - v0.0.178     - origin_visit_add: Adapt
     endpoint signature to return OriginVisit     - origin_visit_upsert:
     Use OriginVisit object as input     - storage/writer: refactor
     JournalWriter.content_add to send model objects
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 16 Mar 2020 11:59:18 +0000
 
 swh-storage (0.0.177-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.177     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2020-03-10 11:37:33
     +0100)
   * Upstream changes:     - v0.0.177     - storage: Identify and provide
     the collision hashes in exception     - Guarantee the order of
     results for revision_get and release_get     - tests: Improve test
     speed     - sql: do not attempt to create the plpgsql lang if
     already exists     - Update requirement on swh.core for RPCClient
     method overrides
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 10 Mar 2020 10:48:11 +0000
 
 swh-storage (0.0.176-1~swh2) unstable-swh; urgency=medium
 
   * Update build dependencies
 
  -- Antoine R. Dumont (@ardumont) <ardumont@softwareheritage.org>  Mon, 02 Mar 2020 14:36:00 +0100
 
 swh-storage (0.0.176-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.176     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2020-02-28 14:44:10 +0100)
   * Upstream changes:     - v0.0.176     - * Accept cassandra-driver >=
     3.22.     - * Make the RPC client and objstorage helper fetch
     Content.data from lazy     - contents.     - * Move ctime out of the
     validation proxy.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 28 Feb 2020 15:21:27 +0000
 
 swh-storage (0.0.175-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.175     - (tagged by Antoine Lambert
     <antoine.lambert@inria.fr> on 2020-02-20 13:51:40 +0100)
   * Upstream changes:     - version 0.0.175
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 20 Feb 2020 13:18:34 +0000
 
 swh-storage (0.0.174-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.174     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2020-02-19 14:18:59 +0100)
   * Upstream changes:     - v0.0.174     - * Fix inconsistent behavior
     of skipped_content_missing across backends.     - * Fix
     FilteringProxy to not drop skipped-contents with a missing sha1_git.
     - * Make storage proxies use swh-model objects instead of dicts.
     - * Add support for (de)serializing swh-model in RPC calls.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 19 Feb 2020 15:00:32 +0000
 
 swh-storage (0.0.172-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.172     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2020-02-12 14:00:04 +0100)
   * Upstream changes:     - v0.0.172     - * Unify exception raised by
     invalid input to API endpoints.     - * Add a validation proxy for
     _add() methods. This proxy is *required*     - in front of all
     backends whose _add() methods may be called or they'll     - crash
     at runtime.     - * Fix RecursionError when storage proxies are
     deepcopied or unpickled.     - * storages: Refactor objstorage
     operations with a dedicated collaborator     - * storages: Refactor
     journal operations with a dedicated writer collab
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 12 Feb 2020 13:13:47 +0000
 
 swh-storage (0.0.171-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.171     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2020-02-06 14:46:05 +0100)
   * Upstream changes:     - v0.0.171     - * Split 'content_add' method
     into 'content_add' and 'skipped_content_add'.     - * Increase
     Cassandra requests timeout to 1 second.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 06 Feb 2020 14:07:37 +0000
 
 swh-storage (0.0.170-1~swh3) unstable-swh; urgency=medium
 
   * Update build dependencies
 
  -- Antoine R. Dumont (@ardumont) <ardumont@softwareheritage.org>  Mon, 03 Feb 2020 17:30:38 +0100
 
 swh-storage (0.0.170-1~swh2) unstable-swh; urgency=medium
 
   * Update build dependencies
 
  -- Antoine R. Dumont (@ardumont) <ardumont@softwareheritage.org>  Mon, 03 Feb 2020 16:00:39 +0100
 
 swh-storage (0.0.170-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.170     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2020-02-03 14:11:53
     +0100)
   * Upstream changes:     - v0.0.170     - swh.storage.cassandra: Add
     Cassandra backend implementation
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 03 Feb 2020 13:23:48 +0000
 
 swh-storage (0.0.169-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.169     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2020-01-30 13:40:00
     +0100)
   * Upstream changes:     - v0.0.169     - retry: Add retry behavior on
     pipeline storage with flushing failure
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 30 Jan 2020 13:26:23 +0000
 
 swh-storage (0.0.168-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.168     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2020-01-30 11:19:31 +0100)
   * Upstream changes:     - v0.0.168     - * Implement content_update
     for the in-mem storage.     - * Remove cur/db arguments from the in-
     mem storage.     - * Move Storage documentation and endpoint paths
     to a new StorageInterface class     - * Rename in_memory.Storage to
     in_memory.InMemoryStorage.     - * CONTRIBUTORS: add Daniele
     Serafini
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 30 Jan 2020 10:25:30 +0000
 
 swh-storage (0.0.167-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.167     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2020-01-24 14:55:57
     +0100)
   * Upstream changes:     - v0.0.167     - pgstorage: Empty temp tables
     instead of dropping them
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 24 Jan 2020 14:01:57 +0000
 
 swh-storage (0.0.166-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.166     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2020-01-24 09:51:52
     +0100)
   * Upstream changes:     - v0.0.166     - storage: Add endpoint to get
     missing content (by sha1_git) and missing snapshot     - Remove
     redundant config checks in load_and_check_config     - Remove 'id'
     and 'object_id' from the output of object_find_by_sha1_git     -
     Make origin_visit_get_random return None instead of {} if there are
     no results     - docs: Fix sphinx warnings
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 24 Jan 2020 09:00:12 +0000
 
 swh-storage (0.0.165-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.165     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2020-01-17 14:04:53
     +0100)
   * Upstream changes:     - v0.0.165     - storage.retry: Fix objects
     loading when using generator parameters
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 17 Jan 2020 13:09:39 +0000
 
 swh-storage (0.0.164-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.164     - (tagged by Antoine Lambert
     <antoine.lambert@inria.fr> on 2020-01-16 17:54:40 +0100)
   * Upstream changes:     - version 0.0.164
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 16 Jan 2020 17:05:02 +0000
 
 swh-storage (0.0.163-1~swh2) unstable-swh; urgency=medium
 
   * Fix test dependency
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Tue, 14 Jan 2020 17:26:08 +0100
 
 swh-storage (0.0.163-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.163     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2020-01-14 17:12:03
     +0100)
   * Upstream changes:     - v0.0.163     - retry: Improve proxy storage
     for add endpoints     - in_memory: Make directory_get_random return
     None when storage empty     - storage: Change content_get_metadata
     api to return Dict[bytes, List[Dict]]     - storage: Add
     content_get_partition endpoint to replace content_get_range     -
     storage: Add endpoint origin_list to replace origin_get_range
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 14 Jan 2020 16:17:45 +0000
 
 swh-storage (0.0.162-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.162     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2019-12-16 14:37:44 +0100)
   * Upstream changes:     - v0.0.162     - Add
     {content,directory,revision,release,snapshot}_get_random.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 16 Dec 2019 13:41:39 +0000
 
 swh-storage (0.0.161-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.161     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2019-12-10 15:03:28
     +0100)
   * Upstream changes:     - v0.0.161     - storage: Add endpoint to
     randomly pick an origin
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 10 Dec 2019 14:08:15 +0000
 
 swh-storage (0.0.160-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.160     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2019-12-06 11:15:48
     +0100)
   * Upstream changes:     - v0.0.160     - storage.buffer: Buffer
     release objects as well     - storage.tests: Unify tests sample data
     - Implement origin lookup by sha1
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 06 Dec 2019 10:23:44 +0000
 
 swh-storage (0.0.159-1~swh2) unstable-swh; urgency=medium
 
   * Force fast hypothesis profile when running tests
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Tue, 26 Nov 2019 17:08:16 +0100
 
 swh-storage (0.0.159-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.159     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2019-11-22 11:05:41
     +0100)
   * Upstream changes:     - v0.0.159     - Add 'pipeline' storage
     "class" for more readable configurations.     - tests: Improve tests
     environments configuration     - Fix a few typos reported by
     codespell     - Add a pre-commit-hooks.yaml config file     - Remove
     utils/(dump|fix)_revisions scripts
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 22 Nov 2019 10:10:31 +0000
 
 swh-storage (0.0.158-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.158     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2019-11-14 13:33:00
     +0100)
   * Upstream changes:     - v0.0.158     - Drop schemata module
     (migrated back to swh-lister)
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 14 Nov 2019 12:37:18 +0000
 
 swh-storage (0.0.157-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.157     - (tagged by Nicolas Dandrimont
     <nicolas@dandrimont.eu> on 2019-11-13 13:22:39 +0100)
   * Upstream changes:     - Release swh.storage 0.0.157     -
     schemata.distribution: Fix bogus NotImplementedError on
     Area.index_uris
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 13 Nov 2019 12:27:07 +0000
 
 swh-storage (0.0.156-1~swh2) unstable-swh; urgency=medium
 
   * Add version constraint on psycopg2
 
  -- Nicolas Dandrimont <olasd@debian.org>  Wed, 30 Oct 2019 18:21:34 +0100
 
 swh-storage (0.0.156-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.156     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2019-10-30 15:12:10 +0100)
   * Upstream changes:     - v0.0.156     - * Stop supporting origin ids
     in API (except in origin_get_range).     - * Make visit['origin'] a
     string everywhere (instead of a dict).
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 30 Oct 2019 14:29:28 +0000
 
 swh-storage (0.0.155-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.155     - (tagged by David Douard
     <david.douard@sdfa3.org> on 2019-10-30 12:14:14 +0100)
   * Upstream changes:     - v0.0.155
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 30 Oct 2019 11:18:37 +0000
 
 swh-storage (0.0.154-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.154     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2019-10-17 13:47:57
     +0200)
   * Upstream changes:     - v0.0.154     - Fix tests in debian build
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 17 Oct 2019 11:52:46 +0000
 
 swh-storage (0.0.153-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.153     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2019-10-17 13:21:00
     +0200)
   * Upstream changes:     - v0.0.153     - Deploy new test fixture
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 17 Oct 2019 11:26:12 +0000
 
 swh-storage (0.0.152-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.152     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2019-10-08 16:55:43
     +0200)
   * Upstream changes:     - v0.0.152     - swh.storage.buffer: Add
     buffering proxy storage implementation     - swh.storage.filter: Add
     filtering storage implementation     - swh.storage.tests: Improve db
     transaction handling     - swh.storage.tests: Add more tests     -
     swh.storage.storage: introduce a db() context manager
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 08 Oct 2019 15:03:16 +0000
 
 swh-storage (0.0.151-1~swh2) unstable-swh; urgency=medium
 
   * Add missing build-dependency on python3-swh.journal
 
  -- Nicolas Dandrimont <olasd@debian.org>  Tue, 01 Oct 2019 18:28:19 +0200
 
 swh-storage (0.0.151-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.151     - (tagged by Stefano Zacchiroli
     <zack@upsilon.cc> on 2019-10-01 10:04:36 +0200)
   * Upstream changes:     - v0.0.151     - * tox: anticipate mypy run to
     just after flake8     - * mypy.ini: be less flaky w.r.t. the
     packages installed in tox     - * storage.py: ignore typing of
     optional get_journal_writer import     - * mypy: ignore swh.journal
     to work-around dependency loop     - * init.py: switch to documented
     way of extending path     - * typing: minimal changes to make a no-
     op mypy run pass     - * Write objects to the journal only if they
     don't exist yet.     - * Use origin URLs for
     skipped_content['origin'] instead of origin ids.     - * Properly
     mock get_journal_writer for the remote-pg-storage tests.     - *
     journal_writer: use journal writer from swh.journal     - * fix
     typos in docstrings and sample paths     - *
     storage.origin_visit_add: Remove deprecated 'ts' parameter     - *
     click "required" param wants bool, not int
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 01 Oct 2019 08:09:53 +0000
 
 swh-storage (0.0.150-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.150     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2019-09-04 16:09:59
     +0200)
   * Upstream changes:     - v0.0.150     - tests/test_storage: Remove
     failing assertion after swh-model update     - tests/test_storage:
     Fix tests execution with psycopg2 < 2.8
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 04 Sep 2019 14:16:09 +0000
 
 swh-storage (0.0.149-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.149     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2019-09-03 14:00:57
     +0200)
   * Upstream changes:     - v0.0.149     - Add support for origin_url in
     origin_metadata_*     - Make origin_add/origin_visit_update validate
     their input     - Make snapshot_add validate its input     - Make
     revision_add and release_add validate their input     - Make
     directory_add validate its input     - Make content_add validate its
     input using swh-model
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 03 Sep 2019 12:27:51 +0000
 
 swh-storage (0.0.148-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.148     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2019-08-23 10:33:02 +0200)
   * Upstream changes:     - v0.0.148     - Tests improvements:     - *
     Remove 'next_branch' from test input data.     - * Fix off-by-one
     error when using origin_visit_upsert on with an unknown visit id.
     - * Use explicit arguments for origin_visit_add.     - * Remove
     test_content_missing__marked_missing, it makes no sense.     - Drop
     person ids:     - * Stop leaking person ids.     - * Remove
     person_get endpoint.     - Logging fixes:     - * Enforce log level
     for the werkzeug logger.     - * Eliminate warnings about %TYPE.
     - * api: use RPCServerApp and RPCClient instead of deprecated
     classes     - Other:     - * Add support for skipped content in in-
     memory storage
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 23 Aug 2019 08:48:21 +0000
 
 swh-storage (0.0.147-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.147     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2019-07-18 12:11:37 +0200)
   * Upstream changes:     - Make origin_get ignore the `type` argument
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 18 Jul 2019 10:16:16 +0000
 
 swh-storage (0.0.146-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.146     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2019-07-18 10:46:21 +0200)
   * Upstream changes:     - Progress toward getting rid of origin ids
     - * Less dependency on origin ids in the in-mem storage     - * add
     the SWH_STORAGE_IN_MEMORY_ENABLE_ORIGIN_IDS env var     - * Remove
     legacy behavior of snapshot_add
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 18 Jul 2019 08:52:09 +0000
 
 swh-storage (0.0.145-1~swh3) unstable-swh; urgency=medium
 
   * Properly rebuild for unstable-swh
 
  -- Nicolas Dandrimont <olasd@debian.org>  Thu, 11 Jul 2019 14:03:30 +0200
 
 swh-storage (0.0.145-1~swh2) buster-swh; urgency=medium
 
   * Remove useless swh.scheduler dependency
 
  -- Nicolas Dandrimont <olasd@debian.org>  Thu, 11 Jul 2019 13:53:45 +0200
 
 swh-storage (0.0.145-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.145     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2019-07-02 12:00:53 +0200)
   * Upstream changes:     - v0.0.145     - Add an
     'origin_visit_find_by_date' endpoint.     - Add support for origin
     urls in all endpoints
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 02 Jul 2019 10:19:19 +0000
 
 swh-storage (0.0.143-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.143     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2019-06-05 13:18:14 +0200)
   * Upstream changes:     - Add test for snapshot/release counters.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 01 Jul 2019 12:38:40 +0000
 
 swh-storage (0.0.142-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.142     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2019-06-11 15:24:49 +0200)
   * Upstream changes:     - Mark network tests, so they can be disabled.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 11 Jun 2019 13:44:19 +0000
 
 swh-storage (0.0.141-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.141     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2019-06-06 17:05:03 +0200)
   * Upstream changes:     - Add support for using URL instead of ID in
     snapshot_get_latest.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 11 Jun 2019 10:36:32 +0000
 
 swh-storage (0.0.140-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.140     - (tagged by mihir(faux__)
     <karbelkar.mihir@gmail.com> on 2019-03-24 21:47:31 +0530)
   * Upstream changes:     - Changes the output of content_find method to
     a list in case of hash collisions and makes the sql query on python
     side and added test duplicate input, colliding sha256 and colliding
     blake2s256
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 16 May 2019 12:09:04 +0000
 
 swh-storage (0.0.139-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.139     - (tagged by Nicolas Dandrimont
     <nicolas@dandrimont.eu> on 2019-04-18 17:57:57 +0200)
   * Upstream changes:     - Release swh.storage v0.0.139     - Backwards-
     compatibility improvements for snapshot_add     - Better
     transactionality in revision_add/release_add     - Fix backwards
     metric names     - Handle shallow histories properly
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 18 Apr 2019 16:08:28 +0000
 
 swh-storage (0.0.138-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.138     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2019-04-09 16:40:49 +0200)
   * Upstream changes:     - Use the db_transaction decorator on all
     _add() methods.     - So they gracefully release the connection on
     error instead     - of relying on reference-counting to call the
     Db's `__del__`     - (which does not happen in Hypothesis tests)
     because a ref     - to it is kept via the traceback object.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 09 Apr 2019 16:50:48 +0000
 
 swh-storage (0.0.137-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.137     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2019-04-08 15:40:24 +0200)
   * Upstream changes:     - Make test_origin_get_range run faster.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 08 Apr 2019 13:56:16 +0000
 
 swh-storage (0.0.135-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.135     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2019-04-04 20:42:32 +0200)
   * Upstream changes:     - Make content_add_metadata require a ctime
     argument.     - This makes Python set the ctime instead of pgsql.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 05 Apr 2019 14:43:28 +0000
 
 swh-storage (0.0.134-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.134     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2019-04-03 13:38:58 +0200)
   * Upstream changes:     - Don't leak origin ids to the journal.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 04 Apr 2019 10:16:09 +0000
 
 swh-storage (0.0.132-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.132     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2019-04-01 11:50:30 +0200)
   * Upstream changes:     - Use sha1 instead of bigint as FK from
     origin_visit to snapshot (part 1: add new column)
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 01 Apr 2019 13:30:48 +0000
 
 swh-storage (0.0.131-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.131     - (tagged by Nicolas Dandrimont
     <nicolas@dandrimont.eu> on 2019-03-28 17:24:44 +0100)
   * Upstream changes:     - Release swh.storage v0.0.131     - Add
     statsd metrics to storage RPC backend     - Clean up
     snapshot_add/origin_visit_update     - Uniformize RPC backend to use
     POSTs everywhere
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 28 Mar 2019 16:34:07 +0000
 
 swh-storage (0.0.130-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.130     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2019-02-26 10:50:44 +0100)
   * Upstream changes:     - Add an helper function to list all origins
     in the storage.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 13 Mar 2019 14:01:04 +0000
 
 swh-storage (0.0.129-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.129     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2019-02-27 10:42:29 +0100)
   * Upstream changes:     - Double the timeout of revision_get.     -
     Metadata indexers often hit the limit.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 01 Mar 2019 10:11:28 +0000
 
 swh-storage (0.0.128-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.128     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2019-02-21 14:59:22
     +0100)
   * Upstream changes:     - v0.0.128     - api.server: Fix wrong
     exception type     - storage.cli: Fix cli entry point name to the
     expected name (setup.py)
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 21 Feb 2019 14:07:23 +0000
 
 swh-storage (0.0.127-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.127     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2019-02-21 13:34:19
     +0100)
   * Upstream changes:     - v0.0.127     - api.wsgi: Open wsgi
     entrypoint and check config at startup time     - api.server: Make
     the api server load and check its configuration     -
     swh.storage.cli: Migrate the api server startup in swh.storage.cli
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 21 Feb 2019 12:59:48 +0000
 
 swh-storage (0.0.126-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.126     - (tagged by Valentin Lorentz
     <vlorentz@softwareheritage.org> on 2019-02-21 10:18:26 +0100)
   * Upstream changes:     - Double the timeout of snapshot_get_latest.
     - Metadata indexers often hit the limit.
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 21 Feb 2019 11:24:52 +0000
 
 swh-storage (0.0.125-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.125     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2019-02-14 10:13:31
     +0100)
   * Upstream changes:     - v0.0.125     - api/server: Do not read
     configuration at each request
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 14 Feb 2019 16:57:01 +0000
 
 swh-storage (0.0.124-1~swh3) unstable-swh; urgency=low
 
   * New upstream release, fixing the distribution this time
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 14 Feb 2019 17:51:29 +0100
 
 swh-storage (0.0.124-1~swh2) unstable; urgency=medium
 
   * New upstream release for dependency fix reasons
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 14 Feb 2019 09:27:55 +0100
 
 swh-storage (0.0.124-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.124     - (tagged by Antoine Lambert
     <antoine.lambert@inria.fr> on 2019-02-12 14:40:53 +0100)
   * Upstream changes:     - version 0.0.124
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Tue, 12 Feb 2019 13:46:08 +0000
 
 swh-storage (0.0.123-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.123     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2019-02-08 15:06:49
     +0100)
   * Upstream changes:     - v0.0.123     - Make Storage.origin_get
     support a list of origins, like other     - Storage.*_get methods.
     - Stop using _to_bytes functions.     - Use the BaseDb (and friends)
     from swh-core
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 08 Feb 2019 14:14:18 +0000
 
 swh-storage (0.0.122-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.122     - (tagged by Antoine Lambert
     <antoine.lambert@inria.fr> on 2019-01-28 11:57:27 +0100)
   * Upstream changes:     - version 0.0.122
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 28 Jan 2019 11:02:45 +0000
 
 swh-storage (0.0.121-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.121     - (tagged by Antoine Lambert
     <antoine.lambert@inria.fr> on 2019-01-28 11:31:48 +0100)
   * Upstream changes:     - version 0.0.121
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Mon, 28 Jan 2019 10:36:40 +0000
 
 swh-storage (0.0.120-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.120     - (tagged by Antoine Lambert
     <antoine.lambert@inria.fr> on 2019-01-17 12:04:27 +0100)
   * Upstream changes:     - version 0.0.120
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Thu, 17 Jan 2019 11:12:47 +0000
 
 swh-storage (0.0.119-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.119     - (tagged by Antoine R. Dumont
     (@ardumont) <antoine.romain.dumont@gmail.com> on 2019-01-11 11:57:13
     +0100)
   * Upstream changes:     - v0.0.119     - listener: Notify Kafka when
     an origin visit is updated
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Fri, 11 Jan 2019 11:02:07 +0000
 
 swh-storage (0.0.118-1~swh1) unstable-swh; urgency=medium
 
   * New upstream release 0.0.118     - (tagged by Antoine Lambert
     <antoine.lambert@inria.fr> on 2019-01-09 16:59:15 +0100)
   * Upstream changes:     - version 0.0.118
 
  -- Software Heritage autobuilder (on jenkins-debian1) <jenkins@jenkins-debian1.internal.softwareheritage.org>  Wed, 09 Jan 2019 18:51:34 +0000
 
 swh-storage (0.0.117-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.117
   * listener: Adapt decoding behavior depending on the object type
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 20 Dec 2018 14:48:44 +0100
 
 swh-storage (0.0.116-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.116
   * Update requirements to latest swh.core
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Fri, 14 Dec 2018 15:57:04 +0100
 
 swh-storage (0.0.115-1~swh1) unstable-swh; urgency=medium
 
   * version 0.0.115
 
  -- Antoine Lambert <antoine.lambert@inria.fr>  Fri, 14 Dec 2018 15:47:52 +0100
 
 swh-storage (0.0.114-1~swh1) unstable-swh; urgency=medium
 
   * version 0.0.114
 
  -- Antoine Lambert <antoine.lambert@inria.fr>  Wed, 05 Dec 2018 10:59:49 +0100
 
 swh-storage (0.0.113-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.113
   * in-memory storage: Add recursive argument to directory_ls endpoint
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Fri, 30 Nov 2018 11:56:44 +0100
 
 swh-storage (0.0.112-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.112
   * in-memory storage: Align with existing storage
   * docstring: Improvements and adapt according to api
   * doc: update index to match new swh-doc format
   * Increase test coverage for stat_counters + fix its bugs.
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Fri, 30 Nov 2018 10:28:02 +0100
 
 swh-storage (0.0.111-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.111
   * Move generative tests in their own module
   * Open in-memory storage implementation
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Wed, 21 Nov 2018 08:55:14 +0100
 
 swh-storage (0.0.110-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.110
   * storage: Open content_get_range endpoint
   * tests: Start using hypothesis for tests generation
   * Improvements: Remove SQLisms from the tests and API
   * docs: Document metadata providers
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Fri, 16 Nov 2018 11:53:14 +0100
 
 swh-storage (0.0.109-1~swh1) unstable-swh; urgency=medium
 
   * version 0.0.109
 
  -- Antoine Lambert <antoine.lambert@inria.fr>  Mon, 12 Nov 2018 14:11:09 +0100
 
 swh-storage (0.0.108-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.108
   * Add a function to get a full snapshot from the paginated view
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Thu, 18 Oct 2018 18:32:10 +0200
 
 swh-storage (0.0.107-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.107
   * Enable pagination of snapshot branches
   * Drop occurrence-related tables
   * Drop entity-related tables
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Wed, 17 Oct 2018 15:06:07 +0200
 
 swh-storage (0.0.106-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.106
   * Fix origin_visit_get_latest_snapshot logic
   * Improve directory iterator
   * Drop backwards compatibility between snapshots and occurrences
   * Drop the occurrence table
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Mon, 08 Oct 2018 17:03:54 +0200
 
 swh-storage (0.0.105-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.105
   * Increase directory_ls endpoint to 20 seconds
   * Add snapshot to the stats endpoint
   * Improve documentation
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Mon, 10 Sep 2018 11:36:27 +0200
 
 swh-storage (0.0.104-1~swh1) unstable-swh; urgency=medium
 
   * version 0.0.104
 
  -- Antoine Lambert <antoine.lambert@inria.fr>  Wed, 29 Aug 2018 15:55:37 +0200
 
 swh-storage (0.0.103-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.103
   * swh.storage.storage: origin_add returns updated list of dict with id
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Mon, 30 Jul 2018 11:47:53 +0200
 
 swh-storage (0.0.102-1~swh1) unstable-swh; urgency=medium
 
   * Release swh-storage v0.0.102
   * Stop using temporary tables for read-only queries
   * Add timeouts for some read-only queries
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Tue, 05 Jun 2018 14:06:54 +0200
 
 swh-storage (0.0.101-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.101
   * swh.storage.api.client: Permit to specify the query timeout option
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 24 May 2018 12:13:51 +0200
 
 swh-storage (0.0.100-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.100
   * remote api: only instantiate storage once per import
   * add thread-awareness to the storage implementation
   * properly cleanup after tests
   * parallelize objstorage and storage additions
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Sat, 12 May 2018 18:12:40 +0200
 
 swh-storage (0.0.99-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.99
   * storage: Add methods to compute directories/revisions diff
   * Add a new table for "bucketed" object counts
   * doc: update table clusters in SQL diagram
   * swh.storage.content_missing: Improve docstring
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Tue, 20 Feb 2018 13:32:25 +0100
 
 swh-storage (0.0.98-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.98
   * Switch backwards compatibility for snapshots off
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Tue, 06 Feb 2018 15:27:15 +0100
 
 swh-storage (0.0.97-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.97
   * refactor database initialization
   * use a separate thread instead of a temporary file for COPY
     operations
   * add more snapshot-related endpoints
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Tue, 06 Feb 2018 14:07:07 +0100
 
 swh-storage (0.0.96-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.96
   * Add snapshot models
   * Add support for hg revision type
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Tue, 19 Dec 2017 16:25:57 +0100
 
 swh-storage (0.0.95-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.95
   * swh.storage: Rename indexer_configuration to tool
   * swh.storage: Migrate indexer model to its own model
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 07 Dec 2017 09:56:31 +0100
 
 swh-storage (0.0.94-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.94
   * Open searching origins methods to storage
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Tue, 05 Dec 2017 12:32:57 +0100
 
 swh-storage (0.0.93-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.93
   * swh.storage: Open indexer_configuration_add endpoint
   * swh-data: Update content mimetype indexer configuration
   * origin_visit_get: make order repeatable
   * db: Make unique indices actually unique and vice versa
   * Add origin_metadata endpoints (add, get, etc...)
   * cleanup: Remove unused content provenance cache tables
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Fri, 24 Nov 2017 11:14:11 +0100
 
 swh-storage (0.0.92-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.92
   * make swh.storage.schemata work on SQLAlchemy 1.0
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Thu, 12 Oct 2017 19:51:24 +0200
 
 swh-storage (0.0.91-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage version 0.0.91
   * Update packaging runes
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Thu, 12 Oct 2017 18:41:46 +0200
 
 swh-storage (0.0.90-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.90
   * Remove leaky dependency on python3-kafka
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Wed, 11 Oct 2017 18:53:22 +0200
 
 swh-storage (0.0.89-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.89
   * Add new package for ancillary schemata
   * Add new metadata-related entry points
   * Update for new swh.model
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Wed, 11 Oct 2017 17:39:29 +0200
 
 swh-storage (0.0.88-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.88
   * Move the archiver to its own module
   * Prepare building for stretch
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Fri, 30 Jun 2017 14:52:12 +0200
 
 swh-storage (0.0.87-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.87
   * update tasks to new swh.scheduler api
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Mon, 12 Jun 2017 17:54:11 +0200
 
 swh-storage (0.0.86-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.86
   * archiver updates
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Tue, 06 Jun 2017 18:43:43 +0200
 
 swh-storage (0.0.85-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.85
   * Improve license endpoint's unknown license policy
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Tue, 06 Jun 2017 17:55:40 +0200
 
 swh-storage (0.0.84-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.84
   * Update indexer endpoints to use indexer configuration id
   * Add indexer configuration endpoint
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Fri, 02 Jun 2017 16:16:47 +0200
 
 swh-storage (0.0.83-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.83
   * Add blake2s256 new hash computation on content
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Fri, 31 Mar 2017 12:27:09 +0200
 
 swh-storage (0.0.82-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.82
   * swh.storage.listener: Subscribe to new origin notifications
   * sql/swh-func: improve equality check on the three columns for
     swh_content_missing
   * swh.storage: add length to directory listing primitives
   * refactoring: Migrate from swh.core.hashutil to swh.model.hashutil
   * swh.storage.archiver.updater: Create a content updater journal
     client
   * vault: add a git fast-import cooker
   * vault: generic cache to allow multiple cooker types and formats
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Tue, 21 Mar 2017 14:50:16 +0100
 
 swh-storage (0.0.81-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.81
   * archiver improvements for mass injection in azure
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Thu, 09 Mar 2017 11:15:28 +0100
 
 swh-storage (0.0.80-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.80
   * archiver improvements related to the mass injection of contents in
     azure
   * updates to the vault cooker
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Tue, 07 Mar 2017 15:12:35 +0100
 
 swh-storage (0.0.79-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.79
   * archiver: keep counts of objects in each archive
   * converters: normalize timestamps using swh.model
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Tue, 14 Feb 2017 19:37:36 +0100
 
 swh-storage (0.0.78-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.78
   * Refactoring some common code into swh.core + adaptation api calls in
   * swh.objstorage and swh.storage (storage and vault)
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 26 Jan 2017 15:08:03 +0100
 
 swh-storage (0.0.77-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.77
   * Paginate results for origin_visits endpoint
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 19 Jan 2017 14:41:49 +0100
 
 swh-storage (0.0.76-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.76
   * Unify storage and objstorage configuration and instantiation
     functions
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 15 Dec 2016 18:25:58 +0100
 
 swh-storage (0.0.75-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.75
   * Add information on indexer tools (T610)
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Fri, 02 Dec 2016 18:21:36 +0100
 
 swh-storage (0.0.74-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.74
   * Use strict equality for content ctags' symbols search
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Tue, 29 Nov 2016 17:25:29 +0100
 
 swh-storage (0.0.73-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.73
   * Improve ctags search query for edge cases
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Mon, 28 Nov 2016 16:34:55 +0100
 
 swh-storage (0.0.72-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.72
   * Permit pagination on content_ctags_search api endpoint
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 24 Nov 2016 14:19:29 +0100
 
 swh-storage (0.0.71-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.71
   * Open full-text search endpoint on ctags
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Wed, 23 Nov 2016 17:33:51 +0100
 
 swh-storage (0.0.70-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.70
   * Add new license endpoints (add/get)
   * Update ctags endpoints to align update conflict policy
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 10 Nov 2016 17:27:49 +0100
 
 swh-storage (0.0.69-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.69
   * storage: Open ctags entry points (missing, add, get)
   * storage: allow adding several origins at once
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 20 Oct 2016 16:07:07 +0200
 
 swh-storage (0.0.68-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.68
   * indexer: Open mimetype/language get endpoints
   * indexer: Add the mimetype/language add function with conflict_update
     flag
   * archiver: Extend worker-to-backend to transmit messages to another
   * queue (once done)
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 13 Oct 2016 15:30:21 +0200
 
 swh-storage (0.0.67-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.67
   * Fix provenance storage init function
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Wed, 12 Oct 2016 02:24:12 +0200
 
 swh-storage (0.0.66-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.66
   * Improve provenance configuration format
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Wed, 12 Oct 2016 01:39:26 +0200
 
 swh-storage (0.0.65-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.65
   * Open api entry points for swh.indexer about content mimetype and
   * language
   * Update schema graph to latest version
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Sat, 08 Oct 2016 10:00:30 +0200
 
 swh-storage (0.0.64-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.64
   * Fix: Missing incremented version 5 for archiver.dbversion
   * Retrieve information on a content cached
   * sql/swh-func: content cache populates lines in deterministic order
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 29 Sep 2016 21:50:59 +0200
 
 swh-storage (0.0.63-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.63
   * Make the 'worker to backend' destination agnostic (message
     parameter)
   * Improve 'unknown sha1' policy (archiver db can lag behind swh db)
   * Improve 'force copy' policy
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Fri, 23 Sep 2016 12:29:50 +0200
 
 swh-storage (0.0.62-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.62
   * Updates to the provenance cache to reduce churn on the main tables
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Thu, 22 Sep 2016 18:54:52 +0200
 
 swh-storage (0.0.61-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.61
   * Handle copies of unregistered sha1 in archiver db
   * Fix copy to only the targeted destination
   * Update to latest python3-swh.core dependency
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 22 Sep 2016 13:44:05 +0200
 
 swh-storage (0.0.60-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.60
   * Update archiver dependencies
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Tue, 20 Sep 2016 16:46:48 +0200
 
 swh-storage (0.0.59-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.59
   * Unify configuration property between director/worker
   * Deal with potential missing contents in the archiver db
   * Improve get_contents_error implementation
   * Remove dead code in swh.storage.db about archiver
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Sat, 17 Sep 2016 12:50:14 +0200
 
 swh-storage (0.0.58-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.58
   * ArchiverDirectorToBackend reads sha1 from stdin and sends chunks of
     sha1
   * for archival.
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Fri, 16 Sep 2016 22:17:14 +0200
 
 swh-storage (0.0.57-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.57
   * Update swh.storage.archiver
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 15 Sep 2016 16:30:11 +0200
 
 swh-storage (0.0.56-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.56
   * Vault: Add vault implementation (directory cooker & cache
   * implementation + its api)
   * Archiver: Add another archiver implementation (direct to backend)
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 15 Sep 2016 10:56:35 +0200
 
 swh-storage (0.0.55-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.55
   * Fix origin_visit endpoint
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 08 Sep 2016 15:21:28 +0200
 
 swh-storage (0.0.54-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.54
   * Open origin_visit_get_by entry point
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Mon, 05 Sep 2016 12:36:34 +0200
 
 swh-storage (0.0.53-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.53
   * Add cache about content provenance
   * debian: fix python3-swh.storage.archiver runtime dependency
   * debian: create new package python3-swh.storage.provenance
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Fri, 02 Sep 2016 11:14:09 +0200
 
 swh-storage (0.0.52-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.52
   * Package python3-swh.storage.archiver
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 25 Aug 2016 14:55:23 +0200
 
 swh-storage (0.0.51-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.51
   * Add new metadata column to origin_visit
   * Update swh-add-directory script for updated API
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Wed, 24 Aug 2016 14:36:03 +0200
 
 swh-storage (0.0.50-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.50
   * Add a function to pull (only) metadata for a list of contents
   * Update occurrence_add api entry point to properly deal with
     origin_visit
   * Add origin_visit api entry points to create/update origin_visit
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Tue, 23 Aug 2016 16:29:26 +0200
 
 swh-storage (0.0.49-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.49
   * Proper dependency on python3-kafka
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Fri, 19 Aug 2016 13:45:52 +0200
 
 swh-storage (0.0.48-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.48
   * Updates to the archiver
   * Notification support for new object creations
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Fri, 19 Aug 2016 12:13:50 +0200
 
 swh-storage (0.0.47-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.47
   * Update storage archiver to new schemaless schema
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Fri, 22 Jul 2016 16:59:19 +0200
 
 swh-storage (0.0.46-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.46
   * Update archiver bootstrap
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Wed, 20 Jul 2016 19:04:42 +0200
 
 swh-storage (0.0.45-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.45
   * Separate swh.storage.archiver's db from swh.storage.storage
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Tue, 19 Jul 2016 15:05:36 +0200
 
 swh-storage (0.0.44-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.44
   * Open listing visits per origin api
 
  -- Quentin Campos <qcampos@etud.u-pem.fr>  Fri, 08 Jul 2016 11:27:10 +0200
 
 swh-storage (0.0.43-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.43
   * Extract objstorage to its own package swh.objstorage
 
  -- Quentin Campos <qcampos@etud.u-pem.fr>  Mon, 27 Jun 2016 14:57:12 +0200
 
 swh-storage (0.0.42-1~swh1) unstable-swh; urgency=medium
 
   * Add an object storage multiplexer to allow transition between
     multiple versions of object storages.
 
  -- Quentin Campos <qcampos@etud.u-pem.fr>  Tue, 21 Jun 2016 15:03:52 +0200
 
 swh-storage (0.0.41-1~swh1) unstable-swh; urgency=medium
 
   * Refactoring of the object storage in order to allow multiple
     versions of it, as well as a multiplexer for version transition.
 
  -- Quentin Campos <qcampos@etud.u-pem.fr>  Thu, 16 Jun 2016 15:54:16 +0200
 
 swh-storage (0.0.40-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.40:
   * Refactor objstorage to allow for different implementations
   * Updates to the checker functionality
   * Bump swh.core dependency to v0.0.20
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Tue, 14 Jun 2016 17:25:42 +0200
 
 swh-storage (0.0.39-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.39
   * Add run_from_webserver function for objstorage api server
   * Add unique identifier message on default api server route endpoints
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Fri, 20 May 2016 15:27:34 +0200
 
 swh-storage (0.0.38-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.38
   * Add an http api for object storage
   * Implement an archiver to perform backup copies
 
  -- Quentin Campos <qcampos@etud.u-pem.fr>  Fri, 20 May 2016 14:40:14 +0200
 
 swh-storage (0.0.37-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.37
   * Add fullname to person table
   * Add svn as a revision type
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Fri, 08 Apr 2016 16:44:24 +0200
 
 swh-storage (0.0.36-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage v0.0.36
   * Add json-schema documentation for the jsonb fields
   * Overhaul entity handling
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Wed, 16 Mar 2016 17:27:17 +0100
 
 swh-storage (0.0.35-1~swh1) unstable-swh; urgency=medium
 
   * Release swh-storage v0.0.35
   * Factor in temporary tables with only an id (db v059)
   * Allow generic object search by sha1_git (db v060)
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Thu, 25 Feb 2016 16:21:01 +0100
 
 swh-storage (0.0.34-1~swh1) unstable-swh; urgency=medium
 
   * Release swh.storage version 0.0.34
   * occurrence improvements
   * commit metadata improvements
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Fri, 19 Feb 2016 18:20:07 +0100
 
 swh-storage (0.0.33-1~swh1) unstable-swh; urgency=medium
 
   * Bump swh.storage to version 0.0.33
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Fri, 05 Feb 2016 11:17:00 +0100
 
 swh-storage (0.0.32-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.32
   * Let the person's id flow
   * sql/upgrades/051: 050->051 schema change
   * sql/upgrades/050: 049->050 schema change - Clean up obsolete
     functions
   * sql/upgrades/049: Final take for 048->049 schema change.
   * sql: Use a new schema for occurrences
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Fri, 29 Jan 2016 17:44:27 +0100
 
 swh-storage (0.0.31-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.31
   * Deal with occurrence_history.branch, occurrence.branch, release.name
     as bytes
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Wed, 27 Jan 2016 15:45:53 +0100
 
 swh-storage (0.0.30-1~swh1) unstable-swh; urgency=medium
 
   * Prepare swh.storage v0.0.30 release
   * type-agnostic occurrences and revisions
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Tue, 26 Jan 2016 07:36:43 +0100
 
 swh-storage (0.0.29-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.29
   * New:
   * Upgrade sql schema to 041→043
   * Deal with communication downtime between clients and storage
   * Open occurrence_get(origin_id) to retrieve latest occurrences per
     origin
   * Open release_get_by to retrieve a release by origin
   * Open directory_get to retrieve information on directory by id
   * Open entity_get to retrieve information on entity + hierarchy from
     its uuid
   * Open directory_get that retrieve information on directory per id
   * Update:
   * directory_get/directory_ls: Rename to directory_ls
   * revision_log: update to retrieve logs from multiple root revisions
   * revision_get_by: branch name filtering is now optional
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Wed, 20 Jan 2016 16:15:50 +0100
 
 swh-storage (0.0.28-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.28
   * Open entity_get api
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Fri, 15 Jan 2016 16:37:27 +0100
 
 swh-storage (0.0.27-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.27
   * Open directory_entry_get_by_path api
   * Improve get_revision_by api performance
   * sql/swh-schema: add index on origin(type, url) --> improve origin
     lookup api
   * Bump to 039 db version
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Fri, 15 Jan 2016 12:42:47 +0100
 
 swh-storage (0.0.26-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.26
   * Open revision_get_by to retrieve a revision by occurrence criterion
     filtering
   * sql/upgrades/036: add 035→036 upgrade script
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Wed, 13 Jan 2016 12:46:44 +0100
 
 swh-storage (0.0.25-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.25
   * Limit results in swh_revision_list*
   * Create the package to align the current db production version on
     https://archive.softwareheritage.org/
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Fri, 08 Jan 2016 11:33:08 +0100
 
 swh-storage (0.0.24-1~swh1) unstable-swh; urgency=medium
 
   * Prepare swh.storage release v0.0.24
   * Add a limit argument to revision_log
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Wed, 06 Jan 2016 15:12:53 +0100
 
 swh-storage (0.0.23-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.23
   * Protect against overflow, wrapped in ValueError for client
   * Fix relative path import for remote storage.
   * api to retrieve revision_log is now 'parents' aware
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Wed, 06 Jan 2016 11:30:58 +0100
 
 swh-storage (0.0.22-1~swh1) unstable-swh; urgency=medium
 
   * Release v0.0.22
   * Fix relative import for remote storage
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Wed, 16 Dec 2015 16:04:48 +0100
 
 swh-storage (0.0.21-1~swh1) unstable-swh; urgency=medium
 
   * Prepare release v0.0.21
   * Protect the storage api client from overflows
   * Add a get_storage function mapping to local or remote storage
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Wed, 16 Dec 2015 13:34:46 +0100
 
 swh-storage (0.0.20-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.20
   * allow numeric timestamps with offset
   * Open revision_log api
   * start migration to swh.model
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Mon, 07 Dec 2015 15:20:36 +0100
 
 swh-storage (0.0.19-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.19
   * Improve directory listing with content data
   * Open person_get
   * Open release_get data reading
   * Improve origin_get api
   * Effort to unify api output on dict (for read)
   * Migrate backend to 032
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Fri, 27 Nov 2015 13:33:34 +0100
 
 swh-storage (0.0.18-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.18
   * Improve origin_get to permit retrieval per id
   * Update directory_get implementation (add join from
   * directory_entry_file to content)
   * Open release_get : [sha1] -> [Release]
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 19 Nov 2015 11:18:35 +0100
 
 swh-storage (0.0.17-1~swh1) unstable-swh; urgency=medium
 
   * Prepare deployment of swh.storage v0.0.17
   * Add some entity related entry points
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Tue, 03 Nov 2015 16:40:59 +0100
 
 swh-storage (0.0.16-1~swh1) unstable-swh; urgency=medium
 
   * v0.0.16
   * Add metadata column in revision (db version 29)
   * cache http connection for remote storage client
 
  -- Antoine R. Dumont (@ardumont) <antoine.romain.dumont@gmail.com>  Thu, 29 Oct 2015 10:29:00 +0100
 
 swh-storage (0.0.15-1~swh1) unstable-swh; urgency=medium
 
   * Prepare deployment of swh.storage v0.0.15
   * Allow population of fetch_history
   * Update organizations / projects as entities
   * Use schema v028 for directory addition
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Tue, 27 Oct 2015 11:43:39 +0100
 
 swh-storage (0.0.14-1~swh1) unstable-swh; urgency=medium
 
   * Prepare swh.storage v0.0.14 deployment
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Fri, 16 Oct 2015 15:34:08 +0200
 
 swh-storage (0.0.13-1~swh1) unstable-swh; urgency=medium
 
   * Prepare deploying swh.storage v0.0.13
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Fri, 16 Oct 2015 14:51:44 +0200
 
 swh-storage (0.0.12-1~swh1) unstable-swh; urgency=medium
 
   * Prepare deploying swh.storage v0.0.12
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Tue, 13 Oct 2015 12:39:18 +0200
 
 swh-storage (0.0.11-1~swh1) unstable-swh; urgency=medium
 
   * Preparing deployment of swh.storage v0.0.11
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Fri, 09 Oct 2015 17:44:51 +0200
 
 swh-storage (0.0.10-1~swh1) unstable-swh; urgency=medium
 
   * Prepare deployment of swh.storage v0.0.10
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Tue, 06 Oct 2015 17:37:00 +0200
 
 swh-storage (0.0.9-1~swh1) unstable-swh; urgency=medium
 
   * Prepare deployment of swh.storage v0.0.9
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Thu, 01 Oct 2015 19:03:00 +0200
 
 swh-storage (0.0.8-1~swh1) unstable-swh; urgency=medium
 
   * Prepare deployment of swh.storage v0.0.8
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Thu, 01 Oct 2015 11:32:46 +0200
 
 swh-storage (0.0.7-1~swh1) unstable-swh; urgency=medium
 
   * Prepare deployment of swh.storage v0.0.7
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Tue, 29 Sep 2015 16:52:54 +0200
 
 swh-storage (0.0.6-1~swh1) unstable-swh; urgency=medium
 
   * Prepare deployment of swh.storage v0.0.6
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Tue, 29 Sep 2015 16:43:24 +0200
 
 swh-storage (0.0.5-1~swh1) unstable-swh; urgency=medium
 
   * Prepare deploying swh.storage v0.0.5
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Tue, 29 Sep 2015 16:27:00 +0200
 
 swh-storage (0.0.1-1~swh1) unstable-swh; urgency=medium
 
   * Initial release
   * swh.storage.api: Properly escape arbitrary byte sequences in
     arguments
 
  -- Nicolas Dandrimont <nicolas@dandrimont.eu>  Tue, 22 Sep 2015 17:02:34 +0200
diff --git a/debian/control b/debian/control
index ab24c1a4..ddb1279e 100644
--- a/debian/control
+++ b/debian/control
@@ -1,53 +1,55 @@
 Source: swh-storage
 Maintainer: Software Heritage developers <swh-devel@inria.fr>
 Section: python
 Priority: optional
 Build-Depends: debhelper (>= 9),
                dh-python (>= 2),
                cassandra,
                openjdk-11-jre,
                python3-aiohttp,
                python3-all,
                python3-cassandra,
                python3-click,
                python3-dateutil,
                python3-flask,
                python3-hypothesis (>= 3.11.0~),
                python3-kafka,
                python3-mypy-extensions,
                python3-psycopg2 (>= 2.8),
                python3-pytest,
                python3-pytest-mock,
+               python3-pytest-redis,
                python3-requests,
                python3-setuptools,
                python3-setuptools-scm,
                python3-sqlalchemy (>= 1.0),
-               python3-swh.core (>= 0.9),
-               python3-swh.counters,
-               python3-swh.journal (>= 0.2),
-               python3-swh.model (>= 0.4),
-               python3-swh.objstorage (>= 0.0.40~),
-               python3-swh.core.db.pytestplugin (>= 0.9),
+               python3-swh.core (>= 0.14),
+               python3-swh.counters (>= 0.8),
+               python3-swh.journal (>= 0.9),
+               python3-swh.model (>= 2.1),
+               python3-swh.objstorage (>= 0.2.2),
+               python3-swh.core.db.pytestplugin (>= 0.14),
                python3-typing-extensions (>= 3.7.4~),
-               python3-tenacity
+               python3-tenacity,
+               redis-server
 # Only the jre 11 is supported with cassandra. Unfortunately, some other jre packages
 # are pulled, so we prevent those from being installed.
 # Related to https://forge.softwareheritage.org/T3053#58819
 Build-Conflicts: openjdk-17-jre-headless,
                  openjdk-16-jre-headless,
                  openjdk-15-jre-headless,
 Standards-Version: 3.9.6
 Homepage: https://forge.softwareheritage.org/diffusion/DSTO/
 
 Package: python3-swh.storage
 Architecture: all
 Depends: python3-swh.core (>= 0.9),
          python3-swh.model (>= 0.4),
          python3-swh.objstorage (>= 0.0.40~),
          python3-psycopg2 (>= 2.8),
          ${misc:Depends},
          ${python3:Depends}
 Breaks: python3-swh.archiver (<< 0.0.4~),
         python3-swh.indexer (<< 0.0.51~),
         python3-swh.vault (<< 0.0.19~)
 Description: Software Heritage storage utilities
diff --git a/docs/extrinsic-metadata-specification.rst b/docs/extrinsic-metadata-specification.rst
index f1522ec7..c4185651 100644
--- a/docs/extrinsic-metadata-specification.rst
+++ b/docs/extrinsic-metadata-specification.rst
@@ -1,347 +1,343 @@
 :orphan:
 
 .. _extrinsic-metadata-specification:
 
 Extrinsic metadata specification
 ================================
 
 :term:`Extrinsic metadata` is information about software that is not part
 of the source code itself but still closely related to the software.
 Typical sources for extrinsic metadata are: the hosting place of a
 repository, which can offer metadata via its web view or API; external
 registries like collaborative curation initiatives; and out-of-band
 information available at source code archival time.
 
 Since they are not part of the source code, a dedicated mechanism to fetch
 and store them is needed.
 
 This specification assumes the reader is familiar with Software Heritage's
 :ref:`architecture` and :ref:`data-model`.
 
 
 Metadata sources
 ----------------
 
 Authorities
 ^^^^^^^^^^^
 
 Metadata authorities are entities that provide metadata about an
 :term:`origin`. Metadata authorities include: code hosting places,
 :term:`deposit` submitters, and registries (eg. Wikidata).
 
 An authority is uniquely defined by these properties:
 
 * its type, representing the kind of authority, which is one of these values:
 
   * ``deposit_client``, for metadata pushed to Software Heritage at the same time
     as a software artifact
   * ``forge``, for metadata pulled from the same source as the one hosting
     the software artifacts (which includes package managers)
   * ``registry``, for metadata pulled from a third-party
 
 * its URL, which unambiguously identifies an instance of the authority type.
 
 Examples:
 
 =============== =================================
 type            url
 =============== =================================
 deposit_client  https://hal.archives-ouvertes.fr/
 deposit_client  https://hal.inria.fr/
 deposit_client  https://software.intel.com/
 forge           https://gitlab.com/
 forge           https://gitlab.inria.fr/
 forge           https://0xacab.org/
 forge           https://github.com/
 registry        https://www.wikidata.org/
 registry        https://swmath.org/
 registry        https://ascl.net/
 =============== =================================
 
 Metadata fetchers
 ^^^^^^^^^^^^^^^^^
 
 Metadata fetchers are software components used to fetch metadata from
 a metadata authority, and ingest them into the Software Heritage archive.
 
 A metadata fetcher is uniquely defined by these properties:
 
 * its type
 * its version
 
 Examples:
 
 * :term:`loaders <loader>`, which may either discover metadata as a
   side-effect of loading source code, or be dedicated to fetching metadata.
 
 * :term:`listers <lister>`, which may discover metadata as a side-effect
   of discovering origins.
 
 * :term:`deposit` submitters, which push metadata to SWH from a
   third-party; usually at the same time as a :term:`software artifact`
 
 * crawlers, which fetch metadata from an authority in a way that is
   none of the above (eg. by querying a specific API of the origin's forge).
 
 
 Storage API
 -----------
 
 Authorities and metadata fetchers
 ^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^
 
-The :term:`storage` API offers these endpoints to manipulate metadata
-authorities and metadata fetchers:
+Data model
+~~~~~~~~~~
 
-* ``metadata_authority_add(type, url, metadata)``
-  which adds a new metadata authority to the storage.
+The :term:`storage` API uses these structures to represent metadata
+authorities and metadata fetchers (simplified Python code)::
 
-* ``metadata_authority_get(type, url)``
-  which looks up a known authority (there is at most one) and if it is
-  known, returns a dictionary with keys ``type``, ``url``, and ``metadata``.
+   class MetadataAuthorityType(Enum):
+       DEPOSIT_CLIENT = "deposit_client"
+       FORGE = "forge"
+       REGISTRY = "registry"
 
-* ``metadata_fetcher_add(name, version, metadata)``
-  which adds a new metadata fetcher to the storage.
+   class MetadataAuthority(BaseModel):
+       """Represents an entity that provides metadata about an origin or
+       software artifact."""
 
-* ``metadata_fetcher_get(name, version)``
-  which looks up a known fetcher (there is at most one) and if it is
-  known, returns a dictionary with keys ``name``, ``version``, and
-  ``metadata``.
+       object_type = "metadata_authority"
 
-These `metadata` fields contain JSON-encodable dictionaries
-with information about the authority/fetcher, in a format specific to each
-authority/fetcher.
-With authority, the `metadata` field is reserved for information describing
-and qualifying the authority.
-With fetchers, the `metadata` field is reserved for configuration metadata
-and other technical usage.
+       type: MetadataAuthorityType
+       url: str
 
-Origin metadata
-^^^^^^^^^^^^^^^
+   class MetadataFetcher(BaseModel):
+       """Represents a software component used to fetch metadata from a metadata
+       authority, and ingest them into the Software Heritage archive."""
 
-Extrinsic metadata are stored in SWH's :term:`storage database`.
-The storage API offers three endpoints to manipulate origin metadata:
+       object_type = "metadata_fetcher"
 
-* Adding metadata::
+       name: str
+       version: str
 
-      raw_extrinsic_metadata_add(
-         "origin", origin_url, discovery_date,
-         authority, fetcher,
-         format, metadata
-      )
+Storage API
+~~~~~~~~~~~
 
-  which adds a new `metadata` byte string obtained from a given authority
-  and associated to the origin.
-  `discovery_date` is a Python datetime.
-  `authority` must be a dict containing keys `type` and `url`, and
-  `fetcher` a dict containing keys `name` and `version`.
-  The authority and fetcher must be known to the storage before using this
-  endpoint.
-  `format` is a text field indicating the format of the content of the
-  `metadata` byte string, see `extrinsic-metadata-formats`_.
+* ``metadata_authority_add(authorities: List[MetadataAuthority])``
+  which adds a list of ``MetadataAuthority`` to the storage.
 
-* Getting latest metadata::
+* ``metadata_authority_get(type: MetadataAuthorityType, url: str) -> Optional[MetadataAuthority]``
+  which looks up a known authority (there is at most one) and if it is
+  known, returns the corresponding ``MetadataAuthority``
 
-      raw_extrinsic_metadata_get_latest(
-         "origin", origin_url, authority
-      )
+* ``metadata_fetcher_add(fetchers: List[MetadataFetcher])``
+  which adds a list of ``MetadataFetcher`` to the storage.
 
-  where `authority` must be a dict containing keys `type` and `url`,
-  which returns a dictionary corresponding to the latest metadata entry
-  added from this origin, in the format::
+* ``metadata_fetcher_get(name: str, version: str) -> Optional[MetadataFetcher]``
+  which looks up a known fetcher (there is at most one) and if it is
+  known, returns the corresponding ``MetadataFetcher``
 
-      {
-        'origin_url': ...,
-        'authority': {'type': ..., 'url': ...},
-        'fetcher': {'name': ..., 'version': ...},
-        'discovery_date': ...,
-        'format': '...',
-        'metadata': b'...'
-      }
+Artifact metadata
+^^^^^^^^^^^^^^^^^
 
+Data model
+~~~~~~~~~~
 
-* Getting all metadata::
+The storage database stores metadata on origins, and all software artifacts
+supported by the data model.
+They are represented using this structure (simplified Python code)::
 
-      raw_extrinsic_metadata_get(
-         "origin", origin_url,
-         authority,
-         page_token, limit
-      )
+   class RawExtrinsicMetadata(HashableObject, BaseModel):
+       object_type = "raw_extrinsic_metadata"
 
-  where `authority` must be a dict containing keys `type` and `url`
-  which returns a dictionary with keys:
+       # target object
+       target: ExtendedSWHID
 
-  * `next_page_token`, which is an opaque token to be used as
-    `page_token` for retrieving the next page. if absent, there is
-    no more pages to gather.
-  * `results`: list of dictionaries, one for each metadata item
-    deposited, corresponding to the given origin and obtained from the
-    specified authority.
+       # source
+       discovery_date: datetime.datetime
+       authority: MetadataAuthority
+       fetcher: MetadataFetcher
 
-  Each of these dictionaries is in the following format::
+       # the metadata itself
+       format: str
+       metadata: bytes
 
-      {
-        'authority': {'type': ..., 'url': ...},
-        'fetcher': {'name': ..., 'version': ...},
-        'discovery_date': ...,
-        'format': '...',
-        'metadata': b'...'
-      }
-
-The parameters ``page_token`` and ``limit`` are used for pagination based on
-an arbitrary order. An initial query to ``origin_metadata_get`` must set
-``page_token`` to ``None``, and further query must use the value from the
-previous query's ``next_page_token`` to get the next page of results.
-
-``metadata`` is a bytes array (eventually encoded using Base64).
+       # context
+       origin: Optional[str] = None
+       visit: Optional[int] = None
+       snapshot: Optional[CoreSWHID] = None
+       release: Optional[CoreSWHID] = None
+       revision: Optional[CoreSWHID] = None
+       path: Optional[bytes] = None
+       directory: Optional[CoreSWHID] = None
+
+       id: Sha1Git
+
+The ``target`` may be:
+
+* a regular :ref:`core SWHID <persistent-identifiers>`,
+* a SWHID-like string with type ``ori`` and the SHA1 of an origin URL
+* a SWHID-like string with type ``emd`` and the SHA1 of an other
+  ``RawExtrinsicMetadata`` object (to represent metadata on metadata objects)
+
+``id`` is a sha1 hash of the ``RawExtrinsicMetadata`` object itself;
+it may be used in other ``RawExtrinsicMetadata`` as target.
+
+``discovery_date`` is a Python datetime.
+``authority`` must be a dict containing keys ``type`` and ``url``, and
+``fetcher`` a dict containing keys ``name`` and ``version``.
+The authority and fetcher must be known to the storage before using this
+endpoint.
+``format`` is a text field indicating the format of the content of the
+``metadata`` byte string, see `extrinsic-metadata-formats`_.
+
+``metadata`` is a byte array.
 Its format is specific to each authority; and is treated as an opaque value
 by the storage.
 Unifying these various formats into a common language is outside the scope
 of this specification.
 
-Artifact metadata
-^^^^^^^^^^^^^^^^^
-
-In addition to origin metadata, the storage database stores metadata on
-all software artifacts supported by the data model.
-
-This works similarly to origin metadata, with one major difference:
-extrinsic metadata can be given on a specific artifact within a specified
-context (for example: a directory in a specific revision from a specific
+Finally, the remaining fields allow metadata can be given on a specific artifact within
+a specified context (for example: a directory in a specific revision from a specific
 visit on a specific origin) which will be stored along the metadata itself.
 
 For example, two origins may develop the same file independently;
 the information about authorship, licensing or even description may vary
 about the same artifact in a different context.
 This is why it is important to qualify the metadata with the complete
 context for which it is intended, if any.
 
-The same two endpoints as for origin can be used, but with a different
-value for the first argument:
-
-* Adding metadata::
-
-      raw_extrinsic_metadata_add(
-         type, id, context, discovery_date,
-         authority, fetcher,
-         format, metadata
-      )
-
-
-* Getting all metadata::
-
-      raw_extrinsic_metadata_get(
-         type, id,
-         authority,
-         after,
-         page_token, limit
-      )
-
-
-definited similarly to ``origin_metadata_add`` and ``origin_metadata_get``,
-but where ``id`` is a core SWHID (with type matching ``<X>``),
-and with an extra ``context`` (argument when adding metadata, and dictionary
-key when getting them) that is a dictionary with keys
-depending on the artifact ``type``:
+The allowed context fields for each ``target`` type are:
 
-* for ``snapshot``: ``origin`` (a URL) and ``visit`` (an integer)
-* for ``release``: those above, plus ``snapshot``
+* for ``emd`` (extrinsic metadata) and ``ori`` (origin): none
+* for ``snp`` (snapshot): ``origin`` (a URL) and ``visit`` (an integer)
+* for ``rel`` (release): those above, plus ``snapshot``
   (the core SWHID of a snapshot)
-* for ``revision``: all those above, plus ``release``
+* for ``rev`` (revision): all those above, plus ``release``
   (the core SWHID of a release)
-* for ``directory``: all those above, plus ``revision``
+* for ``dir`` (directory): all those above, plus ``revision``
   (the core SWHID of a revision)
   and ``path`` (a byte string), representing the path to this directory
   from the root of the ``revision``
-* for ``content``: all those above, plus ``directory``
+* for ``cnt`` (content): all those above, plus ``directory``
   (the core SWHID of a directory)
 
 All keys are optional, but should be provided whenever possible.
 The dictionary may be empty, if metadata is fully independent from context.
 
 In all cases, ``visit`` should only be provided if ``origin`` is
 (as visit ids are only unique with respect to an origin).
 
+Storage API
+~~~~~~~~~~~
+
+The storage API offers three endpoints to manipulate origin metadata:
+
+* Adding metadata::
+
+      raw_extrinsic_metadata_add(metadata: List[RawExtrinsicMetadata])
+
+  which adds a list of ``RawExtrinsicMetadata`` objects, whose ``metadata`` field
+  is a byte string obtained from a given authority and associated to the ``target``.
+
+
+* Getting all metadata::
+
+      raw_extrinsic_metadata_get(
+          target: ExtendedSWHID,
+          authority: MetadataAuthority,
+          after: Optional[datetime.datetime] = None,
+          page_token: Optional[bytes] = None,
+          limit: int = 1000,
+      ) -> PagedResult[RawExtrinsicMetadata]:
+
+  returns a list of ``RawExtrinsicMetadata`` with the given ``target`` and from
+  the given ``authority``.
+  If ``after`` is provided, only objects whose discovery date is more recent are
+  returnered.
+
+  ``PagedResult`` is a structure containing the results themselves,
+  and a ``next_page_token`` used to fetch the next list of results, if any
+
 
 .. _extrinsic-metadata-formats:
 
-Extrinsic metadata format
--------------------------
+Extrinsic metadata formats
+--------------------------
 
 Here is a list of all the metadata format stored:
 
 ``pypi-project-json``
     The metadata is a release entry from a PyPI project's
     JSON file, extracted and re-serialized.
 ``replicate-npm-package-json``
     ditto, but from a replicate.npmjs.com project
 ``nixguix-sources-json``
     ditto, but from https://nix-community.github.io/nixpkgs-swh/
 ``original-artifacts-json``
     tarball data, see below
 ``sword-v2-atom-codemeta``
     XML Atom document, with Codemeta metadata,
     as sent by a deposit client, see the
     :ref:`Deposit protocol reference <deposit-protocol>`.
 ``sword-v2-atom-codemeta-v2``
     Deprecated alias of ``sword-v2-atom-codemeta``
 ``sword-v2-atom-codemeta-v2-in-json``
     Deprecated, JSON serialization of a ``sword-v2-atom-codemeta`` document.
 ``xml-deposit-info``
     Information about a deposit, to identify the provenance of
     a metadata object sent via swh-deposit, see below
 
 Details on some of these formats:
 
 
 original-artifacts-json
 ^^^^^^^^^^^^^^^^^^^^^^^
 
 This is a loosely defined format, originally used as a ``metadata`` column
 on the ``revision`` table that changed over the years.
 
 It is a JSON array, and each entry is a JSON object representing an archive
 (tarball, zipball, ...) that was unpackaged by the SWH loader
 before loading its content in Software Heritage.
 
 When writing this specification, it was stabilized to this format::
 
    [
       {
          "length": <int>,
          "filename": "<original filename>",
          "checksums": {
              "sha1": "<hex-encoded string>",
              "sha256": "<hex-encoded string>",
          },
          "url": "<URL the archive was downloaded from>"
       },
       ...
    ]
 
 Older ``original-artifacts-json`` were migrated to use this format,
 but may be missing some of the keys.
 
 
 xml-deposit-info
 ^^^^^^^^^^^^^^^^
 
 Deposits with code objects are loaded as their own origin, so we can
 look them up in the deposit database from their metadata (which hold the
 origin as a context).
 
 This is not true for metadata-only deposits, because we don't create an
 origin for them; so we need to store this information somewhere.
 The naive solution would be to insert them in the Atom entry provided by
 the client, but it means altering a document before we archive it, which
 potentially corrupts it or loses part of the data.
 
 Therefore, on each metadata-only deposit, the deposit creates an extra
 "metametadata" object, with the original metadata object as target,
 and using this format::
 
    <deposit xmlns="https://www.softwareheritage.org/schema/2018/deposit">
        <deposit_id>{{ deposit.id }}</deposit_id>
        <deposit_client>{{ deposit.client.provider_url }}</deposit_client>
        <deposit_collection>{{ deposit.collection.name }}</deposit_collection>
    </deposit>
diff --git a/requirements-swh-journal.txt b/requirements-swh-journal.txt
index 9a12e5c8..4d54e902 100644
--- a/requirements-swh-journal.txt
+++ b/requirements-swh-journal.txt
@@ -1 +1 @@
-swh.journal >= 0.6.2
+swh.journal >= 0.9
diff --git a/requirements-test.txt b/requirements-test.txt
index 67e0830d..8662be4f 100644
--- a/requirements-test.txt
+++ b/requirements-test.txt
@@ -1,13 +1,15 @@
 hypothesis >= 3.11.0
 pytest
 pytest-mock
 # pytz is in fact a dep of swh.model[testing] and should not be necessary, but
 # the dep on swh.model in the main requirements-swh.txt file shadows this one
 # adding the [testing] extra.
 swh.model[testing] >= 0.0.50
 pytz
+pytest-redis
 pytest-xdist
 types-python-dateutil
 types-pytz
 types-pyyaml
+types-redis
 types-requests
diff --git a/requirements.txt b/requirements.txt
index 06048429..05531f9d 100644
--- a/requirements.txt
+++ b/requirements.txt
@@ -1,10 +1,11 @@
+aiohttp
+cassandra-driver >= 3.19.0, != 3.21.0
 click
+deprecated
 flask
+iso8601
+mypy_extensions
 psycopg2
-aiohttp
+redis
 tenacity
-cassandra-driver >= 3.19.0, != 3.21.0
-deprecated
 typing-extensions
-mypy_extensions
-iso8601
diff --git a/swh.storage.egg-info/PKG-INFO b/swh.storage.egg-info/PKG-INFO
index 83e34df0..fab895dd 100644
--- a/swh.storage.egg-info/PKG-INFO
+++ b/swh.storage.egg-info/PKG-INFO
@@ -1,250 +1,250 @@
 Metadata-Version: 2.1
 Name: swh.storage
-Version: 0.39.0
+Version: 0.40.0
 Summary: Software Heritage storage manager
 Home-page: https://forge.softwareheritage.org/diffusion/DSTO/
 Author: Software Heritage developers
 Author-email: swh-devel@inria.fr
 License: UNKNOWN
 Project-URL: Bug Reports, https://forge.softwareheritage.org/maniphest
 Project-URL: Funding, https://www.softwareheritage.org/donate
 Project-URL: Source, https://forge.softwareheritage.org/source/swh-storage
 Project-URL: Documentation, https://docs.softwareheritage.org/devel/swh-storage/
 Platform: UNKNOWN
 Classifier: Programming Language :: Python :: 3
 Classifier: Intended Audience :: Developers
 Classifier: License :: OSI Approved :: GNU General Public License v3 (GPLv3)
 Classifier: Operating System :: OS Independent
 Classifier: Development Status :: 5 - Production/Stable
 Requires-Python: >=3.7
 Description-Content-Type: text/markdown
 Provides-Extra: testing
 Provides-Extra: journal
 License-File: LICENSE
 License-File: AUTHORS
 
 swh-storage
 ===========
 
 Abstraction layer over the archive, allowing to access all stored source code
 artifacts as well as their metadata.
 
 See the
 [documentation](https://docs.softwareheritage.org/devel/swh-storage/index.html)
 for more details.
 
 ## Quick start
 
 ### Dependencies
 
 Python tests for this module include tests that cannot be run without a local
 Postgresql database, so you need the Postgresql server executable on your
 machine (no need to have a running Postgresql server). They also expect a
 cassandra server.
 
 #### Debian-like host
 
 ```
 $ sudo apt install libpq-dev postgresql-11 cassandra
 ```
 
 #### Non Debian-like host
 
 The tests expects the path to `cassandra` to either be unspecified, it is then
 looked up at `/usr/sbin/cassandra`, either specified through the environment
 variable `SWH_CASSANDRA_BIN`.
 
 Optionally, you can avoid running the cassandra tests.
 
 ```
 (swh) :~/swh-storage$ tox -- -m 'not cassandra'
 ```
 
 ### Installation
 
 It is strongly recommended to use a virtualenv. In the following, we
 consider you work in a virtualenv named `swh`. See the
 [developer setup guide](https://docs.softwareheritage.org/devel/developer-setup.html#developer-setup)
 for a more details on how to setup a working environment.
 
 
 You can install the package directly from
 [pypi](https://pypi.org/p/swh.storage):
 
 ```
 (swh) :~$ pip install swh.storage
 [...]
 ```
 
 Or from sources:
 
 ```
 (swh) :~$ git clone https://forge.softwareheritage.org/source/swh-storage.git
 [...]
 (swh) :~$ cd swh-storage
 (swh) :~/swh-storage$ pip install .
 [...]
 ```
 
 Then you can check it's properly installed:
 ```
 (swh) :~$ swh storage --help
 Usage: swh storage [OPTIONS] COMMAND [ARGS]...
 
   Software Heritage Storage tools.
 
 Options:
   -h, --help  Show this message and exit.
 
 Commands:
   rpc-serve  Software Heritage Storage RPC server.
 ```
 
 
 ## Tests
 
 The best way of running Python tests for this module is to use
 [tox](https://tox.readthedocs.io/).
 
 ```
 (swh) :~$ pip install tox
 ```
 
 ### tox
 
 From the sources directory, simply use tox:
 
 ```
 (swh) :~/swh-storage$ tox
 [...]
 ========= 315 passed, 6 skipped, 15 warnings in 40.86 seconds ==========
 _______________________________ summary ________________________________
   flake8: commands succeeded
   py3: commands succeeded
   congratulations :)
 ```
 
 Note: it is possible to set the `JAVA_HOME` environment variable to specify the
 version of the JVM to be used by Cassandra. For example, at the time of writing
 this, Cassandra does not support java 14, so one may want to use for example
 java 11:
 
 ```
 (swh) :~/swh-storage$ export JAVA_HOME=/usr/lib/jvm/java-14-openjdk-amd64/bin/java
 (swh) :~/swh-storage$ tox
 [...]
 ```
 
 ## Development
 
 The storage server can be locally started. It requires a configuration file and
 a running Postgresql database.
 
 ### Sample configuration
 
 A typical configuration `storage.yml` file is:
 
 ```
 storage:
   cls: postgresql
   db: "dbname=softwareheritage-dev user=<user> password=<pwd>"
   objstorage:
     cls: pathslicing
     root: /tmp/swh-storage/
     slicing: 0:2/2:4/4:6
 ```
 
 which means, this uses:
 
 - a local storage instance whose db connection is to
   `softwareheritage-dev` local instance,
 
 - the objstorage uses a local objstorage instance whose:
 
   - `root` path is /tmp/swh-storage,
 
   - slicing scheme is `0:2/2:4/4:6`. This means that the identifier of
     the content (sha1) which will be stored on disk at first level
     with the first 2 hex characters, the second level with the next 2
     hex characters and the third level with the next 2 hex
     characters. And finally the complete hash file holding the raw
     content. For example: 00062f8bd330715c4f819373653d97b3cd34394c
     will be stored at 00/06/2f/00062f8bd330715c4f819373653d97b3cd34394c
 
 Note that the `root` path should exist on disk before starting the server.
 
 
 ### Starting the storage server
 
 If the python package has been properly installed (e.g. in a virtual env), you
 should be able to use the command:
 
 ```
 (swh) :~/swh-storage$ swh storage rpc-serve storage.yml
 ```
 
 This runs a local swh-storage api at 5002 port.
 
 ```
 (swh) :~/swh-storage$ curl http://127.0.0.1:5002
 <html>
 <head><title>Software Heritage storage server</title></head>
 <body>
 <p>You have reached the
 <a href="https://www.softwareheritage.org/">Software Heritage</a>
 storage server.<br />
 See its
 <a href="https://docs.softwareheritage.org/devel/swh-storage/">documentation
 and API</a> for more information</p>
 ```
 
 ### And then what?
 
 In your upper layer
 ([loader-git](https://forge.softwareheritage.org/source/swh-loader-git/),
 [loader-svn](https://forge.softwareheritage.org/source/swh-loader-svn/),
 etc...), you can define a remote storage with this snippet of yaml
 configuration.
 
 ```
 storage:
   cls: remote
   url: http://localhost:5002/
 ```
 
 You could directly define a postgresql storage with the following snippet:
 
 ```
 storage:
   cls: postgresql
   db: service=swh-dev
   objstorage:
     cls: pathslicing
     root: /home/storage/swh-storage/
     slicing: 0:2/2:4/4:6
 ```
 
 ## Cassandra
 
 As an alternative to PostgreSQL, swh-storage can use Cassandra as a database backend.
 It can be used like this:
 
 ```
 storage:
   cls: cassandra
   hosts:
     - localhost
   objstorage:
     cls: pathslicing
     root: /home/storage/swh-storage/
     slicing: 0:2/2:4/4:6
 ```
 
 The Cassandra swh-storage implementation supports both Cassandra >= 4.0-alpha2
 and ScyllaDB >= 4.4 (and possibly earlier versions, but this is untested).
 
 While the main code supports both transparently, running tests
 or configuring the schema requires specific code when using ScyllaDB,
 enabled by setting the `SWH_USE_SCYLLADB=1` environment variable.
 
 
diff --git a/swh.storage.egg-info/requires.txt b/swh.storage.egg-info/requires.txt
index 22de944d..25752023 100644
--- a/swh.storage.egg-info/requires.txt
+++ b/swh.storage.egg-info/requires.txt
@@ -1,30 +1,33 @@
+aiohttp
+cassandra-driver!=3.21.0,>=3.19.0
 click
+deprecated
 flask
+iso8601
+mypy_extensions
 psycopg2
-aiohttp
+redis
 tenacity
-cassandra-driver!=3.21.0,>=3.19.0
-deprecated
 typing-extensions
-mypy_extensions
-iso8601
 swh.core[db,http]>=0.14.0
 swh.counters>=v0.8.0
 swh.model>=2.1.0
 swh.objstorage>=0.2.2
 
 [journal]
-swh.journal>=0.6.2
+swh.journal>=0.9
 
 [testing]
 hypothesis>=3.11.0
 pytest
 pytest-mock
 swh.model[testing]>=0.0.50
 pytz
+pytest-redis
 pytest-xdist
 types-python-dateutil
 types-pytz
 types-pyyaml
+types-redis
 types-requests
-swh.journal>=0.6.2
+swh.journal>=0.9
diff --git a/swh/storage/backfill.py b/swh/storage/backfill.py
index b27cc6bb..a97a11fe 100644
--- a/swh/storage/backfill.py
+++ b/swh/storage/backfill.py
@@ -1,658 +1,658 @@
 # Copyright (C) 2017-2021  The Software Heritage developers
 # See the AUTHORS file at the top-level directory of this distribution
 # License: GNU General Public License version 3, or any later version
 # See top-level LICENSE file for more information
 
 """Storage backfiller.
 
 The backfiller goal is to produce back part or all of the objects
 from a storage to the journal topics
 
 Current implementation consists in the JournalBackfiller class.
 
 It simply reads the objects from the storage and sends every object identifier back to
 the journal.
 
 """
 
 import logging
 from typing import Any, Callable, Dict, Optional
 
 from swh.core.db import BaseDb
 from swh.model.model import (
     BaseModel,
     Directory,
     DirectoryEntry,
     ExtID,
     RawExtrinsicMetadata,
     Release,
     Revision,
     Snapshot,
     SnapshotBranch,
     TargetType,
 )
 from swh.model.swhids import ExtendedObjectType
 from swh.storage.postgresql.converters import (
     db_to_extid,
     db_to_raw_extrinsic_metadata,
     db_to_release,
     db_to_revision,
 )
-from swh.storage.replay import object_converter_fn
+from swh.storage.replay import OBJECT_CONVERTERS
 from swh.storage.writer import JournalWriter
 
 logger = logging.getLogger(__name__)
 
 PARTITION_KEY = {
     "content": "sha1",
     "skipped_content": "sha1",
     "directory": "id",
     "extid": "target",
     "metadata_authority": "type, url",
     "metadata_fetcher": "name, version",
     "raw_extrinsic_metadata": "target",
     "revision": "revision.id",
     "release": "release.id",
     "snapshot": "id",
     "origin": "id",
     "origin_visit": "origin_visit.origin",
     "origin_visit_status": "origin_visit_status.origin",
 }
 
 COLUMNS = {
     "content": [
         "sha1",
         "sha1_git",
         "sha256",
         "blake2s256",
         "length",
         "status",
         "ctime",
     ],
     "skipped_content": [
         "sha1",
         "sha1_git",
         "sha256",
         "blake2s256",
         "length",
         "ctime",
         "status",
         "reason",
     ],
     "directory": ["id", "dir_entries", "file_entries", "rev_entries"],
     "extid": ["extid_type", "extid", "extid_version", "target_type", "target"],
     "metadata_authority": ["type", "url"],
     "metadata_fetcher": ["name", "version"],
     "origin": ["url"],
     "origin_visit": ["visit", "type", ("origin.url", "origin"), "date",],
     "origin_visit_status": [
         ("origin_visit_status.visit", "visit"),
         ("origin.url", "origin"),
         ("origin_visit_status.date", "date"),
         "type",
         "snapshot",
         "status",
         "metadata",
     ],
     "raw_extrinsic_metadata": [
         "raw_extrinsic_metadata.type",
         "raw_extrinsic_metadata.target",
         "metadata_authority.type",
         "metadata_authority.url",
         "metadata_fetcher.name",
         "metadata_fetcher.version",
         "discovery_date",
         "format",
         "raw_extrinsic_metadata.metadata",
         "origin",
         "visit",
         "snapshot",
         "release",
         "revision",
         "path",
         "directory",
     ],
     "revision": [
         ("revision.id", "id"),
         "date",
         "date_offset",
         "date_neg_utc_offset",
         "committer_date",
         "committer_date_offset",
         "committer_date_neg_utc_offset",
         "type",
         "directory",
         "message",
         "synthetic",
         "metadata",
         "extra_headers",
         (
             "array(select parent_id::bytea from revision_history rh "
             "where rh.id = revision.id order by rh.parent_rank asc)",
             "parents",
         ),
         ("a.id", "author_id"),
         ("a.name", "author_name"),
         ("a.email", "author_email"),
         ("a.fullname", "author_fullname"),
         ("c.id", "committer_id"),
         ("c.name", "committer_name"),
         ("c.email", "committer_email"),
         ("c.fullname", "committer_fullname"),
     ],
     "release": [
         ("release.id", "id"),
         "date",
         "date_offset",
         "date_neg_utc_offset",
         "comment",
         ("release.name", "name"),
         "synthetic",
         "target",
         "target_type",
         ("a.id", "author_id"),
         ("a.name", "author_name"),
         ("a.email", "author_email"),
         ("a.fullname", "author_fullname"),
     ],
     "snapshot": ["id", "object_id"],
 }
 
 
 JOINS = {
     "release": ["person a on release.author=a.id"],
     "revision": [
         "person a on revision.author=a.id",
         "person c on revision.committer=c.id",
     ],
     "origin_visit": ["origin on origin_visit.origin=origin.id"],
     "origin_visit_status": ["origin on origin_visit_status.origin=origin.id",],
     "raw_extrinsic_metadata": [
         "metadata_authority on "
         "raw_extrinsic_metadata.authority_id=metadata_authority.id",
         "metadata_fetcher on raw_extrinsic_metadata.fetcher_id=metadata_fetcher.id",
     ],
 }
 
 EXTRA_WHERE = {
     # hack to force the right index usage on table extid
     "extid": "target_type in ('revision', 'release', 'content', 'directory')"
 }
 
 
 def directory_converter(db: BaseDb, directory_d: Dict[str, Any]) -> Directory:
     """Convert directory from the flat representation to swh model
        compatible objects.
 
     """
     columns = ["target", "name", "perms"]
     query_template = """
     select %(columns)s
     from directory_entry_%(type)s
     where id in %%s
     """
 
     types = ["file", "dir", "rev"]
 
     entries = []
     with db.cursor() as cur:
         for type in types:
             ids = directory_d.pop("%s_entries" % type)
             if not ids:
                 continue
             query = query_template % {
                 "columns": ",".join(columns),
                 "type": type,
             }
             cur.execute(query, (tuple(ids),))
             for row in cur:
                 entry_d = dict(zip(columns, row))
                 entry = DirectoryEntry(
                     name=entry_d["name"],
                     type=type,
                     target=entry_d["target"],
                     perms=entry_d["perms"],
                 )
                 entries.append(entry)
 
     return Directory(id=directory_d["id"], entries=tuple(entries),)
 
 
 def raw_extrinsic_metadata_converter(
     db: BaseDb, metadata: Dict[str, Any]
 ) -> RawExtrinsicMetadata:
     """Convert a raw extrinsic metadata from the flat representation to swh model
        compatible objects.
 
     """
     return db_to_raw_extrinsic_metadata(metadata)
 
 
 def extid_converter(db: BaseDb, extid: Dict[str, Any]) -> ExtID:
     """Convert an extid from the flat representation to swh model
        compatible objects.
 
     """
     return db_to_extid(extid)
 
 
 def revision_converter(db: BaseDb, revision_d: Dict[str, Any]) -> Revision:
     """Convert revision from the flat representation to swh model
        compatible objects.
 
     """
     revision = db_to_revision(revision_d)
     assert revision is not None, revision_d["id"]
     return revision
 
 
 def release_converter(db: BaseDb, release_d: Dict[str, Any]) -> Release:
     """Convert release from the flat representation to swh model
        compatible objects.
 
     """
     release = db_to_release(release_d)
     assert release is not None, release_d["id"]
     return release
 
 
 def snapshot_converter(db: BaseDb, snapshot_d: Dict[str, Any]) -> Snapshot:
     """Convert snapshot from the flat representation to swh model
        compatible objects.
 
     """
     columns = ["name", "target", "target_type"]
     query = """
     select %s
     from snapshot_branches sbs
     inner join snapshot_branch sb on sb.object_id=sbs.branch_id
     where sbs.snapshot_id=%%s
     """ % ", ".join(
         columns
     )
     with db.cursor() as cur:
         cur.execute(query, (snapshot_d["object_id"],))
         branches = {}
         for name, *row in cur:
             branch_d = dict(zip(columns[1:], row))
             if branch_d["target"] is not None and branch_d["target_type"] is not None:
                 branch: Optional[SnapshotBranch] = SnapshotBranch(
                     target=branch_d["target"],
                     target_type=TargetType(branch_d["target_type"]),
                 )
             else:
                 branch = None
             branches[name] = branch
 
     return Snapshot(id=snapshot_d["id"], branches=branches,)
 
 
 CONVERTERS: Dict[str, Callable[[BaseDb, Dict[str, Any]], BaseModel]] = {
     "directory": directory_converter,
     "extid": extid_converter,
     "raw_extrinsic_metadata": raw_extrinsic_metadata_converter,
     "revision": revision_converter,
     "release": release_converter,
     "snapshot": snapshot_converter,
 }
 
 
 def object_to_offset(object_id, numbits):
     """Compute the index of the range containing object id, when dividing
        space into 2^numbits.
 
     Args:
         object_id (str): The hex representation of object_id
         numbits (int): Number of bits in which we divide input space
 
     Returns:
         The index of the range containing object id
 
     """
     q, r = divmod(numbits, 8)
     length = q + (r != 0)
     shift_bits = 8 - r if r else 0
 
     truncated_id = object_id[: length * 2]
     if len(truncated_id) < length * 2:
         truncated_id += "0" * (length * 2 - len(truncated_id))
 
     truncated_id_bytes = bytes.fromhex(truncated_id)
     return int.from_bytes(truncated_id_bytes, byteorder="big") >> shift_bits
 
 
 def byte_ranges(numbits, start_object=None, end_object=None):
     """Generate start/end pairs of bytes spanning numbits bits and
        constrained by optional start_object and end_object.
 
     Args:
         numbits (int): Number of bits in which we divide input space
         start_object (str): Hex object id contained in the first range
                             returned
         end_object (str): Hex object id contained in the last range
                           returned
 
     Yields:
         2^numbits pairs of bytes
 
     """
     q, r = divmod(numbits, 8)
     length = q + (r != 0)
     shift_bits = 8 - r if r else 0
 
     def to_bytes(i):
         return int.to_bytes(i << shift_bits, length=length, byteorder="big")
 
     start_offset = 0
     end_offset = 1 << numbits
 
     if start_object is not None:
         start_offset = object_to_offset(start_object, numbits)
     if end_object is not None:
         end_offset = object_to_offset(end_object, numbits) + 1
 
     for start in range(start_offset, end_offset):
         end = start + 1
 
         if start == 0:
             yield None, to_bytes(end)
         elif end == 1 << numbits:
             yield to_bytes(start), None
         else:
             yield to_bytes(start), to_bytes(end)
 
 
 def raw_extrinsic_metadata_target_ranges(start_object=None, end_object=None):
     """Generate ranges of values for the `target` attribute of `raw_extrinsic_metadata`
     objects.
 
     This generates one range for all values before the first SWHID (which would
     correspond to raw origin URLs), then a number of hex-based ranges for each
     known type of SWHID (2**12 ranges for directories, 2**8 ranges for all other
     types). Finally, it generates one extra range for values above all possible
     SWHIDs.
     """
     if start_object is None:
         start_object = ""
 
     swhid_target_types = sorted(type.value for type in ExtendedObjectType)
 
     first_swhid = f"swh:1:{swhid_target_types[0]}:"
 
     # Generate a range for url targets, if the starting object is before SWHIDs
     if start_object < first_swhid:
         yield start_object, (
             first_swhid
             if end_object is None or end_object >= first_swhid
             else end_object
         )
 
     if end_object is not None and end_object <= first_swhid:
         return
 
     # Prime the following loop, which uses the upper bound of the previous range
     # as lower bound, to account for potential targets between two valid types
     # of SWHIDs (even though they would eventually be rejected by the
     # RawExtrinsicMetadata parser, they /might/ exist...)
     end_swhid = first_swhid
 
     # Generate ranges for swhid targets
     for target_type in swhid_target_types:
         finished = False
         base_swhid = f"swh:1:{target_type}:"
         last_swhid = base_swhid + ("f" * 40)
 
         if start_object > last_swhid:
             continue
 
         # Generate 2**8 or 2**12 ranges
         for _, end in byte_ranges(12 if target_type == "dir" else 8):
             # Reuse previous uppper bound
             start_swhid = end_swhid
 
             # Use last_swhid for this object type if on the last byte range
             end_swhid = (base_swhid + end.hex()) if end is not None else last_swhid
 
             # Ignore out of bounds ranges
             if start_object >= end_swhid:
                 continue
 
             # Potentially clamp start of range to the first object requested
             start_swhid = max(start_swhid, start_object)
 
             # Handle ending the loop early if the last requested object id is in
             # the current range
             if end_object is not None and end_swhid >= end_object:
                 end_swhid = end_object
                 finished = True
 
             yield start_swhid, end_swhid
 
             if finished:
                 return
 
     # Generate one final range for potential raw origin URLs after the last
     # valid SWHID
     start_swhid = max(start_object, end_swhid)
     yield start_swhid, end_object
 
 
 def integer_ranges(start, end, block_size=1000):
     for start in range(start, end, block_size):
         if start == 0:
             yield None, block_size
         elif start + block_size > end:
             yield start, end
         else:
             yield start, start + block_size
 
 
 RANGE_GENERATORS = {
     "content": lambda start, end: byte_ranges(24, start, end),
     "skipped_content": lambda start, end: [(None, None)],
     "directory": lambda start, end: byte_ranges(24, start, end),
     "extid": lambda start, end: byte_ranges(24, start, end),
     "revision": lambda start, end: byte_ranges(24, start, end),
     "release": lambda start, end: byte_ranges(16, start, end),
     "raw_extrinsic_metadata": raw_extrinsic_metadata_target_ranges,
     "snapshot": lambda start, end: byte_ranges(16, start, end),
     "origin": integer_ranges,
     "origin_visit": integer_ranges,
     "origin_visit_status": integer_ranges,
 }
 
 
 def compute_query(obj_type, start, end):
     columns = COLUMNS.get(obj_type)
     join_specs = JOINS.get(obj_type, [])
     join_clause = "\n".join("left join %s" % clause for clause in join_specs)
     additional_where = EXTRA_WHERE.get(obj_type)
 
     where = []
     where_args = []
     if start:
         where.append("%(keys)s >= %%s")
         where_args.append(start)
     if end:
         where.append("%(keys)s < %%s")
         where_args.append(end)
 
     if additional_where:
         where.append(additional_where)
 
     where_clause = ""
     if where:
         where_clause = ("where " + " and ".join(where)) % {
             "keys": "(%s)" % PARTITION_KEY[obj_type]
         }
 
     column_specs = []
     column_aliases = []
     for column in columns:
         if isinstance(column, str):
             column_specs.append(column)
             column_aliases.append(column)
         else:
             column_specs.append("%s as %s" % column)
             column_aliases.append(column[1])
 
     query = """
 select %(columns)s
 from %(table)s
 %(join)s
 %(where)s
     """ % {
         "columns": ",".join(column_specs),
         "table": obj_type,
         "join": join_clause,
         "where": where_clause,
     }
 
     return query, where_args, column_aliases
 
 
 def fetch(db, obj_type, start, end):
     """Fetch all obj_type's identifiers from db.
 
     This opens one connection, stream objects and when done, close
     the connection.
 
     Args:
         db (BaseDb): Db connection object
         obj_type (str): Object type
         start (Union[bytes|Tuple]): Range start identifier
         end (Union[bytes|Tuple]): Range end identifier
 
     Raises:
         ValueError if obj_type is not supported
 
     Yields:
         Objects in the given range
 
     """
     query, where_args, column_aliases = compute_query(obj_type, start, end)
     converter = CONVERTERS.get(obj_type)
     with db.cursor() as cursor:
         logger.debug("Fetching data for table %s", obj_type)
         logger.debug("query: %s %s", query, where_args)
         cursor.execute(query, where_args)
         for row in cursor:
             record = dict(zip(column_aliases, row))
             if converter:
                 record = converter(db, record)
             else:
-                record = object_converter_fn[obj_type](record)
+                record = OBJECT_CONVERTERS[obj_type](record)
 
             logger.debug("record: %s", record)
             yield record
 
 
 def _format_range_bound(bound):
     if isinstance(bound, bytes):
         return bound.hex()
     else:
         return str(bound)
 
 
 MANDATORY_KEYS = ["storage", "journal_writer"]
 
 
 class JournalBackfiller:
     """Class in charge of reading the storage's objects and sends those
        back to the journal's topics.
 
        This is designed to be run periodically.
 
     """
 
     def __init__(self, config=None):
         self.config = config
         self.check_config(config)
 
     def check_config(self, config):
         missing_keys = []
         for key in MANDATORY_KEYS:
             if not config.get(key):
                 missing_keys.append(key)
 
         if missing_keys:
             raise ValueError(
                 "Configuration error: The following keys must be"
                 " provided: %s" % (",".join(missing_keys),)
             )
 
         if "cls" not in config["storage"] or config["storage"]["cls"] not in (
             "local",
             "postgresql",
         ):
             raise ValueError(
                 "swh storage backfiller must be configured to use a local"
                 " (PostgreSQL) storage"
             )
 
     def parse_arguments(self, object_type, start_object, end_object):
         """Parse arguments
 
         Raises:
             ValueError for unsupported object type
             ValueError if object ids are not parseable
 
         Returns:
             Parsed start and end object ids
 
         """
         if object_type not in COLUMNS:
             raise ValueError(
                 "Object type %s is not supported. "
                 "The only possible values are %s"
                 % (object_type, ", ".join(sorted(COLUMNS.keys())))
             )
 
         if object_type in ["origin", "origin_visit", "origin_visit_status"]:
             if start_object:
                 start_object = int(start_object)
             else:
                 start_object = 0
             if end_object:
                 end_object = int(end_object)
             else:
                 end_object = 100 * 1000 * 1000  # hard-coded limit
 
         return start_object, end_object
 
     def run(self, object_type, start_object, end_object, dry_run=False):
         """Reads storage's subscribed object types and send them to the
            journal's reading topic.
 
         """
         start_object, end_object = self.parse_arguments(
             object_type, start_object, end_object
         )
 
         db = BaseDb.connect(self.config["storage"]["db"])
         writer = JournalWriter({"cls": "kafka", **self.config["journal_writer"]})
         assert writer.journal is not None
 
         for range_start, range_end in RANGE_GENERATORS[object_type](
             start_object, end_object
         ):
             logger.info(
                 "Processing %s range %s to %s",
                 object_type,
                 _format_range_bound(range_start),
                 _format_range_bound(range_end),
             )
 
             objects = fetch(db, object_type, start=range_start, end=range_end)
 
             if not dry_run:
                 writer.write_additions(object_type, objects)
             else:
                 # only consume the objects iterator to check for any potential
                 # decoding/encoding errors
                 for obj in objects:
                     pass
 
 
 if __name__ == "__main__":
     print('Please use the "swh-journal backfiller run" command')
diff --git a/swh/storage/cli.py b/swh/storage/cli.py
index 34fd17e0..881d0f40 100644
--- a/swh/storage/cli.py
+++ b/swh/storage/cli.py
@@ -1,227 +1,276 @@
 # Copyright (C) 2015-2020  The Software Heritage developers
 # See the AUTHORS file at the top-level directory of this distribution
 # License: GNU General Public License version 3, or any later version
 # See top-level LICENSE file for more information
 
 # WARNING: do not import unnecessary things here to keep cli startup time under
 # control
 import logging
 import os
 from typing import Dict, Optional
 
 import click
 
 from swh.core.cli import CONTEXT_SETTINGS
 from swh.core.cli import swh as swh_cli_group
+from swh.storage.replay import ModelObjectDeserializer
+
 
 try:
     from systemd.daemon import notify
 except ImportError:
     notify = None
 
 
 @swh_cli_group.group(name="storage", context_settings=CONTEXT_SETTINGS)
 @click.option(
     "--config-file",
     "-C",
     default=None,
     type=click.Path(exists=True, dir_okay=False,),
     help="Configuration file.",
 )
 @click.option(
     "--check-config",
     default=None,
     type=click.Choice(["no", "read", "write"]),
     help=(
         "Check the configuration of the storage at startup for read or write access; "
         "if set, override the value present in the configuration file if any. "
         "Defaults to 'read' for the 'backfill' command, and 'write' for 'rpc-server' "
         "and 'replay' commands."
     ),
 )
 @click.pass_context
 def storage(ctx, config_file, check_config):
     """Software Heritage Storage tools."""
     from swh.core import config
 
     if not config_file:
         config_file = os.environ.get("SWH_CONFIG_FILENAME")
 
     if config_file:
         if not os.path.exists(config_file):
             raise ValueError("%s does not exist" % config_file)
         conf = config.read(config_file)
     else:
         conf = {}
 
     if "storage" not in conf:
         ctx.fail("You must have a storage configured in your config file.")
 
     ctx.ensure_object(dict)
     ctx.obj["config"] = conf
     ctx.obj["check_config"] = check_config
 
 
 @storage.command(name="rpc-serve")
 @click.option(
     "--host",
     default="0.0.0.0",
     metavar="IP",
     show_default=True,
     help="Host ip address to bind the server on",
 )
 @click.option(
     "--port",
     default=5002,
     type=click.INT,
     metavar="PORT",
     show_default=True,
     help="Binding port of the server",
 )
 @click.option(
     "--debug/--no-debug",
     default=True,
     help="Indicates if the server should run in debug mode",
 )
 @click.pass_context
 def serve(ctx, host, port, debug):
     """Software Heritage Storage RPC server.
 
     Do NOT use this in a production environment.
     """
     from swh.storage.api.server import app
 
     if "log_level" in ctx.obj:
         logging.getLogger("werkzeug").setLevel(ctx.obj["log_level"])
     ensure_check_config(ctx.obj["config"], ctx.obj["check_config"], "write")
     app.config.update(ctx.obj["config"])
     app.run(host, port=int(port), debug=bool(debug))
 
 
 @storage.command()
 @click.argument("object_type")
 @click.option("--start-object", default=None)
 @click.option("--end-object", default=None)
 @click.option("--dry-run", is_flag=True, default=False)
 @click.pass_context
 def backfill(ctx, object_type, start_object, end_object, dry_run):
     """Run the backfiller
 
     The backfiller list objects from a Storage and produce journal entries from
     there.
 
     Typically used to rebuild a journal or compensate for missing objects in a
     journal (eg. due to a downtime of this later).
 
     The configuration file requires the following entries:
 
     - brokers: a list of kafka endpoints (the journal) in which entries will be
       added.
     - storage_dbconn: URL to connect to the storage DB.
     - prefix: the prefix of the topics (topics will be <prefix>.<object_type>).
     - client_id: the kafka client ID.
 
     """
     ensure_check_config(ctx.obj["config"], ctx.obj["check_config"], "read")
 
     # for "lazy" loading
     from swh.storage.backfill import JournalBackfiller
 
     try:
         from systemd.daemon import notify
     except ImportError:
         notify = None
 
     conf = ctx.obj["config"]
     backfiller = JournalBackfiller(conf)
 
     if notify:
         notify("READY=1")
 
     try:
         backfiller.run(
             object_type=object_type,
             start_object=start_object,
             end_object=end_object,
             dry_run=dry_run,
         )
     except KeyboardInterrupt:
         if notify:
             notify("STOPPING=1")
         ctx.exit(0)
 
 
 @storage.command()
 @click.option(
     "--stop-after-objects",
     "-n",
     default=None,
     type=int,
     help="Stop after processing this many objects. Default is to " "run forever.",
 )
+@click.option(
+    "--type",
+    "-t",
+    "object_types",
+    default=[],
+    type=click.Choice(
+        # use a hardcoded list to prevent having to load the
+        # replay module at cli loading time
+        [
+            "origin",
+            "origin_visit",
+            "origin_visit_status",
+            "snapshot",
+            "revision",
+            "release",
+            "directory",
+            "content",
+            "skipped_content",
+            "metadata_authority",
+            "metadata_fetcher",
+            "raw_extrinsic_metadata",
+            "extid",
+        ]
+    ),
+    help="Object types to replay",
+    multiple=True,
+)
 @click.pass_context
-def replay(ctx, stop_after_objects):
+def replay(ctx, stop_after_objects, object_types):
     """Fill a Storage by reading a Journal.
 
     There can be several 'replayers' filling a Storage as long as they use
     the same `group-id`.
     """
     import functools
 
     from swh.journal.client import get_journal_client
     from swh.storage import get_storage
     from swh.storage.replay import process_replay_objects
 
     ensure_check_config(ctx.obj["config"], ctx.obj["check_config"], "write")
 
     conf = ctx.obj["config"]
     storage = get_storage(**conf.pop("storage"))
 
+    if "error_reporter" in conf:
+        from redis import Redis
+
+        reporter = Redis(**conf["error_reporter"]).set
+    else:
+        reporter = None
+    validate = conf.get("privileged", False)
+
+    if not validate and reporter:
+        ctx.fail(
+            "Invalid configuration: you cannot have 'error_reporter' set if "
+            "'privileged' is False; we cannot validate anonymized objects."
+        )
+
+    deserializer = ModelObjectDeserializer(reporter=reporter, validate=validate)
+
     client_cfg = conf.pop("journal_client")
+    client_cfg["value_deserializer"] = deserializer.convert
+    if object_types:
+        client_cfg["object_types"] = object_types
     if stop_after_objects:
         client_cfg["stop_after_objects"] = stop_after_objects
+
     try:
         client = get_journal_client(**client_cfg)
     except ValueError as exc:
         ctx.fail(exc)
 
     worker_fn = functools.partial(process_replay_objects, storage=storage)
 
     if notify:
         notify("READY=1")
 
     try:
         client.process(worker_fn)
     except KeyboardInterrupt:
         ctx.exit(0)
     else:
         print("Done.")
     finally:
         if notify:
             notify("STOPPING=1")
         client.close()
 
 
 def ensure_check_config(storage_cfg: Dict, check_config: Optional[str], default: str):
     """Helper function to inject the setting of check_config option in the storage config
     dict according to the expected default value (default value depends on the command,
     eg. backfill can be read-only).
 
     """
     if check_config is not None:
         if check_config == "no":
             storage_cfg.pop("check_config", None)
         else:
             storage_cfg["check_config"] = {"check_write": check_config == "write"}
     else:
         if "check_config" not in storage_cfg:
             storage_cfg["check_config"] = {"check_write": default == "write"}
 
 
 def main():
     logging.basicConfig()
     return serve(auto_envvar_prefix="SWH_STORAGE")
 
 
 if __name__ == "__main__":
     main()
diff --git a/swh/storage/fixer.py b/swh/storage/fixer.py
index 14b21c5e..1c29df49 100644
--- a/swh/storage/fixer.py
+++ b/swh/storage/fixer.py
@@ -1,331 +1,65 @@
 # Copyright (C) 2020 The Software Heritage developers
 # See the AUTHORS file at the top-level directory of this distribution
 # License: GNU General Public License version 3, or any later version
 # See top-level LICENSE file for more information
 
-import copy
-import datetime
 import logging
-from typing import Any, Dict, List, Optional
+from typing import Any, Callable, Dict, List
 
-from swh.model.model import Origin, TimestampWithTimezone
+from swh.model.model import Origin
 
 logger = logging.getLogger(__name__)
 
 
 def _fix_content(content: Dict[str, Any]) -> Dict[str, Any]:
     """Filters-out invalid 'perms' key that leaked from swh.model.from_disk
     to the journal.
 
     >>> _fix_content({'perms': 0o100644, 'sha1_git': b'foo'})
     {'sha1_git': b'foo'}
 
     >>> _fix_content({'sha1_git': b'bar'})
     {'sha1_git': b'bar'}
 
     """
     content = content.copy()
     content.pop("perms", None)
     return content
 
 
-def _fix_revision_pypi_empty_string(rev):
-    """PyPI loader failed to encode empty strings as bytes, see:
-    swh:1:rev:8f0095ee0664867055d03de9bcc8f95b91d8a2b9
-    or https://forge.softwareheritage.org/D1772
-    """
-    rev = {
-        **rev,
-        "author": rev["author"].copy(),
-        "committer": rev["committer"].copy(),
-    }
-    if rev["author"].get("email") == "":
-        rev["author"]["email"] = b""
-    if rev["author"].get("name") == "":
-        rev["author"]["name"] = b""
-    if rev["committer"].get("email") == "":
-        rev["committer"]["email"] = b""
-    if rev["committer"].get("name") == "":
-        rev["committer"]["name"] = b""
-    return rev
-
-
-def _fix_revision_transplant_source(rev):
-    if rev.get("metadata") and rev["metadata"].get("extra_headers"):
-        rev = copy.deepcopy(rev)
-        rev["metadata"]["extra_headers"] = [
-            [key, value.encode("ascii")]
-            if key == "transplant_source" and isinstance(value, str)
-            else [key, value]
-            for (key, value) in rev["metadata"]["extra_headers"]
-        ]
-    return rev
-
-
-def _check_date(date):
-    """Returns whether the date can be represented in backends with sane
-    limits on timestamps and timezones (resp. signed 64-bits and
-    signed 16 bits), and that microseconds is valid (ie. between 0 and 10^6).
-    """
-    if date is None:
-        return True
-    try:
-        TimestampWithTimezone.from_dict(date)
-    except ValueError:
-        return False
-    else:
-        return True
-
-
-def _check_revision_date(rev):
-    """Exclude revisions with invalid dates.
-    See https://forge.softwareheritage.org/T1339"""
-    return _check_date(rev["date"]) and _check_date(rev["committer_date"])
-
-
-def _fix_revision(revision: Dict[str, Any]) -> Optional[Dict]:
-    """Fix various legacy revision issues.
-
-    Fix author/committer person:
-
-    >>> from pprint import pprint
-    >>> date = {
-    ...     'timestamp': {
-    ...         'seconds': 1565096932,
-    ...         'microseconds': 0,
-    ...     },
-    ...     'offset': 0,
-    ... }
-    >>> rev0 = _fix_revision({
-    ...     'id': b'rev-id',
-    ...     'author': {'fullname': b'', 'name': '', 'email': ''},
-    ...     'committer': {'fullname': b'', 'name': '', 'email': ''},
-    ...     'date': date,
-    ...     'committer_date': date,
-    ...     'type': 'git',
-    ...     'message': '',
-    ...     'directory': b'dir-id',
-    ...     'synthetic': False,
-    ... })
-    >>> rev0['author']
-    {'fullname': b'', 'name': b'', 'email': b''}
-    >>> rev0['committer']
-    {'fullname': b'', 'name': b'', 'email': b''}
-
-    Fix type of 'transplant_source' extra headers:
-
-    >>> rev1 = _fix_revision({
-    ...     'id': b'rev-id',
-    ...     'author': {'fullname': b'', 'name': '', 'email': ''},
-    ...     'committer': {'fullname': b'', 'name': '', 'email': ''},
-    ...     'date': date,
-    ...     'committer_date': date,
-    ...     'metadata': {
-    ...         'extra_headers': [
-    ...             ['time_offset_seconds', b'-3600'],
-    ...             ['transplant_source', '29c154a012a70f49df983625090434587622b39e']
-    ...     ]},
-    ...     'type': 'git',
-    ...     'message': '',
-    ...     'directory': b'dir-id',
-    ...     'synthetic': False,
-    ... })
-    >>> pprint(rev1['metadata']['extra_headers'])
-    [['time_offset_seconds', b'-3600'],
-     ['transplant_source', b'29c154a012a70f49df983625090434587622b39e']]
-
-    Revision with invalid date are filtered:
-
-    >>> from copy import deepcopy
-    >>> invalid_date1 = deepcopy(date)
-    >>> invalid_date1['timestamp']['microseconds'] = 1000000000  # > 10^6
-    >>> rev = _fix_revision({
-    ...     'author': {'fullname': b'', 'name': '', 'email': ''},
-    ...     'committer': {'fullname': b'', 'name': '', 'email': ''},
-    ...     'date': invalid_date1,
-    ...     'committer_date': date,
-    ... })
-    >>> rev is None
-    True
-
-    >>> invalid_date2 = deepcopy(date)
-    >>> invalid_date2['timestamp']['seconds'] = 2**70  # > 10^63
-    >>> rev = _fix_revision({
-    ...     'author': {'fullname': b'', 'name': '', 'email': ''},
-    ...     'committer': {'fullname': b'', 'name': '', 'email': ''},
-    ...     'date': invalid_date2,
-    ...     'committer_date': date,
-    ... })
-    >>> rev is None
-    True
-
-    >>> invalid_date3 = deepcopy(date)
-    >>> invalid_date3['offset'] = 2**20  # > 10^15
-    >>> rev = _fix_revision({
-    ...     'author': {'fullname': b'', 'name': '', 'email': ''},
-    ...     'committer': {'fullname': b'', 'name': '', 'email': ''},
-    ...     'date': date,
-    ...     'committer_date': invalid_date3,
-    ... })
-    >>> rev is None
-    True
-
-    """  # noqa
-    rev = _fix_revision_pypi_empty_string(revision)
-    rev = _fix_revision_transplant_source(rev)
-    if not _check_revision_date(rev):
-        logger.warning(
-            "Invalid revision date detected: %(revision)s", {"revision": rev}
-        )
-        return None
-    return rev
-
-
-def _fix_origin(origin: Dict) -> Dict:
-    """Fix legacy origin with type which is no longer part of the model.
-
-    >>> from pprint import pprint
-    >>> pprint(_fix_origin({
-    ...     'url': 'http://foo',
-    ... }))
-    {'url': 'http://foo'}
-    >>> pprint(_fix_origin({
-    ...     'url': 'http://bar',
-    ...     'type': 'foo',
-    ... }))
-    {'url': 'http://bar'}
-
-    """
-    o = origin.copy()
-    o.pop("type", None)
-    return o
-
-
-def _fix_origin_visit(visit: Dict) -> Dict:
-    """Fix various legacy origin visit issues.
-
-    `visit['origin']` is a dict instead of an URL:
-
-    >>> from datetime import datetime, timezone
-    >>> from pprint import pprint
-    >>> date = datetime(2020, 2, 27, 14, 39, 19, tzinfo=timezone.utc)
-    >>> pprint(_fix_origin_visit({
-    ...     'origin': {'url': 'http://foo'},
-    ...     'date': date,
-    ...     'type': 'git',
-    ...     'status': 'ongoing',
-    ...     'snapshot': None,
-    ... }))
-    {'date': datetime.datetime(2020, 2, 27, 14, 39, 19, tzinfo=datetime.timezone.utc),
-     'origin': 'http://foo',
-     'type': 'git'}
-
-    `visit['type']` is missing , but `origin['visit']['type']` exists:
-
-    >>> pprint(_fix_origin_visit(
-    ...     {'origin': {'type': 'hg', 'url': 'http://foo'},
-    ...     'date': date,
-    ...     'status': 'ongoing',
-    ...     'snapshot': None,
-    ... }))
-    {'date': datetime.datetime(2020, 2, 27, 14, 39, 19, tzinfo=datetime.timezone.utc),
-     'origin': 'http://foo',
-     'type': 'hg'}
-
-    >>> pprint(_fix_origin_visit(
-    ...     {'origin': {'type': 'hg', 'url': 'http://foo'},
-    ...     'date': '2020-02-27 14:39:19+00:00',
-    ...     'status': 'ongoing',
-    ...     'snapshot': None,
-    ... }))
-    {'date': datetime.datetime(2020, 2, 27, 14, 39, 19, tzinfo=datetime.timezone.utc),
-     'origin': 'http://foo',
-     'type': 'hg'}
-
-    Old visit format (origin_visit with no type) raises:
-
-    >>> _fix_origin_visit({
-    ...     'origin': {'url': 'http://foo'},
-    ...     'date': date,
-    ...     'status': 'ongoing',
-    ...     'snapshot': None
-    ... })
-    Traceback (most recent call last):
-    ...
-    ValueError: Old origin visit format detected...
-
-    >>> _fix_origin_visit({
-    ...     'origin': 'http://foo',
-    ...     'date': date,
-    ...     'status': 'ongoing',
-    ...     'snapshot': None
-    ... })
-    Traceback (most recent call last):
-    ...
-    ValueError: Old origin visit format detected...
-
-    """  # noqa
-    visit = visit.copy()
-    if "type" not in visit:
-        if isinstance(visit["origin"], dict) and "type" in visit["origin"]:
-            # Very old version of the schema: visits did not have a type,
-            # but their 'origin' field was a dict with a 'type' key.
-            visit["type"] = visit["origin"]["type"]
-        else:
-            # Very old schema version: 'type' is missing, stop early
-
-            # We expect the journal's origin_visit topic to no longer reference
-            # such visits. If it does, the replayer must crash so we can fix
-            # the journal's topic.
-            raise ValueError(f"Old origin visit format detected: {visit}")
-    if isinstance(visit["origin"], dict):
-        # Old version of the schema: visit['origin'] was a dict.
-        visit["origin"] = visit["origin"]["url"]
-    date = visit["date"]
-    if isinstance(date, str):
-        visit["date"] = datetime.datetime.fromisoformat(date)
-    # Those are no longer part of the model
-    for key in ["status", "snapshot", "metadata"]:
-        visit.pop(key, None)
-    return visit
-
-
 def _fix_raw_extrinsic_metadata(obj_dict: Dict) -> Dict:
     """Fix legacy RawExtrinsicMetadata with type which is no longer part of the model.
 
     >>> _fix_raw_extrinsic_metadata({
     ...     'type': 'directory',
     ...     'target': 'swh:1:dir:460a586d1c95d120811eaadb398d534e019b5243',
     ... })
     {'target': 'swh:1:dir:460a586d1c95d120811eaadb398d534e019b5243'}
     >>> _fix_raw_extrinsic_metadata({
     ...     'type': 'origin',
     ...     'target': 'https://inria.halpreprod.archives-ouvertes.fr/hal-01667309',
     ... })
     {'target': 'swh:1:ori:155291d5b9ada4570672510509f93fcfd9809882'}
 
     """
     o = obj_dict.copy()
     if o.pop("type", None) == "origin":
         o["target"] = str(Origin(o["target"]).swhid())
     return o
 
 
+object_fixers: Dict[str, Callable[[Dict], Dict]] = {
+    "content": _fix_content,
+    "raw_extrinsic_metadata": _fix_raw_extrinsic_metadata,
+}
+
+
 def fix_objects(object_type: str, objects: List[Dict]) -> List[Dict]:
     """
     Fix legacy objects from the journal to bring them up to date with the
     latest storage schema.
     """
-    if object_type == "content":
-        return [_fix_content(v) for v in objects]
-    elif object_type == "revision":
-        revisions = [_fix_revision(v) for v in objects]
-        return [rev for rev in revisions if rev is not None]
-    elif object_type == "origin":
-        return [_fix_origin(v) for v in objects]
-    elif object_type == "origin_visit":
-        return [_fix_origin_visit(v) for v in objects]
-    elif object_type == "raw_extrinsic_metadata":
-        return [_fix_raw_extrinsic_metadata(v) for v in objects]
-    else:
-        return objects
+    if object_type in object_fixers:
+        fixer = object_fixers[object_type]
+        objects = [fixer(v) for v in objects]
+    return objects
diff --git a/swh/storage/replay.py b/swh/storage/replay.py
index a022c566..eb86a16b 100644
--- a/swh/storage/replay.py
+++ b/swh/storage/replay.py
@@ -1,184 +1,231 @@
 # Copyright (C) 2019-2020 The Software Heritage developers
 # See the AUTHORS file at the top-level directory of this distribution
 # License: GNU General Public License version 3, or any later version
 # See top-level LICENSE file for more information
 
+from collections import Counter
+from functools import partial
 import logging
-from typing import Any, Callable, Container, Dict, List
+from typing import Any, Callable
+from typing import Counter as CounterT
+from typing import Dict, List, Optional, TypeVar, Union, cast
 
 try:
     from systemd.daemon import notify
 except ImportError:
     notify = None
 
 from swh.core.statsd import statsd
+from swh.journal.serializers import kafka_to_value
+from swh.model.hashutil import hash_to_hex
 from swh.model.model import (
     BaseContent,
     BaseModel,
     Content,
     Directory,
     ExtID,
+    HashableObject,
     MetadataAuthority,
     MetadataFetcher,
     Origin,
     OriginVisit,
     OriginVisitStatus,
     RawExtrinsicMetadata,
     Release,
     Revision,
     SkippedContent,
     Snapshot,
 )
-from swh.storage.exc import HashCollision
-from swh.storage.fixer import fix_objects
+from swh.storage.exc import HashCollision, StorageArgumentException
 from swh.storage.interface import StorageInterface
+from swh.storage.utils import remove_keys
 
 logger = logging.getLogger(__name__)
 
 GRAPH_OPERATIONS_METRIC = "swh_graph_replayer_operations_total"
 GRAPH_DURATION_METRIC = "swh_graph_replayer_duration_seconds"
 
 
-object_converter_fn: Dict[str, Callable[[Dict], BaseModel]] = {
+OBJECT_CONVERTERS: Dict[str, Callable[[Dict], BaseModel]] = {
     "origin": Origin.from_dict,
     "origin_visit": OriginVisit.from_dict,
     "origin_visit_status": OriginVisitStatus.from_dict,
     "snapshot": Snapshot.from_dict,
     "revision": Revision.from_dict,
     "release": Release.from_dict,
     "directory": Directory.from_dict,
     "content": Content.from_dict,
     "skipped_content": SkippedContent.from_dict,
     "metadata_authority": MetadataAuthority.from_dict,
     "metadata_fetcher": MetadataFetcher.from_dict,
     "raw_extrinsic_metadata": RawExtrinsicMetadata.from_dict,
     "extid": ExtID.from_dict,
 }
+# Deprecated, for BW compat only.
+object_converter_fn = OBJECT_CONVERTERS
+
+
+OBJECT_FIXERS = {
+    "revision": partial(remove_keys, keys=("metadata",)),
+}
+
+
+class ModelObjectDeserializer:
+
+    """A swh.journal object deserializer that checks object validity and reports
+    invalid objects
+
+    The deserializer will directly produce BaseModel objects from journal
+    objects representations.
+
+    If validation is activated and the object is hashable, it will check if the
+    computed hash matches the identifier of the object.
+
+    If the object is invalid and a 'reporter' function is given, it will be
+    called with 2 arguments::
+
+      reporter(object_id, journal_msg)
+
+    Where 'object_id' is a string representation of the object identifier (from
+    the journal message), and 'journal_msg' is the row message (bytes)
+    retrieved from the journal.
+
+    If 'raise_on_error' is True, a 'StorageArgumentException' exception is
+    raised.
+
+    Typical usage::
+
+      deserializer = ModelObjectDeserializer(validate=True, reporter=reporter_cb)
+      client = get_journal_client(
+          cls="kafka", value_deserializer=deserializer, **cfg)
+
+    """
+
+    def __init__(
+        self,
+        validate: bool = True,
+        raise_on_error: bool = False,
+        reporter: Optional[Callable[[str, bytes], None]] = None,
+    ):
+        self.validate = validate
+        self.reporter = reporter
+        self.raise_on_error = raise_on_error
+
+    def convert(self, object_type: str, msg: bytes) -> Optional[BaseModel]:
+        dict_repr = kafka_to_value(msg)
+        if object_type in OBJECT_FIXERS:
+            dict_repr = OBJECT_FIXERS[object_type](dict_repr)
+        obj = OBJECT_CONVERTERS[object_type](dict_repr)
+        if self.validate:
+            if isinstance(obj, HashableObject):
+                cid = obj.compute_hash()
+                if obj.id != cid:
+                    error_msg = (
+                        f"Object has id {hash_to_hex(obj.id)}, "
+                        f"but it should be {hash_to_hex(cid)}: {obj}"
+                    )
+                    logger.error(error_msg)
+                    self.report_failure(msg, obj)
+                    if self.raise_on_error:
+                        raise StorageArgumentException(error_msg)
+                    return None
+        return obj
+
+    def report_failure(self, msg: bytes, obj: BaseModel):
+        if self.reporter:
+            oid: str = ""
+            if hasattr(obj, "swhid"):
+                swhid = obj.swhid()  # type: ignore[attr-defined]
+                oid = str(swhid)
+            elif isinstance(obj, HashableObject):
+                uid = obj.compute_hash()
+                oid = f"{obj.object_type}:{uid.hex()}"  # type: ignore[attr-defined]
+            if oid:
+                self.reporter(oid, msg)
 
 
 def process_replay_objects(
-    all_objects: Dict[str, List[Dict[str, Any]]], *, storage: StorageInterface
+    all_objects: Dict[str, List[BaseModel]], *, storage: StorageInterface
 ) -> None:
     for (object_type, objects) in all_objects.items():
         logger.debug("Inserting %s %s objects", len(objects), object_type)
         with statsd.timed(GRAPH_DURATION_METRIC, tags={"object_type": object_type}):
             _insert_objects(object_type, objects, storage)
         statsd.increment(
             GRAPH_OPERATIONS_METRIC, len(objects), tags={"object_type": object_type}
         )
     if notify:
         notify("WATCHDOG=1")
 
 
+ContentType = TypeVar("ContentType", bound=BaseContent)
+
+
 def collision_aware_content_add(
-    content_add_fn: Callable[[List[Any]], Dict[str, int]], contents: List[BaseContent],
-) -> None:
+    contents: List[ContentType],
+    content_add_fn: Callable[[List[ContentType]], Dict[str, int]],
+) -> Dict[str, int]:
     """Add contents to storage. If a hash collision is detected, an error is
        logged. Then this adds the other non colliding contents to the storage.
 
     Args:
         content_add_fn: Storage content callable
         contents: List of contents or skipped contents to add to storage
 
     """
     if not contents:
-        return
+        return {}
     colliding_content_hashes: List[Dict[str, Any]] = []
+    results: CounterT[str] = Counter()
     while True:
         try:
-            content_add_fn(contents)
+            results.update(content_add_fn(contents))
         except HashCollision as e:
             colliding_content_hashes.append(
                 {
                     "algo": e.algo,
                     "hash": e.hash_id,  # hex hash id
                     "objects": e.colliding_contents,  # hex hashes
                 }
             )
             colliding_hashes = e.colliding_content_hashes()
             # Drop the colliding contents from the transaction
             contents = [c for c in contents if c.hashes() not in colliding_hashes]
         else:
             # Successfully added contents, we are done
             break
     if colliding_content_hashes:
         for collision in colliding_content_hashes:
             logger.error("Collision detected: %(collision)s", {"collision": collision})
-
-
-def dict_key_dropper(d: Dict, keys_to_drop: Container) -> Dict:
-    """Returns a copy of the dict d without any key listed in keys_to_drop"""
-    return {k: v for (k, v) in d.items() if k not in keys_to_drop}
+    return dict(results)
 
 
 def _insert_objects(
-    object_type: str, objects: List[Dict], storage: StorageInterface
+    object_type: str, objects: List[BaseModel], storage: StorageInterface
 ) -> None:
     """Insert objects of type object_type in the storage.
 
     """
-    objects = fix_objects(object_type, objects)
-
-    if object_type == "content":
-        # for bw compat, skipped content should now be delivered in the skipped_content
-        # topic
-        contents: List[BaseContent] = []
-        skipped_contents: List[BaseContent] = []
-        for content in objects:
-            c = BaseContent.from_dict(content)
-            if isinstance(c, SkippedContent):
-                logger.warning(
-                    "Received a series of skipped_content in the "
-                    "content topic, this should not happen anymore"
-                )
-                skipped_contents.append(c)
-            else:
-                contents.append(c)
-        collision_aware_content_add(storage.skipped_content_add, skipped_contents)
-        collision_aware_content_add(storage.content_add_metadata, contents)
-    elif object_type == "skipped_content":
-        skipped_contents = [SkippedContent.from_dict(obj) for obj in objects]
-        collision_aware_content_add(storage.skipped_content_add, skipped_contents)
+    if object_type not in OBJECT_CONVERTERS:
+        logger.warning("Received a series of %s, this should not happen", object_type)
+        return
+
+    method = getattr(storage, f"{object_type}_add")
+    if object_type == "skipped_content":
+        method = partial(collision_aware_content_add, content_add_fn=method)
+    elif object_type == "content":
+        method = partial(
+            collision_aware_content_add, content_add_fn=storage.content_add_metadata
+        )
     elif object_type in ("origin_visit", "origin_visit_status"):
         origins: List[Origin] = []
-        converter_fn = object_converter_fn[object_type]
-        model_objs = []
-        for obj in objects:
-            origins.append(Origin(url=obj["origin"]))
-            model_objs.append(converter_fn(obj))
+        for obj in cast(List[Union[OriginVisit, OriginVisitStatus]], objects):
+            origins.append(Origin(url=obj.origin))
         storage.origin_add(origins)
-        method = getattr(storage, f"{object_type}_add")
-        method(model_objs)
     elif object_type == "raw_extrinsic_metadata":
-        converted = [RawExtrinsicMetadata.from_dict(o) for o in objects]
-        authorities = {emd.authority for emd in converted}
-        fetchers = {emd.fetcher for emd in converted}
+        emds = cast(List[RawExtrinsicMetadata], objects)
+        authorities = {emd.authority for emd in emds}
+        fetchers = {emd.fetcher for emd in emds}
         storage.metadata_authority_add(list(authorities))
         storage.metadata_fetcher_add(list(fetchers))
-        storage.raw_extrinsic_metadata_add(converted)
-    elif object_type == "revision":
-        # drop the metadata field from the revision (is any); this field is
-        # about to be dropped from the data model (in favor of
-        # raw_extrinsic_metadata) and there can be bogus values in the existing
-        # journal (metadata with \0000 in it)
-        method = getattr(storage, object_type + "_add")
-        method(
-            [
-                object_converter_fn[object_type](dict_key_dropper(o, ("metadata",)))
-                for o in objects
-            ]
-        )
-    elif object_type in (
-        "directory",
-        "extid",
-        "revision",
-        "release",
-        "snapshot",
-        "origin",
-        "metadata_fetcher",
-        "metadata_authority",
-    ):
-        method = getattr(storage, object_type + "_add")
-        method([object_converter_fn[object_type](o) for o in objects])
-    else:
-        logger.warning("Received a series of %s, this should not happen", object_type)
+    method(objects)
diff --git a/swh/storage/tests/test_backfill.py b/swh/storage/tests/test_backfill.py
index b1a90520..787cad12 100644
--- a/swh/storage/tests/test_backfill.py
+++ b/swh/storage/tests/test_backfill.py
@@ -1,298 +1,301 @@
 # Copyright (C) 2019-2021 The Software Heritage developers
 # See the AUTHORS file at the top-level directory of this distribution
 # License: GNU General Public License version 3, or any later version
 # See top-level LICENSE file for more information
 
 import functools
 import logging
 from unittest.mock import patch
 
 import pytest
 
 from swh.journal.client import JournalClient
 from swh.model.tests.swh_model_data import TEST_OBJECTS
 from swh.storage import get_storage
 from swh.storage.backfill import (
     PARTITION_KEY,
     JournalBackfiller,
     byte_ranges,
     compute_query,
     raw_extrinsic_metadata_target_ranges,
 )
 from swh.storage.in_memory import InMemoryStorage
-from swh.storage.replay import process_replay_objects
+from swh.storage.replay import ModelObjectDeserializer, process_replay_objects
 from swh.storage.tests.test_replay import check_replayed
 
 TEST_CONFIG = {
     "journal_writer": {
         "brokers": ["localhost"],
         "prefix": "swh.tmp_journal.new",
         "client_id": "swh.journal.client.test",
     },
     "storage": {"cls": "postgresql", "db": "service=swh-dev"},
 }
 
 
 def test_config_ko_missing_mandatory_key():
     """Missing configuration key will make the initialization fail
 
     """
     for key in TEST_CONFIG.keys():
         config = TEST_CONFIG.copy()
         config.pop(key)
 
         with pytest.raises(ValueError) as e:
             JournalBackfiller(config)
 
         error = "Configuration error: The following keys must be provided: %s" % (
             ",".join([key]),
         )
         assert e.value.args[0] == error
 
 
 def test_config_ko_unknown_object_type():
     """Parse arguments will fail if the object type is unknown
 
     """
     backfiller = JournalBackfiller(TEST_CONFIG)
     with pytest.raises(ValueError) as e:
         backfiller.parse_arguments("unknown-object-type", 1, 2)
 
     error = (
         "Object type unknown-object-type is not supported. "
         "The only possible values are %s" % (", ".join(sorted(PARTITION_KEY)))
     )
     assert e.value.args[0] == error
 
 
 def test_compute_query_content():
     query, where_args, column_aliases = compute_query("content", "\x000000", "\x000001")
 
     assert where_args == ["\x000000", "\x000001"]
 
     assert column_aliases == [
         "sha1",
         "sha1_git",
         "sha256",
         "blake2s256",
         "length",
         "status",
         "ctime",
     ]
 
     assert (
         query
         == """
 select sha1,sha1_git,sha256,blake2s256,length,status,ctime
 from content
 
 where (sha1) >= %s and (sha1) < %s
     """
     )
 
 
 def test_compute_query_skipped_content():
     query, where_args, column_aliases = compute_query("skipped_content", None, None)
 
     assert where_args == []
 
     assert column_aliases == [
         "sha1",
         "sha1_git",
         "sha256",
         "blake2s256",
         "length",
         "ctime",
         "status",
         "reason",
     ]
 
     assert (
         query
         == """
 select sha1,sha1_git,sha256,blake2s256,length,ctime,status,reason
 from skipped_content
 
 
     """
     )
 
 
 def test_compute_query_origin_visit():
     query, where_args, column_aliases = compute_query("origin_visit", 1, 10)
 
     assert where_args == [1, 10]
 
     assert column_aliases == [
         "visit",
         "type",
         "origin",
         "date",
     ]
 
     assert (
         query
         == """
 select visit,type,origin.url as origin,date
 from origin_visit
 left join origin on origin_visit.origin=origin.id
 where (origin_visit.origin) >= %s and (origin_visit.origin) < %s
     """
     )
 
 
 def test_compute_query_release():
     query, where_args, column_aliases = compute_query("release", "\x000002", "\x000003")
 
     assert where_args == ["\x000002", "\x000003"]
 
     assert column_aliases == [
         "id",
         "date",
         "date_offset",
         "date_neg_utc_offset",
         "comment",
         "name",
         "synthetic",
         "target",
         "target_type",
         "author_id",
         "author_name",
         "author_email",
         "author_fullname",
     ]
 
     assert (
         query
         == """
 select release.id as id,date,date_offset,date_neg_utc_offset,comment,release.name as name,synthetic,target,target_type,a.id as author_id,a.name as author_name,a.email as author_email,a.fullname as author_fullname
 from release
 left join person a on release.author=a.id
 where (release.id) >= %s and (release.id) < %s
     """  # noqa
     )
 
 
 @pytest.mark.parametrize("numbits", [2, 3, 8, 16])
 def test_byte_ranges(numbits):
     ranges = list(byte_ranges(numbits))
 
     assert len(ranges) == 2 ** numbits
     assert ranges[0][0] is None
     assert ranges[-1][1] is None
 
     bounds = []
     for i, (left, right) in enumerate(zip(ranges[:-1], ranges[1:])):
         assert left[1] == right[0], f"Mismatched bounds in {i}th range"
         bounds.append(left[1])
 
     assert bounds == sorted(bounds)
 
 
 def test_raw_extrinsic_metadata_target_ranges():
     ranges = list(raw_extrinsic_metadata_target_ranges())
 
     assert ranges[0][0] == ""
     assert ranges[-1][1] is None
 
     bounds = []
     for i, (left, right) in enumerate(zip(ranges[:-1], ranges[1:])):
         assert left[1] == right[0], f"Mismatched bounds in {i}th range"
         bounds.append(left[1])
 
     assert bounds == sorted(bounds)
 
 
 RANGE_GENERATORS = {
     "content": lambda start, end: [(None, None)],
     "skipped_content": lambda start, end: [(None, None)],
     "directory": lambda start, end: [(None, None)],
     "extid": lambda start, end: [(None, None)],
     "metadata_authority": lambda start, end: [(None, None)],
     "metadata_fetcher": lambda start, end: [(None, None)],
     "revision": lambda start, end: [(None, None)],
     "release": lambda start, end: [(None, None)],
     "snapshot": lambda start, end: [(None, None)],
     "origin": lambda start, end: [(None, 10000)],
     "origin_visit": lambda start, end: [(None, 10000)],
     "origin_visit_status": lambda start, end: [(None, 10000)],
     "raw_extrinsic_metadata": lambda start, end: [(None, None)],
 }
 
 
 @patch("swh.storage.backfill.RANGE_GENERATORS", RANGE_GENERATORS)
 def test_backfiller(
     swh_storage_backend_config,
     kafka_prefix: str,
     kafka_consumer_group: str,
     kafka_server: str,
     caplog,
 ):
     prefix1 = f"{kafka_prefix}-1"
     prefix2 = f"{kafka_prefix}-2"
 
     journal1 = {
         "cls": "kafka",
         "brokers": [kafka_server],
         "client_id": "kafka_writer-1",
         "prefix": prefix1,
     }
     swh_storage_backend_config["journal_writer"] = journal1
     storage = get_storage(**swh_storage_backend_config)
-
     # fill the storage and the journal (under prefix1)
     for object_type, objects in TEST_OBJECTS.items():
         method = getattr(storage, object_type + "_add")
         method(objects)
 
     # now apply the backfiller on the storage to fill the journal under prefix2
     backfiller_config = {
         "journal_writer": {
             "brokers": [kafka_server],
             "client_id": "kafka_writer-2",
             "prefix": prefix2,
         },
         "storage": swh_storage_backend_config,
     }
 
     # Backfilling
     backfiller = JournalBackfiller(backfiller_config)
     for object_type in TEST_OBJECTS:
         backfiller.run(object_type, None, None)
 
     # Trace log messages for unhandled object types in the replayer
     caplog.set_level(logging.DEBUG, "swh.storage.replay")
 
     # now check journal content are the same under both topics
     # use the replayer scaffolding to fill storages to make is a bit easier
     # Replaying #1
+    deserializer = ModelObjectDeserializer()
     sto1 = get_storage(cls="memory")
     replayer1 = JournalClient(
         brokers=kafka_server,
         group_id=f"{kafka_consumer_group}-1",
         prefix=prefix1,
         stop_on_eof=True,
+        value_deserializer=deserializer.convert,
     )
+
     worker_fn1 = functools.partial(process_replay_objects, storage=sto1)
     replayer1.process(worker_fn1)
 
     # Replaying #2
     sto2 = get_storage(cls="memory")
     replayer2 = JournalClient(
         brokers=kafka_server,
         group_id=f"{kafka_consumer_group}-2",
         prefix=prefix2,
         stop_on_eof=True,
+        value_deserializer=deserializer.convert,
     )
     worker_fn2 = functools.partial(process_replay_objects, storage=sto2)
     replayer2.process(worker_fn2)
 
     # Compare storages
     assert isinstance(sto1, InMemoryStorage)  # needed to help mypy
     assert isinstance(sto2, InMemoryStorage)
     check_replayed(sto1, sto2)
 
     for record in caplog.records:
         assert (
             "this should not happen" not in record.message
         ), "Replayer ignored some message types, see captured logging"
diff --git a/swh/storage/tests/test_cli.py b/swh/storage/tests/test_cli.py
index 87121b93..f472dbf1 100644
--- a/swh/storage/tests/test_cli.py
+++ b/swh/storage/tests/test_cli.py
@@ -1,111 +1,125 @@
 # Copyright (C) 2020  The Software Heritage developers
 # See the AUTHORS file at the top-level directory of this distribution
 # License: GNU General Public License version 3, or any later version
 # See top-level LICENSE file for more information
 
 import copy
 import logging
 import re
 import tempfile
 from unittest.mock import patch
 
 from click.testing import CliRunner
 from confluent_kafka import Producer
 import pytest
 import yaml
 
 from swh.journal.serializers import key_to_kafka, value_to_kafka
 from swh.model.model import Snapshot, SnapshotBranch, TargetType
 from swh.storage import get_storage
 from swh.storage.cli import storage as cli
+from swh.storage.replay import OBJECT_CONVERTERS
 
 logger = logging.getLogger(__name__)
 
 
 CLI_CONFIG = {
     "storage": {"cls": "memory",},
 }
 
 
 @pytest.fixture
 def swh_storage():
     """An swh-storage object that gets injected into the CLI functions."""
     storage = get_storage(**CLI_CONFIG["storage"])
     with patch("swh.storage.get_storage") as get_storage_mock:
         get_storage_mock.return_value = storage
         yield storage
 
 
 @pytest.fixture
 def monkeypatch_retry_sleep(monkeypatch):
     from swh.journal.replay import copy_object, obj_in_objstorage
 
     monkeypatch.setattr(copy_object.retry, "sleep", lambda x: None)
     monkeypatch.setattr(obj_in_objstorage.retry, "sleep", lambda x: None)
 
 
 def invoke(*args, env=None, journal_config=None):
     config = copy.deepcopy(CLI_CONFIG)
     if journal_config:
         config["journal_client"] = journal_config.copy()
         config["journal_client"]["cls"] = "kafka"
 
     runner = CliRunner()
     with tempfile.NamedTemporaryFile("a", suffix=".yml") as config_fd:
         yaml.dump(config, config_fd)
         config_fd.seek(0)
         args = ["-C" + config_fd.name] + list(args)
         ret = runner.invoke(cli, args, obj={"log_level": logging.DEBUG}, env=env,)
         return ret
 
 
 def test_replay(
     swh_storage, kafka_prefix: str, kafka_consumer_group: str, kafka_server: str,
 ):
     kafka_prefix += ".swh.journal.objects"
 
     producer = Producer(
         {
             "bootstrap.servers": kafka_server,
             "client.id": "test-producer",
             "acks": "all",
         }
     )
 
     snapshot = Snapshot(
         branches={
             b"HEAD": SnapshotBranch(
                 target_type=TargetType.REVISION, target=b"\x01" * 20,
             )
         },
     )
     snapshot_dict = snapshot.to_dict()
 
     producer.produce(
         topic=kafka_prefix + ".snapshot",
         key=key_to_kafka(snapshot.id),
         value=value_to_kafka(snapshot_dict),
     )
     producer.flush()
 
     logger.debug("Flushed producer")
 
     result = invoke(
         "replay",
         "--stop-after-objects",
         "1",
         journal_config={
             "brokers": [kafka_server],
             "group_id": kafka_consumer_group,
             "prefix": kafka_prefix,
         },
     )
 
     expected = r"Done.\n"
     assert result.exit_code == 0, result.output
     assert re.fullmatch(expected, result.output, re.MULTILINE), result.output
 
     assert swh_storage.snapshot_get(snapshot.id) == {
         **snapshot_dict,
         "next_branch": None,
     }
+
+
+def test_replay_type_list():
+    result = invoke("replay", "--help",)
+    assert result.exit_code == 0, result.output
+    types_in_help = re.findall("--type [[]([a-z_|]+)[]]", result.output)
+    assert len(types_in_help) == 1
+    types = types_in_help[0].split("|")
+
+    assert sorted(types) == sorted(list(OBJECT_CONVERTERS.keys())), (
+        "Make sure the list of accepted types in cli.py "
+        "matches implementation in replay.py"
+    )
diff --git a/swh/storage/tests/test_replay.py b/swh/storage/tests/test_replay.py
index 6562f2aa..8c13fa95 100644
--- a/swh/storage/tests/test_replay.py
+++ b/swh/storage/tests/test_replay.py
@@ -1,375 +1,544 @@
 # Copyright (C) 2019-2020 The Software Heritage developers
 # See the AUTHORS file at the top-level directory of this distribution
 # License: GNU General Public License version 3, or any later version
 # See top-level LICENSE file for more information
 
 import dataclasses
 import datetime
 import functools
 import logging
-from typing import Any, Container, Dict, Optional
+import re
+from typing import Any, Container, Dict, Optional, cast
 
 import attr
 import pytest
 
 from swh.journal.client import JournalClient
-from swh.journal.serializers import key_to_kafka, value_to_kafka
+from swh.journal.serializers import kafka_to_value, key_to_kafka, value_to_kafka
 from swh.model.hashutil import DEFAULT_ALGORITHMS, MultiHash, hash_to_bytes, hash_to_hex
 from swh.model.model import Revision, RevisionType
 from swh.model.tests.swh_model_data import (
     COMMITTERS,
     DATES,
     DUPLICATE_CONTENTS,
     REVISIONS,
 )
 from swh.model.tests.swh_model_data import TEST_OBJECTS as _TEST_OBJECTS
 from swh.storage import get_storage
 from swh.storage.cassandra.model import ContentRow, SkippedContentRow
+from swh.storage.exc import StorageArgumentException
 from swh.storage.in_memory import InMemoryStorage
-from swh.storage.replay import process_replay_objects
+from swh.storage.replay import ModelObjectDeserializer, process_replay_objects
 
 UTC = datetime.timezone.utc
 
 TEST_OBJECTS = _TEST_OBJECTS.copy()
+# add a revision with metadata to check this later is dropped while being replayed
 TEST_OBJECTS["revision"] = list(_TEST_OBJECTS["revision"]) + [
     Revision(
-        id=hash_to_bytes("a569b03ebe6e5f9f2f6077355c40d89bd6986d0c"),
+        id=hash_to_bytes("51d9d94ab08d3f75512e3a9fd15132e0a7ca7928"),
         message=b"hello again",
         date=DATES[1],
         committer=COMMITTERS[1],
         author=COMMITTERS[0],
         committer_date=DATES[0],
         type=RevisionType.GIT,
         directory=b"\x03" * 20,
         synthetic=False,
         metadata={"something": "interesting"},
         parents=(REVISIONS[0].id,),
     ),
 ]
+WRONG_ID_REG = re.compile(
+    "Object has id [0-9a-f]{40}, but it should be [0-9a-f]{40}: .*"
+)
 
 
 def nullify_ctime(obj):
     if isinstance(obj, (ContentRow, SkippedContentRow)):
         return dataclasses.replace(obj, ctime=None)
     else:
         return obj
 
 
 @pytest.fixture()
 def replayer_storage_and_client(
     kafka_prefix: str, kafka_consumer_group: str, kafka_server: str
 ):
     journal_writer_config = {
         "cls": "kafka",
         "brokers": [kafka_server],
         "client_id": "kafka_writer",
         "prefix": kafka_prefix,
     }
     storage_config: Dict[str, Any] = {
         "cls": "memory",
         "journal_writer": journal_writer_config,
     }
     storage = get_storage(**storage_config)
+    deserializer = ModelObjectDeserializer()
     replayer = JournalClient(
         brokers=kafka_server,
         group_id=kafka_consumer_group,
         prefix=kafka_prefix,
         stop_on_eof=True,
+        value_deserializer=deserializer.convert,
     )
 
     yield storage, replayer
 
 
 def test_storage_replayer(replayer_storage_and_client, caplog):
     """Optimal replayer scenario.
 
     This:
     - writes objects to a source storage
     - replayer consumes objects from the topic and replays them
     - a destination storage is filled from this
 
     In the end, both storages should have the same content.
     """
     src, replayer = replayer_storage_and_client
 
     # Fill Kafka using a source storage
     nb_sent = 0
     for object_type, objects in TEST_OBJECTS.items():
         method = getattr(src, object_type + "_add")
         method(objects)
         if object_type == "origin_visit":
             nb_sent += len(objects)  # origin-visit-add adds origin-visit-status as well
         nb_sent += len(objects)
 
     caplog.set_level(logging.ERROR, "swh.journal.replay")
 
     # Fill the destination storage from Kafka
     dst = get_storage(cls="memory")
     worker_fn = functools.partial(process_replay_objects, storage=dst)
     nb_inserted = replayer.process(worker_fn)
     assert nb_sent == nb_inserted
 
     assert isinstance(src, InMemoryStorage)  # needed to help mypy
     assert isinstance(dst, InMemoryStorage)
     check_replayed(src, dst)
 
     collision = 0
     for record in caplog.records:
         logtext = record.getMessage()
         if "Colliding contents:" in logtext:
             collision += 1
 
     assert collision == 0, "No collision should be detected"
 
 
-def test_storage_play_with_collision(replayer_storage_and_client, caplog):
+def test_storage_replay_with_collision(replayer_storage_and_client, caplog):
     """Another replayer scenario with collisions.
 
     This:
     - writes objects to the topic, including colliding contents
     - replayer consumes objects from the topic and replay them
     - This drops the colliding contents from the replay when detected
 
     """
     src, replayer = replayer_storage_and_client
 
     # Fill Kafka using a source storage
     nb_sent = 0
     for object_type, objects in TEST_OBJECTS.items():
         method = getattr(src, object_type + "_add")
         method(objects)
         if object_type == "origin_visit":
             nb_sent += len(objects)  # origin-visit-add adds origin-visit-status as well
         nb_sent += len(objects)
 
     # Create collision in input data
     # These should not be written in the destination
     producer = src.journal_writer.journal.producer
     prefix = src.journal_writer.journal._prefix
     for content in DUPLICATE_CONTENTS:
         topic = f"{prefix}.content"
         key = content.sha1
         now = datetime.datetime.now(tz=UTC)
         content = attr.evolve(content, ctime=now)
         producer.produce(
             topic=topic, key=key_to_kafka(key), value=value_to_kafka(content.to_dict()),
         )
         nb_sent += 1
 
     producer.flush()
 
     caplog.set_level(logging.ERROR, "swh.journal.replay")
 
     # Fill the destination storage from Kafka
     dst = get_storage(cls="memory")
     worker_fn = functools.partial(process_replay_objects, storage=dst)
     nb_inserted = replayer.process(worker_fn)
     assert nb_sent == nb_inserted
 
     # check the logs for the collision being properly detected
     nb_collisions = 0
     actual_collision: Dict
     for record in caplog.records:
         logtext = record.getMessage()
         if "Collision detected:" in logtext:
             nb_collisions += 1
             actual_collision = record.args["collision"]
 
     assert nb_collisions == 1, "1 collision should be detected"
 
     algo = "sha1"
     assert actual_collision["algo"] == algo
     expected_colliding_hash = hash_to_hex(DUPLICATE_CONTENTS[0].get_hash(algo))
     assert actual_collision["hash"] == expected_colliding_hash
 
     actual_colliding_hashes = actual_collision["objects"]
     assert len(actual_colliding_hashes) == len(DUPLICATE_CONTENTS)
     for content in DUPLICATE_CONTENTS:
         expected_content_hashes = {
             k: hash_to_hex(v) for k, v in content.hashes().items()
         }
         assert expected_content_hashes in actual_colliding_hashes
 
     # all objects from the src should exists in the dst storage
     assert isinstance(src, InMemoryStorage)  # needed to help mypy
     assert isinstance(dst, InMemoryStorage)  # needed to help mypy
     check_replayed(src, dst, exclude=["contents"])
     # but the dst has one content more (one of the 2 colliding ones)
     assert (
         len(list(src._cql_runner._contents.iter_all()))
         == len(list(dst._cql_runner._contents.iter_all())) - 1
     )
 
 
 def test_replay_skipped_content(replayer_storage_and_client):
     """Test the 'skipped_content' topic is properly replayed."""
     src, replayer = replayer_storage_and_client
     _check_replay_skipped_content(src, replayer, "skipped_content")
 
 
-def test_replay_skipped_content_bwcompat(replayer_storage_and_client):
-    """Test the 'content' topic can be used to replay SkippedContent objects."""
-    src, replayer = replayer_storage_and_client
-    _check_replay_skipped_content(src, replayer, "content")
-
-
 # utility functions
 
 
 def check_replayed(
     src: InMemoryStorage,
     dst: InMemoryStorage,
     exclude: Optional[Container] = None,
     expected_anonymized=False,
 ):
     """Simple utility function to compare the content of 2 in_memory storages"""
 
     def fix_expected(attr, row):
         if expected_anonymized:
             if attr == "releases":
                 row = dataclasses.replace(
                     row, author=row.author and row.author.anonymize()
                 )
             elif attr == "revisions":
                 row = dataclasses.replace(
                     row,
                     author=row.author.anonymize(),
                     committer=row.committer.anonymize(),
                 )
         if attr == "revisions":
             # the replayer should now drop the metadata attribute; see
             # swh/storgae/replay.py:_insert_objects()
             row.metadata = "null"
 
         return row
 
     for attr_ in (
         "contents",
         "skipped_contents",
         "directories",
         "extid",
         "revisions",
         "releases",
         "snapshots",
         "origins",
         "origin_visits",
         "origin_visit_statuses",
         "raw_extrinsic_metadata",
     ):
         if exclude and attr_ in exclude:
             continue
         expected_objects = [
             (id, nullify_ctime(fix_expected(attr_, obj)))
             for id, obj in sorted(getattr(src._cql_runner, f"_{attr_}").iter_all())
         ]
         got_objects = [
             (id, nullify_ctime(obj))
             for id, obj in sorted(getattr(dst._cql_runner, f"_{attr_}").iter_all())
         ]
         assert got_objects == expected_objects, f"Mismatch object list for {attr_}"
 
 
 def _check_replay_skipped_content(storage, replayer, topic):
     skipped_contents = _gen_skipped_contents(100)
     nb_sent = len(skipped_contents)
     producer = storage.journal_writer.journal.producer
     prefix = storage.journal_writer.journal._prefix
 
     for i, obj in enumerate(skipped_contents):
         producer.produce(
             topic=f"{prefix}.{topic}",
             key=key_to_kafka({"sha1": obj["sha1"]}),
             value=value_to_kafka(obj),
         )
     producer.flush()
 
     dst_storage = get_storage(cls="memory")
     worker_fn = functools.partial(process_replay_objects, storage=dst_storage)
     nb_inserted = replayer.process(worker_fn)
 
     assert nb_sent == nb_inserted
     for content in skipped_contents:
         assert not storage.content_find({"sha1": content["sha1"]})
 
     # no skipped_content_find API endpoint, so use this instead
     assert not list(dst_storage.skipped_content_missing(skipped_contents))
 
 
 def _updated(d1, d2):
     d1.update(d2)
     d1.pop("data", None)
     return d1
 
 
 def _gen_skipped_contents(n=10):
     # we do not use the hypothesis strategy here because this does not play well with
     # pytest fixtures, and it makes test execution very slow
     algos = DEFAULT_ALGORITHMS | {"length"}
     now = datetime.datetime.now(tz=UTC)
     return [
         _updated(
             MultiHash.from_data(data=f"foo{i}".encode(), hash_names=algos).digest(),
             {
                 "status": "absent",
                 "reason": "why not",
                 "origin": f"https://somewhere/{i}",
                 "ctime": now,
             },
         )
         for i in range(n)
     ]
 
 
 @pytest.mark.parametrize("privileged", [True, False])
-def test_storage_play_anonymized(
+def test_storage_replay_anonymized(
     kafka_prefix: str, kafka_consumer_group: str, kafka_server: str, privileged: bool,
 ):
     """Optimal replayer scenario.
 
     This:
     - writes objects to the topic
     - replayer consumes objects from the topic and replay them
 
     This tests the behavior with both a privileged and non-privileged replayer
     """
     writer_config = {
         "cls": "kafka",
         "brokers": [kafka_server],
         "client_id": "kafka_writer",
         "prefix": kafka_prefix,
         "anonymize": True,
     }
     src_config: Dict[str, Any] = {"cls": "memory", "journal_writer": writer_config}
 
     storage = get_storage(**src_config)
 
     # Fill the src storage
     nb_sent = 0
     for obj_type, objs in TEST_OBJECTS.items():
         if obj_type in ("origin_visit", "origin_visit_status"):
             # these are unrelated with what we want to test here
             continue
         method = getattr(storage, obj_type + "_add")
         method(objs)
         nb_sent += len(objs)
 
     # Fill a destination storage from Kafka, potentially using privileged topics
     dst_storage = get_storage(cls="memory")
+    deserializer = ModelObjectDeserializer(
+        validate=False
+    )  # we cannot validate an anonymized replay
     replayer = JournalClient(
         brokers=kafka_server,
         group_id=kafka_consumer_group,
         prefix=kafka_prefix,
         stop_after_objects=nb_sent,
         privileged=privileged,
+        value_deserializer=deserializer.convert,
     )
     worker_fn = functools.partial(process_replay_objects, storage=dst_storage)
 
     nb_inserted = replayer.process(worker_fn)
     replayer.consumer.commit()
 
     assert nb_sent == nb_inserted
     # Check the contents of the destination storage, and whether the anonymization was
     # properly used
     assert isinstance(storage, InMemoryStorage)  # needed to help mypy
     assert isinstance(dst_storage, InMemoryStorage)
     check_replayed(storage, dst_storage, expected_anonymized=not privileged)
+
+
+def test_storage_replayer_with_validation_ok(
+    replayer_storage_and_client, caplog, redisdb
+):
+    """Optimal replayer scenario
+
+    with validation activated and reporter set to a redis db.
+
+    - writes objects to a source storage
+    - replayer consumes objects from the topic and replays them
+    - a destination storage is filled from this
+    - nothing has been reported in the redis db
+    - both storages should have the same content
+    """
+    src, replayer = replayer_storage_and_client
+    replayer.deserializer = ModelObjectDeserializer(validate=True, reporter=redisdb.set)
+
+    # Fill Kafka using a source storage
+    nb_sent = 0
+    for object_type, objects in TEST_OBJECTS.items():
+        method = getattr(src, object_type + "_add")
+        method(objects)
+        if object_type == "origin_visit":
+            nb_sent += len(objects)  # origin-visit-add adds origin-visit-status as well
+        nb_sent += len(objects)
+
+    # Fill the destination storage from Kafka
+    dst = get_storage(cls="memory")
+    worker_fn = functools.partial(process_replay_objects, storage=dst)
+    nb_inserted = replayer.process(worker_fn)
+    assert nb_sent == nb_inserted
+
+    # check we do not have invalid objects reported
+    invalid = 0
+    for record in caplog.records:
+        logtext = record.getMessage()
+        if WRONG_ID_REG.match(logtext):
+            invalid += 1
+    assert invalid == 0, "Invalid objects should not be detected"
+    assert not redisdb.keys()
+    # so the dst should be the same as src storage
+    check_replayed(cast(InMemoryStorage, src), cast(InMemoryStorage, dst))
+
+
+def test_storage_replayer_with_validation_nok(
+    replayer_storage_and_client, caplog, redisdb
+):
+    """Replayer scenario with invalid objects
+
+    with validation and reporter set to a redis db.
+
+    - writes objects to a source storage
+    - replayer consumes objects from the topic and replays them
+    - the destination storage is filled with only valid objects
+    - the redis db contains the invalid (raw kafka mesg) objects
+    """
+    src, replayer = replayer_storage_and_client
+    replayer.value_deserializer = ModelObjectDeserializer(
+        validate=True, reporter=redisdb.set
+    ).convert
+
+    caplog.set_level(logging.ERROR, "swh.journal.replay")
+
+    # Fill Kafka using a source storage
+    nb_sent = 0
+    for object_type, objects in TEST_OBJECTS.items():
+        method = getattr(src, object_type + "_add")
+        method(objects)
+        if object_type == "origin_visit":
+            nb_sent += len(objects)  # origin-visit-add adds origin-visit-status as well
+        nb_sent += len(objects)
+
+    # insert invalid objects
+    for object_type in ("revision", "directory", "release", "snapshot"):
+        method = getattr(src, object_type + "_add")
+        method([attr.evolve(TEST_OBJECTS[object_type][0], id=b"\x00" * 20)])
+        nb_sent += 1
+
+    # Fill the destination storage from Kafka
+    dst = get_storage(cls="memory")
+    worker_fn = functools.partial(process_replay_objects, storage=dst)
+    nb_inserted = replayer.process(worker_fn)
+    assert nb_sent == nb_inserted
+
+    # check we do have invalid objects reported
+    invalid = 0
+    for record in caplog.records:
+        logtext = record.getMessage()
+        if WRONG_ID_REG.match(logtext):
+            invalid += 1
+    assert invalid == 4, "Invalid objects should be detected"
+    assert set(redisdb.keys()) == {
+        f"swh:1:{typ}:{'0'*40}".encode() for typ in ("rel", "rev", "snp", "dir")
+    }
+
+    for key in redisdb.keys():
+        # check the stored value looks right
+        rawvalue = redisdb.get(key)
+        value = kafka_to_value(rawvalue)
+        assert isinstance(value, dict)
+        assert "id" in value
+        assert value["id"] == b"\x00" * 20
+
+    # check that invalid objects did not reach the dst storage
+    for attr_ in (
+        "directories",
+        "revisions",
+        "releases",
+        "snapshots",
+    ):
+        for id, obj in sorted(getattr(dst._cql_runner, f"_{attr_}").iter_all()):
+            assert id != b"\x00" * 20
+
+
+def test_storage_replayer_with_validation_nok_raises(
+    replayer_storage_and_client, caplog, redisdb
+):
+    """Replayer scenario with invalid objects
+
+    with raise_on_error set to True
+
+    This:
+    - writes both valid & invalid objects to a source storage
+    - a StorageArgumentException should be raised while replayer consumes
+      objects from the topic and replays them
+    """
+    src, replayer = replayer_storage_and_client
+    replayer.value_deserializer = ModelObjectDeserializer(
+        validate=True, reporter=redisdb.set, raise_on_error=True
+    ).convert
+
+    caplog.set_level(logging.ERROR, "swh.journal.replay")
+
+    # Fill Kafka using a source storage
+    nb_sent = 0
+    for object_type, objects in TEST_OBJECTS.items():
+        method = getattr(src, object_type + "_add")
+        method(objects)
+        if object_type == "origin_visit":
+            nb_sent += len(objects)  # origin-visit-add adds origin-visit-status as well
+        nb_sent += len(objects)
+
+    # insert invalid objects
+    for object_type in ("revision", "directory", "release", "snapshot"):
+        method = getattr(src, object_type + "_add")
+        method([attr.evolve(TEST_OBJECTS[object_type][0], id=b"\x00" * 20)])
+        nb_sent += 1
+
+    # Fill the destination storage from Kafka
+    dst = get_storage(cls="memory")
+    worker_fn = functools.partial(process_replay_objects, storage=dst)
+    with pytest.raises(StorageArgumentException):
+        replayer.process(worker_fn)
+
+    # check we do have invalid objects reported
+    invalid = 0
+    for record in caplog.records:
+        logtext = record.getMessage()
+        if WRONG_ID_REG.match(logtext):
+            invalid += 1
+    assert invalid == 1, "One invalid objects should be detected"
+    assert len(redisdb.keys()) == 1