The project is divided into four main aspects that concern the large-scale collection of data:
-
An infrastructure (hardware and software) to store, share and process millions of annotated pathology images that can be gigabytes each.
-
Legal and ethical requirements to ensure adequate usage of data while fully respecting patients’ privacy and data confidentiality.
-
An initial set of 3 million digital pathology slides collected and stored into the repository to provide data for the development of AI tools.
-
Functionalities to aid the use of the repository as well as the processing of images for diagnostic and research purposes.
Learn more on how our repository will work.