PII data field identification and mapping tools

Job ID: 36979560

Budget: $30 – $250 USD

Data inventory and data mapping are processes of identifying and documenting the types of personal
data that Client collects, processes, and stores, as well as their locations and flows. Data inventory involves
creating an inventory of all the data assets and categorizing them based on their nature, sensitivity, and
purpose of processing. Data mapping, on the other hand, involves identifying the data flows and
relationships between different data elements, systems, and processes. Together, data inventory and data
mapping help Client to better understand our data landscape, assess the risks associated with data
processing, and ensure compliance with data protection regulations.
Solution Requirements:
 Facilitate the classification of data
 Facilitate automated data dictionary
 Allow users to detail technical lineage with automatically populated details, with ability to export the
diagram in a readable format (e.g. PDF)
 Connect to varied platforms to include Oracle, MySQL , MS SQL , PostgreSQL, and Mongo DB
 Utilize AI or machine learning to understand patterns in structured and unstructured documents to
identify PII
 Enable customizable classification and categorization
 Integrate with major cloud service providers, such as Amazon Web Services (AWS), Microsoft Azure,
or Google Cloud Platform
 Support seamless data discovery and mapping for cloud-based data sources and services, including
Software-as-a-Service (SaaS), Platform-as-a-Service (PaaS), and Infrastructure-as-a-Service (IaaS)
offerings
 Monitor and track data movement, transformations, and data access within cloud and hybrid cloud
environments, capturing relevant metadata for data governance and audit purposes



Data Classification and Cataloguing is the process of organizing and labeling data according to its level of
sensitivity and criticality, as well as its format and location. It involves identifying and classifying
structured and unstructured data, such as personal data, financial data, or confidential business
information, and creating a catalog of all the data assets within Client. This helps Client better understand
the nature and value of our data, and enables us to establish appropriate data protection controls and
access policies based on the level of sensitivity of the data.
Solution Requirements:
 Capability to automatically classify structured and unstructured data via appropriate technology (e.g.
Machine Learning) to categorize and tag data. Please specify the technology
 Enable automatic and classification of data when necessary, with the option to add custom categories
and metadata to enhance data classification and cataloguing
 Enable manual classification of data when necessary, with the option to add custom categories and
metadata to enhance data classification and cataloguing
 Provide a centralized catalog of all data assets, with the ability to search and filter data by attributes
such as sensitivity level, location, and data type
 Capability to set retention policies for classified data, and generate alert when the retention period is
approaching and reached
 Provide appropriate access controls for classified data, ensuring that only authorized personnel can
access sensitive data
 seamlessly integrate with major cloud service providers, such as Amazon Web Services (AWS),
Microsoft Azure, Google Cloud Platform, and hybrid cloud environments
 Support the discovery and cataloguing of cloud-based data sources, including Software-as-a-Service
(SaaS), Platform-as-a-Service (PaaS), and Infrastructure-as-a-Service (IaaS) offerings
 Provide connectors or APIs to extract metadata and relevant information from cloud-based data
repositories and systems for inclusion in the data catalog
 Mechanisms to synchronize and update the data catalog as new data sources are added, modified, or
removed in cloud and hybrid cloud environments
Related categories: SQL Oracle MySQL Database Administration SQLite