What is big data desensitization?
What is big data desensitization?
Big data desensitization, also known as data bleaching, data deprivatization or data deformation, refers to the deformation of certain sensitive information through desensitization rules to achieve reliable protection of sensitive private data. to safely use desensitized real data sets in development, testing and other non-production and outsourced environments.
Privacy data desensitization technology
Usually in big data platforms, data is stored in a structured format. Each table consists of many rows, and each row of data has Composed of many columns. According to the data attributes of the column, data columns can usually be divided into the following types:
Columns that can accurately locate a person are called identifiable columns, such as ID number, address, and Name etc.
A single column cannot locate an individual, but multiple columns of information can be used to potentially identify a person. These columns are called semi-identifying columns, such as postal code, birthday and gender. A US research paper claims that 87% of Americans can be identified using only postal code, birthday and gender information.
Columns containing sensitive user information, such as transaction amounts, illnesses, and income.
Other columns that do not contain user sensitive information.
Privacy data leakage types
Privacy data leakage can be divided into many types. Depending on the type, different privacy data can usually be used Leakage risk model to measure the risk of preventing privacy data leakage and desensitize data corresponding to different data desensitization algorithms. Generally speaking, types of privacy data leaks include:
Personal identity leakage. When a data user confirms through any means that a piece of data in a data table belongs to a certain person, it is called a personal identity leak. Personal identity leakage is the most serious, because once personal identity leakage occurs, data users can obtain sensitive information about specific individuals.
Attribute leakage, when data users learn new attribute information about a person based on the data table they access, it is called attribute leakage. Personal identity leakage will certainly lead to attribute leakage, but attribute leakage can also occur independently.
Member relationship leaked. When a data user can confirm that a person's data exists in a data table, it is called membership disclosure. The risk of membership relationship leakage is relatively small. Personal identity leakage and attribute leakage definitely mean membership relationship leakage, but membership relationship leakage may also occur independently.
Recommended tutorial: "PHP"
The above is the detailed content of What is big data desensitization?. For more information, please follow other related articles on the PHP Chinese website!

Hot AI Tools

Undresser.AI Undress
AI-powered app for creating realistic nude photos

AI Clothes Remover
Online AI tool for removing clothes from photos.

Undress AI Tool
Undress images for free

Clothoff.io
AI clothes remover

AI Hentai Generator
Generate AI Hentai for free.

Hot Article

Hot Tools

Notepad++7.3.1
Easy-to-use and free code editor

SublimeText3 Chinese version
Chinese version, very easy to use

Zend Studio 13.0.1
Powerful PHP integrated development environment

Dreamweaver CS6
Visual web development tools

SublimeText3 Mac version
God-level code editing software (SublimeText3)

Hot Topics

In this chapter, we will understand the Environment Variables, General Configuration, Database Configuration and Email Configuration in CakePHP.

PHP 8.4 brings several new features, security improvements, and performance improvements with healthy amounts of feature deprecations and removals. This guide explains how to install PHP 8.4 or upgrade to PHP 8.4 on Ubuntu, Debian, or their derivati

To work with date and time in cakephp4, we are going to make use of the available FrozenTime class.

To work on file upload we are going to use the form helper. Here, is an example for file upload.

In this chapter, we are going to learn the following topics related to routing ?

CakePHP is an open-source framework for PHP. It is intended to make developing, deploying and maintaining applications much easier. CakePHP is based on a MVC-like architecture that is both powerful and easy to grasp. Models, Views, and Controllers gu

Validator can be created by adding the following two lines in the controller.

Visual Studio Code, also known as VS Code, is a free source code editor — or integrated development environment (IDE) — available for all major operating systems. With a large collection of extensions for many programming languages, VS Code can be c