Validating language experience approach

For example, harmonization of short codes (st, rd, etc.) to actual words (street, road, etcetera).

Standardization of data is a means of changing a reference data set to a new standard, ex, use of standard codes.

(For example, "referential integrity" is a term used to refer to the enforcement of foreign-key constraints above.) Good quality source data has to do with “Data Quality Culture” and must be initiated at the top of the organization.



For instance, if the addresses are inconsistent, the company will suffer the cost of resending mail or even losing customers.

The inconsistencies detected or removed may have been originally caused by user entry errors, by corruption in transmission or storage, or by different data dictionary definitions of similar entities in different stores.

Data cleaning differs from data validation in that validation almost invariably means data is rejected from the system at entry and is performed at the time of entry, rather than on batches of data.

Integrating such validation into an authoring tool looks, shall we say, challenging.

Data cleansing or data cleaning is the process of detecting and correcting (or removing) corrupt or inaccurate records from a record set, table, or database and refers to identifying incomplete, incorrect, inaccurate or irrelevant parts of the data and then replacing, modifying, or deleting the dirty or coarse data.

