> ## Knowledge Base Index
> Fetch the complete knowledge base index at: https://pipecorn.crisp.help/sitemap.xml
> Use this file to discover available pages before exploring further.
> Pure-Markdown content can be obtained by appending a '.md' suffix to the content URLs listed in the sitemap (without the trailing slash).

# How Does Data Cleaning Works?

For every extraction, Pronto cleans:  

* First Names
* Last Names
* Company Names

By cleaning, we mean:  

* Deleting emojis
* Correcting typos
* Normalizing names (capital letters)
  
Here are a few examples:  
  


[![](https://storage.crisp.chat/users/helpdesk/website/ed5e34ca65d2f800/38d46de0-f6fa-4d6d-865d-a8ef9f_6y3jae.png)](https://downloads.intercomcdn.com/i/o/1183967612/e6c19738bc7f319875d50460/cleanshot-2022-06-24-at-092615_1eruk70.png?expires=1738161000&signature=a9586007b5de31aec1a81aeb4df9e7333f6a196db40d52cddd52538536762368&req=dSEvFcB4modeW%2FMW1HO4zexArNORYz7uAhI3ZCqnUd2tUFIfDUg76RZEIeRA%0AHcicqGubpp2tQ3sZ%2BhY%3D%0A)