* Cantinho Satkeys

Refresh History
  • JP: try65hytr A Todos  4tj97u<z 2dgh8i k7y8j0 r4v8p
    12 de Junho de 2026, 05:28
  • JP: try65hytr Pessoal  2dgh8i k7y8j0 yu7gh8
    10 de Junho de 2026, 03:47
  • j.s.: passem por aqui [link]
    09 de Junho de 2026, 20:57
  • j.s.: um anonimo contribuiu com €10,00  h7t45
    09 de Junho de 2026, 20:56
  • j.s.: try65hytr a todos  49E09B4F
    09 de Junho de 2026, 20:56
  • m1957: Vamos todos colaborar para que o forum continue! Bom fim de semana.
    06 de Junho de 2026, 02:24
  • cereal killa: dgtgtr pessoal  49E09B4F
    04 de Junho de 2026, 14:49
  • j.s.: [link]
    03 de Junho de 2026, 10:01
  • j.s.: fica aqui a descrição do numero da conta
    03 de Junho de 2026, 10:00
  • j.s.: podem fazer, como tem sido sempre feito, por transferencia bancaria
    03 de Junho de 2026, 10:00
  • j.s.: por lapso não foi indicado  como podem ajudar o  forum
    03 de Junho de 2026, 09:58
  • j.s.: bo ghyt74 a todos  49E09B4F
    03 de Junho de 2026, 09:57
  • JP: try65hytr Pessoal  4tj97u<z 2dgh8i k7y8j0 classic
    02 de Junho de 2026, 04:05
  • FELISCUNHA: Bom dia , votos de um santo domingo para todo o auditório  4tj97u<z
    31 de Maio de 2026, 11:40
  • bruno mirandela: boa tarde a todos
    30 de Maio de 2026, 18:04
  • j.s.: [link]
    30 de Maio de 2026, 17:41
  • j.s.: tenham um bom fim de semana  49E09B4F
    30 de Maio de 2026, 17:38
  • j.s.: dgtgtr a todos  49E09B4F
    30 de Maio de 2026, 17:38
  • FELISCUNHA: ghyt74   49E09B4F  e bom fim de semana   4tj97u<z
    30 de Maio de 2026, 12:02
  • cereal killa: try65hytr pessoal  wwd46l0'
    29 de Maio de 2026, 21:14

Autor Tópico: Data Cleaning in Python (Updated 7/2020)  (Lida 329 vezes)

0 Membros e 1 Visitante estão a ver este tópico.

Online mitsumi

  • Sub-Administrador
  • ****
  • Mensagens: 133304
  • Karma: +0/-0
Data Cleaning in Python (Updated 7/2020)
« em: 09 de Agosto de 2020, 10:46 »

Data Cleaning in Python
Video: .mp4 (1280x720, 30 fps(r)) | Audio: aac, 48000 Hz, 2ch | Size: 1.61 GB
Genre: eLearning Video | Duration: 55 lectures (4 hour, 22 mins) | Language: English

 Preprocessing, structuring and normalizing data

What you'll learn

    Data cleaning or cleansing as a preprocessing step towards making the data more consistent and high quality before training predictive models.

Requirements

    Basics of Python

Description

Data cleaning or Data cleansing is very important from the perspective of building intelligent automated systems. Data cleansing is a preprocessing step that improves the data validity, accuracy, completeness, consistency and uniformity. It is essential for building reliable machine learning models that can produce good results. Otherwise, no matter how good the model is, its results cannot be trusted. Beginners with machine learning starts working with the publicly available datasets that are thoroughly analyzed with such issues and are therefore, ready to be used for training models and getting good results. But it is far from how the data is, in real world. Common problems with the data may include missing values, noise values or univariate outliers, multivariate outliers, data duplication, improving the quality of data through standardizing and normalizing it, dealing with categorical features. The datasets that are in raw form and have all such issues cannot be benefited from, without knowing the data cleaning and preprocessing steps. The data directly acquired from multiple online sources, for building useful application, are even more exposed to such problems. Therefore, learning the data cleansing skills help users make useful analysis with their business data. Otherwise, the term 'garbage in garbage out' refers to the fact that without sorting out the issues in the data, no matter how efficient the model is, the results would be unreliable.

In this course, we discuss the common problems with data, coming from different sources. We also discuss and implement how to resolve these issues handsomely. Each concept has three components that are theoretical explanation, mathematical evaluation and code. The lectures *.1.* refers to the theory and mathematical evaluation of a concept while the lectures *.2.* refers to the practical code of each concept.  In *.1.*, the first (*) refers to the Section number, while the second (*) refers to the lecture number within a section. All the codes are written in Python using Jupyter Notebook.

Who this course is for:

    The target students are beginners to data science and machine learning.

Download link:
Só visivel para registados e com resposta ao tópico.

Only visible to registered and with a reply to the topic.

Links are Interchangeable - No Password - Single Extraction