Universidad Austral
Limpieza de datos para el procesamiento de lenguaje natural
Universidad Austral

Limpieza de datos para el procesamiento de lenguaje natural

This course is part of Procesamiento de Lenguaje Natural Specialization

Taught in Spanish

Included with Coursera Plus

Course

Gain insight into a topic and learn the fundamentals

4.3

(14 reviews)

Beginner level

Recommended experience

12 hours (approximately)
Flexible schedule
Learn at your own pace

Details to know

Shareable certificate

Add to your LinkedIn profile

Assessments

12 quizzes

Course

Gain insight into a topic and learn the fundamentals

4.3

(14 reviews)

Beginner level

Recommended experience

12 hours (approximately)
Flexible schedule
Learn at your own pace

See how employees at top companies are mastering in-demand skills

Placeholder

Build your subject-matter expertise

This course is part of the Procesamiento de Lenguaje Natural Specialization
When you enroll in this course, you'll also be enrolled in this Specialization.
  • Learn new concepts from industry experts
  • Gain a foundational understanding of a subject or tool
  • Develop job-relevant skills with hands-on projects
  • Earn a shareable career certificate
Placeholder
Placeholder

Earn a career certificate

Add this credential to your LinkedIn profile, resume, or CV

Share it on social media and in your performance review

Placeholder

There are 4 modules in this course

Este módulo te permitirá obtener los conocimientos necesarios para la construcción de un programa de extracción de datos de páginas Web basadas en HTML.

What's included

5 videos6 readings3 quizzes2 discussion prompts1 plugin

En este módulo se describen un conjunto de pasos necesarios para el pre procesar páginas HTML y extraer información de ellas. Además, se detallarán distintos tipos de aproximación al mismo.

What's included

3 videos4 readings3 quizzes1 discussion prompt

En este módulo se presentarán las técnicas avanzadas de scraping para extracción de datos de páginas HTML que utilizan diversas librerías de JavaScript para su construcción

What's included

3 videos4 readings3 quizzes1 discussion prompt1 plugin

Una vez estriado el texto de las paginas HTML que es una fuente habitual de extracción de información, se pueden sumar distintas fuentes de tipos de datos, como ser PDF, DOC, XLS e imágenes. En este módulo se verán diversas técnicas que pueden servir para recolectar la información de ellas y unificarlas en un mismo conjunto de documentos.

What's included

4 videos4 readings3 quizzes1 discussion prompt1 plugin

Instructor

Instructor ratings
4.3 (5 ratings)
Hernán Daniel Merlino
Universidad Austral
4 Courses4,193 learners

Offered by

Recommended if you're interested in Machine Learning

Why people choose Coursera for their career

Felipe M.
Learner since 2018
"To be able to take courses at my own pace and rhythm has been an amazing experience. I can learn whenever it fits my schedule and mood."
Jennifer J.
Learner since 2020
"I directly applied the concepts and skills I learned from my courses to an exciting new project at work."
Larry W.
Learner since 2021
"When I need courses on topics that my university doesn't offer, Coursera is one of the best places to go."
Chaitanya A.
"Learning isn't just about being better at your job: it's so much more than that. Coursera allows me to learn without limits."

Learner reviews

Showing 3 of 14

4.3

14 reviews

  • 5 stars

    64.28%

  • 4 stars

    14.28%

  • 3 stars

    7.14%

  • 2 stars

    14.28%

  • 1 star

    0%

LV
5

Reviewed on Nov 7, 2022

New to Machine Learning? Start here.

Placeholder

Open new doors with Coursera Plus

Unlimited access to 7,000+ world-class courses, hands-on projects, and job-ready certificate programs - all included in your subscription

Advance your career with an online degree

Earn a degree from world-class universities - 100% online

Join over 3,400 global companies that choose Coursera for Business

Upskill your employees to excel in the digital economy

Frequently asked questions