Applied Natural Language Processing in Engineering Part 2

Ce cours n'est pas disponible en Français (France)

Nous sommes actuellement en train de le traduire dans plus de langues.

Applied Natural Language Processing in Engineering Part 2

Instructeur : Ramin Mohammadi

Inclus avec Coursera Plus

7 modules

Obtenez un aperçu d'un sujet et apprenez les principes fondamentaux.

3 semaines à compléter

à 10 heures par semaine

Planning flexible

Apprenez à votre propre rythme

7 modules

Obtenez un aperçu d'un sujet et apprenez les principes fondamentaux.

3 semaines à compléter

à 10 heures par semaine

Planning flexible

Apprenez à votre propre rythme

Compétences que vous acquerrez

Catégorie : Artificial Neural Networks
Catégorie : Natural Language Processing
Catégorie : PyTorch (Machine Learning Library)
Catégorie : Statistical Machine Learning
Catégorie : Applied Machine Learning
Catégorie : Algorithms
Catégorie : Deep Learning
Catégorie : Large Language Modeling
Catégorie : Machine Learning Methods

Détails à connaître

Certificat partageable

Ajouter à votre profil LinkedIn

Récemment mis à jour !

octobre 2025

Évaluations

21 devoirs

Enseigné en Anglais

Découvrez comment les employés des entreprises prestigieuses maîtrisent des compétences recherchées

En savoir plus sur Coursera pour les affaires

logos de Petrobras, TATA, Danone, Capgemini, P&G et L'Oreal

Il y a 7 modules dans ce cours

This course is best suited for software engineers, data scientists, and graduate students in computer science or engineering fields who wish to develop expertise in building and deploying natural language processing systems to solve real-world language understanding challenges.

You will master core NLP tasks such as Part-of-Speech tagging, Named Entity Recognition, sentiment analysis, and Neural Machine Translation while implementing various neural architectures from Recurrent Neural Networks and bidirectional RNNs to Conditional Random Fields and state-of-the-art transformer models. The course emphasizes practical application through extensive laboratory work and projects, where you will develop complete NLP pipelines using frameworks like PyTorch and Hugging Face, learning to preprocess data, train models, and evaluate performance using industry-standard metrics. By the end of the course, you will be equipped with both theoretical understanding and practical skills to design, implement, and optimize NLP solutions for real-world engineering applications, from chatbots and translation systems to information extraction and text analysis tools. The curriculum culminates in a comprehensive capstone project where you will apply multiple techniques learned throughout the course to solve a complex language processing challenge. You will be equipped with both theoretical knowledge to tackle complex language processing problems in industry settings, enabling you to build production-ready NLP applications that can understand, interpret, and generate human language effectively.

This module delves into the critical preprocessing step of tokenization in NLP, where text is segmented into smaller units called tokens. You will explore various tokenization techniques, including character-based, word-level, Byte Pair Encoding (BPE), WordPiece, and Unigram tokenization. Then you’ll examine the importance of normalization and pre-tokenization processes to ensure text uniformity and improve tokenization accuracy. Through practical examples and hands-on exercises, students will learn to handle out-of-vocabulary (OOV) issues, manage large vocabularies efficiently, and understand the computational complexities involved. By the end of the module, you will be equipped with the knowledge to implement and optimize tokenization methods for diverse NLP applications.

Inclus

1 vidéo13 lectures2 devoirs1 élément d'application

1 vidéoTotal 1 minute

Meet Your Faculty1 minute

13 lecturesTotal 69 minutes

Course Introduction2 minutes
Syllabus - Applied Natural Language Processing in Engineering Part 210 minutes
Academic Integrity1 minute
Week 8 Overview2 minutes
Introduction5 minutes
Pre-Tokenization5 minutes
Character-based Tokenization5 minutes
Word-level Tokenization5 minutes
Byte Pair Encoding (BPE)10 minutes
WordPiece Tokenization10 minutes
Unigram Tokenization10 minutes
Vocabulary Pruning in Unigram Tokenization2 minutes
Summary and Final Thoughts2 minutes

2 devoirsTotal 75 minutes

Assess Your Learning: Tokenization30 minutes
Module 8 Quiz45 minutes

1 élément d'applicationTotal 10 minutes

The Viterbi Algorithm for Tokenization10 minutes

In this module, we will explore foundational models in natural language processing (NLP), focusing on language models, feedforward neural networks (FFNNs), and Hidden Markov Models (HMMs). Language models are crucial in predicting and generating sequences of text by assigning probabilities to words or phrases within a sentence, allowing for applications such as autocomplete and text generation. FFNNs, though limited to fixed-size contexts, are foundational neural architectures used in language modeling, learning complex word relationships through non-linear transformations. In contrast, HMMs model sequences based on hidden states, which influence observable outcomes. They are particularly useful in tasks like part-of-speech tagging and speech recognition. As the module progresses, we will also examine modern advancements like neural transition-based parsing and the evolution of language models into sophisticated architectures such as transformers and large-scale pre-trained models like BERT and GPT. This module provides a comprehensive view of how language modeling has developed from statistical methods to cutting-edge neural architectures.

Inclus

2 vidéos19 lectures4 devoirs

2 vidéosTotal 7 minutes

Language Models4 minutes
Hidden Markov Models3 minutes

19 lecturesTotal 183 minutes

Week 9 Overview2 minutes
Introduction to Language Models5 minutes
Probability Assignment in Language Model2 minutes
Evolution of Language Models10 minutes
State-of-the-Art Models2 minutes
N-Gram5 minutes
Probabilities in Language Models10 minutes
Example: The Cat Sat on the Mat10 minutes
Limitations of N-Gram Models5 minutes
FFNN in Language Modeling20 minutes
Pros and Cons of FFNNs5 minutes
Introduction to HMM10 minutes
Hidden Markov Models2 minutes
Mathematical Representation of HMMs15 minutes
Likelihood Problem: Forward Algorithm10 minutes
Decoding Problem: Viterbi Algorithm15 minutes
Learning Problem: Baum-Welch Algorithm15 minutes
Example of HMM20 minutes
HMMs in Speech Recognition20 minutes

4 devoirsTotal 120 minutes

Assess Your Learning: Language Models30 minutes
Assess Your Learning: FFNNs15 minutes
Assess Your Learning: HMMs30 minutes
Module 9 Quiz 45 minutes

In this module, we will explore Recurrent Neural Networks (RNNs), a fundamental architecture in deep learning designed for sequential data. RNNs are particularly well-suited for tasks where the order of inputs matters, such as time series prediction, language modeling, and speech recognition. Unlike traditional neural networks, RNNs have connections that allow them to “remember” information from previous steps by sharing parameters across time steps. This ability enables them to capture temporal dependencies in data, making them powerful for sequence-based tasks. However, RNNs come with challenges like vanishing and exploding gradients which affect their ability to learn long-term dependencies. Throughout the module, you will explore different RNN variants such as Long Short-Term Memory (LSTM) and Gated Recurrent Units (GRUs), which address these challenges. You will also delve into advanced training techniques and applications of RNNs in real-world NLP and time series problems.

Inclus

2 vidéos22 lectures2 devoirs1 élément d'application

2 vidéosTotal 4 minutes

Recurrent Neural Networks3 minutes
The RNN Process0 minutes

22 lecturesTotal 221 minutes

Week 10 Overview2 minutes
Recurrent Neural Networks (RNNs)2 minutes
Challenges & Applications in RNN5 minutes
Parameter Sharing in RNN5 minutes
Dynamic Systems5 minutes
Dynamic Systems to RNN10 minutes
Computing Gradient in RNN10 minutes
RNN Advantages and Disadvantages5 minutes
Training an RNN Language Model20 minutes
Problems with RNN15 minutes
How to Solve these Issues?15 minutes
Gated RNN15 minutes
LSTM Equations15 minutes
Gated Recurrent Unit (GRU)10 minutes
Residual Neural Networks20 minutes
Skip Connection: The Key to Learning Residuals15 minutes
Conventions Used2 minutes
Step-by-Step Breakdown 1 - 25 minutes
Step-by-Step Breakdown 3 A - G15 minutes
Step-by-Step Breakdown 3 H - N15 minutes
Step-by-Step Breakdown 4 - 610 minutes
Perplexity Calculation5 minutes

2 devoirsTotal 75 minutes

Assess Your Learning: RNNs30 minutes
Module 10 Quiz 45 minutes

1 élément d'applicationTotal 10 minutes

Introduction to LSTM, GRU, and Residual Networks10 minutes

This module introduces students to advanced Natural Language Processing (NLP) techniques, focusing on foundational tasks such as Part-of-Speech (PoS) tagging, sentiment analysis, and sequence modeling with recurrent neural networks (RNNs). Students will examine how PoS tagging helps in understanding grammatical structures, enabling applications such as machine translation and named entity recognition (NER). The module delves into sentiment analysis, highlighting various approaches from traditional machine learning models (e.g., Naive Bayes) to advanced deep learning techniques (e.g., bidirectional RNNs and transformers). Students will learn to implement both forward and backward contextual understanding using bidirectional RNNs, which improves accuracy in tasks where sequence order impacts meaning. By the end of the course, students will gain hands-on experience building NLP models for real-world applications, equipping them to handle sequential data and capture complex dependencies in text analysis.

Inclus

1 vidéo15 lectures4 devoirs

1 vidéoTotal 5 minutes

Introduction to PoS Tagging, Bidirectional RNNs, and Sentiment Analysis5 minutes

15 lecturesTotal 113 minutes

Week 11 Overview2 minutes
Introduction to PoS Tagging10 minutes
How does PoS Tagging Works?10 minutes
Challenges in & Advantages of PoS Tagging5 minutes
Using Recurrent Neural Networks (RNNs) for PoS Tagging10 minutes
Steps in PoS Tagging with RNN5 minutes
Using LSTM or GRU in Place of Simple RNNs10 minutes
Conclusion10 minutes
Motivation2 minutes
Bidirectional RNNs10 minutes
Multi-layer RNNs10 minutes
Introduction5 minutes
Approaches with RNNs20 minutes
Other Approaches for Sentiment Analysis2 minutes
Conclusion2 minutes

4 devoirsTotal 135 minutes

Assess Your Learning: PoS30 minutes
Assess Your Learning: Bidirectional RNNs30 minutes
Assess Your Learning: Sentiment Analysis30 minutes
Module 11 Quiz 45 minutes

This module introduces you to core tasks and advanced techniques in Natural Language Processing (NLP), with a focus on structured prediction, machine translation, and sequence labeling. You will explore foundational topics such as Named Entity Recognition (NER), Part-of-Speech (PoS) tagging, and sentiment analysis and use neural network architectures like Recurrent Neural Networks (RNNs), Long Short-Term Memory (LSTM) networks, and Conditional Random Fields (CRFs). The module will cover key concepts in sequence modeling, such as bidirectional and multi-layer RNNs, which capture both past and future context to enhance the accuracy of tasks like NER and PoS tagging. Additionally, you will delve into Neural Machine Translation (NMT), examining encoder-decoder models with attention mechanisms to address challenges in translating long sequences. Practical implementations will involve integrating these models into real-world applications, focusing on handling complex language structures, rare words, and sequential dependencies. By the end of this module, you will be proficient in building and optimizing deep learning models for a variety of NLP tasks.

Inclus

3 vidéos18 lectures4 devoirs

3 vidéosTotal 6 minutes

Introduction to CRF2 minutes
Introduction to NER and NMT3 minutes
Visualization of the NMT Process 0 minutes

18 lecturesTotal 164 minutes

Week 12 Overview2 minutes
Definition of CRF10 minutes
CRF Model with LSTM10 minutes
Combining LSTM with CRF20 minutes
Calculating the Probability of a Sequence, Log-Probability & Training Objective15 minutes
Decoding: Finding the Best Label Sequence5 minutes
Details on LSTM-CRF Components15 minutes
Summary of the Transition Matrix in CRF5 minutes
Named Entity Recognition (NER)10 minutes
NER Using RNNs/LSTMs10 minutes
BiLSTM for NER10 minutes
CRF Layer for Sequencing Labeling10 minutes
Attention in NER10 minutes
Table: Alphabetical List of PoS Tags used in the Penn Treebank Project5 minutes
Machine Translation Overview5 minutes
Sequence-to-Sequence Model for NMT10 minutes
Learning in NMT: Optimization and Loss Function10 minutes
Byte Pair Encoding (BPE) for Handling Rare Words2 minutes

4 devoirsTotal 135 minutes

Assess Your Learning: CRFs30 minutes
Assess Your Learning: NERs30 minutes
Assess Your Learning: NMTs30 minutes
Module 12 Quiz45 minutes

In this module we’ll focus on attention mechanisms and explore the evolution and significance of attention in neural networks, starting with its introduction in neural machine translation. We’ll cover the challenges of traditional sequence-to-sequence models and how attention mechanisms, particularly in Transformer architectures, address issues like long-range dependencies and parallelization, which enhances the model's ability to focus on relevant parts of the input sequence dynamically. Then, we’ll turn our attention to Transformers and delve into the revolutionary architecture introduced by Vaswani et al. in 2017, which has significantly advanced natural language processing. We’ll cover the core components of Transformers, including self-attention, multi-head attention, and positional encoding to explain how these innovations address the limitations of traditional sequence models and enable efficient parallel processing and handling of long-range dependencies in text.

Inclus

2 vidéos25 lectures3 devoirs2 éléments d'application

2 vidéosTotal 9 minutes

Attention Mechanisms3 minutes
Transformers5 minutes

25 lecturesTotal 239 minutes

Week 13 Overview2 minutes
Introduction and Motivation5 minutes
Sequence-to-Sequence Models5 minutes
Challenges of Seq2Seq Models15 minutes
Attention Mechanisms5 minutes
General Seq2Seq Models10 minutes
Detailed Attention Process in Seq2Seq15 minutes
Introduction and Transformer Architecture2 minutes
Applications of Transformer Architectures5 minutes
Key, Query, Value3 minutes
Self-Attention15 minutes
Self-Attention as Routing5 minutes
Computing and Weighting Values10 minutes
Self-Attention in Matrix Form10 minutes
Position Representations 10 minutes
The Intuition15 minutes
Elementwise Nonlinearity20 minutes
Multi-head Attention10 minutes
Sequence-Tensor Form10 minutes
Transformers15 minutes
Types of Transformers20 minutes
Cross-Attention15 minutes
Decoder Process with Cross-Attention10 minutes
Drawbacks of Transformers5 minutes
Conclusion2 minutes

3 devoirsTotal 105 minutes

Assess Your Learning: Attention30 minutes
Assess Your Learning: Transformer30 minutes
Module 13 Quiz45 minutes

2 éléments d'applicationTotal 40 minutes

Multi-Head Visualization20 minutes
Encoder-Decoder Example20 minutes

In this module, we’ll hone in on pre-training and explore the foundational role of pre-training in modern NLP models, highlighting how models are initially trained on large, general datasets to learn language structures and semantics. This pre-training phase, often involving tasks like masked language modeling, equips models with broad linguistic knowledge, which can then be fine-tuned on specific tasks, enhancing performance and reducing the need for extensive task-specific data.

Inclus

1 vidéo19 lectures2 devoirs

1 vidéoTotal 5 minutes

Pre-Training5 minutes

19 lecturesTotal 209 minutes

Week 14 Overview2 minutes
Introduction to Pre-Training15 minutes
Pretrained Word Embeddings10 minutes
Learning from Reconstructing Input10 minutes
Pretraining Through Language Modeling20 minutes
Pretraining for Three Types of Architectures10 minutes
BERT: Bidirectional Encoder Representations from Transformers15 minutes
BERT Pre-training 10 minutes
Fine-tuning15 minutes
Full fine-tuning vs Parameter-Efficient Fine-tuning15 minutes
Limitations of Pre-trained Encoders and Extensions of BERT10 minutes
Pretraining Decoders10 minutes
Generative Pretrained Transformer (GPT)10 minutes
Scaling Laws15 minutes
What kinds of things does pretraining teach?10 minutes
Pretraining encoder-decoders: What pretraining objective to use?15 minutes
Span Corruption: T5 model10 minutes
Transfer Learning to Downstream Tasks5 minutes
Congratulations! 2 minutes

2 devoirsTotal 75 minutes

Assess Your Learning: Pre-training30 minutes
Module 14 Quiz45 minutes

Instructeur

Ramin Mohammadi

Northeastern University

4 Cours531 apprenants

Offert par

Northeastern University

En savoir plus sur Machine Learning

Statut : Essai gratuit
Packt
Natural Language Processing with Real-World Projects
Spécialisation
Statut : Essai gratuit
DeepLearning.AI
Natural Language Processing
Spécialisation
Statut : Prévisualisation
Northeastern University
NLP in Engineering: Concepts & Real-World Applications
Cours
Statut : Essai gratuit
Packt
Applied Generative AI & NLP with Python
Cours

Pour quelles raisons les étudiants sur Coursera nous choisissent-ils pour leur carrière ?

Felipe M.

Étudiant(e) depuis 2018

’Pouvoir suivre des cours à mon rythme à été une expérience extraordinaire. Je peux apprendre chaque fois que mon emploi du temps me le permet et en fonction de mon humeur.’

Jennifer J.

Étudiant(e) depuis 2020

’J'ai directement appliqué les concepts et les compétences que j'ai appris de mes cours à un nouveau projet passionnant au travail.’

Larry W.

Étudiant(e) depuis 2021

’Lorsque j'ai besoin de cours sur des sujets que mon université ne propose pas, Coursera est l'un des meilleurs endroits où se rendre.’

Chaitanya A.

’Apprendre, ce n'est pas seulement s'améliorer dans son travail : c'est bien plus que cela. Coursera me permet d'apprendre sans limites.’

Ouvrez de nouvelles portes avec Coursera Plus

Accès illimité à 10,000+ cours de niveau international, projets pratiques et programmes de certification prêts à l'emploi - tous inclus dans votre abonnement.

Faites progresser votre carrière avec un diplôme en ligne

Obtenez un diplôme auprès d’universités de renommée mondiale - 100 % en ligne

Découvrir les diplômes

Rejoignez plus de 3 400 entreprises mondiales qui ont choisi Coursera pour les affaires

Améliorez les compétences de vos employés pour exceller dans l’économie numérique

Foire Aux Questions

To access the course materials, assignments and to earn a Certificate, you will need to purchase the Certificate experience when you enroll in a course. You can try a Free Trial instead, or apply for Financial Aid. The course may offer 'Full Course, No Certificate' instead. This option lets you see all course materials, submit required assessments, and get a final grade. This also means that you will not be able to purchase a Certificate experience.

When you purchase a Certificate you get access to all course materials, including graded assignments. Upon completing the course, your electronic Certificate will be added to your Accomplishments page - from there, you can print your Certificate or add it to your LinkedIn profile.

Yes. In select learning programs, you can apply for financial aid or a scholarship if you can’t afford the enrollment fee. If fin aid or scholarship is available for your learning program selection, you’ll find a link to apply on the description page.