DLND-NMT-Extended

NMLT (EN to FR ) TensorFlow

In this project, I am going to build language translation model called seq2seq model or encoder-decoder model in TensorFlow. The objective of the model is translating English sentences to French sentences. I am going to show the detailed steps, and they will answer to the questions like how to preprocess the dataset, how to define inputs, how to define encoder model, how to define decoder model, how to build the entire seq2seq model, how to calculate the loss and clip gradients, and how to train and get prediction. Please open the IPython notebook file to see the full workflow and detailed descriptions.

This is a part of extended version of one of the projects of Udacity's Deep Learning Nanodegree. Some codes/functions (save, load, measuring accuracy, etc) are provided by Udacity. However, majority part is implemented by myself along with much richer explanations and references on each section.

In this repo I have also experiemented with Bidirectional LSTM architecture with the same data. I have also made a little different Keras implementation of the same model architecture in Keras-Implementation folder.

This Repo is implemented in Tensorflow 1.X.

Brief Overview of the Contents (For the Tensorflow 1.X models)

Data preprocessing

In this section, you will see how to get the data, how to create lookup table, and how to convert raw text to index based array with the lookup table.

Build model

In short, this section will show how to define the Seq2Seq model in TensorFlow. The below steps (implementation) will be covered.

(1) define input parameters to the encoder model
- enc_dec_model_inputs
(2) build encoder model
- encoding_layer
(3) define input parameters to the decoder model
- enc_dec_model_inputs, process_decoder_input, decoding_layer
(4) build decoder model for training
- decoding_layer_train
(5) build decoder model for inference
- decoding_layer_infer
(6) put (4) and (5) together
- decoding_layer
(7) connect encoder and decoder models
- seq2seq_model
(8) train and estimate loss and accuracy

Training

This section is about putting previously defined functions together to build an actual instance of the model. Furthermore, it will show how to define cost function, how to apply optimizer to the cost function, and how to modify the value of the gradients in the TensorFlow's optimizer module to perform gradient clipping.

Prediction

Nothing special but showing the prediction result.

Name		Name	Last commit message	Last commit date
Latest commit History 3 Commits
Keras-Implementation		Keras-Implementation
checkpoints-bi		checkpoints-bi
checkpoints		checkpoints
data		data
.DS_Store		.DS_Store
.gitignore		.gitignore
README.md		README.md
conversion.png		conversion.png
decoder_shift.png		decoder_shift.png
decoder_training_shift.png		decoder_training_shift.png
dlnd_language_translationv2-BI.ipynb		dlnd_language_translationv2-BI.ipynb
dlnd_language_translationv2.ipynb		dlnd_language_translationv2.ipynb
encoding_model.PNG		encoding_model.PNG
go_insert.png		go_insert.png
gradient_clipping.PNG		gradient_clipping.PNG
helper.py		helper.py
inference_phase.PNG		inference_phase.PNG
lookup.png		lookup.png
pad_insert.png		pad_insert.png
params-bi.p		params-bi.p
params.p		params.p
preprocess.p		preprocess.p
problem_unittests.py		problem_unittests.py
training_phase.PNG		training_phase.PNG

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

DLND-NMT-Extended

NMLT (EN to FR ) TensorFlow

Brief Overview of the Contents (For the Tensorflow 1.X models)

Data preprocessing

Build model

Training

Prediction

About

Releases

Packages

Languages

Sayan-Banerjee/DLND-NMT-Extended

Folders and files

Latest commit

History

Repository files navigation

DLND-NMT-Extended

NMLT (EN to FR ) TensorFlow

Brief Overview of the Contents (For the Tensorflow 1.X models)

Data preprocessing

Build model

Training

Prediction

About

Resources

Stars

Watchers

Forks

Releases

Packages 0

Languages

Packages