Facial Emotion Recognition System
Facial Emotion Recognition System
ORG
INTRODUCTION
1.1 Introduction
The project operates within the domain of Emotion Recognition and Analysis
through Facial Expressions. It is concerned with the development of a system that
can automatically detect and classify a wide range of emotional states by analyzing
facial expressions. This domain finds applications in various fields, such as eLearning
and effective computing, where understanding and interpreting human emotions
based on their facial expressions is of significant importance.
The scope of this project is to develop a robust system for facial emotion
recognition using deep learning, particularly CNNs. It involves data collection,
preprocessing, CNN architecture design, feature extraction, and classification to
accurately identify a range of emotions. The project aims to explore cross-cultural
considerations and practical applications in real-world deployment. Its primary focus
is to contribute to the field of affective computing by creating a tool that can
understand and respond to human emotions based on facial expressions.
1.5 Methodology
• Data Collection and Pre-processing: Diverse dataset of facial images that depict a
wide range of emotional states, including Sad, Anger, Fear, Joy, Disgust, Confused,
Frustrated and Surprise is gathered. These images have been carefully pre-processed
to improve their quality, standardize lighting conditions, and ensure that facial
features are aligned consistently for uniform input.
• Feature Extraction: Employ CNNs to automatically extract relevant features from the
pre-processed facial images. CNNs excel at capturing spatial and temporal
dependencies within images, crucial for understanding complex emotional
expressions.
• Training and Learning: The CNN is trained using the pre-processed dataset, where
the network learns to differentiate and classify facial expressions into predefined
emotion categories. The architecture’s re-usability of weights and ability to capture
spatial and temporal dependencies contribute to better fitting the image dataset.
• Classification and Emotion Identification: The trained CNN model is used to classify
facial expressions and identify an individual’s emotion based on primary emotional
categories (Sad, Anger, Fear, Joy, Disgust, Surprise). The architecture’s ability to
differentiate and recognize subtle features in facial expressions enhances the
accuracy of emotion identification.
• Evaluation and Testing: The CNN based facial emotion recognition system’s
performance using suitable metrics is evaluated, which include accuracy, precision,
recall, and the F1 score. This assessment gauges the system’s capability to accurately
and effectively identify and distinguish among a diverse set of emotions.
• Application: The developed system is put into practical use in real-world applications,
including areas like mental health support, Retail and Customer Service, Emotion-
aware Gaming, Human Resource and Recruitment, Cognitive Load Monitoring in
Education, human-computer interaction and enhancing user experiences and its
actively refined through an iterative process, taking into account valuable user
feedback and staying updated with the latest advancements in deep learning and
emotion recognition research to make it even better.
Chapter 2
LITERATURE REVIEW
Jyoti Kumari.,et al,[1] discussed the Detection of mental disorders, and synthetic
human expressions. The author mentioned that the two common methods used
predominantly in the literature for Facial Emotion Recognition (FER) automatic
systems depend on geometry and appearance. The author provides a quick scan for
facial expression recognition. A comparative study was also performed using various
feature extraction techniques in the Japanese Female Facial Expression (JAFFE)
dataset. The limitations of Analyzing facial expressions has a major drawback humans
can control the simulation to some extent, so recognition results may be falsified,
intentionally or unintentionally.
G. Kalaivani et. al,[2] made use of the Viola Jones techniques and image cropping
techniques for extracting and delineating areas of the [Link] proposed
segmentation techniques are applied and compared to the method found which is
suitable for segmentation of the oral region, and then the oral region can be
extracted by means of contrast extension and image segmentation techniques. After
extracting the mouth area, the facial feelings are ranked based on the white pixel
values in the mouth area extracted from the face image. The limitations are that
traditional image segmentation techniques are more fragmentation and have high
noise sensitivity.
Ravichandra Ginne, et al,[4] have surveyed various FER techniques which has
become an active research area that finds many applications in areas such as human
and computer interfaces, human emotion analysis, psychoanalysis, medical
IJNRD2402169 International Journal of Novel Research and Development ([Link]) b558
© 2024 IJNRD | Volume 9, Issue 2 February 2024| ISSN: 2456-4184 | [Link]
diagnostics, etc. The common techniques used for this purpose depend on geometry
and appearance. CNN have been shown to outperform traditional approaches for
various visual recognition tasks including recognition of facial expressions. Despite
efforts to improve the accuracy of FER systems using CNN, current methods may not
be sufficient for practical [Link] study includes a general review FER of
systems using CNN and their strengths and limitations which helps us to further
understand and improve FER [Link] limitations are Improper encoding of
object’s position and orientation. Lack of the ability to be spatially unchanged for the
input data, in International Journal of Advances in Electronics and Computer Science.
object into their predictions and Dynamic FER has a higher recognition rate than
static.
Elzbieta kukla, et al[11] introduced a method that uses a series of neural networks
to recognize facial [Link] an input, the algorithm receives a natural image of
the face and returns the emotion expressed by the face. To determine the best
classifiers for recognizing specific emotions, single- and multiple-layered networks
IJNRD2402169 International Journal of Novel Research and Development ([Link]) b560
© 2024 IJNRD | Volume 9, Issue 2 February 2024| ISSN: 2456-4184 | [Link]
were tested. The experiments covered different resolutions of the images displaying
the faces as well as the images, including the areas of the mouths and [Link] the
basis of the results of the tests, a series of neural networks are proposed. The series
introduces six basic emotions and a neutral [Link] Limitations are Black box,
development period, amount of data and calculation cost.
Raut et al,[12] displayed Facial Emotion Recognition Using Machine Learning. The
author in this research stated that the subtle emotions in Eulerian Motion
Magnification (EMM) are difficult to detect. Movement characteristics such as speed
and acceleration can be used to zoom [Link] image is transformed as a whole by
enlarging changes in the amplitude and phase properties. Depending on the
characteristics, there are A-EMM (capacitive based) and P-EMM (phase based)
motion amplification. Oriented FAST and Rotated BRIEF (ORB) and Speeded-Up
Robust Features (SURF), Scale-Invariant Feature Transform (SIFT) are used and also
feature descriptor algorithms [Link] dataset used in this experiment was the iBug-
300W dataset containing over 7,000 images as well as the CK + dataset containing
593 facial expression sequences from 123 different subjects. The limitation is that if
the motions are large, this manipulation can introduce artefacts,in Master’s Projects.
Chapter 3
PROJECT DESCRIPTION
3.1 Existing System
Existing systems for facial emotion recognition often use various techniques,
including deep learning and traditional machine learning methods. They typically
involve collecting facial data, extracting features, training models, and then using
these models to recognize emotions in real-time or from static images. Some
methods heavily rely on detecting facial landmarks and extracting geometric
features, such as distances between facial points, angles, and ratios, to recognize
emotions. Facial Action Coding System (FACS) involves manual coding of facial
movements and expressions based on anatomically defined facial actions. While it’s
a detailed and comprehensive system, it requires expert annotation and might be
labor-intensive.
training data can result in less effective models, and collecting diverse and balanced
datasets can be challenging.
This project is focused on improving the accuracy and reliability of facial emotion
recognition by addressing the challenge of controlling and falsifying facial
expressions by humans. The key differentiators in this project include:
Quick Scan Approach: This project introduces a quick scan approach, which
represents an efficient and rapid method for recognizing facial expressions. This
approach could be particularly valuable for real-time applications where speed is
essential.
Comparative Study: The project also includes a comparative study using various
feature extraction techniques on the (Japanese Female Facial Expression dataset)
JAFFE dataset, which means that the model is actively evaluating and improving the
performance of the system, which is a valuable aspect of research and development
in this field.
The project would require an initial investment in technology and data acquisition.
However, the potential benefits in terms of mental health support and overall
wellbeing could justify the costs.
The project aligns with the growing societal awareness of mental health issues and
the need for innovative solutions.
• Processor - i5
• Speed - 3 GHz
• RAM - 8 GB(min)
• Monitor - SVGA
• Health Insurance Portability and Accountability Act (HIPAA): If the project involves
health-related data, such as medical records, HIPAA compliance is essential.
• Informed Consent: Obtain explicit and informed consent from individuals whose data
is used in the project.
• Mental Health Support and Counseling: National and International Mental Health
Guidelines: Consult and align with established guidelines and best practices in
mental health support and counseling, as recommended by organizations like the
World Health Organization (WHO) and national health agencies.
• Clinical Validation: If the project involves clinical trials or validation with mental
health professionals, adhere to recognized clinical trial standards.
• AI and Facial Recognition Regulations: Local AI and Data Privacy Laws: Stay informed
about regional laws and regulations related to AI, facial recognition, and data privacy.
• Research and Academic Standards: Ethics Review Boards: If the project is associated
with an academic institution, obtain approvals from ethics review boards for
research involving human subjects.
• User Data Handling Policies: Develop and communicate clear policies regarding how
user data is collected, stored, and used. Users should be informed about their datas
purpose and retention.
Chapter 4
METHODOLOGY
4.1 Architecture Diagram
In the above Figure 4.1, initially, the user provides an input in the form of an image,
image pre-processing is applied to eliminate any visual defects or unwanted artifacts.
Subsequently, feature extraction techniques are employed to capture relevant
information from the image. Following this, a CNN model is trained using a dataset of
emotions, where the model learns to recognize emotional cues in the images. Once the
model is trained and tested to minimize errors, it is then employed to process the input
image. Finally, the system provides an output that signifies the detected emotion,
effectively translating visual data into emotional insights.
The Fig 4.2, begins with an Test Image given by the user as its primary data source
and this image is directed to the Image Pre-processing task, where any visual defects
or unwanted artifacts are removed. After pre-processing, the data flows into the
Feature Extraction activity, which captures relevant information from the image. The
output from feature extraction is then directed to the CNN Model Training process.
Here, a Dataset of Emotions is used to train the CNN Model. The model learns to
recognize emotional cues in the images provided in the dataset. Once the model is
trained and tested to minimize errors, it moves to the Emotion Analysis phase. In
Emotion Classification, the trained CNN model is employed to process the input test
image. The analysis results in the Detected Emotion data, which signifies the emotion
detected in the input image. Finally, the Detected Emotion is the output of the
system, providing a valuable insight into the emotions conveyed by the input image.
In the above Fig 4.3, the primary scenario involves a user uploading an image into
the system. The system then undertakes a series of actions, including image
processing, feature extraction, and emotion identification. Once these processes are
complete, the system generates an output, which is subsequently presented to the
user for viewing. This use case illustrates the user’s interaction with the system,
demonstrating its ability to analyze and convey emotions based on the uploaded
image.
In the Fig 4.4, The Mental Health System Represents the core class orchestrating the
system & Manages interaction with Face Detection Module, Emotion Module, and
User Database. The Face Detection Module Detects faces in an image using the
detect Faces method & returns a list of detected faces, each represented by the Face
class. The Emotion Module recognizes emotions in a detected face using the
recognize Emotion method & returns the recognized emotion, represented by the
Emotion class. The Face represents a detected face with attributes such as bounding
box and facial landmarks. The Emotion represents a specific emotion detected in a
facial expression & provides methods to retrieve the type of emotion. The Emotion
Type is an enumeration class enumerating different types of emotions (e.g., happy,
sad, angry).
In above Figure 4.5 ,the application begins when the user uploads an image featuring
a human face. The system promptly employs facial recognition to detect the
presence of a face in the image. Following this, the device conducts image pre-
processing to refine the image, extracting critical facial features. Finally, the output,
which encompasses both the pre-processed image and the extracted features, is
securely stored in a database for future use. This process seamlessly combines user
interaction, facial detection, image enhancement, feature extraction, and database
management.
4.3 Module Description
Critical points of the Face are very important and can be used for facial recognition
and detection. Here a total of 68 facial critical points are represented to the
discoverer in the [Link] (a set of tools for developing machine learning and data
analysis applications in the real world.)
All these 68 critical point on the face are depicted in the above figure. The (x,y)
coordinates of every facial point can be retrieved by using [Link] tools. All these
points can be broken down into categories such as face, nose, eyebrow and jaw. The
proposed method is based on a two-level CNN framework. The first level
recommended is background removal, used to extract emotions from an image, as
shown in the figure below.
All the convolutional layers used are capable of pattern detection. Within each
convolutional layer, four filters were used. The input image fed to the first-part CNN
(used for background removal) generally consists of shapes, edges, textures, and
objects along with the face. The edge detector, circle detector, and corner detector
filters are used at the start of the convolutional layer 1. Once the face has been
detected, the second-part CNN filter catches facial features, such as eyes, ears, lips,
nose, and cheeks. The edge detection filters used in this layer are shown in Fig.4.9.
The second-part CNN consists of layers with kernel matrix, e.g., [0.25, 0.17, 0.9;
0.89, 0.36, 0.63; 0.7, 0.24, 0.82]. These numbers are selected between 0 and 1
initially. These numbers are optimized for EV detection, based on the ground truth
IJNRD2402169 International Journal of Novel Research and Development ([Link]) b572
© 2024 IJNRD | Volume 9, Issue 2 February 2024| ISSN: 2456-4184 | [Link]
we had, in the supervisory training dataset. Here, we used minimum error decoding
to optimize filter values. Once the filter is tuned by supervisory learning, it is then
applied to the background-removed face (i.e., on the output image of the first-part
CNN), for detection of different facial parts (e.g., eye, lips. nose, ears, etc.)
Deep neural network with transfer learning approach is employed to extract bottle
features from the input images and save these features. At a later stage, a network
of fully connected layers is used where the bottle features are loaded back and
images are then passed to the model for prediction.
Deep neural network with transfer learning is employed to extract the bottleneck
features from JAFFE images dataset and fed these features to a set of fully connected
layers to predict the facial emotions of these images. Finally, complete content and
organizational editing is done before formatting.
Transfer learning is used as follows, Reused VGG16 Model, Loaded with Image net
weights. Extract bottleneck feature from it for the testing scenario. Tuned the model
with bottleneck weights to achieve higher accuracy.
Initially, the Cohn–Kanade dataset with 486 sequences yielded a maximum accuracy
of 45%. To improve efficiency, additional datasets were gathered, including
internetsourced data and the users own pictures. The algorithm’s accuracy increased
with the number of images in the dataset, reaching 96% with a 10,000-image
dataset. The study involved 25 iterations, optimizing the number of layers and filters
for a CNN. The optimal configuration was found to be four layers and four filters. The
algorithm successfully detected emotions in front-facing images but faced challenges
IJNRD2402169 International Journal of Novel Research and Development ([Link]) b573
© 2024 IJNRD | Volume 9, Issue 2 February 2024| ISSN: 2456-4184 | [Link]
with grayscale images and orientation. Despite limitations such as high computing
power requirements during tuning and issues with facial hair, the algorithm’s
accuracy compared favorably with other studies. However, limitations were noted in
cases of missing facial features and multiple faces in an image. The algorithm’s
performance was further tested on Caltech faces, CMU, and NIST databases, with
accuracy decreasing with an increasing number of images due to overfitting. The
ideal number of images for proper functioning was determined to be in the range of
2000–10,000.
4.3.7 Cross-Validation
Chapter 5
IMPLEMENTATION AND TESTING
5.0.1 Image Insertion & Error Testing
The procedure level testing is made first by giving improper inputs, the errors
occurred are noted and eliminated. This is the final step in system life cycle. Here the
tested error-free system is implemented into real-life environment and make
necessary changes, which runs in an online fashion. The system maintenance is done
every months or year based on company policies, and is checked for errors like
runtime errors, long run errors and other maintanances like table verification and
reports.
• Command Prompt Output:In addition to the visual output, the software generates
textual feedback and commands in the command prompt or terminal. These textual
outputs can be useful for users who prefer programmatic interaction.
• User Interaction: Users have the flexibility to interact with the software in real-time
as the program processes the input data. They can observe the visual outputs and
simultaneously receive textual feedback.
• File Export and Import: Users might have the option to export data or results to
various file formats, such as CSV, Excel, PDF, or other specialized file types. This
output can be saved locally or shared for external use.
• Input Validation: Testing the module to correctly handles various image formats,
sizes, and color spaces.
• Face Detection: Verification that the face detection component accurately locates
faces within the input image.
• Prediction Accuracy: Unit tests should assess the model’s accuracy in recognizing
emotions based on sample inputs.
• Boundary Testing: Include test cases with a variety of facial expressions, including
extreme and subtle emotions.
• Error Handling: Unit tests should include scenarios that intentionally cause errors to
ensure the software handles exceptions appropriately.
• Image Processing and Feature Extraction: This integration point involves testing the
proper flow of data from image capture to the feature extraction module. It checks
if facial features and emotions are correctly identified from images.
• Real-time Processing: For live camera feeds, this tests the integration between the
camera input, real-time processing, and the output visualization. It ensures that the
software can handle continuous data streams.
• Graphical User Interface (GUI): Integration testing ensures that the GUI components
(if applicable) communicate effectively with the back-end components. This includes
displaying results and accepting user inputs.
• Visual Studio Environment: Testing includes the interaction between the software
and the Visual Studio IDE, as command prompts and other development tools are
used for feedback and control.
• User Interface Testing: Functional testing evaluates the graphical user interface (GUI)
components if applicable. It verifies that users can interact with the software as
expected, input images, receive results, and navigate the user interface without
issues.
• Input Validation: Testing ensures that the software validates input correctly. Invalid
or corrupt image data should be handled gracefully, and the software should provide
appropriate error messages or responses
• Code Review: Performing a thorough review of the source code that makes up the
different components of the project including examining the code for the image
IJNRD2402169 International Journal of Novel Research and Development ([Link]) b577
© 2024 IJNRD | Volume 9, Issue 2 February 2024| ISSN: 2456-4184 | [Link]
processing, feature extraction, machine learning models, and any other custom
algorithms. The objective is to ensure that the code follows coding standards, is well-
documented, and adheres to best practices.
• Performance Testing: Assessing the efficiency of the code. Verify that the algorithms
and models run within acceptable time frames and do not lead to performance
bottlenecks. This may involve profiling tools to identify areas of optimization.
• Code Duplication: Checking for duplicated code, which can lead to maintenance
problems. Ensure that code is modular and reusable when necessary.
Scalability Testing: Examining how the software behaves when the number of users
or the volume of input data increases. This is important if the user is planning to
deploy the software in a real-world scenario.
Chapter 6
RESULTS & DISCUSSIONS
6.0.1 Efficiency of the Proposed System
In summary, the proposed model exhibits a high proficiency level by addressing the
nuances of mental health recognition through diverse training data, real-time
prediction capabilities, and an integrated system for practical recommendations. Its
potential application in employee monitoring and client recommendations positions
it as a valuable tool for promoting mental well-being in both individual and
organizational contexts.
• Training Data Diversity:In previous versions, they may have been trained on limited
datasets with less diversity in terms of angles, lighting conditions, and colors but in
the proposed model it is specifically trained on a more diverse dataset, capturing a
wider range of real-world conditions.
• Real-time Prediction Capability: In previous versions, they might not have included
real-time prediction capabilities, limiting their applicability to static images but in the
proposed model it is designed to work on video clippings, demonstrating an
enhanced ability to predict facial expressions in real-time, catering to dynamic and
evolving scenarios.
• Integrated System and Recommendations: In previous versions, they may not have
included an integrated system for displaying predictions and recommendations but
in the proposed model it integrates with a system that not only displays predictions
but also provides recommendations, offering a more actionable and user-friendly
interface.
f r o n t a l f a c e \ d e f a u l t . xml ’ )
roi \gray = gray [ y : y+h , x : x+w] roi \ gray = cv2 . r e s i z e ( roi \ gray , (48 , 48) , i n t e r p o l a t i
r e c t s . append ( ( x , w, y, h))
i =0
for face in a l l f a c e s :
roi = face . astype ( ” f l o a t ” ) / 255.0
. argmax ( ) ] l a b e l \ p o s i t i o n = ( r e c t s [ i ] [ 0 ] + i n t ( ( r e c t s [ i ] [ 1 ]
/2)),
abs ( r e c t s [ i ] [ 2 ] − 10) )
i =+1
cv2 . putText ( img , label , l a b e l \ position , 0) , 2)
The Fig 6.1, involves the initiation of the facial emotion recognition system, an
input image capturing the visage of a test subject is obtained, ensuring clarity and
absence of distortions in their facial expression. This input image then undergoes a
process designed to detect and recognize the nuances of the displayed emotions.
Further undergoing pre-processing steps, meticulously adjusting the image for
optimal quality, involving resizing, normalization, and noise reduction. The facial
feature detection algorithms come into play, meticulously identifying key facial
landmarks and regions of interest that contribute to the unique expression being
conveyed. The core of the system lies in deep neural network, enriched with transfer
learning capabilities, previously trained on comprehensive datasets like JAFFE. This
neural network extracts bottleneck features from the input image. Employing a pre-
trained model enhances the model’s ability to recognize intricate facial expressions.
The final result, elegantly displayed on the output interface, provides viewers with a
nuanced understanding of the detected emotion.
The Figure 6.2 displays the emotion detected which provides a comprehensive
visual representation, where the detected emotion is superimposed on the original
input image. In this specific output, a detailed analysis of the subject’s facial
expression reveals that the predominant emotion is one of confusion. This
conclusion is substantiated by the subject’s facial features, including furrowed brows,
a tightly pressed mouth, and a gaze suggesting uncertainty, all indicative of the
underlying emotional state of confusion.
The Fig 6.3, shows The actual values and predicted values are displayed in a report
format. This allows the user to determine how many emotions are appropriate.
1: Fury – 8/11 – 72%
2: Disdain – 3/4 - 75%
3: Disgust – 11/17 - 65%
4: Anxiety – 5/5 - 100%
5: Joy – 14/16 - 88%
6: Melancholy – 8/11 - 73%
7: Unexpected – 11/18 – 61%
It’s important to note that while certain emotions achieved high accuracy rates
(e.g., Surprise and Sadness), others, such as Happiness, had a lower accuracy rate.
Additionally, it is mentioned that numerous samples were incorrectly assigned to
other classes, including fear, surprise, anger, and sadness. However, fear was
accurately detected in every sample. In summary, the evaluation results provide
insights into the system’s performance for different emotions, highlighting areas of
high accuracy as well as areas that may require improvement, especially for emotions
with lower accuracy rates. The information presented in the report allows users to
assess the effectiveness of the facial emotion recognition system across various
emotional categories.
Chapter 7
7.1 Conclusion
Mental health problems are a few things which if gone unnoticed, will not only
lead to personal downfall but are also directly linked with the efficiency and energy
that a person can give to his/her work. To make this model more effective, the model
is trained with images taken on varied conditions like angles, light conditions, colors,
etc. And it recognizes the faces, which will enhance the performance of the model
so that in the future it can work on video clippings which will predict expressions in
real time and prepare an integrated system such that the predictions made by the
model would be displayed on a screen with best recommendations. The goal is to
design an end system such that when this system captures images of the employees
over a predefined period of time then based on the classifications the system could
propose some exposure to the clients who are proven to have any critical condition.
• Personalized Recommendations:
Enhance the system’s ability to provide personalized recommendations based on an
individual’s historical data and responses. This could involve tailoring interventions
and suggestions to specific needs and preferences, making the system more user-
centric.
Chapter 8
PLAGIARISM REPORT
Chapter 9
SOURCE CODE & POSTER
PRESENTATION
9.1 Source Code
37 filepath = os . path . join ( ” . / emotion\ d e t e c t o r \ models / model\ v{epoch }. hdf5 ” ) checkpoint = tf.
keras .
callbacks . ModelCheckpoint ( fi le pa th , monitor= ’ val \ accuracy ’ , verbose =1 , save\ best \ only=True ,
mode= ’max ’ )
38 callbacks = [ checkpoint ]
39 nb\ t r a i n \ samples = 28709
40 nb\ v a l i d a t i o n \ samples = 717)
41 batch\ s i z e =512
42 h i s t o r y =model . f i t \ generator ( t r a i n \ generator , steps \ per \ epoch=nb\ t r a i n \ samples batch\ size , epochs =50 , v a l i d a t i o n \
v a l i d a t i o n \ samples // batch\ s i z e )
43 fig , ( ax1 , ax2 )= p l t . subplots ( nrows =1 , ncols =2 , f i g s i z e =(20 ,6) )
44 ax1 . plot ( h i s t o r y . h i s t o r y [ ’ accuracy ’ ] , l a b e l = ’ t r a i n \ accuracy ’ )
52 importtensorflow as tf
69 roi \ gray = cv2 . r e s i z e ( roi \ gray , (48 , 48) , i n t e r p o l a t i o n =cv2 . INTER\ AREA)
72 i =0
73 for face in a l l f a c e s :
74 roi = face . astype ( ” f l o a t ” ) / 255.0 75roi = img\
to \ array ( roi )
76 roi = np . expand\ dims ( roi , axis =0)
References
[2] Kalaivani, G., & Anitha, S. Sathyapriya,DrD. (2016). A Literature review on Emotion
Recognition for Various Facial Emotional Extraction. IOSR Journal of Computer
Engineering (IOSR-JCE).
[3] Ravichandra Ginne, K., Jariwala, “FACIAL EXPRESSION RECOGNITION USING CNN: A
SURVEY”, Mar 2018, International Journal of Advances in Electronics and Computer
Science, Vol 5, No. 3, 2018.
[4] Turetsky, B. I., Kohler, C. G., Indersmitten, T., Bhati, M. T., Charbonnier, D., & Gur, R.
C. (2007). Facial emotion recognition in schizophrenia: when and why does it go
awry? Schizophrenia research, 94(1– 3), 253–263.
[5] Gibbons, R. D., Weiss, D. J., Frank, E., & Kupfer, D. (2016). Computerized Adaptive
Diagnosis and Testing of Mental Health Disorders. Annu Rev Clin Psychol, 12, 83–104,
Epub 2015 Nov 20. PMID: 26651865.
[6] Shojaeilangari, W., Yau, K., Nandakumar, J., Li, & Teoh, E. K. (July 2015). ”Robust
Representation and Recognition of Facial Emotions Using Extreme Sparse Learning”.
IEEE Transactions on Image Processing, 24(7), 2140–2152.
doi:10.1109/TIP.2015.2416634.
[9] Zhang, L., & Tjondronegoro, D. (2011). Facial expression recognition using facial
movement features. IEEE Trans Affect Comput, 2(4), 219–229.
[10] Kukla, E., & Nowak, P. (2015). Facial Emotion Recognition Based on Cascade of
Neural Networks. In A. Zgrzywa, K. Choros & A. Siemi´ nski (Eds.), New´ Research in
Multimedia and Internet Systems. Advances in Intelligent Systems and Computing
(Vol. 314). Cham: Springer.
[12] T. Ahonen, A. Hadid, and M. Pietikainen, ”Face description with local binary¨
patterns: Application to face recognition,” IEEE Transactions on Pattern Analysis and
Machine Intelligence, vol. 28, no. 12, pp. 2037-2041, 2006.
[13] P. Viola and M. Jones, ”Rapid object detection using a boosted cascade of
simple features,” in Proceedings of the 2001 IEEE Computer Society Conference on
Computer Vision and Pattern Recognition, 2001, vol. 1, pp. I-511-I-518.