House Price Prediction with ML Techniques
House Price Prediction with ML Techniques
I have a hypothesis!
fi
Example: house price prediction
• I think house price can be predicted with an unknown equation:
House Price = funknown (Size, Bedrooms, Location, etc)
<latexit sha1_base64="aW3BUaOJ591Sk0DwRh41yBGc2UI=">AAACSnicbZDPSxtBFMdnU2ttbDXq0ctgEFIIYVcFe6gg9eLBQ0SjQhLC7ORtHDI7s8y8rcZl/75eevLmH+HFQ0W8OMlGsKYPBr7zeT/mzTdMpLDo+3de6cPcx/lPC5/Li1++Li1XVlbPrE4NhxbXUpuLkFmQQkELBUq4SAywOJRwHg4PxvnzX2Cs0OoURwl0YzZQIhKcoUO9CusgXGN2qFMLtGkEh5zu0aiXFTxVQ6WvVJ7TWgFOxA3k9UL/hL7ROrav9yNdTM3rtACAPP/Wq1T9hj8JOiuCqaiSaTR7ldtOX/M0BoVcMmvbgZ9gN2MGBZeQlztu1YTxIRtA20nFYrDdbGJFTjcd6dNIG3cU0gl925Gx2NpRHLrKmOGlfZ8bw//l2ilG37uZUEmKoHjxUJRKipqOfaV9YYCjHDnBuBFuV8ovmWEcnftlZ0Lw/suz4myrEWw3to53qvs/pnYskHWyQWokILtknxySJmkRTn6Te/KXPHp/vAfvyXsuSkvetGeN/BOluRfzeLXL</latexit>
Problem: I don’t know the weights of the factors are But I have data
Example: house price prediction
• I think house price can be predicted with an unknown equation:
House Price = funknown (Size, Bedrooms, Location, etc)
<latexit sha1_base64="aW3BUaOJ591Sk0DwRh41yBGc2UI=">AAACSnicbZDPSxtBFMdnU2ttbDXq0ctgEFIIYVcFe6gg9eLBQ0SjQhLC7ORtHDI7s8y8rcZl/75eevLmH+HFQ0W8OMlGsKYPBr7zeT/mzTdMpLDo+3de6cPcx/lPC5/Li1++Li1XVlbPrE4NhxbXUpuLkFmQQkELBUq4SAywOJRwHg4PxvnzX2Cs0OoURwl0YzZQIhKcoUO9CusgXGN2qFMLtGkEh5zu0aiXFTxVQ6WvVJ7TWgFOxA3k9UL/hL7ROrav9yNdTM3rtACAPP/Wq1T9hj8JOiuCqaiSaTR7ldtOX/M0BoVcMmvbgZ9gN2MGBZeQlztu1YTxIRtA20nFYrDdbGJFTjcd6dNIG3cU0gl925Gx2NpRHLrKmOGlfZ8bw//l2ilG37uZUEmKoHjxUJRKipqOfaV9YYCjHDnBuBFuV8ovmWEcnftlZ0Lw/suz4myrEWw3to53qvs/pnYskHWyQWokILtknxySJmkRTn6Te/KXPHp/vAfvyXsuSkvetGeN/BOluRfzeLXL</latexit>
···
480000 = 1100 ⇥ w1 + 2 ⇥ w2 · · · can approximate it with small error
<latexit sha1_base64="b8tqFIe9KZOzgMh64Beo6cZY+1c=">AAADAXicjZLLahsxFIY101s6vTnpIotsRE1DoWBGY+eyCYRmk2UCdRLwGKPRHDsiGs0gnUkwg7vJq2STRUvptm+RXd+m8qVu63iRA0K/zvmEfh0pKZS0GIa/PP/R4ydPn608D168fPX6TW117cTmpRHQFrnKzVnCLSipoY0SFZwVBniWKDhNLg7G9dNLMFbm+jMOC+hmfKBlXwqOLtVb9dbjBAZSV1zJgYZ0FDS3Qhd0c48yN8coM7D0qsfoRxrNVxGNRZqjjeOgtTPno4fwUesPv72As2X49l87zQW+udTO7pxnD7GzOVXUyRh0Ou9Dr1YPG+Ek6H3BZqJOZnHUq93FaS7KDDQKxa3tsLDAbsUNSqFgFMSlhYKLCz6AjpOaOyvdavKCI/reZVLaz40bGukk+++OimfWDrPEkRnHc7tYGyeX1Tol9ne7ldRFiaDF9KB+qSjmdPwdaCoNCFRDJ7gw0nml4pwbLtB9msA1gS1e+b44iRqs2YiOW/X9T7N2rJAN8o58IIzskH1ySI5Imwjvi3fjffW++df+rf/d/zFFfW+25y35L/yfvwGHKeGu</latexit>
···
Example: house price prediction
• When we want to estimate the house price for a new house on market
• Feed the request into our function with estimated parameters
A set of methods that can automatically detect patterns in data and then
use the uncovered patterns to predict future data .
Predict or inference
• Given a training
1 X
dataset Dtrain = {(xi , yi )} , we aim to learn a model fw such <latexit sha1_base64="p7j4gVT2OC/ll+6aLeXes9lgK/4=">AAAB9XicbVDLSsNAFL2pr1pfUZduBovgqiRV0GXRjcsK9gFtLJPppB06mYSZiaWE/IcbF4q49V/c+TdO2iy09cDA4Zx7uWeOH3OmtON8W6W19Y3NrfJ2ZWd3b//APjxqqyiRhLZIxCPZ9bGinAna0kxz2o0lxaHPacef3OZ+54lKxSLxoGcx9UI8EixgBGsjPQaDtB9iPfaDdJplA7vq1Jw50CpxC1KFAs2B/dUfRiQJqdCEY6V6rhNrL8VSM8JpVuknisaYTPCI9gwVOKTKS+epM3RmlCEKImme0Giu/t5IcajULPTNZB5RLXu5+J/XS3Rw7aVMxImmgiwOBQlHOkJ5BWjIJCWazwzBRDKTFZExlphoU1TFlOAuf3mVtOs196JWv7+sNm6KOspwAqdwDi5cQQPuoAktICDhGV7hzZpaL9a79bEYLVnFzjH8gfX5A0a+kwQ=</latexit>
L(fw (xi ), yi )
|D| 2
i
3
<latexit sha1_base64="kk1CPlkNU32b6cZnq08pBq5USlo=">AAACJ3icbVDLSsNAFJ3UV62vqEs3wSJUkJJUQVdS1IULFxXsA5oSJtNJO3TyYGaihjR/48ZfcSOoiC79EydpBG09MMOZc+5l7j12QAkXuv6pFObmFxaXisulldW19Q11c6vF/ZAh3EQ+9VnHhhxT4uGmIILiTsAwdG2K2/boPPXbt5hx4ns3Igpwz4UDjzgEQSElSz01HQZRbCTx+GKcmDx0LWK6UAwRpPFVUnGs7GU78V1S+aH3iUX2DyJ5WWpZr+oZtFli5KQMcjQs9cXs+yh0sScQhZx3DT0QvRgyQRDFSckMOQ4gGsEB7krqQRfzXpztmWh7Uulrjs/k8YSWqb87YuhyHrm2rEwn5dNeKv7ndUPhnPRi4gWhwB6afOSEVBO+loam9QnDSNBIEogYkbNqaAhlcEJGW5IhGNMrz5JWrWocVmvXR+X6WR5HEeyAXVABBjgGdXAJGqAJEHgAT+AVvCmPyrPyrnxMSgtK3rMN/kD5+gZy8qd6</latexit>
•
Question: For house price
Murphy, K. P. (2012). Machine learning: a probabilistic perspective. MIT press.
prediction problem, what is
the type of the task?
Classi cation or Regression?
fi
fi
fi
fi
Linear Regression
• Linear projection from feature space to target space
• Example:
Preference level
Queens
Age
User info
Linear Regression
• Training Objective
SLIM / Linear Recommender
• Well, we don’t have user info
• User user rating matrix R as input X
• Predict user ratings matrix R as Y
Queens
Beetles
1 ?
1 ? ? Features
1 ?
1 ? ?1 Labels
? 1 ?
1 ? Data point
? ? 1 ? ?
SLIM / Linear Recommender
• Well, we don’t have user info
• User user rating matrix A as input X
• Predict user ratings matrix A as Y
Queens
Beetles
Distributed training!
• The diagonal entries of A are called the singular values of the matrix A and
T T
they are the square roots of the eigenvalues of A A and AA
T
• U are orthonormal eigenvectors of AA
T
• V are orthonormal eigenvectors of A A
Projected Linear Recommender
• Reduce feature dimension for linear recommender
Encoder
Decoder
fi
Denoised Autoencoder
900+ citations
Variational Autoencoder
• Control the user representation to t gaussian distribution
fi
Variational Autoencoder
• Control the user representation to t gaussian distribution
fi
Performance comparison so far
Take home points
• Linear Recommender — Simple, interpretable
• Projected Linear Recommender — More e cient due to reducing feature dim
• Autoencoder based Recommender — The simplest neural recommender
• Denoise Autoencoder for Recommendation — High citation, doesn’t work
well in practice
ffi