官术网_书友最值得收藏!

Feature learning – using AI to better our AI

The cherry on top, a cherry powered by the most sophisticated algorithms used today in the automatic construction of features for the betterment of machine learning and AI pipelines.

The previous chapter dealt with automatic feature creation using mathematical formulas, but once again, in the end, it is us, the humans, that choose the formulas and reap the benefits of them. This chapter will outline algorithms that are not in and of themselves a mathematical formula, but an architecture attempting to understand and model data in such a way that it will exploit patterns in data in order to create new data. This may sound vague at the moment, but we hope to get you excited about it!

We will focus mainly on neural algorithms that are specially designed to use a neural network design (nodes and weights). These algorithms will then impose features onto the data in such a way that can sometimes be unintelligible to humans, but extremely useful for machines. Some of the topics we'll look at are:

  • Restricted Boltzmann machines
  • Word2Vec/GLoVe for word embedding

Word2Vec and GLoVe are two ways of adding large dimensionality data to seemingly word tokens in the text. For example, if we look at a visual representation of the results of a Word2Vec algorithm, we might see the following:

By representing words as vectors in Euclidean space, we can achieve mathematical-esque results. In the previous example, by adding these automatically generated features we can add and subtract words by adding and subtracting their vector representations as given to us by Word2Vec. We can then generate interesting conclusions, such as king+man-woman=queen. Cool!

主站蜘蛛池模板: 吉隆县| 西盟| 蕉岭县| 四平市| 温州市| 栾城县| 武冈市| 巫溪县| 永登县| 昂仁县| 连平县| 秦皇岛市| 洪江市| 彭山县| 温泉县| 绥阳县| 孝昌县| 涟源市| 莒南县| 高要市| 任丘市| 凌源市| 梨树县| 武威市| 百色市| 迁西县| 横峰县| 肇源县| 台东市| 阿拉善左旗| 台东市| 柳林县| 遂溪县| 克什克腾旗| 天长市| 安阳市| 磐安县| 友谊县| 孟连| 江陵县| 伊春市|