官术网_书友最值得收藏!

Feature learning – using AI to better our AI

The cherry on top, a cherry powered by the most sophisticated algorithms used today in the automatic construction of features for the betterment of machine learning and AI pipelines.

The previous chapter dealt with automatic feature creation using mathematical formulas, but once again, in the end, it is us, the humans, that choose the formulas and reap the benefits of them. This chapter will outline algorithms that are not in and of themselves a mathematical formula, but an architecture attempting to understand and model data in such a way that it will exploit patterns in data in order to create new data. This may sound vague at the moment, but we hope to get you excited about it!

We will focus mainly on neural algorithms that are specially designed to use a neural network design (nodes and weights). These algorithms will then impose features onto the data in such a way that can sometimes be unintelligible to humans, but extremely useful for machines. Some of the topics we'll look at are:

  • Restricted Boltzmann machines
  • Word2Vec/GLoVe for word embedding

Word2Vec and GLoVe are two ways of adding large dimensionality data to seemingly word tokens in the text. For example, if we look at a visual representation of the results of a Word2Vec algorithm, we might see the following:

By representing words as vectors in Euclidean space, we can achieve mathematical-esque results. In the previous example, by adding these automatically generated features we can add and subtract words by adding and subtracting their vector representations as given to us by Word2Vec. We can then generate interesting conclusions, such as king+man-woman=queen. Cool!

主站蜘蛛池模板: 北海市| 花莲县| 阳春市| 邯郸市| 弋阳县| 肥城市| 锦屏县| 越西县| 绥阳县| 光山县| 茶陵县| 印江| 河津市| 安徽省| 达孜县| 海南省| 霍林郭勒市| 丰城市| 建阳市| 永顺县| 扎鲁特旗| 巍山| 宁阳县| 江北区| 常德市| 定襄县| 松滋市| 哈尔滨市| 盈江县| 淮北市| 蕉岭县| 韶山市| 夹江县| 新野县| 炉霍县| 玉树县| 化州市| 呼和浩特市| 阳谷县| 商南县| 伊通|