Machine Learning, Knowledge Discovery in Databases (KDD) and Data Mining
Two terms are commonly confused, as they often employ the same methods and overlap significantly. They can be roughly defined as follows:
- Machine learning focuses on prediction, based on known properties learned from the training data.
- Data mining (which is the analysis step of Knowledge Discovery in Databases) focuses on the discovery of (previously) unknown properties on the data.
The two areas overlap in many ways: data mining uses many machine learning methods, but often with a slightly different goal in mind. On the other hand, machine learning also employs data mining methods as "unsupervised learning" or as a preprocessing step to improve learner accuracy. Much of the confusion between these two research communities (which do often have separate conferences and separate journals, ECML PKDD being a major exception) comes from the basic assumptions they work with: in machine learning, performance is usually evaluated with respect to the ability to reproduce known knowledge, while in KDD the key task is the discovery of previously unknown knowledge. Evaluated with respect to known knowledge, an uninformed (unsupervised) method will easily be outperformed by supervised methods, while in a typical KDD task, supervised methods cannot be used due to the unavailability of training data.
Read more about this topic: Machine Learning
Famous quotes containing the words machine, knowledge, discovery, data and/or mining:
“The white man regards the universe as a gigantic machine hurtling through time and space to its final destruction: individuals in it are but tiny organisms with private lives that lead to private deaths: personal power, success and fame are the absolute measures of values, the things to live for. This outlook on life divides the universe into a host of individual little entities which cannot help being in constant conflict thereby hastening the approach of the hour of their final destruction.”
—Policy statement, 1944, of the Youth League of the African National Congress. pt. 2, ch. 4, Fatima Meer, Higher than Hope (1988)
“At no time in history ... have the people who are not fit for society had such a glorious opportunity to pretend that society is not fit for them. Knowledge of the slums is at present a passport to societyso much the parlor philanthropists have achievedand all they have to do is to prove that they know their subject. It is an odd qualification to have pitched on; but gentlemen and ladies are always credulous, especially if you tell them that they are not doing their duty.”
—Katharine Fullerton Gerould (18791944)
“One of the laudable by-products of the Freudian quackery is the discovery that lying, in most cases, is involuntary and inevitablethat the liar can no more avoid it than he can avoid blinking his eyes when a light flashes or jumping when a bomb goes off behind him.”
—H.L. (Henry Lewis)
“This city is neither a jungle nor the moon.... In long shot: a cosmic smudge, a conglomerate of bleeding energies. Close up, it is a fairly legible printed circuit, a transistorized labyrinth of beastly tracks, a data bank for asthmatic voice-prints.”
—Susan Sontag (b. 1933)
“Any relation to the land, the habit of tilling it, or mining it, or even hunting on it, generates the feeling of patriotism. He who keeps shop on it, or he who merely uses it as a support to his desk and ledger, or to his manufactory, values it less.”
—Ralph Waldo Emerson (18031882)