Practical algorithms for on-line sampling

dc.creatorDomingo, Carlos
dc.creatorGavalda, Ricard
dc.creatorWatanabe, Osamu
dc.date1998-09-30
dc.date.accessioned2026-07-07T03:23:42Z
dc.date.available2026-07-07T03:23:42Z
dc.descriptionOne of the core applications of machine learning to knowledge discovery consists on building a function (a hypothesis) from a given amount of data (for instance a decision tree or a neural network) such that we can use it afterwards to predict new instances of the data. In this paper, we focus on a particular situation where we assume that the hypothesis we want to use for prediction is very simple, and thus, the hypotheses class is of feasible size. We study the problem of how to determine which of the hypotheses in the class is almost the best one. We present two on-line sampling algorithms for selecting hypotheses, give theoretical bounds on the number of necessary examples, and analize them exprimentally. We compare them with the simple batch sampling approach commonly used and show that in most of the situations our algorithms use much fewer number of examples.
dc.descriptionTo appear in the Proc. of Discovery Science '98, Dec. 1998
dc.identifierhttps://arxiv.org/abs/cs/9809122
dc.identifierhttp://arxiv.org/abs/cs/9809122
dc.identifier.urihttp://salesiana.dossiersoluciones.com/handle/123456789/33049
dc.subjectMachine Learning
dc.subjectData Structures and Algorithms
dc.subjectI.2.6;H.2.8
dc.titlePractical algorithms for on-line sampling
dc.typetext

Files

Collections