Meta tags:
Headings (most frequently used words):
density, estimation, contents, example, application, and, purpose, kernel, see, also, references, external, links,
Text of the page (most frequently used words):
the (76), density (60), and (52), #estimation (36), data (34), diabetes (26), glu (26), statistics (22), for (19), probability (17), analysis (15), kernel (15), this (13), from (13), model (13), test (12), function (12), statistical (11), edit (11), are (11), mbox (11), regression (10), estimates (10), conditional (9), distribution (9), with (8), population (8), correlation (8), plot (8), wikipedia (7), multivariate (7), transformation (7), chart (7), see (7), isbn (7), that (7), using (6), non (6), estimator (6), time (6), likelihood (6), interval (6), estimated (6), nonparametric (5), log (5), rank (5), linear (5), variance (5), standard (5), sample (5), mean (5), random (5), histogram (5), pima (5), functions (5), used (5), table (4), contents (4), search (4), additional (4), may (4), page (4), articles (4), references (4), least (4), spectral (4), frequency (4), box (4), statistic (4), smoothing (4), series (4), distributions (4), anova (4), bayesian (4), inference (4), parametric (4), power (4), method (4), experiment (4), normalization (4), processing (4), deviation (4), range (4), set (4), learning (4), links (4), 978 (4), doi (4), parzen (4), application (4), also (4), estimate (4), example (4), shown (4), cases (4), will (4), article (4), hide (4), move (4), sidebar (4), view (3), about (3), use (3), description (3), multiple (3), index (3), econometrics (3), control (3), design (3), applications (3), models (3), product (3), survival (3), squares (3), vector (3), cross (3), partial (3), tests (3), exponential (3), general (3), factor (3), adaptive (3), variable (3), posterior (3), bayes (3), way (3), friedman (3), median (3), prediction (3), unbiased (3), distance (3), moments (3), point (3), family (3), order (3), theory (3), study (3), scatter (3), one (3), records (3), new (3), york (3), sources (3), some (3), rosenblatt (3), novelty (3), detection (3), signal (3), based (3), can (3), blue (3), figure (3), including (3), purpose (3), they (3), gaussian (3), each (3), were (3), account (3), unobservable (3), tools (3), main (3), languages (2), toggle (2), code (2), contact (2), privacy (2), policy (2), terms (2), organization (2), commons (2), was (2), categories (2), all (2), needing (2), august (2), 2012 (2), short (2), different (2), wikidata (2), cs1 (2), maint (2), names (2), authors (2), list (2), densities (2), https (2), org (2), portal (2), information (2), system (2), environmental (2), national (2), engineering (2), medical (2), studies (2), clinical (2), first (2), limit (2), domain (2), autoregressive (2), fuller (2), specific (2), structural (2), seasonal (2), adjustment (2), stationarity (2), normal (2), cluster (2), principal (2), contingency (2), categorical (2), generalized (2), simple (2), equations (2), validation (2), coefficient (2), determination (2), pearson (2), moment (2), maximum (2), sum (2), lehmann (2), ratio (2), squared (2), score (2), most (2), bootstrap (2), theorem (2), minimum (2), estimating (2), optimal (2), location (2), scale (2), shape (2), sampling (2), natural (2), designs (2), trial (2), randomized (2), controlled (2), error (2), methodology (2), size (2), missing (2), collection (2), reduction (2), cleaning (2), unit (2), scaling (2), transform (2), summary (2), dispersion (2), skewness (2), mode (2), geometric (2), arithmetic (2), software (2), two (2), dimensional (2), content (2), free (2), research (2), external (2), chapman (2), hall (2), 1986 (2), silverman (2), wiley (2), practice (2), university (2), press (2), chapter (2), jerome (2), springer (2), 2001 (2), 387 (2), 95284 (2), elements (2), robert (2), tibshirani (2), trevor (2), hastie (2), ripley (2), cambridge (2), jstor (2), 1214 (2), aoms (2), annals (2), mathematical (2), illustration (2), clifton (2), 2014 (2), june (2), 2013 (2), mass (2), link (2), cite (2), journal (2), 1988 (2), mellitus (2), indian (2), women (2), self (2), 2011 (2), pages (2), kde (2), made (2), such (2), who (2), usually (2), its (2), current (2), form (2), when (2), which (2), improve (2), 100 (2), distributed (2), very (2), anomaly (2), more (2), important (2), conclusions (2), obtained (2), other (2), give (2), then (2), true (2), second (2), shows (2), associated (2), construct (2), according (2), red (2), black (2), underlying (2), thought (2), gaussians (2), centered (2), curve (2), learn (2), help (2), citations (2), appearance (2), upload (2), file (2), changes (2), history (2), read (2), create (2), donate (2), menu (2), add, topic, mobile, cookie, statement, developers, conduct, legal, safety, contacts, disclaimers, text, available, under, apply, site, you, agree, registered, trademark, profit, wikimedia, foundation, inc, creative, attribution, sharealike, license, last, edited, april, 2026, utc, hidden, excerpts, retrieved, php, title, density_estimation, oldid, 1347599606, wikiproject, mathematics, category, kriging, geostatistics, geographic, cartography, spatial, psychometrics, official, accounts, jurimetrics, demography, crime, census, actuarial, science, social, identification, reliability, quality, process, probabilistic, methods, chemometrics, epidemiology, trials, bioinformatics, biostatistics, nelson, aalen, hazard, hitting, accelerated, failure, aft, proportional, hazards, kaplan, meier, whittle, wavelet, fourier, autoregression, var, heteroskedasticity, arch, arima, jenkins, arma, xcf, pacf, autocorrelation, acf, breusch, godfrey, durbin, watson, ljung, johansen, dickey, granger, causality, break, cointegration, trend, decomposition, elliptical, equation, classification, discriminant, canonical, components, manova, cochran, mantel, haenszel, mcnemar, graphical, cohen, kappa, degrees, freedom, covariance, partition, poisson, regressions, binomial, logistic, bernoulli, families, homoscedasticity, heteroscedasticity, robust, isotonic, semiparametric, nonlinear, predictors, ordinary, template, splines, mars, simultaneous, mixed, effects, errors, residuals, confounding, credible, prior, van, der, waerden, ordered, alternative, jonckheere, terpstra, kruskal, wallis, mann, whitney, hodges, signed, wilcoxon, sign, bic, aic, selection, normality, shapiro, wilk, jarque, bera, lilliefors, anderson, darling, kolmogorov, smirnov, chi, goodness, fit, student, wald, lagrange, multiplier, comparisons, randomization, permutation, uniformly, powerful, tails, testing, hypotheses, jackknife, resampling, tolerance, pivot, confidence, plug, scheffé, rao, blackwellization, estimators, frequentist, robustness, asymptotics, divergence, efficiency, loss, decision, functional, sufficiency, completeness, monotone, parameter, space, specification, empirical, quasi, sectional, cohort, observational, down, stochastic, approximation, scientific, assignment, interaction, factorial, blocking, experiments, questionnaire, opinion, poll, stratified, survey, replication, effect, detrending, differencing, preprocessing, component, dimensionality, truncation, winsorizing, outlier, min, max, standardization, feature, fisher, anscombe, stabilizing, yeo, johnson, cox, transformations, line, ecdf, matrix, heatmap, violin, stem, leaf, display, run, radar, pie, forest, fan, correlogram, biplot, bar, graphics, spearman, kendall, dependence, grouped, tables, count, kurtosis, central, percentile, interquartile, variation, average, absolute, lehmer, heinz, heronian, harmonic, cubic, contraharmonic, center, continuous, descriptive, outline, libagf, matlab, indians, database, original, 732, notes, uci, machine, repository, downloads, packages, wildlife, assessment, ruwpa, wisp, creem, centre, into, ecological, modelling, london, 412, 24620, scott, 1992, visualization, jeffrey, racine, princeton, 2007, 691, 12161, brian, 1996, 0521460866, pattern, recognition, neural, networks, 46809224, oclc, mining, 200, full, color, illustrations, 1962, 1076, 2237880, 1177704472, 1065, 1956, 837, 1177728190, 832, remarks, histograms, pimentel, marco, david, lei, tarassenko, lionel, january, review, 249, 1016, sigpro, 026, 215, geof, givens, computational, 330, 470, 53331, calculator, 0412246203, support, datasets, venables, smith, everhart, dickson, knowler, johannes, greenes, los, alamitos, 265, 2245318, pmc, 261, proceedings, symposium, computer, care, washington, adap, algorithm, forecast, onset, documentation, alberto, bernacchia, simone, pigolotti, consistent, royal, society, volume, issue, 407, 422, 1111, 1467, 9868, 00772, fitting, generative, embedding, integrated, answers, fundamental, problem, where, inferences, finite, fields, termed, window, after, credited, independently, creating, famous, class, accuracy, naive, classifier, marginal, murray, emanuel, weights, kernels, bandwidths, numbers, normally, section, excerpt, rainfall, river, discharge, analysed, gain, insight, their, behaviour, occurrence, hydrology, frequently, observation, lies, low, region, likely, examples, illustrating, exploratory, presentational, purposes, case, bivariate, aspect, often, presentation, back, client, provide, explanation, possibly, have, been, means, ideal, reason, fairly, easily, comprehensible, mathematicians, gumbel, informal, investigation, properties, given, valuable, indication, features, multimodality, yield, regarded, evidently, while, others, further, these, appears, increased, level, displaystyle, frac, obtain, via, brevity, abbreviated, formula, rule, placed, computed, over, 143, 110, greater, levels, clearer, plots, package, within, programming, language, three, concentration, presence, absence, third, not, glucose, plasma, years, old, heritage, living, near, phoenix, arizona, tested, criteria, collected, institute, digestive, kidney, diseases, 532, complete, world, health, consider, incidence, following, quoted, verbatim, variety, approaches, techniques, basic, rescaled, quantization, clustering, windows, simply, construction, observed, large, demonstration, mixture, around, solid, frame, samples, generated, drawn, gray, averaging, yields, dashed, how, remove, message, please, unsourced, material, challenged, removed, scholar, books, newspapers, news, find, adding, reliable, needs, concept, encyclopedia, item, projects, printable, version, download, pdf, print, export, get, shortened, url, permanent, related, what, here, actions, english, talk, українська, sunda, فارسی, català, top, personal, special, recent, community, contribute, events, navigation, jump,
Text of the page (random words):
ly density estimation is the construction of an estimate based on observed data of an unobservable underlying probability density function the unobservable density function is thought of as the density according to which a large population is distributed the data are usually thought of as a random sample from that population 1 a variety of approaches to density estimation are used including parzen windows and a range of data clustering techniques including vector quantization the most basic form of density estimation is a rescaled histogram example edit estimated density of p glu diabetes 1 red p glu diabetes 0 blue and p glu black estimated probability of p diabetes 1 glu estimated probability of p diabetes 1 glu we will consider records of the incidence of diabetes the following is quoted verbatim from the data set description a population of women who were at least 21 years old of pima indian heritage and living near phoenix arizona was tested for diabetes mellitus according to world health organization criteria the data were collected by the us national institute of diabetes and digestive and kidney diseases we used the 532 complete records 2 3 in this example we construct three density estimates for glu plasma glucose concentration one conditional on the presence of diabetes the second conditional on the absence of diabetes and the third not conditional on diabetes the conditional density estimates are then used to construct the probability of diabetes conditional on glu the glu data were obtained from the mass package 4 of the r programming language within r pima tr and pima te give a fuller account of the data the mean of glu in the diabetes cases is 143 1 and the standard deviation is 31 26 the mean of glu in the non diabetes cases is 110 0 and the standard deviation is 24 29 from this we see that in this data set diabetes cases are associated with greater levels of glu this will be made clearer by plots of the estimated density functions the first figure shows density estimates of p glu diabetes 1 p glu diabetes 0 and p glu the density estimates are kernel density estimates using a gaussian kernel that is a gaussian density function is placed at each data point and the sum of the density functions is computed over the range of the data from the density of glu conditional on diabetes we can obtain the probability of diabetes conditional on glu via bayes rule for brevity diabetes is abbreviated db in this formula p diabetes 1 glu p glu db 1 p db 1 p glu db 1 p db 1 p glu db 0 p db 0 displaystyle p mbox diabetes 1 mbox glu frac p mbox glu mbox db 1 p mbox db 1 p mbox glu mbox db 1 p mbox db 1 p mbox glu mbox db 0 p mbox db 0 the second figure shows the estimated posterior probability p diabetes 1 glu from these data it appears that an increased level of glu is associated with diabetes application and purpose edit a very natural use of density estimates is in the informal investigation of the properties of a given set of data density estimates can give a valuable indication of such features as skewness and multimodality in the data in some cases they will yield conclusions that may then be regarded as self evidently true while in others all they will do is to point the way to further analysis and or data collection 5 histogram and density function for a gumbel distribution 6 an important aspect of statistics is often the presentation of data back to the client in order to provide explanation and illustration of conclusions that may possibly have been obtained by other means density estimates are ideal for this purpose for the simple reason that they are fairly easily comprehensible to non mathematicians more examples illustrating the use of density estimates for exploratory and presentational purposes including the important case of bivariate data 7 density estimation is also frequently used in anomaly detection or novelty detection 8 if an observation lies in a very low density region it is likely to be an anomaly or a novelty in hydrology the histogram and estimated density function of rainfall and river discharge data analysed with a probability distribution are used to gain insight in their behaviour and frequency of occurrence 9 an example is shown in the blue figure kernel density estimation edit this section is an excerpt from kernel density estimation edit kernel density estimation of 100 normally distributed random numbers using different smoothing bandwidths in statistics kernel density estimation kde is the application of kernel smoothing for probability density estimation i e a non parametric method to estimate the probability density function of a random variable based on kernels as weights kde answers a fundamental data smoothing problem where inferences about the population are made based on a finite data sample in some fields such as signal processing and econometrics it is also termed the parzen rosenblatt window method after emanuel parzen and murray rosenblatt who are usually credited with independently creating it in its current form 10 11 one of the famous applications of kernel density estimation is in estimating the class conditional marginal densities of data when using a naive bayes classifier which can improve its prediction accuracy 12 see also edit frequency distribution kernel density estimation mean integrated squared error histogram multivariate kernel density estimation spectral density estimation kernel embedding of distributions generative model application of order statistics non parametric density estimation probability distribution fitting references edit alberto bernacchia simone pigolotti self consistent method for density estimation journal of the royal statistical society series b statistical methodology volume 73 issue 3 june 2011 pages 407 422 https doi org 10 1111 j 1467 9868 2011 00772 x diabetes in pima indian women r documentation smith j w everhart j e dickson w c knowler w c and johannes r s 1988 r a greenes ed using the adap learning algorithm to forecast the onset of diabetes mellitus proceedings of the symposium on computer applications in medical care washington 1988 los alamitos ca 261 265 pmc 2245318 cite journal cs1 maint multiple names authors list link support functions and datasets for venables and ripley s mass silverman b w 1986 density estimation for statistics and data analysis chapman and hall isbn 978 0412246203 a calculator for probability distributions and density functions geof h givens 2013 computational statistics wiley p 330 isbn 978 0 470 53331 4 pimentel marco a f clifton david a clifton lei tarassenko lionel 2 january 2014 a review of novelty detection signal processing 99 june 2014 215 249 doi 10 1016 j sigpro 2013 12 026 an illustration of histograms and probability density functions rosenblatt m 1956 remarks on some nonparametric estimates of a density function the annals of mathematical statistics 27 3 832 837 doi 10 1214 aoms 1177728190 parzen e 1962 on estimation of a probability density function and mode the annals of mathematical statistics 33 3 1065 1076 doi 10 1214 aoms 1177704472 jstor 2237880 hastie trevor tibshirani robert friedman jerome h 2001 the elements of statistical learning data mining inference and prediction with 200 full color illustrations new york springer isbn 0 387 95284 5 oclc 46809224 sources brian d ripley 1996 pattern recognition and neural networks cambridge cambridge university press isbn 978 0521460866 trevor hastie robert tibshirani and jerome friedman the elements of statistical learning new york springer 2001 isbn 0 387 95284 5 see chapter 6 qi li and jeffrey s racine nonparametric econometrics theory and practice princeton university press 2007 isbn 0 691 12161 3 see chapter 1 d w scott multivariate density estimation theory practice and visualization new york wiley 1992 b w silverman density estimation london chapman and hall 1986 isbn 978 0 412 24620 3 external links edit creem centre for research into ecological and environmental modelling downloads for free density estimation software packages distance 4 from research unit for wildlife population assessment ruwpa and wisp uci machine learning repository content summary see pima indians diabetes database for the original data set of 732 records and additional notes matlab code for one dimensional and two dimensional density estimation libagf c software for variable kernel density estimation v t e statistics outline index descriptive statistics continuous data center mean arithmetic arithmetic geometric contraharmonic cubic generalized power geometric harmonic heronian heinz lehmer median mode dispersion average absolute deviation coefficient of variation interquartile range percentile range standard deviation variance shape central limit theorem moments kurtosis l moments skewness count data index of dispersion summary tables contingency table frequency distribution grouped data dependence partial correlation pearson product moment correlation rank correlation kendall s τ spearman s ρ scatter plot graphics bar chart biplot box plot control chart correlogram fan chart forest plot histogram pie chart q q plot radar chart run chart scatter plot stem and leaf display violin plot heatmap scatter plot matrix ecdf plot line chart statistical data processing transformations data transformation log transformation power transform box cox transformation yeo johnson transformation variance stabilizing transformation anscombe transform fisher transformation scaling and normalization feature scaling normalization standardization z score min max normalization unit vector normalization data cleaning data cleaning outlier winsorizing truncation missing data data reduction dimensionality reduction principal component analysis factor analysis time series preprocessing differencing detrending seasonal adjustment stationarity transformation data collection study design effect size missing data optimal design population replication sample size determination statistic statistical power survey methodology sampling cluster stratified opinion poll questionnaire standard error controlled experiments blocking factorial experiment interaction random assignment randomized controlled trial randomized experiment scientific control adaptive designs adaptive clinical trial stochastic approximation up and down designs observational studies cohort study cross sectional study natural experiment quasi experiment statistical inference statistical theory population statistic probability distribution sampling distribution order statistic empirical distribution density estimation statistical model model specification l p space parameter location scale shape parametric family likelihood monotone location scale family exponential family completeness sufficiency statistical functional bootstrap u v optimal decision loss function efficiency statistical distance divergence asymptotics robustness frequentist inference point estimation estimating equations maximum likelihood method of moments m estimator minimum distance unbiased estimators mean unbiased minimum variance rao blackwellization lehmann scheffé theorem median unbiased plug in interval estimation confidence interval pivot likelihood interval prediction interval tolerance interval resampling bootstrap jackknife testing hypotheses 1 2 tails power uniformly most powerful test permutation test randomization test multiple comparisons parametric tests likelihood ratio score lagrange multiplier wald specific tests z test normal student s t test f test goodness of fit chi squared g test kolmogorov smirnov anderson darling lilliefors jarque bera normality shapiro wilk likelihood ratio test model selection cross validation aic bic rank statistics sign sample median signed rank wilcoxon hodges lehmann estimator rank sum mann whitney nonparametric anova 1 way kruskal wallis 2 way friedman ordered alternative jonckheere terpstra van der waerden test bayesian inference bayesian probability prior posterior credible interval bayes factor bayesian estimator maximum posterior estimator correlation regression analysis correlation pearson product moment partial correlation confounding variable coefficient of determination regression analysis errors and residuals regression validation mixed effects models simultaneous equations models multivariate adaptive regression splines mars template least squares and regression analysis linear regression simple linear regression ordinary least squares general linear model bayesian regression non standard predictors nonlinear regression nonparametric semiparametric isotonic robust homoscedasticity and heteroscedasticity generalized linear model exponential families logistic bernoulli binomial poisson regressions partition of variance analysis of variance anova anova analysis of covariance multivariate anova degrees of freedom categorical multivariate time series survival analysis categorical cohen s kappa contingency table graphical model log linear model mcnemar s test cochran mantel haenszel statistics multivariate regression manova principal components canonical correlation discriminant analysis cluster analysis classification structural equation model factor analysis multivariate distributions elliptical distributions normal time series general decomposition trend stationarity seasonal adjustment exponential smoothing cointegration structural break granger causality specific tests dickey fuller johansen q statistic ljung box durbin watson breusch godfrey time domain autocorrelation acf partial pacf cross correlation xcf arma model arima model box jenkins autoregressive conditional heteroskedasticity arch vector autoregression var autoregressive model ar frequency domain spectral density estimation fourier analysis least squares spectral analysis wavelet whittle likelihood survival survival function kaplan meier estimator product limit proportional hazards models accelerated failure time aft model first hitting time hazard function nelson aalen estimator test log rank test applications biostatistics bioinformatics clinical trials studies epidemiology medical statistics engineering statistics chemometrics methods engineering probabilistic design process quality control reliability system identification social statistics actuarial science census crime statistics demography econometrics jurimetrics national accounts official statistics population statistics psychometrics spatial statistics cartography environmental statistics geographic information system geostatistics kriging category mathematics portal commons wikiproject retrieved from https en wikipedia org w index php title density_estimation oldid 1347599606 categories estimation of densities nonparametric statistics hidden categories cs1 maint multiple names authors list articles with short description short description is different from wikidata articles needing additional references from august 2012 all articles needing additional references articles w...
|