Statistics

Domain — 397 words · page 1 of 4

This page gathers the 397 dictionary entries belonging to the domain “Statistics”, as labelled by Wiktionary. Each word leads to its full entry: definitions, etymology, pronunciation, examples. Page 1 of 4: from “-ile” to “deviate”.

  1. -ile suffix Any of the values in a sorted data set that splits it into a specified number of equally sized groups.
  2. aa noun Initialism of arithmetic average.
  3. absolute deviation noun The absolute value of the difference between a given value (such as mean or expected value) and a variate value (usually an observed value).
  4. absorbing adj Allowing a process to enter it, but not to leave it.
  5. algebraic statistics noun A discipline within mathematical statistics in which algebra is used to describe and analyse problems;
  6. alpha noun The significance level of a statistical test; the alpha level.
  7. alternative hypothesis noun A rival hypothesis to the null hypothesis. Their likelihoods are compared by a statistical hypothesis test.
  8. analysis of variance noun A collection of statistical models, and their associated procedures, in which the observed variance is partitioned into components due to…
  9. Anderson-Darling test noun A statistical test of whether a given sample of data is drawn from a given probability distribution.
  10. anentropy noun A measure of the simplicity or purity of a distribution.
  11. anticorrelation noun Negative or inverse correlation, a relationship in which one value increases as the other decreases.
  12. antithetic noun Ellipsis of antithetic variate.
  13. arithmetic mean noun The measure of central tendency of a set of values computed by dividing the sum of the values by their number.
  14. association noun Any relationship between two measured quantities that renders them statistically dependent (but not necessarily causal or a correlation).
  15. atomless adj Not assigning strictly positive probability to any single outcome.
  16. autocorrelation noun The cross-correlation of a signal with itself: the correlation between values of a signal in successive time periods.
  17. average noun Any measure of central tendency, especially any mean, the median, or the mode.
  18. backshift noun A function on a stochastic process whose value on any random variable is the preceding random variable in the process.
  19. bagplot noun A method for visualizing two- or three-dimensional statistical data, analogous to the one-dimensional box plot.
  20. balance noun The remainder.
  21. base year noun a year chosen for comparison with other years, in statistical analysis and various indices, such as a price index. It usually has an…
  22. Bayesian noun A proponent of Bayesianism.
  23. Bayesian network noun A directed acyclic graph whose vertices represent random variables and whose directed edges represent conditional dependencies. Each random…
  24. Bayesianism noun An approach to probability which quantifies expectations based on given data, following Bayes' theorem.
  25. Behrens-Fisher problem noun The problem of interval estimation and hypothesis testing concerning the difference between the means of two normally distributed…
  26. Bernoulli distribution noun A discrete probability distribution that represents the result of a single trial, taking value 1 with "success" probability p and value 0…
  27. Bernoulli process noun A series of independent trials, analogous to repeatedly flipping a coin to determine whether it is fair.
  28. beta distribution noun Any of a family of continuous probability distributions defined on the interval [0, 1] whose shape is parametrised by two positive…
  29. bias noun The difference between the expectation of the sample estimator and the true population value, which reduces the representativeness of the…
  30. bias-robust adj Describing statistics that are relatively unbiased even when some assumptions about the data or the model do not hold true.
  31. biased adj Exhibiting a systematic distortion of results due to a factor not allowed for in its derivation; skewed.
  32. bin noun Any of the discrete intervals in a histogram, etc
  33. binarize verb To dichotomize a variable.
  34. binary distribution noun Bernoulli distribution.
  35. binning noun A data pre-processing technique in which original data values fall into a small interval ("bin") and are replaced by a value representative…
  36. binomial adj Of or relating to the binomial distribution.
  37. binomial distribution noun The discrete probability distribution of the number of successes in a sequence of n independent trials, each of which yields success with…
  38. binomial test noun A test used to determine if a proportion of a binary variable is equivalent to a hypothesized value.
  39. bipower noun An alternative to variance; a measure of variability based on summing the product of adjacent data points rather than summing the squares…
  40. birthday effect noun A statistical phenomenon where an individual's likelihood of death appears to increase on or close to their birthday, variously attributed…
  41. biserial adj Having a correlation between one series and another that is divided into two types
  42. bivariate normal distribution noun The joint probability distribution of two variables that are both individually normally distributed.
  43. Bonferroni correction noun A method of counteracting the multiple comparisons problem, and thus reducing the risk of type I error, by testing each individual…
  44. Bonferroni inequality noun Any of certain upper and lower bounds on the probability of finite unions of events, serving as a generalization of Boole's inequality.
  45. bootstrap noun Any method or instance of estimating properties of an estimator (such as its variance) by measuring those properties when sampling from an…
  46. bootstrapper noun One who uses bootstrap methods.
  47. box and whiskers plot noun A graphical summary of a numerical data sample through five statistics — median, upper quartile, lower quartile, and upper extreme and…
  48. box plot noun A graphical summary of a numerical data sample through five statistics: median, lower quartile, upper quartile, and some indication of more…
  49. Box-Cox distribution noun The distribution of a random variable X for which the Box-Cox transformation on X follows a truncated normal distribution.
  50. breakdown point noun The number or proportion of arbitrarily large or small extreme values that must be introduced into a batch or sample to cause the estimator…
  51. case fatality rate noun A percentage derived by the division of the number of fatalities by the number of confirmed (diagnosed) cases (sufferers) of a disease.
  52. Cauchy distribution noun A symmetric continuous probability distribution with fat tails, with probability density function that is
  53. ceiling effect noun The phenomenon where an independent variable no longer has an effect on a dependent variable, or the level above which variance in an…
  54. cell noun The unit in a statistical array (a spreadsheet, for example) where a row and a column intersect.
  55. censor verb To partially obscure an observation.
  56. censoring noun A partial obscuring of data points.
  57. centerpoint noun A generalization of the median to data in higher-dimensional Euclidean space.
  58. centroid noun the arithmetic mean (alternatively, median) position of a cluster of points in a coordinate system based on some application-dependent…
  59. Chapman-Robbins bound noun A lower bound on the variance of estimators of a deterministic parameter; a generalization of the Cramér-Rao bound, it is both tighter and…
  60. chunklet noun A group of data points belonging to the same constraint cluster.
  61. class noun A grouping of data values in an interval, often used for computation of a frequency distribution.
  62. cluster noun In cluster analysis: a subset of a population whose members are similar enough to each other and distinct from others as to be considered a…
  63. cluster analysis noun The classification of objects into different groups, or more precisely, the partitioning of a data set into subsets (clusters), so that the…
  64. clustering illusion noun The tendency to erroneously perceive random streaks and clusters arising in small samples from random distributions as being non-random.
  65. coefficient of determination noun The proportion of the variation in the dependent variable that is predictable from the independent variable(s), providing a measure of how…
  66. cohort noun A demographic grouping of people, especially those in a defined age group, or having a common characteristic.
  67. cohort effect noun The phenomenon in which research outcomes may depend on the characteristics and experiences shared within each particular cohort.
  68. cointegrate verb To exhibit cointegration.
  69. compressed sensing noun A signal processing technique for efficiently acquiring and reconstructing a signal, by finding solutions to underdetermined linear systems.
  70. concept drift noun The tendency of a statistical model to become less and less accurate over time in real-world conditions.
  71. confidence interval noun A range defined such that there is a fixed probability that the value of the measured parameter will fall within the range.
  72. confidence limits noun For a statistical sample a pair of values that delimit the interval for which there is a certain probability that the true value of some…
  73. confound noun A confounding variable.
  74. confounding variable noun An extraneous variable in a statistical model that correlates (positively or negatively) with both the dependent variable and the…
  75. constraint cluster noun a cluster of data points in a set conforming to the must-link and cannot-link constraints specified for each pair of data points
  76. contactive adj Contiguous but not coordinated
  77. contingency table noun A table presenting the joint distribution of two categorical variables.
  78. control verb (construed with for) To design (an experiment) so that the effects of one or more variables are reduced or eliminated.
  79. copula noun A function that represents the association between two or more variables, independent of the individual marginal distributions of the…
  80. correlation noun One of the several measures of the linear statistical relationship between two random variables, indicating both the strength and direction…
  81. correlation coefficient noun Any of the several measures indicating the strength and direction of a linear relationship between two random variables.
  82. correlator noun A correlation function
  83. covariable adj Possibly predictive of the outcome under study.
  84. covariance noun A statistical measure defined as scriptstyle operatorname Cov(X,Y)= operatorname E((X-μ)(Y-ν)) given two real-valued random variables X and…
  85. Cramér-Rao bound noun A lower bound on the variance of unbiased estimators of a deterministic (fixed, though unknown) parameter.
  86. cross section noun A sample meant to be representative of a whole population.
  87. cross tabulation noun A presentation of data about categorical variable in a tabular form to aid in identifying a relationship between the variables.
  88. cross-validation noun Any technique or instance of assessing how the results of a statistical analysis will generalize to an independent dataset.
  89. crude adj Not adjusted or further analyzed.
  90. cumulative error noun A statistical or measurement error that compounds with further calculation, analysis, etc.
  91. CV noun Initialism of cross-validation.
  92. data set noun A set of data to be analyzed.
  93. decile noun Any of the values in a series that divides the distribution of individuals in that series into ten groups of equal frequency.
  94. demean verb To subtract the mean from (a value, or every observation in a data set).
  95. demeaned adj Having had the mean subtracted from all values.
  96. density noun The probability that an outcome will fall into a given range, per unit of that range; the relative likelihood of possible values of a…
  97. dependent adj Having a probability that is affected by the outcome of a separate event.
  98. depth noun the lower of the two ranks of a value in an ordered set of values
  99. descriptive statistics noun A branch of statistics dealing with summarization and description of collections of data—data sets, including the concepts of arithmetic…
  100. deviate noun A value equal to the difference between a measured variable factor and a fixed or algorithmic reference value.

All domains · Search for a word