Uniting Psychometric Modelling and Poisson Distributions: A Metrological Study of Elementary Counting
Abstract
1. Introduction
2. PMFs and Measurement
2.1. Probability Mass Functions
- ProbabilitiesThe probabilities can be either derived from actual frequencies (occupancy per bin) or estimated in terms of prior knowledge or even judgment when scores are set. Examples of PMFs are given: for (i) an analytical case, (Section 3.3); (ii) for a binomial distribution (Section 4.1); (iii) for a Poisson distribution (Section 4.2); and (iv) for psychometric quantities (Section 5.4). When setting response scores as probabilities P, choices have to be made about how to define the scale-end scores of 0 and 100. Inferentially stable measurement also requires that the score varies monotonically across the scale in a manner allowing summarisation, which exhausts the data of all available information in a minimally sufficient statistic [19,20,21]. (These requirements [22] are subsequently checked (Section 5.5).)
- Ranges
- –
- Intrinsically discrete ranges, (Equation (1)) (such as when counting dots (Section 3.3)).
- –
- Discrete ranges chosen for convenience (such as in sensory panel responses), where the observed variable is on a continuous scale but the large uncertainties make it practical to round off the score, say, to the nearest integer.
- –
- Ranges can be as short as two, such as for the binomial distribution from binary Bernoulli trials (Section 4.1), as well as long, polytomous ranges spanning several categories (Section 4.3.2).
- ScalesIn different application areas, the scales associated with both X (abscissa axis) and P (ordinate axis) of a PMF can be
- –
- Fully quantitative (Section 3.3);
- –
- More qualitative (Section 5).
These scales can range, respectively, from ratio and interval scales to ordinal and nominal scales.
2.2. Accuracy of Classification
- Analytical accuracy.In the ‘analytical’ scenario 1, the ‘accuracy’ estimation of the ‘measurand’, the quantitative number of each set of discrete objects (with the number of dots increasing from 1 to 10 in the present counting case), can be expressed in terms of
- –
- (i) Trueness—estimated as the analytical difference between the measured and true dot count, ;
- –
- (ii) Precision—estimated in terms of analytical dispersion, such as dots (Section 3.3).
- Clinical accuracy.In the ‘clinical’ scenario 2, a definition of ‘error’ in classification—such as that needed when estimating accuracy (Section 2.2)—is the difference ‘distance’ between the ‘response categorisation’ and ‘input (true) categorisation’, as given in ([2] Equation (2.8)). The challenge with clinical data is that the distances between different categories of classification on the abscissa of a PMF (such as when measuring the difference before or after an intervention or when estimating dispersion measures) may not be fully known or even meaningful (such as on nominal scales (Section 4.1)). Such effects can arise when measurement quality is limited, as dealt with in Measurement System Analysis, described further in (Section 3.1). Appropriate methods for dealing with such scales include log-odds ratio transformations, including GLMM and the Rasch psychometric theory with which measurands, such as counter ability and task difficulty, can be identified, as dealt with later in the paper in our clinical counting example (Section 5).
3. Analytical and Clinical Performance Metrics
3.1. PMF and Measurement System Analysis (MSA)
3.2. Analytical and Clinical Performance Criteria
- ‘Analytical’ performance criteria for determining, e.g., how much (quality characteristic: concentration) of a particular analyte (MSA object) is present in a sampled object (by ‘variable’), such as, first, analytical method accuracy (trueness and precision [24]) (Section 2.2) and, second, sensitivity, such as instrument limit of detection (MSA measurement instrument), as exemplified in the analytical interpretation of the elementary counting case of the present study (Section 3.3). According to [29], “…Analytical performance focuses on the gathering of evidence that the measurement instrument in question reliably, accurately and consistently measures and or detects an analyte”. This is closely related to terminology in acceptance sampling standards [31] (Section 3.1), where inspection by variables is inspection by measuring the magnitude(s) of a characteristic(s) of an item.
- ‘Clinical’ performance, according to [29], “aims to demonstrate that the measurement instrument can achieve clinically relevant outputs through predictable and reliable use by the intended users”. This is closely related to terminology in the definition in Section 3.1.3 of [32], where inspection by attribute is “inspection whereby either the item is classified simply as conforming or nonconforming with respect to a specified requirement or set of specified requirements, or the number of nonconformities in the item is counted”. Commonly used clinical performance metrics of the measurement systems include: selectivity (Equation (23)) and sensitivity (Equation (22), not to be confused with analytical instrument sensitivity, bullet 1), which are plotted against each other on receiver operating characteristic curves [33], when sampling by ‘attribute’. A psychometric treatment of these clinical performance metrics [34,35] yields quantitative estimates for quality characteristics, such as task difficulty (MSA object) and agent ability (MSA measurement instrument), as attributes of the different elements of the measurement system illustrated in Figure 1, corresponding to the top-right entry in Table 1.
3.3. PMFs for Counting and Related Tasks: ‘Analytical’ and ‘Sampling by Variable’
- (i) known exactly;
- (ii) Conceptually simple.
4. Case Studies of PMFs: Quantitative Statistical Process Control
4.1. Binomial Distribution and Dichotomous Bernoulli Trials in SPC
4.1.1. Dichotomous Decision-Making with Uncertainty
4.1.2. Distances on Categorical Scales: Counted Fractions
4.1.3. Compensating for Counted-Fraction Scale Non-Linearity
- The intercept at .
- The slope.at .
4.2. Poisson Distributions in SPC
4.3. Rasch Psychometric Model: Principle of Specific Objectivity
- ‘Agent’ ability, .
- ‘Task’ difficulty, .
4.3.1. Agnostic Rasch Models
4.3.2. Polytomous Measurement Models
4.3.3. Poisson PMFs for Counting
- In contrast to Rasch’s [48] psychometric model, it is important to note that Poisson [81] placed important restrictions on the applicability of his model: particularly that the mis-classification events were to be “very rare”, as should apply to the PMFs shown in Figure 7. Although there appears in the literature to be no exact threshold where the Poisson approximation breaks down, a ‘rule of thumb’ is as follows: is moderate (typically for good accuracy) (Figures 2–24 in [36]). Our psychometric simulations (Section 5) of task difficulty and counter ability do not suffer from the same restrictions and allow the choice of response level, , over a complete range (shown later in Figure 9): from the most mis-classifications (for a low-ability counter attempting a difficult counting task) to the least number of mis-classifications, where the Poisson approximation should be valid in the latter case (i.e., for a high-ability counter performing an easy counting task) [82]. The Poisson ‘rule of thumb’ range can be seen to be comparable with the mis-classification rates shown in Figure 7. However, to make a full comparison between the classic ‘rule of thumb’ limit to the Poisson distribution and the present psychometric simulations requires a proper account of the applicability of the ergodic principle: that is, to what extent the classic group-statistic Poisson approach [36,81] corresponds to the individual statistics of Rasch [57] psychometric modelling. For instance, the Poisson rate has in some way to be interpreted in psychometric cases where , i.e., just one counting agent (MSA: Instrument), while, at the same time, the overall number of degrees of freedom needs to be sufficiently large to achieve adequate reliability and validity (Section 5.5).
- The Poisson PMF number of mis-classifications shown in Figure 7 has no information about which numbers are perceived correctly or not, but should in principle correspond—for each item (j)—to the occupancies of the analytical PMFs shown in Figure 3. Similarly, the Poisson PMF number of mis-classifications has no obvious relation to the corresponding ‘clinical’ metrics, that is, the counting task difficulty and the counter ability. (In principle, one can estimate the expected count by inverting Equation (21) since—in the present case of an elementary construct—the relation between task difficulty and the number of dots is known (Equation (26))). Future work is intended on this novel approach to benchmarking of the Rasch Poisson Counts Model against full psychometric Rasch modelling, as mentioned in the final Discussion (Section 6).
5. ‘Clinical’ Psychometric Study: Counting Dots
5.1. Dichotomous and Polytomous Mis-Classification Probabilities. CTT and Rasch Measurement Theory
- False Acceptance (or positive, FPR) Probability;
- False Rejection (or negative, FNR) Probability.
- (i) Compensates for counted-fraction scale non-linearity (Section 4.1.3);
- (ii) Exchanges the analytical measurands (such as the number of dots, Section 3.3) for the clinical measurands: counting task difficulty and counter ability.
5.2. Counted-Fraction Scale Non-Linearity When Counting Dots: ICC

5.3. Communication of Measurement Information Throughout the Measurement Process: Amount of Entropy
5.4. Simulated PMFs for the Elementary Counting Case
5.4.1. Counting Task Difficulty Simulated: Object (Task) Entropy
5.4.2. Counting Classifier Ability Simulated
- Excel:returns a vector of random numbers having the Normal distribution of the measurand counter abilities across the cohort shown with the blue histograms in Figure 9.
5.5. Reliability and Validity
5.6. Validation of Simulated Logistic Regressions
6. Discussion and Conclusions
Author Contributions
Funding
Data Availability Statement
Acknowledgments
Conflicts of Interest
Glossary of Symbols
| Explanation | |
| discrete random variable | X |
| range | |
| event, category c | |
| probability of | |
| probabilities for different categories, c | |
| occupancy of bin for category c | |
| difference between the perceived and true dot count | |
| true value, e.g., number of dots | |
| dispersion, e.g., in number of dots | |
| measurement object (MSA) | entity, A |
| instrument (MSA) | B |
| operator (MSA) | C |
| object measurand | Z |
| response | Y |
| restitution via uncertainty model | R |
| object (A) probability of attribute z | |
| response (y; z) probability | Q = |
| entropy | H |
| Kullback–Leibler (K-L) metric | |
| calibration function relates the instrument response to the input from the measurement object | f |
| “… focuses on the gathering of evidence that the measurement instrument in question reliably, accurately and consistently measures and or detects an analyte” [29]. | analytical performance |
| “…aims to demonstrate that the measurement instrument can achieve clinically relevant outputs through predictable and reliable use by the intended users” [29]. | |
| clinical performance | |
| log-odds ratio | |
| expected number of events per interval | |
| expectation of Poisson distribution | E(X) = |
| variance of Poisson distribution | V(X) = |
| sample size | |
| ordinal or nominal response scores | |
| probability of response of instrument (agent) i; object (item) j, over k categories | |
| False acceptance (or Positive, FPR) Probability | |
| False rejection (or Negative, FNR) Probability | |
| True positive rate (TPR) a.k.a. sensitivity | |
| True negative rate (TNR) a.k.a. specificity | |
| probability of observing category c | |
| probability of successful classification per category/class | |
| instrument ability, (standard) uncertainty | , u() |
| task difficulty, (standard) uncertainty | , u() |
| explanatory variables | |
| number of, e.g., integer dots | G |
| Standard Error, i.e., standard uncertainty | |
| reliability coefficient for a Rasch variable , (for either Rasch attribute: or ) including an error term, |
Abbreviations
| CTT | Classical Test Theory |
| EMPIR | European Metrology Programme for Innovation and Research |
| EPM | European Programme for Metrology |
| FAP | False acceptance probability |
| FRP | False rejection probability |
| GLMM | Generalised Linear Measurement Model |
| GUM | Guide to the expression of uncertainty of measurement |
| ICC | Item characteristic curve |
| IRT | Item Response Theory |
| JCGM | Joint Committee for Guides in Metrology |
| K-L | Kullback–Leibler |
| MSA | Measurement System Analysis |
| PMF | Probability Mass Function |
| Probability Density Function | |
| RISE | Research Institutes of Sweden |
| RMT | Rasch Measurement Theory |
| SE | Standard Error |
| VIM | International Metrology Vocabulary |
References
- Bureau International des Poids et Mesures. The International System of Units (SI), 9th ed.; BIPM: Sèvres, France, 2019. [Google Scholar]
- Pendrill, L.R. Chapter 2. In Quality Assured Measurement—Unification across Social and Physical Sciences; Springer: Cham, Switzerland, 2019. [Google Scholar] [CrossRef] [Scilit]
- Fisher, W.P., Jr.; Pendrill, L. (Eds.) Models, Measurement, and Metrology Extending the SI: Trust and Quality Assured Knowledge Infrastructures; De Gruyter Series in Measurement Sciences; De Gruyter Oldenbourg: Berlin, Germany; Boston, MA, USA, 2024. [Google Scholar] [CrossRef] [Scilit]
- Pradella, M. Measurement Uncertainty: New Definition, Viewpoints, and Laboratories. Laboratories 2026, 3, 4. [Google Scholar] [CrossRef] [Scilit]
- Fisher, W.P., Jr.; Massengill, P.J. (Eds.) Explanatory Models, Unit Standards, and Personalized Learning in Educational Measurement: Selected Papers by A. Jackson Stenner; Springer: Singapore, 2023. [Google Scholar] [CrossRef] [Scilit]
- Fisher, W.P., Jr. Measure and manage: Intangible assets metric standards for sustainability. In Business Administration Education: Changes in Management and Leadership Strategies; Marques, J., Dhiman, S., Holt, S., Eds.; Palgrave Macmillan: New York, NY, USA, 2012; pp. 43–63. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- Barney, M.; Barney, F. Transdisciplinary Measurement through AI: Hybrid Metrology and Psychometrics Powered by Large Language Models. In Models, Measurement, and Metrology Extending the SI: Trust and Quality Assured Knowledge Infrastructures; Fisher, W.P., Jr., Pendrill, L.R., Eds.; De Gruyter Series in Measurement Sciences; De Gruyter Oldenbourg: Berlin, Germany; Boston, MA, USA, 2024; pp. 103–132. [Google Scholar] [CrossRef] [Scilit]
- Dehaene, S.; Izard, V.; Spelke, E.; Pica, P. Log or Linear? Distinct Intuitions of the Number Scale in Western and Amazonian Indigene Cultures. Science 2008, 320, 1217–1220. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- Pendrill, L.R.; Fisher, W.P., Jr. Counting and Quantification: Comparing Psychometric and Metrological Perspectives on Visual Perceptions of Number. Measurement 2015, 71, 46–55. [Google Scholar] [CrossRef] [Scilit]
- Fisher, W.P., Jr. Bateson and Wright on number and quantity: How to not separate thinking from its relational context. Symmetry 2021, 13, 1415. [Google Scholar] [CrossRef] [Scilit]
- Helliwell, T.; Giles, T.; Fensome-Rimmer, S. ISO 15189:2012—An Approach to the Assessment of Uncertainty of Measurement for Cellular Pathology Laboratories; Guidance Document G146; The Royal College of Pathologists: London, UK, 2015. [Google Scholar]
- Pennecchi, F.; Bich, W. A Bayesian Model for Measurements by Counting. Qual. Quant. 2026. [Google Scholar] [CrossRef] [Scilit]
- Mallinson, T. Extending the justice-oriented, anti-racist framework for validity testing to the application of measurement theory in re(developing) rehabilitation assessments. In Models, Measurement, and Metrology Extending the SI; Fisher, W.P., Jr., Pendrill, L., Eds.; De Gruyter: Berlin, Germany; Boston, MA, USA, 2024; pp. 401–428. [Google Scholar]
- Sul, D.; Dominguez, D.G. Culturally responsive evaluation with Latinx communities through culturally specific assessment: Building the Latinx immigration trauma assessment. New Dir. Eval. 2024, 2024, 103–112. [Google Scholar] [CrossRef] [Scilit]
- Sul, D. Situating culturally specific assessment development within the disjuncture-response dialectic. In Models, Measurement, and Metrology Extending the SI; Fisher, W.P., Jr., Pendrill, L., Eds.; De Gruyter: Berlin, Germany, 2024; pp. 475–500. [Google Scholar]
- Sul, D.; Blackmon, A.T. Enacting culturally specific assessment by constructing a STEM leadership assessment framework. J. Educ. Meas. 2026, 63, e70048. [Google Scholar] [CrossRef] [Scilit]
- JCGM. Evaluation of Measurement Data—The Role of Measurement Uncertainty in Conformity Assessment; Technical Report JCGM 106:2012; Joint Committee for Guides in Metrology: Sèvres, France, 2012. [Google Scholar]
- Mitcham, C. (Ed.) Encyclopedia of Science, Technology, and Ethics; Macmillan Reference: New York, NY, USA, 2005. [Google Scholar]
- Andersen, E.B. Sufficient statistics and latent trait models. Psychometrika 1977, 42, 69–81. [Google Scholar] [CrossRef] [Scilit]
- Andrich, D. Sufficiency and conditional estimation of person parameters in the polytomous Rasch model. Psychometrika 2010, 75, 292–308. [Google Scholar] [CrossRef] [Scilit]
- Fischer, G.H. On the existence and uniqueness of maximum-likelihood estimates in the Rasch model. Psychometrika 1981, 46, 59–77. [Google Scholar] [CrossRef] [Scilit]
- Andrich, D. Distinctions between assumptions and requirements in measurement in the social sciences. In Mathematical and Theoretical Systems: Proceedings of the 24th International Congress of Psychology of the International Union of Psychological Science, Vol. 4; Keats, J.A., Taft, R., Heath, R.A., Lovibond, S.H., Eds.; Elsevier Science Publishers: Amsterdam, The Netherlands, 1989; pp. 7–16. [Google Scholar]
- Bhatt, U.; Antorán, J.; Zhang, Y.; Liao, Q.V.; Sattigeri, P.; Fogliato, R.; Melançon, G.G.; Krishnan, R.; Stanley, J.; Tickoo, O.; et al. Uncertainty as a Form of Transparency: Measuring, Communicating, and Using Uncertainty. In AIES ’21: Proceedings of the 2021 AAAI/ACM Conference on AI, Ethics, and Society (AIES ’21); ACM: New York, NY, USA, 2021; pp. 401–413. [Google Scholar] [CrossRef] [Scilit]
- ISO 5725-1:2023(en); Accuracy (trueness and precision) of measurement methods and results—Part 1: General principles and definitions. International Standard. ISO: Geneva, Switzerland, 2023.
- Brown, R.J.C.; Güttler, B.; Neyezhmakov, P.; Stock, M.; Wielgosz, R.I.; Kück, S.; Vasilatou, K. Report of the CCU/CCQM Workshop on ‘The Metrology of Quantities Which Can Be Counted’. Metrology 2023, 3, 309–324. [Google Scholar] [CrossRef] [Scilit]
- JCGM GUM. Joint Committee for Guides in Metrology – Part 6: Developing and Using Measurement Models; Number JCGM GUM-6:2020; JCGM: Sèvres, France, 2020. [Google Scholar]
- Rossi, G.B. Measurement and Probability: A Probabilistic Theory of Measurement with Applications; Springer: Dordrecht, The Netherlands, 2014. [Google Scholar] [CrossRef] [Scilit]
- Hofmann, B. “My Biomarkers Are Fine, Thank You”: On the Biomarkerization of Modern Medicine. J. Gen. Intern. Med. 2025, 40, 453–457. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- European Commission. Factsheet for Manufacturers of In Vitro Diagnostic Medical Devices; European Commission: Luxembourg, 2020; Available online: https://health.ec.europa.eu/publications/factsheet-manufacturers-vitro-diagnostic-medical-devices_en (accessed on 13 February 2026).
- European Union. Regulation (EU) 2017/746 on In Vitro Diagnostic Medical Devices; European Union: Brussels, Belgium, 2017. [Google Scholar]
- ISO 3951-1:2022; Sampling Procedures for Inspection by Variables. International Standard. ISO: Geneva, Switzerland, 2022.
- ISO 2859-1:2026; Sampling Schemes Indexed by Acceptance Quality Limit (AQL) for Lot-by-Lot Inspection. International Standard. ISO: Geneva, Switzerland, 2026.
- Birdsall, T.G. The Theory of Signal Detectability: ROC Curves and Their Character; University of Michigan Library: Ann Arbor, MI, USA, 1973. [Google Scholar]
- Linacre, J.M. Bernoulli Trials, Fisher Information, Shannon Information and Rasch. Rasch Meas. Trans. 2006, 20, 1062–1063. [Google Scholar]
- Pendrill, L.R.; Melin, J.; Stavelin, A.; Nordin, G. Modernising Receiver Operating Characteristic (ROC) Curves. Algorithms 2023, 16, 253. [Google Scholar] [CrossRef] [Scilit]
- Montgomery, D.C. Introduction to Statistical Quality Control, 3rd ed.; Wiley: New York, NY, USA, 1996. [Google Scholar]
- NIST. Uncertainty Machine. Available online: https://uncertainty.nist.gov/ (accessed on 13 February 2026).
- Bernoulli, J. Ars Conjectandi; Opus posthumum; OCLC 7073795; Thurneysen Brothers: Basel, Switzerland, 1713. [Google Scholar]
- ISO/IEC 17025:2017; General Requirements for the Competence of Testing and Calibration Laboratories. International Organization for Standardization: Geneva, Switzerland, 2017.
- ISO 15189:2022; Medical Laboratories—Requirements for Quality and Competence. International Organization for Standardization: Geneva, Switzerland, 2022.
- Mari, L.; Brown, R.J.C.; Narduzzi, C.; Nordin, G.; Trapmann, S.; Varela Magalhães, D. What Is Measurement Uncertainty? A Discussion. Metrologia 2025, 62, 062101. [Google Scholar] [CrossRef] [Scilit]
- Joint Committee for Guides in Metrology. Evaluation of Measurement Data—Guide to the Expression of Uncertainty in Measurement; Technical Report JCGM 100:2008; JCGM: Paris, France, 2008; GUM 1995 with minor corrections. [Google Scholar] [CrossRef] [Scilit]
- Pendrill, L.R. Man as a Measurement Instrument. NCSLI Meas. 2014, 9, 24–35. [Google Scholar] [CrossRef] [Scilit]
- Tukey, J.A. Data Analysis and Behavioural Science. In The Collected Works of John A. Tukey, Volume III; Jones, L.V., Ed.; Chapman and Hall: London, UK, 1984. [Google Scholar]
- Pearson, K. Mathematical contributions to the theory of evolution: On a form of spurious correlation which may arise when indices are used in the measurements of organs. Proc. R. Soc. 1897, 60, 489–498. [Google Scholar] [CrossRef] [Scilit]
- Filzmoser, P.; Hron, K.; Reimann, C. Principal Component Analysis for Compositional Data with Outliers. Environmetrics 2009, 20, 621–632. [Google Scholar] [CrossRef] [Scilit]
- Wright, B.D. Thinking with Raw Scores. Rasch Meas. Trans. 1993, 7, 299–300. [Google Scholar]
- Rasch, G. Probabilistic Models for Some Intelligence and Attainment Tests; Danmarks Paedagogiske Institut: Copenhagen, Denmark, 1960. [Google Scholar]
- Andrich, D. Rating Scales and Rasch Measurement. Expert Rev. Pharmacoecon. Outcomes Res. 2011, 11, 571–585. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- Aitchison, J. The Statistical Analysis of Compositional Data. J. R. Stat. Soc. 1982, 44, 139–177. [Google Scholar] [CrossRef] [Scilit]
- Bland, J.M.; Altman, D.G. The Odds Ratio. Br. Med. J. 2000, 320, 1468. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- McCullagh, P. Regression Models for Ordinal Data. J. R. Stat. Soc. 1980, 42, 109–142. [Google Scholar] [CrossRef] [Scilit]
- Wright, B.D.; Stone, M.H. Best Test Design: Rasch Measurement; MESA Press: Chicago, IL, USA, 1979. [Google Scholar]
- Student, S.R.; Briggs, D.C.; Davis, L. Growth Across Grades and Common Item Grade Alignment in Vertical Scaling Using the Rasch Model. Educ. Meas. Issues Pract. 2025, 44, 84–95. [Google Scholar] [CrossRef] [Scilit]
- Montgomery, D.C.; Runger, G.C. Applied Statistics and Probability for Engineers, 5th ed.; John Wiley & Sons: Hoboken, NJ, USA, 2011. [Google Scholar]
- Meredith, W.M. The Poisson Distribution and Poisson Process in Psychometric Theory; ETS Research Bulletin Series RB-68-42; Educational Testing Service: Princeton, NJ, USA, 1968. [Google Scholar]
- Rasch, G. On General Laws and the Meaning of Measurement in Psychology. In Proceedings of the Fourth Berkeley Symposium on Mathematical Statistics and Probability, Volume 4: Contributions to Biology and Problems of Medicine; Neyman, J., Ed.; University of California Press: Berkeley, CA, USA, 1961; pp. 321–333. [Google Scholar]
- Fisher, W.P., Jr. Invariance and traceability for measures of human, social, and natural capital: Theory and application. Measurement 2009, 42, 1278–1287. [Google Scholar] [CrossRef] [Scilit]
- Bashkansky, E.; Turetsky, V. Ability Evaluation by Binary Tests: Problems, Challenges and Recent Advances. J. Phys. Conf. Ser. 2016, 772, 012012. [Google Scholar] [CrossRef] [Scilit]
- Liu, R.; Liu, H.; Shi, D.; Jiang, Z. Poisson Diagnostic Classification Models: A Framework and an Exploratory Example. Educ. Psychol. Meas. 2022, 82, 506–516. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- Rasch, G. Retirement Lecture of 9 March 1972: Objectivity in Social Sciences: A Method Problem. Rasch Meas. Trans. 2010, 24, 1252–1272, Translated by Cecilie Kreiner. [Google Scholar]
- Stone, M.; Stenner, J. From Ordinality to Quantity. Rasch Meas. Trans. 2014, 27, 1435–1437. [Google Scholar]
- Joint Committee for Guides in Metrology (JCGM). International Vocabulary of Metrology—Basic and General Concepts and Associated Terms (VIM), 3rd ed.; JCGM 200:2008 with minor corrections (2012); Joint Committee for Guides in Metrology (JCGM): Paris, France, 2008. [Google Scholar]
- Pendrill, L.; Espinoza, A.; Wadman, J.; Nilsask, F.; Wretborn, J.; Ekelund, U.; Pahlm, U. Reducing Search Times and Entropy in Hospital Emergency Departments with Real-Time Location Systems. IISE Trans. Healthc. Syst. Eng. 2021, 11, 305–315. [Google Scholar] [CrossRef] [Scilit]
- Pendrill, L.R. Category-based interlaboratory comparisons: Psychometric Rasch analyses defining reference values and statistical weighting in a clinical example. Educ. Methods Psychom. 2026, 4, 26, SAMC 2024 Special Issue. [Google Scholar] [CrossRef] [Scilit]
- Massof, R.W.; Fisher, W.P., Jr. Psychophysics and the Measurement of Sensory Magnitudes. Measurement 2026. Manuscript in review. [Google Scholar]
- Rice, S.; Pendrill, L.R.; Petersson, N.; Nordlinder, J.; Farbrot, A. Rationale and Design of a Novel Method to Assess the Usability of Body-Worn Absorbent Incontinence Care Products by Caregivers. J. Wound Ostomy Cont. Nurs. 2018, 45, 456–464. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- Adams, R.J.; Wilson, M.; Wu, M. Multilevel item response models: An approach to errors in variables regression. J. Educ. Behav. Stat. 1997, 22, 47–76. [Google Scholar] [CrossRef] [Scilit]
- Beretvas, S.N.; Kamata, A. Part II. Multi-level Measurement Rasch Models. In Rasch Measurement: Advanced and Specialized Applications; Smith, E.V., Jr., Smith, R.M., Eds.; JAM Press: Maple Grove, MN, USA, 2007; pp. 291–470. [Google Scholar]
- Briggs, D.C.; Wilson, M. An introduction to multidimensional measurement using Rasch models. J. Appl. Meas. 2003, 4, 87–100. [Google Scholar] [PubMed]
- Linacre, J.M. A User’s Guide to FACETS Rasch-Model Computer Program, Version 4.4.5. 2026. Available online: http://www.Winsteps.com (accessed on 9 April 2026).
- Linacre, J.M.; Engelhard, G.; Tatum, D.S.; Myford, C.M. Measurement with judges: Many-faceted conjoint measurement. Int. J. Educ. Res. 1994, 21, 569–577. [Google Scholar] [CrossRef] [Scilit]
- von Davier, M.; Carstensen, C.H. Multivariate and Mixture Distribution Rasch Models: Extensions and Applications; Springer: Berlin/Heidelberg, Germany, 2007. [Google Scholar] [CrossRef] [Scilit]
- Smith, R.M. Guessing and the Rasch Model. Rasch Meas. Trans. 1993, 6, 262–263. [Google Scholar]
- Masters, G.N. A Rasch model for partial credit scoring. Psychometrika 1982, 47, 149–174. [Google Scholar] [CrossRef] [Scilit]
- Masters, G.N.; Wright, B.D. The Partial Credit Model. In Handbook of Modern Item Response Theory; van der Linden, W.J., Hambleton, R.K., Eds.; Springer: New York, NY, USA, 1996; pp. 101–121. [Google Scholar]
- Andrich, D. A Rating Formulation for Ordered Response Categories. Psychometrika 1978, 43, 561–573. [Google Scholar] [CrossRef] [Scilit]
- Rasch, G. On specific objectivity: An attempt at formalizing the request for generality and validity of scientific statements. Dan. Yearb. Philos. 1977, 14, 58–94. [Google Scholar] [CrossRef] [Scilit]
- Andrich, D. Models for measurement: Precision and the non-dichotomization of graded responses. Psychometrika 1995, 60, 7–26. [Google Scholar] [CrossRef] [Scilit]
- Pendrill, L.R. Quantities and units: Order amongst complexity. In Models, Measurement, and Metrology: Extending the SI—Trust and Quality Assured Knowledge Infrastructures; Fisher, W.P., Jr., Pendrill, L.R., Eds.; De Gruyter: Berlin, Germany; Boston, MA, USA, 2024. [Google Scholar] [CrossRef] [Scilit]
- Poisson, S.D. Recherches sur la Probabilité des Jugements en Matière Criminelle et en Matière Civile; Original Work in French: Foundational in Probability Theory; Bachelier: Paris, France, 1837. [Google Scholar]
- Bookstein, A. Informetric distributions, part II: Resilience to ambiguity. J. Am. Soc. Inf. Sci. 1990, 41, 376–386. [Google Scholar] [CrossRef] [Scilit]
- Akkerhuis, T. Measurement System Analysis for Binary Tests. Ph.D. Thesis, University of Groningen, Groningen, The Netherlands, 2016. [Google Scholar]
- Akkerhuis, T.; de Mast, J.; Erdmann, T. The statistical evaluation of binary test without gold standard: Robustness of latent variable approaches. Measurement 2017, 95, 473–479. [Google Scholar] [CrossRef] [Scilit]
- Linacre, J.M. Evaluating a Screening Test. Rasch Meas. Trans. 1994, 7, 317–318. [Google Scholar]
- Cipriani, D.; Fox, C.; Khuder, S.; Boudreau, N. Comparing Rasch analyses probability estimates to sensitivity, specificity and likelihood ratios when examining the utility of medical diagnostic tests. J. Appl. Meas. 2005, 6, 180–201. [Google Scholar] [PubMed]
- Fisher, W.P., Jr.; Burton, E. Embedding measurement within existing computerized data systems: Scaling clinical laboratory and medical records heart failure data to predict ICU admission. J. Appl. Meas. 2010, 11, 271–287. [Google Scholar] [PubMed]
- Linacre, J.M. Expected Score ICC, IRF (Rasch-Half-Point Thresholds), n.d. Available online: https://www.winsteps.com/winman/expectedscoreicc.htm (accessed on 27 March 2026).
- Baker, F.B.; Kim, S.H. Item Response Theory: Parameter Estimation Techniques, 2nd ed.; CRC Press: Boca Raton, FL, USA, 2004. [Google Scholar]
- Linacre, J.M. How to Simulate Rasch Data. Rasch Meas. Trans. 2007, 21, 1125. [Google Scholar]
- Weaver, W.; Shannon, C.E. The Mathematical Theory of Communication; University of Illinois Press: Champaign, IL, USA, 1963. [Google Scholar]
- Benish, W.A. A Review of the Application of Information Theory to Clinical Diagnostic Testing. Entropy 2020, 22, 97. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- Pele, O.; Werman, M. The Quadratic-Chi Histogram Distance Family. In Proceedings of the European Conference on Computer Vision (ECCV); Springer: Berlin/Heidelberg, Germany, 2010; pp. 749–762. [Google Scholar]
- Melin, J.; Cano, S.J.; Gillman, A.; Marquis, S.; Flöel, A.; Göschel, L.; Pendrill, L.R. NeuroMET Memory Metric: Traceability and Comparability through Crosswalks. Sci. Rep. 2023, 13, 5179. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- Melin, J.; Cano, S.J.; Flöel, A.; Göschel, L.; Pendrill, L.R. The Role of Entropy in Construct Specification Equations (CSE) to Improve the Validity of Memory Tests. Entropy 2022, 24, 934. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- Stenner, A.J.; Smith, M. Testing Construct Theories. Percept. Mot. Ski. 1982, 55, 415–426. [Google Scholar] [CrossRef] [Scilit]
- Stenner, A.J.; Smith, M.I.; Burdick, D.S. Toward a Theory of Construct Definition. J. Educ. Meas. 1983, 20, 305–316. [Google Scholar] [CrossRef] [Scilit]
- Fisher, W.P., Jr.; Stenner, A.J. Theory-based metrological traceability in education: A reading measurement network. Measurement 2016, 92, 489–496. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- Fisher, W.P., Jr.; Stenner, A.J. Theory-based metrological traceability in education: A reading measurement network. In Explanatory Models, Unit Standards, and Personalized Learning in Educational Measurement: Selected Papers by A. Jackson Stenner; Fisher, W.P., Jr., Massengill, P.J., Eds.; Reprint of Fisher and Stenner (2016); Springer: Berlin/Heidelberg, Germany, 2023; pp. 275–293. [Google Scholar] [CrossRef] [Scilit]
- Klir, G.J.; Folger, T.A. Fuzzy Sets, Uncertainty, and Information; Prentice Hall: Upper Saddle River, NJ, USA, 1988. [Google Scholar]
- Linacre, J.M. Data variance: Explained, modeled, and empirical. Rasch Meas. Trans. 2003, 17, 942–943. [Google Scholar]
- Melin, J.; Melin, J.; Cano, S.J.; Flöel, A.; Göschel, L.; Pendrill, L.R. Construct Specification Equations: ‘Recipes’ for Certified Reference Materials in Cognitive Measurement. Meas. Sens. 2021, 18, 100290. [Google Scholar] [CrossRef] [Scilit]
- Shannon, C.E. A Mathematical Theory of Communication. Bell Syst. Tech. J. 1948, 27, 379–423. [Google Scholar] [CrossRef] [Scilit]
- Brillouin, L. Science and Information Theory, 2nd ed.; Academic Press: Cambridge, MA, USA, 1962. [Google Scholar] [CrossRef] [Scilit]
- De Boeck, P.; Wilson, M. (Eds.) Explanatory Item Response Models: A Generalized Linear and Nonlinear Approach; Statistics for Social and Behavioral Sciences; Springer: Berlin/Heidelberg, Germany, 2004. [Google Scholar]
- Embretson, S.E. (Ed.) Measuring Psychological Constructs: Advances in Model-Based Approaches; American Psychological Association: Washington, DC, USA, 2010. [Google Scholar]
- Ru, D.; Qiu, L.; Hu, X.; Zhang, T.; Shi, P.; Chang, S.; Jiayang, C.; Wang, C.; Sun, S.; Li, H.; et al. RAGChecker: A Fine-grained Framework for Diagnosing Retrieval-Augmented Generation. arXiv 2024. [Google Scholar] [CrossRef] [Scilit]
- Pendrill, L.R. Using Measurement Uncertainty in Decision-Making & Conformity Assessment. Metrologia 2014, 51, S206. Available online: https://iopscience.iop.org/article/10.1088/0026-1394/51/4/S206 (accessed on 10 April 2026). [CrossRef] [Scilit]
- Center for AI Standards and Innovation (CAISI). CAISI Evaluation of DeepSeek V4 Pro; National Institute of Standards and Technology (NIST): Gaithersburg, MD, USA, 2026. [Google Scholar]
- Adel, T.; Bilson, S.; Levene, M.; Thompson, A. Trustworthy Artificial Intelligence in the Context of Metrology. arXiv 2024. [Google Scholar] [CrossRef] [Scilit]
- Letizia, N.A.; Novello, N.; Tonello, A.M. Copula Density Neural Estimation. IEEE Trans. Neural Netw. Learn. Syst. 2025, 36, 19452–19459. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- Decruyenaere, A.; Dehaene, H.; Rabaey, P.; Polet, C.; Decruyenaere, J.; Demeester, T.; Vansteelandt, S. Debiasing Synthetic Data Generated by Deep Generative Models. arXiv 2024. [Google Scholar] [CrossRef] [Scilit]
- Glorot, X.; Bordes, A.; Bengio, Y. Deep Sparse Rectifier Neural Networks. In Proceedings of the Fourteenth International Conference on Artificial Intelligence and Statistics (AISTATS), Fort Lauderdale, FL, USA, 11–13 April 2011; Gordon, G., Dunson, D., Dudík, M., Eds.; Proceedings of Machine Learning Research; Journal of Machine Learning Research Inc.: Cambridge, MA, USA, 2011; Volume 15, pp. 315–323. [Google Scholar]
- Sklar, M. Fonctions de répartition à N dimensions et leurs marges. Ann. De L’ISUP 1959, 8, 229–231. [Google Scholar]











| Discrete | Continuous | |
|---|---|---|
| Qualitative | Instrument response (Section 5.4): per category/class | Clinical performance (bullet 2): Instrument ability, , u() Task difficulty, , u() |
| Quantitative | Analytical counting (Section 3.3) How many dots in object? Counting errors Limit of detection | Analytical measure (bullet 1): How much of a quantity in object? Measurement errors and uncertainties Trueness & precision |
| x | 1 | 2 | 3 | 4 | 5 | 6 | 7 | 8 | 9 | 10 |
|---|---|---|---|---|---|---|---|---|---|---|
| 33.72 | 10.77 | 2.29 | 0.37 | 0.05 | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 3.09 | 7.89 | 13.43 | 17.15 | 17.53 | 14.92 | 10.89 | 6.95 | 3.95 | 2.02 | |
| 1.83 | 5.27 | 10.10 | 14.51 | 16.68 | 15.97 | 13.12 | 9.42 | 6.02 | 3.46 |
Disclaimer/Publisher’s Note: The statements, opinions and data contained in all publications are solely those of the individual author(s) and contributor(s) and not of MDPI and/or the editor(s). MDPI and/or the editor(s) disclaim responsibility for any injury to people or property resulting from any ideas, methods, instructions or products referred to in the content. |
© 2026 by the authors. Licensee MDPI, Basel, Switzerland. This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution (CC BY) license.
Share and Cite
Pendrill, L.R.; Fisher, W.P., Jr. Uniting Psychometric Modelling and Poisson Distributions: A Metrological Study of Elementary Counting. Foundations 2026, 6, 26. https://doi.org/10.3390/foundations6030026
Pendrill LR, Fisher WP Jr. Uniting Psychometric Modelling and Poisson Distributions: A Metrological Study of Elementary Counting. Foundations. 2026; 6(3):26. https://doi.org/10.3390/foundations6030026
Chicago/Turabian StylePendrill, Leslie R., and William P. Fisher, Jr. 2026. "Uniting Psychometric Modelling and Poisson Distributions: A Metrological Study of Elementary Counting" Foundations 6, no. 3: 26. https://doi.org/10.3390/foundations6030026
APA StylePendrill, L. R., & Fisher, W. P., Jr. (2026). Uniting Psychometric Modelling and Poisson Distributions: A Metrological Study of Elementary Counting. Foundations, 6(3), 26. https://doi.org/10.3390/foundations6030026
