Joseph L Fleiss
Biographic Data
| ID | 5971026 |
|---|---|
| NAME | Joseph L Fleiss |
| GIVEN NAMES | Joseph L |
| FAMILY NAME | Fleiss |
| SIGNATURE | FLEISS J L |
| AFFILIATIONS | New York State Department of State |
| VERIFIED | No |
| TOTAL WORKS | 19 |
| TOTAL CITATIONS | 927 |
| AUTHOR COUNT | 19 |
| EDITOR COUNT | 0 |
| FIRST PUBLICATION YEAR | 1967 |
| LATEST PUBLICATION YEAR | 2003 |
| H-INDEX | 5 |
Statistical Methods for Rates and Proportions
Preface.Preface to the Second Edition.Preface to the First Edition.1. An Introduction to Applied Probability.2. Statistical Inference for a Single Proportion.3. Assessing Significance in a Fourfold Table.4. Determining Sample Sizes Needed to Detect a Difference Between Two Proportions.5. How to Randomize.6. Comparative Studies: Cross-Sectional, Naturalistic, or Multinomial Sampling.7. Comparative Studies: Prospective and Retrospective Sampling.8.…
Significance tests have a role in epidemiologic research: Reactions to A. M. Walker
Significance tests have a role in epidemiologic research: reactions to A. M. Walker. J L FleissCopyRight https://doi.org/10.2105/AJPH.76.5.559 Published Online: October 07, 2011
Dr. Fleiss Responds
Dr. Fleiss Responds Joseph L. Fleiss CopyRight https://doi.org/10.2105/AJPH.76.8.1033-a Published Online: October 07, 2011
Confidence intervals vs significance tests: Quantitative interpretation
Confidence intervals vs significance tests: quantitative interpretation. J L FleissCopyRight https://doi.org/10.2105/AJPH.76.5.587 Published Online: October 07, 2011
Balanced Incomplete Block Designs for Inter-Rater Reliability Studies
Occasionally, an inter-rater reliability study must be designed so that each subject is rated by fewer than all the participating raters. If there is interest in comparing the raters' mean levels of rating, and if it is desired that each mean be estimated with the same precision, then a balanced incomplete block design for the reliability study is indicated. Methods for executing the design and for analyzing the resulting data are presented, usin…
Large sample variance of kappa in the case of different sets of raters
Large sample variance of kappa in the case of different sets of raters
Published formulas for the large sample variance of the kappa statistic that are appropriate for the case of different sets of raters for different Ss, when each set of raters is selected at random from a larger pool of available raters, are determined to be incorrect. New formulas are derived and c
Intraclass correlations: Uses in assessing rater reliability
Reliability coefficients often take the form of intraclass correlation coefficients. In this article, guidelines are given for choosing among six different forms of the intraclass correlation for reliability studies in which n target are rated by k judges. Relevant to the choice of the coefficient are the appropriate statistical model for the reliability and the application to be made of the reliability results. Confidence intervals for each of t…
The effects of measurement errors on some multivariate procedures
The effects of random errors of measurement on a variety of basic multivariate procedures are discussed. For some procedures, the effect is one of attenuation (i.e., simply reducing the estimated strength of association). For others, the effect may be to alter the apparent direction of association. Some proposed solutions to the problems caused by errors of measurement are presented, and areas requiring further research are pointed out
Pathways to the hospital for the geriatric psychiatric patient in New York and London
This communication examines the pathways of geriatric psychiatric patients in New York and London from the time of onset of a psychiatric episode to hospitalization. Informants of 50 patients in each city were interviewed with a semi-structured interview covering the events and the patient's activities prior to hospitalization. The results show that the time from the onset of the episode to hospitalization is significantly shorter in London than …
The Equivalence of Weighted Kappa and the Intraclass Correlation Coefficient as Measures of Reliability
Cross-National Study of Diagnosis of the Mental Disorders: Some Demographic Correlates of Hospital Diagnosis in New York and London
The Study of the Psychiatric Symptoms of Systemic Lupus Erythematosus: A Critical Review
This paper critically reviews studies of psychiatric symptoms in systemic lupus erythematosus. Reported findings vary by chronologic order of publication, clinical orientation of the author, mode of selecting cases and method of examination. This variation in findings reflects errors of method, the most common of which are biased selection of subjects, lack of comparison groups, unreliable recording of observations on the patient and inconsistent…
Measuring nominal scale agreement among many raters
On the asserted invariance of the odds ratio
Group 1 Px 1-Px Group 2 P2 l-P2 The measure of association for the four-fold table most frequently employed by epidemiologists is the odds ratio (also termed the cross-product ratio and the approximate relative risk), to = Px (l-P2)IP2 (1-Pi). The odds ratio has been studied by Cornfield (1956) and by Mantel and Haenszel (1959). Edwards (1966) claims a certain invariance property for the odds ratio. He asserts that if a characteristic is normally…
On the Methods and Theory of Clustering
The need for methods of clustering individuals into homogeneous groups seems clear. One hopes, by applying them to his data, to discover clusterings which may prove to be important. This aim appears straightforward, but the methods which exist do not necessarily satisfy them. The procedures which employ the correlation measure of profile similarity, and those which employ the distance measure are discussed. Technical and logical problems are show…
Estimating the magnitude of experimental effects
Large sample standard errors of kappa and weighted kappa
The statistics kappa (Cohen, 1960) and weighted kappa (Cohen, 1968) were introduced to provide coefficients of agreement between two raters for nominal scales. Kappa is appropriate when all disagreements may be considered equally serious, and weighted kappa is appropriate when the relative seriousness of the different possible disagreements can be specified. The papers describing these two statistics also present expressions for their standard er…
A Fortran IV Program for the Analysis of Demographic, Item, and Scale Data
Intraclass correlations: Uses in assessing rater reliability
Reliability coefficients often take the form of intraclass correlation coefficients. In this article, guidelines are given for choosing among six different forms of the intraclass correlation for reliability studies in which n target are rated by k judges. Relevant to the choice of the coefficient are the appropriate statistical model for the reliability and the application to be made of the reliability results. Confidence intervals for each of t…
Measuring nominal scale agreement among many raters
Significance tests have a role in epidemiologic research: Reactions to A. M. Walker
Significance tests have a role in epidemiologic research: reactions to A. M. Walker. J L FleissCopyRight https://doi.org/10.2105/AJPH.76.5.559 Published Online: October 07, 2011
Large sample standard errors of kappa and weighted kappa
The statistics kappa (Cohen, 1960) and weighted kappa (Cohen, 1968) were introduced to provide coefficients of agreement between two raters for nominal scales. Kappa is appropriate when all disagreements may be considered equally serious, and weighted kappa is appropriate when the relative seriousness of the different possible disagreements can be specified. The papers describing these two statistics also present expressions for their standard er…
Confidence intervals vs significance tests: Quantitative interpretation
Confidence intervals vs significance tests: quantitative interpretation. J L FleissCopyRight https://doi.org/10.2105/AJPH.76.5.587 Published Online: October 07, 2011
The effects of measurement errors on some multivariate procedures
The effects of random errors of measurement on a variety of basic multivariate procedures are discussed. For some procedures, the effect is one of attenuation (i.e., simply reducing the estimated strength of association). For others, the effect may be to alter the apparent direction of association. Some proposed solutions to the problems caused by errors of measurement are presented, and areas requiring further research are pointed out
Large sample variance of kappa in the case of different sets of raters
Published formulas for the large sample variance of the kappa statistic that are appropriate for the case of different sets of raters for different Ss, when each set of raters is selected at random from a larger pool of available raters, are determined to be incorrect. New formulas are derived and c
Dr. Fleiss Responds
Dr. Fleiss Responds Joseph L. Fleiss CopyRight https://doi.org/10.2105/AJPH.76.8.1033-a Published Online: October 07, 2011
A Fortran IV Program for the Analysis of Demographic, Item, and Scale Data
On the Methods and Theory of Clustering
The need for methods of clustering individuals into homogeneous groups seems clear. One hopes, by applying them to his data, to discover clusterings which may prove to be important. This aim appears straightforward, but the methods which exist do not necessarily satisfy them. The procedures which employ the correlation measure of profile similarity, and those which employ the distance measure are discussed. Technical and logical problems are show…
Estimating the magnitude of experimental effects
Large sample standard errors of kappa and weighted kappa
The statistics kappa (Cohen, 1960) and weighted kappa (Cohen, 1968) were introduced to provide coefficients of agreement between two raters for nominal scales. Kappa is appropriate when all disagreements may be considered equally serious, and weighted kappa is appropriate when the relative seriousness of the different possible disagreements can be specified. The papers describing these two statistics also present expressions for their standard er…
On the asserted invariance of the odds ratio
Group 1 Px 1-Px Group 2 P2 l-P2 The measure of association for the four-fold table most frequently employed by epidemiologists is the odds ratio (also termed the cross-product ratio and the approximate relative risk), to = Px (l-P2)IP2 (1-Pi). The odds ratio has been studied by Cornfield (1956) and by Mantel and Haenszel (1959). Edwards (1966) claims a certain invariance property for the odds ratio. He asserts that if a characteristic is normally…
Measuring nominal scale agreement among many raters
The Study of the Psychiatric Symptoms of Systemic Lupus Erythematosus: A Critical Review
This paper critically reviews studies of psychiatric symptoms in systemic lupus erythematosus. Reported findings vary by chronologic order of publication, clinical orientation of the author, mode of selecting cases and method of examination. This variation in findings reflects errors of method, the most common of which are biased selection of subjects, lack of comparison groups, unreliable recording of observations on the patient and inconsistent…
The Equivalence of Weighted Kappa and the Intraclass Correlation Coefficient as Measures of Reliability
Cross-National Study of Diagnosis of the Mental Disorders: Some Demographic Correlates of Hospital Diagnosis in New York and London
Pathways to the hospital for the geriatric psychiatric patient in New York and London
This communication examines the pathways of geriatric psychiatric patients in New York and London from the time of onset of a psychiatric episode to hospitalization. Informants of 50 patients in each city were interviewed with a semi-structured interview covering the events and the patient's activities prior to hospitalization. The results show that the time from the onset of the episode to hospitalization is significantly shorter in London than …
The effects of measurement errors on some multivariate procedures
The effects of random errors of measurement on a variety of basic multivariate procedures are discussed. For some procedures, the effect is one of attenuation (i.e., simply reducing the estimated strength of association). For others, the effect may be to alter the apparent direction of association. Some proposed solutions to the problems caused by errors of measurement are presented, and areas requiring further research are pointed out
Large sample variance of kappa in the case of different sets of raters
Large sample variance of kappa in the case of different sets of raters
Published formulas for the large sample variance of the kappa statistic that are appropriate for the case of different sets of raters for different Ss, when each set of raters is selected at random from a larger pool of available raters, are determined to be incorrect. New formulas are derived and c
Intraclass correlations: Uses in assessing rater reliability
Reliability coefficients often take the form of intraclass correlation coefficients. In this article, guidelines are given for choosing among six different forms of the intraclass correlation for reliability studies in which n target are rated by k judges. Relevant to the choice of the coefficient are the appropriate statistical model for the reliability and the application to be made of the reliability results. Confidence intervals for each of t…
Balanced Incomplete Block Designs for Inter-Rater Reliability Studies
Occasionally, an inter-rater reliability study must be designed so that each subject is rated by fewer than all the participating raters. If there is interest in comparing the raters' mean levels of rating, and if it is desired that each mean be estimated with the same precision, then a balanced incomplete block design for the reliability study is indicated. Methods for executing the design and for analyzing the resulting data are presented, usin…
Significance tests have a role in epidemiologic research: Reactions to A. M. Walker
Significance tests have a role in epidemiologic research: reactions to A. M. Walker. J L FleissCopyRight https://doi.org/10.2105/AJPH.76.5.559 Published Online: October 07, 2011
Dr. Fleiss Responds
Dr. Fleiss Responds Joseph L. Fleiss CopyRight https://doi.org/10.2105/AJPH.76.8.1033-a Published Online: October 07, 2011
Confidence intervals vs significance tests: Quantitative interpretation
Confidence intervals vs significance tests: quantitative interpretation. J L FleissCopyRight https://doi.org/10.2105/AJPH.76.5.587 Published Online: October 07, 2011
Statistical Methods for Rates and Proportions
Preface.Preface to the Second Edition.Preface to the First Edition.1. An Introduction to Applied Probability.2. Statistical Inference for a Single Proportion.3. Assessing Significance in a Fourfold Table.4. Determining Sample Sizes Needed to Detect a Difference Between Two Proportions.5. How to Randomize.6. Comparative Studies: Cross-Sectional, Naturalistic, or Multinomial Sampling.7. Comparative Studies: Prospective and Retrospective Sampling.8.…
Mathematics (14 works) · Statistics (13 works) · Psychology (12 works) · Medicine (7 works) · Reliability and Agreement in Measurement (7 works) · Computer Science (5 works) · Econometrics (5 works) · Advanced Statistical Methods and Models (4 works) · Hemodynamic Monitoring and Therapy (4 works) · Kappa (4 works)