| MEASUREMENT PROPERTIES | |
| Limitations | >=5 European languages (English) |
| Observations | |
| 1. RELIABILITY | |
| A. Internal Consistency | Tested |
| Cronbach's (Describe) |
The exploratory factor analysis (EFA) of the 'now' function data and 'now' and 'at diagnosis' emotional and psychological well-being data showed that these scales were unidimensional and had high internal consistency (Cronbach's alpha >0.9). However, EFA of the 'at diagnosis' function data resulted in two factors, and there was no clinically meaningful distinction between the groups of items loading onto each factor. Therefore, Rasch analysis was employed to aid item reduction and assess unidimensionality more rigorously. |
| B. Reliability intraobserver or test-retest | Tested |
| Continuous scores: intraclass correlation coefficient (ICC) Dichotomus: Cohen kappa (Describe) |
The validation study included a sample size of more than 50 for each domain, and the intraclass correlation coefficient (ICCagreement) was greater than 0.8 in each domain, indicating good reliability. The standard error of measurement (SEMagreement) for each domain ranged from 9.3 to 11.9 on a scale out of 100. |
| C. Reliability interobserver or Measurement error | Not Tested |
| Standard error of measurement (SEM), smallest detectable change (SDC) or Limits of agreement (LoA) (Describe) |
|
| 2. VALIDITY | |
| A. Content validity: face validity | Tested |
| Expert opinion (relevance and comprehensiveness) (Describe) |
The candidate items for the PMR-IS were derived from a conceptual framework and a list of symptoms and effects of PMR identified through literature synthesis and qualitative research. The PMR questionnaire was completed by 28 participants with PMR, along with the QQ-10 questionnaire used to assess the face validity, feasibility, and utility of patient healthcare questionnaires. The results showed a high mean value score of 79% (SD 12) and a low burden score of 21% (SD 18), indicating good face validity and acceptability to patients. |
| B. Construct Validity: Structural validity |
Tested |
| Hypotheses-testing | Tested |
| Cross-cultural validity | Not Tested |
| Brief Description |
The Rasch analysis used a partial credit model and iteratively deleted the least well-fitting items until satisfactory fit statistics were achieved for unidimensional scales. At the end of the process, a 9-item functional scale and a 4-item psychological and emotional well-being scale were created. The only item showing differential item functioning (DIF) was 'take your shoes or socks on or off', which showed DIF for gender in the 'now' dataset. |
| C. Criterion validity | Not Tested |
| Comparison with a 'gold standard' Continuous scores: correlations, ROC curves Dichotomus: Sensitivity & Specificity (Describe) |
|