I'm spinning this off of the Process Map thread, just for clarity. It also goes along with the other thread of siloed forensic research because there's a myriad of other disciplines that have already wrestled with this.
We'll go from this:
I respectfully disagree. The discipline tends to act first, think second. Just like the whole likelihood ratio....'we have to do something' mentality, followed by mea culpas and journal publications that directly contradict the assertions that were being made at the time.ER wrote: ↑Thu Jan 16, 2020 2:37 pm To those that are argue against the OSAC 5 Conclusion scale because...
- There are no instructions on HOW to use it;
- There has been no validation, testing, or studies using it;
- The conclusions don't protect against errors; or
- Any of the other myriad reasons:
…
Preventing progress on these very basic first steps because subsequent steps haven't been completed is insanity. It's asking the baby to run before she crawls.
What we have here is a time tested debate between Dichotomous Scales and Likert Scales. Dichotomous simply meaning yes/no, so for us it'd be ID/EXC. Likert scales are just psychometric scales named after the guy who invented them in 1932 and are usually a 5 scale system that attempt to measure attitudes around a certain phenomenon. For us, that is the 5 conclusion scale that is being offered up by the OSAC. Those are: Source ID; Support for Same Source; Inconclusive; Support for Different Source; Source Exclusion. (I Would argue the adoption of Inconclusive to the ID/EXC framework puts us in a likert scenario)
Each has it's disadvantages. Dichotomous approaches can force a choice, bias to the extremes and distort the fact that there are true neutral responses. Likert scales create more variability which in turn diminishes reliability of the scale and falsely imply precision. Even numbered Likert scales (one without a neutral option) can also force a choice (Take Busey's/Vanderkolk's workshop at an IAI some time, you'll see.)
You could argue that the critique that we are overstating our conclusions is just the phenomenon of forcing choices which bias to the extreme.
This debate centers around what we've been talking about, validity and reliability. Here's a great discussion in plain english that shows that concern around validity/reliability of likert scales is real, long standing (since the 1930s),exists in domains outside of fingerprints, and is directly applicable.
And this brings us into the instruction aspect of what ER wrote above because there cannot be instructions on how to use it. Likert scales are, by definition, non-parameterized. Meaning it is ranked, not numerical (read: subject to statistical methods) and they are not derived from an equal-interval scale. The difference between Inconclusive and what counts as Support for Identification and Support for Exclusion are not anywhere near the same distance if we consider the scale as a measurement. This is what gives rise to the variability. Just keep reading and you'll see.
So, for me, the thing that really counts is performance testing. That's what should be done up front. Whitebox, blackbox, greybox...you name it. Adopting these prior to showing them to have value will be problematic. So, let's do a micro performance test.
All of that being said, I bring you the prints. Please rate on the poll above which OSAC conclusion you would use.