Skip to content

Expert Evidence and Statistical Interpretation in Court

How wildlife forensic scientists present species identification evidence in court, meet admissibility standards, and communicate likelihood ratios and Bayesian statistics to judges and juries.

Last updated:

Share

Wildlife forensic scientists must satisfy two distinct standards to be effective in court: the admissibility threshold set by the jurisdiction (Daubert in US federal courts, Criminal Procedure Rules Part 19 in England and Wales, or equivalent provisions elsewhere), and the communication standard required to convey statistical conclusions accurately to non-specialist judges and juries. The core statistical tool is the likelihood ratio, which compares the probability of the forensic evidence under the prosecution hypothesis against the probability under the defence hypothesis. Reference database size is the most predictable vulnerability in wildlife DNA evidence: smaller databases produce wider uncertainty bounds, and any expert who does not acknowledge this openly will face more damaging cross-examination than one who does.

A wildlife forensic scientist who identifies a piece of ivory as African elephant, or matches a feather to a protected raptor, has completed the laboratory work. The harder test comes in the witness box. There, every methodological choice made in the laboratory becomes a target: the reference database used, the software that called the match, the statistical framework that converted a probability into a statement about guilt. Courts in different countries apply different admissibility thresholds, and a finding that convinces a biologist can fall apart under a skilled cross-examination if the scientist cannot explain the numbers clearly.

Two admissibility frameworks dominate the case law that wildlife prosecutors encounter. In the United States, federal courts apply the Daubert standard, which places the judge in the role of gatekeeper. In many other common-law jurisdictions, including the United Kingdom and Australia, courts historically used the older Frye general-acceptance test, though England and Wales now use the Criminal Procedure Rules, which achieve much the same result. Both frameworks push forensic scientists toward the same behaviour: document the method, know its error rate, cite the peer-reviewed literature, and never overstate what the data says.

Beyond admissibility, wildlife cases bring a specific statistical challenge. The reference databases that underpin DNA barcoding and population genetics are far smaller and patchier than their human-forensic equivalents. A random-match probability calculated from a database of 200 specimens cannot be quoted with the same confidence as one from a database of millions. This topic covers the frameworks, the landmark cases, and the practical tools, including likelihood ratios, Bayesian barcoding inference, and minimum-number-of-individuals calculations, that let wildlife scientists give evidence that is both honest and useful to the court.

By the end of this topic you will be able to:

  • Explain the Daubert standard's four criteria and describe how each applies to wildlife DNA barcoding and morphological identification methods.
  • Calculate the minimum number of individuals (MNI) in a mixed-element seizure and articulate why MNI is a floor rather than an estimate.
  • Construct a likelihood ratio statement for a wildlife DNA match and identify the transposition fallacy (prosecutor's fallacy) in a given expert opinion.
  • Apply SWFS minimum reporting standards to a draft forensic report, identifying where conclusions are overstated or database limitations are not disclosed.
  • Describe how Bayesian barcoding inference uses trade-context prior probabilities to produce a posterior probability of species identity, and explain when a court may find this framing difficult to follow.
Key terms
Daubert standard
A US Federal Rules of Evidence standard (from Daubert v. Merrell Dow Pharmaceuticals, 1993) requiring the trial judge to assess whether scientific evidence is based on sufficient facts, is the product of reliable principles and methods, and has been applied reliably to the case facts before it is presented to a jury.
Frye test
An older admissibility standard requiring that a scientific technique be generally accepted by the relevant scientific community. Still used in some US state courts; functionally similar tests operate in several other common-law countries.
Likelihood ratio (LR)
The ratio of the probability of the evidence under the prosecution hypothesis to the probability under the defence hypothesis. An LR greater than 1 supports the prosecution; the size of the ratio tells the court by how much.
Bayesian barcoding inference
A method for assigning a DNA barcode sequence to a species that uses prior probabilities (the known composition of species in a trade network) and a likelihood function (how well the sequence fits reference data for each candidate species) to produce a posterior probability of species identity.
Minimum number of individuals (MNI)
The smallest number of individual animals logically required to account for all biological material in a seizure. Calculated from the most frequent repeated element, such as the left mandible or a unique genotype.
SWFS reporting standards
Minimum standards set by the Society for Wildlife Forensic Science specifying what a forensic report must state, including the question, the methods, the known limitations, and conclusions expressed as degrees of support rather than absolute statements.

Admissibility frameworks: Daubert, Frye, and beyond

The Daubert decision (1993) changed the relationship between science and US courts. Before Daubert, the Frye test required only that a technique be generally accepted in the relevant scientific community. Daubert replaced general acceptance with a structured, multi-factor inquiry: Has the theory been tested? Has it been subjected to peer review? Is there a known or potential error rate? Is it generally accepted? A trial judge must run through these questions as a gatekeeper before scientific testimony reaches a jury.

For wildlife forensics, these questions carry real bite. DNA barcoding using COI sequences satisfies all Daubert factors in well-established taxa: the method is published, tested, has a known success rate by taxonomic group, and is accepted by the systematic biology community. But barcoding of recently described species, or species where the reference database is thin, sits in a grayer zone. A knowledgeable defence attorney can challenge the database coverage and force the scientist to concede that the match probability is conditional on an incomplete reference.

Outside the United States, the rules vary. England and Wales now use Criminal Procedure Rules Part 19, which require experts to help the court, state the limits of their expertise, and indicate where their opinion is disputed. Australia follows similar principles under the Evidence Act uniform provisions. The practical effect across jurisdictions is the same: an expert who overstates certainty or conceals methodological limitations faces serious credibility damage in cross-examination, and the wildlife forensic community's own reporting standards, through SWFS, align with these legal expectations.

Likelihood ratios and the Bayesian framework for wildlife DNA

The logical framework that underpins modern forensic DNA interpretation is Bayesian. Two hypotheses compete: under the prosecution's hypothesis (Hp), the evidence DNA comes from the species, individual, or population in question; under the defence hypothesis (Hd), it comes from some other source. The likelihood ratio is the ratio of P(evidence | Hp) to P(evidence | Hd). When that ratio is high, the evidence strongly favours the prosecution account; when it is near 1, the evidence is neutral.

Likelihood ratio logic: the probability of the evidence under the prosecution hypothesis divided by the probability under the
Likelihood ratio logic: the probability of the evidence under the prosecution hypothesis divided by the probability under the defence hypothesis, producing a value that tells the court which way and how strongly the data points.

In human forensics, the denominator is typically the random match probability from a large population database. Wildlife forensics faces a harder problem. For many traded species, population genetic reference data is patchy. The SWFS standards address this: the expert must state the reference database size and composition, and if the database is incomplete, the opinion should be bounded accordingly. A match to African elephant using a well-sampled database of hundreds of loci and thousands of individuals is far more defensible than a match using twelve microsatellites from fifty samples.

LR valueVerbal scale (ILAC)Practical meaning in court
1–10Weak supportEvidence barely favours the hypothesis; often not worth stressing
10–100Moderate supportWorth noting but not compelling on its own
100–10,000Moderately strong supportUseful contribution when combined with other evidence
10,000–1,000,000Strong supportSignificant; challenges defence account without other explanation
>1,000,000Very strong supportNear-conclusive for species-level identity from well-sampled taxa

Bayesian barcoding takes the LR concept and combines it with a prior probability drawn from what is known about the composition of trade in a given market or seizure context. If 80% of ivory confiscated on a particular trade route is from African forest elephant rather than savanna elephant, that prior changes the posterior probability of species assignment from a barcoding match. Courts that understand Bayesian reasoning accept this framing; courts that do not can find it confusing, which places an obligation on the expert to explain it in terms the judge and jury can follow.

Minimum number of individuals in seizure cases

Wildlife trafficking cases often involve fragmentary material: hundreds of bear paws, thousands of turtle shells, dozens of dried seahorses. The charge almost always depends on a count of individual animals, not just a weight of material. But the same animal contributes multiple parts, and duplicate counting inflates the number. The minimum number of individuals calculation provides the floor.

  1. Inventory and sort
    Separate all biological material into element types: left forelimb, right forelimb, skull, pelvis. A single animal contributes one of each paired element on each side.
  2. Identify the most frequent single-side element
    Count the most numerous element from one side (left or right). Forty-five left front bear paws means at least 45 individual bears, regardless of how many right paws, spines, or skulls are present.
  3. Apply DNA typing to refine
    Where budget permits, genotyping distinguishes individuals even from the same element type. Two samples of the same element with different genotypes confirm two individuals. This raises MNI above the morphological floor.
  4. Report as a floor, not an estimate
    MNI is conservative. The actual number of animals in the shipment may be higher. Courts should be told explicitly that MNI understates, not overestimates, the scale of the crime.

Landmark cases: R v Bailey and US v Kapp

Two cases in the English-language wildlife prosecution literature illustrate how courts have tested forensic expert evidence in practice.

R v Bailey (UK) involved raptor persecution, specifically peregrine falcons and their eggs. The prosecution relied on forensic identification of eggshell fragments recovered from a gamekeeper's property, chemical analysis of residues, and feather morphology. The case was significant because it required the court to assess whether ornithological and ecotoxicological expertise met the threshold for admissibility under English rules. The defence challenged the chain of custody and the uniqueness of the shell fragments. Expert witnesses for the prosecution had to demonstrate that their identification methods, drawn from published comparative collections, were reliable and not merely opinion. The case established a precedent for treating specialist natural-history identification as properly scientific expert evidence rather than mere conjecture.

United States v. Kapp involved the trafficking of tigers and leopards and their meat, hides, and parts in violation of the Endangered Species Act and the Lacey Act. The defence contested aspects of the forensic identification evidence. The court accepted the testimony, relying on the laboratory's accreditation and the analyst's documented training. The case illustrates that accreditation of the laboratory and the analyst's qualifications carry independent weight alongside the methodological factors that Daubert requires.

Key factors assessed when wildlife forensic evidence is challenged in court, showing the four Daubert criteria alongside two
Key factors assessed when wildlife forensic evidence is challenged in court, showing the four Daubert criteria alongside two additional practical factors that frequently arise in wildlife cases.

SWFS reporting standards in practice

The Society for Wildlife Forensic Science developed its minimum reporting standards precisely because wildlife forensic scientists work across many laboratory and agency contexts without the centralised accreditation structure that governs human forensic science in most countries. The standards require that a report answer the question the investigator actually asked, describe the evidence examined, state the methods clearly enough for a competent scientist to evaluate them, acknowledge limitations, and express the conclusion as a degree of support for the hypothesis rather than a categorical declaration.

  • State the question: what was the forensic question? Species identity? Geographic origin? Individuality? The answer to a narrower question than what the prosecution needs can be a problem if the gap is not stated.
  • Describe the evidence: item number, condition received, any continuity concerns, and the type of biological material examined.
  • State the method and its limitations: which markers, which reference database, what the database coverage is for the species group, and the method's published error rate in comparable taxa.
  • Express the conclusion proportionately: use a verbal scale tied to the likelihood ratio, or state the posterior probability explicitly. Avoid 'the sample is definitely from species X' when what the data supports is 'the evidence is very strongly consistent with species X given current reference data.'
  • Note any competing explanations: if another species could produce the same barcode due to reference database gaps, say so. Courts expect the expert to address the defence hypothesis, not only the prosecution's.

Communicating statistics to non-specialist courts

Wildlife cases are often heard by judges and juries with no scientific background. The forensic scientist's job is to translate statistical conclusions into language that is accurate and comprehensible without being misleading. Three persistent problems arise in practice.

The first is the transposition fallacy, also called the prosecutor's fallacy. The probability of the evidence given innocence is not the same as the probability of innocence given the evidence. A scientist who says 'there is only a one in ten thousand chance this is not from the endangered species' has stated the LR in reverse and may be giving the court a misleading impression about the probability of guilt. The correct statement is 'the evidence is ten thousand times more probable if it came from the species than if it came from a non-target source.'

The second is overconfidence from a matching result when the reference database is small. A unique genotype match to a reference population of thirty animals is statistically interesting but far from conclusive. The third is treating class-level evidence as individual-level: a species-level identification does not prove a specific individual animal was taken from a specific location, and conflating the two inflates the strength of the evidence.

Check your understanding
Question 1 of 4· 0 answered

Under the Daubert standard, which of the following is NOT one of the factors a judge must consider?

Key Takeaways

  • Daubert (US) and equivalent rules elsewhere require wildlife forensic methods to be testable, peer-reviewed, have a known error rate, and be generally accepted before a judge allows them before a jury.
  • The likelihood ratio framework evaluates evidence under competing hypotheses and avoids the prosecutor's fallacy of inverting the probability.
  • Bayesian barcoding inference combines a prior probability from trade-context knowledge with the DNA match likelihood to produce a posterior probability of species identity.
  • Minimum number of individuals is a conservative floor calculated from the most frequent single-side element or unique genotypes, and should always be reported as an undercount rather than an estimate.
  • SWFS reporting standards require proportionate conclusions with database limitations explicitly acknowledged, because incomplete reference databases are a specific and predictable weakness in wildlife forensic testimony.
What is the Daubert standard and how does it apply to wildlife forensics?
Daubert is the US federal rule requiring a judge to assess scientific evidence before it reaches a jury. The judge asks whether the method has been tested, whether it has a known error rate, whether it is peer-reviewed, and whether it is generally accepted in the relevant scientific community. Wildlife DNA barcoding, morphological identification, and trace-evidence comparisons must all satisfy these criteria when challenged in US federal court.
What is a likelihood ratio in the context of wildlife DNA evidence?
A likelihood ratio compares two hypotheses: the probability of the DNA profile if it came from the species in question against the probability if it came from another species or individual. A ratio of 10,000 means the evidence is 10,000 times more probable under the prosecution's hypothesis than the defence's. Courts prefer this framing to a simple match percentage because it explicitly states the competing explanation.
What is the minimum number of individuals (MNI) calculation?
MNI is the smallest number of individual animals that could account for all the biological material in a seizure. For bones you count the most frequent element on one side; for tusks you count pairs and singles; for DNA you identify distinct genotypes. MNI is conservative by design and tends to undercount, so it is a floor, not an estimate of the true number.
What is the Society for Wildlife Forensic Science reporting standard?
The SWFS minimum reporting standards specify what a wildlife forensic report must contain: the question posed, the evidence examined, the methods used, the known limitations of those methods, and the conclusion stated as a degree of support for the hypothesis rather than a bare match. Reports should not overstate certainty and should note when a reference database gap weakens a conclusion.
What happened in R v Bailey in the UK?
R v Bailey is a landmark UK case involving the persecution of raptors, specifically peregrine falcons. Forensic evidence including eggshell fragments, feather identification, and chemical residue analysis was presented under the Frye-equivalent test applied in English courts at the time. The case helped establish that specialised ornithological and chemical analysis meets the threshold for expert evidence in wildlife crime prosecution.

Test yourself on Wildlife Forensics with free, timed mocks.

Practice Wildlife Forensics questions

Found this useful? Pass it along.

Share

Spotted an error in this page? Report a correction or read our editorial standards.

Your journey to becoming a forensic professional starts here.

Practice with mock tests, learn from structured notes, and get your questions answered by a global forensic community, all in one place.