Deep diveMetals & Mining

How machine-learning mineral prospectivity models are built and validated

A prospectivity model ranks ground by how closely its evidence resembles the footprint of a target deposit type, so an exploration team can spend a limited drilling budget where the geology argues hardest. Building one worth trusting means translating a mineral systems model into mappable evidence, choosing a method that suits the labels you actually have, and validating it in a way that spatial data cannot flatter. Its output guides targeting; it is not a resource.

Reviewed 8 min read

On this page
  1. Target ranking and drill budget allocation
  2. From a mineral systems model to ranked drill targets
  3. Terms that come up in prospectivity work
  4. Preparing evidence layers before any modelling starts
  5. Knowledge-driven and data-driven methods side by side
  6. Scarce labels and the negative-sampling problem
  7. Validating a prospectivity model so it cannot flatter itself
  8. Where a prospectivity map stops and public reporting begins
  9. Ranking porphyry copper targets across a hypothetical district
  10. Questions and answers
  11. Sources

Target ranking and drill budget allocation

Exploration is a sequence of spending decisions made under uncertainty: which licences to keep, which areas to map in detail, which anomalies deserve drill holes. A prospectivity model supports those decisions by scoring every cell of a grid on how strongly the available evidence matches what a deposit of the target style would leave behind. The useful output is a ranking, ideally with uncertainty attached, that geologists can interrogate and argue with.

The model does not find deposits; it concentrates attention. It earns its place if, in ground it has not seen, it would have led the team to the known occurrences while flagging only a small share of the area, and if the geologists can say in plain terms why a given cell scores high.

From a mineral systems model to ranked drill targets

01Deposit model02Critical processes03Mappable proxies04Evidence layers05Model and validate06Ranked targets
  1. Deposit model

    The ore-forming process for the commodity and deposit style, such as porphyry copper.

  2. Critical processes

    Source, fluid pathway, trap and preservation: each a necessary ingredient of the system.

  3. Mappable proxies

    Observable features standing in for each process, such as intrusive contacts, faults or alteration.

  4. Evidence layers

    Gridded maps of each proxy built from geology, geophysics, geochemistry and remote sensing.

  5. Model and validate

    Combine the layers with a knowledge-driven or data-driven method and test it by area.

  6. Ranked targets

    A shortlist with scores, uncertainty and the evidence behind each, ready for field checks.

Conceptual workflow based on the mineral systems approach to exploration targeting; it illustrates the logic and is not a fixed procedure.

Terms that come up in prospectivity work

Mineral systems approach
Treating a deposit as the outcome of geological processes acting across scales, then targeting the evidence each process leaves rather than the deposit alone1.
Targeting criterion (mappable proxy)
A feature that can be mapped across the whole study area and represents one critical process, for example proximity to a fault set that could have channelled fluids.
Evidence layer
One proxy rendered as a grid at a shared resolution and projection, carrying its own uncertainty and source scale.
Positive and negative labels
Cells known to host an occurrence of the target type, and cells assumed barren. The second set is usually inferred rather than observed.
Spatial autocorrelation
The tendency of neighbouring cells to resemble one another, which lets a model score well by memorising neighbourhoods instead of learning geology.
Prediction-area plot
A chart comparing the share of known occurrences captured with the share of ground flagged, used to judge how efficiently a map concentrates targets2.

Preparing evidence layers before any modelling starts

0 of 6 checked

Knowledge-driven and data-driven methods side by side

CriterionFuzzy logic or index overlayWeights of evidenceRandom forest or gradient boostingNeural networks
What sets the weightsExpert judgement on each proxyStatistics from known occurrences, layer by layerLearned from labels, including interactions between layersLearned from labels, including spatial patterns
Labels neededNone, which suits frontier groundA modest set of occurrencesMore occurrences plus credible negativesThe most; risky in sparse districts
Correlated layersHandled only as well as the expert handles themPoorly, because it assumes conditional independenceReasonably wellWell, given enough data
ExplainabilityFully transparentTransparent per layerGood with feature-attribution methodsHardest to explain to a geologist
Typical failureEncodes the expert's blind spotsInflated scores from dependent layersMemorises clustered training areasOverfits small datasets
Where it fitsGreenfield belts with few known depositsMature districts with a clear deposit modelDistricts with enough occurrences and even coverageLarge, dense datasets such as continental grids

Hybrid workflows are common: expert-built proxies feed a data-driven model, and an expert overlay acts as a sense check on its ranking.

Scarce labels and the negative-sampling problem

A district may contain only a handful of known occurrences of the target style, often clustered where outcrop is good or where earlier explorers happened to look. Data-driven methods also need negatives, and genuinely barren ground is rarely known. Teams typically sample negatives at random from areas far from known occurrences, sample them from areas an expert rates as unfavourable, or use positive-unlabelled learning, which treats everything without a label as uncertain.

Every choice biases the result. Random negatives can include undiscovered deposits; expert-filtered negatives import the expert's assumptions; clustered positives teach the model what well-explored ground looks like rather than what mineralised ground looks like. Test how much the ranking moves under different negative strategies and report that sensitivity, instead of presenting one run as the answer.

Validating a prospectivity model so it cannot flatter itself

  1. Split by area, not at random

    Divide the study area into spatial blocks or hold out whole sub-districts, so test cells never sit next to training cells. Random splits on autocorrelated data report accuracy the model will not achieve on new ground3.

  2. Hold some occurrences out of everything

    Keep a few known deposits out of every stage, including the choice of layers, and check where they land in the final ranking.

  3. Plot success-rate and prediction-area curves

    Show what share of held-out occurrences fall inside the top-ranked share of the area. A model that needs most of the map to capture them adds little2.

  4. Beat a simple baseline

    Benchmark against distance to known mineralisation or an expert overlay. A complex model has to outperform the obvious answer to justify itself.

  5. Map uncertainty, not only the score

    Rerun the model across resampled labels and parameter choices, then show where the ranking is stable and where it flips between runs.

Where a prospectivity map stops and public reporting begins

Ranking porphyry copper targets across a hypothetical district

Questions and answers

How much data is enough for a machine-learning prospectivity model?

There is no fixed threshold. What matters is whether the known occurrences are numerous and spread out enough to survive a spatial hold-out, and whether every evidence layer covers the whole area evenly. With only a few clustered occurrences, a knowledge-driven or weights-of-evidence approach is usually more honest than a complex learner, and data-driven methods can be added as drilling produces more labels.

Do prospectivity models transfer from one region to another?

Only with care. A model trained in one belt learns that belt's data coverage, mapping conventions and erosion level as much as its geology. The mineral systems logic and the choice of proxies usually transfer well; the trained weights often do not. Retrain or recalibrate on local data, and validate on held-out local occurrences before using a transferred model to rank ground.

How should drilling data be managed for prospectivity work?

Keep collars, surveys, lithology logs and assays in one versioned database with original document references, quality-control results and the date each record was added. Record holes that found nothing as carefully as those that hit mineralisation. Freeze the exact extract used for each model run, so a ranking can be reproduced and audited when it is challenged by management or a partner.

Should barren drill holes be used as negative labels?

They are among the best negatives available, because they are observed rather than assumed. Use them with their limits in mind: a hole tests a small volume, may have stopped short of the target, or may have been drilled for a different deposit style. Treat them as strong but local evidence, and combine them with a sampling strategy for the large areas no one has drilled.

Sources

  1. Translating the mineral systems approach into an effective exploration targeting system (McCuaig, Beresford and Hronsky), Ore Geology Reviews — Elsevier · checked 10 October 2026
  2. Prediction-area (P-A) plot and C-A fractal analysis to classify and evaluate evidential maps for mineral prospectivity modeling (Yousefi and Carranza), Computers & Geosciences — Elsevier · checked 10 October 2026
  3. Cross-validation strategies for data with temporal, spatial, hierarchical, or phylogenetic structure (Roberts et al.), Ecography — Wiley · checked 10 October 2026
  4. JORC Code review: update and timeline — Joint Ore Reserves Committee · checked 10 October 2026
  5. NI 43-101 Standards of Disclosure for Mineral Projects — Ontario Securities Commission · checked 10 October 2026
  6. CRIRSCO International Reporting Template and member reporting standards — Committee for Mineral Reserves International Reporting Standards · checked 10 October 2026

More in Metals & Mining

Back to Metals & Mining

Next step

Send us an inventory of your exploration data layers

List the layers you hold, their coverage and the deposit model you are targeting. We will reply with a view on which modelling approach your data can support and how to validate it before it shapes a drilling budget.

Discuss a prospectivity model