Predicted distributions of 65 groundfish species in Canadian Pacific waters
Description:
This dataset contains layers of predicted occurrence for 65 groundfish species as well as overall species richness (i.e., the total number of species present) in Canadian Pacific waters, and the median standard error per grid cell across all species. They cover all seafloor habitat depths between 10 and 1400 m that have a mean summer salinity above 28 PSU. Two layers are provided for each species: 1) predicted species occurrence (prob_occur) and 2) the probability that a grid cell is an occurrence hotspot for that species (hotspot_prob; defined as being in the lower of: 1) 0.8, or 2) the 80th percentile of the predicted probability of occurrence values across all grid cells that had a probability of occurrence greater than 0.05.). The first measure provides an overall prediction of the distribution of the species while the second metric identifies areas where that species is most likely to be found, accounting for uncertainty within our model. All layers are provided at a 1 km resolution.
Methods:
These layers were developed using a species distribution model described in Thompson et al. 2023. This model integrates data from three fisheries-independent surveys: the Fisheries and Oceans Canada (DFO) Groundfish Synoptic Bottom Trawl Surveys (Sinclair et al. 2003; Anderson et al. 2019), the DFO Groundfish Hard Bottom Longline Surveys (Lochead and Yamanaka 2006, 2007; Doherty et al. 2019), and the International Pacific Halibut Commission Fisheries Independent Setline Survey (IPHC 2021). Further details on the methods are found in the metadata PDF available with the dataset.
Abstract from Thompson et al. 2023:
Predictions of the distribution of groundfish species are needed to support ongoing marine spatial planning initiatives in Canadian Pacific waters. Data to inform species distribution models are available from several fisheries-independent surveys. However, no single survey covers the entire region and different gear types are required to survey the range of habitats that are occupied by groundfish. Bottom trawl gear is used to sample soft bottom habitat, predominantly on the continental shelf and slope, whereas longline gear often focuses on nearshore and hardbottom habitats where trawling is not possible. Because data from these two gear types are not directly comparable, previous species distribution models in this region have been limited to using data from one survey at a time, restricting their spatial extent and usefulness at a regional scale. Here we demonstrate a method for integrating presence-absence data across surveys and gear types that allows us to predict the coastwide distributions of 66 groundfish species in British Columbia. Our model leverages the use of available data from multiple surveys to estimate how species respond to environmental gradients while accounting for differences in catchability by the different surveys. Overall, we find that this integrated method has two main benefits: 1) it increases the accuracy of predictions in data-limited surveys and regions while having negligible impacts on the accuracy when data are already sufficient to make predictions, 2) it reduces uncertainty, resulting in tighter confidence intervals on predicted species occurrences. These benefits are particularly relevant in areas of our coast where our understanding of habitat suitability is limited due to a lack of spatially comprehensive long-term groundfish research surveys.
Data Sources:
Research data was provided by Pacific Science’s Groundfish Data Unit for research surveys from the GFBio database between 2003 and 2020 for all species which had at least 150 observations, across all gear type and survey datasets available.
Uncertainties:
These are modeled results based on species observations at sea and their related environmental covariate predictions that may not always accurately reflect real-world groundfish distributions though methods that integrate different data types/sources have been demonstrated to improve model inference by increasing the accuracy of the predictions and reducing uncertainty.
Jeux de données disponibles au téléchargement
Informations supplémentaires
| Champ | Valeur |
|---|---|
| Dernière mise à jour | octobre 20, 2025, 02:40 (TU) |
| Créé | octobre 20, 2025, 02:40 (TU) |
|
Domaine / Sujet
Domaine ou sujet du jeu de données catalogué.
|
Biota |
|
Titre
Titre du jeu de données.
|
Predicted distributions of 65 groundfish species in Canadian Pacific waters |
|
Description
Une description du jeu de données.
|
Description: This dataset contains layers of predicted occurrence for 65 groundfish species as well as overall species richness (i.e., the total number of species present) in Canadian Pacific waters, and the median standard error per grid cell across all species. They cover all seafloor habitat depths between 10 and 1400 m that have a mean summer salinity above 28 PSU. Two layers are provided for each species: 1) predicted species occurrence (prob_occur) and 2) the probability that a grid cell is an occurrence hotspot for that species (hotspot_prob; defined as being in the lower of: 1) 0.8, or 2) the 80th percentile of the predicted probability of occurrence values across all grid cells that had a probability of occurrence greater than 0.05.). The first measure provides an overall prediction of the distribution of the species while the second metric identifies areas where that species is most likely to be found, accounting for uncertainty within our model. All layers are provided at a 1 km resolution. Methods: These layers were developed using a species distribution model described in Thompson et al. 2023. This model integrates data from three fisheries-independent surveys: the Fisheries and Oceans Canada (DFO) Groundfish Synoptic Bottom Trawl Surveys (Sinclair et al. 2003; Anderson et al. 2019), the DFO Groundfish Hard Bottom Longline Surveys (Lochead and Yamanaka 2006, 2007; Doherty et al. 2019), and the International Pacific Halibut Commission Fisheries Independent Setline Survey (IPHC 2021). Further details on the methods are found in the metadata PDF available with the dataset. Abstract from Thompson et al. 2023: Predictions of the distribution of groundfish species are needed to support ongoing marine spatial planning initiatives in Canadian Pacific waters. Data to inform species distribution models are available from several fisheries-independent surveys. However, no single survey covers the entire region and different gear types are required to survey the range of habitats that are occupied by groundfish. Bottom trawl gear is used to sample soft bottom habitat, predominantly on the continental shelf and slope, whereas longline gear often focuses on nearshore and hardbottom habitats where trawling is not possible. Because data from these two gear types are not directly comparable, previous species distribution models in this region have been limited to using data from one survey at a time, restricting their spatial extent and usefulness at a regional scale. Here we demonstrate a method for integrating presence-absence data across surveys and gear types that allows us to predict the coastwide distributions of 66 groundfish species in British Columbia. Our model leverages the use of available data from multiple surveys to estimate how species respond to environmental gradients while accounting for differences in catchability by the different surveys. Overall, we find that this integrated method has two main benefits: 1) it increases the accuracy of predictions in data-limited surveys and regions while having negligible impacts on the accuracy when data are already sufficient to make predictions, 2) it reduces uncertainty, resulting in tighter confidence intervals on predicted species occurrences. These benefits are particularly relevant in areas of our coast where our understanding of habitat suitability is limited due to a lack of spatially comprehensive long-term groundfish research surveys. Data Sources: Research data was provided by Pacific Science’s Groundfish Data Unit for research surveys from the GFBio database between 2003 and 2020 for all species which had at least 150 observations, across all gear type and survey datasets available. Uncertainties: These are modeled results based on species observations at sea and their related environmental covariate predictions that may not always accurately reflect real-world groundfish distributions though methods that integrate different data types/sources have been demonstrated to improve model inference by increasing the accuracy of the predictions and reducing uncertainty. |
|
Étiquettes / Mots-clés
Mots-clés/étiquettes catégorisant le jeu de données.
|
|
|
Format (CSV, XLS, TXT, PDF, etc.)
Format de fichier du jeu de données.
|
|
|
Taille du jeu de données
Taille du jeu de données en mégaoctets.
|
|
|
Identifiant des métadonnées
Identifiant des métadonnées — peut être utilisé comme identifiant unique pour l’entrée du catalogue.
|
|
|
Date de publication
Date de publication du jeu de données.
|
2023-02-27 |
|
Période temporelle des données (date de début)
Date de début des données dans le jeu de données.
|
|
|
Période temporelle des données (date de fin)
Date de fin des données temporelles du jeu de données.
|
|
|
Zone géospatiale couverte
Région spatiale ou lieu nommé couvert par le jeu de données.
|
| Champ | Valeur |
|---|---|
|
Catégorie d’accès
Type d’accès accordé pour le jeu de données (ouvert, restreint, service, etc.).
|
|
|
Licence
Licence utilisée pour accéder au jeu de données.
|
Open Government Licence - Canada |
|
Restrictions d’utilisation
Restrictions d’utilisation des données.
|
|
|
Emplacement
Emplacement du jeu de données.
|
https://open.canada.ca/data/en/dataset/51c60d88-c6ac-4e1c-9724-83b6048aeccd |
|
Service de données
Service de données pour accéder à un jeu de données.
|
|
|
Propriétaire
Propriétaire du jeu de données.
|
Fisheries and Oceans Canada | Pêches et Océans Canada |
|
Point de contact
Qui contacter concernant l’accès ?
|
Government of Canada; Fisheries and Oceans Canada; Pacific Science, 604-822-8419, [email protected] |
|
Courriel du point de contact
Courriel à utiliser pour l’accès ?
|
|
|
Éditeur
Éditeur du jeu de données.
|
|
|
Courriel de l’éditeur
Adresse courriel de l’éditeur.
|
[email protected] |
|
Auteur
Auteur du jeu de données.
|
|
|
Courriel de l'auteur
Adresse courriel de l’auteur.
|
|
|
Date d’accès
Date à laquelle les données et métadonnées ont été consultées.
|
| Champ | Valeur |
|---|---|
|
Identifiant
Identifiant unique du jeu de données.
|
|
|
Langue
Langue(s) du jeu de données
|
|
|
Lien vers la description du jeu de données
Une URL vers un document externe décrivant le jeu de données.
|
|
|
Identifiant persistant
Les données sont identifiées par un identifiant persistant.
|
|
|
Identifiant globalement unique
Les données sont identifiées par un identifiant persistant et globalement unique.
|
|
|
Contient des données sur des individus
Le jeu de données contient-il des données sur des individus ?
|
|
|
Contient des données identifiables sur des individus
Le jeu de données contient-il des données identifiables sur des individus ?
|
|
|
Contient des données autochtones
Le jeu de données contient-il des données sur des communautés autochtones ?
|
|
|
Type de portail
Type de plateforme du portail source.
|
| Champ | Valeur |
|---|---|
|
Version
Version du jeu de données
|
None |
|
Source
Source du jeu de données.
|
None |
|
Notes de version
Notes de version à propos du jeu de données.
|
|
|
Est une version d’un autre jeu de données
Lien vers le jeu de données dont celui-ci est une version.
|
|
|
Autres versions
Lien vers les jeux de données qui en sont des versions.
|
|
|
Texte de provenance
Texte de provenance des données.
|
|
|
URL de provenance
URL de provenance des données.
|
|
|
Résolution temporelle
Décrit la granularité des données temporelles dans le jeu de données.
|
|
|
Résolution géospatiale en mètres
Décrit la granularité (en mètres) des données géospatiales dans le jeu de données.
|
|
|
Résolution géospatiale (par régions)
Décrit la granularité (par régions) des données géospatiales dans le jeu de données.
|
| Champ | Valeur |
|---|---|
|
Permission de la communauté autochtone
Qui détient la permission de la communauté autochtone. Qui contacter pour l’accès à un jeu de données concernant des communautés autochtones.
|
|
|
Permission communautaire
Permission communautaire (qui a accordé la permission).
|
|
|
Les communautés autochtones concernées par le jeu de données
Communautés autochtones dont sont issues les données.
|
| Champ | Valeur |
|---|---|
|
Nombre de lignes de données
Pour un jeu de données tabulaire, nombre total de lignes.
|
|
|
Nombre de colonnes de données
Pour un jeu de données tabulaire, nombre total de colonnes uniques.
|
|
|
Nombre de cellules de données
Pour un jeu de données tabulaire, nombre total de cellules contenant des données.
|
|
|
Nombre de relations de données
Pour un jeu de données RDF, nombre total de triplets.
|
|
|
Nombre d’entités
Pour un jeu de données RDF, nombre total d’entités.
|
|
|
Nombre de propriétés de données
Pour un jeu de données RDF, nombre total de propriétés uniques utilisées par les triplets.
|
|
|
Qualité des données
Décrit la qualité des données du jeu de données.
|
|
|
Métrique de qualité des données
Une métrique utilisée pour mesurer la qualité des données, telle que les valeurs manquantes ou les formats invalides.
|
0 Commentaires