arXiv ScienceSearch

arXiv subjects

Rasmus Astrup

Publications and source records attributed to Rasmus Astrup.

At least 19 recordsLinked to original sources

SegmentAnyTreeV2: Scaling Transformer-Based Tree Instance Segmentation Across Sensors, Platforms, and Forests

We present SegmentAnyTreeV2, a sensor- and platform-agnostic framework for semantic and instance segmentation of forest point clouds. The model combines a serialization-based Point Transformer v3 backbone with a lightweight semantic head and a tree-focused cross-attention mask decoder. Semantic predictions restrict instance decoding to tree-class voxels, while instance-aware query initialization, one-to-many seed supervision, and asymmetric mask scoring improve separation in dense and structurally complex stands. We further introduce FOR-instance v3, an expanded benchmark comprising 427 scenes and 26,496 annotated trees across diverse biomes, forest structures, and LiDAR platforms. On the FOR-instanceV2 test split, SegmentAnyTreeV2 achieves 90.5% precision, 80.2% recall, 85.0% F1, 90.7% coverage, and 87.6% semantic mIoU, outperforming previous learning-based methods in both instance detection and mask completeness. Zero-shot evaluation on independent sites further demonstrates strong cross-domain generalization.

cs.CV

ForestFormer3D: A Unified Framework for End-to-End Segmentation of Forest LiDAR 3D Point Clouds

The segmentation of forest LiDAR 3D point clouds, including both individual tree and semantic segmentation, is fundamental for advancing forest management and ecological research. However, current approaches often struggle with the complexity and variability of natural forest environments. We present ForestFormer3D, a new unified and end-to-end framework designed for precise individual tree and semantic segmentation. ForestFormer3D incorporates ISA-guided query point selection, a score-based block merging strategy during inference, and a one-to-many association mechanism for effective training. By combining these new components, our model achieves state-of-the-art performance for individual tree segmentation on the newly introduced FOR-instanceV2 dataset, which spans diverse forest types and regions. Additionally, ForestFormer3D generalizes well to unseen test sets (Wytham woods and LAUTx), showcasing its robustness across different forest conditions and sensor modalities. The FOR-instanceV2 dataset and the ForestFormer3D code are publicly available at https://bxiang233.github.io/FF3D/.

cs.CV

Multispectral airborne laser scanning for tree species classification: a benchmark of machine learning and deep learning algorithms

Climate-smart and biodiversity-preserving forestry demands precise information on forest resources, extending to the individual tree level. Multispectral airborne laser scanning (ALS) has shown promise in automated point cloud processing, but challenges remain in leveraging deep learning techniques and identifying rare tree species in class-imbalanced datasets. This study addresses these gaps by conducting a comprehensive benchmark of deep learning and traditional shallow machine learning methods for tree species classification. For the study, we collected high-density multispectral ALS data ($>1000$ $\mathrm{pts}/\mathrm{m}^2$) at three wavelengths using the FGI-developed HeliALS system, complemented by existing Optech Titan data (35 $\mathrm{pts}/\mathrm{m}^2$), to evaluate the species classification accuracy of various algorithms in a peri-urban study area located in southern Finland. We established a field reference dataset of 6326 segments across nine species using a newly developed browser-based crowdsourcing tool, which facilitated efficient data annotation. The ALS data, including a training dataset of 1065 segments, was shared with the scientific community to foster collaborative research and diverse algorithmic contributions. Based on 5261 test segments, our findings demonstrate that point-based deep learning methods, particularly a point transformer model, outperformed traditional machine learning and image-based deep learning approaches on high-density multispectral point clouds. For the high-density ALS dataset, a point transformer model provided the best performance reaching an overall (macro-average) accuracy of 87.9% (74.5%) with a training set of 1065 segments and 92.0% (85.1%) with a larger training set of 5000 segments.

cs.CV

BranchPoseNet: Characterizing tree branching with a deep learning-based pose estimation approach

This paper presents an automated pipeline for detecting tree whorls in proximally laser scanning data using a pose-estimation deep learning model. Accurate whorl detection provides valuable insights into tree growth patterns, wood quality, and offers potential for use as a biometric marker to track trees throughout the forestry value chain. The workflow processes point cloud data to create sectional images, which are subsequently used to identify keypoints representing tree whorls and branches along the stem. The method was tested on a dataset of destructively sampled individual trees, where the whorls were located along the stems of felled trees. The results demonstrated strong potential, with accurate identification of tree whorls and precise calculation of key structural metrics, unlocking new insights and deeper levels of information from individual tree point clouds.

cs.CV

Benchmarking tree species classification from proximally-sensed laser scanning data: introducing the FOR-species20K dataset

Proximally-sensed laser scanning offers significant potential for automated forest data capture, but challenges remain in automatically identifying tree species without additional ground data. Deep learning (DL) shows promise for automation, yet progress is slowed by the lack of large, diverse, openly available labeled datasets of single tree point clouds. This has impacted the robustness of DL models and the ability to establish best practices for species classification. To overcome these challenges, the FOR-species20K benchmark dataset was created, comprising over 20,000 tree point clouds from 33 species, captured using terrestrial (TLS), mobile (MLS), and drone laser scanning (ULS) across various European forests, with some data from other regions. This dataset enables the benchmarking of DL models for tree species classification, including both point cloud-based (PointNet++, MinkNet, MLP-Mixer, DGCNNs) and multi-view image-based methods (SimpleView, DetailView, YOLOv5). 2D image-based models generally performed better (average OA = 0.77) than 3D point cloud-based models (average OA = 0.72), with consistent results across different scanning platforms and sensors. The top model, DetailView, was particularly robust, handling data imbalances well and generalizing effectively across tree sizes. The FOR-species20K dataset, available at https://zenodo.org/records/13255198, is a key resource for developing and benchmarking DL models for tree species classification using laser scanning data, providing a foundation for future advancements in the field.

cs.CV

SegmentAnyTree: A sensor and platform agnostic deep learning model for tree segmentation using laser scanning data

This research advances individual tree crown (ITC) segmentation in lidar data, using a deep learning model applicable to various laser scanning types: airborne (ULS), terrestrial (TLS), and mobile (MLS). It addresses the challenge of transferability across different data characteristics in 3D forest scene analysis. The study evaluates the model's performance based on platform (ULS, MLS) and data density, testing five scenarios with varying input data, including sparse versions, to gauge adaptability and canopy layer efficacy. The model, based on PointGroup architecture, is a 3D CNN with separate heads for semantic and instance segmentation, validated on diverse point cloud datasets. Results show point cloud sparsification enhances performance, aiding sparse data handling and improving detection in dense forests. The model performs well with >50 points per sq. m densities but less so at 10 points per sq. m due to higher omission rates. It outperforms existing methods (e.g., Point2Tree, TLS2trees) in detection, omission, commission rates, and F1 score, setting new benchmarks on LAUTx, Wytham Woods, and TreeLearn datasets. In conclusion, this study shows the feasibility of a sensor-agnostic model for diverse lidar data, surpassing sensor-specific approaches and setting new standards in tree segmentation, particularly in complex forests. This contributes to future ecological modeling and forest management advancements.

cs.CV

Automated forest inventory: analysis of high-density airborne LiDAR point clouds with 3D deep learning

Detailed forest inventories are critical for sustainable and flexible management of forest resources, to conserve various ecosystem services. Modern airborne laser scanners deliver high-density point clouds with great potential for fine-scale forest inventory and analysis, but automatically partitioning those point clouds into meaningful entities like individual trees or tree components remains a challenge. The present study aims to fill this gap and introduces a deep learning framework, termed ForAINet, that is able to perform such a segmentation across diverse forest types and geographic regions. From the segmented data, we then derive relevant biophysical parameters of individual trees as well as stands. The system has been tested on FOR-Instance, a dataset of point clouds that have been acquired in five different countries using surveying drones. The segmentation back-end achieves over 85% F-score for individual trees, respectively over 73% mean IoU across five semantic categories: ground, low vegetation, stems, live branches and dead branches. Building on the segmentation results our pipeline then densely calculates biophysical features of each individual tree (height, crown diameter, crown volume, DBH, and location) and properties per stand (digital terrain model and stand density). Especially crown-related features are in most cases retrieved with high accuracy, whereas the estimates for DBH and location are less reliable, due to the airborne scanning setup.

cs.CV

Multi-Sensor Terrestrial SLAM for Real-Time, Large-Scale, and GNSS-Interrupted Forest Mapping

Forests, as critical components of our ecosystem, demand effective monitoring and management. However, conducting real-time forest inventory in large-scale and GNSS-interrupted forest environments has long been a formidable challenge. In this paper, we present a novel solution that leverages robotics and sensor-fusion technologies to overcome these challenges and enable real-time forest inventory with higher accuracy and efficiency. The proposed solution consists of a new SLAM algorithm to create an accurate 3D map of large-scale forest stands with detailed estimation about the number of trees and the corresponding DBH, solely with the consecutive scans of a 3D lidar and an imu. This method utilized a hierarchical unsupervised clustering algorithm to detect the trees and measure the DBH from the lidar point cloud. The algorithm can run simultaneously as the data is being recorded or afterwards on the recorded dataset. Furthermore, due to the proposed fast feature extraction and transform estimation modules, the recorded data can be fed to the SLAM with higher frequency than common SLAM algorithms. The performance of the proposed solution was tested through filed data collection with hand-held sensor platform as well as a mobile forestry robot. The accuracy of the results was also compared to the state-of-the-art SLAM solutions.

cs.RO

FOR-instance: a UAV laser scanning benchmark dataset for semantic and instance segmentation of individual trees

The FOR-instance dataset (available at https://doi.org/10.5281/zenodo.8287792) addresses the challenge of accurate individual tree segmentation from laser scanning data, crucial for understanding forest ecosystems and sustainable management. Despite the growing need for detailed tree data, automating segmentation and tracking scientific progress remains difficult. Existing methodologies often overfit small datasets and lack comparability, limiting their applicability. Amid the progress triggered by the emergence of deep learning methodologies, standardized benchmarking assumes paramount importance in these research domains. This data paper introduces a benchmarking dataset for dense airborne laser scanning data, aimed at advancing instance and semantic segmentation techniques and promoting progress in 3D forest scene segmentation. The FOR-instance dataset comprises five curated and ML-ready UAV-based laser scanning data collections from diverse global locations, representing various forest types. The laser scanning data were manually annotated into individual trees (instances) and different semantic classes (e.g. stem, woody branches, live branches, terrain, low vegetation). The dataset is divided into development and test subsets, enabling method advancement and evaluation, with specific guidelines for utilization. It supports instance and semantic segmentation, offering adaptability to deep learning frameworks and diverse segmentation strategies, while the inclusion of diameter at breast height data expands its utility to the measurement of a classic tree variable. In conclusion, the FOR-instance dataset contributes to filling a gap in the 3D forest research, enhancing the development and benchmarking of segmentation algorithms for dense airborne laser scanning data.

cs.CV

Towards accurate instance segmentation in large-scale LiDAR point clouds

Panoptic segmentation is the combination of semantic and instance segmentation: assign the points in a 3D point cloud to semantic categories and partition them into distinct object instances. It has many obvious applications for outdoor scene understanding, from city mapping to forest management. Existing methods struggle to segment nearby instances of the same semantic category, like adjacent pieces of street furniture or neighbouring trees, which limits their usability for inventory- or management-type applications that rely on object instances. This study explores the steps of the panoptic segmentation pipeline concerned with clustering points into object instances, with the goal to alleviate that bottleneck. We find that a carefully designed clustering strategy, which leverages multiple types of learned point embeddings, significantly improves instance segmentation. Experiments on the NPM3D urban mobile mapping dataset and the FOR-instance forest dataset demonstrate the effectiveness and versatility of the proposed strategy.

cs.CV

Point2Tree(P2T) -- framework for parameter tuning of semantic and instance segmentation used with mobile laser scanning data in coniferous forest

This article introduces Point2Tree, a novel framework that incorporates a three-stage process involving semantic segmentation, instance segmentation, optimization analysis of hyperparemeters importance. It introduces a comprehensive and modular approach to processing laser points clouds in Forestry. We tested it on two independent datasets. The first area was located in an actively managed boreal coniferous dominated forest in V{\aa}ler, Norway, 16 circular plots of 400 square meters were selected to cover a range of forest conditions in terms of species composition and stand density. We trained a model based on Pointnet++ architecture which achieves 0.92 F1-score in semantic segmentation. As a second step in our pipeline we used graph-based approach for instance segmentation which reached F1-score approx. 0.6. The optimization allowed to further boost the performance of the pipeline by approx. 4 \% points.

cs.CV

Assessing and mitigating systematic errors in forest attribute maps utilizing harvester and airborne laser scanning data

Cut-to-length harvesters collect useful information for modeling relationships between forest attributes and airborne laser scanning (ALS) data. However, harvesters operate in mature forests, which may introduce selection biases that can result in systematic errors in harvester data-based forest attribute maps. We fitted regression models (harvester models) for volume (V), height (HL), stem frequency (N), above-ground biomass, basal area, and quadratic mean diameter (QMD) using harvester and ALS data. Performances of the harvester models were evaluated using national forest inventory plots in an 8.7 Mha study area. We estimated biases of large-area synthetic estimators and compared efficiencies of model-assisted (MA) estimators with field data-based direct estimators. The harvester models performed better in productive than unproductive forests, but systematic errors occurred in both. The use of MA estimators resulted in efficiency gains that were largest for HL (relative efficiency, RE=6.0) and smallest for QMD (RE=1.5). The bias of the synthetic estimator was largest for N (39%) and smallest for V (1%). The latter was due to an overestimation of deciduous and an underestimation of spruce forests that by chance balanced. We conclude that a probability sample of reference observations may be required to ensure the unbiasedness of estimators utilizing harvester data.

stat.AP

Prediction of butt rot volume in Norway spruce forest stands using harvester, remotely sensed and environmental data

Butt rot (BR) damages associated with Norway spruce (Picea abies [L.] Karst.) account for considerable economic losses in timber production across the northern hemisphere. While information on BR damages is critical for optimal decision-making in forest management, the maps of BR damages are typically lacking in forest information systems. We predicted timber volume damaged by BR at the stand-level in Norway using harvester information of 186,026 stems (clear-cuts), remotely sensed, and environmental data (e.g. climate and terrain characteristics). We utilized random forest (RF) models with two sets of predictor variables: (1) predictor variables available after harvest (theoretical case) and (2) predictor variables available prior to harvest (mapping case). We found that forest attributes characterizing the maturity of forest, such as remote sensing-based height, harvested timber volume and quadratic mean diameter at breast height, were among the most important predictor variables. Remotely sensed predictor variables obtained from airborne laser scanning data and Sentinel-2 imagery were more important than the environmental variables. The theoretical case with a leave-stand-out cross-validation achieved an RMSE of 11.4 $m^3ha^{-1}$ (pseudo $R^2$: 0.66) whereas the mapping case resulted in a pseudo $R^2$ of 0.60. When the spatially distinct k-means clusters of harvested forest stands were used as units in the cross-validation, the RMSE value and pseudo $R^2$ associated with the mapping case were 15.6 $m^3ha^{-1}$ and 0.37, respectively. This indicates that the knowledge about the BR status of spatially close stands is of high importance for obtaining satisfactory error rates in the mapping of BR damages.

stat.AP

Improving living biomass C-stock loss estimates by combining optical satellite, airborne laser scanning, and NFI data

Policy measures and management decisions aiming at enhancing the role of forests in mitigating climate-change require reliable estimates of C-stock dynamics in greenhouse gas inventories (GHGIs). Aim of this study was to assemble design-based estimators to provide estimates relevant for GHGIs using national forest inventory (NFI) data. We improve basic expansion (BE) estimates of living-biomass C-stock loss using field-data only, by leveraging with remotely-sensed auxiliary data in model-assisted (MA) estimates. Our case studies from Norway, Sweden, Denmark, and Latvia covered an area of >70 Mha. Landsat-based Forest Cover Loss (FCL) and one-time wall-to-wall airborne laser scanning (ALS) data served as auxiliary data. ALS provided information on the C-stock before a potential disturbance indicated by FCL. The use of FCL in MA estimators resulted in considerable efficiency gains which in most cases were further increased by using ALS in addition. A doubling of efficiency was possible for national estimates and even larger efficiencies were observed at the sub-national level. Average annual estimates were considerably more precise than pooled estimates using NFI data from all years at once. The combination of remotely-sensed with NFI field data yields reliable estimates which is not necessarily the case when using remotely-sensed data without reference observations.

stat.AP

Timber Volume Estimation Based on Airborne Laser Scanning -- Comparing the Use of National Forest Inventory and Forest Management Inventory Data

Large-scale forest resource maps based on national forest inventory (NFI) data and airborne laser scanning may facilitate synergies between NFIs and forest management inventories (FMIs). A comparison of models used in such a NFI-based map and a FMI indicate that NFI-based maps can directly be used in FMIs to estimate timber volume of mature spruce forests. Traditionally, FMIs and NFIs have been separate activities. The increasing availability of detailed NFI-based forest resource maps provides the possibility to eliminate or reduce the need of field sample plot measurements in FMIs if their accuracy is similar. We aim to 1) compare a timber volume model used in a NFI-based map and models used in a FMI, and 2) evaluate utilizing additional local sample plots in the model of the NFI-based map. Accuracies of timber volume estimates using models from an existing NFI-based map and a FMI were compared at plot and stand level. Estimates from the NFI-based map were similar to or more accurate than the FMI. The addition of local plots to the modeling data did not clearly improve the model of the NFI-based map.The comparison indicates that NFI-based maps can directly be used in FMIs for timber volume estimation in mature spruce stands, leading to potentially large cost savings.

stat.AP

Above-ground biomass change estimation using national forest inventory data with Sentinel-2 and Landsat 8

This study aimed at estimating total forest above-ground net change (Delta AGB, Mt) over five years (2014-2019) based on model-assisted estimation utilizing freely available satellite imagery. The study was conducted for a boreal forest area (approx. 1.4 Mill hectares) in Norway where bi-temporal national forest inventory (NFI), Sentinel-2, and Landsat data were available. Biomass change was modelled based on a direct approach. The precision of estimates using only the NFI data in a basic expansion estimator were compared to four different alternative model-assisted estimates using 1) Sentinel-2 or Landsat data, and 2) using bi- or uni-temporal remotely sensed data. We found that the use of remotely sensed data improved the precision of the purely field-based estimates by a factor of up to three. The most precise estimates were found for the model-assisted estimation using bi-temporal Sentinel-2 (standard error; SE= 1.7 Mt). However, the decrease in precision when using Landsat data was small (SE= 1.92 Mt). In addition, we found that Delta AGB could be precisely estimated also when remotely sensed data were available only at the end of the monitoring period. We conclude that satellite optical data can considerably improve Delta AGB estimates, even in those cases where repeated and coincident NFI data are available. The free availability, global coverage, frequent update, and long-term time horizon make data from programs such as Sentinel-2 and Landsat a valuable data source for a consistent and durable monitoring of forest carbon dynamics.

stat.AP

Prediction and model-assisted estimation of diameter distributions using Norwegian national forest inventory and airborne laser scanning data

Diameter at breast height (DBH) distributions offer valuable information for operational and strategic forest management decisions. We predicted DBH distributions using Norwegian national forest inventory and airborne laser scanning data and compared the predictive performances of linear mixed-effects (PPM), generalized linear-mixed (GLM) and k nearest neighbor (NN) models. While GLM resulted in smaller prediction errors than PPM, both were clearly outperformed by NN. We therefore studied the ability of the NN model to improve the precision of stem frequency estimates by DBH classes in the 8.7 Mha study area using a model-assisted (MA) estimator suitable for systematic sampling. MA estimates yielded greater than or approximately equal efficiencies as direct estimates using field data only. The relative efficiencies (REs) associated with the MA estimates ranged between 0.95-1.47 and 0.96-1.67 for 2 and 6 cm DBH class widths, respectively, when dominant tree species were assumed to be known. The use of a predicted tree species map, instead of the observed information, decreased the REs by up to 10%.

stat.AP

Mapping forest age using National Forest Inventory, airborne laser scanning, and Sentinel-2 data

The age of forest stands is critical information for many aspects of forest management and conservation but area-wide information about forest stand age often does not exist. In this study, we developed regression models for large-scale area-wide prediction of age in Norwegian forests. For model development we used more than 4800 plots of the Norwegian National Forest Inventory (NFI) distributed over Norway between 58{\deg} and 65{\deg} northern latitude in a 181,773 km2 study area. Predictor variables were based on airborne laser scanning (ALS), Sentinel-2, and existing public map data. We performed model validation on an independent data set consisting of 63 spruce stands with known age. The best modelling strategy was to fit independent linear regression models to each observed site index (SI) level and using a SI prediction map in the application of the models. The most important predictor variable was an upper percentile of the ALS heights, and root-mean-squared-errors (RMSE) ranged between 3 and 31 years (6% to 26%) for SI-specific models, and 21 years (25%) on average. Mean deviance (MD) ranged between -1 and 3 years. The models improved with increasing SI and the RMSE were largest for low SI stands older than 100 years. Using a mapped SI, which is required for practical applications, RMSE and MD on plot-level ranged from 19 to 56 years (29% to 53%), and 5 to 37 years (5% to 31%), respectively. For the validation stands, the RMSE and MD were 12 (22%) and 2 years (3%). Tree height estimated from airborne laser scanning and predicted site index were the most important variables in the models describing age. Overall, we obtained good results, especially for stands with high SI, that could be considered for practical applications but see considerable potential for improvements, if better SI maps were available.

stat.AP