NobleBlocks

ARC Centre of Excellence for Mathematical and Statistical Frontiers

facilityParkville, Victoria, Australia

Research output, citation impact, and the most-cited recent papers from ARC Centre of Excellence for Mathematical and Statistical Frontiers (Australia). Aggregated across the NobleBlocks index of 300M+ scholarly works.

Total works
885
Citations
37.1K
h-index
84
i10-index
817
Also known as
ARC Centre of Excellence for Mathematical and Statistical Frontiers

Top-cited papers from ARC Centre of Excellence for Mathematical and Statistical Frontiers

Outstanding Challenges in the Transferability of Ecological Models
Katherine L. Yates, Phil J. Bouchet, M. Julian Caley, Kerrie Mengersen +4 more
2018· Trends in Ecology & Evolution779doi:10.1016/j.tree.2018.08.001

Predictive models are central to many scientific disciplines and vital for informing management in a rapidly changing world. However, limited understanding of the accuracy and precision of models transferred to novel conditions (their 'transferability') undermines confidence in their predictions. Here, 50 experts identified priority knowledge gaps which, if filled, will most improve model transfers. These are summarized into six technical and six fundamental challenges, which underlie the combined need to intensify research on the determinants of ecological predictability, including species traits and data quality, and develop best practices for transferring models. Of high importance is the identification of a widely applicable set of transferability metrics, with appropriate tools to quantify the sources and impacts of prediction uncertainty under novel conditions.

Unmanned Aerial Vehicles (UAVs) and Artificial Intelligence Revolutionizing Wildlife Monitoring and Conservation
Felipé Gonzalez, Glen Montes, Eduard Puig, Sandra Johnson +2 more
2016· Sensors467doi:10.3390/s16010097

Surveying threatened and invasive species to obtain accurate population estimates is an important but challenging task that requires a considerable investment in time and resources. Estimates using existing ground-based monitoring techniques, such as camera traps and surveys performed on foot, are known to be resource intensive, potentially inaccurate and imprecise, and difficult to validate. Recent developments in unmanned aerial vehicles (UAV), artificial intelligence and miniaturized thermal imaging systems represent a new opportunity for wildlife experts to inexpensively survey relatively large areas. The system presented in this paper includes thermal image acquisition as well as a video processing pipeline to perform object detection, classification and tracking of wildlife in forest or open areas. The system is tested on thermal video data from ground based and test flight footage, and is found to be able to detect all the target wildlife located in the surveyed area. The system is flexible in that the user can readily define the types of objects to classify and the object characteristics that should be considered during classification.

Airborne particles in indoor environment of homes, schools, offices and aged care facilities: The main routes of exposure
Lídia Morawska, Godwin A. Ayoko, Gwi–Nam Bae, Giorgio Buonanno +4 more
2017· Environment International376doi:10.1016/j.envint.2017.07.025

It has been shown that the exposure to airborne particulate matter is one of the most significant environmental risks people face. Since indoor environment is where people spend the majority of time, in order to protect against this risk, the origin of the particles needs to be understood: do they come from indoor, outdoor sources or both? Further, this question needs to be answered separately for each of the PM mass/number size fractions, as they originate from different sources. Numerous studies have been conducted for specific indoor environments or under specific setting. Here our aim was to go beyond the specifics of individual studies, and to explore, based on pooled data from the literature, whether there are generalizable trends in routes of exposure at homes, schools and day cares, offices and aged care facilities. To do this, we quantified the overall 24 h and occupancy weighted means of PM 10 , PM 2.5 and PN - particle number concentration. Based on this, we developed a summary of the indoor versus outdoor origin of indoor particles and compared the means to the WHO guidelines (for PM 10 and PM 2.5 ) and to the typical levels reported for urban environments (PN). We showed that the main origins of particle metrics differ from one type of indoor environment to another. For homes, outdoor air is the main origin of PM 10 and PM 2.5 but PN originate from indoor sources; for schools and day cares, outdoor air is the source of PN while PM 10 and PM 2.5 have indoor sources; and for offices, outdoor air is the source of all three particle size fractions. While each individual building is different, leading to differences in exposure and ideally necessitating its own assessment (which is very rarely done), our findings point to the existence of generalizable trends for the main types of indoor environments where people spend time, and therefore to the type of prevention measures which need to be considered in general for these environments.

A Review of Modern Computational Algorithms for Bayesian Optimal Design
Elizabeth Ryan, Christopher Drovandi, James McGree, A. N. Pettitt
2015· International Statistical Review323doi:10.1111/insr.12107

Summary Bayesian experimental design is a fast growing area of research with many real‐world applications. As computational power has increased over the years, so has the development of simulation‐based design methods, which involve a number of algorithms, such as Markov chain Monte Carlo, sequential Monte Carlo and approximate Bayes methods, facilitating more complex design problems to be solved. The Bayesian framework provides a unified approach for incorporating prior information and/or uncertainties regarding the statistical model with a utility function which describes the experimental aims. In this paper, we provide a general overview on the concepts involved in Bayesian experimental design, and focus on describing some of the more commonly used Bayesian utility functions and methods for their estimation, as well as a number of algorithms that are used to search over the design space to find the Bayesian optimal design. We also discuss other computational strategies for further research in Bayesian optimal design.

The NorWeST Summer Stream Temperature Model and Scenarios for the Western U.S.: A Crowd‐Sourced Database and New Geospatial Tools Foster a User Community and Predict Broad Climate Warming of Rivers and Streams
Daniel J. Isaak, Seth J. Wenger, Erin E. Peterson, Jay M. Ver Hoef +4 more
2017· Water Resources Research307doi:10.1002/2017wr020969

Abstract Thermal regimes are fundamental determinants of aquatic ecosystems, which makes description and prediction of temperatures critical during a period of rapid global change. The advent of inexpensive temperature sensors dramatically increased monitoring in recent decades, and although most monitoring is done by individuals for agency‐specific purposes, collectively these efforts constitute a massive distributed sensing array that generates an untapped wealth of data. Using the framework provided by the National Hydrography Dataset, we organized temperature records from dozens of agencies in the western U.S. to create the NorWeST database that hosts >220,000,000 temperature recordings from >22,700 stream and river sites. Spatial‐stream‐network models were fit to a subset of those data that described mean August water temperatures (AugTw) during 63,641 monitoring site‐years to develop accurate temperature models ( r 2 = 0.91; RMSPE = 1.10°C; MAPE = 0.72°C), assess covariate effects, and make predictions at 1 km intervals to create summer climate scenarios. AugTw averaged 14.2°C (SD = 4.0°C) during the baseline period of 1993–2011 in 343,000 km of western perennial streams but trend reconstructions also indicated warming had occurred at the rate of 0.17°C/decade (SD = 0.067°C/decade) during the 40 year period of 1976–2015. Future scenarios suggest continued warming, although variation will occur within and among river networks due to differences in local climate forcing and stream responsiveness. NorWeST scenarios and data are available online in user‐friendly digital formats and are widely used to coordinate monitoring efforts among agencies, for new research, and for conservation planning.

Spatial autoregressive models for statistical inference from ecological data
Jay M. Ver Hoef, Erin E. Peterson, Mevin B. Hooten, Ephraim M. Hanks +1 more
2017· Ecological Monographs243doi:10.1002/ecm.1283

Abstract Ecological data often exhibit spatial pattern, which can be modeled as autocorrelation. Conditional autoregressive (CAR) and simultaneous autoregressive (SAR) models are network‐based models (also known as graphical models) specifically designed to model spatially autocorrelated data based on neighborhood relationships. We identify and discuss six different types of practical ecological inference using CAR and SAR models, including: (1) model selection, (2) spatial regression, (3) estimation of autocorrelation, (4) estimation of other connectivity parameters, (5) spatial prediction, and (6) spatial smoothing. We compare CAR and SAR models, showing their development and connection to partial correlations. Special cases, such as the intrinsic autoregressive model (IAR), are described. Conditional autoregressive and SAR models depend on weight matrices, whose practical development uses neighborhood definition and row‐standardization. Weight matrices can also include ecological covariates and connectivity structures, which we emphasize, but have been rarely used. Trends in harbor seals ( Phoca vitulina ) in southeastern Alaska from 463 polygons, some with missing data, are used to illustrate the six inference types. We develop a variety of weight matrices and CAR and SAR spatial regression models are fit using maximum likelihood and Bayesian methods. Profile likelihood graphs illustrate inference for covariance parameters. The same data set is used for both prediction and smoothing, and the relative merits of each are discussed. We show the nonstationary variances and correlations of a CAR model and demonstrate the effect of row‐standardization. We include several take‐home messages for CAR and SAR models, including (1) choosing between CAR and IAR models, (2) modeling ecological effects in the covariance matrix, (3) the appeal of spatial smoothing, and (4) how to handle isolated neighbors. We highlight several reasons why ecologists will want to make use of autoregressive models, both directly and in hierarchical models, and not only in explicit spatial settings, but also for more general connectivity models.

Monitoring of Coral Reefs Using Artificial Intelligence: A Feasible and Cost-Effective Approach
Manuel González‐Rivero, Oscar Beijbom, Alberto Rodriguez‐Ramirez, Dominic E. P. Bryant +4 more
2020· Remote Sensing211doi:10.3390/rs12030489

Ecosystem monitoring is central to effective management, where rapid reporting is essential to provide timely advice. While digital imagery has greatly improved the speed of underwater data collection for monitoring benthic communities, image analysis remains a bottleneck in reporting observations. In recent years, a rapid evolution of artificial intelligence in image recognition has been evident in its broad applications in modern society, offering new opportunities for increasing the capabilities of coral reef monitoring. Here, we evaluated the performance of Deep Learning Convolutional Neural Networks for automated image analysis, using a global coral reef monitoring dataset. The study demonstrates the advantages of automated image analysis for coral reef monitoring in terms of error and repeatability of benthic abundance estimations, as well as cost and benefit. We found unbiased and high agreement between expert and automated observations (97%). Repeated surveys and comparisons against existing monitoring programs also show that automated estimation of benthic composition is equally robust in detecting change and ensuring the continuity of existing monitoring data. Using this automated approach, data analysis and reporting can be accelerated by at least 200x and at a fraction of the cost (1%). Combining commonly used underwater imagery in monitoring with automated image annotation can dramatically improve how we measure and monitor coral reefs worldwide, particularly in terms of allocating limited resources, rapid reporting and data integration within and across management areas.

Bayesian Synthetic Likelihood
L. F. Price, Christopher Drovandi, Anthony Lee, David J. Nott
2017· Journal of Computational and Graphical Statistics203doi:10.1080/10618600.2017.1302882

L. F. Pricea* http://orcid.org/0000-0002-5646-2963, C. C. Drovandia http://orcid.org/0000-0001-9222-8763, A. Leeb http://orcid.org/0000-0001-7765-0616 & D. J. Nottca School of Mathematical Sciences, Queensland University of Technology, Australia and Australian Research Council Centre of Excellence for Mathematical and Statistical Frontiers (ACEMS)b Department of Statistics, University of Warwick, Coventry, UKc Department of Statistics and Applied Probability, National University of Singapore, SingaporeCONTACT L. F. Price leah.south@hdr.qut.edu.au School of Mathematical Sciences, Queensland University of Technology, Brisbane City, QLD 4000, Australia; and Australian Research Council Centre of Excellence for Mathematical and Statistical Frontiers (ACEMS)Color versions of one or more of the figures in the article can be found online at www.tandfonline.com/r/JCGS.Supplementary materials for this article are available online. Please go to www.tandfonline.com/r/JCGS.ABSTRACTHaving the ability to work with complex models can be highly beneficial. However, complex models often have intractable likelihoods, so methods that involve evaluation of the likelihood function are infeasible. In these situations, the benefits of working with likelihood-free methods become apparent. Likelihood-free methods, such as parametric Bayesian indirect likelihood that uses the likelihood of an alternative parametric auxiliary model, have been explored throughout the literature as a viable alternative when the model of interest is complex. One of these methods is called the synthetic likelihood (SL), which uses a multivariate normal approximation of the distribution of a set of summary statistics. This article explores the accuracy and computational efficiency of the Bayesian version of the synthetic likelihood (BSL) approach in comparison to a competitor known as approximate Bayesian computation (ABC) and its sensitivity to its tuning parameters and assumptions. We relate BSL to pseudo-marginal methods and propose to use an alternative SL that uses an unbiased estimator of the SL, when the summary statistics have a multivariate normal distribution. Several applications of varying complexity are considered to illustrate the findings of this article. Supplemental materials are available online. Computer code for implementing the methods on all examples is available at https://github.com/cdrovandi/Bayesian-Synthetic-Likelihood.

Variability in cardiac electrophysiology: Using experimentally-calibrated populations of models to move beyond the single virtual physiological human paradigm
Anna Muszkiewicz, Oliver J. Britton, Philip Gemmell, Elisa Passini +4 more
2015· Progress in Biophysics and Molecular Biology192doi:10.1016/j.pbiomolbio.2015.12.002

Physiological variability manifests itself via differences in physiological function between individuals of the same species, and has crucial implications in disease progression and treatment. Despite its importance, physiological variability has traditionally been ignored in experimental and computational investigations due to averaging over samples from multiple individuals. Recently, modelling frameworks have been devised for studying mechanisms underlying physiological variability in cardiac electrophysiology and pro-arrhythmic risk under a variety of conditions and for several animal species as well as human. One such methodology exploits populations of cardiac cell models constrained with experimental data, or experimentally-calibrated populations of models. In this review, we outline the considerations behind constructing an experimentally-calibrated population of models and review the studies that have employed this approach to investigate variability in cardiac electrophysiology in physiological and pathological conditions, as well as under drug action. We also describe the methodology and compare it with alternative approaches for studying variability in cardiac electrophysiology, including cell-specific modelling approaches, sensitivity-analysis based methods, and populations-of-models frameworks that do not consider the experimental calibration step. We conclude with an outlook for the future, predicting the potential of new methodologies for patient-specific modelling extending beyond the single virtual physiological human paradigm.

GHOST: Recovering Historical Signal from Heterotachously Evolved Sequence Alignments
Stephen Crotty, Bùi Quang Minh, Nigel Bean, Barbara R. Holland +3 more
2019· Systematic Biology192doi:10.1093/sysbio/syz051

Molecular sequence data that have evolved under the influence of heterotachous evolutionary processes are known to mislead phylogenetic inference. We introduce the General Heterogeneous evolution On a Single Topology (GHOST) model of sequence evolution, implemented under a maximum-likelihood framework in the phylogenetic program IQ-TREE (http://www.iqtree.org). Simulations show that using the GHOST model, IQ-TREE can accurately recover the tree topology, branch lengths, and substitution model parameters from heterotachously evolved sequences. We investigate the performance of the GHOST model on empirical data by sampling phylogenomic alignments of varying lengths from a plastome alignment. We then carry out inference under the GHOST model on a phylogenomic data set composed of 248 genes from 16 taxa, where we find the GHOST model concurs with the currently accepted view, placing turtles as a sister lineage of archosaurs, in contrast to results obtained using traditional variable rates-across-sites models. Finally, we apply the model to a data set composed of a sodium channel gene of 11 fish taxa, finding that the GHOST model is able to elucidate a subtle component of the historical signal, linked to the previously established convergent evolution of the electric organ in two geographically distinct lineages of electric fish. We compare inference under the GHOST model to partitioning by codon position and show that, owing to the minimization of model constraints, the GHOST model offers unique biological insights when applied to empirical data.

Muscle networks: Connectivity analysis of EMG activity during postural control
Tjeerd W. Boonstra, Alessander Danna‐dos‐Santos, Hong-Bo Xie, Melvyn Roerdink +2 more
2015· Scientific Reports184doi:10.1038/srep17830

Understanding the mechanisms that reduce the many degrees of freedom in the musculoskeletal system remains an outstanding challenge. Muscle synergies reduce the dimensionality and hence simplify the control problem. How this is achieved is not yet known. Here we use network theory to assess the coordination between multiple muscles and to elucidate the neural implementation of muscle synergies. We performed connectivity analysis of surface EMG from ten leg muscles to extract the muscle networks while human participants were standing upright in four different conditions. We observed widespread connectivity between muscles at multiple distinct frequency bands. The network topology differed significantly between frequencies and between conditions. These findings demonstrate how muscle networks can be used to investigate the neural circuitry of motor coordination. The presence of disparate muscle networks across frequencies suggests that the neuromuscular system is organized into a multiplex network allowing for parallel and hierarchical control structures.

Ancient genome-wide DNA from France highlights the complexity of interactions between Mesolithic hunter-gatherers and Neolithic farmers
Maïté Rivollat, Choongwon Jeong, Stephan Schiffels, İşil Küçükkalıpçı +4 more
2020· Science Advances183doi:10.1126/sciadv.aaz5344

= 98) (7000-3000 BCE). Using the genetic substructure observed in European hunter-gatherers, we characterize diverse patterns of admixture in different regions, consistent with both routes of expansion. Early western European farmers show a higher proportion of distinctly western hunter-gatherer ancestry compared to central/southeastern farmers. Our data highlight the complexity of the biological interactions during the Neolithic expansion by revealing major regional variations.

DESPOT: Online POMDP Planning with Regularization
Nan Ye, Adhiraj Somani, David Hsu, Wee Sun Lee
2017· Journal of Artificial Intelligence Research180doi:10.1613/jair.5328

The partially observable Markov decision process (POMDP) provides a principled general framework for planning under uncertainty, but solving POMDPs optimally is computationally intractable, due to the "curse of dimensionality" and the "curse of history". To overcome these challenges, we introduce the Determinized Sparse Partially Observable Tree (DESPOT), a sparse approximation of the standard belief tree, for online planning under uncertainty. A DESPOT focuses online planning on a set of randomly sampled scenarios and compactly captures the "execution" of all policies under these scenarios. We show that the best policy obtained from a DESPOT is near-optimal, with a regret bound that depends on the representation size of the optimal policy. Leveraging this result, we give an anytime online planning algorithm, which searches a DESPOT for a policy that optimizes a regularized objective function. Regularization balances the estimated value of a policy under the sampled scenarios and the policy size, thus avoiding overfitting. The algorithm demonstrates strong experimental results, compared with some of the best online POMDP algorithms available. It has also been incorporated into an autonomous driving system for real-time vehicle control. The source code for the algorithm is available online.

Global CO2 emissions from dry inland waters share common drivers across ecosystems
Philipp S. Keller, Núria Catalán, Daniel von Schiller, Hans‐Peter Grossart +4 more
2020· Nature Communications179doi:10.1038/s41467-020-15929-y

Abstract Many inland waters exhibit complete or partial desiccation, or have vanished due to global change, exposing sediments to the atmosphere. Yet, data on carbon dioxide (CO 2 ) emissions from these sediments are too scarce to upscale emissions for global estimates or to understand their fundamental drivers. Here, we present the results of a global survey covering 196 dry inland waters across diverse ecosystem types and climate zones. We show that their CO 2 emissions share fundamental drivers and constitute a substantial fraction of the carbon cycled by inland waters. CO 2 emissions were consistent across ecosystem types and climate zones, with local characteristics explaining much of the variability. Accounting for such emissions increases global estimates of carbon emissions from inland waters by 6% (~0.12 Pg C y −1 ). Our results indicate that emissions from dry inland waters represent a significant and likely increasing component of the inland waters carbon cycle.

Dynamic changes in genomic and social structures in third millennium BCE central Europe
Luka Papac, Michal Ernée, Miroslav Dobeš, Michaela Langová +4 more
2021· Science Advances169doi:10.1126/sciadv.abi6941

Europe's prehistory oversaw dynamic and complex interactions of diverse societies, hitherto unexplored at detailed regional scales. Studying 271 human genomes dated ~4900 to 1600 BCE from the European heartland, Bohemia, we reveal unprecedented genetic changes and social processes. Major migrations preceded the arrival of "steppe" ancestry, and at ~2800 BCE, three genetically and culturally differentiated groups coexisted. Corded Ware appeared by 2900 BCE, were initially genetically diverse, did not derive all steppe ancestry from known Yamnaya, and assimilated females of diverse backgrounds. Both Corded Ware and Bell Beaker groups underwent dynamic changes, involving sharp reductions and complete replacements of Y-chromosomal diversity at ~2600 and ~2400 BCE, respectively, the latter accompanied by increased Neolithic-like ancestry. The Bronze Age saw new social organization emerge amid a ≥40% population turnover.

Interventions to help coral reefs under global change—A complex decision challenge
Kenneth R. N. Anthony, Kate J. Helmstedt, Line K. Bay, Pedro Fidelman +4 more
2020· PLoS ONE156doi:10.1371/journal.pone.0236399

Climate change is impacting coral reefs now. Recent pan-tropical bleaching events driven by unprecedented global heat waves have shifted the playing field for coral reef management and policy. While best-practice conventional management remains essential, it may no longer be enough to sustain coral reefs under continued climate change. Nor will climate change mitigation be sufficient on its own. Committed warming and projected reef decline means solutions must involve a portfolio of mitigation, best-practice conventional management and coordinated restoration and adaptation measures involving new and perhaps radical interventions, including local and regional cooling and shading, assisted coral evolution, assisted gene flow, and measures to support and enhance coral recruitment. We propose that proactive research and development to expand the reef management toolbox fast but safely, combined with expedient trialling of promising interventions is now urgently needed, whatever emissions trajectory the world follows. We discuss the challenges and opportunities of embracing new interventions in a race against time, including their risks and uncertainties. Ultimately, solutions to the climate challenge for coral reefs will require consideration of what society wants, what can be achieved technically and economically, and what opportunities we have for action in a rapidly closing window. Finding solutions that work for coral reefs and people will require exceptional levels of coordination of science, management and policy, and open engagement with society. It will also require compromise, because reefs will change under climate change despite our best interventions. We argue that being clear about society's priorities, and understanding both the opportunities and risks that come with an expanded toolset, can help us make the most of a challenging situation. We offer a conceptual model to help reef managers frame decision problems and objectives, and to guide effective strategy choices in the face of complexity and uncertainty.

Spatial resilience of the Great Barrier Reef under cumulative disturbance impacts
Camille Mellin, Samuel A. Matthews, Kenneth R. N. Anthony, Stuart C. Brown +4 more
2019· Global Change Biology145doi:10.1111/gcb.14625

Abstract In the face of increasing cumulative effects from human and natural disturbances, sustaining coral reefs will require a deeper understanding of the drivers of coral resilience in space and time. Here we develop a high‐resolution, spatially explicit model of coral dynamics on Australia's Great Barrier Reef (GBR). Our model accounts for biological, ecological and environmental processes, as well as spatial variation in water quality and the cumulative effects of coral diseases, bleaching, outbreaks of crown‐of‐thorns starfish ( Acanthaster cf. solaris ), and tropical cyclones. Our projections reconstruct coral cover trajectories between 1996 and 2017 over a total reef area of 14,780 km 2 , predicting a mean annual coral loss of −0.67%/year mostly due to the impact of cyclones, followed by starfish outbreaks and coral bleaching. Coral growth rate was the highest for outer shelf coral communities characterized by digitate and tabulate Acropora spp. and exposed to low seasonal variations in salinity and sea surface temperature, and the lowest for inner‐shelf communities exposed to reduced water quality. We show that coral resilience (defined as the net effect of resistance and recovery following disturbance) was negatively related to the frequency of river plume conditions, and to reef accessibility to a lesser extent. Surprisingly, reef resilience was substantially lower within no‐take marine protected areas, however this difference was mostly driven by the effect of water quality. Our model provides a new validated, spatially explicit platform for identifying the reefs that face the greatest risk of biodiversity loss, and those that have the highest chances to persist under increasing disturbance regimes.

A framework for automated anomaly detection in high frequency water-quality data from in situ sensors
Catherine Leigh, Omar Alsibai, Rob J. Hyndman, Sevvandi Kandanaarachchi +4 more
2019· The Science of The Total Environment133doi:10.1016/j.scitotenv.2019.02.085

Monitoring the water quality of rivers is increasingly conducted using automated in situ sensors, enabling timelier identification of unexpected values or trends. However, the data are confounded by anomalies caused by technical issues, for which the volume and velocity of data preclude manual detection. We present a framework for automated anomaly detection in high-frequency water-quality data from in situ sensors, using turbidity, conductivity and river level data collected from rivers flowing into the Great Barrier Reef. After identifying end-user needs and defining anomalies, we ranked anomaly importance and selected suitable detection methods. High priority anomalies included sudden isolated spikes and level shifts, most of which were classified correctly by regression-based methods such as autoregressive integrated moving average models. However, incorporation of multiple water-quality variables as covariates reduced performance due to complex relationships among variables. Classifications of drift and periods of anomalously low or high variability were more often correct when we applied mitigation, which replaces anomalous measurements with forecasts for further forecasting, but this inflated false positive rates. Feature-based methods also performed well on high priority anomalies and were similarly less proficient at detecting lower priority anomalies, resulting in high false negative rates. Unlike regression-based methods, however, all feature-based methods produced low false positive rates and have the benefit of not requiring training or optimization. Rule-based methods successfully detected a subset of lower priority anomalies, specifically impossible values and missing observations. We therefore suggest that a combination of methods will provide optimal performance in terms of correct anomaly detection, whilst minimizing false detection rates. Furthermore, our framework emphasizes the importance of communication between end-users and anomaly detection developers for optimal outcomes with respect to both detection performance and end-user application. To this end, our framework has high transferability to other types of high frequency time-series data and anomaly detection applications.

Transferring biodiversity models for conservation: Opportunities and challenges
Ana M. M. Sequeira, Phil J. Bouchet, Katherine L. Yates, Kerrie Mengersen +1 more
2018· Methods in Ecology and Evolution132doi:10.1111/2041-210x.12998

Abstract After decades of extensive surveying, knowledge of the global distribution of species still remains inadequate for many purposes. In the short to medium term, such knowledge is unlikely to improve greatly given the often prohibitive costs of surveying and the typically limited resources available. By forecasting biodiversity patterns in time and space, predictive models can help fill critical knowledge gaps and prioritise research to support better conservation and management. The ability of a model to predict biodiversity metrics in novel environments is termed “transferability,” and models with high transferability will be the most useful in this context. Despite their potentially broad utility, little guidance exists on what confers high transferability to biodiversity models. We synthesise recent advances in biodiversity model transfers to facilitate increased understanding of what underpins successful model transferability, demonstrating that a consistent approach has so far been lacking but is essential for achieving high levels of repeatability, transparency and accountability of model transfers. We provide a set of guidelines to support efficient learning and the improvement of model transferability.

Ten millennia of hepatitis B virus evolution
Arthur Kocher, Luka Papac, Rodrigo Barquera, Felix M. Key +4 more
2021· Science130doi:10.1126/science.abi5658

Hepatitis B virus (HBV) has been infecting humans for millennia and remains a global health problem, but its past diversity and dispersal routes are largely unknown. We generated HBV genomic data from 137 Eurasians and Native Americans dated between ~10,500 and ~400 years ago. We date the most recent common ancestor of all HBV lineages to between ~20,000 and 12,000 years ago, with the virus present in European and South American hunter-gatherers during the early Holocene. After the European Neolithic transition, Mesolithic HBV strains were replaced by a lineage likely disseminated by early farmers that prevailed throughout western Eurasia for ~4000 years, declining around the end of the 2nd millennium BCE. The only remnant of this prehistoric HBV diversity is the rare genotype G, which appears to have reemerged during the HIV pandemic.