What do we gain from simplicity versus complexity in species distribution models?

Cory Merow; Matthew J. Smith; Thomas C. Edwards Jr.; Antoine Guisan; Sean M. McMahon; Signe Normand; Wilfried Thuiller; Rafael O. Wuest; Niklaus E. Zimmermann; Jane Elith

doi:10.1111/ecog.00845

What do we gain from simplicity versus complexity in species distribution models?

Ecography

By: Cory Merow, Matthew J. Smith, Thomas C. Edwards Jr., Antoine Guisan, Sean M. McMahon, Signe Normand, Wilfried Thuiller, Rafael O. Wuest, Niklaus E. Zimmermann, and Jane Elith

https://doi.org/10.1111/ecog.00845

Links

More information: Publisher Index Page (via DOI)
Open Access Version: External Repository
Download citation as: RIS | Dublin Core

Abstract

Species distribution models (SDMs) are widely used to explain and predict species ranges and environmental niches. They are most commonly constructed by inferring species' occurrence–environment relationships using statistical and machine-learning methods. The variety of methods that can be used to construct SDMs (e.g. generalized linear/additive models, tree-based models, maximum entropy, etc.), and the variety of ways that such models can be implemented, permits substantial flexibility in SDM complexity. Building models with an appropriate amount of complexity for the study objectives is critical for robust inference. We characterize complexity as the shape of the inferred occurrence–environment relationships and the number of parameters used to describe them, and search for insights into whether additional complexity is informative or superfluous. By building ‘under fit’ models, having insufficient flexibility to describe observed occurrence–environment relationships, we risk misunderstanding the factors shaping species distributions. By building ‘over fit’ models, with excessive flexibility, we risk inadvertently ascribing pattern to noise or building opaque models. However, model selection can be challenging, especially when comparing models constructed under different modeling approaches. Here we argue for a more pragmatic approach: researchers should constrain the complexity of their models based on study objective, attributes of the data, and an understanding of how these interact with the underlying biological processes. We discuss guidelines for balancing under fitting with over fitting and consequently how complexity affects decisions made during model building. Although some generalities are possible, our discussion reflects differences in opinions that favor simpler versus more complex models. We conclude that combining insights from both simple and complex SDM building approaches best advances our knowledge of current and future species ranges.

Additional publication details
Publication type	Article
Publication Subtype	Journal Article
Title	What do we gain from simplicity versus complexity in species distribution models?
Series title	Ecography
DOI	10.1111/ecog.00845
Volume	37
Issue	12
Year Published	2014
Language	English
Publisher	Wiley
Contributing office(s)	Coop Res Unit Seattle
Description	15 p.
First page	1267
Last page	1281
Google Analytic Metrics	Metrics page