Ensemble with a stacked meta-learner
Seven learners with deliberately different inductive biases — bagged trees, boosted trees, and linear models including a MaxEnt-style formulation — are combined by a meta-learner trained on out-of-fold predictions. When a species has too little data to train the stacker honestly, it degrades to a performance-weighted average rather than overfitting the blend.