mlr3mbo
mlr3mbo picked its defaults from a benchmark study, not from taste
A side-by-side editorial comparison of desirability2 and dials — release velocity, themes, recent moves, and the top alternatives to consider.
desirability2 is making multi-metric model selection a first-class tidymodels step.
desirability2 implements desirability functions, which map several metrics onto a common 0-1 scale so they can be combined into a single objective. The package is young: three releases, the first of which only added a NEWS file. Its substance arrived in 0.1.0 with hooks into tidymodels' tune package.
dials is quietly registering the tuning parameters for tidymodels' deep-learning push
dials defines the parameter objects and grid constructors that tidymodels tunes over, which makes its release notes a reliable early read on what the rest of the stack is about to support. The last two releases are dominated by attention-model parameters — SAINT and tabular deep learning via brulee, TabPFN via parsnip's tab_pfn() — alongside catboost parameters for bonsai and calibration parameters for tailor.
desirability2 implements desirability functions, which map several metrics onto a common 0-1 scale so they can be combined into a single objective. The package is young: three releases, the first of which only added a NEWS file. Its substance arrived in 0.1.0 with hooks into tidymodels' tune package.
The direction is integration rather than standalone use. Version 0.1.0 added select_best_desirability() and show_best_desirability() to resolve a tuning run against several metrics at once; 0.2.0 exported make_desirability_cols() so other packages can build on it and made data-driven limits the default, removing the need to state ranges by hand. Both releases move work from the user into the package.
The exported helper and the developer-facing desirability() API point to adoption by other tidymodels packages as the next step rather than new functionality here.
dials defines the parameter objects and grid constructors that tidymodels tunes over, which makes its release notes a reliable early read on what the rest of the stack is about to support. The last two releases are dominated by attention-model parameters — SAINT and tabular deep learning via brulee, TabPFN via parsnip's tab_pfn() — alongside catboost parameters for bonsai and calibration parameters for tailor.
The grid machinery itself is settled: grid_space_filling() consolidated the older designs, and the grid_*() functions now error rather than warn on the wrong size argument. What keeps moving is the parameter catalog, and it is moving toward neural and foundation-model territory that tidymodels historically left alone. Error-message quality is a steady secondary theme.
Expect further parameter objects to land ahead of the parsnip and brulee releases that use them — the attention and tabular-foundation-model work in flight is the clearest thing the entries point to.
Other Analytics products tracked by Sparkpulse, ranked by recent ship velocity. Each card links to a full editorial trajectory and lets you pivot into a head-to-head comparison with either desirability2 or dials.
mlr3mbo picked its defaults from a benchmark study, not from taste
loo keeps rewriting the diagnostics Bayesian modellers read off model comparison
mlr3fselect turned feature selection into an asynchronous, distributable job
lime survives on compatibility patches years after its research moment
mlr3measures is systematically retrofitting sample weights across every metric
mlr3cluster went from a handful of clusterers to covering the field
See all desirability2 alternatives → · See all dials alternatives →
Latest ship moves from both products, interleaved chronologically. ⚡ = editorial spark.
Both compete on the same themes — tidymodels — within Analytics. desirability2 and dials are shipping at a similar cadence (velocity 0.0 vs 0.0, both within Sparkpulse's "active" band). See the at-a-glance table above for a side-by-side breakdown of velocity, recent sparks, and editorial themes.
Sparkpulse doesn't pick a winner — we score release velocity, not feature parity. desirability2 and dials are shipping at a similar cadence (velocity 0.0 vs 0.0, both within Sparkpulse's "active" band). For your specific use case, the alternatives sections above list other Analytics products to evaluate alongside.
Top desirability2 alternatives in Analytics are ranked by recent ship velocity. Browse the "desirability2 alternatives" section above for the current picks, or visit /alternatives/desirability2 for the full list with editorial commentary on each.
Top dials alternatives in Analytics are ranked by recent ship velocity. Browse the "dials alternatives" section above for the current picks, or visit /alternatives/dials for the full list with editorial commentary on each.