pr2database
The protist reference database keeps widening past the rRNA gene it was built on.
A side-by-side editorial comparison of mice and valr — release velocity, themes, recent moves, and the top alternatives to consider.
mice can finally predict, not just estimate, from multiply imputed data.
mice is the reference implementation of multiple imputation by chained equations, and the default answer to missing data in R. The releases here follow a consistent shape: one or two substantive additions per version, most contributed by outside authors, plus fixes to methods that have been in the package for years. The current 3.19.0 adds predict_mi(), which pools predictions across imputations under Rubin's rules and can return prediction intervals.
valr's interval verbs now read genomic files in place instead of demanding a loaded tibble.
valr reimplements bedtools-style genome interval arithmetic as tidyverse verbs backed by C++. Its long project has been closing the behavioural gap with bedtools — the book-ended interval semantics finally match in 0.10.0, three releases after the deprecation began. The July release also ends the assumption that intervals must be in memory: bed_map(), bed_intersect(), bed_subtract(), bed_coverage() and bed_window() accept a bigWig or bigBed path or URL where an interval table used to go.
mice is the reference implementation of multiple imputation by chained equations, and the default answer to missing data in R. The releases here follow a consistent shape: one or two substantive additions per version, most contributed by outside authors, plus fixes to methods that have been in the package for years. The current 3.19.0 adds predict_mi(), which pools predictions across imputations under Rubin's rules and can return prediction intervals.
Two things are happening. The imputation method catalogue keeps widening — lasso variants, multivariate PMM, categorical PMM via canonical correlation — while the pooling side is being extended past its original purpose, first to synthetic data, now to predictions on held-out sets. That second thread points at predictive modelling workflows rather than the inferential ones mice was built for. Meanwhile the maintainers keep finding consequential old bugs: the augment() ordered-factor defect in 3.18.0 had been silently degrading ordinal imputations for years.
predict_mi() is framed around evaluating predictive performance on test sets, and the ignore argument added in 3.12.0 already exists to hold out rows from the imputation model. Expect the next work to join those up into a fuller train/test story for imputed data, since the pieces are now in place but not yet connected.
valr reimplements bedtools-style genome interval arithmetic as tidyverse verbs backed by C++. Its long project has been closing the behavioural gap with bedtools — the book-ended interval semantics finally match in 0.10.0, three releases after the deprecation began. The July release also ends the assumption that intervals must be in memory: bed_map(), bed_intersect(), bed_subtract(), bed_coverage() and bed_window() accept a bigWig or bigBed path or URL where an interval table used to go.
Two arcs converge here. One is compatibility: min_overlap arrived with a deprecation warning in 0.9.0 and its default flipped from 0 to 1 in 0.10.0, so book-ended intervals are excluded by default as bedtools does, with the internal calculations in bed_closest() and friends deliberately left counting them. The other is the file-backed path, which grew out of the cpp11bigwig dependency adopted in 0.8.3 for read_bigwig() and re-exported in 0.9.0 — reading a file became querying one. Underneath, the C++ base keeps getting lighter: Rcpp swapped for cpp11, rlang cut to a single function, per-group memory copies removed from three verbs.
Only five verbs take a file argument today and bed_closest(), bed_glyph() and the statistical verbs do not, so extending the file-backed path across the rest of the API is the obvious follow-up. The deprecated tibble re-exports and the now-defunct n_fields argument suggest continued removal of the compatibility layer in the next minor release.
Other Infra & APIs products tracked by Sparkpulse, ranked by recent ship velocity. Each card links to a full editorial trajectory and lets you pivot into a head-to-head comparison with either mice or valr.
The protist reference database keeps widening past the rRNA gene it was built on.
Composable aligned layouts, rebuilt on S7 while ggplot2 4.0 lands underneath.
Conservation planning absorbs the literature's target-setting rules as code.
Joint species distribution models in Gibbs-sampled C++, quiet since 2023.
An ecosystem model starts tracking carbon isotopes and land-use change.
Ten years in, US mapping splits its data out and finally adds Puerto Rico.
Latest ship moves from both products, interleaved chronologically. ⚡ = editorial spark.
They serve adjacent needs but don't currently overlap on shipped themes. mice and valr are shipping at a similar cadence (velocity 0.0 vs 0.0, both within Sparkpulse's "active" band). See the at-a-glance table above for a side-by-side breakdown of velocity, recent sparks, and editorial themes.
Sparkpulse doesn't pick a winner — we score release velocity, not feature parity. mice and valr are shipping at a similar cadence (velocity 0.0 vs 0.0, both within Sparkpulse's "active" band). For your specific use case, the alternatives sections above list other Infra & APIs products to evaluate alongside.
Top mice alternatives in Infra & APIs are ranked by recent ship velocity. Browse the "mice alternatives" section above for the current picks, or visit /alternatives/mice for the full list with editorial commentary on each.
Top valr alternatives in Infra & APIs are ranked by recent ship velocity. Browse the "valr alternatives" section above for the current picks, or visit /alternatives/valr for the full list with editorial commentary on each.