n2kanalysis
n2kanalysis has spent eight years wiring INLA models to an S3 bucket.
A side-by-side editorial comparison of omock and spanishoddata — release velocity, themes, recent moves, and the top alternatives to consider.
A mock-data generator for OMOP studies that keeps widening what it can fake.
omock builds synthetic OMOP Common Data Model tables so packages in the darwin-eu and OHDSI ecosystem can be tested without touching patient data. The 0.7.0 release adds concept set subsetting with its own vignette, unit and value support in mockMeasurement, observation date validation and cohort attrition initialization, while deprecating mockConcepts. Releases arrive as PR digests rather than written notes.
spanishoddata spent a year finding out its 2020-2021 data was quietly incomplete.
spanishoddata provides access to Spain's open mobility origin-destination datasets from the Ministry of Transport, converting them into DuckDB and parquet for analysis at scale. Nearly every release in this window is a data-fidelity fix rather than a feature: district-to-municipal reaggregation was wrong for the 2020-2021 vintage, literal 'NA' strings in the source CSVs broke DuckDB enum casting, and the Amazon S3 metadata bucket turned out to be truncated at March 2021, silently hiding data.
omock builds synthetic OMOP Common Data Model tables so packages in the darwin-eu and OHDSI ecosystem can be tested without touching patient data. The 0.7.0 release adds concept set subsetting with its own vignette, unit and value support in mockMeasurement, observation date validation and cohort attrition initialization, while deprecating mockConcepts. Releases arrive as PR digests rather than written notes.
The package has been moving from generating tables to shipping and managing reference datasets — mockDatasets arrived in 0.4.0, mockCdmFromDataset gained a source argument in 0.5.0, and 0.6.1 added download retry handling plus an internal GiBleed dataset after the hosted files moved. The recent work is filling in CDM fidelity: type concepts, measurement units, attrition, and guards for degenerate cases like an empty person table.
Expect coverage to keep extending table by table toward the parts of the CDM that omock still cannot mock, with mockConcepts removed outright once the concept set subsetting path settles.
spanishoddata provides access to Spain's open mobility origin-destination datasets from the Ministry of Transport, converting them into DuckDB and parquet for analysis at scale. Nearly every release in this window is a data-fidelity fix rather than a feature: district-to-municipal reaggregation was wrong for the 2020-2021 vintage, literal 'NA' strings in the source CSVs broke DuckDB enum casting, and the Amazon S3 metadata bucket turned out to be truncated at March 2021, silently hiding data.
The package is in a trust-building phase. The pattern across 0.2.1 through 0.2.6 is the maintainers repeatedly discovering that upstream metadata and the package's own aggregation were misrepresenting what data existed, then fixing it and adding a check so it surfaces next time. That is now backed by infrastructure: comprehensive unit tests plus weekly live-data runs on GitHub workers that alert maintainers when the upstream ministry changes something. The last feature release sits outside the six-entry window, which is itself the story.
Expect continued upstream-tracking fixes as the ministry's API and S3 layout shift, with the experimental quick-access and checksum functions the most likely candidates for promotion to stable.
Other Analytics products tracked by Sparkpulse, ranked by recent ship velocity. Each card links to a full editorial trajectory and lets you pivot into a head-to-head comparison with either omock or spanishoddata.
n2kanalysis has spent eight years wiring INLA models to an S3 bucket.
A Fortran-descended optimizer got thread-safe, then found two flags that never worked.
ggstatsplot reached 1.0 by adding tests, having outsourced its statistics years ago.
collapse got a JSS paper and a 7x fmean speedup in the same release.
gtsummary is quietly rebuilding itself around analysis results data, one table verb at a time.
broadcast is filling in NumPy-style array broadcasting for R, operator by operator.
See all omock alternatives → · See all spanishoddata alternatives →
Latest ship moves from both products, interleaved chronologically. ⚡ = editorial spark.
They serve adjacent needs but don't currently overlap on shipped themes. omock and spanishoddata are shipping at a similar cadence (velocity 0.0 vs 0.0, both within Sparkpulse's "active" band). See the at-a-glance table above for a side-by-side breakdown of velocity, recent sparks, and editorial themes.
Sparkpulse doesn't pick a winner — we score release velocity, not feature parity. omock and spanishoddata are shipping at a similar cadence (velocity 0.0 vs 0.0, both within Sparkpulse's "active" band). For your specific use case, the alternatives sections above list other Analytics products to evaluate alongside.
Top omock alternatives in Analytics are ranked by recent ship velocity. Browse the "omock alternatives" section above for the current picks, or visit /alternatives/omock-r for the full list with editorial commentary on each.
Top spanishoddata alternatives in Analytics are ranked by recent ship velocity. Browse the "spanishoddata alternatives" section above for the current picks, or visit /alternatives/spanishoddata for the full list with editorial commentary on each.