exametrika
A test-theory package that grew into a graphical-model toolkit, now spending its releases paying down the API debt that growth created.
A side-by-side editorial comparison of eratosthenes and Statsig — release velocity, themes, recent moves, and the top alternatives to consider.
eratosthenes spends 0.1.0 hardening inputs rather than adding chronology methods.
eratosthenes does Bayesian estimation of archaeological chronologies from relative sequences, absolute constraints and artifact assemblages. The 0.0.9 line built out the inference diagnostics — traceplots, histograms, batch-means MCSE reporting, displacement estimation — and then consolidated artifact probability-density estimation into a single gibbs_ad_type(). The 0.1.0 tag turns outward instead, adding validators for every user-supplied structure and replacing seq_check() with a more informative seq_diag().
Statsig is packaging its own workflows as skills for someone else's agent to run.
Statsig's experimentation and feature-flag platform opened a public agent-skills repository holding reusable workflows — create dashboard, create cloud metric — written as instructions an AI agent can execute. Its MCP server extended to Segments and Layers, covering user targeting and experiment configuration, and Metrics Explorer gained the ability to abort long-running queries. The crawled feed captures marketing pages as entries, so several rows carry navigation boilerplate rather than release notes.
eratosthenes does Bayesian estimation of archaeological chronologies from relative sequences, absolute constraints and artifact assemblages. The 0.0.9 line built out the inference diagnostics — traceplots, histograms, batch-means MCSE reporting, displacement estimation — and then consolidated artifact probability-density estimation into a single gibbs_ad_type(). The 0.1.0 tag turns outward instead, adding validators for every user-supplied structure and replacing seq_check() with a more informative seq_diag().
The package is moving from research code to something a non-author can run. Consolidating estimation behind one function, then wrapping every input class in a validator, are the two steps that make failures legible instead of cryptic, and the diagnostics added earlier serve the same end for the sampler itself. Nothing in the window changes the underlying model; the work is all about making it usable and its output checkable.
With inputs validated and diagnostics in place, the next release is more likely to extend the constraint or assemblage modelling than to keep reworking the interface, though the feed's three sparse tags give little to read a cadence from.
Statsig's experimentation and feature-flag platform opened a public agent-skills repository holding reusable workflows — create dashboard, create cloud metric — written as instructions an AI agent can execute. Its MCP server extended to Segments and Layers, covering user targeting and experiment configuration, and Metrics Explorer gained the ability to abort long-running queries. The crawled feed captures marketing pages as entries, so several rows carry navigation boilerplate rather than release notes.
The direction is agent-operated experimentation. MCP covering segments and layers means an agent can configure targeting, and the skills repository means the multi-step workflows around that configuration are distributable artifacts rather than documentation. Statsig is betting the console is not where experiments get set up much longer.
Expect the skills catalog to grow toward the analysis side — reading experiment results and proposing rollout decisions, not just creating objects. The query-abort work suggests warehouse cost control is a live concern as agent-driven usage increases.
Other Infra & APIs products tracked by Sparkpulse, ranked by recent ship velocity. Each card links to a full editorial trajectory and lets you pivot into a head-to-head comparison with either eratosthenes or Statsig.
A test-theory package that grew into a graphical-model toolkit, now spending its releases paying down the API debt that growth created.
nuggets keeps compounding on the 2.0 rewrite — more pattern families, lighter install.
projoint spent a year on CRAN paperwork, then shipped a correctness fix it flagged itself.
dqcheckr adds drift analysis, then removes the YAML a user had to hand-write.
An actuarial mainstay spends its releases on CI plumbing, not on new mathematics.
EDAForge is a data-quality auditor renamed mid-flight, still finding its CRAN footing.
See all eratosthenes alternatives → · See all Statsig alternatives →
Latest ship moves from both products, interleaved chronologically. ⚡ = editorial spark.
They serve adjacent needs but don't currently overlap on shipped themes. eratosthenes is currently shipping more aggressively (velocity 2.5 vs 0.0), with 0 editorial sparks in the last 30 days against 0. See the at-a-glance table above for a side-by-side breakdown of velocity, recent sparks, and editorial themes.
Sparkpulse doesn't pick a winner — we score release velocity, not feature parity. eratosthenes is currently shipping more aggressively (velocity 2.5 vs 0.0), with 0 editorial sparks in the last 30 days against 0. For your specific use case, the alternatives sections above list other Infra & APIs products to evaluate alongside.
Top eratosthenes alternatives in Infra & APIs are ranked by recent ship velocity. Browse the "eratosthenes alternatives" section above for the current picks, or visit /alternatives/eratosthenes for the full list with editorial commentary on each.
Top Statsig alternatives in Infra & APIs are ranked by recent ship velocity. Browse the "Statsig alternatives" section above for the current picks, or visit /alternatives/statsig for the full list with editorial commentary on each.