LangGraph
A checkpoint-persistence maintenance train, with the tracing API still being argued over.
A side-by-side editorial comparison of Alhena AI and LiveKit Agents — release velocity, themes, recent moves, and the top alternatives to consider.
A vendor running a public benchmark on its own category, and publishing where everyone fails.
Alhena AI's feed is a research blog, not a changelog, but it is unusually structured for one: since late July it has run a single continuing study in which 15 live AI shopping agents are tested as ordinary shoppers on real storefronts. The findings are consistent and unflattering to the category - all 15 can answer questions, 9 can sell, 4 can complete a return or order change, and 1 remembers a shopper across sessions. Recent instalments break the results down by 11 verticals and by a specific task, foundation shade matching from a selfie, where five agents ignored the image entirely.
After months of vendor plugins and turn-detection fixes, LiveKit Agents ships PII redaction.
1.7.0 is the first release in this window that is not provider breadth or failure-path repair. It adds PII redaction to Agent Observability — semantic redaction of detected entities from chat history and audio recordings, with sensitive fields filtered out of logs and traces while diagnostic context survives — and renames trace attributes and log fields to support it, which breaks third-party observability queries on upgrade. The same release adds expressive mode, where a voice agent's prosody and emotion are set by emotion tags the model generates from conversation context rather than by configuration. Underneath, the usual run of turn-taking fixes continues: tool events emitted after interruption, adaptive interruption preserved across tool calls, transcripts kept when TTS returns no word timings.
Alhena AI's feed is a research blog, not a changelog, but it is unusually structured for one: since late July it has run a single continuing study in which 15 live AI shopping agents are tested as ordinary shoppers on real storefronts. The findings are consistent and unflattering to the category - all 15 can answer questions, 9 can sell, 4 can complete a return or order change, and 1 remembers a shopper across sessions. Recent instalments break the results down by 11 verticals and by a specific task, foundation shade matching from a selfie, where five agents ignored the image entirely.
The blog is building a capability ladder - Answer, Recommend, Sell, Act, Remember - and using it to argue that architecture, not category difficulty, decides where an agent stops. That framing does competitive work: it defines the axis on which agents are compared, places memory and task completion at the top, and reports that almost nothing on the market reaches them. Nothing here describes Alhena's own product releases, so the feed shows the argument the company is making rather than what it is shipping.
The benchmark series looks set to continue with further vertical and task cuts against the same 15-agent panel. A refreshed run showing movement on the Act and Remember rungs would be the natural next instalment, though these entries do not say when it is due.
1.7.0 is the first release in this window that is not provider breadth or failure-path repair. It adds PII redaction to Agent Observability — semantic redaction of detected entities from chat history and audio recordings, with sensitive fields filtered out of logs and traces while diagnostic context survives — and renames trace attributes and log fields to support it, which breaks third-party observability queries on upgrade. The same release adds expressive mode, where a voice agent's prosody and emotion are set by emotion tags the model generates from conversation context rather than by configuration. Underneath, the usual run of turn-taking fixes continues: tool events emitted after interruption, adaptive interruption preserved across tool calls, transcripts kept when TTS returns no word timings.
The train has been a breadth-plus-correctness operation — add speech and avatar vendors, then fix the ways conversations go wrong, with endpointing recurring constantly. 1.7.0 points somewhere else: at what a voice agent is allowed to record and how it is allowed to sound. Both are properties of the platform rather than of a plugin, and the trace-attribute rename shows the observability layer being treated as a product surface with its own contract. Cadence stays roughly weekly with a largely external contributor list.
Redaction policy will need to become configurable — which entity classes, retained or dropped at capture — since a single semantic default will not satisfy both debugging and compliance. Expect expressive mode to grow explicit overrides once developers find the model choosing the wrong tone.
Other ai-assistants products tracked by Sparkpulse, ranked by recent ship velocity. Each card links to a full editorial trajectory and lets you pivot into a head-to-head comparison with either Alhena AI or LiveKit Agents.
A checkpoint-persistence maintenance train, with the tracing API still being argued over.
AutoGPT's experts now get hired, fired, given private memory — and a wallet that pays merchants.
Qodo is arguing its way from AI code review up to governing the whole SDLC.
Comet writes the observability textbook while Opik quietly becomes the product.
Snorkel is building the scoreboard for agents that have to keep working, not just answer.
DataRobot keeps shipping infrastructure, then writing essays about why you need it.
See all Alhena AI alternatives → · See all LiveKit Agents alternatives →
Latest ship moves from both products, interleaved chronologically. ⚡ = editorial spark.
They serve adjacent needs but don't currently overlap on shipped themes. LiveKit Agents is currently shipping more aggressively (velocity 6.3 vs 5.0), with 1 editorial sparks in the last 30 days against 0. See the at-a-glance table above for a side-by-side breakdown of velocity, recent sparks, and editorial themes.
Sparkpulse doesn't pick a winner — we score release velocity, not feature parity. LiveKit Agents is currently shipping more aggressively (velocity 6.3 vs 5.0), with 1 editorial sparks in the last 30 days against 0. For your specific use case, the alternatives sections above list other ai-assistants products to evaluate alongside.
Top Alhena AI alternatives in ai-assistants are ranked by recent ship velocity. Browse the "Alhena AI alternatives" section above for the current picks, or visit /alternatives/alhena for the full list with editorial commentary on each.
Top LiveKit Agents alternatives in ai-assistants are ranked by recent ship velocity. Browse the "LiveKit Agents alternatives" section above for the current picks, or visit /alternatives/livekit-agents for the full list with editorial commentary on each.