Property-Level Reconstructability of Agent Decisions: An Anchor-Level Pilot Across Vendor SDK Adapter Regimes

Agentic AI failures need post-hoc reconstruction: what the agent did, on whose authority, against which policy, and from what reasoning. Cross-regime feasibility remains unmeasured under one property-level schema. We apply the Decision Trace Reconstructor unmodified to pinned worked-example anchors from six public vendor SDK regimes spanning cloud-agent, observability, tool-use, telemetry, and protocol traces, plus two comparator columns. Each Decision Event Schema (DES) property is classified as fully fillable, partially fillable, structurally unfillable, or opaque. Per-property reconstructability of an agent decision already varies between regimes at this anchor scale. Strict-governance-completeness separates into three tiers ranging from 42.9% to 85.7%, yielding one regime-independent gap (reasoning trace), four regime-dependent gaps, and one Mixed property; the pilot is single-annotator, one anchor per cell, descriptive, with outputs checksum-verifiable from a deposited reproducibility package.

Paper

References (12)

06Tracing - OpenAI Agents SDK Documentation2025 · openai.github.io/openai-agents-python/tracing/
07Incident 2025-07-19-1eb1: Replit AI agent deletes production database during code freeze2025 · OECD AI Incidents Monitor
08Vertex AI Agent Engine overview - Google Cloud Documentation2025 · Google Cloud Documentation (Tier A vendor primary doc)
09Semantic conventions for AWS Bedrock operations - OpenTelemetry. OpenTelemetry Specification (Tier A standards spec)2025 · openteleme try.io/docs/specs/semconv/gen-ai/aws-bedrock/
10Anchor-Level Reconstructability PilotZenodo
11Decision Event SchemaZenodo / GitHub
12Tool use with Claude - Claude API Documentation (Anthropic)Anthropic Documentation (Tier A vendor primary doc)

Similar papers

© 2026 NYSGPT2525 LLC