PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 21, 20260 citationsOpen Access

Trajectory Drift and Execution Validity in Multi-Step LLM Workflows

View Full Paper
VPV. P.

Key Points

  • This research aims to analyze execution trajectory behaviors in multi-step large language model workflows using a deterministic framework.
  • Developed a framework to analyze execution trajectories with deterministic measurements.
  • Utilized cross-provider data captured from OpenAI and Anthropic models.
  • Identified metrics such as drift velocity and transition stability.
  • Local continuity in execution steps remains high while structural persistence weakens over time.
  • Observed distinct transition behaviors and divergence characteristics across workflows.
  • Findings indicate that continued execution impacts performance beyond just token efficiency.

Abstract

Large language model (LLM) systems increasingly operate through iterative multi-step execution involving retries, branching, refinement, orchestration, convergence, and continuation behaviors. Existing runtime instrumentation systems primarily expose request-level telemetry such as latency, token consumption, execution traces, and workflow events, but provide limited visibility into how execution trajectories evolve over time.This paper introduces a deterministic framework for analyzing execution trajectory behavior in multi-step LLM workflows using replayable lexical and structural signals. The analysis separates continuation, drift, branching, and convergence execution behaviors using deterministic trajectory-relative measurements. Rather than evaluating semantic correctness or reasoning quality, the framework analyzes structural persistence relative to originating execution conditions.Across a controlled cross-provider corpus of replayable traces captured from OpenAI and Anthropic models, we observe a repeatable local-versus-global mismatch phenomenon: local continuity between adjacent execution steps can remain high while persistence to the originating trajectory progressively weakens. This creates measurable regimes in which execution appears locally coherent despite structural divergence over longer execution horizons.The paper further introduces deterministic runtime diagnostics including drift velocity, transition stability, branch divergence, and branch convergence using replayable structural primitives only. Results show that execution families exhibit distinguishable transition behaviors and divergence characteristics across multi-step workflows.The findings suggest that continuation in iterative LLM systems has runtime implications beyond token efficiency alone. Multi-step execution increasingly requires analysis of execution-state evolution across extended continuation horizons rather than request-level telemetry in isolation.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

V. P. (2026) studied this question.

synapsesocial.com/papers/6a0ea1c1be05d6e3efb60844https://doi.org/10.5281/zenodo.20290420
Ask AI
Helpful
Bookmark
Share
View Full Paper