Sheaf-Guarded Updates: Streaming Structural Verification for Evolving Agent State

저자: Jason Leonard Volk | 날짜: 2026 | URL: https://openreview.net/forum?id=6NkQDugCm7 📄 PDF


⚠️ 이 페이지의 요약·평가·해설은 생성형 AI(Claude)가 자동 생성한 2차적 분석물입니다. 논문 원문의 저작권은 원저작자에게 있으며, 정확한 내용은 원문(위 DOI·arXiv 등 출처)을 확인하세요.

라이선스: OpenReview 공개(오픈액세스)

Essence

에이전트의 내부 상태를 growing graph 위의 cellular sheaf로 모델링하고, coboundary check, H1 obstruction, Dirichlet energy, Laplacian spectra를 이용해 self-modification이 구조적 일관성을 깨는지 커밋 전에 검증하는 스트리밍 검증 시스템을 제시한다.

Motivation

Achievement

Figure 4

Coboundary norm distributions for 490 valid generaliza-

  1. 조건부 Lyapunov 안정성 증명: edge-vertex ratio와 Laplacian spectral gap에 대해 Borkar stochastic approximation을 이용한 두 개의 conditional Lyapunov stability theorem과, near-identity sheaf에 대한 restricted stability corollary를 제시했다.
  2. O(1) amortized 검증: bounded-cell 분해와 cached-assembly 가정 하에 cellular decomposition을 통해 편집당 재계산 비용을 O(1) amortized로 낮추고, V=5M vertex, 35 µs median per-edit latency, zero assembled-cohomology drift를 단일 상용 머신에서 달성했다.
  3. ProofDAG 벤치마크에서 완전 분리: exact representation 하에서 F1=1.000, 14% relative noise 하에서도 F1≥0.94의 coboundary 기반 판별 성능을 보였다.
  4. structural aliasing 규명: 990개 lemma·10개 수학 도메인에 걸친 live LLM 제안 연구를 통해 자연어 loading의 조악함이 valid/contradictory coboundary 분포를 겹치게 만드는 structural aliasing을 유발함을 보이고, loader fidelity가 자율 커밋 결정의 binding constraint임을 규명했다.
  5. Lean/mathlib 검증: 507개 declaration, 954개 kernel-certified dependency edge, 6개 namespace group에서 exact-loader 경로가 zero clean residual(max 2.61×10⁻¹⁶)과 62개 controlled corruption 전량 탐지를 달성했다.

How

Originality

Limitation & Further Study

Evaluation

Novelty: 4/5 Technical Soundness: 4/5 Significance: 4/5 Clarity: 4/5 Overall: 4/5

총평: Sheaf 이론을 growing graph 기반 에이전트 상태 검증에 창의적으로 결합하고 이론적 안정성 증명과 대규모 실증(5M vertex, Lean/mathlib 실증)을 모두 갖춘 견고한 연구이나, semantic truth와 structural coherence의 간극 및 loader fidelity 문제라는 실질적 한계를 스스로 명확히 드러낸 균형 잡힌 워크숍 논문이다.

같이 보면 좋은 논문

기반 연구SPECTER2 유사도 0.92로 LLM Reasoning and Safety Benchmarks와 LLM Benchmarking and Agent Evaluation가 맞닿아, 'How Claude Code is used in practice \ Anthropic'가 이 ICML 2026 논문의 배경·대안·응용 맥락을 보완한다.
기반 연구SPECTER2 유사도 0.91로 LLM Reasoning and Safety Benchmarks와 Formal Methods and Computational Reasoning가 맞닿아, 'Accelerating Scientific Research with Gemini: Case Studies and Common Techniques'가 이 ICML 2026 논문의 배경·대안·응용 맥락을 보완한다.
기반 연구SPECTER2 유사도 0.91로 LLM Reasoning and Safety Benchmarks와 Formal Methods and Computational Reasoning가 맞닿아, 'MerLean: An Agentic Framework for Autoformalization in Quantum Computation'가 이 ICML 2026 논문의 배경·대안·응용 맥락을 보완한다.
다른 접근streaming 환경에서의 구조적 검증 문제를 다른 방식으로 다루는 것으로 보인다.
다른 접근LLM 추론 및 에이전트 구조 분석과 관련된 SHAPE와 유사한 방법론적 관심사를 공유한다.
다른 접근formal specification 기반 agent 검증이라는 공통 목표를 가진 유사 연구이다.
다른 접근LLM 추론/에이전트 구조를 분석·검증하는 관점에서 상호 연관된 연구이다.
다른 접근agent 상태 검증/명세를 다루는 SEVerA와 유사한 formal verification 접근법을 공유한다.
다른 접근에이전트 신뢰성 검증에 대한 다른 구조적 접근법을 제시하는 것으로 추정된다.
후속 연구coboundary 검사 및 spectral gap 개념을 확장하는 연구로 추정된다.
← 목록으로 돌아가기

🎧 Audio Overview

이 논문 리뷰를 팟캐스트형 오디오로 생성합니다. (Gemini · 키는 브라우저에만 저장 · 완성본은 이메일로도 전송)
▸ 고급: 구성 방향(대본 작성 지침) 직접 수정
속도 1.0x
⬇ MP3 다운로드