FlowLens AgentOps is an evidence-driven observability and deterministic failure attribution tool for multi-agent runtimes, built on DeerFlow.

Stars

28

7-day growth

No data

Forks

0

Open issues

0

License

MIT

Last updated

2026-07-15

AI repository intelligence
FR-AI / ANALYSIS

Why it is worth attention

It provides rule-based root cause analysis with confidence scores and redacted diagnostic export, all without mutating runtime state, and supports both replay (no backend) and connected modes.

Who it is for

  • Developers building multi-agent systems with DeerFlow
  • Engineers debugging complex agent failures and dependencies
  • DevOps or SRE teams needing observability into agent runtimes
  • Researchers or educators demonstrating agent diagnostics

Use cases

  • Identifying the primary cause and contributing factors of a multi-agent run failure
  • Inspecting a normalized execution timeline across all agent phases
  • Replaying seven predefined failure scenarios without any backend or API keys
  • Exporting redacted diagnostic data for offline analysis or LLM explanation

Strengths

  • Deterministic failure attribution using rule-based evidence, with confidence and suggested actions
  • Read-only diagnostics path that never modifies checkpoints, prompts, or runtime state
  • Comprehensive privacy and security measures including redaction of secrets and bounded tool content
  • Verified performance benchmark: median 24.8 ms for 2,000 events (local Python processing)

Considerations

  • Focused on developer diagnostics in v0.1; not a production alert manager or billing platform
  • Connected mode requires a fully configured DeerFlow runtime with authentication and model providers
  • Replay data is synthetic and may not reflect real-world failure patterns

README quick start

Quick Start

Description

Observable diagnostics, failure attribution, metrics, and no-key replay for multi-agent runtimes.

Related repositories

Similar projects matched by category, topics, and programming language.

HezaoHezao
Featured
HezaoHezao GitHub avatar

poirot

Poirot is a deep research agent kernel with a middleware-first architecture, five-layer long-term memory, multi-agent orchestration, sandbox isolation, and a three-layer skill self-evolution system, backed by 2400+ tests.

AI & Machine LearningLarge Language Models
25
S40911120
Featured
S40911120 GitHub avatar

recensa

Recensa is a self-hosted web viewer that indexes Claude Code session transcripts into a local SQLite database, enabling full-text search, replay, and audit of all past agent conversations without uploading data anywhere.

AI & Machine LearningLarge Language Models
67
lopopolo
Featured
lopopolo GitHub avatar

harness-engineering

Harness Engineering is a methodology for improving coding agent outputs by carefully crafting the environment around them—providing curated context, tools, and executable constraints that encode an organization’s nonfunctional requirements and cumulative lessons.

AI & Machine LearningAI Agents
2,390