-
Notifications
You must be signed in to change notification settings - Fork 0
Benchmarks Methodology
Devrajsinh Gohil edited this page Aug 30, 2026
·
1 revision
To prevent experimental bias, thermal throttling artifacts, and WAN noise, benchmarks follow strict scientific protocols.
-
Paired Difference Testing (
$N=50$ ): Each trial pairs LangGraph and AgentMesh on identical inputs. -
Interleaved ABBA Order: Trial execution order alternates (
A-B-B-A) to distribute thermal and background OS variance equally. - Trace Replay Mode: Replays pre-recorded real LLM responses under calibrated micro-sleeps to isolate control-plane latency from cloud API jitter.
-
Hypothesis Testing: Evaluated with two-tailed paired
$t$ -tests ($lpha = 0.001$ ) and 95% confidence intervals.\n
Getting Started
How-To Guides
- Compile a Graph
- Annotated Reducers
- Parallel Fanout
- Send() Map-Reduce
- Command() Routing
- Nested Subgraphs
- Async & Streaming
- Checkpointing & State
- Financial Swarm Example
Architecture
- System Overview
- C++ Engine Internals
- O(1) Scheduler
- Dual-Tier Graph
- Persistence & WAL
- Zero-Copy Pybind Bridge
- SOLID Design Principles
API Reference
Benchmarks
Contributing