OxAlpha Agent Latency & Throughput Simulator Z.ai GLM Multi-Step Inference

Effective Speedup
3.77x
OxAlpha Runtime
39.4s
Standard Runtime
148.6s
Prefill Latency Saved
64.2s
Peak VRAM Saved
14.8 GB
Step 1 / 12: AST Parsing & Symbol Resolution
Cumulative Context: 2,350 tokens
OxAlpha Step Latency: 3.12s (Std: 11.8s)

Multi-Step Execution Pipeline Log

Enjoy this tool? Build your own with Super