alphaXiv

The Griffin Institute

12 Aug 2025

CARES: Collaborative Agentic Reasoning for Error Detection in Surgery

National University of Singapore (NUS)University College London (UCL)The First Affiliated Hospital of Guangzhou Medical University The Griffin Institute Gloucestershire Hospitals NHS Foundation Trust

CARES is a zero-shot, multi-agent framework designed for multi-class error detection in robotic-assisted surgery, leveraging a risk-stratified approach and clinically-informed Chain-of-Thought (CoT) prompting. The system achieved an average of 54.3 mF1 on the RARP dataset and 52.0 mF1 on the MERP dataset, outperforming baseline Vision-Language Models by up to 14% in mF1 score and remaining competitive with supervised models without requiring prior training.

29 Nov 2024

computer-science artificial-intelligence computer-vision-and-pattern-recognition

SEDMamba: Enhancing Selective State Space Modelling with Bottleneck Mechanism and Fine-to-Coarse Temporal Fusion for Efficient Error Detection in Robot-Assisted Surgery

University College London The Griffin Institute

Automated detection of surgical errors can improve robotic-assisted surgery. Despite promising progress, existing methods still face challenges in capturing rich temporal context to establish long-term dependencies while maintaining computational efficiency. In this paper, we propose a novel hierarchical model named SEDMamba, which incorporates the selective state space model (SSM) into surgical error detection, facilitating efficient long sequence modelling with linear complexity. SEDMamba enhances selective SSM with a bottleneck mechanism and fine-to-coarse temporal fusion (FCTF) to detect and temporally localize surgical errors in long videos. The bottleneck mechanism compresses and restores features within their spatial dimension, thereby reducing computational complexity. FCTF utilizes multiple dilated 1D convolutional layers to merge temporal information across diverse scale ranges, accommodating errors of varying duration. Our work also contributes the first-of-its-kind, frame-level, in-vivo surgical error dataset to support error detection in real surgical cases. Specifically, we deploy the clinically validated observational clinical human reliability assessment tool (OCHRA) to annotate the errors during suturing tasks in an open-source radical prostatectomy dataset (SAR-RARP50). Experimental results demonstrate that our SEDMamba outperforms state-of-the-art methods with at least 1.82% AUC and 3.80% AP performance gains with significantly reduced computational complexity. The corresponding error annotations, code and models are released at this https URL

There are no more papers matching your filters at the moment.

Events

Personalize Your Feed

Install Browser Extension

We're hiring

alphaXiv

Explore

State of the Art

Sign In

Labs

Feedback

Dark mode

CARES: Collaborative Agentic Reasoning for Error Detection in Surgery

SEDMamba: Enhancing Selective State Space Modelling with Bottleneck Mechanism and Fine-to-Coarse Temporal Fusion for Efficient Error Detection in Robot-Assisted Surgery

Events

AI for Law

Personalize Your Feed