• 学术搜索
  • 科研智能体
    • Research Labs
    • AI 阅读
    • AI 文库
    • 深度研究
    • 学者亮点
  • 学术资源
    • AI2000
    • 期刊/会议
    • 学者库
    • 学术API
    • 溯源树
    • 数据集
  • 知识沉淀
    • 学术空间
订阅小程序
旧版功能
aminer vip
开通会员低至0.73元/天
一次搞定AI科研
立即登录
  • English
  • 联系方式
    直

    直觉手术

    Intuitive Surgical Inc.
    企业
    281论文总数
    9,255引用总数

    论文量&引用量时间轴

    机构学者

    排序
    Jonathan M. Sorger
    Jonathan M. Sorger
    Department of Biomedical Engineering, Johns Hopkins University School of Medicine
    论文:21引用:0H-index:0
    Omid Mohareri
    Omid Mohareri
    Department of Electrical and Computer Engineering, University of British Columbia
    论文:19引用:0H-index:0
    Adam Schmidt
    Adam Schmidt
    Institute of Control and Information Engineering, Poznań University of Technology
    论文:11引用:0H-index:0
    Max Allan
    Max Allan
    Intuitive Surgical Inc
    论文:10引用:0H-index:0
    Anthony Jarc
    Anthony Jarc
    Intuitive Surgical Inc.
    论文:10引用:0H-index:0
    Simon DiMaio
    Simon DiMaio
    Intuitive Surgical, Inc.
    论文:8引用:0H-index:0
    Danail Stoyanov
    Danail Stoyanov
    Surgical Robot Vision Research Group, Department of Computer Science, Faculty of Engineering Sciences, University College London
    论文:8引用:0H-index:0
    Azizian Mahdi
    Azizian Mahdi
    Intuitive Surgical Inc
    论文:8引用:0H-index:0
    Tim Salcudean
    Tim Salcudean
    Department of Electrical and Computer Engineering, Faculty of Applied Science, University of British Columbia
    论文:6引用:0H-index:0

    论文(281)

    年份
    起
    –
    止
    排序
    1Qualitative Perspectives of Early Surgeon Users on the Value of the Davinci 5 Surgical System
    Derek J. Erstad, Zahra A. Fazal, Karlis Draulis,Feibi Zheng, Gretchen Jackson, Christy Y. Chai

    To understand the value and experience of using the da Vinci 5 (dV-5) robotic surgical system among early-adopting surgeons in the United States. In March 2024, da Vinci 5 was released with new features such as a haptic technology (called Force Feedback) and Case Insights, a tool leveraging artificial intelligence (AI) to deliver video recordings of cases with objective metrics of performance. Few studies have assessed the value and challenges experienced by surgeons using this new system. Twenty-three semi-structured qualitative interviews were completed with surgeon-participants over video conferencing software representing a selection of surgical specialties, case volumes, and practice types. Interviews were recorded, transcribed verbatim, and deidentified. Results were analyzed by one reviewer using an inductive-deductive thematic approach and further verified by another reviewer. Themes were mapped to value domains and challenges with adoption. Among the participants, there were a higher proportion of males, surgeons who practiced at community hospitals, and those with medium to high volumes of robotic cases. Thematic analysis revealed two main themes with seven subthemes exploring either the value beliefs or barriers/challenges with adoption to the new system. Participants found the most value in dV-5 with its ergonomic comfort, ability to support future training of surgeons, and an economic benefit in reducing operative time. There were mixed findings around its impact for improving clinical outcomes given the early system maturity. Thematic analysis of interviews with early-adopting surgeons indicated that the dV5 system supports better ergonomics, surgeon training and may have implications for some key surgical metrics such as tissue tearing and operative time.

    2026Journal of Robotic Surgery(2026)引用:24
    引用
    AI阅读
    加入学术空间
    2SurgLaVi: Large-Scale Hierarchical Dataset for Surgical Vision-Language Representation Learning
    Alejandra Perez, Chinedu Nwoye, Ramtin Raji Kermani,Omid Mohareri, Muhammad Abdullah Jamal

    Vision-language pre-training (VLP) offers unique advantages for surgery by aligning language with surgical videos, enabling workflow understanding and transfer across tasks without relying on expert-labeled datasets. However, progress in surgical VLP remains constrained by the limited scale, procedural diversity, semantic quality, and hierarchical structure of existing datasets. In this work, we present SurgLaVi, the largest and most diverse surgical vision-language dataset to date, comprising nearly 240k clip-caption pairs from more than 200 procedures, and featuring hierarchical levels at coarse-, mid-, and fine-level. At the core of SurgLaVi lies a fully automated pipeline that systematically generates fine-grained transcriptions of surgical videos and segments them into coherent procedural units. To ensure high-quality annotations, it applies dual-modality filtering to remove irrelevant and noisy samples. Within this framework, the resulting captions are enriched with contextual detail, producing annotations that are both semantically rich and easy to interpret. To ensure accessibility, we release SurgLaVi-β, an open-source derivative of 113k clip-caption pairs constructed entirely from public data, which is over four times larger than existing surgical VLP datasets. To demonstrate the value of the SurgLaVi datasets, we introduce SurgCLIP, a CLIP-style video-text contrastive framework with dual encoders, as a representative base model. SurgCLIP achieves consistent improvements across phase, step, action, and tool recognition, surpassing prior state-of-the-art methods, often by large margins. These results validate that large-scale, semantically rich, and hierarchically structured datasets directly translate into stronger and more generalizable representations, establishing SurgLaVi as a key resource for developing surgical foundation models.

    2026Medical image analysis(2026)引用:23
    引用
    AI阅读
    加入学术空间
    3Learning Where to Look: Scaling Parkland Grade Prediction from Surgical Videos
    Sreeram Kamabattula, Sue Kulason, Busisiwe Mlambo, Lilia Purvis, Kiran Bhattacharyya

    The Parkland Grading Scale (PGS) is widely used to quantify operative difficulty in cholecystectomy, with higher grades associated with worse post-operative outcomes. However, consistent, scalable PGS assessment is limited by the reliance on two manual steps: determining where to look in the surgical video for key evidence, and assigning a grade. Previous machine learning approaches have either depended on manual selection of where to look, or approximated it with fixed-duration video segments, leaving it unclear whether models can accurately predict PGS without explicit guidance on where to look. To address this, we evaluate 287 robotic cholecystectomy videos annotated with PGS and a standardized key-segment. Using a temporal convolution network and attention-based framework, we compare the performance of a fully automated model using full surgical videos without key-segment supervision to a model provided with the key-segment (where to look). Providing the key-segment yields substantial performance gains (weighted F1 +0.25 and Krippendorff’s α (KA) +0.29). We further introduce ParkNet _LEARN , which learns to where to look and predicts PGS from full surgical videos, achieving significant improvements over the no-supervision automation (weighted F1 +0.18 and KA +0.23), and a KA = 0.60–within 0.06 of the model with key-segment provided. These findings highlight the importance of attending to where to look for automating operative difficulty assessment, and is a valuable step toward supporting large-scale research on surgical performance and post-operative outcomes.

    2026International Journal of Computer Assisted Radiology and Surgery(2026)引用:9
    引用
    AI阅读
    加入学术空间
    4SUREON: A Benchmark and Vision-Language-Model for Surgical Reasoning
    Alejandra Perez, Anita Rau, Lee White, Busisiwe Mlambo, Chinedu Nwoye, Muhammad Abdullah Jamal,Omid Mohareri

    Surgeons don't just see – they interpret. When an expert observes a surgical scene, they understand not only what instrument is being used, but why it was chosen, what risk it poses, and what comes next. Current surgical AI cannot answer such questions, largely because training data that explicitly encodes surgical reasoning is immensely difficult to annotate at scale. Yet surgical video lectures already contain exactly this – explanations of intent, rationale, and anticipation, narrated by experts for the purpose of teaching. Though inherently noisy and unstructured, these narrations encode the reasoning that surgical AI currently lacks. We introduce SUREON, a large-scale video QA dataset that systematically harvests this training signal from surgical academic videos. SUREON defines 12 question categories covering safety assessment, decision rationale, and forecasting, and uses a multi-agent pipeline to extract and structure supervision at scale. Across 134.7K clips and 170 procedure types, SUREON yields 206.8k QA pairs and an expert-validated benchmark of 354 examples. To evaluate the extent to which this supervision translates to surgical reasoning ability, we introduce two models: SureonVLM, a vision-language model adapted through supervised fine-tuning, and SureonVLM-R1, a reasoning model trained with Group Relative Policy Optimization. Both models can answer complex questions about surgery and substantially outperform larger general-domain models, exceeding 84

    2026引用:6
    引用
    AI阅读
    加入学术空间
    5Semantic Taxonomy-Driven Instrument Classification Streamlines Kinematic Analysis of Objective Performance Indicators in Robotic Surgery
    Mattia Ballo, Elizabeth W. Tindal, Jeffrey Nussbaum, Vikrom Dhar, Valery Dronsky,Rebecca Kowalski, Rachel Webman, Andrew Yee,Filippo Filicori

    Objective performance indicators (OPIs) derived from robotic surgery are showing potential for automated skill assessment, but their high dimensionality, data sparsity, and lack of functional context limit their clinical utility and interpretability. Here, we introduce and validate a semantic taxonomy that automatically classifies surgical instruments into functional roles—such as ‘Dominant’, ‘Active Retractor’, and ‘Passive Retractor’—based on their kinematic signatures. Applied to 462 cholecystectomies, hernia repairs, and sleeve gastrectomies, this framework drastically reduced data dimensionality. In predictive modeling for surgical experience and task efficiency, taxonomy-structured OPIs achieved superior performance to conventional metrics while requiring substantially fewer features to reach optimal results (mean, 12.4 vs. 19.5; P = 0.025). By providing functional context, this approach streamlines kinematic analysis, creating a more scalable and interpretable foundation for objective skill assessment, actionable feedback, and data-driven surgical training, ultimately enhancing surgical quality and safety.

    2026npj Digital Surgery(2026)引用:2
    引用
    AI阅读
    加入学术空间
    立即登录,查看全部 281 篇论文

    合作机构(100)

    斯坦福大学合作论文 24
    约翰斯·霍普金斯大学合作论文 24
    伦敦大学学院合作论文 9
    华盛顿大学合作论文 8
    范德比尔特大学合作论文 7
    斯特拉斯堡大学合作论文 6
    不列颠哥伦比亚大学合作论文 6
    俄亥俄州立大学合作论文 5
    慕尼黑工业大学合作论文 4
    加州理工学院合作论文 4

    机构统计