Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:26:44.013878Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 88 of 88 outbound references and 9 inbound Pith citation observations for arXiv:2506.02555.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:26:44.013878Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:16:27.620334Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
88 of 88 outbound references displayed
External citation measurements
2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation e2152a53-338f-49bd-8695-837846db1a4e · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45ff216c-6d84-423a-9fe1-6ccfa2ebe7b7 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Cholecinstanceseg: A tool instance segmen- tation dataset for laparoscopic surgery, 2024
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b138aaf-e228-4ad3-b1ac-264190f59f5d · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Flamingo: a visual language model for few-shot learning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b90920a1-3940-4c9a-b1e2-072359f503ec · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence 2017 Robotic Instrument Segmentation Challenge
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8c2fae3-b9d2-486c-a79b-252e93bc6e4a · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Pixel-wise recognition for holistic surgical scene under- standing
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3424534-8919-4178-ba63-55d565b9471b · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Surgical-vqla: Transformer with gated vision-language embedding for visual ques- tion localized-answering in robotic surgery
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81d338e4-99aa-45dd-bdaa-f976de9bf6eb · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Qwen2.5-VL Technical Report
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edd8eaff-da8d-496b-9c76-6e12bbcbb9d9 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Curriculum learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c2311f1-407c-4093-91db-99a330891050 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence De- tecting surgical tools by modelling local appearance and global shape
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc1b2c2d-3bcc-4134-90a3-a7a14aa8d16a · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Rinner, Sebastian Bo- denstedt, Alexander C
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6151acee-2918-4f6e-b934-76869327c956 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48c87e8e-92d1-41c2-b8f8-6d8ceaadcceb · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3a1cad3-6030-4dd9-a9be-c9ebd0a92195 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Med-gemma: Medical vision-language models from google deepmind
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8409f445-d0b5-4c3a-81f9-500b8304525d · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Multimodal Whole Slide Foundation Model for Pathology
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf731aed-20fa-401a-b483-f16678ab00ec · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Llm-assisted multi-teacher continual learning for visual question an- swering in robotic surgery, 2024
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ba975baa-651d-4a12-8d5d-61eb617d0b4f · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Data Filtering Networks
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8816cfd9-36a0-4356-bf22-927d8f697fe5 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Cataract-1k dataset for deep-learning-assisted analysis of cataract surgery videos
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d157c05a-fe16-4a1a-a5cc-1efea293e759 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e3635fb-bf4a-4274-8078-043c29dbf1d7 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Khan, Sophia Bano, Hani J
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e5b88ed3-f961-4ec7-ac6b-7447b5c8f650 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Lora: Low-rank adaptation of large lan- guage models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0b7c7212-552c-4c6b-87c7-fe667eed7bff · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Ophnet: A large-scale video benchmark for ophthalmic surgical workflow under- standing, 2024
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 72c15374-cf1e-4575-bb0a-cf5fc9cf838c · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence GPT-4o System Card
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e86a02e-5d78-4ae5-96ec-fce440c2db37 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Le, Yun-Hsuan Sung, Zhen Li, and Tom Duerig
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 43f07c56-a1ed-44c3-a26a-f1a430b89cc7 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Surgical visual question answering: A new frontier for interpretable computer-assisted intervention
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7f37a43d-772c-4ba9-aa06-1272b97cebfb · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Segcol challenge: Semantic segmentation for tools and fold edges in colonoscopy data, 2024
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2b5e16c2-9cbe-4a6d-bfc5-83c374d70b32 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Lavanchy, Sanat Ramesh, Diego Dall’Alba, Cris- tians Gonzalez, Paolo Fiorini, Beat P
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ddec11f9-f134-4240-b184-da74eb21ff06 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Llava- med: Training a large language-and-vision assistant for biomedicine in one day
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2e9ac32d-3cfe-4e47-9d2f-6177640ccc9f · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4f8ac4da-2b74-420d-892c-f15ddde04a2d · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Baichuan-omni-1.5 technical re- port
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ada36601-b84d-4de9-a244-940f00c0b219 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Visual instruction tuning, 2023
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7975da19-cd78-4200-b52d-6eb0741a478e · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Surgical SAM 2: Real-time Segment Anything in Surgical Video by Efficient Frame Pruning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1fb5cbd-c443-4561-849e-445d06238e5d · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Expert-level vision-language foundation model for real-world radiology and comprehensive evaluation
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fa6782a-ec65-4a5b-a889-3da3744fec18 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Radiology- llama2: Best-in-class large language model for radiol- ogy, 2023
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7d582b4a-1ddb-4e22-97d6-4a2bd4e79a67 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Surgraw: Multi-agent workflow with chain- of-thought reasoning for surgical intelligence
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b50918d9-6f48-46a4-a744-ac8b86a687b9 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence A visual-language foundation model for com- putational pathology
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation da303f06-6ec7-4a33-a44c-65f49b416f34 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence A multimodal generative ai copilot for human pathology
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0896b01f-5e43-4777-a983-fbe38842fd8c · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Nunez Do Rio, Lyn- don da Cruz, Christos Bergeles, Hongyu Chen, Fu- cang Jia, Nikhil KumarTomar, Debesh Jha, Michael A
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7268eb00-2c1f-4f08-8a34-8688a2b4645e · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Endoscapes2023, a critical view of safety and surgical scene segmentation dataset for laparoscopic cholecys- tectomy, 2024
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 85942856-c5ab-4d19-9385-ecb6ad53eda9 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence National institutes of health
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4abf143f-baee-4d37-9b79-776c1c3b260b · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Rendezvous: Attention mechanisms for the recognition of surgical action triplets in endoscopic videos
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 74b71bd7-76c1-4abb-947b-34a0ce74fa75 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Cholectrack20: A multi- perspective tracking dataset for surgical tools
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 998de0c0-c352-4ae4-9003-7e60d681112c · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Foundation models in radiology: What, how, why, and why not
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f64cea09-1498-41e3-a433-a442078a277c · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Skywork R1V: Pioneering Multimodal Reasoning with Chain-of-Thought
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 348aa646-bc5a-43c1-ae92-3b9cee713bf7 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Competence-based curriculum learning for neural ma- chine translation
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ccfb18a3-e4cb-4cea-81c8-92f8e8689768 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Sar-rarp50: Segmentation of surgical in- strumentation and action recognition on robot-assisted radical prostatectomy challenge, 2024
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9661da0e-b4a7-4ba5-94a7-2ff49fcff1f0 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Learning transferable visual models from natural lan- guage supervision
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c0be85b4-7d9c-42e6-8137-eeb4f147380d · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Rios, M.A
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f349c760-3b2a-47e3-94a3-234259a019f1 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Surgical-VQA: Visual question answering in surgical scenes using trans- former
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 64318212-7e97-408e-bde5-2478ea61c927 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Surgicalgpt: end-to-end language-vision gpt for visual question answering in surgery
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e126565e-efbe-4488-8df9-4ff007250a4e · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Think step by step: Chain-of-gesture prompting for error detection in robotic surgical videos
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 997de449-0da0-4a15-8cdc-c77b13a30317 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Gemini: A Family of Highly Capable Multimodal Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d1e4581-143d-4903-8455-9375e40445d9 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Gemma 3 Technical Report
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a28b9b2-abfa-4019-aeb2-bd20e66ba847 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Kimi-VL Technical Report
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c09be85-8ae3-45d5-8f53-fa0ac36d52bd · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Minicpm-o 2.6: A gpt- 4o level mllm for vision, speech, and multimodal live streaming on your phone, 2025
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 46da65d6-b462-495e-999b-06483925c7ee · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e2068ef-3aab-4765-9979-53bdc9d23ce9 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a28e57e-f308-4adb-bac1-9c28549b9ffd · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Twinanda, Sherif Shehata, Didier Mutter, Jacques Marescaux, Michel De Mathelin, and Nicolas Padoy
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dbdd323f-ac16-470b-b460-f4b1e75f49fe · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Molecular-driven Foundation Model for Oncologic Pathology
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9518bd0a-af75-4f71-b35b-9ade359d7634 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence A foundation model for clinical- grade computational pathology and rare cancers detec- tion
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 44826df3-3d5f-476b-928c-6eb3a2d710fa · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Copesd: A multi-level surgical motion dataset for training large vision-language models to co- pilot endoscopic submucosal dissection, 2024
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6bd54c96-1e80-4e08-98e4-00190f63cb7d · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence EndoChat: Grounded Multimodal Large Language Model for Endoscopic Surgery
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc6f7719-a6cf-4710-a705-22fdb94db794 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84b5f96f-d5eb-4e4a-9609-597e91f0fd1b · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence A pathology foundation model for cancer diagnosis and prognosis prediction
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0a840cea-e449-4a49-9cb7-134eb8dc46f9 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Auto- laparo: A new dataset of integrated multi-tasks for image-guided surgical automation in laparoscopic hys- terectomy, 2022
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0b3d9e2c-71b3-461d-9504-3d16f3746b14 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d594965-ff03-42c5-ba27-4ad8dbb124c4 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f7a165f-6b33-4b4a-8d28-4283ce774e65 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Qwen2.5 Technical Report
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cb83604-b9c2-4455-aad2-ecbef0664fc1 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd352e31-d82c-4085-874a-5d2512ec83de · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d20ba40d-23a4-455e-afc3-51a00dc46498 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Learning Multi-modal Representations by Watching Hundreds of Surgical Video Lectures
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b81bb05c-8a99-4764-94ac-9532dc5dba02 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Advanc- ing surgical vqa with scene graph knowledge
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f4c384da-9b6f-4b37-b5da-d662d749e87c · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Cognition guided human-object re- lationship detection
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a5ce6b65-fbcd-4a22-bbc6-7a1523932501 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Sigmoid loss for language image pre- training
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e7381476-3277-4121-a530-b252841a80b1 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0a8dfb3-6cc0-40ee-a393-04b366ffd805 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Long Context Transfer from Language to Vision
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09b1edc2-b9e3-47bb-bfe4-8544bf68d961 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Knowledge-enhanced visual-language pre-training on chest radiology images.Nature Commu- nications, 14(1):4542, 2023
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f5aa36f7-b1ca-45f4-a42b-58e1f7303da1 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Large-scale long-tailed disease diagnosis on radiology images
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ee514e3a-760c-49f7-8dfb-3f8be28bc8ec · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 662e2651-e10d-4bae-8b15-ee495186f028 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence The demonstrated step is gastrojejunal defect closure, and during the step, the surgeon is closing the orifice left by the stapler, creating the gastrojejunostomy
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 322359df-0745-4d0a-b14a-3a1720686582 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Unresolved cited work
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 19dd6fe9-d7ce-4153-90b9-96e24f698e31 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Unresolved cited work
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 09f2ff85-34ee-4589-90d1-b371c2a4430c · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence LLM Decoder V V V V V V TTTT Multimodal Fusion Vision Encoder Text Tokenizer �� �푀 �� �� 1
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2f842d4a-97bb-4980-b285-7654b121a2df · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Unresolved cited work
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c78ae143-3429-4f59-9ac1-1f8da21575dc · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Unresolved cited work
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 73e7d5eb-ba27-476c-bfd4-13010016698e · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Overall illustration of proposed SurgVLM
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4ba1e0a0-a4da-46b9-a1f6-e636befa89e9 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Unresolved cited work
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1e56be53-a7d8-45c5-ab9e-e085d6a491ec · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence Unresolved cited work
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 214ffcbe-8112-4362-9033-9ef928783ac0 · outbound
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence present” vs. “absent
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dae74437-f337-4167-8bbc-1122a3242d7f · inbound
SiMing-Bench: Evaluating Procedural Correctness from Continuous Interactions in Clinical Skill Videos SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 96144497-895a-42ca-8828-9ccef4c52098 · inbound
Adapting 2D Multi-Modal Large Language Model for 3D CT Image Analysis SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fbaaae2f-b5e0-4082-bee6-8fcb0a7c7c43 · inbound
SurgCoT: Advancing Spatiotemporal Reasoning in Surgical Videos through a Chain-of-Thought Benchmark SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c65108b3-6832-4304-a2ad-88a4e996827d · inbound
Watch, Remember, Reason: Human-View Video Understanding with MLLMs SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence
Reference 268
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0c4774a1-6508-40b4-b2da-5a4aacc9d219 · inbound
SurgiQ: A Large-Scale Multi-Domain Benchmark for Evaluating Surgical Understanding in Large Language Models SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1fa95e6b-2896-41eb-af80-f37239e7d809 · inbound
SurgAtlas: A Large-Scale Surgical Video-Language Dataset with 2,391 Hours of Open and Minimally Invasive Surgery SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 61ddf5ee-050e-4002-8d1d-b67130108209 · inbound
LAVIFT: Latent-Action-Guided Vision Fine-Tuning for Surgical Interaction Recognition SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05a2e1cd-382f-428c-9eff-a8a348db8818 · inbound
MarineEVT: Advancing Event-Centric Marine Video Understanding via Visual Tool Reasoning SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b4d878f-5b6e-40f6-9069-78358e2fbd1e · inbound
SurgNarrator: A Generative Retrieval Framework for Surgical Video Understanding SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.