EXPLORE CONNECTIONS
Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery
Follow the relationships that help you find your next source.
Based on 525 indexed works selected for this snapshot; the candidate window is limited. Counts describe this index, not the complete source archives. Prepared from the PostgreSQL corpus; source versions are checked before display. Snapshot 2026-09-26. Up to 32 works or names per graph.
Subjects & research connections
Connections use shared source subjects, names, places, and explicitly mentioned entities. A shared label is not evidence of a citation, experimental result, or verified species identification.
Relationships as a list (31)
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV · cs.AI · cs.RO 2. VLANeXt: Recipes for Building Strong VLA Models
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.RO · cs.AI · cs.CV 3. Evolve Vision-Language-Action Model into an Agent with On-the-fly Tool-use
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV · cs.AI 4. Temporal and Contextual Transformer for Multi-Camera Editing of TV Shows
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV · cs.AI 5. Band-Attention Modulation Network for Robust Face Forgery Detection
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery eess.IV · cs.CV 6. Bridging the Inter-Domain Gap through Low-Level Features for Cross-Modal Medical Image Segmentation
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV · cs.AI 7. Cross-Task Generalization in Handwriting-Based Alzheimer's Screening via Vision Language Adaptation
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.RO · cs.AI 8. HERMES: A Holistic End-to-End Risk-Aware Multimodal Embodied System with Vision-Language Models for Long-Tail Autonomous Driving
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV · cs.AI 9. MPFlow: Multi-modal Posterior-Guided Flow Matching for Zero-Shot MRI Reconstruction
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV · cs.AI 10. PaLMR: Towards Faithful Visual Reasoning via Multimodal Process Alignment
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.RO · cs.CV 11. RotVLA: Rotational Latent Action for Vision-Language-Action Model
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV · cs.AI 12. Interpreting and Enhancing Emotional Circuits in Large Vision-Language Models via Cross-Modal Information Flow
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV · cs.AI 13. A Multimodal 3D Foundation Model for Light Sheet Fluorescence Microscopy Enables Few-Shot Segmentation, Classification, and Deblurring
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV · cs.AI 14. Deep Learning-based 3D Oral Cavity Reconstruction Using 2D Intraoral Images
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV · cs.AI 15. 3D Oral Modelling with Improved Vertex Distribution Using Matching-Based Learning
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV · cs.AI 16. JoyAI-VL-Interaction: Real-Time Vision-Language Interaction Intelligence
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV · cs.AI 17. Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV · cs.AI 18. SOV-CAD: Stepwise Orthographic Views Guided CAD Modeling Sequence Reconstruction
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.RO · cs.AI 19. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.RO · cs.CV 20. Learning to Navigate with Minimal Parameters: Decomposing Visual Navigation Through Closed-Form Geometric Interfaces
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV · cs.AI 21. Q-CueGraph: Query-Conditioned Visual Evidence Graphs for Multimodal Reasoning
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.RO · cs.AI 22. TacSushi: Tactile-Grounded World-Action Modeling for Dexterous Sushi Manipulation
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV · eess.IV 23. TAPe+ML: A Compact Structured Representation for Multi-Task Computer Vision
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV 24. Improving Binary Neural Networks through Fully Utilizing Latent Weights
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV 25. FlowFace++: Explicit Semantic Flow-supervised End-to-End Face Swapping
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.AI 26. Joint Prediction and Denoising for Large-scale Multilingual Self-supervised Learning
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV 27. To Generate or Not? Safety-Driven Unlearned Diffusion Models Are Still Easy To Generate Unsafe Images ... For Now
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.AI 28. ReLU$^2$ Wins: Discovering Efficient Activation Functions for Sparse LLMs
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.AI 29. ELiSe: Efficient Learning of Sequences in Structured Recurrent Networks
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.AI 30. Generating Interesting Scientific Ideas using Knowledge Graphs and LLMs: Evaluations with 100 Research Group Leaders
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.AI 31. Probing many-body Bell correlation depth with superconducting qubits
- 1. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery cs.CV 32. AIR: Analytic Imbalance Rectifier for Continual Learning
Citation graph
Arrows run from the citing work to its reference. Only explicit source/provider reference lists are used. Incoming links cover this candidate window; this is not a global citation count.
No supported relationships are available in this snapshot. This does not mean that no relationships exist.
Collaborator network
Shared authorship within up to 200 candidate works (1 examined); up to 32 names shown. Names are matched as supplied, without verified person disambiguation. Shared credit does not necessarily establish personal collaboration.
Relationships as a list (45)
- Guankun Wang 1 shared work Long Bai
- Guankun Wang 1 shared work Wan Jun Nah
- Guankun Wang 1 shared work Jie Wang
- Guankun Wang 1 shared work Zhaoxi Zhang
- Guankun Wang 1 shared work Zhen Chen
- Guankun Wang 1 shared work Jinlin Wu
- Guankun Wang 1 shared work Mobarakol Islam
- Guankun Wang 1 shared work Hongbin Liu
- Guankun Wang 1 shared work Hongliang Ren
- Long Bai 1 shared work Wan Jun Nah
- Jie Wang 1 shared work Long Bai
- Long Bai 1 shared work Zhaoxi Zhang
- Long Bai 1 shared work Zhen Chen
- Jinlin Wu 1 shared work Long Bai
- Long Bai 1 shared work Mobarakol Islam
- Hongbin Liu 1 shared work Long Bai
- Hongliang Ren 1 shared work Long Bai
- Jie Wang 1 shared work Wan Jun Nah
- Wan Jun Nah 1 shared work Zhaoxi Zhang
- Wan Jun Nah 1 shared work Zhen Chen
- Jinlin Wu 1 shared work Wan Jun Nah
- Mobarakol Islam 1 shared work Wan Jun Nah
- Hongbin Liu 1 shared work Wan Jun Nah
- Hongliang Ren 1 shared work Wan Jun Nah
- Jie Wang 1 shared work Zhaoxi Zhang
- Jie Wang 1 shared work Zhen Chen
- Jie Wang 1 shared work Jinlin Wu
- Jie Wang 1 shared work Mobarakol Islam
- Hongbin Liu 1 shared work Jie Wang
- Hongliang Ren 1 shared work Jie Wang
- Zhaoxi Zhang 1 shared work Zhen Chen
- Jinlin Wu 1 shared work Zhaoxi Zhang
- Mobarakol Islam 1 shared work Zhaoxi Zhang
- Hongbin Liu 1 shared work Zhaoxi Zhang
- Hongliang Ren 1 shared work Zhaoxi Zhang
- Jinlin Wu 1 shared work Zhen Chen
- Mobarakol Islam 1 shared work Zhen Chen
- Hongbin Liu 1 shared work Zhen Chen
- Hongliang Ren 1 shared work Zhen Chen
- Jinlin Wu 1 shared work Mobarakol Islam
- Hongbin Liu 1 shared work Jinlin Wu
- Hongliang Ren 1 shared work Jinlin Wu
- Hongbin Liu 1 shared work Mobarakol Islam
- Hongliang Ren 1 shared work Mobarakol Islam
- Hongbin Liu 1 shared work Hongliang Ren
Publication timeline
Publication years for these 32 related works; 0 have no source publication date. This is a discovery sample, not a measure of research output or growth.
2023 · 2 works
2024 · 3 works
2026 · 24 works
- VLANeXt: Recipes for Building Strong VLA Models
- Evolve Vision-Language-Action Model into an Agent with On-the-fly Tool-use
- Band-Attention Modulation Network for Robust Face Forgery Detection
- Bridging the Inter-Domain Gap through Low-Level Features for Cross-Modal Medical Image Segmentation
- Cross-Task Generalization in Handwriting-Based Alzheimer's Screening via Vision Language Adaptation
- HERMES: A Holistic End-to-End Risk-Aware Multimodal Embodied System with Vision-Language Models for Long-Tail Autonomous Driving
- MPFlow: Multi-modal Posterior-Guided Flow Matching for Zero-Shot MRI Reconstruction
- PaLMR: Towards Faithful Visual Reasoning via Multimodal Process Alignment
- RotVLA: Rotational Latent Action for Vision-Language-Action Model
- Interpreting and Enhancing Emotional Circuits in Large Vision-Language Models via Cross-Modal Information Flow
- A Multimodal 3D Foundation Model for Light Sheet Fluorescence Microscopy Enables Few-Shot Segmentation, Classification, and Deblurring
- Deep Learning-based 3D Oral Cavity Reconstruction Using 2D Intraoral Images
- 3D Oral Modelling with Improved Vertex Distribution Using Matching-Based Learning
- JoyAI-VL-Interaction: Real-Time Vision-Language Interaction Intelligence
- Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning
- SOV-CAD: Stepwise Orthographic Views Guided CAD Modeling Sequence Reconstruction
- More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning
- Learning to Navigate with Minimal Parameters: Decomposing Visual Navigation Through Closed-Form Geometric Interfaces
- Q-CueGraph: Query-Conditioned Visual Evidence Graphs for Multimodal Reasoning
- TacSushi: Tactile-Grounded World-Action Modeling for Dexterous Sushi Manipulation
- TAPe+ML: A Compact Structured Representation for Multi-Task Computer Vision
- ELiSe: Efficient Learning of Sequences in Structured Recurrent Networks
- Generating Interesting Scientific Ideas using Knowledge Graphs and LLMs: Evaluations with 100 Research Group Leaders
- AIR: Analytic Imbalance Rectifier for Continual Learning
Semantic map
Model: all-minilm. Positions are a two-dimensional approximation of embedding similarity; proximity is not a citation or proof of agreement. Results come from the snapshot’s candidate window.
No current, compatible embeddings are available for related works in this snapshot. A semantic map appears after background embedding and snapshot generation.