Search arXiv⌕ Search

EXPLORE CONNECTIONS

PaLMR: Towards Faithful Visual Reasoning via Multimodal Process Alignment

Follow the relationships that help you find your next source.

Based on 525 indexed works selected for this snapshot; the candidate window is limited. Counts describe this index, not the complete source archives. Prepared from the PostgreSQL corpus; source versions are checked before display. Snapshot 2026-09-26. Up to 32 works or names per graph.

Subjects & research connections

Connections use shared source subjects, names, places, and explicitly mentioned entities. A shared label is not evidence of a citation, experimental result, or verified species identification.

cs.CV · Kai Wangcs.CV · cs.AIcs.CV · cs.AIcs.CV · cs.AIcs.CV · cs.AIcs.CV · cs.AIcs.CV · cs.AIcs.CV · cs.AIcs.CV · cs.AIcs.CV · cs.AIcs.CV · cs.AIcs.CV · cs.AIcs.CV · cs.AIcs.CV · cs.AIcs.CV · cs.AIcs.AI · cs.CVcs.CVcs.CVcs.AIcs.CVcs.AIcs.AIcs.AIcs.AIcs.CVcs.CVcs.CVcs.CVcs.AIcs.AIcs.AIPaLMR: Towards Faithful Visual Reasoning via Multimodal Process AlignmentPaLMR: Towards Faithful…Refinement Is Inherently Editable: Training-Free Prompt-to-Prompt Image Editing with Generative Refinement Network2Temporal and Contextual Transformer for Multi-Camera Editing of TV Shows3Band-Attention Modulation Network for Robust Face Forgery Detection4Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery5Cross-Task Generalization in Handwriting-Based Alzheimer's Screening via Vision Language Adaptation6VLANeXt: Recipes for Building Strong VLA Models7MPFlow: Multi-modal Posterior-Guided Flow Matching for Zero-Shot MRI Reconstruction8Interpreting and Enhancing Emotional Circuits in Large Vision-Language Models via Cross-Modal Information Flow9A Multimodal 3D Foundation Model for Light Sheet Fluorescence Microscopy Enables Few-Shot Segmentation, Classification, and Deblurring10Deep Learning-based 3D Oral Cavity Reconstruction Using 2D Intraoral Images113D Oral Modelling with Improved Vertex Distribution Using Matching-Based Learning12JoyAI-VL-Interaction: Real-Time Vision-Language Interaction Intelligence13Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning14SOV-CAD: Stepwise Orthographic Views Guided CAD Modeling Sequence Reconstruction15Q-CueGraph: Query-Conditioned Visual Evidence Graphs for Multimodal Reasoning16Evolve Vision-Language-Action Model into an Agent with On-the-fly Tool-use17Improving Binary Neural Networks through Fully Utilizing Latent Weights18FlowFace++: Explicit Semantic Flow-supervised End-to-End Face Swapping19Joint Prediction and Denoising for Large-scale Multilingual Self-supervised Learning20To Generate or Not? Safety-Driven Unlearned Diffusion Models Are Still Easy To Generate Unsafe Images ... For Now21ReLU$^2$ Wins: Discovering Efficient Activation Functions for Sparse LLMs22ELiSe: Efficient Learning of Sequences in Structured Recurrent Networks23Generating Interesting Scientific Ideas using Knowledge Graphs and LLMs: Evaluations with 100 Research Group Leaders24Probing many-body Bell correlation depth with superconducting qubits25AIR: Analytic Imbalance Rectifier for Continual Learning26LaneTCA: Enhancing Video Lane Detection with Temporal Context Aggregation27MultiViewDx: Evidence-Linked Multi-View Clinical Diagnosis28Comparing YOLOv11 and YOLOv8 for instance segmentation of occluded and non-occluded immature green fruits in complex orchard environment29Addressing Uncertainty in LLMs to Enhance Reliability in Generative AI30How Do Users Negotiate Harmful Value Conflicts with AI Companions? A Study with Minion, a Technology Probe for In-Situ Human-AI Conflict Response31Foundations of Large Language Models32
Relationships as a list (31)

Citation graph

Arrows run from the citing work to its reference. Only explicit source/provider reference lists are used. Incoming links cover this candidate window; this is not a global citation count.

No supported relationships are available in this snapshot. This does not mean that no relationships exist.

Collaborator network

Shared authorship within up to 200 candidate works (2 examined); up to 32 names shown. Names are matched as supplied, without verified person disambiguation. Shared credit does not necessarily establish personal collaboration.

1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared workKai WangKai WangChao Tan2Chenyang Yan3Fang Zhao4Huanling Gao5Jianbing Zhang6Kanzhi Cheng7Qiang Hui8Shiguo Lian9Xinyu Dai10Yantao Li11Ao He12Haoyu Zhang13Senmao Li14Yaxing Wang15Yulong Chen16Ziqian Zhang17
Relationships as a list (61)

Publication timeline

Publication years for these 32 related works; 0 have no source publication date. This is a discovery sample, not a measure of research output or growth.

2021 · 1 work
2022 · 1 work
2023 · 2 works
2024 · 5 works
2025 · 1 work
2026 · 22 works

Semantic map

Model: all-minilm. Positions are a two-dimensional approximation of embedding similarity; proximity is not a citation or proof of agreement. Results come from the snapshot’s candidate window.

No current, compatible embeddings are available for related works in this snapshot. A semantic map appears after background embedding and snapshot generation.