EXPLORE CONNECTIONS
More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning
Follow the relationships that help you find your next source.
Based on 525 indexed works selected for this snapshot; the candidate window is limited. Counts describe this index, not the complete source archives. Prepared from the PostgreSQL corpus; source versions are checked before display. Snapshot 2026-09-26. Up to 32 works or names per graph.
Subjects & research connections
Connections use shared source subjects, names, places, and explicitly mentioned entities. A shared label is not evidence of a citation, experimental result, or verified species identification.
Relationships as a list (31)
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI · cs.RO 2. Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.RO · cs.AI 3. HERMES: A Holistic End-to-End Risk-Aware Multimodal Embodied System with Vision-Language Models for Long-Tail Autonomous Driving
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI · cs.RO 4. VLANeXt: Recipes for Building Strong VLA Models
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.RO · cs.AI 5. Evolve Vision-Language-Action Model into an Agent with On-the-fly Tool-use
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.RO · cs.AI 6. TacSushi: Tactile-Grounded World-Action Modeling for Dexterous Sushi Manipulation
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 7. Temporal and Contextual Transformer for Multi-Camera Editing of TV Shows
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 8. Joint Prediction and Denoising for Large-scale Multilingual Self-supervised Learning
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 9. ReLU$^2$ Wins: Discovering Efficient Activation Functions for Sparse LLMs
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 10. ELiSe: Efficient Learning of Sequences in Structured Recurrent Networks
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 11. Band-Attention Modulation Network for Robust Face Forgery Detection
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 12. Generating Interesting Scientific Ideas using Knowledge Graphs and LLMs: Evaluations with 100 Research Group Leaders
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 13. Probing many-body Bell correlation depth with superconducting qubits
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 14. Addressing Uncertainty in LLMs to Enhance Reliability in Generative AI
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 15. How Do Users Negotiate Harmful Value Conflicts with AI Companions? A Study with Minion, a Technology Probe for In-Situ Human-AI Conflict Response
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 16. Foundations of Large Language Models
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 17. Multimodal AI predicts clinical outcomes of drug combinations from preclinical data
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 18. Search-Based Software Engineering and AI Foundation Models: Current Landscape and Future Roadmap
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 19. SheetMind: Actions Set Accuracy, Agents Set the Failure Mode
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 20. WebArxiv: A Reproducible Benchmark for Evaluating Multimodal Web Agents on arXiv Tasks
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 21. RedCoder: Automated Multi-Turn Red Teaming for Code LLMs
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 22. Conversational DNA: A Visual Language and Interactive Atlas of Human and AI Dialogue
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.RO 23. Physics-Guided Residual Reinforcement Learning for Humanoid Narrow-Path Traversal
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 24. Cross-Task Generalization in Handwriting-Based Alzheimer's Screening via Vision Language Adaptation
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 25. LOGIC: Efficient and Robust Contextual Biasing for Speech LLMs via Logit-Space Integration
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 26. TIDE: Temporal Incremental Draft Engine for Self-Improving LLM Inference
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 27. TabSieve: Explicit In-Table Evidence Selection for Tabular Prediction
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 28. Decoding ML Decision: An Agentic Reasoning Framework for Large-Scale Ranking System
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 29. SHINE: Sequential Hierarchical Integration Network for EEG and MEG
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 30. MPFlow: Multi-modal Posterior-Guided Flow Matching for Zero-Shot MRI Reconstruction
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 31. Learning Causal Structure of Time Series using Best Order Score Search
- 1. More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning cs.AI 32. PaLMR: Towards Faithful Visual Reasoning via Multimodal Process Alignment
Citation graph
Arrows run from the citing work to its reference. Only explicit source/provider reference lists are used. Incoming links cover this candidate window; this is not a global citation count.
No supported relationships are available in this snapshot. This does not mean that no relationships exist.
Collaborator network
Shared authorship within up to 200 candidate works (1 examined); up to 32 names shown. Names are matched as supplied, without verified person disambiguation. Shared credit does not necessarily establish personal collaboration.
Relationships as a list (10)
- Alexandre Chapin 1 shared work Yi Li
- Liming Chen 1 shared work Yi Li
- Jan Peters 1 shared work Yi Li
- Alap Kshirsagar 1 shared work Yi Li
- Alexandre Chapin 1 shared work Liming Chen
- Alexandre Chapin 1 shared work Jan Peters
- Alap Kshirsagar 1 shared work Alexandre Chapin
- Jan Peters 1 shared work Liming Chen
- Alap Kshirsagar 1 shared work Liming Chen
- Alap Kshirsagar 1 shared work Jan Peters
Publication timeline
Publication years for these 32 related works; 0 have no source publication date. This is a discovery sample, not a measure of research output or growth.
2024 · 3 works
2026 · 26 works
- More Structure, Not More Capacity: Object-Centric Representations for Visuomotor Imitation Learning
- HERMES: A Holistic End-to-End Risk-Aware Multimodal Embodied System with Vision-Language Models for Long-Tail Autonomous Driving
- VLANeXt: Recipes for Building Strong VLA Models
- Evolve Vision-Language-Action Model into an Agent with On-the-fly Tool-use
- TacSushi: Tactile-Grounded World-Action Modeling for Dexterous Sushi Manipulation
- ELiSe: Efficient Learning of Sequences in Structured Recurrent Networks
- Band-Attention Modulation Network for Robust Face Forgery Detection
- Generating Interesting Scientific Ideas using Knowledge Graphs and LLMs: Evaluations with 100 Research Group Leaders
- How Do Users Negotiate Harmful Value Conflicts with AI Companions? A Study with Minion, a Technology Probe for In-Situ Human-AI Conflict Response
- Foundations of Large Language Models
- Multimodal AI predicts clinical outcomes of drug combinations from preclinical data
- Search-Based Software Engineering and AI Foundation Models: Current Landscape and Future Roadmap
- SheetMind: Actions Set Accuracy, Agents Set the Failure Mode
- WebArxiv: A Reproducible Benchmark for Evaluating Multimodal Web Agents on arXiv Tasks
- RedCoder: Automated Multi-Turn Red Teaming for Code LLMs
- Conversational DNA: A Visual Language and Interactive Atlas of Human and AI Dialogue
- Physics-Guided Residual Reinforcement Learning for Humanoid Narrow-Path Traversal
- Cross-Task Generalization in Handwriting-Based Alzheimer's Screening via Vision Language Adaptation
- LOGIC: Efficient and Robust Contextual Biasing for Speech LLMs via Logit-Space Integration
- TIDE: Temporal Incremental Draft Engine for Self-Improving LLM Inference
- TabSieve: Explicit In-Table Evidence Selection for Tabular Prediction
- Decoding ML Decision: An Agentic Reasoning Framework for Large-Scale Ranking System
- SHINE: Sequential Hierarchical Integration Network for EEG and MEG
- MPFlow: Multi-modal Posterior-Guided Flow Matching for Zero-Shot MRI Reconstruction
- Learning Causal Structure of Time Series using Best Order Score Search
- PaLMR: Towards Faithful Visual Reasoning via Multimodal Process Alignment
Semantic map
Model: all-minilm. Positions are a two-dimensional approximation of embedding similarity; proximity is not a citation or proof of agreement. Results come from the snapshot’s candidate window.
No current, compatible embeddings are available for related works in this snapshot. A semantic map appears after background embedding and snapshot generation.