PaperScope
LIVE · 2026-09-09 05:40 UTC

multi-view 11 papers this week · +0% WoW

papers mentioning "multi-view" in title/abstract · 30d window

Latestcs.CLcs.LGcs.AIcs.CV

Mentions per Day (30d)

Latest Papers

#TitleAuthorsCatDate
1Prior-free relative 6D pose estimation of multiple object instancesBehdad Khodabandehloo, Andrea Caraffa, Davide Boscaini +1cs.CV2026-09-08
2Leveraging Visual and Geometric Priors for Metric-scale and Complete Vehicle Gaussian Reconstruction from Limited ViewsJinyu Miao, Jiusi Li, Yifei He +4cs.CV2026-09-08
3Towards Embodied Air-Ground Cooperative Object Search: Benchmark, Dataset and Agentic MethodBoao Yu, Zimo Chen, Junreng Rao +4cs.CV2026-09-08
#4CoVeR: Coverage-Based Token Pruning for Multi-View 3D Reasoning in VLMsNhat-Tan Bui, Varshini Elangovan, Arun Reddy Anugu +7cs.CV2026-09-08
#53DWay: Generalizing Robot Manipulation via 3D Consistent WaypointsZiqin Huang, Yingyue Li, Chenyangguang Zhang +6cs.RO2026-09-08
#6WSPolypNet: Weakly Supervised Polyp Localization in Colonoscopy VideosGiseong Hwang, Minjae Jo, Yeonghyeon Park +9cs.CV2026-09-08
#7Zero-Shot 3D Plant Organ Segmentation with SAM3 and Semantic NeRFsAndreas Gilson, Laura Hennig, Peter Pietrzykcs.CV2026-09-07
#8I Don't Miss You, but I Do: Self-Explanation Faithfulness of Modality Missingness in Vision-Language ModelsAydin Javadov, Daniel Schoess, Florian von Wangenheimcs.LG2026-09-07
#9Heat Kernel Textures: the Geodesic Gaussians That Do Not SplatSimone Foti, Caner Korkmaz, Stefanos Zafeiriou +1cs.CV2026-09-07
#10When Semantically Consistent Encoding Meets View-Label Heterogeneity Modeling: A Unified Framework for Incomplete Multi-View Multi-Label LearningChengliang Liu, Bo Li, Bob Zhang +3cs.CV2026-09-07
#11Self-Supervised Multi-View 3D Gaze Target Estimation via Probabilistic Ray MarchingKeqi Chen, Vinkle Srivastav, Nicolas Padoycs.CV2026-09-07
#12RelightFormer: Feed-forward Generative Transformer for Multiview Object RelightingHejun Wang, Jinxi Li, Junwei Jiang +4cs.CV2026-09-07
#13KODAMA: Multimodal Digital Twin Reconstruction for Urban RF Propagation ModellingMaximiliano Wardle, A. Ryo Koblitzcs.CV2026-09-07
#14MV-STRIDE: Enabling MLLMs to Master Multi-View Spatial Reasoning via Hierarchical Capability ModelingJin Xu, Xiaojian Huang, Zhuodong Luo +6cs.CV2026-09-07
#15SkillAlign: Aligning Skill Interfaces for LLM-based AgentsShuo Ren, Xiaomian Kang, Jiajun Zhangcs.AI2026-09-07
#16CARDEA: Auditable Reasoning Grounded in Spatial Evidence for End-to-End Coronary Angiography InterpretationJia-Jen Lee, Shih-Yen Hou, Kee Koon Ng +2cs.CV2026-09-07
#17Contextual Observer Grounding: Evaluating Situated Spatial Reasoning in Vision-Language ModelsMimo Shirasaka, Haochen Zhang, Yonatan Biskcs.CV2026-09-07
#18DrugReason: Dynamic Multi-View Reasoning over Knowledge Graph and Language Evidence for Drug RepurposingZijie Liu, Hongxuan Li, Zhen Tan +5cs.LG2026-09-06
#193DHarnessBench: Probing Agentic 3D-to-Code Capabilities of Frontier Vision-Language ModelsLing Liu, Bingchen Gong, Amal Dev Parakkat +1cs.CV2026-09-06
#20Scaling 3D Generative Priors to Large-Scale Scene Meshes from Multi-View ImagesSangEun Lee, Wonseok Chae, Hoyoung Yoo +3cs.CV2026-09-06
#21DualPathOcc: Dual-Resolution BEV Encoder for 3D Occupancy PredictionLihao Qiu, Jian Chen, Ruihao Wang +3cs.CV2026-09-06
#22WorldSculpt: Generating Compositional Worlds from Grounded VideosMuyao Niu, Jixuan He, Ruihan Yu +9cs.CV2026-09-04
#23CrossDepth: Geometry-Constrained Attention for Generalizable Multi-View Surround Depth EstimationSamer Abualhanud, Max Mehltrettercs.CV2026-09-04
#24Reflection-aware Generative Novel View SynthesisGeonU Kim, Shin Dong-Yeon, Tae-Hyun Ohcs.CV2026-09-04
#25BLASt3R: Bundle Adjustment of Any Image Set with Multi-View Matching and Monocular PriorsVincent Leroy, Philippe Weinzaepfel, Lojze Zust +2cs.CV2026-09-04
#26Learning 3D Editing without Paired Supervision via Generative Prior DistillationHao Wen, Weibin Yun, Hongxing Fan +4cs.CV2026-09-04
#27An Evaluation Framework for Generating Multi-View Images of a Person in a SceneMahir Majid, Young Kyung Kim, Guillermo Sapirocs.CV2026-09-04
#28BooM-VVT: Boosting Mask-Free Video Virtual Try-On with Image-Level Pseudo DataWei Zhang, Xin Li, Peishu Shi +4cs.CV2026-09-03
#29Continuous Actions from Discrete Minds: Latent-Aligned Planning for End-to-End Autonomous DrivingRuoyu Yao, Yusen Xie, Qingzhao Liu +5cs.CV2026-09-03
#30Sparse auto-regressive modeling for scene generation from multi-view imagesThomas Lucas, Maxime Pietrantoni, Philippe Weinzaepfel +4cs.CV2026-09-03
#31Unfold The World: Factorize 4D Properties in Reinforcing Spatial ReasoningYijun Yang, Shenghe Zheng, Wenbo Li +8cs.CV2026-09-03
#32Building Pretraining Data for World Models: An Unreal Engine-Based Pipeline for Action-Conditioned Video GenerationHaoyu Wang, Songchun Zhang, Haoran Li +3cs.CV2026-09-03
#33TraveL: Transformer-based Multi-view Path Distributional Representation LearningFang He, Tao-yang Fu, Wang-chien Leecs.LG2026-09-03
#34P-CORE: Self-Supervised Surface Consistency for Point-Based Neural EditingYanshu Zhang, Shichong Peng, Mehran Aghabozorgi +2cs.CV2026-09-03
#35MuyBridge: Mobile Human Center-of-Mass Estimation from Monocular Video via Sparse FusionAidan Bradshaw, Marco Giordano, David Rode +8cs.CV2026-09-02
#36InceptionGS: Generative Bootstrapping for Large-Scale Gaussian Splatting under Unstructured View SamplingTianheng Lu, Guangyu Wang, Ruqi Huang +1cs.CV2026-09-02
#37MV-dVRK: A Multi-Viewpoint Benchmark for Spatial Surgical PerceptionGuido Caccianiga, Sergey Prokudin, Yutong Chen +9cs.CV2026-09-02
#38Learning from Scarce Labels: Multi-View Echocardiography for Ejection Fraction PredictionZhiyuan Gao, Dominic Yurk, Yaser S. Abu-Mostafaeess.IV2026-09-02
#39TAPVid-MV: A Benchmark for Tracking Any Point in 3D Across Multiple ViewsSkanda Koppula, Frano Rajic, Abdullah Faiz Ur Rahman +9cs.CV2026-09-01
#40Cross-Model Distillation of a Human-Pose Foundation Model from Unannotated Infant Video for Markerless 3D Pose EstimationR. James Cotton, Divya Joshi, Colleen Peytoncs.CV2026-09-01
#41iPINN for Broadband CARS Phase Retrieval: A Framework for Function Approximation and Inverse Modeling Problems in Nonlinear SpectroscopyRavi Teja Vulchi, Carl Messerschmidt, Mohammadsadegh Vafaeinezhad +4cs.LG2026-09-01
#42MADS: A Multiview Acoustic Descriptor Set Beyond Standard Spectral SummariesUtsab Ghosh, Roshni Chakrabortycs.SD2026-09-01
#43Feed-Forward Multi-view Multi-person Reconstruction with Contrastive Human-Aware 3D RepresentationYuanwang Yang, Buzhen Huang, Zongxuan Ren +2cs.CV2026-09-01
#44Inverse Rendering for Modeling with Line PrimitivesKenji Tojo, Ariel Shamir, Nobuyuki Umetani +1cs.GR2026-09-01
#45Streaming4D: Accelerate 4D World Models via Block-wise Video Generation and Incremental ReconstructionXiaoyan Liu, Jiaxin Liu, Kangrui Li +1cs.CV2026-09-01
#46Not All Agreement Counts as Corroboration: Provenance-Conserving Multi-View Fusion for Typed Action Admission in Human-Robot CollaborationZekai Jin, Hanrong Zhang, Yihong Tang +3cs.RO2026-08-31
#47FaceSnap: Real-Time Personalized Lightstage Facial Performance CaptureRukhshanda Hussain, Noé Artru, Emeline Got +6cs.CV2026-08-31
#48DARP: A Calibrated Dual-Arm RGB-D-IR Dataset for Multi-View Robotic PerceptionManish Kansana, Mohammed Yusuf Mujawar, Sudip Mittal +2cs.RO2026-08-31
#49Multi-View Reflective Surface Inspection via Semantic-Saliency Cross-VerificationVan-Giang Nguyen, Thanh-Tuan Tran, Xuan-Hieu Phan +1cs.CV2026-08-31
#50MR-JEPA: A General Purpose Video Foundation Model for Cardiac MRIAthira J. Jacob, Puneet Sharma, Dorin Comaniciu +1cs.CV2026-08-31

all trends · matching is case-insensitive substring after tokenization