PaperScope
LIVE · 2026-09-17 05:40 UTC

perception 13 papers this week · +75% WoW

papers mentioning "perception" in title/abstract · 30d window

Latestcs.CLcs.LGcs.AIcs.CV

Mentions per Day (30d)

Latest Papers

#TitleAuthorsCatDate
1MUSE: Benchmarking Large Vision-Language Models on Multi-Modal Understanding in Situated EducationLuyao Zhu, Xun Wei Yee, Wei Li +2cs.AI2026-09-16
2ReFigBench: Benchmarking Scientific Figure Reconstruction as Editable PowerPoint ArtifactsLiyang Fan, Chi Wei, Yitai Li +8cs.CL2026-09-16
3RankGround: Efficient High-Resolution GUI Grounding via Lightweight Reranker-Guided Crop SelectionLiyang Fan, Xinping Bi, Yitai Li +3cs.CV2026-09-16
#4The Uneven Impact of Generative AI on Student Learning: Examining the Roles of Reliance, Evaluation Literacy, and Course Policy in AI-related CoursesLydia Manikonda, Mei Si, Sirajam Munira +2cs.AI2026-09-16
#5HAP: A Hand-Driven Active Perception Framework for Egocentric Head Motion PredictionYunji Feng, Junyi Ma, Guanzhong Sun +2cs.CV2026-09-16
#6Accuracy- and Real-Time-Aware 4D Radar Preprocessing for Autonomous Driving Perception SystemsWoo-Jin Jung, Dong-Hee Paek, Jeong-Su Park +1cs.CV2026-09-16
#7AeroWeaver: An Embodied-Agent Harness for Weaving Aerial Skills into Distributed, Adaptive Swarm ExecutionJiabin Lou, Yirong Yang, Haopeng Wang +6cs.AI2026-09-16
#8ActiveScale: Scaling Active Perception for Robots across Model, Data, and HardwareShuai Zhou, Kaisheng Pang, Wenxuan Song +3cs.RO2026-09-16
#9Learning from Distributed Eyes: Leveraging Collaborative Perception for Automated Model AdaptationYanan Ma, Yihang Tao, Zhengru Fang +4cs.CV2026-09-16
#10Semantic-ITC: A Frame-wise Indoor Mobile Laser Scanning Dataset and Benchmark for Semantic SegmentationHaiyang Wu, Muhammad Affan, George Vosselman +1cs.CV2026-09-16
#11Emotion Experience, Expression, and Perception: Emotion Analysis on Multimodal Social Media PostsChristopher Bagdon, Carina Silberer, Roman Klingercs.CL2026-09-16
#12Understanding Dynamic Scenes at Gigapixel Scale: Wide-Area Spatio-Temporal Perception from UAVsYuhang Zhu, Meiyi Zhu, Yunkai Dang +4cs.CV2026-09-16
#13Multi-View Mixture-of-Experts with Vision-Language Reranking for Cross-View Object Geo-LocalizationXuyu Fan, Qi Ming, Zhu Han +6cs.CV2026-09-16
#14Stealthy in Semantics, Antagonistic in Space: Attacking Visible-Infrared Object Detectors via Object-Level MisalignmentYueqi Zhu, Qi Ming, Guo Cheng +6cs.CV2026-09-16
#15A Comprehensive Review of Generative Physical Artificial IntelligenceSatyam Gaba, Krutiksinh Rana, Siva Sai +2cs.RO2026-09-16
#16Beyond Pixel Similarity: Task-Aware Evaluation of GAN-Based Synthetic Sonar Data for Robotic PerceptionHannan Ejaz Keen, Muhammad Moazam Fraz, Karsten Bernscs.RO2026-09-16
#17Finder: Agentic Closed-Loop Object Finding for Embodied GroundingShixiong Xu, Zhiyuan Chen, Song Ding +4cs.CV2026-09-16
#18Anchoring What Matters: A Dual-Level Learning Framework for Visually-Grounded Multimodal ReasoningXinxin Song, Siyuan Li, Tingxiong Xiao +1cs.AI2026-09-16
#19Collaborative Memory for Multi-Agent VLM SystemsHuixin Zhang, Shao-Jun Xia, Di Wang +3cs.AI2026-09-15
#20Investigating Adversarial Robustness of Heterogeneous Cooperative PerceptionChenyi Wang, Yutong Liu, Qingzhao Zhang +1cs.CV2026-09-15
#21Semantic-Spatial Agreement Verification for Mitigating Object Hallucination in Multimodal Large Language ModelsZiheng Ren, Qian Gao, Jun Fan +3cs.CV2026-09-15
#22HuMemSLAM: Efficient Human-Inspired Semantic Place Recognition for Robust Visual SLAMMayowa Adebambo, Sebastian Donnelly, Armand Amaritei +2cs.RO2026-09-15
#23Intrinsic Robot Rewarding: Reusing VLA Representations for Autonomous Evaluation and Policy ImprovementTobias Schaffer, Mohab Elkhayat, Daniela Nicklas +2cs.RO2026-09-15
#24High-Fidelity Video Quality Assessment with VQA-Specific SaliencyHakan Emre Gedik, Shashank Gupta, Alan Bovikcs.CV2026-09-15
#25NeuroSymbEAD: A Large Scale Neuro-Symbolic Caption Dataset for Omni-Directional Embodied Autonomous DrivingMuhammad Ahmed Ullah Khan, Mohammed Elamine, Sheikh Talha Uddin +3cs.CV2026-09-15
#26Bridging Learned Visual Perception and Symbolic Belief-Space PlanningGuy Azran, Michael Navat, Sarah Kerencs.AI2026-09-15
#27VOR-Bench: A Human Perception-Driven Benchmark for Video Object RemovalHaonan Huang, Tianrui Qiu, Xianghao Zang +10cs.CV2026-09-15
#28TEMPO: Learning Temporal Context for Dynamic Robot ManipulationZhenyang Feng, Jimin Heo, Erik B. Sudderth +1cs.RO2026-09-15
#29CoAdapt: An LLM-based Framework for Adaptive Collaborative Perception in IIoT Robotic SwarmsHoussam Hajj Hassan, Antonia Maria Masucci, Lynda Zitoune +1cs.AI2026-09-15
#30Hyper-RED: Scalable Event Pre-training via Semantic Hypergraph DistillationMeisen Wang, Zhiqiang Tian, Wei Bao +3cs.CV2026-09-15
#31LEAP: Learning Emergent Active Perception for Quadruped NavigationÜ. Bora Gökbakan, Stéphane Caron, Philippe Souèrescs.RO2026-09-15
#32World Models for Embodied Intelligence: From Plausible to Controllable to ActionableNanjie Yao, Hao Wang, Chong Cheng +10cs.RO2026-09-15
#33Bridging the Perceptual Gap: Residual-Enhanced Downscaling and Manifold-Aware Perception Alignment Adaptation for NR-IQAYu Li, Zhengran Shen, Yachun Mi +2cs.CV2026-09-15
#34Can Knowledge Transfer Parameters Be Learned? LePoKet for Efficient Robotic VisionYanick C. Tchenko, Felix Mohr, Hicham Hadj-Abdelkader +1cs.CV2026-09-15
#35OPD-Aha: From Linguistic Momentum to Visual Reflection in Multimodal On-Policy DistillationChenhao Qiu, Dawei Li, Yechao Zhang +2cs.LG2026-09-15
#36Can a Neural Encoding Model Replicate an fMRI Visualization Study?Erfan Nasirzadeh Orang, Zack Whilecs.HC2026-09-14
#37Diversified and Perceptible Counterfactual Examples Leveraging Expert KnowledgeAkram Bensalem, Fahima Djelil, Marie-Jeanne Lesot +1cs.AI2026-09-14
#38Authorship attribution and aesthetic evaluation of AI poetry: a case study with HaikuLivia Oddi, Simone Scardapane, Toru Sugimoto +1cs.CL2026-09-14
#39Augmenting Large Audio-Language Models with Frame-Level Grounding for Fine-Grained Temporal PerceptionYanfeng Shi, Yan Song, Junhui Li +4cs.SD2026-09-14
#40Legislating World-Model-Based Planning with Legal ReasoningDylan Waldner, Yiannis Kantaros, Guido Governatori +2cs.RO2026-09-14
#41LG-VLN: A Zero-Shot Vision-and-Language Navigation Framework with LangGraph State OrchestrationJianhe Zhao, Yanhua Qiu, Zhiyu Zhang +2cs.CV2026-09-14
#42ER-EDF: A Psychology-Grounded Emotion Regulation Framework for Speech Empathetic Dialogue Generation in Large Audio-Language ModelsHongyu Jin, Wenda Zhang, Runqiu Fei +3cs.AI2026-09-14
#43LLMs as Oracles: Reliance on LLMs for Subjective Personal QuestionsMyra Cheng, Lujain Ibrahim, Grace Liu +5cs.CY2026-09-13
#44Func-R1: Incentivizing Mathematical Function Reasoning in Multimodal Large Language ModelsMingze Yin, Xiaohan Wang, Dian Li +8cs.CL2026-09-13
#45Perceive, Refine, Reason: A Calibrated Pipeline for Measuring Indicators in Strategic Visual Communication on Social MediaWeihong Qi, Chen Lingcs.CV2026-09-13
#46Investigating the Impacts of Generative AI on Information SeekingAlexi Orchard, Shannon Lodoencs.HC2026-09-13
#47CGGT: Curve-Grounded Geometry Transformer for 3D Parametric Curve ReconstructionZhirui Gao, Renjiao Yi, Yunfan Ye +4cs.CV2026-09-13
#48OptoAgent: A Trustworthy Multi-Agent Framework for Opportunistic Vision Micro-Screening in Classroom EnvironmentsToqeer Ali Syed, Ali Akarma, Adeel Ahmad +1cs.AI2026-09-13
#49Beyond Scene Description: Multi-Agent Orchestration for Non-visual Access to Virtual WorldsToqeer Ali Syed, Ali Akarma, Adeel Ahmad +1cs.AI2026-09-13
#50AlayaVista: Streaming World Modeling from Panoramic States to Perspective VideoJiaming Tan, Mingliang Zhai, Zhen Li +3cs.CV2026-09-13

all trends · matching is case-insensitive substring after tokenization