Kailun Yang
収録論文 139本 ・ フィジカルAI/ロボット学習
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- HGeo-TopoMap: Boosting Topological Mapping with Hierarchical Geometric Priors2026/7/1
- Label-Free Target-Domain Adaptation for Unconstrained Event-Image Feature Matching via Dual-Stage Distillation2026/7/1
- DexPIE: Stable Dexterous Policy Improvement from Real-World Experience2026/6/1
- CylindTrack: Depth-Aware Cylindrical Motion Modeling for Panoramic Multi-Object Tracking2026/6/1
- PS-MOT: Cultivating Instance Awareness from Point Seeds for Multi-Object Tracking2026/6/1
- EgoEV-HandPose: Egocentric 3D Hand Pose Estimation and Gesture Recognition with Stereo Event Cameras2026/5/1
- Towards Multi-Source Domain Generalization for Sleep Staging with Noisy Labels2026/4/1
- BLaDA: Bridging Language to Functional Dexterous Actions within 3DGS Fields2026/4/1
- AnyUser: Translating Sketched User Intent into Domestic Robots2026/4/1
- E-VLA: Event-Augmented Vision-Language-Action Model for Dark and Blurred Scenes2026/4/1
- ProOOD: Prototype-Guided Out-of-Distribution 3D Occupancy Prediction2026/4/1
- InterEdit: Navigating Text-Guided 3D Dyadic Human Motion Editing2026/3/1
- OccTrack360: 4D Panoptic Occupancy Tracking from Surround-View Fisheye Cameras2026/3/1
- O3N: Omnidirectional Open-Vocabulary Occupancy Prediction for Urban Autonomous Agents2026/3/1
- PanoAffordanceNet: Towards Holistic Affordance Grounding in 360{\deg} Indoor Environments2026/3/1
- Spherical-GOF: Geometry-Aware Panoramic Gaussian Opacity Fields for 3D Scene Reconstruction2026/3/1
- Can we Trust Unreliable Voxels? Exploring 3D Semantic Occupancy Prediction under Label Noise2026/3/1
- NOVA: Next-step Open-Vocabulary Autoregression for 3D Multi-Object Tracking in Autonomous Driving2026/3/1
- Panoramic Multimodal Semantic Occupancy Prediction for Quadruped Robots2026/3/1
- Not an Obstacle for Dog, but a Hazard for Human: A Co-Ego Navigation System for Guide Dog Robots2026/3/1
- Seeing Beyond: Extrapolative Domain Adaptive Panoramic Segmentation2026/3/1
- Towards Universal Computational Aberration Correction in Photographic Cameras: A Comprehensive Benchmark Analysis2026/3/1
- $M^2$-Occ: Resilient 3D Semantic Occupancy Prediction for Autonomous Driving with Incomplete Camera Inputs2026/3/1
- TRACER: Texture-Robust Affordance Chain-of-Thought for Deformable-Object Refinement2026/1/1
- Learning Fine-Grained Correspondence with Cross-Perspective Perception for Open-Vocabulary 6D Object Pose Estimation2026/1/1
- OneOcc: Semantic Occupancy Prediction for Legged Robots with a Single Panoramic Camera2025/11/1
- OmniTrack++: Omnidirectional Multi-Object Tracking by Learning Large-FoV Trajectory Feedback2025/11/1
- Seeing Clearly and Deeply: An RGBD Imaging Approach with a Bio-inspired Monocentric Design2025/10/1
- RefAtomNet++: Advancing Referring Atomic Video Action Recognition using Semantic Retrieval based Multi-Trajectory Mamba2025/10/1
- EReLiFM: Evidential Reliability-Aware Residual Flow Meta-Learning for Open-Set Domain Generalization under Noisy Labels2025/10/1
- DepTR-MOT: Unveiling the Potential of Depth-Informed Trajectory Refinement for Multi-Object Tracking2025/9/1
- Event-guided 3D Gaussian Splatting for Dynamic Human and Scene Reconstruction2025/9/1
- Segment-to-Act: Label-Noise-Robust Action-Prompted Video Segmentation Towards Embodied Intelligence2025/9/1
- CoBEVMoE: Heterogeneity-aware Feature Fusion with Dynamic Mixture-of-Experts for Collaborative Perception2025/9/1
- UniFucGrasp: Human-Hand-Inspired Unified Functional Grasp Annotation Strategy and Dataset for Diverse Dexterous Hands2025/8/1
- QuaDreamer: Controllable Panoramic Video Generation for Quadruped Robots2025/8/1
- RoHOI: Robustness Benchmark for Human-Object Interaction Detection2025/7/1
- Hallucinating 360{\deg}: Panoramic Street-View Generation via Local Scenes Diffusion and Probabilistic Prompting2025/7/1
- NRSeg: Noise-Resilient Learning for BEV Semantic Segmentation via Driving World Models2025/7/1
- HopaDIFF: Holistic-Partial Aware Fourier Conditioned Diffusion for Referring Human Action Segmentation in Multi-Person Scenarios2025/6/1
- Unlocking Constraints: Source-Free Occlusion-Aware Seamless Segmentation2025/6/1
- Out-of-Distribution Semantic Occupancy Prediction2025/6/1
- Language-Driven Dual Style Mixing for Single-Domain Generalized Object Detection2025/5/1
- Panoramic Out-of-Distribution Segmentation2025/5/1
- Exploring Video-Based Driver Activity Recognition under Noisy Labels2025/4/1
- Detecting Heel Strike and toe off Events Using Kinematic Methods and LSTM Models2025/3/1
- LFX: Towards Unified Light Field Dense Semantic Segmentation and Salient Object Detection2025/3/1
- One-Shot Affordance Grounding of Deformable Objects in Egocentric Organizing Scenes2025/3/1
- Unveiling the Potential of Segment Anything Model 2 for RGB-Thermal Semantic Segmentation with Language Guidance2025/3/1
- Resource-Efficient Affordance Grounding with Complementary Depth and Semantic Prompts2025/3/1
- EgoEvGesture: Gesture Recognition Based on Egocentric Event Camera2025/3/1
- Omnidirectional Multi-Object Tracking2025/3/1
- HierDAMap: Towards Universal Domain Adaptive BEV Mapping via Hierarchical Perspective Priors2025/3/1
- TS-CGNet: Temporal-Spatial Fusion Meets Centerline-Guided Diffusion for BEV Mapping2025/3/1
- DeProPose: Deficiency-Proof 3D Human Pose Estimation via Adaptive Multi-View Fusion2025/2/1
- Event-aided Semantic Scene Completion2025/2/1
- Multi-Keypoint Affordance Representation for Functional Dexterous Grasping2025/2/1
- CT-UIO: Continuous-Time UWB-Inertial-Odometer Localization Using Non-Uniform B-spline with Fewer Anchors2025/2/1
- Benchmarking the Robustness of Optical Flow Estimation to Corruptions2024/11/1
- E-3DGS: Gaussian Splatting with Exposure and Motion Events2024/10/1
- EI-Nexus: Towards Unmediated and Flexible Inter-Modality Local Feature Extraction and Matching for Event-Image Data2024/10/1
- Towards Single-Lens Controllable Depth-of-Field Imaging via Depth-Aware Point Spread Functions2024/9/1
- P2U-SLAM: A Monocular Wide-FoV SLAM System Based on Point Uncertainty and Pose Uncertainty2024/9/1
- GenMapping: Unleashing the Potential of Inverse Perspective Mapping for Robust Online HD Map Construction2024/9/1
- SF-TIM: A Simple Framework for Enhancing Quadrupedal Robot Jumping Agility by Combining Terrain Imagination and Measurement2024/8/1
- Referring Atomic Video Action Recognition2024/7/1
- Learning Granularity-Aware Affordances from Human-Object Interaction for Tool-Based Functional Dexterous Grasping2024/7/1
- Occlusion-Aware Seamless Segmentation2024/7/1
- Label-efficient Semantic Scene Completion with Scribble Annotations2024/5/1
- DTCLMapper: Dual Temporal Consistent Learning for Vectorized HD Map Construction2024/5/1
- Towards Consistent Object Detection via LiDAR-Camera Synergy2024/5/1
- Design, analysis, and manufacturing of a glass-plastic hybrid minimalist aspheric panoramic annular lens2024/5/1
- Exploring Quasi-Global Solutions to Compound Lens Based Computational Imaging Systems2024/4/1
- CFMW: Cross-modality Fusion Mamba for Robust Object Detection under Adverse Weather2024/4/1
- MambaMOS: LiDAR-based 3D Moving Object Segmentation with Motion-aware State Space Model2024/4/1
- Skeleton-Based Human Action Recognition with Noisy Labels2024/3/1
- Representing Domain-Mixing Optical Degradation for Real-World Computational Aberration Correction via Vector Quantization2024/3/1
- Offboard Occupancy Refinement with Hybrid Propagation for Autonomous Driving2024/3/1
- EchoTrack: Auditory Referring Multi-Object Tracking for Autonomous Driving2024/2/1
- Towards Precise 3D Human Pose Estimation with Multi-Perspective Spatial-Temporal Relational Transformers2024/1/1
- LF Tracy: A Unified Single-Pipeline Approach for Salient Object Detection in Light Field Cameras2024/1/1
- Fourier Prompt Tuning for Modality-Incomplete Scene Segmentation2024/1/1
- Navigating Open Set Scenarios for Skeleton-based Action Recognition2023/12/1
- Exploring Event-based Human Pose Estimation with 3D Event Representations2023/11/1
- CoBEV: Elevating Roadside 3D Object Detection with Depth and Height Complementarity2023/10/1
- S$^3$-MonoDETR: Supervised Shape&Scale-perceptive Deformable Transformer for Monocular 3D Object Detection2023/9/1
- Exploring Self-supervised Skeleton-based Action Recognition in Occluded Environments2023/9/1
- Elevating Skeleton-Based Action Recognition with Efficient Multi-Modality Self-Supervision2023/9/1
- FocusFlow: Boosting Key-Points Optical Flow Estimation for Autonomous Driving2023/8/1
- Open Scene Understanding: Grounded Situation Recognition Meets Segment Anything for Helping People with Visual Impairments2023/7/1
- OAFuser: Towards Omni-Aperture Fusion for Light Field Semantic Segmentation2023/7/1
- Towards Anytime Optical Flow Estimation with Event Cameras2023/7/1
- Tightly-Coupled LiDAR-Visual SLAM Based on Geometric Features for Mobile Agents2023/7/1
- LF-PGVIO: A Visual-Inertial-Odometry Framework for Large Field-of-View Cameras using Points and Geodesic Segments2023/6/1
- PVPUFormer: Probabilistic Visual Prompt Unified Transformer for Interactive Image Segmentation2023/6/1
- Towards Source-free Domain Adaptive Semantic Segmentation via Importance-aware and Prototype-contrast Learning2023/6/1
- Bi-Mapper: Holistic BEV Semantic Mapping for Autonomous Driving2023/5/1
- SSD-MonoDETR: Supervised Scale-aware Deformable Transformer for Monocular 3D Object Detection2023/5/1
- Exploring Few-Shot Adaptation for Activity Recognition on Diverse Domains2023/5/1
- PanoVPR: Towards Unified Perspective-to-Equirectangular Visual Place Recognition via Sliding Windows across the Panoramic View2023/3/1
- Towards Activated Muscle Group Estimation in the Wild2023/3/1
- FishDreamer: Towards Fisheye Semantic Completion via Unified Image Outpainting and Segmentation2023/3/1
- MateRobot: Material Recognition in Wearable Robotics for People with Visual Impairments2023/2/1
- LF-VISLAM: A SLAM Framework for Large Field-of-View Cameras with Negative Imaging Plane on Mobile Agents2022/9/1
- Multi-modal Depression Estimation based on Sub-attentional Fusion2022/7/1
- Behind Every Domain There is a Shift: Adapting Distortion-aware Vision Transformers for Panoramic Semantic Segmentation2022/7/1
- Trans4Map: Revisiting Holistic Bird's-Eye-View Mapping from Egocentric Images to Allocentric Semantics with Vision Transformers2022/7/1
- Efficient Human Pose Estimation via 3D Event Point Cloud2022/6/1
- Panoramic Panoptic Segmentation: Insights Into Surrounding Parsing for Mobile Agents via Unsupervised Contrastive Learning2022/6/1
- Review on Panoramic Imaging and Its Applications in Scene Understanding2022/5/1
- Indoor Navigation Assistance for Visually Impaired People via Dynamic SLAM and Panoptic Segmentation with an RGB-D Sensor2022/4/1
- MatchFormer: Interleaving Attention in Transformers for Feature Matching2022/3/1
- CMX: Cross-Modal Fusion for RGB-X Semantic Segmentation with Transformers2022/3/1
- TransDARC: Transformer-based Driver Activity Recognition with Latent Space Feature Calibration2022/3/1
- Towards Robust Semantic Segmentation of Accident Scenes via Multi-Source Mixed Sampling and Meta-Learning2022/3/1
- Bending Reality: Distortion-aware Transformers for Adapting to Panoramic Semantic Segmentation2022/3/1
- LF-VIO: A Visual-Inertial-Odometry Framework for Large Field-of-View Cameras with Negative Plane2022/2/1
- Delving Deep into One-Shot Skeleton-based Action Recognition with Diverse Occlusions2022/2/1
- CSFlow: Learning Optical Flow via Cross Strip Correlation for Autonomous Driving2022/2/1
- PanoFlow: Learning 360{\deg} Optical Flow for Surrounding Temporal Understanding2022/2/1
- TransKD: Transformer Knowledge Distillation for Efficient Semantic Segmentation2022/2/1
- Transfer beyond the Field of View: Dense Panoramic Semantic Segmentation via Unsupervised Domain Adaptation2021/10/1
- Trans4Trans: Efficient Transformer for Transparent Object and Semantic Scene Segmentation in Real-World Navigation Assistance2021/8/1
- Flying Guide Dog: Walkable Path Discovery for the Visually Impaired Utilizing Drones and Transformer-based Semantic Segmentation2021/8/1
- Panoramic Depth Estimation via Supervised and Unsupervised Learning in Indoor Scenes2021/8/1
- DensePASS: Dense Panoramic Semantic Segmentation via Unsupervised Domain Adaptation with Attention-Augmented Context Exchange2021/8/1
- MASS: Multi-Attentional Semantic Segmentation of LiDAR Data for Dense Top-View Understanding2021/7/1
- Trans4Trans: Efficient Transformer for Transparent Object Segmentation to Help Visually Impaired People Navigate in the Real World2021/7/1
- HIDA: Towards Holistic Indoor Understanding for the Visually Impaired via Semantic Instance Segmentation with a Wearable Solid-State LiDAR Sensor2021/7/1
- Aerial-PASS: Panoramic Annular Scene Segmentation in Drone Videos2021/5/1
- Perception Framework through Real-Time Semantic Segmentation and Scene Recognition on a Wearable System for the Visually Impaired2021/3/1
- Panoramic Panoptic Segmentation: Towards Complete Surrounding Understanding via Unsupervised Contrastive Learning2021/3/1
- DR-TANet: Dynamic Receptive Temporal Attention Network for Street Scene Change Detection2021/3/1
- Capturing Omni-Range Context for Omnidirectional Segmentation2021/3/1
- Panoptic Lintention Network: Towards Efficient Navigational Perception for the Visually Impaired2021/3/1
- Panoramic annular SLAM with loop closure and global optimization2021/2/1
- Polarization-driven Semantic Segmentation via Efficient Attention-bridged Fusion2020/11/1
- Real-time Fusion Network for RGB-D Semantic Segmentation Incorporating Unexpected Obstacle Detection for Road-driving Images2020/2/1
- DS-PASS: Detail-Sensitive Panoramic Annular Semantic Segmentation through SwaftNet for Surrounding Sensing2019/9/1