E-MoFlow: Learning Egomotion and Optical Flow from Event Data via Implicit Regularization

Kavli Affiliate: Yi Zhou | First 5 Authors: Wenpu Li, Wenpu Li, , , | Summary: The estimation of optical flow and 6-DoF ego-motion, two fundamental tasks in 3D vision, has typically been addressed independently. For neuromorphic vision (e.g., event cameras), however, the lack of robust data association makes solving the two problems separately an […]


Continue.. E-MoFlow: Learning Egomotion and Optical Flow from Event Data via Implicit Regularization

E-MoFlow: Learning Egomotion and Optical Flow from Event Data via Implicit Regularization

Kavli Affiliate: Yi Zhou | First 5 Authors: Wenpu Li, Wenpu Li, , , | Summary: The estimation of optical flow and 6-DoF ego-motion, two fundamental tasks in 3D vision, has typically been addressed independently. For neuromorphic vision (e.g., event cameras), however, the lack of robust data association makes solving the two problems separately an […]


Continue.. E-MoFlow: Learning Egomotion and Optical Flow from Event Data via Implicit Regularization

Spinons, solitons and random singlets in the spin-chain compound copper benzoate

Kavli Affiliate: Long Zhang | First 5 Authors: Ying Chen, Ying Chen, , , | Summary: The $S=1/2$ antiferromagnetic Heisenberg chain is a paradigmatic quantum system hosting exotic excitations such as spinons and solitons, and forming random singlet state in the presence of quenched disorder. Realizing and distinguishing these excitations in a single material remains […]


Continue.. Spinons, solitons and random singlets in the spin-chain compound copper benzoate

R2RGEN: Real-to-Real 3D Data Generation for Spatially Generalized Manipulation

Kavli Affiliate: Zheng Zhu | First 5 Authors: Xiuwei Xu, Xiuwei Xu, , , | Summary: Towards the aim of generalized robotic manipulation, spatial generalization is the most fundamental capability that requires the policy to work robustly under different spatial distribution of objects, environment and agent itself. To achieve this, substantial human demonstrations need to […]


Continue.. R2RGEN: Real-to-Real 3D Data Generation for Spatially Generalized Manipulation

Dimension- and Facet-Dependent Altermagnetic Triferroics and Biferroics in CrSb

Kavli Affiliate: Long Zhang | First 5 Authors: Long Zhang, Long Zhang, , , | Summary: Altermagnets have recently garnered significant interest due to their vanishing net magnetic moment and non-relativistic momentum-dependent spin splitting. However, altermagnetic (AM) multiferroics especially triferroics remain scarce. We investigate the experimentally synthesized non-van der Waals CrSb as a model system […]


Continue.. Dimension- and Facet-Dependent Altermagnetic Triferroics and Biferroics in CrSb

Anomalous proteinaceous shells with octagonal local order

Kavli Affiliate: Rudolf Podgornik | Summary:Proteinaceous shells useful for various biomedical applications exhibit a wide range of anomalous structures that are fundamentally different from icosahedral viral capsids described by the Caspar-Klug paradigmatic model. Exploring the Protein Data Bank, we have identified nine different types of anomalous shells structurally close to flat octagonal quasicrystals. As we […]


Continue.. Anomalous proteinaceous shells with octagonal local order

Character Mixing for Video Generation

Kavli Affiliate: Yi Zhou | First 5 Authors: Tingting Liao, Tingting Liao, , , | Summary: Imagine Mr. Bean stepping into Tom and Jerry–can we generate videos where characters interact naturally across different worlds? We study inter-character interaction in text-to-video generation, where the key challenge is to preserve each character’s identity and behaviors while enabling […]


Continue.. Character Mixing for Video Generation

Multi-Modal Multi-Task Semantic Communication: A Distributed Information Bottleneck Perspective

Kavli Affiliate: Cheng Peng | First 5 Authors: Yujie Zhou, Yujie Zhou, , , | Summary: Semantic communication (SemCom) shifts the focus from data transmission to meaning delivery, enabling efficient and intelligent communication. Existing AI-based coding schemes for multi-modal multi-task SemCom often require transmitters with full-modal data to participate in all receivers’ tasks, which leads […]


Continue.. Multi-Modal Multi-Task Semantic Communication: A Distributed Information Bottleneck Perspective

Multi-Modal Multi-Task Semantic Communication: A Distributed Information Bottleneck Perspective

Kavli Affiliate: Cheng Peng | First 5 Authors: Yujie Zhou, Yujie Zhou, , , | Summary: Semantic communication (SemCom) shifts the focus from data transmission to meaning delivery, enabling efficient and intelligent communication. Existing AI-based coding schemes for multi-modal multi-task SemCom often require transmitters with full-modal data to participate in all receivers’ tasks, which leads […]


Continue.. Multi-Modal Multi-Task Semantic Communication: A Distributed Information Bottleneck Perspective

Taming Text-to-Sounding Video Generation via Advanced Modality Condition and Interaction

Kavli Affiliate: Cheng Peng | First 5 Authors: Kaisi Guan, Kaisi Guan, , , | Summary: This study focuses on a challenging yet promising task, Text-to-Sounding-Video (T2SV) generation, which aims to generate a video with synchronized audio from text conditions, meanwhile ensuring both modalities are aligned with text. Despite progress in joint audio-video training, two […]


Continue.. Taming Text-to-Sounding Video Generation via Advanced Modality Condition and Interaction