Spinons, solitons and random singlets in the spin-chain compound copper benzoate

Kavli Affiliate: Long Zhang | First 5 Authors: Ying Chen, Ying Chen, , , | Summary: The $S=1/2$ antiferromagnetic Heisenberg chain is a paradigmatic quantum system hosting exotic excitations such as spinons and solitons, and forming random singlet state in the presence of quenched disorder. Realizing and distinguishing these excitations in a single material remains […]


Continue.. Spinons, solitons and random singlets in the spin-chain compound copper benzoate

R2RGEN: Real-to-Real 3D Data Generation for Spatially Generalized Manipulation

Kavli Affiliate: Zheng Zhu | First 5 Authors: Xiuwei Xu, Xiuwei Xu, , , | Summary: Towards the aim of generalized robotic manipulation, spatial generalization is the most fundamental capability that requires the policy to work robustly under different spatial distribution of objects, environment and agent itself. To achieve this, substantial human demonstrations need to […]


Continue.. R2RGEN: Real-to-Real 3D Data Generation for Spatially Generalized Manipulation

Dimension- and Facet-Dependent Altermagnetic Triferroics and Biferroics in CrSb

Kavli Affiliate: Long Zhang | First 5 Authors: Long Zhang, Long Zhang, , , | Summary: Altermagnets have recently garnered significant interest due to their vanishing net magnetic moment and non-relativistic momentum-dependent spin splitting. However, altermagnetic (AM) multiferroics especially triferroics remain scarce. We investigate the experimentally synthesized non-van der Waals CrSb as a model system […]


Continue.. Dimension- and Facet-Dependent Altermagnetic Triferroics and Biferroics in CrSb

Anomalous proteinaceous shells with octagonal local order

Kavli Affiliate: Rudolf Podgornik | Summary:Proteinaceous shells useful for various biomedical applications exhibit a wide range of anomalous structures that are fundamentally different from icosahedral viral capsids described by the Caspar-Klug paradigmatic model. Exploring the Protein Data Bank, we have identified nine different types of anomalous shells structurally close to flat octagonal quasicrystals. As we […]


Continue.. Anomalous proteinaceous shells with octagonal local order

Character Mixing for Video Generation

Kavli Affiliate: Yi Zhou | First 5 Authors: Tingting Liao, Tingting Liao, , , | Summary: Imagine Mr. Bean stepping into Tom and Jerry–can we generate videos where characters interact naturally across different worlds? We study inter-character interaction in text-to-video generation, where the key challenge is to preserve each character’s identity and behaviors while enabling […]


Continue.. Character Mixing for Video Generation

Multi-Modal Multi-Task Semantic Communication: A Distributed Information Bottleneck Perspective

Kavli Affiliate: Cheng Peng | First 5 Authors: Yujie Zhou, Yujie Zhou, , , | Summary: Semantic communication (SemCom) shifts the focus from data transmission to meaning delivery, enabling efficient and intelligent communication. Existing AI-based coding schemes for multi-modal multi-task SemCom often require transmitters with full-modal data to participate in all receivers’ tasks, which leads […]


Continue.. Multi-Modal Multi-Task Semantic Communication: A Distributed Information Bottleneck Perspective

Multi-Modal Multi-Task Semantic Communication: A Distributed Information Bottleneck Perspective

Kavli Affiliate: Cheng Peng | First 5 Authors: Yujie Zhou, Yujie Zhou, , , | Summary: Semantic communication (SemCom) shifts the focus from data transmission to meaning delivery, enabling efficient and intelligent communication. Existing AI-based coding schemes for multi-modal multi-task SemCom often require transmitters with full-modal data to participate in all receivers’ tasks, which leads […]


Continue.. Multi-Modal Multi-Task Semantic Communication: A Distributed Information Bottleneck Perspective

Taming Text-to-Sounding Video Generation via Advanced Modality Condition and Interaction

Kavli Affiliate: Cheng Peng | First 5 Authors: Kaisi Guan, Kaisi Guan, , , | Summary: This study focuses on a challenging yet promising task, Text-to-Sounding-Video (T2SV) generation, which aims to generate a video with synchronized audio from text conditions, meanwhile ensuring both modalities are aligned with text. Despite progress in joint audio-video training, two […]


Continue.. Taming Text-to-Sounding Video Generation via Advanced Modality Condition and Interaction

VLA-R1: Enhancing Reasoning in Vision-Language-Action Models

Kavli Affiliate: Zheng Zhu | First 5 Authors: Angen Ye, Angen Ye, , , | Summary: Vision-Language-Action (VLA) models aim to unify perception, language understanding, and action generation, offering strong cross-task and cross-scene generalization with broad impact on embodied AI. However, current VLA models often lack explicit step-by-step reasoning, instead emitting final actions without considering […]


Continue.. VLA-R1: Enhancing Reasoning in Vision-Language-Action Models

EvoWorld: Evolving Panoramic World Generation with Explicit 3D Memory

Kavli Affiliate: Cheng Peng | First 5 Authors: Jiahao Wang, Jiahao Wang, , , | Summary: Humans possess a remarkable ability to mentally explore and replay 3D environments they have previously experienced. Inspired by this mental process, we present EvoWorld: a world model that bridges panoramic video generation with evolving 3D memory to enable spatially […]


Continue.. EvoWorld: Evolving Panoramic World Generation with Explicit 3D Memory