Hierarchical Self-Prompting SAM: A Prompt-Free Medical Image Segmentation Framework

Kavli Affiliate: Jing Wang | First 5 Authors: Mengmeng Zhang, Xingyuan Dai, Yicheng Sun, Jing Wang, Yueyang Yao | Summary: Although the Segment Anything Model (SAM) is highly effective in natural image segmentation, it requires dependencies on prompts, which limits its applicability to medical imaging where manual prompts are often unavailable. Existing efforts to fine-tune […]


Continue.. Hierarchical Self-Prompting SAM: A Prompt-Free Medical Image Segmentation Framework

Low-velocity precessing jets can explain observed morphologies in the Twin Radio Galaxy TRG J104454+354055

Kavli Affiliate: Luis C. Ho | First 5 Authors: Santanu Mondal, Gourab Giri, Ravi Joshi, Paul J. Wiita, Gopal-Krishna | Summary: Our understanding of large-scale radio jets in merger systems has been drastically improved in the era of VLA, VLBA/EVN, uGMRT, and MeerKAT. Twin Radio Galaxies (TRGs) are the rare interacting galaxy pairs where both […]


Continue.. Low-velocity precessing jets can explain observed morphologies in the Twin Radio Galaxy TRG J104454+354055

Photometric redshift estimation for emission line galaxies of DESI Legacy Imaging Surveys by CNN-MLP

Kavli Affiliate: Xuebing Wu | Summary:Emission Line Galaxies (ELGs) are crucial for cosmological studies, particularly in understanding the large-scale structure of the Universe and the role of dark energy. ELGs form an essential component of the target catalogue for the Dark Energy Spectroscopic Instrument (DESI), a major astronomical survey. However, the accurate selection of ELGs […]


Continue.. Photometric redshift estimation for emission line galaxies of DESI Legacy Imaging Surveys by CNN-MLP

VidText: Towards Comprehensive Evaluation for Video Text Understanding

Kavli Affiliate: Jing Wang | First 5 Authors: Zhoufaran Yang, Zhoufaran Yang, , , | Summary: Visual texts embedded in videos carry rich semantic information, which is crucial for both holistic video understanding and fine-grained reasoning about local human actions. However, existing video understanding benchmarks largely overlook textual information, while OCR-specific benchmarks are constrained to […]


Continue.. VidText: Towards Comprehensive Evaluation for Video Text Understanding

Soliton resolution for the coupled complex short pulse equation

Kavli Affiliate: Ran Wang | First 5 Authors: Nan Liu, Ran Wang, , , | Summary: We address the long-time asymptotics of the solution to the Cauchy problem of ccSP (coupled complex short pulse) equation on the line for decaying initial data that can support solitons. The ccSP system describes ultra-short pulse propagation in optical […]


Continue.. Soliton resolution for the coupled complex short pulse equation

Spa-VLM: Stealthy Poisoning Attacks on RAG-based VLM

Kavli Affiliate: Jing Wang | First 5 Authors: Lei Yu, Yechao Zhang, Ziqi Zhou, Yang Wu, Wei Wan | Summary: With the rapid development of the Vision-Language Model (VLM), significant progress has been made in Visual Question Answering (VQA) tasks. However, existing VLM often generate inaccurate answers due to a lack of up-to-date knowledge. To […]


Continue.. Spa-VLM: Stealthy Poisoning Attacks on RAG-based VLM

Revisiting the Li abundances of Stars with and without Detected Planets from the High Resolution Spectroscopy

Kavli Affiliate: Subo Dong | Summary:Whether the presence of planets affects the lithium (Li) abundance of their host stars is still an open question. To investigate the difference of the Li abundance between planet-host stars (HS) and isolated stars (IS) with no detected planets, we analyze a large sample of stars with temperatures ranging from […]


Continue.. Revisiting the Li abundances of Stars with and without Detected Planets from the High Resolution Spectroscopy

Multi-objective Large Language Model Alignment with Hierarchical Experts

Kavli Affiliate: Zhuo Li | First 5 Authors: Zhuo Li, Guodong Du, Weiyang Guo, Yigeng Zhou, Xiucheng Li | Summary: Aligning large language models (LLMs) to simultaneously satisfy multiple objectives remains a significant challenge, especially given the diverse and often conflicting nature of human preferences. Existing alignment methods struggle to balance trade-offs effectively, often requiring […]


Continue.. Multi-objective Large Language Model Alignment with Hierarchical Experts