9.0CLJun 22
KaLM-Reranker-V1: Fast but Not Late Interaction for Compressed Document RerankingXinping Zhao, Jiaxin Xu, Ziqi Dai et al.
As retrieval systems scale, high-quality reranking becomes increasingly important. However, most existing rerankers, whether encoder-based or decoder-based, jointly encode the query and passage, tightly coupling their computation and limiting deployment efficiency as well as flexibility. We present KaLM-Reranker-V1, a fast but not late-interaction (FBNL) reranker that decouples query and passage computation while retaining expressive relevance modeling. Built on an encoder-decoder architecture, KaLM-Reranker-V1 uses the encoder to pre-encode passages with Matryoshka embedding pooling, while the decoder models the system instruction, user instruction, and query intent; cross-attention then captures relevance between the query context and passage representations. This design makes KaLM-Reranker-V1 efficient through decoupled passage encoding, yet not late interaction, by preserving rich relevance modeling through cross-attention. We instantiate KaLM-Reranker-V1 in three sizes, Nano, Small, and Large, with 0.27B, 1B, and 4B activated parameters, respectively. Extensive experiments on BEIR, MIRACL, and LMEB demonstrate that KaLM-Reranker-V1 achieves strong reranking performance with superior efficiency. On BEIR, KaLM-Reranker-V1 achieves state-of-the-art performance, on par with strong industrial models such as the Qwen3-Reranker series; on MIRACL, despite not being extensively trained on multilingual data, KaLM-Reranker-V1 still shows excellent reranking performance. Moreover, on LMEB, reranking models demonstrate a clear advantage, with even the 0.27B Nano model remaining competitive with 7-12B embedding models.
7.4SOC-PHJun 20
Perceiving exposure segregation with open urban imageryYunke Zhang, Ruolong Ma, Xin Zhang et al.
Socioeconomic exposure segregation -- the lack of daily interaction between income groups -- erodes social capital and entrenches inequality, yet the specific physical features that drive these behavioral restrictions remain poorly understood. Prior research has quantified where segregation occurs using mobility data, but has not identified how the built environment facilitates or inhibits these interactions. Here we introduce VISAGE, a large multi-modal model-enabled framework that perceives exposure segregation directly from open satellite and street-level imagery across 10,030 communities in 31 U.S. cities. Moving beyond black-box correlations, we operationalize cross-disciplinary sociological theory into an interpretable visual codebook to detect physical regulators of social mixing. We find that the built environment encodes a legible grammar of segregation: "defensible" architectural forms (e.g., fences, gated enclosures) and monofunctional zoning systematically predict higher social isolation, whereas mixed-use infrastructure fosters interaction, explaining substantial variance in mobility-derived segregation patterns (Pearson $r=0.770$). Crucially, we show that inclusionary housing policies manifest in distinct visual signatures associated with higher mixing, suggesting that policy interventions successfully alter the physical landscape to encourage diversity. Our findings offer a scalable pathway to decipher the social production of space, providing a mechanism-based lens to understand how the built environment shapes social behavior.