Abdelrahman Abdallah retweeted
DSPy 3.4.0 out, with: (1) native support for Jev and System One models - in the timeless DSPy syntax. (2) a brand new optimizer, ReAnchor, specifically for calibrating outputs with confidence. (3) lightning-fast import speeds for the LLM abstraction, via sibling library LM15)
DSPy 3.4.0 was just released! This release includes native support for Jev and System one models inside of DSPy! Use it with compatible signatures. This release also includes a brand new optimizer, ReAnchor, specifically for calibrating outputs with confidence.
9
32
202
11,348
Abdelrahman Abdallah retweeted
SmallReason-ColBERT: An Ultra-Small Late-Interaction Retriever for Reasoning Intensive Retrieval @perdactor et al. present a 32M retriever for reasoning-heavy queries, adding a learned token-weighting head. ๐Ÿ“arxiv.org/abs/2609.29652 ๐Ÿ‘จ๐Ÿฝโ€๐Ÿ’ปgithub.com/DataScienceUIBK/Sโ€ฆ
18
122
5,170
Abdelrahman Abdallah retweeted
OBLIQ-IR: Training a Dense Retriever for Oblique Queries @perdactor et al. train a dense retriever for oblique queries, where relevance depends on latent traits like stance or style rather than topic overlap. ๐Ÿ“ arxiv.org/abs/2609.29649 ๐Ÿ‘จ๐Ÿฝโ€๐Ÿ’ป github.com/DataScienceUIBK/oโ€ฆ
1
5
33
2,073
Abdelrahman Abdallah retweeted
EXCISE: Query-Side Exclusion for Late-Interaction Retrieval Presents a query-time fix for "exclusion inversion" in ColBERT retrievers, where excluded topics get ranked higher, not lower. ๐Ÿ“ arxiv.org/abs/2608.05497
4
13
847
Abdelrahman Abdallah retweeted
imo a significant part of the problem is the insistence of the community on misnaming the paradigm as โ€œmulti-vector retrievalโ€, we called it late interaction for a reason the problem isnโ€™t in having one vector; the problem is in the scoring function! piped.video/Z2TmdcylyEc?si=hrtqโ€ฆ
5
5
79
5,358
Abdelrahman Abdallah retweeted
most interesting thing happening rn is that you are allowed to use ai for everything. everything except writing. if you use ai for writing, they will unalive you in public
62
24
846
35,980
๐ŸŽ‰ Our lab has 10 papers accepted at #EMNLP2026 โ€” 9 Main Conference + 1 Findings! This is ~1.5 years of work, and almost none of it landed on the first try. Several of these papers went through 4โ€“6 submission cycles before they were accepted. Rejection was part of the process, not the end of it. Main Conference 1๏ธโƒฃ OBLIQ-IR: Training a Dense Retriever for Oblique Queries Mahmoud Abdalla, ๐—”๐—ฏ๐—ฑ๐—ฒ๐—น๐—ฟ๐—ฎ๐—ต๐—บ๐—ฎ๐—ป ๐—”๐—ฏ๐—ฑ๐—ฎ๐—น๐—น๐—ฎ๐—ต, Shaimaa Sedek, Adam Jatowt ๐Ÿ“„ Coming soon ยท ๐Ÿ’ป Coming soon 2๏ธโƒฃ MEMORA: A Memory-Enhanced Multimodal Committee Reranking Agent for Reasoning-Intensive Retrieval Mohamed Mahmoud, Mostafa Farouk Senussi, Mahmoud Abdalla, Mahmoud SalahEldin Kasem, ๐—”๐—ฏ๐—ฑ๐—ฒ๐—น๐—ฟ๐—ฎ๐—ต๐—บ๐—ฎ๐—ป ๐—”๐—ฏ๐—ฑ๐—ฎ๐—น๐—น๐—ฎ๐—ต, Hyun Soo Kang ๐Ÿ“„ Coming soon ยท ๐Ÿ’ป Coming soon 3๏ธโƒฃ SustainableQA: A Comprehensive Question Answering Dataset for Corporate Sustainability and EU Taxonomy Reporting Mohammed Ali, ๐—”๐—ฏ๐—ฑ๐—ฒ๐—น๐—ฟ๐—ฎ๐—ต๐—บ๐—ฎ๐—ป ๐—”๐—ฏ๐—ฑ๐—ฎ๐—น๐—น๐—ฎ๐—ต, Adam Jatowt ๐Ÿ“„ arxiv.org/abs/2508.03000 ยท ๐Ÿ’ป github.com/DataScienceUIBK/Sโ€ฆ 4๏ธโƒฃ TEMPO: Realistic Multi-Domain Benchmark for Temporal Reasoning-Intensive Retrieval ๐—”๐—ฏ๐—ฑ๐—ฒ๐—น๐—ฟ๐—ฎ๐—ต๐—บ๐—ฎ๐—ป ๐—”๐—ฏ๐—ฑ๐—ฎ๐—น๐—น๐—ฎ๐—ต, Mohammed Ali, Muhammad Abdul-Mageed, Adam Jatowt ๐Ÿ“„ arxiv.org/abs/2601.09523 ยท ๐ŸŒ tempo-bench.github.io 5๏ธโƒฃ Argus-Retriever: Vision-LLM Late-Interaction Retrieval with Region-Aware Query-Conditioned MoE for Visual Document Retrieval ๐—”๐—ฏ๐—ฑ๐—ฒ๐—น๐—ฟ๐—ฎ๐—ต๐—บ๐—ฎ๐—ป ๐—”๐—ฏ๐—ฑ๐—ฎ๐—น๐—น๐—ฎ๐—ต, Mahmoud Abdalla, Mohammed Ali, Adam Jatowt ๐Ÿ“„ arxiv.org/abs/2606.04300 ยท ๐Ÿ’ป github.com/DataScienceUIBK/Aโ€ฆ 6๏ธโƒฃ Large Language Models Systematically Favor Popular Options: Evidence and Mitigation Across MCQs ๐—”๐—ฏ๐—ฑ๐—ฒ๐—น๐—ฟ๐—ฎ๐—ต๐—บ๐—ฎ๐—ป ๐—”๐—ฏ๐—ฑ๐—ฎ๐—น๐—น๐—ฎ๐—ต, Mohammed Ali, Bhawna Piryani, Mahmoud Abdalla, Adam Jatowt ๐Ÿ“„ Coming soon ยท ๐Ÿ’ป Coming soon 7๏ธโƒฃ SmallReason-ColBERT: An Ultra-Small Late-Interaction Retriever for Reasoning-Intensive Retrieval ๐—”๐—ฏ๐—ฑ๐—ฒ๐—น๐—ฟ๐—ฎ๐—ต๐—บ๐—ฎ๐—ป ๐—”๐—ฏ๐—ฑ๐—ฎ๐—น๐—น๐—ฎ๐—ต, Mohammed Ali, Adam Jatowt ๐Ÿ“„ Coming soon ยท ๐Ÿ’ป Coming soon 8๏ธโƒฃ MoCA-Agent: A Market-of-Claims Code Agent for Financial and Numerical Reasoning ๐—”๐—ฏ๐—ฑ๐—ฒ๐—น๐—ฟ๐—ฎ๐—ต๐—บ๐—ฎ๐—ป ๐—”๐—ฏ๐—ฑ๐—ฎ๐—น๐—น๐—ฎ๐—ต, AbdelRahim A. Elmadany, Sameh Al Natour, Hasan Cavusoglu, Adam Jatowt, Muhammad Abdul-Mageed ๐Ÿ“„ arxiv.org/abs/2606.11537 ยท ๐Ÿ’ป github.com/UBC-NLP/MoCA-Agenโ€ฆ 9๏ธโƒฃ The Magnitude Mirage: Rethinking Confidence for Reasoning-Intensive Retrieval Jamie Holdcroft, ๐—”๐—ฏ๐—ฑ๐—ฒ๐—น๐—ฟ๐—ฎ๐—ต๐—บ๐—ฎ๐—ป ๐—”๐—ฏ๐—ฑ๐—ฎ๐—น๐—น๐—ฎ๐—ต, Adam Jatowt ๐Ÿ“„ Coming soon ยท ๐Ÿ’ป Coming soon Findings ๐Ÿ”Ÿ REGREACT: Self-Correcting Multi-Agent Pipelines for Structured Regulatory Information Extraction Mohammed Ali, ๐—”๐—ฏ๐—ฑ๐—ฒ๐—น๐—ฟ๐—ฎ๐—ต๐—บ๐—ฎ๐—ป ๐—”๐—ฏ๐—ฑ๐—ฎ๐—น๐—น๐—ฎ๐—ต, Adam Jatowt ๐Ÿ“„ arxiv.org/abs/2604.12054 ยท ๐Ÿ’ป Coming soon Huge congratulations to everyone in the lab, and thanks to our supervisor @adammo . To anyone sitting on a pile of rejections right now: keep resubmitting. It works. ๐Ÿ™
3
445
1/๐Ÿ‘๏ธ Meet Argus-Retriever: the first late-interaction visual doc retriever where the document representation adapts to the query: D(q). - 86.0 NDCG@5 on ViDoRe V1+V2 (SOTA open model) - ๐Ÿ“ฆ 1024-dim head, 4.5ร— smaller index
Argus-Retriever: Vision-LLM Late-Interaction Retrieval with Region-Aware Query-Conditioned MoE for Visual Document Retrieval Introduces a query-conditioned late-interaction visual document retriever. ๐Ÿ“ arxiv.org/abs/2606.04300 ๐Ÿ‘จ๐Ÿฝโ€๐Ÿ’ป github.com/DataScienceUIBK/Aโ€ฆ
1
7
53
8,711
2/ The idea: a region-aware Mixture-of-Experts inside the document encoder. A router reads each region's content, its 2D position, and a pooled query context z_q, then mixes 4 latent experts (+1 shared). The page is now encoded differently per query โ†’ D(q). Still MaxSim.
2
318
Abdelrahman Abdallah retweeted
Argus-Retriever: Vision-LLM Late-Interaction Retrieval with Region-Aware Query-Conditioned MoE for Visual Document Retrieval Introduces a query-conditioned late-interaction visual document retriever. ๐Ÿ“ arxiv.org/abs/2606.04300 ๐Ÿ‘จ๐Ÿฝโ€๐Ÿ’ป github.com/DataScienceUIBK/Aโ€ฆ
1
5
20
9,583
Thrilled to share two paper acceptances! ๐ŸŽ‰ ๐Ÿ“„ MM-BRIGHT โ€” Multimodal Benchmark for Reasoning-Intensive Retrieval โ†’ #KDD2026 ๐Ÿ› ๏ธ Rankify โ€” Python toolkit for Retrieval, Re-Ranking & RAG (670+ โญ) โ†’ #ACL2026 Demo Thanks to @ supervisor & collaborators ๐Ÿš€ #NLP #RAG #IR
3
7
52
2,901
Abdelrahman Abdallah retweeted
another late interaction w
[1/3] tiny 32M reasoning Colber retriever, just dropped. ๐Ÿค— Reason-mxbai-colbert-v0-32m a late-interaction ColBERT fine-tuned for BRIGHT-style retrieval. 19.0 avg nDCG@10 on BRIGHT, ~5x smaller than Reason-ModernColBERT. huggingface.co/DataScience-Uโ€ฆ
1
2
14
1,479
[1/3] tiny 32M reasoning Colber retriever, just dropped. ๐Ÿค— Reason-mxbai-colbert-v0-32m a late-interaction ColBERT fine-tuned for BRIGHT-style retrieval. 19.0 avg nDCG@10 on BRIGHT, ~5x smaller than Reason-ModernColBERT. huggingface.co/DataScience-Uโ€ฆ
4
14
92
7,403
[1/3] Built on @mixedbreadai's mxbai-edge-colbert-v0-32m with one key change: we widened the projection head from 64โ†’128 dim using small-random init (zero-pad = dead neurons under L2 MaxSim, classic gotcha).
1
6
469
[1/3] Still limited on splits (leetcode/theorem) โ€” base model's sans_pos + lowercased tokeniser caps code retrieval. But Imporved Pony split from 8.0 NDCG@10 by Reason-ModernColBERT โ†’ 20 NDCG@10
4
366
Abdelrahman Abdallah retweeted
For the people not on Reddit, yesterday ICLR desk-rejected an accepted oral paper from a US-sanctioned institution. The conference is in 4 days; this could probably have been done better. openreview.net/forum?id=w2tnโ€ฆ
7
21
179
50,626
I am happy to announce that this week was great news. 2 accept paper main ACL 2026 1 findings ACL 2026 1 Sigir 2026 #acl #acl2026 #sigir
1
1
26
1,515
Abdelrahman Abdallah retweeted
The paper acceptance notifications will be out by the 6th of April, AoE. The PCs are working hard throughout the holiday season to finalize the decisions. Apologies for the delay!
35
19
137
43,685
#acl 2026 Is SAC still sleeping, or what is the problem?
3
16
3,656
Ss from rednote #acl
1
12
2,450