view article Article Native-speed vLLM transformers modeling backend hmellor, lysandre • 17 days ago • 58
RuleChef: Grounding LLM Task Knowledge in Human-Editable Rules Paper • 2607.01293 • Published 24 days ago • 3
LettuceDetect v2 Collection SOTA hallucination detection for agentic workflows, multilingual, long context • 5 items • Updated 19 days ago • 1
view article Article Introducing North Mini Code: Cohere’s First Model For Developers CohereLabs • Jun 9 • 83
view article Article ACL-Verbatim: Hallucination-Free Question Answering for NLP Researchers adaamko • Jun 9 • 2
Latent Terms: Dense Retrievers Contain Trivially Extractable BM25-ready Zipfian Vocabularies Paper • 2605.29384 • Published May 28 • 2
ACL-Verbatim: hallucination-free question answering for research Paper • 2605.21102 • Published May 20 • 8
Verbatim RAG v1 Collection Hallucination free RAG and out SOTA state-of-the-art extractors • 8 items • Updated Jun 2 • 9
LettuceDetect: A Hallucination Detection Framework for RAG Applications Paper • 2502.17125 • Published Feb 24, 2025 • 14
view article Article Squeez: Task-Conditioned Tool-Output Pruning for Coding Agents adaamko • Apr 10 • 3
view article Article Training and Finetuning Multimodal Embedding & Reranker Models with Sentence Transformers tomaarsen • Apr 16 • 75
view article Article Multimodal Embedding & Reranker Models with Sentence Transformers tomaarsen • Apr 9 • 67
Squeez: Task-Conditioned Tool-Output Pruning for Coding Agents Paper • 2604.04979 • Published Apr 4 • 11
view article Article Welcome Gemma 4: Frontier multimodal intelligence on device +5 merve, pcuenq, sergiopaniego, burtenshaw, Steveeeeeeen, alvarobartt, SaylorTwift • Apr 2 • 919
view article Article How We Built a Semantic Highlight Model To Save Token Cost for RAG zilliz • Jan 15 • 67