arXiv cs.AI / cs.LG / cs.CL·4d agoThe Sirens' Song: When Proximal Background Context Overshadows Distant Evidence#long-context#llm#lyraAI research
Hugging Face daily papers·5d agoKnowledge Pull Requests for Continual Document Authoring#knowledge-pull-requests#document-authoring#wikipedia
arXiv cs.AI / cs.LG / cs.CL·5d agoThe Copy Ceiling: An Input-Exposure Control for Ontology-Grounded Generation over Curated Corpora#evaluation#rag#copy-ceilingAI research
Hugging Face daily papers·6d agoJev-Mem: System-One-Controlled Agentic Memory for Efficient AI Agents#agentic-memory#jev-mem#locomo
Hugging Face daily papers·6d agoOvis-Embedding: Pushing the Frontiers of Universal Omni-Modal Embeddings#ovis-embedding#embeddings#omni-modal
MarkTechPost·8d agoLinkup Research Releases SPARSEUP: A 149M-Parameter Open-Source Sparse Embedding Model#beir#linkup#modernbert 4 min1
arXiv cs.AI / cs.LG / cs.CL·8d agoPredictable Failure in Multi-Hop Retrieval: Score-Distributional Confidence Scoring and Abstention#retrieval#multi-hop#ragAI research
arXiv cs.CR·8d agoCIPL: A Channel-Aware Framework for Recoverable Privacy Leakage in LLM Agents#llm-agents#privacy-leakage#evaluationAI safety & security1
arXiv cs.CR·8d agoMicro-Collaborative Poisoning: A Distributed Attack on RAG Systems#rag#data-poisoning#llmAI safety & security
arXiv cs.AI / cs.LG / cs.CL·9d agoRAFT: A Stateful Retrieval-Augmented Framework for Troubleshooting Agents#apache-jira#enterprise-support#graphragAI research1
MarkTechPost·10d agoGoogle Research Introduces Retrieve-for-Train (R4T): An RL-Compiled Diffusion Retriever for 12× to 20× Faster Query Fan-Out#diffusion-models#embeddings#google-research 4 min1
arXiv cs.AI / cs.LG / cs.CL·15d agoAutonomous Research for Open-Ended Problems: A Case Study on Telecom Ticket Retrieval#ai-for-science#autonomous-research#cursorAI research
Hugging Face daily papers·17d agoGenerative Late-Interaction Embeddings For Visual Document Retrieval#compression#embeddings#late-interaction1
Hugging Face daily papers·18d agoThink Before You Link: Rarity, Reasoning, and Retrieval in Multilingual Entity Linking#benchmark#entity-linking#knowledge-graph1
arXiv cs.AI / cs.LG / cs.CL·18d agoMeClear: Cooperative Game-Theoretic Attribution and Risk-Aware Memory Clearance for Long-Horizon LLM Agents#agents#benchmark#llm-agentsAI research
arXiv cs.CR·19d agoRevoked but Still Authoritative: An Empirical Study of Revocation Enforcement in Agent-Memory Systems#agent-memory#guard#llm-agentsAI safety & security
MarkTechPost·21d agoPerplexity Details Its GPU Embedding Stack: How Ivy, Tulip and ROSE Serve pplx-embed#cuda#embeddings#flashattention 5 min2
Hugging Face daily papers·21d agoPARSER: Read in Parallel, Reason in Depth for Long-Context LLM Agents#llm-agents#long-context#multi-hop-qa1
Hugging Face Blog·Aug 26, 2026Training and Finetuning Multi-Vector Embedding Models with Sentence Transformers#embeddings#hugging-face#multi-vectorAI tools & infra1
Hugging Face Blog·Aug 18, 2026Multi-Vector (Late Interaction) Embedding Models with Sentence Transformers#embeddings#hugging-face#late-interactionAI tools & infra1