When Failures Propagate: Causal Failure Attribution in Agentic Retrieval-Augmented Generation
A new benchmark, AgenticRAG-FP, has been introduced to evaluate the ability of AI systems to identify the source of errors in retrieval-augmented gene...
A new benchmark, AgenticRAG-FP, has been introduced to evaluate the ability of AI systems to identify the source of errors in retrieval-augmented gene...
Researchers have developed AgentMercury, a framework for generating large-scale, realistic business environments that can be used to train AI agents. ...
Researchers have developed a framework called ARQ that refines CodeQL queries for detecting vulnerabilities in C/C++ programs. ARQ uses synthesized pr...
Researchers have studied how the Adam optimization algorithm behaves near a point where it stops improving. They focused on a simple mathematical prob...
Researchers propose a single hierarchical system for product IDs that can be used in both discovery and search. This system, called Semantic ID, repre...
Researchers have introduced a new model called RiskTraf for predicting traffic flow. Unlike existing models that often ignore or misrepresent raw meas...
Researchers have developed a way to improve the quality of galaxy images taken by ground-based telescopes using generative AI. By training AI on high-...
Researchers have developed a new framework called C-Score to assess the robustness of semi-supervised learning (SSL) models in real-world environments...
Researchers have proposed LA-ReduNet, a lightweight version of the neural network ReduNet. Unlike traditional deep networks, ReduNet explicitly derive...
Researchers have developed a method to eliminate 'stale-fact errors' in code-assistant memory, where models like Retrieval-Augmented Generation (RAG) ...
Researchers have proposed a new task for human-object interaction (HOI) motion captioning that requires generated captions to specify both the subject...
Researchers have developed a new attack method called Vis-Poison that can compromise multimodal large language models (MLLMs) by introducing poisoned ...