ConceptGuard: Benchmarking Context-Sensitive Unlearning in Large Language Models
Researchers have developed a benchmark called ConceptGuard to evaluate the ability of large language models to remove harmful or sensitive knowledge w...
Researchers have developed a benchmark called ConceptGuard to evaluate the ability of large language models to remove harmful or sensitive knowledge w...
Researchers propose a new framework called HARP for prioritizing vulnerabilities based on specific operational preferences. Unlike existing methods th...
Researchers propose DeltaMomentum, a new method for updating momentum in deep learning optimizers. Unlike traditional exponential moving average (EMA)...
Researchers from Japan investigated whether adding listening behaviors to an AI clone can improve its perceived authenticity. They integrated verbal b...
Researchers have developed StreamSoccer, a system for live soccer commentary that uses event-driven memory to process and generate commentary in real-...
Researchers have proposed a framework for improving the accuracy of medical image captioning. Medical image captioning is a technique that helps docto...
Researchers have developed a method to improve the efficiency of communication topologies in multi-agent systems. The approach, called Reward-Guided A...
Researchers have proposed a hybrid framework for autonomous driving that combines reinforcement learning and PID control with the common-sense reasoni...
Researchers have developed a new framework called Explain-MDRC for recognizing depression in clinical interviews. It combines text, audio, and facial ...
Researchers have created an interactive benchmark called Build What I Mean (BWIM) to test how well language models can follow instructions in a collab...
Researchers have developed a new AI framework called Doc-V* that can answer questions about multi-page documents without needing to recognize text fir...
A new method called Token-to-Mask (T2M) has been proposed to improve the performance of diffusion language models. T2M identifies low-confidence posit...