CLEAR: Continuous Latent Adapter Routing for Utility-Preserving LLM Safety Alignment
Researchers propose a new framework called CLEAR to improve the safety of large language models (LLMs) without sacrificing their utility. The approach...
Researchers propose a new framework called CLEAR to improve the safety of large language models (LLMs) without sacrificing their utility. The approach...
Researchers have developed AUSO (Action-Level Unified Skill Optimization), a method for learning and using skills in AI agents. Unlike existing method...
The European Union's regulations on sustainability and privacy require companies to create documentation artifacts. However, creating these artifacts ...
Researchers have proposed a new approach to solving the Steiner Traveling Salesman Problem (Steiner-TSP) on Graphs of Convex Sets. The problem involve...
Researchers have developed a new type of neural network called Anatomy-Informed Neural Networks (AINN). Unlike traditional deep-learning models, AINNs...
Researchers have created a benchmark called VIALS for evaluating AI's ability to interpret visual artifacts in the life sciences. These artifacts incl...
Researchers have evaluated the safety of conversational AI systems used by Generation Alpha (born 2010-2024) for mental health support. The study foun...
Researchers analyzed language models to see if they still hold biases towards certain groups, even when their outputs appear neutral. They found that ...
Researchers have identified a problem with large language models processing electronic health records (EHRs), where information in the middle of long ...
Researchers have found that large language models (LLMs) are highly sensitive to minor changes in prompts. They analyzed a dataset of 132,000 prompt v...
Researchers propose a new approach to building industrial agents, called OneModel. Unlike traditional modular systems that break down complex tasks in...
Researchers have developed a method to detect bias in mental health natural language processing (NLP) models. They propose the Divergence Hypothesis, ...