LongDocBench: Benchmarking TOC Hierarchy and Contextual Relationship Recovery in Long Documents
Researchers have created a new benchmark called LongDocBench to evaluate the ability of AI systems t...
Researchers have created a new benchmark called LongDocBench to evaluate the ability of AI systems t...
Researchers have developed a method called Funnel of Thoughts (FoT) to improve the efficiency of lar...
Researchers have developed a method called Evo-Harness that allows self-improving large language mod...
Researchers have developed a decision-making framework for IoT systems in cold chain logistics that ...
Researchers have developed a new agent-native runtime called StateM that improve...
Researchers propose a new method for choosing the best state representation when...
Researchers have developed a new benchmark for evaluating how policies affect co...
Researchers have developed a method to generate synthetic tabular data that adhe...