Beyond Control-Flow: Integrating the Resource Perspective into Multi-Collaborative Process Modeling from Text
The paper discusses a novel approach to process modeling in Business Process Management (BPM) that integrates resource…
Recent ai-research headlines from arXiv cs.AI.

The paper discusses a novel approach to process modeling in Business Process Management (BPM) that integrates resource…

The paper introduces PALoRA, a framework designed to enhance the integration of new knowledge into Large Language…

The paper discusses a new framework for safe fine-tuning of large language models (LLMs) called Buffer-and-Reinforce.…

A new paper proposes a method to mitigate look-ahead bias in financial backtesting using large language models. The…

A recent study explored the relationship between echocardiographic traits and AI-ECG predictions of heart failure. The…

HeartBeatAI is a new deep learning framework designed for multi-label ECG arrhythmia detection. It addresses…

A recent study explores the use of A* search algorithms to improve reasoning in large language models (LLMs). The…

The article discusses a new approach called Hera for coordinating device-cloud collaborative large language model…

The article presents a new framework called Agent-as-Peer-Debriefer designed to enhance qualitative data analysis…

A new paper presents an algebraic framework for deep convolutional learning based on lattice theory and mathematical…

GlobalDentBench is introduced as the first multinational benchmark for evaluating large language models (LLMs) in…

AVBench is a newly introduced benchmark aimed at improving the evaluation of audio-video generative models,…

The article discusses a new approach to deploying large language models (LLMs) that goes beyond inference-only…

A new study proposes a multi-dimensional framework for evaluating reasoning quality in large language models (LLMs).…

The paper discusses the limitations of mean cross-entropy (CE) as a metric for evaluating language model quality. It…

A new study explores the use of perceptual speech features to support clinical decision-making in mental health care.…

A recent study highlights the fragmented nature of emotional intelligence in large language models (LLMs). The…

The article discusses the MDIA, a Multi-Agent Diagnostic Intelligence Pipeline designed for clinical reasoning. It…

A recent paper discusses the inherent limitations in explaining AI systems, particularly large-scale models. The…

The paper titled 'Hylos: Operability Contracts for Model-Native Spatial Intelligence' introduces a new systems…

A new study presents an automated pipeline using multi-agent language models to detect and classify delusion-related…

The paper introduces the Trajectory Proper Score (TPS) for evaluating agentic uncertainty quantification in AI. It…

The paper discusses a novel approach to uncertainty decomposition in subjective natural language processing (NLP). It…

The paper titled 'PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and…

The paper introduces GRAIL, an AI translation system designed to assist scientists in converting Python geospatial…

The paper introduces PANDO, a framework designed to enhance the efficiency of multimodal AI agents through online…

The paper introduces CoRe-Code, a framework for collaborative reinforcement learning aimed at improving code…

The article discusses a new paradigm in manufacturing called Agent Manufacturing, which focuses on the role of…

A new framework called Test-Time Exploration (TTExplore) aims to improve the performance of intelligent agents in…

The paper introduces Geo-Expert, a series of parameter-efficient geological language models designed to improve…

The paper presents a new approach to solving combinatorial counting problems using a language called Cofola. This…

The paper presents a novel approach to Chain-of-Thought (CoT) graph learning by interpreting it through the lens of…

A new framework called POLARIS has been introduced to enhance safety testing for Large Language Models (LLMs). This…

The paper presents TaBIIC2, a tool designed for the interactive construction of ontological taxonomies using weighted…

The paper introduces ProActor, a framework for proactive task scheduling using timing-aware reinforcement learning. It…

The paper introduces a new method called NORA for improving the understanding of financial numerical entities in…

The paper titled 'Energy Shields for Fairness' introduces a novel approach to ensuring runtime fairness in…

The paper presents a multi-turn dialog system tailored for industrial asset operations and maintenance. This system…

A new paper addresses the issue of object hallucination in Large Vision-Language Models (LVLMs). The authors propose a…

The paper presents NeurIPS, a framework designed to enhance surface-based brain decoding by utilizing neuro-anatomical…

A study evaluated the use of a privacy-preserving small language model (SLM) for retrieving clinical data in pemphigus…

The paper titled 'AION: Next-Generation Tasks and Practical Harness for Time Series' presents a new framework for time…

The paper presents a new framework for multi-agent reinforcement learning in cooperative air combat scenarios. It…

The paper presents RECTOR, a rule-based reranking system designed for autonomous driving trajectory selection. It…

The paper introduces a new protocol called prover-verifier deliberation (PVD) for improving the reliability of…

The paper introduces a method called stochastic backtracking for improving test-time scaling in language models. This…

The paper titled 'Representation Without Control: Testing the Realization Effect in Language Models' explores the…

The paper introduces SimuWoB, a synthetic benchmark designed for evaluating mobile GUI agents. It addresses the…

The paper introduces SpecAlign, a framework designed to enhance the semantic alignment of SystemVerilog Assertions…

The paper presents DarkForest, a new framework aimed at improving the accuracy of multi-agent large language models…
WeSearch's declared handling of arXiv cs.AI's content. Indexing, snippets, summaries, retrieval and training are separate questions — see the rights registry or read this source's machine-readable record.