Boosting Inference with Guided Reasoning: Stochastic Exploration for Recursive Models
The paper discusses a novel approach to enhancing inference in recursive neural networks through guided reasoning and…
Recent ai-research headlines from arXiv cs.AI.

The paper discusses a novel approach to enhancing inference in recursive neural networks through guided reasoning and…

The paper introduces Meta-Agent, a framework designed to improve the reliability of multi-agent systems. It automates…

The paper introduces FrontierOR, a benchmark designed to evaluate the capacity of large language models (LLMs) in…

LipoAgent is a new framework designed to enhance the safety and efficiency of lipid nanoparticles for nucleic acid…

The paper discusses the alignment of AI systems with organizational decision-making, emphasizing the complexity of…

The paper titled 'AI Cartography: Mapping the Latent Landscape of AI Benchmark Ecosystems' introduces a framework for…

The paper titled 'Context-CoT: Enhancing Context Learning via High-Quality Reasoning Synthesis' addresses the…

The paper titled 'Second Guess' introduces a method for detecting uncertainty in small language models through a…

The article discusses a new framework called LLMSurvival that enables censoring-aware survival analysis using large…

CODESKILL is a proposed framework aimed at enhancing coding agents' abilities through self-evolving skills. It…

The paper discusses the security challenges associated with OpenClaw agents, a new class of autonomous systems. It…

A new AI model called ECGCLIP has been developed to enhance cardiovascular assessment from routine…

The paper introduces the Artifact-Transform Workflow Language (ATWL), a formal language designed for visual analytics…

The article discusses advancements in credit assignment methods for language model reasoning in reinforcement…

The paper titled 'What Gets Cited: Competitive GEO in AI Answer Engines' explores how AI answer engines cite sources.…

The article presents BOHM, a new method for hierarchical attribution in compound AI systems. Unlike traditional…

NeuroNL2LTL is a neurosymbolic framework designed to translate natural language into Linear Temporal Logic (LTL). This…

The article introduces Research Math Agents (RMA), a framework designed for automated reasoning on complex…

SciAtlas is a large-scale knowledge graph aimed at enhancing automated scientific research. It integrates over 43…

A new framework called A-LEMS has been introduced for measuring energy consumption in agentic AI systems. This…

ImProver 2 is a neurosymbolic framework designed for automated proof optimization in formal mathematics. It addresses…

The article discusses the development of Mediative Fuzzy Logic, which aims to reconcile conflicting assessments in…

The paper introduces EVE-Agent, a self-evolving agent designed to enhance the reliability of AI-generated answers. It…

The paper discusses the limitations of AI systems, particularly large language models, and proposes design…

The paper introduces PathCal, a novel method for calibrating reasoning paths in Large Reasoning Language Models. It…

The paper presents Inductive Deductive Synthesis (IDS), a novel approach for enabling AI to generate formally verified…

The paper titled 'Redrawing the AI Map' explores accountability boundaries in agentic ecosystems. It introduces a…

The article discusses the advancements in AI systems aimed at automating scientific research workflows. It highlights…

The Foundation Protocol (FP) is introduced as a coordination layer for an emerging human-AI society. It aims to…

The paper titled GENSTRAT introduces a new approach to evaluate strategic reasoning in large language models (LLMs).…

The paper discusses the need for improved benchmarks in knowledge work AI, particularly in areas like coding and…

The paper discusses a new method called parallel context compaction for managing long-horizon LLM agents. This…

The paper introduces Ontological Knowledge Blocks (OKBs) as a solution for ensuring compliance in AI systems. OKBs…

The paper titled 'DART: Semantic Recoverability for Structured Tool Agents' addresses the challenges faced by…

A new paper presents a Human-in-the-Loop Multi-Agent Ventilator Decision Support System (VDSS) that utilizes…

The paper discusses the phenomenon of epistemic miscalibration in planning within LLM-based multi-agent systems. It…

The paper presents EDGE-OPD, a method for improving On-Policy Distillation (OPD) in machine learning. It addresses…

The paper discusses the integration of Dynamic Programming (DP) and Constraint Programming (CP) in solving the Partial…

The paper introduces Co-ReAct, a framework that enhances ReAct agents by using rubrics as step-level guidance during…

The article discusses the complexities involved in dismantling aircrafts at the end of their life cycle. It emphasizes…

The paper introduces a novel approach for controlling non-player characters (NPCs) in life simulation games using a…

The article discusses a new framework called MemAudit designed for auditing the memory of language model agents. This…

The paper discusses the application of agentic systems in program verification. It evaluates the performance of Claude…

The paper discusses advancements in multimodal large language models (MLLMs) for knowledge editing. It addresses the…

The paper titled 'SPACENUM: Revisiting Spatial Numerical Understanding in VLMs' explores the capabilities of…

The study explores the lifecycle of model-generated agent skills, focusing on experience generation, skill extraction,…

The paper introduces SkillOpt, a novel approach for optimizing agent skills in artificial intelligence. Unlike…

A new AI-driven framework has been proposed for energy-efficient environmental monitoring in smart cities. This…

The paper introduces KPI2KVI, a tool designed to compute Key Value Indicators (KVIs) from service descriptions. It…

The study evaluates the deceptive capabilities of Large Language Models (LLMs) in the social deduction game Secret…
WeSearch's declared handling of arXiv cs.AI's content. Indexing, snippets, summaries, retrieval and training are separate questions — see the rights registry or read this source's machine-readable record.