TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning
The paper introduces TaskGround, a framework designed for structured executable task inference in household settings.…
Recent ai-research headlines from arXiv cs.AI.

The paper introduces TaskGround, a framework designed for structured executable task inference in household settings.…

The article discusses a new framework for Multivariate Time Series Anomaly Detection (MTSAD) that integrates…

A recent study explores the impact of Generative AI on productivity in education. The findings indicate that while…

A new system called pArticleMap has been introduced to enhance research in nanomedicine by mapping literature and…

The article discusses a new framework called ConceptAgent designed to address limitations in current concept erasure…

The paper introduces TRACE, a new algorithm designed to reduce hallucinations in language models by utilizing…

The paper introduces Generative Visual Grounding (GVG), a framework aimed at enhancing understanding of EEG signals…

The paper discusses the importance of scalable environments for developing generalizable agents in artificial…

The paper introduces a new reinforcement learning method called Pairwise Preference Reward and Group-based Diversity…

The paper explores the limitations of Multi-Modal Large Language Models (MLLMs) in spatial reasoning, particularly…

DARE-EEG is a new foundation model designed for mining dual-aligned representations of EEG data. It addresses the…

The paper titled 'SD-Search: On-Policy Hindsight Self-Distillation for Search-Augmented Reasoning' introduces a novel…

The paper introduces Causely, a causal intelligence layer designed to enhance AI agents in Site Reliability…

Researchers have introduced QSTRBench, a new benchmark designed to evaluate the reasoning capabilities of language…

A new variant of the Firefly Algorithm has been developed to enhance data clustering capabilities. This algorithm…

The paper introduces OCCAM, a framework designed for open-set causal concept explanation and ontology induction in…

A new denoising pipeline for high-throughput Raman spectroscopy has been developed. This method utilizes a…

The paper introduces Asymmetric Meta-Reflective Self-Distillation (AMR-SD) as a solution for token-level credit…

The article discusses a new framework called VISAFF for Emotion Recognition in Conversation (ERC). This framework aims…

The paper presents a novel approach called Query-Conditioned Entity Alignment (QCEA) for improving cross-system…

The paper discusses the limitations of outcome-only evaluations in artificial intelligence, particularly in the…

DeepSlide is a new multi-agent system designed to enhance the presentation delivery process. It focuses on improving…

The article discusses the SDOF framework designed to enhance multi-agent orchestration by enforcing state constraints.…

The paper explores the impact of improving Theory of Mind (ToM) capabilities in Large Language Models (LLMs) on…

SkillSmith is a new framework designed to optimize the execution of skills in large language model-based agent…

A recent study investigates the latent biases in instruction-tuned language models used for high-stakes decisions.…

The paper introduces CAX-Agent, a lightweight agent harness designed for reliable automation in MAPDL finite-element…

The paper titled 'NOVA: Fundamental Limits of Knowledge Discovery Through AI' explores the capabilities and…

The paper presents a novel framework called ICRL, which aims to enhance the self-improvement capabilities of language…

The NIMO Controller is a proposed self-driving laboratory orchestrator that utilizes the Model Context Protocol to…

The paper introduces a Distributed Trust Framework (DTF) designed to enhance authorization for sovereign AI systems.…

The article introduces Solvita, a new framework designed to enhance large language models for competitive programming.…

The paper introduces SMCEvolve, a new framework for automated scientific discovery using Sequential Monte Carlo…

A new framework called LaMR has been proposed for improving coding agents by enhancing context pruning. This method…

The paper discusses the evaluation of large language models (LLMs) in the context of zero-shot goal recognition. It…

The article introduces the Belief Engine (BE), a new framework designed for multi-agent deliberation using large…

A recent study highlights the importance of ensemble monitoring for AI control, demonstrating that diverse signals are…

The paper presents a new framework called Influence-Based Team Steering (IBTS) for enhancing human-machine teaming in…

A new framework called NSPI has been proposed for automated polynomial inequality proving. This method combines large…

The paper presents X-SYNTH, a framework for enterprise context synthesis based on observed human attention. It…

The article introduces CAPS, a new framework for efficient parallel reasoning in large language models. CAPS utilizes…

The paper introduces RTL-BenchMT, an automated framework for maintaining RTL generation benchmarks. It addresses…

The paper presents DRS-GUI, a training-free framework for GUI grounding that enhances the performance of Multimodal…

The paper discusses the necessity of metacognition in artificial intelligence design. It proposes that AI systems…

The article discusses the STAR framework, which aims to enhance the reliability of root cause analysis (RCA) agents in…

A new framework called OmniManim has been developed to improve the generation of educational animations from code.…

The article introduces TopoEvo, a new framework designed for root cause analysis in microservices. This framework…

ColPackAgent is a new framework designed to autonomously run Monte Carlo simulations for colloidal packing. It…

The article introduces PRISM, a framework designed to enhance the reliability of prompts used in enterprise…

The article discusses a new framework called NudgeRL for improving exploration in reinforcement learning with…
WeSearch's declared handling of arXiv cs.AI's content. Indexing, snippets, summaries, retrieval and training are separate questions — see the rights registry or read this source's machine-readable record.