Memory-Augmented Reinforcement Learning Agent for CAD Generation
A new paper presents a memory-augmented reinforcement learning framework aimed at improving computer-aided design…
Recent ai-research headlines from arXiv cs.AI.

A new paper presents a memory-augmented reinforcement learning framework aimed at improving computer-aided design…

The paper introduces CogScale, a benchmark designed to evaluate the cognitive and memory abilities of various AI…

A recent study investigates the role of code in enhancing mathematical reasoning. The findings suggest that while code…

The GroupAffect-4 dataset introduces a multimodal approach to studying four-person collaborative interactions. It…

The article discusses a new algorithm for reinforcement learning in multinomial logistic Markov Decision Processes…

OpenComputer is a new framework designed to create verifiable software worlds for computer-use agents. It includes…

The paper presents a method for uncertainty quantification in continuous AI agent evaluation. It introduces split…

The article introduces a new adaptive optimization framework that utilizes Schatten-p norms for deep neural networks.…

A recent study investigates the effectiveness of LLM agents in hardware-aware code optimization. The research reveals…

The article discusses a study on temporal grounding in scene-to-plan reasoning for autonomous vehicles. It highlights…

The article discusses the development of an explainable digital twin for wastewater treatment plants. This simulator…

The paper presents a novel approach to streamline constraint reasoning using Convolutional Neural Networks (CNN) for…

The paper introduces PEEK, a system designed to enhance long-context LLM agents by maintaining reusable orientation…

The paper discusses the need for improved guardrails for foundation models used in socially sensitive areas like…

The article discusses the introduction of a new framework called Probabilistic Tiny Recursive Model (PTRM) designed to…

GeoX is a new framework designed to enhance geospatial reasoning through self-play and verifiable rewards. It operates…

A recent study challenges the effectiveness of procedural knowledge in tool-grounded agents for offensive…

AutoResearchClaw is a new multi-agent autonomous research pipeline designed to enhance scientific discovery through…

A recent study explores the performance of embodied large language models (LLMs) in robotic tasks. The research…

The article discusses a new framework called inference-time argumentation (ITA) for claim verification in uncertain…

The article discusses a case study on using the Aristotle API for AI-assisted theorem proving in Lean 4, focusing on…

The paper introduces POW3R, a policy-aware rubric reward framework for reinforcement learning with verifiable rewards.…

A new machine learning model called HaorFloodAlert has been developed to predict flash floods in Bangladesh's haor…

The paper presents a methodology for selecting and composing runtime architecture patterns for production LLM agents.…

A new method called MAS-PNCG has been developed to improve the efficiency of Incremental Potential Contact…

The paper introduces OmniGUI, a new benchmark for evaluating graphical user interface (GUI) agents in omni-modal…

The paper discusses the divergence between human and AI aesthetic evaluations. It highlights that while AI can…

The paper introduces DotRAG, a new framework for Graph Retrieval-Augmented Generation that reformulates retrieval as a…

The paper presents ALDEN, a new method for enhancing private data extraction from Retrieval-Augmented Generation (RAG)…

The article discusses a new framework called Wearable As Graph (WAG) designed for analyzing personalized wearable data…

The paper introduces Domain-Driven Adaptable AI Pipelines (DDAP), a framework designed to assist non-expert scientists…

The article introduces STAR, a new retriever designed to enhance Graph Retrieval Augmented Generation for multi-hop…

The article discusses a new adaptive table retrieval method designed to improve the retrieval of relevant tables from…

The paper presents a new framework called DualView for multi-hop document reranking, which is essential for effective…

The paper introduces ClusterRAG, a novel approach for Personalized Retrieval-Augmented Generation (RAG). It emphasizes…

The article introduces the Agentic GraphRAG framework designed for analyzing unstructured financial data. This system…

The paper presents a new approach to improving Retrieval-Augmented Generation (RAG) without relying on taxonomy-based…

A new framework for community-based facilities management has been proposed, utilizing blockchain technology and…

The article discusses a new method called M3DocDep for processing long, multi-page documents using large…

The paper introduces Query-Aware Flow Diffusion for Graph-Based Retrieval-Augmented Generation (QAFD-RAG), a new…

The article discusses a new framework called Mask-to-Correct$^+$ designed to improve automated fact correction in the…

The article presents a reproducibility analysis of the PO4ISR model for session-based recommendations. It identifies…

A recent study tested whether large language models (LLMs) can simulate human belief dynamics in social networks. The…

The paper discusses the challenges of insuring risks associated with artificial intelligence (AI). It identifies…

The paper discusses the life history of features in language models, highlighting their emergence, persistence, and…

A new framework has been proposed to enhance spatiotemporal prediction performance by addressing the limitations of…

A new paper introduces a B-spline-based decoupling framework for compressing transformer models. This method aims to…

The paper introduces HELLoRA, a method for efficient fine-tuning of Mixture-of-Experts models using Low-Rank…

The paper presents the Fully Looped Transformer, a model designed to enhance training stability and performance in…

The paper introduces ReCrit, a transition-aware reinforcement learning framework aimed at improving scientific…
WeSearch's declared handling of arXiv cs.AI's content. Indexing, snippets, summaries, retrieval and training are separate questions — see the rights registry or read this source's machine-readable record.