60 stories tagged with #harness, in publish-time order across the WeSearch catalog. Tag pages update as new stories ingest.
⌘ RSS feed for this tag → or search "Harness"
A look at OpenAI's open-source agent harness that now powers Codex and ChatGPT Work, as the company works to optimize the harness to cut runaway token usage (Jason Hiner/The Deep View)
Visa used Mythos to hunt for bugs in its own payment network, then open-sourced the harness that made it possible
Birgitta Boeckeler on Harness Engineering for AI Agents [audio]
Microsoft says MAI-Cyber-1-Flash and MDASH, its vulnerability identification harness, deliver "world-class performance at 50% of the cost of leading models" (Microsoft AI)
customized Color sample Logo And Packaging Meanings In Pet Harness Sets
customized Color sample Logo And Packaging Meanings In Pet Harness Sets…
AI can fuel biological weapons. We must harness its power for defense | Annie Jacobsen
The offensive potential is no longer theoretical. We need to develop systems to strengthen public health as quickly as AI is accelerating biological design…
I built an AI dev harness that isn't allowed to trust itself. Then I checked the part doing the not-trusting.
A follow-up to I built an AI dev harness that isn't allowed to trust itself . The harness I wrote about isn't allowed to trust itself. That was the whole point, and it rested on on…
NVIDIA Harnesses Vera CPU to Speed Up Design of Next-Generation CPUs and GPUs
The complexity of modern chip design continues to grow as engineering teams work to develop increasingly sophisticated CPUs, GPUs and AI systems. To help meet that challenge, NVIDI…
AI For Mental Health Relies Critically On Leading-Edge Harness Engineering
AI for mental health relies on a hidden factor, namely the use of leading-edge harness engineering. This is critical. An AI Insider analysis and scoop.…
Laserbrain – the adaptive recursion harness for AI agents
A fixed reference your AI agent checks itself against to catch its own drift and stop. Provable, free grammar, one-line MCP.…
Open Source AI Harness Profiler – discover where tf your tokens are going
Contribute to TryRekon/Rekon development by creating an account on GitHub.…
Show HN: Cross-Harness self hosted registry and analytics for AI Agents
Observal is a local registry and analytics platform for your AI components. Setup Observal, define the scope and share your Skills, MCPs and Agents. - Observal/Observal…
Beyond grep: The case for a context-rich AI coding harness
Augment Code's Vinay Perneti talks models, harnesses, and context.…
Keeping teams on one AI harness
Every team I talk to asks the same question: how do we keep everyone on the same AI setup? Wiki pages and sync scripts fail at this. What holds is a versioned baseline, team packs,…
Self-testing AI harness finds its own bugs
Evaluating performance and efficiency of the GitHub Copilot agentic harness
Explore how the GitHub Copilot agentic harness delivers strong results across multiple benchmarks and leading token efficiency.…
Woman, 21, dies after being thrown from Brazil rope jump bridge without harness
Instructors hurled Maria Eduarda Rodrigues de Freitas into 40-metre abyss without attaching safety equipment A 21-year-old woman who died when two rope jumping instructors threw he…
The Harness Has a Token Budget
Our project CLAUDE.md crossed 4,000 tokens last quarter, and the agent started missing rules it had...…
Show HN: Aura, an LLM coding harness that dogfooded itself
An AI coding harness that dogfooded itself into shape: Planner/Worker agents, repo awareness, surgical edits, validation, recovery, and safe diff approvals. - CarpseDeam/Aura-IDE…
A harness for every task: dynamic workflows in Claude Code
https://t.co/R6exTuF7P8…
EvoTrainer: Co-Evolving LLM Policies and Training Harnesses for Autonomous Agentic Reinforcement Learning
Autonomous LLM training is often framed as recipe search, which leaves the training harness largely static. This limitation sharpens in agentic RL, where shifting bottlenecks and s…
System Boundaries: The Difference Between ChatBot, Workflow, Agent, and Harness
When people first build Agent systems, they often naturally read them as an upgrade path:…
Harness Base Definition: The Control System Outside the Model
Previously, we split Agent into several minimal parts:…
Show HN: LiteHarness – One SDK for Claude Agent, OpenAI Agent, Pi AI
Unified Server for running OpenCode, Claude Code, Codex agents - LiteLLM-Labs/lite-harness…
Harness Acquires Codecov from Sentry
Harness acquires Codecov from Sentry to bring code coverage intelligence into software delivery governance for AI-accelerated engineering teams. | Harness Press…
Stop Shipping AI Slop: Build an Anti-Slop Harness Around Your LLM
"AI slop" is not a model problem. It's an engineering problem you decided not to solve. The slop is...…
Harness Engineering Course
Agent Harness Explained: Build Production-Ready AI Agents with Microsoft Agent Framework
Learn what an agent harness is, why it matters for production AI systems, and how to implement one step-by-step using Microsoft Agent Framework's create_harness_agent — with real P…
Scaling Laws for Agent Harnesses via Effective Feedback Compute
Agent harnesses increasingly determine the performance of language-model systems by deciding how models call tools, receive feedback, verify intermediate states, store memory, and …
Show HN: theta-spec - a humble harness agnostic configuration spec
harness agnostic configuration standard . Contribute to tamarillo-ai/theta-spec development by creating an account on GitHub.…
The Agent Harness Taught Me Why I Used to Fail
On building AI agents and accidentally understanding yourself Introduction We tend to...…
Stop Upgrading the Model. Start Engineering the Harness.
When a team hits a ceiling with their coding agent, the first instinct is to reach for a better...…
OpenAI Agents SDK: Building with Model-Native Harnesses - StartupHub.ai
Comprehensive up-to-date news coverage, aggregated from sources all over the world by Google News.…
Show HN: What 1k Harness Experiments Taught Me About Self-Improving Agents
Project Repository: https://github.com/workofart/harness-experiment So I recently wanted to see whether an AI agent could self-improve a harness to solve terminal bench tasks. To a…
Show HN: VAEN – Package and import portable AI coding-agent Harnesses
Package your AI coding harness into a portable .agent file, and share it across repos, teams, & the community without ever having to copy-paste instructions, skills, MCP config…
Show HN: Open-Source AI Racing Harness
Cadence v8.4: a multi-model coding harness where Claude writes, Codex reviews, and Bugbot triages
Claude writes. Codex reviews. Bugbot triages. Gemini sits on the council. ...…
Show HN: CoreTex – An Open-Source, Unix-like, biomimetic, flat-file AI Harness
A UNIX-inspired, biomimetic, flat-file AI harness and knowledge engine. - mrdanielcasper/CoreTex…
AI has slashed coding time in 2026, but it’s sacrificed software stability
AI accelerates coding, but slows innovation at scale…
Impeccable: Design skills for AI harnesses
1 skill, 23 commands, and curated anti-patterns for impeccable frontend design. Works with Cursor, Claude Code, Gemini CLI, and Codex CLI.…
The AI Agent Harness: The Glue That Turns LLMs into Digital Workers
AI models have plateaued on raw intelligence. The next gains come from what you build around them.…
Designing a Modular Wiring Harness for Multi-Function Vehicle Trackers
How to allocate a 9-pin connector across 6 swappable modules for fleet, cold chain, security, and e-vehicle tracking.…
Harness Engineering: The New DevOps Layer for AI Agents
SIA: Self Improving AI with Harness & Weight Updates
Humans are the bottleneck in building and improving AI. Both the models and the agents that wrap them are written, tuned, and corrected by people. The long-horizon goal of an AI th…
It's Not the Capability: Harness Sensitivity Is Non-Monotone Across LLM Agent Tiers
A prevalent assumption in LLM agent deployment holds that more structured harnesses universally improve reliability, and that higher-capability models need proportionally less st…
humble attempt at building a manifest centered configuration surface for arbitrary harnesses
theta: a humble approach to harness agnostic configuration
canonical implementation of the theta-spec. Contribute to tamarillo-ai/theta development by creating an account on GitHub.…
Polar: Agentic RL on Any Harness at Scale
Reinforcement learning for language agents increasingly depends on custom harnesses that manage long-running context, multi-turn tool use and multi-agent orchestration. However, po…
Token-level eval harness for tool-calling agents: what we wired up
TL;DR: We replaced our "did the agent finish the task" pass/fail eval with a token-level harness that...…
Building the harness around our coding agents: eight failure modes, eight pillars
Building the harness around our coding agents. Eight failure modes and pillars
Notes on the harness we built around Claude Code and Codex, organized as eight coding agent failure modes and eight harness pillars.…
Stord raises $250M to harness AI for e-commerce logistics
Harness, Scaffold, and the AI Agent Terms Worth Getting Right
We’re on a journey to advance and democratize artificial intelligence through open source and open science.…
Harness Engineering: Stop Re-Prompting Your Coding Agent Every Session
Every time I started a new agent session, I was re-explaining the same things. The architecture...…
Ask HN: Is Codex serving worse models or is it just the harness getting worse?
AION: Next-Generation Tasks and Practical Harness for Time Series
Time series research is moving beyond fixed forecasting benchmarks toward realistic tasks that combine prediction, contextual reasoning, tool use, and structured decision support. …
DemoEvolve: Overcoming Sparse Feedback in Agentic Harness Evolution with Demonstrations
Agent harness evolution improves frozen language-model agents by modifying the executable structures around them. We study this paradigm as a form of sample-efficient fast adaptati…
Stop Comparing LLM Agents Without Disclosing the Harness
This position paper argues that, for long-horizon tasks evaluated across models with comparable frontier capability, the agent execution harness, namely the infrastructure layer th…
Moved from Harness to Revolte for delivery automation, what's the difference?
French Spider-Man rests mid-cliff at 60, no harness needed
Watch legendary rock and urban climber Alain Robert (often called "The French Spider-Man") casually stop to rest his arms in the middle of a climb at age 60 while in……