Name: Operator
Rating: 24.1 (1 reviews)
Author: OpenAI

Operator by OpenAI | AI Market Cap

HF PapersOpenAIresearch1w ago

How Much Static Structure Do Code Agents Need? A Study of Deterministic Anchoring

LLM-based code agents navigate repositories through keyword search but miss the structural relationships, such as call graphs, inheritance hierarchies, and configuration dependencies, that define how software actually works. This makes agent navigation stochastic and difficult to reproduce across runs. We investigate whether lightweight static analysis can provide deterministic anchors for these agents: stable structural facts injected as plain-text comments that constrain probabilistic exploration and make navigation more predictable. Starting from a strong baseline, Codex from OpenAI, we systematically inject varying granularities of structural annotations and measure their effects on localization, trajectory behavior, and run-to-run stability. Our study identifies what we call the deterministic anchoring effect: static structure helps less by making agents "smarter" and more by making their navigation disciplined and reproducible. Three observations support this finding: (1) Anchoring works: lightweight call/inheritance topology improves function-level localization (+2.2pp Func@5) and shortens trajectories (-1.6 interaction rounds); (2) Anchoring is scale-sensitive: the optimal granularity and directionality depend on repository characteristics, where denser semantics show diminishing returns and hub-heavy projects benefit from inverse-only links that expose "who-calls-me" without forward edges; (3) Anchoring stabilizes: tags raise link-following rate from 0.15-0.18 to 0.21-0.24, roughly halve run-to-run variance, and improve single-run reliability (Pass@1 +3.4 pp) on medium-scale repositories, at the cost of roughly 10% more input tokens. These observations suggest practical guidelines: default to lightweight topology on medium projects, prune forward edges in large repositories, and reserve dense tags for implicit-dependency cases.

View Source

#huggingface#daily-papers

Operator

Similar Models

Introducing GeneBench-Pro

Social & Blog Posts6

Research Papers22

Other

Introducing GPT‑5 for developers

We’re introducing GeneBench-Pro, a research-level benchmark for a harder kind of AI progress: how well agents can navigate messy biological data, choose the right analysis path, and make judgment call

Introducing GPT‑5 for developers

We’ve designed and built our first AI chip: Jalapeño. Designed from the ground up by OpenAI and brought to production with @Broadcom, Jalapeño is purpose-built for the LLM workloads powering ChatGPT,

WARP: Weight-Space Analysis for Recovering Training Data Portfolios

We’re introducing GeneBench-Pro, a research-level benchmark for a harder kind of AI progress: how well agents can navigate messy biological data, choose the right analysis path, and make judgment call

How ChatGPT adoption has expanded

Introducing GeneBench-Pro

HP Inc. launches Frontier strategic partnership with OpenAI

We’ve designed and built our first AI chip: Jalapeño. Designed from the ground up by OpenAI and brought to production with @Broadcom, Jalapeño is purpose-built for the LLM workloads powering ChatGPT,

OpenAI and Broadcom unveil LLM-optimized inference chip

WARP: Weight-Space Analysis for Recovering Training Data Portfolios

How Much Static Structure Do Code Agents Need? A Study of Deterministic Anchoring

To Run or Not to Run: Analyzing the Cost-Effectiveness of Code Execution in LLM-Based Program Repair

ResAdapt: Adaptive Resolution for Efficient Multimodal Reasoning

HISA: Efficient Hierarchical Indexing for Fine-Grained Sparse Attention

Kernel-Smith: A Unified Recipe for Evolutionary Kernel Optimization

Working Notes on Late Interaction Dynamics: Analyzing Targeted Behaviors of Late Interaction Models

A Neural Score-Based Particle Method for the Vlasov-Maxwell-Landau System

AVO: Agentic Variation Operators for Autonomous Evolutionary Search

AI Lifecycle-Aware Feasibility Framework for Split-RIC Orchestration in NTN O-RAN

A Schrödinger Eigenfunction Method for Long-Horizon Stochastic Optimal Control

BOOST-RPF: Boosted Sequential Trees for Radial Power Flow

SparseDVFS: Sparse-Aware DVFS for Energy-Efficient Edge Inference

All elementary functions from a single binary operator

GIST: Gauge-Invariant Spectral Transformers for Scalable Graph Neural Operators

High-dimensional estimation with missing data: Statistical and computational limits

Physics-Informed Neural Systems for the Simulation of EUV Electromagnetic Wave Diffraction from a Lithography Mask

Building Trust in PINNs: Error Estimation through Finite Difference Methods

Factorized Neural Implicit DMD for Parametric Dynamics

Differentiable Zero-One Loss via Hypersimplex Projections

The logic of KM belief update is contained in the logic of AGM belief revision

Learning Physical Operators using Neural Operators