A System for Production-Ready AI Agents

OpenAI Presence moves beyond simple model implementation by providing a structured environment for deploying AI agents in high-stakes enterprise workflows. The core challenge addressed is reliability: ensuring agents can perform specific tasks—such as billing resolution, insurance claims, or IT support—while adhering to strict company policies and guardrails.

Presence functions as a comprehensive platform that integrates model reasoning with:

  • Policy and Guardrail Enforcement: Defining what an agent is permitted to do, when it requires human approval, and when to escalate to a human representative.
  • Simulation and Evaluation: Before deployment, agents are tested against edge cases and high-risk scenarios to verify accuracy and adherence to operational procedures.
  • Continuous Improvement Loop: Post-launch, the system uses production data to identify performance gaps. A Codex-powered process suggests updates, which teams can test and approve for a controlled rollout, allowing the agent to evolve alongside changing business requirements.

Operational Integration and Performance

Presence is not a self-serve tool; it is a managed service deployed in collaboration with OpenAI Forward Deployed Engineers (FDEs) and systems integrators. The product is designed to be modular, allowing enterprises to maintain consistent policies and evaluation frameworks across different channels while customizing specific workflows.

Evidence of its effectiveness is demonstrated by OpenAI's own internal use: their English-language phone support channel now resolves 75% of inbound issues without human intervention. By utilizing the Codex-powered improvement loop, the team reduced human handoffs by 15 percentage points in just 10 days. Early enterprise adopters, including BBVA, SoftBank, and IAG, are currently testing the platform for financial services, customer support, and disaster-response claims processing.