Not-so-hot take: The company running an AI agent should not also be the only company responsible for proving that agent followed the rules.
OpenAI’s new enterprise product, Presence, makes that problem pretty immediate.
Presence helps connect OpenAI's enterprise agents to company systems, defines what they can do, tests them before launch, evaluates their responses, controls when they escalates to a human, and improves itself over time.
When building the permissions, policies, evaluations, and records around every action it takes lives inside the same vendor’s stack, the vendor running the agent is also enforcing the rules and producing the evidence that those rules were followed.
That may be acceptable for low-risk tasks, but it becomes much harder to justify when agents can access customer accounts, issue refunds, change subscriptions, move money, or make decisions that affect real people.
Enterprise agents need a neutral layer underneath them that records what happened and enforces what the agent is allowed to do, regardless of which model or agent platform a company uses.
That is what we are building at @Sahara_AI - verifiable execution records and usage policies that do not depend on the vendor being verified.
OpenAI’s new enterprise product, Presence, makes that problem pretty immediate.
Presence helps connect OpenAI's enterprise agents to company systems, defines what they can do, tests them before launch, evaluates their responses, controls when they escalates to a human, and improves itself over time.
When building the permissions, policies, evaluations, and records around every action it takes lives inside the same vendor’s stack, the vendor running the agent is also enforcing the rules and producing the evidence that those rules were followed.
That may be acceptable for low-risk tasks, but it becomes much harder to justify when agents can access customer accounts, issue refunds, change subscriptions, move money, or make decisions that affect real people.
Enterprise agents need a neutral layer underneath them that records what happened and enforces what the agent is allowed to do, regardless of which model or agent platform a company uses.
That is what we are building at @Sahara_AI - verifiable execution records and usage policies that do not depend on the vendor being verified.
