← All Journal articles

Property Operations

AI Proficiency Means Finished, Verified Work

AI proficiency is not mastery of one tool. For property managers, it means using the right capability for a defined job, verifying the finished work, and expanding authority only after measuring time, completeness, errors, rework, and exceptions.

A hand checks a metal lever handle on a repaired interior door.

OpenAI launched the Agents API in public beta on September 10, 2026. In its announcement, OpenAI said the API was available to all developers and provided managed infrastructure for building cloud agents with the harness behind Codex.

As reviewed on September 11, 2026, the Agents API supports long-running sessions and tools that include MCP, custom functions, and built-in tools. That is a meaningful platform development. It is not a finished property-management workflow.

AI is an umbrella for different capabilities: drafting, retrieving context, classifying requests, preparing work packets, using authorized tools, and handling narrow routine tasks. A team can be effective with one capability and unready for another.

For this operating standard, proficiency means using an appropriate capability for a defined job and verifying the finished work against the prior process. Did the work take less time? Was it more complete? How often did someone correct or redo it? Did exceptions reach the right person?

Courses and certifications can support learning. The operating test is whether the team can produce accurate, reviewable work inside clear boundaries.

The Agents API may reduce some infrastructure work in a custom project. OpenAI still leaves the developer responsible for the tools, knowledge, and workflows that make the agent useful. The API does not establish AppFolio permissions, assign ownership of a maintenance decision, or prove that a completed tool call produced the right operating result.

The five sources reviewed on September 11, 2026, do not establish a packaged Codex-to-AppFolio connector, an approved AppFolio MCP server, or a production-ready Codex-to-AppFolio maintenance workflow. That is an evidence gap, not proof that no private or later integration exists. Confirm current AppFolio access, operations, and implementation requirements directly before proceeding.

Start With Three Levels of Email Work

Email is a practical place to coach the progression because the task is familiar and the boundaries can be stated plainly. These levels are an illustrative operating framework, not a documented native OpenAI-AppFolio workflow.

Level 1: Draft From Supplied Facts

A manager supplies the facts and intended outcome to ChatGPT and requests a draft. A person checks the result, corrects it if needed, and sends it.

For example, a coordinator supplies a resident's maintenance request, the known access window, and the information still needed. The coordinator then verifies the resident, unit, issue, and requested next step before using the draft.

At this stage, the tool does not search connected systems, select a vendor, change a work order, or make an operating decision. It turns supplied information into a message for human review.

Measure whether the finished draft is accurate, complete, and faster to review than writing from scratch. Do not assume improvement merely because a draft appeared quickly.

Level 2: Prepare a Context-Aware Response

At the next level, an agent built with the Agents API could be configured to use explicitly authorized tools. In an illustrative maintenance workflow, it might retrieve approved context such as an existing work-order status, property and unit identifiers, access instructions, or prior communications.

It could then prepare a response or work packet and flag missing or conflicting information.

The important word is prepare. The designated employee approves the response and decides whether any record should change.

This level relies on custom implementation. It is not a claim that Codex has a native AppFolio connection.

Level 3: Schedule Bounded Routine Work

The third level is not permission to run the department. It is a scheduled process limited to a defined queue, approved tools, and explicit rules.

For example, a pilot could review nonemergency requests that lack access instructions and prepare a standard request for the missing information. A later version might send that narrow class of message only after testing and approval.

The process should escalate an emergency indicator, apparent duplicate, conflicting record, spending question, resident-sensitive issue, or anything else outside the rule set.

The question is not whether the technology can compose an email. The question is whether the task is routine enough, the boundaries are narrow enough, and the exception owner is clear enough to support limited execution.

Apply the Same Progression to Maintenance

A maintenance workflow may include intake, triage, access instructions, vendor assignment, spending authority, vendor follow-up, resident updates, completion evidence, and documented closeout.

AppFolio's public Stack API catalog, reviewed September 11, 2026, lists operations to get, create, or update work orders for a property, unit, or occupancy. The listed attributes include property, unit, status, job description, assigned users, completion date, permission to enter, priority, scheduled dates, vendor, and vendor trade.

The catalog establishes categories of available API operations. It does not establish a particular organization's entitlement, authorization, security configuration, or field-level permissions.

Before connecting an agent, answer the operating questions:

  • Which system is the system of record at each step?
  • Which fields may the agent read, prepare, or change?
  • Who decides whether an issue is an emergency?
  • Who approves vendor selection or a spending exception?
  • What happens when details are incomplete or contradictory?
  • Who accepts completion evidence and closes the work order?

A staged authority model gives the agent more access only after the prior stage produces acceptable evidence.

Read-only comparison

Agent role: Checks requests against a required-field checklist and flags gaps or conflicts.

Human role: Retains triage, priority, record changes, and escalation decisions.

Drafting and work packets

Agent role: Prepares clarification messages, triage summaries, vendor follow-up drafts, and resident updates.

Human role: Reviews communications and retains system changes.

Human-approved writes

Agent role: Prepares a narrowly defined create or update action for explicit approval.

Human role: Approves the exact change and handles exceptions.

Limited routine execution

Agent role: Performs only preapproved, low-consequence actions within tested rules.

Human role: Retains emergency judgment, spending decisions, exceptions, and closeout acceptance.

This is how an owner distinguishes useful assistance from unattended work with unclear authority.

A Before-and-After Maintenance Pilot

Consider a nonemergency request with a vague description and no access instructions.

Before: The Coordinator Reconstructs the Case

The coordinator reads the request, looks up the property and unit, checks for an existing work order, reviews prior communications, sends a clarification message, and later checks whether the resident replied.

If a vendor becomes involved, the coordinator tracks outreach and searches for completion evidence before closing the record. Each handoff must be documented for the next person to understand the case.

After: A Read-Only Packet Supports the Coordinator

In an illustrative pilot, an agent compares the request against a checklist:

  • Confirmed property and unit
  • Description of the issue
  • Possible emergency indicators
  • Access or permission-to-enter information
  • Existing open work orders that may be duplicates
  • Facts needed for triage
  • Known vendor, schedule, or status context, if authorized and available

The agent prepares a missing-information message for the resident and a structured triage summary for the coordinator.

If it finds possible emergency conditions, conflicting unit information, an apparent duplicate, or missing facts that affect priority, the packet identifies the exception rather than resolving it.

The coordinator remains responsible for priority, vendor assignment, spending authority, and the decision to create or update a work order.

This scenario is a proposed pilot design. It is not evidence of a deployed Codex-AppFolio integration or measured savings. The finished work would be a verified packet with the correct property and unit, visible gaps, a usable draft, and a documented human decision.

Do Not Automate the Same Decision Twice

Before building a custom maintenance agent, determine whether AppFolio's native maintenance automation already handles the proposed function.

As reviewed September 11, 2026, AppFolio describes Realm-X Maintenance Performer as handling resident intake, troubleshooting, work-order creation, emergency classification, and vendor dispatch under configured rules and human supervision.

Those documented functions may overlap with a proposed custom agent. That overlap is an inference operators should investigate, not proof that every implementation will conflict.

A custom agent and a native product should not independently apply emergency rules, create records from the same request, route the same issue to different vendors, or send conflicting resident updates. Assign one owner to each decision and field transition.

Document ownership for:

  • Resident intake and required information
  • Emergency classification and after-hours routing
  • Work-order creation and duplicate detection
  • Priority and schedule changes
  • Vendor assignment and spending authority
  • Access instructions and resident communications
  • Completion evidence
  • Final closeout acceptance

If the team cannot identify which system owns a step, do not expand automation at that step.

Require Evidence Before Write Authority

A maintenance agent should earn additional authority through test results, not a polished demonstration.

Test representative cases before allowing the agent to create or update records. Include normal requests, incomplete requests, duplicates, after-hours activity, possible emergencies, conflicting data, vendor exceptions, and requests that should be escalated.

Verify results at the field level:

  • Correct property and unit
  • Correct priority or escalation to a named person
  • Correct handling of permission to enter
  • Appropriate vendor context without unauthorized assignment
  • Correct schedule and status treatment
  • Clear resident communication
  • Completion evidence captured or flagged as missing
  • No duplicate work order created
  • Correct exception routing

OWASP guidance reviewed September 11, 2026, recommends minimum necessary tool access, separate read and write scopes, explicit authorization for sensitive operations, human involvement in high-risk actions, output validation, structured adversarial testing, regression testing after material changes, and retained validation evidence.

NIST's AI RMF Core, reviewed the same date, calls for documented application scope, potential costs from errors, operator proficiency processes, and human-oversight processes.

For a maintenance pilot, retain human judgment over emergency decisions, nonstandard resident situations, vendor and spending exceptions, operating-rule changes, and final acceptance of completed work. Narrower delegation should follow explicit testing and approval, not precede it.

Coaching Exercise: Prove One Recurring Workflow

Do not begin with a portfolio-wide agent. Begin with one recurring task that is common enough to measure and bounded enough to supervise.

A practical candidate is missing-information follow-up for nonemergency maintenance requests.

Define Done

Write the required outcome before testing. It might include the correct resident and unit, a clear request for specific missing information, defined escalation triggers, a saved review record, and no work-order change without human approval.

Establish a Baseline

Use a representative sample of the current process. Measure:

  • Median handling time
  • Required-field completeness
  • Corrections or rework
  • Exception rate and exception type

Do not claim savings before local results support them.

Test a Bounded Workflow

Begin with supplied facts or read-only access. Allow the tool to draft a message or prepare a packet. Do not grant write authority because it makes a demonstration look more complete.

Include Difficult Cases

Test wrong unit numbers, duplicate requests, incomplete access information, emergency language, and records that conflict with the current request.

Easy cases show whether the tool can follow a pattern. Difficult cases show whether the operating boundary holds.

Review the Finished Work

Compare the output against the baseline:

  • Was the work faster?
  • Was it more complete?
  • How often did a person correct it?
  • What rework did it create?
  • Which exceptions appeared?
  • Did each exception reach the right owner?

Then decide whether to expand, hold, revise, or stop. Expansion is not the default. Ending a weak pilot is also evidence of sound operating judgment.

The Practical Next Step

Choose one recurring communication or maintenance task and create a one-page pilot charter before connecting more tools.

Include the system of record, definition of done, baseline measures, permitted tools and fields, prohibited actions, exception triggers, named human approver, test cases, and evidence to retain.

Start with read-only comparison or drafting. Require demonstrated results before expanding authority to create, update, or send.

The Agents API may make custom agent construction more accessible. It does not remove the need to prove each handoff. A disciplined pilot lets an owner assign authority, name exception owners, test outputs, and require evidence before an agent takes on a larger role across the portfolio.

Keep exploring

Ask Jane about this article.

Connect the article to GroundHaven consulting or current market evidence.

Continue with Jane
AI Management Consulting

Put perspective to work.

Discuss your priorities and budget. Build a plan around the support your team needs.

Explore consulting options