Skip to main content
Explore Products
Orbyt Jobs
Orbyt Intelligence
Orbyt One
Orbyt Jobs
Overview
The job search CRM. Free forever.
Features
Every tool in the CRM
Compare
Against the alternatives
Pricing
Free forever, paid when you outgrow it
API
23 endpoints, MCP native
Salaries
Comp data inside the CRM
By your situation
Job Search Tracks
15 tracks for your exact moment
For Recruiters
Hiring and comp benchmarking
Orbyt Intelligence
Overview
The salary dataset, and its API.
Features
What the platform does
Compare
Against the alternatives
Pricing
Free tier, then Pro and Ultra
API
20 endpoints, Decision-Ready
Start without a card
Playground
Run a live query
MCP server
Three steps into Claude Code
API docs
Endpoints, auth, and limits
Orbyt One
Overview
One account. Every Orbyt product.
Pricing
What one account costs
Explore Research
Orbyt Collective
Orbyt Collective
Overview
An agent leadership team.
How It Works
The machinery, end to end
Process
How the work actually moves
Leadership
The agent officers
Seat Board
Every seat, and what it is doing
Autonomy Ledger
What it decides without us
Influences
The minds we build like
Published
Research Hub
Papers and field notes.
Agent-Native Dataset Design
Preprint, DOI 10.5281/zenodo.19754393
Governing an Agent Leadership Team
Preprint, Zenodo deposit pending
Explore Developers
Orbyt Jobs
Orbyt Intelligence
Orbyt One
Build
Jobs API Docs
23 endpoints, MCP native
MCP Integrations
Claude Desktop
Connect over MCP.
ChatGPT GPT Actions
Connect as a custom GPT action.
Apple Shortcuts
Connect from Shortcuts.
Zapier / Make.com / n8n
Connect with no code.
OpenClaw
Setup in under a minute.
Across products
Developer Hub
Start here
Orbyt API
The platform API
Build
Intelligence API
20 endpoints, Decision-Ready
Webhooks
Events and delivery
CLI
The terminal client
API Changelog
Every version, dated
Try
MCP Server
Wired into Claude Code in three steps
Playground
Engine response shapes with cURL
Try It Live
One call, one real response
Reference
Reference
The full index
Methodology
How the numbers are made
Engines
What computes each answer
Dataset
What is in it, and where from
Glossary
Every term, defined
Status
Live service health
Across products
Developer Hub
Start here
Orbyt API
The platform API
Explore Resources
Orbyt Jobs
Orbyt Intelligence
Orbyt One
Learn
Interview Prep
Company-by-company question sets
AI Skills Lab
The skills that pay in 2026
Career Guides
Long-form career playbooks
Job Search Articles
Every article on the search itself
Job Board
Curated AI-era roles
Arcade
The job search, as games
Salary data
Salary Explorer
3,445 roles across 81 cities
AI Role Salaries
AI roles, by category
Cities
Comp by metro
Industries
Comp by sector
Compare Salaries
Two roles, side by side
Compare Offers
Side-by-side offer math
Skills Impact
What each skill adds to pay
Salary Projections
Five-year pay forecasts
Free tools
All Free Tools
Every calculator and generator
Cover Letter Generator
Tailored in one pass
Unemployment Calculator
What you are owed, by state
Salary Widget
Embed salary data anywhere
Resume Score
Grade your resume against a role
Salary Calculator
Base, bonus, equity in minutes
Take-Home Calculator
After federal and state tax
Total Comp Calculator
Full compensation math
Data
Data Catalog
Every role, city, and engine
Companies
54 leveling frameworks
Reports
Compensation Reports
Free summary PDF
International
The US, UK, and Canada
United Kingdom
UK salary data
Canada
Canadian salary data
Trust
Trust Center
How the data is governed
Security
Controls and posture
SLA
Uptime and support commitments
Help
Support
Help center and contact
Compare
Orbyt against the alternatives
Explore Books
The Books
Start reading
Cold Start
Read the opening, free.
Unfair Advantage
Read the opening, free.
The series
Book 1: Cold Start
Reviewing
Book 2: Unfair Advantage
Reviewing
Book 3: Human Heartbeat
Writing
Book 4: Without Me
Future
Book 5: Observer Effect
Future
Explore Blog
The Machine Speaks
Categories
AI-Native
AI Agents
AI Engineering
AI Product
AI Design
AI Strategy
AI Leadership
AI Build
Jobs in the AI Era
Latest
Hello, Agent Smith.
Sep 3, 2026
Nobody Pays You for the Code Anymore.
Sep 2, 2026
Self-Healing Is a Euphemism.
Sep 1, 2026
The Machine. It Runs the Company.
Aug 30, 2026
One Lab to Rule Them All
Aug 28, 2026
Explore Pricing
Orbyt Jobs
Orbyt Intelligence
Orbyt One
Plans
Jobs pricing
What each plan includes
All plans
Every product, side by side
Plans
Intelligence pricing
Free, Pro, and Ultra
All plans
Every product, side by side
Billing
All plans
Every product, side by side
Explore Company
About
Who we are
Leadership
One human decides. AI agents advise.
Values
The principles behind the work.
Creed
The company creed.
The story
Building in Public
The numbers behind the work
Skunkworks
iOS, Apple Watch, and Vision Pro.
Contact
Email the team
Orbyt Labs
Products
Research
Developers
Resources
Books
Blog
Pricing
Company
Log inStart
Products
Orbyt JobsOrbyt IntelligenceOrbyt One
Orbyt Jobs
OverviewFeaturesComparePricingAPISalaries
By your situation
Job Search TracksFor Recruiters
Orbyt Intelligence
OverviewFeaturesComparePricingAPI
Start without a card
PlaygroundMCP server
Orbyt One
OverviewPricing
Research
Orbyt Collective
Orbyt Collective
OverviewHow It WorksProcessLeadershipSeat BoardAutonomy LedgerInfluences
Published
Research HubAgent-Native Dataset DesignGoverning an Agent Leadership Team
Developers
Orbyt JobsOrbyt IntelligenceOrbyt One
Build
Jobs API Docs
MCP Integrations
Claude DesktopChatGPT GPT ActionsApple ShortcutsZapier / Make.com / n8nOpenClaw
Across products
Developer HubOrbyt API
Build
Intelligence APIWebhooksCLIAPI Changelog
Try
MCP ServerPlaygroundTry It Live
Reference
ReferenceMethodologyEnginesDatasetGlossaryStatus
Resources
Orbyt JobsOrbyt IntelligenceOrbyt One
Learn
Interview PrepAI Skills LabCareer GuidesJob Search ArticlesJob BoardArcade
Salary data
Salary ExplorerAI Role SalariesCitiesIndustriesCompare SalariesCompare OffersSkills ImpactSalary Projections
Free tools
All Free ToolsCover Letter GeneratorUnemployment CalculatorSalary WidgetResume ScoreSalary CalculatorTake-Home CalculatorTotal Comp Calculator
Data
Data CatalogCompanies
Reports
Compensation ReportsInternationalUnited KingdomCanada
Trust
Trust CenterSecuritySLA
Help
SupportCompare
Calculators and tools
Free ToolsSalary CalculatorTake-Home CalculatorTotal Comp CalculatorCompare OffersSkills ImpactSalary Projections 2030Resume ScoreCover Letter GeneratorSalary WidgetUnemployment CalculatorAI Skills Assessment
Books
The Books
Start reading
Cold StartUnfair Advantage
The series
Book 1: Cold StartBook 2: Unfair AdvantageBook 3: Human HeartbeatBook 4: Without MeBook 5: Observer Effect
Blog
The Machine Speaks
Categories
AI-NativeAI AgentsAI EngineeringAI ProductAI DesignAI StrategyAI LeadershipAI BuildJobs in the AI Era
Latest
Hello, Agent Smith.Nobody Pays You for the Code Anymore.Self-Healing Is a Euphemism.The Machine. It Runs the Company.One Lab to Rule Them All
Pricing
Orbyt JobsOrbyt IntelligenceOrbyt One
Plans
Jobs pricingAll plans
Plans
Intelligence pricing
Company
About
Who we are
LeadershipValuesCreed
The story
Building in PublicSkunkworksContact
StartAlready have an account? Log in
  1. Home/
  2. Research/
  3. Governing an Agent Leadership Team
Preprint. CC BY 4.0. Sep, 3 2026

Governing an Agent Leadership Team: Guardrail, Kill-Switch, and Liveness Patterns from a One-Human Autonomous Organization.

By Justin Bartak. ORCID 0009-0005-2615-3624

DOI: Zenodo deposit pending. A DOI will be added on deposit.

Download PDF →Failure corpus (JSON)Autonomy ledger (JSON)
12
Named agent seats, registry read 2026-09-02
66
Documented failures, corpus read 2026-09-02
65
Seat runs, 2026-07-21 to 2026-09-01
45
Logged decisions, 2026-07-13 to 2026-09-02
1
Global halt, set and lifted 2026-08-20

Abstract

Half of the documented engineering failures at one company running an agent organization, 33 of 66 in a single-rater, agent-coded corpus, were failures of its own instruments: tests, guards, gates or metrics that could not fail, could not see, or measured a proxy. We operate the organization studied: twelve named seats, nine running a language model, hold a reporting, proposing and publishing cadence under one accountable human. The corpus (`public/failure-corpus.json`, 2026-09-02) spans 2026-06-09 to 2026-09-02 with eleven undated items; 42 of 66 carry a named countermeasure, and verification gaps are the largest class at 33.

The ledger (`public/autonomy-ledger.json`, 2026-09-02) records 65 seat runs over six weeks, 2026-07-21 to 2026-09-01, at 84.6 percent completion, six of the 55 completions being panel-written observer rows; the frozen 80.0 percent baseline shares its 30 rows, and the 4.6-point gap is within two runs' movement. Forty-five decisions went to an append-only log (`docs/org/decision-log.md`); the organization kill switch was pulled once, on 2026-08-20. The corpus has no agent-failure class and one rater assigned every class, so the share is exploratory rather than a comparison.

The design law that survived is positive observation: a control is live only when its refusal has been observed, and a monitor that could not look reports NOT OBSERVED rather than clean. That verdict is printed; at the layer that escalates, a monitor that could not look and one that saw health are the same event. We contribute a three-tier guardrail model, positive-observation liveness, the published corpus, ten design principles, and a retrofit checklist.

Key findings

Five results from the record, ranked by load-bearing weight on the paper’s thesis. Every number carries the date it was measured.

1. Half of the documented failures were failures of the instruments.
Of 66 failures in the corpus read on 2026-09-02, 33 are verification gaps: tests, guards, gates or metrics that could not fail, could not see, or measured a proxy. The corpus carries no class for agent failure; the agents' own failures sit in the run ledger, on a different denominator.
2. Verification gaps are the largest class, and 42 of 66 carry a named countermeasure.
The three-class taxonomy splits the corpus into 33 verification gaps, 18 operator-process failures and 15 external-behavior failures. 42 items carry a named countermeasure (13 guard scripts, 22 tests, one workflow and six other files) that now stands against the failure, which is 63.6 percent (public/failure-corpus.json, measured 2026-09-02).
3. The seats completed 84.6 percent of 65 runs against a frozen 80.0 percent baseline.
Between 2026-07-21 and 2026-09-01 the ledger recorded 65 runs, 55 completions, 10 failures and 47 discards, printed and never netted out. The baseline of 30 runs and 24 completions was frozen on 2026-08-11 and never regenerated (public/autonomy-ledger.json, read 2026-09-02).
4. The organization kill switch was pulled exactly once, and 45 decisions went to an append-only log.
The kill switch is a file. Its whole history is five commits, with one global halt set and lifted on 2026-08-20 during a domain change. The decision log held 45 entries on 2026-09-02, D-0001 on 2026-07-13 to D-0045 on 2026-09-02, and permits no edits: a correction is a new entry.
5. The design law that survived is positive observation.
A control counts as live only when its refusal has been observed, and a monitor that could not look reports NOT OBSERVED rather than clean. The fence fails unless it watches a real denial, a canary watches the fence, and the dead-man switch judges run conclusions rather than the ledger the watched system writes.

Reproducibility package

The paper reads three published data files, each generated from the organization’s own records and served from this site at a stable URL under CC BY 4.0. The source repository is private, so these three files are the whole surface a reader can re-count without access, and every table in the paper names the command that produced it.

Failure corpus
public/failure-corpus.json
66 documented failures with a three-class taxonomy and the mechanism that now prevents each one where it exists. CC BY 4.0.
Autonomy ledger
public/autonomy-ledger.json
Every recorded seat run: completions, failures and discards per seat, with the frozen baseline and the metric definition verbatim.
Decision ledger
public/decision-ledger.json
The public projection of the append-only decision log, ids and dates parsed mechanically, prose hand reviewed.
Preprint PDF (self-hosted)
/research/agent-leadership-governance/agent-leadership-governance.pdf
The full paper with its appendices. CC BY 4.0. Mirrored to Zenodo on deposit.
Live autonomy ledger page
/orbyt-collective/autonomy-ledger
The same ledger rendered as markup, with the completion rate and its frozen baseline.
Author ORCID
0009-0005-2615-3624
Justin Bartak (Purecraft LLC). Linked Wikidata entity Q139551829.

Cite this paper

Permissive (CC BY 4.0) attribution. Cite the site URL for now. A DOI will be added here and in both citation forms on Zenodo deposit; the URL does not change.

Plain text (APA-style)

Bartak, J. (2026). Governing an Agent Leadership Team: Guardrail, Kill-Switch, and Liveness Patterns from a One-Human Autonomous Organization. Preprint. https://www.orbytlabs.ai/research/agent-leadership-governance

BibTeX

@misc{bartak2026governing,
  author    = {Bartak, Justin},
  title     = {Governing an Agent Leadership Team: Guardrail, Kill-Switch,
              and Liveness Patterns from a One-Human Autonomous Organization},
  year      = {2026},
  note      = {Preprint},
  url       = {https://www.orbytlabs.ai/research/agent-leadership-governance}
}

Where this paper lives

This page is the canonical owned URL with full Schema.org structured data. The permanent DOI arrives with the Zenodo deposit, which is a human step because a DOI cannot be withdrawn.

Orbyt Research (this page)
https://www.orbytlabs.ai/research/agent-leadership-governance
live
PDF (self-hosted)
https://www.orbytlabs.ai/research/agent-leadership-governance/agent-leadership-governance.pdf
live
Published data files
https://www.orbytlabs.ai/failure-corpus.json (and two siblings)
live
Zenodo (primary, DOI)
Deposit pending
pending
ORCID Works
Added on deposit
pending

License

Released under Creative Commons Attribution 4.0 International (CC BY 4.0). You may copy, redistribute, remix, transform, and build on the paper for any purpose, including commercially, with attribution.

The three published data files the paper reads are also CC BY 4.0. Quote, redistribute, build on, with attribution to Bartak, J. (2026) for the paper and Orbyt Labs for the data.

Related on this site

Agent-Native Dataset Design
The first paper in the series: the data product this organization operates. DOI 10.5281/zenodo.19754393.
How the Collective works
The machinery end to end, with the ten-term vocabulary the paper uses.
Autonomy ledger
What the seats decide without us, counted per run and never netted out.
Process
How the work actually moves from a seat run to a committed change.
Leadership
The agent officers, their charters, and the one human who holds the line.

Read the full paper.

The full preprint covers the machine, the guardrail tiers, the watch chain, the empirical record, ten design principles and a retrofit checklist, with three appendices.

Download PDF →All research

Keep Exploring

What we have learned building an AI-native company, what works and what breaks.

Products

  • Orbyt Jobs
  • Orbyt Intelligence
  • Orbyt One

Research

  • Orbyt Collective

Developers

  • Orbyt API
  • Jobs API
  • Intelligence API

Publishing

  • Books
  • Blog

Help

  • Support
  • Contact
  • Status

Company

  • Leadership
  • Values
  • Creed
Products
  • Orbyt Jobs
  • Orbyt Intelligence
  • Orbyt One
Research
  • Orbyt Collective
  • Research
Developers
  • Developer Hub
  • Orbyt API
  • Jobs API
  • Intelligence API
Publishing
  • Books
  • Blog
Help
  • Support
  • Contact
  • Status
Company
  • About
  • Leadership
  • Values
  • Creed
Orbyt Labs™

© 2026 Purecraft LLC  All rights reserved.

Privacy·Terms·Security·Trademark·Accessibility·DPA·Refund·Status·Sitemap

Orbyt Labs, the Orbyt Labs logo, and the Orbyt product names (Orbyt Jobs, Orbyt Intelligence, Orbyt Collective, Orbyt One, Orbyt Books, Orbyt Arcade) are trademarks of Purecraft LLC. Product names, logos, and brands of others are the property of their respective owners. Orbyt Labs is not affiliated with, sponsored by, or endorsed by any third party referenced on this site.