Public draft · scoring disabledOntology 0.1.0How this benchmark is built

The title is spreading.
The operating record still has to hold up.

The benchmark for AI-native Forward Deployed Engineering.

FDE is the operating model. AI-native deployment is the specialization this release examines first.

Benchmark the work. Not the résumé.

A title goes in. A reviewable record comes out.
01“I’m an AI-native FDE.”The claim. A title is only a starting assertion.
  1. 02Do the workHands-on task under constraint
  2. 03Leave the trailCode, tests, and decisions
  3. 04Show the changeAn attributable operating outcome
  4. 05Name the witnessA reference who can corroborate it
06
Reviewable recordWhat did this person actually do?
  • 55 capability familiesprofiled
  • 10 operating domainscovered
  • Headline scorenone
A profile, not a number.

Synthetic illustration of the proposed sequence. No task has been administered and no record has been issued; the status register says exactly what exists.

Field notes

Provisional editorial series · primary-source desk research

01 / 05
Field standard

Everyone is an FDE now. The title tells you almost nothing.

The line is intentionally rhetorical, not a measured prevalence claim. As employers stretch the label across software engineering, deployment, reinvention, and frontline delivery, the defensible question is what work a person can actually prove.

Primary recordsOpenAI · Accenture · Wipro · CognizantEditorial thesis · title expansion is observed semantically, not estimated statistically
7 minRead field note
Editorial method

These field notes interpret first-party records to frame benchmark questions. They do not estimate how many people use the title, how many would qualify, or whether the proposed benchmark predicts performance.

Inspect the research program
NextThe work, mapped.Ten domains and 55 capability families, the vocabulary every field note above is written against.
Capability explorer · ontology 0.1.0

The work, mapped.

Browse the AI-native FDE construct: a portable forward-deployed core, explicit AI-system delivery requirements, and separately versioned context packs. Each domain opens into its released capability families and proposed observable signals.

Open the full benchmark explorer
10 of 10 domains
Domain 01

Forward-Deployed Discovery and Problem Framing

Discover, frame, prioritize, and contract consequential enterprise problems under ambiguity.

Released families
6
Ontology share
11%
Release
0.1.0
Capability familiesReleased vocabulary
  1. 01.01Stakeholder Discovery
  2. 01.02Workflow and Constraint Discovery
  3. 01.03Problem Framing and Decomposition
  4. 01.04Outcome Contract and Acceptance Criteria
  5. 01.05Opportunity Sizing and Prioritization
  6. 01.06Customer Trust and Expectation Management
Observable signal

“Rejects an attractive AI use case when the operating constraint makes it unsafe or uneconomic.”

At a glance

The benchmark by the numbers.

Open the method paper
01

Capability coverage

55families

Released families across the benchmark’s ten domains.

6DISC
7SYST
8AI
6ARCH
7ENT
7PROD
4OUT
4LEAD
2MOD
4PACK
Ontology release 0.1.0
02

Evidence maturity

6levels

An ordinal path from assertion to independent field corroboration.

E0Claim
E1Know
E2Do
E3Artifact
E4Prod
E5Field
E0

Claim strength increases only when the record changes.

E5
Ordinal taxonomy · not a score
03

Source inventory

v0.1release

Current, inspectable assets bundled with the public release.

Capability families55
Bundled public documents99
Released JSON contracts25
Counts resolve from bundled release sources
Outcome architecture · proposed design contract

Three results. Three kinds of authority.

Facts, measurement, and credentialing need different evidence and different decision authority. The benchmark keeps them visibly separate, and any single score stays a research question behind a published authorization gate.

Open the outcome architecture
  1. 01
    Factual layerVerified production record

    What happened, where it happened, and what this person actually contributed.

    Evidence operations
  2. 02
    Measurement layerCapability profile

    A versioned, multidimensional interpretation of observed work—not a résumé score.

    Measurement panel
  3. 03
    Issuance layerGoverned credential tier

    A shareable, verifiable status with an issuer, scope, version, expiry, and recourse path.

    Future non-conflicted credential authority

Future scalar gateA single score is a research question, not a v0.1 product.Read the authorization gate

Research desk8 instruments

The instruments behind the benchmark.

The working product, its public method, and the records that say what is implemented, proposed, or still waiting for evidence.Open the research program

  1. 01
    Observe the workEvaluation workspace
    Prototype
  2. 02
    Inspect the recordEmployer verification
    Prototype
  3. 03
    Three outputs. Three authorities.Outcome architecture
    Proposed
  4. 04
    The benchmark publicationMethod + construct draft
    Public draft
  5. 05
    Where forward deployment is movingField intelligence · 15 sources
    Desk research
  6. 06
    How a chart earns trustData product contract
    Governance spec
  7. 07
    Six views of the workField research observatory
    Pre-fieldwork
  8. 08
    What exists in source todayRelease ledger
    Living record
Release disclosure · evidence cutoff 1 Sep 2026

What this release can support.

Transparency belongs in the record—not on a victory scoreboard. Current absences remain inspectable below.

Public draftv0.1.0
ConstructPublished10 domains · 55 families
Human studyNot startedNo participant evidence admitted
Official scoringNot authorizedNo cut score, tier, or rank
External validationNot recordedIndependent review remains a gate
Open the current evidence boundary

No admitted human-participant record supports a capability, hiring, ranking, tier, or credential interpretation.

No independent validation, accreditation, external-majority commission approval, or reproduction is represented.

Publication readiness is tracked separately from measurement validity; neither can be advanced by interface polish.

Reach a person

Corrections, security reports, study interest, and review offers all go through one form.

Open the contact form