Deming’s List

The FDE market map: companies, verified roles, and specialist search.

Employer evaluation guide

The FDE hiring scorecard.

Six job-related signals, direct evidence, explicit unknowns, and a structured interview loop—without reducing the role to title, pedigree, or charisma.

By Andre Becker NortonPublished and reviewed

A scorecard cannot rescue an uncalibrated role.

Start with the six-month outcome, code boundary, customer phase, operating constraints, and definition of reusable leverage.

Before reviewing candidates, write one observable six-month outcome; the signals that require direct rather than adjacent evidence; which constraints are essential, learnable, or merely preferred; and the interview stage responsible for resolving each unknown.

If the hiring team disagrees about whether the job is FDE, Solutions Engineering, Implementation Engineering, Applied AI Engineering, or deployment leadership, use the role-design matrix first.

Score evidence, not confidence.

1. Production engineering

Look for: attributable code, architecture, integrations, reliability work, debugging, or operational ownership.

Test: ask for a system boundary, failure mode, tradeoff, and what changed after launch.

2. Problem framing

Look for: turning an ambiguous request into a bounded technical problem and explicit success criteria.

Test: give an under-specified customer problem and observe the questions asked before proposing a stack.

3. Customer judgment

Look for: domain discovery, stakeholder translation, expectation management, and honest communication under pressure.

Test: explore a moment when technical truth and customer urgency conflicted.

4. End-to-end ownership

Look for: responsibility spanning discovery, design, build, rollout, adoption, stabilization, or operation.

Test: map exactly where ownership began, ended, and transferred.

5. Learning under constraints

Look for: sensible decisions amid legacy systems, security, regulation, sparse data, travel, or compressed timelines.

Test: ask what constraint changed the design and what evidence justified the response.

6. Reusable leverage

Look for: product fixes, components, tooling, documentation, evaluation systems, or operating patterns that outlast one engagement.

Test: ask what the second deployment could do faster because of the first.

Separate missing evidence from weak evidence.

Use the same anchored scale for every candidate on the same role.
RatingMeaningWhat to record
0 · Not assessedThe interview did not produce evidence relevant to this dimension.Which stage will resolve it; do not turn absence into a negative score.
1 · WeakEvidence is indirect, overly assisted, or contradicts the job requirement.The exact behavior or artifact and why it matters to the role.
2 · AdjacentRelevant potential exists, but the candidate has not yet demonstrated the required scope.Transferable evidence and the risk that remains.
3 · DirectThe candidate has demonstrated the required behavior at comparable scope.Specific situation, action, outcome, and verification source.
4 · ExceptionalEvidence exceeds the role need and shows repeatable judgment across contexts.Why the evidence is exceptional rather than merely impressive.

No compensating averages by default

If production ownership is essential, a high communication score should not quietly cancel a failure on production engineering. Define non-compensable requirements before interviews begin.

Give each stage a job.

  1. Evidence screen

    Map one shipped system, the candidate’s exact ownership, customer context, outcome, and unresolved claims. Avoid generic résumé walkthroughs.

  2. Technical working session

    Use a realistic deployment problem. Assess decomposition, interfaces, data, reliability, evaluation, security, and tradeoffs—not trivia detached from the work.

  3. Customer discovery simulation

    Provide incomplete requirements and a stakeholder with conflicting incentives. Score questions, boundary setting, translation, and response to pushback.

  4. Deployment retrospective

    Deep-dive a real project: what failed, what was learned, how rollout changed, and what became reusable.

  5. Constraint and motivation conversation

    State travel, location, customer presence, on-call, compensation, and operating reality plainly; assess fit without hiding the burden.

  6. Independent written debrief

    Interviewers record evidence and a rating before group discussion, then resolve contradictions and unknowns explicitly.

Questions that expose the ownership boundary.

System

  • What did you personally design, build, and operate?
  • What broke after the demo?
  • Which technical tradeoff would you reverse now?

Customer

  • What did the customer ask for, and what did they actually need?
  • When did you say no or narrow scope?
  • How did you know adoption was real?

Learning

  • What did the first deployment teach the product team?
  • What became reusable?
  • What remained a one-off, and why?

Unknowns

  • Which claim still needs verification?
  • What context made the outcome possible?
  • Where did someone else own the hard part?

What the process should prevent.

  • Title matching: assuming every current FDE is qualified and every adjacent title is irrelevant.
  • Charisma substitution: rewarding polished customer presence without testing production depth.
  • Algorithm substitution: over-weighting generic coding puzzles that do not resemble deployment work.
  • Hidden burden: disclosing travel, onsite cadence, on-call, clearance, or customer intensity late.
  • Memory scoring: writing impressions after the group discussion instead of independent evidence.
  • Unstructured references: asking only whether someone was “great” rather than verifying ownership and outcomes with permission.

Role evidence plus structured assessment guidance.

Hiring an FDE?

Turn the role into an inspectable search.

Bring the six-month outcome and the signal your current funnel fails to test.