Evidence first
Scores are anchored to answer-specific evidence and turn IDs, not impressionistic prose.
Scoring methodology · v2.0
HireEGG converts observable interview evidence into role-weighted competency estimates. The number is reproducible, the uncertainty is visible, and every conclusion can be traced back to the candidate's answers.
The short answer for enterprise teams
AI finds evidence. Versioned code calculates the score. People make the hiring decision.
Scores are anchored to answer-specific evidence and turn IDs, not impressionistic prose.
The same evidence, weights, and methodology version produce the same score.
Thin evidence widens intervals and can withhold a rating instead of creating false precision.
From answer to score
No hidden end-to-end model decides employability. Each layer has a defined job and a reviewable output.
Each answer is stored as a turn-level evidence item: the question, intended competency, a short evidence quote, structured strong and weak signals, and evidence strength.
Evidence is mapped to technical depth, communication, problem solving, and domain knowledge. One answer can contribute to more than one competency, but its relevance is explicit.
Specific, relevant answer evidence receives more weight. Thin or missing evidence receives less weight; it does not become a confident low or high score.
Versioned code applies configured competency weights and Bayesian shrinkage. The language model cannot set, edit, or override the score.
Every score is paired with evidence coverage, reliability, cited turn IDs, and a 90% interval. If coverage is too low, HireEGG withholds the final readiness rating.
The mathematical core
For each competency, HireEGG combines evidence strength and relevance into an effective evidence weight. The observed estimate is then shrunk toward a neutral prior of 50 according to reliability.
competency = reliability × observed evidence
+ (1 − reliability) × 50
overall = Σ(normalized weight × competency)
This structure prevents a sparse interview from appearing more certain than the evidence supports. The report also publishes a 90% interval around the estimate.
Default role profile
Technical depth
35%
Role-relevant technical judgment and implementation detail
Communication
25%
Clarity, structure, precision, and stakeholder awareness
Problem solving
20%
Trade-offs, failure analysis, validation, and decision quality
Domain knowledge
20%
Applied understanding of the role's working context
What the report exposes
A role-weighted performance estimate, or withheld when evidence is insufficient.
Turn IDs, evidence counts, and answer excerpts behind each competency.
A visible range communicating uncertainty rather than a falsely exact number.
Separate measures showing how much relevant evidence was captured and how stable it is.
Integrity is separate
Browser and companion-app engines produce their own event timeline, graph briefs, sample counts, flagged rates, data-quality labels, limitations, and cross-device correlations.
Integrity risk never changes a performance or competency score.
A sensor event can warrant review; it cannot establish intent by itself.
Single-device events are distinguished from time-correlated browser and phone evidence.
Elevated risk routes the session to evidence review rather than automatic rejection.
Enterprise questions
Deterministic, versioned code decides every number. AI is limited to extracting and summarizing evidence from the interview. It cannot alter competency weights, thresholds, reliability, intervals, or the final score.
It is the weighted estimate of demonstrated competency evidence for the configured role and level—not a personality score and not a guarantee of future performance. HR also sees its interval, coverage, reliability, and the exact cited turns behind it.
Uncertainty should not masquerade as confidence. Bayesian shrinkage pulls thin-evidence estimates toward a neutral prior, while stronger independent evidence allows the estimate to move farther away from neutral.
No. Performance and integrity are separate tracks. Integrity telemetry can route a session to human review, but it cannot change an answer score, competency score, or overall performance score.
Not yet. HireEGG will not publish a percentile until it has a sufficiently large, role-and-level-specific cohort linked to validated outcomes. A fabricated or mixed-cohort percentile would be misleading.
No. HireEGG provides decision support. Low evidence, low performance, or elevated integrity signals require human review of the cited evidence before an employment decision.
Validation status
This is HireEGG's internal evidence-based methodology. HireEGG does not currently claim independent predictive-validity certification, IO-psychology certification, or validated cross-employer percentiles. The validation program includes frozen rubrics, double-scored calibration sets, inter-rater agreement, role-specific outcome studies, and subgroup monitoring where lawful and ethical.