GS-0501 Financial Administration and Program — existing vs. agent-built
Financial Administration and Program
Broad federal financial management: policy, systems, reporting, internal controls, and program-level financial administration.
USA Hire general-competencies assessment measuring non-technical competencies (Reading, Reasoning, Problem Solving, Decision Making, etc.) plus an Occupational Questionnaire on self-reported experience. Technical financial-management KSAs (FFMIA, A-123, TFM, USSGL) are not directly measured.
Job analysis anchored to O*NET / MOSAIC + FFMIA, CFO Act of 1990. Every item carries signed task → KSAO → item provenance. Gated by licensed I-O psychologist sign-off before any applicant sees it.
Build this assessment →Eight-dimension comparison
Each axis is scored 0–100 using the rubric in comparisonMetrics.ts. Existing numbers come from the researched record; agent-built numbers reflect the swarm's uniform quality gates.
Chance to Compete Act (Pub. L. 118-19) three-prong test
- §3(2)(A) Allows applicants to demonstrate job-related skills
§3(2)(A) — the assessment must permit an applicant to demonstrate job-related knowledge, skills, abilities, or competencies, not merely attest to them.
Existing: partialAgent-built: yes - §3(2)(B) Is based on a job analysis
§3(2)(B) — content must be derived from a current job analysis that identifies the tasks and the KSAOs required to perform them.
Existing: yesAgent-built: yes - §3(2)(C) Is not principally reliant on a self-assessment
§3(2)(C) — the assessment "does not solely include or principally rely upon a self-assessment from an automated examination."
Existing: partialAgent-built: yes
Technical KSAO coverage gap
Of the 5 critical technical KSAOs identified in the job analysis for this series, how many does each assessment directly measure (not self-report)?
Per-dimension detail
No series-specific technical KSAOs are directly measured — assessment is non-technical or inferred.
14-agent swarm maps every task → KSAO → item; ≥90% of critical technical KSAOs directly measured, no self-report proxies.
DIF analysis: yes; adverse impact studied.
Per-item bias sensitivity review + pilot DIF (Mantel-Haenszel / logistic regression) + adverse-impact 4/5ths pre-deployment simulation.
content / criterion / construct validity evidence; no public technical report.
Auto-drafted 29 CFR §1607.15 packet: job analysis, content validity matrix (Lawshe CVR), technical report, signed audit log.
3 modalities: Situational Judgment, Reading / reasoning items, Self-report questionnaire.
Blueprint enforces minimum 3 of 5 RFI modalities where content supports it (JKT + Work Sample + SJT / Simulation / SI).
Partial traceability — validation evidence exists but item-level provenance is not public.
Every item carries a cryptographically signed provenance record: task statement → KSA → criticality/frequency → item.
Three-prong test flags: (A) partial, (C) partial.
Three-prong compliance enforced at the blueprint gate: (A) skill demonstration, (B) job-analysis anchored, (C) ≤10% self-rating by weight.
mobile-compatible; Section 508 VPAT available; FK grade 11.
Tailwind-responsive, Section 508 AA verified, Flesch-Kincaid target grade 9–11, accommodation paths pre-wired.
Aggregate score returned; sub-score breakdown and rationale generally not available to applicants.
Each score includes IRT theta + SEM + KSAO sub-scores + plain-English rationale (NIST AI RMF explainability).
Bias & fairness findings
Drawn from the researched record for this series. Severity follows EEOC / NCME Standards guidance.
- No public technical report
External defensibility review requires a public §1607.15 report. Agent-built runs auto-publish the technical report with the final assessment.
- Historical concern
USA Hire replaced ACWA, which replaced PACE (struck down in Luevano v. Campbell, 1981, 29 EPD ¶32,860 — consent decree for adverse impact against Black and Hispanic applicants)
- Historical concern
Subsequent general-cognitive screens have repeatedly faced adverse-impact scrutiny; general cognitive ability tests produce ~1 SD subgroup differences in meta-analytic research
