Proprietary Evaluation Methodology

LEX-EVAL™

A human-led Legal AI Evaluation Framework designed to distinguish fluent language from actual legal reliability.

“High linguistic quality is not evidence of legal reliability.”

Decision Layer

90–100

Expert Grade

80–89

Strong

70–79

Conditional

50–69

High Risk

0–49

Unreliable

CLF

Overrides score

Reliability outcome: PASS · PASS WITH REVIEW · FAIL-CRITICAL.

100-Point Architecture

Seven dimensions. One auditable evaluation record.

Legal Accuracy

25 pts

Is the substantive legal answer correct?

Authority & Citation Integrity

15 pts

Are authorities real, relevant, current, and properly represented?

Issue Identification

10 pts

Does the output identify the legal issues that actually matter?

Legal Reasoning

20 pts

Does the analysis connect facts, rules, authorities, and conclusions coherently?

Jurisdictional Alignment

10 pts

Does the answer remain within the correct legal system and applicable framework?

Hallucination & Fabrication Control

10 pts

Does it avoid invented authorities, facts, rules, and unsupported certainty?

Professional Usefulness

10 pts

Can a professional reviewer use the output efficiently and safely?

Critical Legal Failures

A numerical score cannot neutralize a Critical Legal Failure. LEX-EVAL™ separates aggregate quality from failures that can make professional reliance unsafe.

CLF-01Fabricated Authority
CLF-02Material Misstatement of Law
CLF-03Wrong Jurisdiction
CLF-04Invalid or Obsolete Authority
CLF-05Material Fact Distortion
CLF-06Unsafe Professional Reliance

Evaluation Deliverables

From score to corrective intelligence.

Weighted scorecard and reliability outcome

Authority and citation verification

Error taxonomy and severity classification

Critical Legal Failure assessment

Corrected or gold-standard answer

Aggregate benchmark analytics

Evaluate a legal AI workflow