Linguistic AI Architecture • Scoring Specification 2026

How Trio's 4-Dimensional AI Engine Calculates English Scores

A comprehensive technical breakdown of Trio's 7-part 63-question assessment architecture, acoustic neural diagnostics, NLP syntax tree parsing, configurable decision matrix, and CEFR-aligned composite weighting algorithms.

63 Items
7 Distinct Parts
Calibrated CEFR & standardized assessment battery
4 Pillars
Core Competencies
Speaking, Listening, Grammar & Lexicon
0 – 100
Composite Scale
Granular precision mapped to A1–C2
< 30s
Scoring Latency
Instantaneous multi-modal AI evaluation
Diagnostic Pillars

The 4 Core Scoring Dimensions

Trio evaluates English proficiency not as an arbitrary flat grade, but across four mathematical linguistic dimensions calibrated to international CEFR workplace standards.

35% COMPOSITE WEIGHT

1. Speaking Fluency & Acoustic Phonetics

Neural speech models process raw audio spectrograms at 16kHz to diagnose pronunciation clarity, rhythm, cadence, and vocal hesitation.

  • Phoneme-Level Acoustic Matching: Formant frequencies (F1/F2) and vowel lengths are compared against native acoustic models.
  • Speech Cadence & WPM: Optimal professional speaking rate calibrated between 120–160 Words Per Minute.
  • Hesitation & Pause Penalties: Unnatural pauses (>1.2s) and excessive filler sounds ("uh", "um") are mathematically deducted.
  • Fed by: Part C (Read Aloud) and oral response fluency components.
25% COMPOSITE WEIGHT

2. Listening Comprehension & Auditory Memory

Measures auditory attention, phonological loop retention, and conversational parsing using strict single-play locked audio dialogues.

  • Single-Play Strictness: Eliminates artificial replay advantage to evaluate real-time cognitive absorption.
  • Acoustic Sentence Repeat: Word Error Rate (WER) scoring on dictation typing directly measures phonological storage.
  • Accent Resilience: Audio corpus incorporates standard US, UK, Australian, and global workplace dialects.
  • Fed by: Part D (Listening Dictation) and Part E (Listening Comprehension).
20% COMPOSITE WEIGHT

3. Grammar, Syntax Trees & Morphological Mastery

Evaluates syntactic hierarchy, subordinate clauses, subject-verb agreement, and sentence structure building under time constraints.

  • Levenshtein & LCS Sequence Scoring: Mathematical evaluation of word permutations against target grammar parse trees.
  • Tense & Morphological Agreement: Flags verb conjugation, modal auxiliaries, and dependent clauses.
  • Reading Syntax Extraction: Measures comprehension of formal corporate policy clauses and conditions.
  • Fed by: Part A (Sentence Builds) and Part F (Reading Comprehension).
20% COMPOSITE WEIGHT

4. Vocabulary, Lexical Breadth & Pragmatics

Measures the sophistication, contextual suitability, Type-Token Ratio (TTR), and pragmatic suitability of candidate expression.

  • CEFR Vocabulary Profiler: Measures proportion of A1–A2 basic words vs B2–C2 high-value enterprise vocabulary.
  • Type-Token Ratio (TTR): Evaluates lexical diversity in written responses to prevent repetitive phrasing.
  • Pragmatic Register & Keywords: Evaluates polite, professional workplace communication in written emails.
  • Fed by: Part B (Vocabulary & Synonyms) and Part G (Written Expression).
Assessment Structure

The 7 Test Parts (63-Question Battery)

Every candidate completes a standard 63-question assessment battery divided into 7 psychometrically calibrated parts, designed to systematically test every layer of spoken and written English.

A

Part A: Sentence Builds

Grammar & Syntactic Mastery
10 Questions Weight: 20% ~3.5 Mins

Candidates are presented with 3 scrambled phrase groups (e.g. ["was delayed", "the morning dispatch", "by heavy traffic"]) and must tap the blocks in the exact sequence to construct a grammatically correct, meaningful English sentence.

Evaluation Algorithm

Longest Common Subsequence (70%) + Levenshtein Word Distance (30%) with non-linear strict power curve.

Core Competency

Syntactic order, dependency parsing, phrase structure rules.

Scoring Range

Exact match: 100 pts • Partial order: 50–90 pts • Incorrect: 0 pts.

B

Part B: Vocabulary & Synonym Selection

Lexical Resource & Precision
10 Questions Weight: 15% ~3.0 Mins

Candidates evaluate a target vocabulary prompt and select the closest semantic match from 4 carefully calibrated options spanning business terminology, academic verbs, and workplace collocations.

Evaluation Algorithm

Normalized exact choice match against CEFR word-frequency tiered key.

Core Competency

Contextual semantic precision, collocation awareness, advanced lexicon.

Scoring Range

Correct option: 100 pts • Incorrect option: 0 pts.

C

Part C: Read Aloud & Pronunciation

Speaking, Acoustic Cadence & Phonetics
8 Questions Weight: 15% ~4.0 Mins

Candidates read aloud full operational policy passages and customer interaction texts into their microphone. The AI engine captures acoustic telemetry, evaluates phone-level alignment, and measures speech tempo.

Evaluation Algorithm

Acoustic Word Error Rate (WER $\le 0.15$), WPM cadence (120–160), and neural acoustic confidence multiplier.

Core Competency

Pronunciation clarity, connected speech rhythm, vowel formants, stress.

Scoring Range

Continuously scored 0–100 based on $(1 - \text{WER} \times 2.2)^{1.5} \times \text{Confidence}$.

D

Part D: Listening Dictation & Sentence Repeat

Auditory Memory & Typing Transcription
12 Questions Weight: 10% ~4.5 Mins

Candidates listen to short spoken sentences with strict single-play enforcement (Audio plays ONLY ONCE). Candidates must immediately type the exact sentence with correct spelling and punctuation.

Evaluation Algorithm

String Levenshtein edit distance & Word Error Rate: $(1 - \text{WER} \times 2.0)^{1.6} \times 100$.

Core Competency

Phonological loop storage, auditory decoding, orthographic transcription.

Scoring Range

Threshold $\ge 75$ for full credit; scaled continuously based on edit distance.

E

Part E: Listening Comprehension

Conversational Dialogues & Gist Inference
10 Questions Weight: 15% ~4.0 Mins

Candidates listen to audio recordings of workplace consultations, customer escalations, and dispatch announcements (1 playback only), then answer multiple-choice questions assessing both main ideas and granular details.

Evaluation Algorithm

Objective key matching with single-play timestamp validation.

Core Competency

Gist inference, rapid information filtering, international accent tolerance.

Scoring Range

Correct option: 100 pts • Incorrect option: 0 pts.

F

Part F: Reading Comprehension

Policy Passages & Logical Deduction
8 Questions Weight: 15% ~3.5 Mins

Candidates read technical operations memos, privacy regulations, customer dispute policies, and workflow procedures, answering contextual questions that test comprehension depth beyond simple keyword search.

Evaluation Algorithm

Multi-option contextual verification engine.

Core Competency

Skimming, scanning, logical deduction, nuanced business policy interpretation.

Scoring Range

Correct option: 100 pts • Incorrect option: 0 pts.

G

Part G: Written Expression & Situational Response

Business Pragmatics, Lexical TTR & Email Composition
5 Questions Weight: 10% ~5.0 Mins

Candidates are given real-world customer service scenarios (e.g. "A customer expresses frustration regarding a delayed pickup due to unexpected rain. Write a polite 2-3 sentence email response...") and compose formal, polite responses.

Evaluation Algorithm

Combined ratio of Length Volume (50%) + Type-Token Ratio (30%) + Domain Keyword Coverage (20%).

Core Competency

Professional register, diplomatic tone, syntactic complexity, lexical variety.

Scoring Range

Graded 0–100 via $(\text{Length} \times 0.50 + \text{TTR} \times 0.30 + \text{Keywords} \times 0.20)^{1.6} \times 100$.

Calculation Engine

The Scoring Matrix & Decision Engine

How raw item scores transform into CEFR certifications and automated hiring decisions via Trio's multi-tenant configurable decision matrix.

STEP 01

Item-Level Strict Scoring

Every response across all 63 items is evaluated by specialized micro-engines (Speech WER, Levenshtein, MCQ verification, NLP TTR) to generate exact raw scores from 0 to 100.

STEP 02

Section Subscore Aggregation

Scores for each of the 7 sections ($S_A, S_B, \dots, S_G$) are averaged and combined into the 4 primary competency pillars (Speaking 35%, Listening 25%, Grammar 20%, Vocabulary 20%).

STEP 03

Decision Matrix & Scaling

The weighted composite score is mapped against standardized psychometric scales (Trio Scaled Score 20–80, CEFR Tiers A1–C2, Global Proficiency Bands) and recruiter-tuned hiring recommendation cutoffs.

Mathematical Formula for Composite Score

Composite Score = ∑(Weightk × SubScorek) / ∑Weightk
Default Weights: Sentence Build (20%) + Vocab (15%) + Reading (15%) + Listening MCQ (15%) + Read Aloud (15%) + Dictation (10%) + Written (10%) = 100%
Formula Engine

Interactive Scoring Simulator

Test the composite formula in real time: Adjust section sliders to observe how Trio's weighting algorithm calculates composite scores, standardized equivalencies, and assigns hiring recommendations.

Speaking & Phonetics (35% Composite Weight)
88%
Listening Comprehension (25% Composite Weight)
82%
Grammar & Syntax (20% Composite Weight)
90%
Lexical Breadth & Vocabulary (20% Composite Weight)
85%
LIVE FORMULA: (0.35 × Speaking) + (0.25 × Listening) + (0.20 × Grammar) + (0.20 × Lexical)
CALCULATED CEFR TIER
C1
86.3 / 100
Proficient User — Advanced
Trio Scaled 72 / 80
Global Band Band 8.0
Decision Strong Hire
Spontaneous, fluent expression with sophisticated syntax and accurate pronunciation in executive and technical domains.
Global Standardization

CEFR Calibration & Benchmark Matrix

How Trio composite score ranges translate into Common European Framework of Reference (CEFR) levels and corporate hiring readiness.

CEFR Level Score Range (0–100) Trio Scaled (20–80) Global Band Hiring Recommendation Workplace Proficiency Profile
C2 95 – 100 78 – 80 Band 8.5 Strong Hire Mastery — Flawless, native-like spontaneous expression, idiomatic precision, and executive nuance.
C1 85 – 94 68 – 77 Band 8.0 Strong Hire Effective Operational Proficiency — Complex arguments, fluent discourse, high grammatical accuracy.
B2 70 – 84 54 – 67 Band 6.5 – 7.0 Hire Vantage / Upper Intermediate — Independent workplace communication with native speakers without strain.
B1 55 – 69 42 – 53 Band 5.0 – 6.0 Conditional Hire Threshold / Intermediate — Routine workplace tasks, basic client support, simple descriptive email drafting.
A2 40 – 54 30 – 41 Band 4.0 Do Not Hire Waystage / Elementary — Basic routine interactions, slow speech comprehension, frequent grammatical breakdowns.
A1 0 – 39 20 – 29 < Band 4.0 Do Not Hire Breakthrough / Beginner — Isolated memorized words, severe hesitation, unable to sustain workplace dialogue.