VLAP · Vasl Language Analysis Platform

The clinical AI
the field was
missing.

Standard NLP models are trained on majority-White internet text. They were never designed to read how Black and Latino, LGBTQ+, and first-generation youth communicate distress — and they consistently fail to. VLAP was built specifically for those communities. Not adapted. Built.

High-Distress Signal Sensitivity
90%

On high-distress signal detection in active IRB validation with a university research partner. Specificity and predictive-value results will be published with the completed study — because half a statistic isn’t transparency.

Active validation · Results pending · Non-diagnostic output
2,400+
AAVE and youth vernacular tokens added beyond standard BERT-family vocabulary
47
Clinical signals across 6 categories
5
Cultural communities represented in the training annotation cohort
Chapter 01 — The Problem Standard NLP Cannot Solve

The model
never learned
their language.

General-purpose NLP models — including large language models — are trained on internet text that skews heavily toward majority-White, educated, English-speaking populations. The result is a systematic blind spot: the specific ways Black and Latino, LGBTQ+, and first-generation youth signal emotional distress are consistently misread, deprioritized, or missed entirely.

This isn't a model failure in the traditional sense. These models perform well on the populations they were trained for. The failure is in deploying them, uncorrected, for populations they were never trained on — and expecting accurate clinical signal detection to follow.

The gap isn't theoretical. In clinical practice, it means a youth saying "lowkey been struggling fr" is read as casual. A pre-disclosure minimization pattern ("it's not that deep but...") reduces the model's confidence instead of increasing it. Community-developed coded language — terms built specifically to avoid content filters — registers as noise. Standard models consistently miss these signals in the communities that most need them caught.

"lowkey been struggling fr, ain't nobody understand what I'm going through"
Standard NLP reads
Low confidence signal. "lowkey" reduces severity weight. Grammatical irregularity reduces confidence further. Output: insufficient signal for flagging.
VLAP reads
CCM-02 AAVE minimization construction. "fr" authenticity escalator contradicting the minimization hedge. HOP-03 futility framing in AAVE. Elevated signal — surfaces to care team.
"it's not that deep but lowkey been struggling since school started."
Standard NLP reads
Negation ("not that deep") reduces signal weight. Mild concern flagged, low priority.
VLAP reads
CCM-04 classic pre-disclosure frame. Negation before authentic disclosure is the signal, not evidence against it. ISO-04 temporal onset marker — distress tied to a specific starting point.
"been thinking about unaliving lately ngl."
Standard NLP reads
"unaliving" not in vocabulary. Processed as unknown token or misclassified. Crisis signal missed entirely.
VLAP reads
SHA-02 coded suicidal ideation — "unaliving" is in the extended vocabulary. "ngl" is a CCM sincerity marker confirming non-performative disclosure. Per escalation rules, this combination triggers immediate clinical supervisor review.
Chapter 02 — How VLAP Works

Language is culture,
not pathology.
Interpretation is care.

What VLAP reads

VLAP analyzes the language members choose to share on the platform — check-ins, peer-group messages, coach conversations. It looks for 47 clinical signals across 6 categories: patterns like pre-disclosure minimization, withdrawal, and escalating distress, expressed the way young people actually express them.

What VLAP separates

Every detection is sorted into four distinct layers: cultural markers (expression that’s normal in a community’s context), linguistic patterns, behavioral indicators, and clinical relevance. Cultural expression is never treated as a symptom. That separation is the whole point.

What VLAP never does

It never diagnoses. It never labels or scores a young person. It never sends alerts to police, principals, or parents. It never shows individual conversations to schools or funders. Every escalation goes to one place only: a licensed clinician, who reviews the full context before anything happens.

Who sees what

Members see their own experience. Coaches see the context they need to reach out well. Licensed clinicians see signal context before sessions. Organizations see only aggregate, de-identified trends — never an individual young person.

How we know it works

Our signal accuracy and member outcomes are being independently validated through an IRB-approved study with a university research partner. We publish our measurement definitions, and when the study concludes, we’ll publish what it finds — including accuracy across demographic groups.

Written for a parent as much as a clinician — see the plain-language version for families.

Chapter 03 — VLAP Architecture

Built on BERT-family.
Fine-tuned for
cultural fluency.

VLAP is a fine-tuned BERT-family transformer model, trained on a purpose-built corpus of culturally specific mental health language. The BERT-family architecture was selected for its bidirectional context processing — essential for reading the layered meaning in code-switching, minimization patterns, and culturally framed expressions where individual words carry different weight depending on surrounding context. The model was then extended with a culturally specific vocabulary and fine-tuned against a clinically annotated training corpus.

Foundation
BERT-Family Base Architecture — Bidirectional Transformer

The bidirectional encoding layer processes the full context window simultaneously in both directions — unlike unidirectional models that read left-to-right. This is architecturally necessary for cultural signal detection: the meaning of "lowkey" depends entirely on what follows it, and the clinical significance of a minimization hedge ("it's not that deep but") only becomes clear in context of what comes after the conjunction.

BERT-family base · fine-tuned · behavioral health corpus
Extension
Vocabulary Extension — 2,400+ Cultural Tokens

Standard BERT-family vocabulary was extended with 2,400+ AAVE terms, youth vernacular expressions, code-switching patterns, and community-developed coded language — including terms built specifically to circumvent content filters (e.g., "unaliving"). This extension was compiled through community engagement with BIPOC and LGBTQ+ youth populations and validated by licensed clinicians with relevant community competency. Without this extension, approximately 23% of the training corpus would be processed as unknown tokens by the base model.

Community-sourced · Clinically validated · Continuously updated
Training
Fine-Tuning on Culturally Specific Mental Health Corpus

The extended BERT-family model was fine-tuned on annotated language samples drawn from the communities VLAP serves — not web-scraped text, but specifically collected and clinically annotated mental health communication. Each sample was labeled by licensed clinicians with community competency training against the 47-signal taxonomy across six clinical categories. The fine-tuning process was conducted in multiple phases with inter-annotator agreement validation.

Multi-phase fine-tuning · IAA validation across annotation cohort
Output
Dimensional Signal Profiling — Non-Diagnostic Output Layer

The model output layer generates dimensional signal profiles across the six-category taxonomy — not diagnostic outputs, not severity scores, and not clinical recommendations. The output is an interpretive context package: which signal patterns are present, their dimensional classification, and cultural interpretation notes that help clinicians understand what the patterns typically indicate in the communities that produce them. Every output includes a non-diagnostic disclaimer and is designed to support clinical judgment, not substitute for it.

Non-diagnostic · Interpretive context · 47-signal taxonomy · Clinician-surfaced only
Processing
In-Memory Processing — No Verbatim Storage

VLAP processes member language in-memory. Verbatim input text is not stored after signal profile generation. The retained output is a dimensional signal profile — not a transcript, not a quote, not a record of the member's exact language. This architectural constraint is a HIPAA technical safeguard, not a configurable setting. It cannot be disabled by organizational administrators or overridden by clinical staff.

In-memory inference · No PHI retention post-processing · HIPAA technical safeguard
Chapter 04 — Training Methodology

A corpus built
from inside the
communities.

The training data for VLAP was not web-scraped, crowd-sourced generically, or assembled from existing open datasets. It was specifically collected from and with the communities the model serves — and annotated by licensed clinicians who have worked within those communities.

Corpus Scale
2,400+
Extended vocabulary tokens beyond base BERT-family vocabulary
5
Cultural community cohorts represented in annotation
47
Distress signal patterns across 6 clinical categories
Methodology
Community-partnered data collection

Language samples were collected in partnership with community organizations serving Black, Latino, LGBTQ+, first-generation, and youth in urban and rural communities — not scraped from public platforms. Consent protocols, anonymization procedures, and community review were built into the collection process.

Clinical annotation with community competency requirement

All training samples were annotated by licensed clinicians with documented community competency training in the relevant population cohorts. Inter-annotator agreement was validated across the annotation cohort before samples were included in training data.

Vocabulary development through community engagement

The 2,400+ vocabulary extension was compiled through structured engagement with youth from the target communities — not through generic web scraping. Vocabulary candidates were reviewed for clinical accuracy by the annotation cohort before inclusion.

Bias monitoring integrated into training

False positive and false negative rates were disaggregated by demographic subgroup throughout training — not as a post-hoc audit but as an operational gate. Signal detection performance must meet minimum parity thresholds across Black, Latino, LGBTQ+, first-generation, and limited English proficiency subgroups for a model version to be deployed.

Chapter 05 — The 47-Signal Distress Signal Taxonomy

47 signals.
Six categories.

The VLAP Distress Signal Taxonomy (v1.0) is the clinical reference document behind VLAP — it defines the 47 discrete linguistic, behavioral, and contextual signals VLAP is trained to detect, organized into six clinically grounded categories. It serves as the model specification, the annotation codebook for IRB study data labeling, and the signal-level evidence base for VLAP’s NIMH SBIR Phase I application. The taxonomy is not static — it is updated as language evolves and as clinical review proceeds.

Critical Design Constraint — Applies to Every Signal in This Taxonomy

VLAP outputs are clinical decision support signals. No signal in this taxonomy, individually or in combination, triggers an automated clinical action. All VLAP outputs are reviewed by a licensed clinician, certified coach, or trained care coordinator before any clinical response. VLAP flags; humans act.

HOP
9 signals
Hopelessness & Futility

The Beck Hopelessness Scale and Columbia Suicide Severity Rating Scale both identify hopelessness as the strongest independent predictor of suicidal ideation. Generic NLP models perform adequately on Standard American English expressions of hopelessness (HOP-01). VLAP’s differentiation is in detecting culturally coded and vernacular expressions of futility (HOP-02 through HOP-09) that generic models systematically misclassify as frustration or hyperbole.

HOP-01 Absolute futurity negation High
“There's no point anymore”
“Nothing's ever going to change”
“What's even the use”

Standard SAE expression of hopelessness. Well-detected by generic NLP models. Included for completeness and as inter-rater reliability baseline in validation study.

HOP-02 Future self-erasure Moderate
“I won't be here when that happens”
“By the time that matters I'll be gone”
“Doesn't apply to me anymore”

Ambiguous temporal framing. Generic models frequently classify as planning or travel intent rather than ideation. Clinical significance depends on context — requires co-occurring signal for High risk classification.

HOP-03 AAVE futility construction High
“Can't keep doing this no more fr”
“Ain't no way out of this”
“I'm done fr fr”
“This ain't it no more”

CRITICAL — highest false negative rate in generic NLP models. The intensifier 'fr' (for real) and double negation construction 'no more' are systematic misclassification sources. Generic models return 'frustration' or 'venting'. VLAP training data includes 340+ AAVE futility constructions. This single signal accounts for the largest share of VLAP's sensitivity advantage over baseline.

HOP-04 Coded exhaustion Variable
“I'm just so tired”
“Tired of everything honestly”
“Tired of waking up and doing this again”

In isolation: Low risk, common expression. In combination with ISO or SHA signals: Moderate-High. 'Tired of waking up' is a clinically significant formulation — the exhaustion is specifically about existing. Requires attention to the object of the exhaustion, not just the emotion.

HOP-05 Temporal foreclosure Low
“I used to care about that stuff”
“I don't really think about the future anymore”
“Stopped making plans”

Shift from future-orientation to present-only or past-only temporal framing. Requires longitudinal signal comparison — most meaningful when contrasted against previous session content. Flags for V2 cross-session modeling. In single-session context: Low-Moderate.

HOP-06 Third-person self-distancing Moderate
“I was the kind of person who used to...”
“That's not me anymore”
“The old me would have cared about this”

Speaking about oneself in the past tense in a living context. Dissociative marker documented in pre-suicidal ideation literature. Distinct from healthy reflection — the distancing is from a valued self-identity. Particularly significant in combination with PFE signals.

HOP-07 Burden narrative High
“Everyone would honestly be better off”
“I'm just a problem for everyone”
“They don't actually need me around”
“I make everything harder for people I love”

Classic pre-suicidal cognition — perceived burdensomeness is one of the two primary components of Joiner's Interpersonal Theory of Suicide (alongside thwarted belongingness). High independent risk. Documented as having the highest correlation with suicide attempt in the clinical literature for this population.

HOP-08 Queer futility narrative High
“It doesn't get better for people like me”
“I've seen what happens to people like us”
“That kind of life isn't for someone like me”

Specifically references the cultural narrative of LGBTQIA+ suffering — 'people like me' and 'people like us' are key referents. Generic models do not detect the referent. For LGBTQIA+ youth, this expression carries the weight of observed community trauma, not just personal pessimism. Requires LGBTQIA+ identity context flag from member profile.

HOP-09 Immigrant & first-gen foreclosure Moderate
“All that sacrifice and for what”
“My parents went through all that and I can't even”
“I was supposed to be the one who made it”

Specific to first-generation and immigrant youth — the weight of intergenerational sacrifice creates a distinct form of hopelessness in which the member perceives themselves as failing a familial obligation. Not detectable by generic models without cultural context. Requires first-gen or immigrant identity flag from member profile.

SHA
8 signals
Coded Suicidal Ideation

The SHA category captures both explicit and coded expressions of suicidal ideation. Generic models perform well on explicit ideation (SHA-01) and poorly on all other signals. SHA-02 through SHA-08 represent the core of VLAP’s clinical differentiation — these are the expressions young people actually use, developed specifically to evade platform content moderation and the perceived stigma of explicit disclosure.

SHA-01 Explicit suicidal ideation High
“I want to kill myself”
“I'm thinking about suicide”
“I've been having thoughts of ending my life”

Well-detected by all NLP models. Included as validation baseline. VLAP does not differentiate meaningfully from generic models on this signal. The 47-signal taxonomy is justified by performance on SHA-02 through SHA-08 and the other five categories, not SHA-01.

SHA-02 Unaliving and derivatives High
“Been thinking about unaliving”
“I want to unalive myself honestly”
“Thought about unaliving last night”

CRITICAL — youth-specific euphemism developed explicitly to evade content moderation on TikTok, Instagram, and other platforms. First documented in widespread use approximately 2020. Most generic NLP models trained before 2022 do not have 'unaliving' in clinical context in training data. VLAP's training data includes 180+ documented uses with clinical annotation. This is one of the clearest demonstrations of generic model failure and VLAP's advantage.

SHA-03 Sleep-death conflation Moderate
“I just want to go to sleep and not wake up”
“Wish I didn't have to open my eyes tomorrow”
“I just want it to stop so I can rest”

Common expression chronically underweighted in generic models. The sleep framing is literal — the desire is for non-existence, not rest. 'I just want it to stop' requires object identification — what 'it' refers to determines clinical significance. Requires co-occurring signal for High classification unless the expression is highly specific.

SHA-04 Fantasy of absence Moderate
“Imagine if I just wasn't here anymore”
“What would happen if I disappeared”
“Sometimes I think about just being gone”

Passive ideation — the member is not describing an action but imagining a state of non-existence. Generic models frequently classify as daydreaming or escapism. In clinical literature, passive ideation is a documented precursor to active ideation. Escalate to Moderate when combined with any HOP signal.

SHA-05 Method inquiry with minimization wrapper High
“I was just curious, like how many pills would it actually take”
“Not that I'd do it but what does it feel like when...”
“Just wondering what happens if someone...”

CRITICAL — the minimization wrapper ('just wondering', 'not that I'd do it') is the disclosure suppression mechanism, not evidence of low risk. Generic models weight the minimization wrapper and reduce risk classification. VLAP is trained to detect the method inquiry independent of the wrapper. This signal should always trigger human review regardless of stated intent.

SHA-06 Goodbye behavior references High
“Told my best friend I loved her, like really told her”
“I gave away some stuff I don't need anymore”
“Finally finished that thing I'd been putting off”
“Made sure everything was in order”

Behavioral markers described in text — the member is describing actions consistent with pre-suicidal behavior. Often expressed with a sense of resolution or relief, which is a clinically significant affect shift. Requires careful contextual interpretation — these expressions can be innocent. Combination with HOP-01 through HOP-08 significantly elevates risk.

SHA-07 Anniversary and date markers with inverted framing High
“I just need to get through [date]”
“After [event] it won't matter anymore”
“I told myself if things aren't better by [date]...”

References to specific dates or events with framing that implies a terminus — a point after which the member does not expect or want to continue. Distinct from goal-setting. The inversion is key: the date is a deadline, not a milestone.

SHA-08 'Catching a fade' and community-specific death euphemisms High
“I've been thinking about catching a fade fr”
“Ready to check out for real”
“Done with this whole run honestly”

Community-specific coded language for death or suicide. 'Catching a fade' is documented in AAVE; 'checking out' and 'ending the run' are youth-vernacular constructions. These do not appear in generic NLP training data in clinical context. VLAP's training data includes 95+ community-specific death euphemisms with clinical annotation.

CCM
8 signals
Minimization & Disclosure Suppression

CCM signals are the category most specific to Vasl’s target population and most absent from generic NLP training data. They represent culturally conditioned patterns of downplaying distress before disclosing it — a documented phenomenon in Black, Latino, and LGBTQIA+ communities where mental health help-seeking carries stigma. Generic models classify CCM signals as low risk. VLAP treats them as pre-disclosure markers that warrant gentle engagement prompts and session logging.

CCM-01 Classic minimization opener Variable
“It's probably not a big deal but...”
“This is probably stupid to even say but...”
“I don't want to be dramatic, but...”

The minimization opener predicts significant disclosure in the sentence that follows. Generic models weight the minimizer and under-score the disclosure. VLAP is trained to detect the opener as a pre-disclosure flag and analyze the following content at elevated weight. The minimizer is a social safety mechanism — the member is testing for a non-judgmental response before fully disclosing.

CCM-02 AAVE minimization constructions Variable
“Lowkey been struggling lately”
“Kinda going through it ngl”
“It's whatever though, I'm good”
“Ion even know how to explain it”

AAVE-specific minimization — 'lowkey', 'kinda', and 'ion' (I don't) are qualifiers that reduce apparent severity without reducing actual severity. 'It's whatever though' is a disclosure-withdrawal marker — the member raised something and then minimized. VLAP detects the content of the disclosure, not the qualifier. High false negative rate in generic models.

CCM-03 Permission-seeking before disclosure Low
“Can I tell you something kind of personal?”
“Is it weird if I say something?”
“Don't judge me, okay, but...”

Pre-disclosure social negotiation — the member is seeking explicit permission and a non-judgmental response before disclosing. Particularly significant in youth who have previously disclosed to an adult and experienced a negative response. VLAP flags this for an immediate warm, non-judgmental coach response.

CCM-04 Gallows humor as disclosure mechanism Moderate
“lol I'm literally such a mess”
“haha I hate myself sometimes tho”
“not to be unalive about it but lmao”

CRITICAL — humor framing is the most common disclosure suppression mechanism in youth aged 13-22. The joke is the test: if the listener laughs or deflects, the disclosure is abandoned. VLAP is trained to detect the clinical content independent of the humor wrapper. 'Not to be unalive about it' is a specific formulation that combines SHA-02 with CCM-04 — extremely high combined risk.

CCM-05 Asking for a friend construction Moderate
“My friend is going through something, like what would you even do if...”
“This is for someone I know, but like how do you handle...”
“Not me but what does it mean when someone...”

Distance-creating framing used to test response before self-disclosure — a documented pre-disclosure mechanism in adolescents. VLAP flags all 'asking for a friend' formulations for gentle direct engagement rather than a literal, accusatory response.

CCM-06 Topic drop without resolution Moderate
“Anyway... never mind, it's fine”
“Actually forget I said anything”
“I shouldn't have brought it up”

The member raised a distress topic and then withdrew from it. This is a failed disclosure — the member attempted to disclose and retreated, likely due to fear of the response. VLAP flags topic drops for coach follow-up: acknowledging the withdrawal and creating a safe reopening.

CCM-07 Internalized stigma disclosure pattern Low
“I know I'm probably just being sensitive”
“Other people have it so much worse than me”
“I shouldn't be this upset, it's not even that serious”

Internalized stigma — the member has accepted the message that their distress is not legitimate or warranted. Particularly documented in Black, Latino, and LGBTQIA+ youth. The comparison to 'other people' with more severe situations is a disclosure suppression mechanism. VLAP flags for validation-focused coach response.

CCM-08 Performed resilience Low
“I'm good, I always figure it out”
“I don't let things get to me like that”
“I'm strong, I've been through worse”

Strength-performance in communities where emotional vulnerability is stigmatized as weakness — documented in Black male youth in particular, and in communities with strong 'we don't ask for help' cultural norms. VLAP flags for gentle, direct engagement that validates strength while creating space for vulnerability.

ISO
7 signals
Social Isolation & Withdrawal

Social isolation is a documented independent risk factor for suicidal ideation and is one of the two primary components of Joiner’s Interpersonal Theory of Suicide (thwarted belongingness). ISO signals are most clinically significant in combination with HOP or SHA signals, and as longitudinal markers of increasing withdrawal across sessions.

ISO-01 Direct withdrawal statement Moderate
“I've basically been staying in my room”
“Stopped going out or seeing people”
“Not really talking to anyone these days”

Explicit social withdrawal. Well-detected by generic models in standard formulation. VLAP adds detection of AAVE and youth-vernacular variants. Moderate risk in isolation; elevates significantly in combination with HOP or SHA signals.

ISO-02 Relationship severance language High
“Told people to stop checking on me”
“Blocked a bunch of people”
“Deleted my accounts”
“Pushed everyone away honestly”

Active severance of social support — distinct from passive withdrawal. The member is taking action to remove supportive contacts. High combined risk when co-occurring with any SHA signal.

ISO-03 Perceived burdensomeness — social expression Moderate
“I don't want to bring everyone down with my stuff”
“Everyone's dealing with their own things, I can't add to that”
“I don't want to be a burden”

Perceived burdensomeness expressed through social withdrawal rationale. Distinct from HOP-07 (self-directed) — ISO-03 is expressed as a reason for social withdrawal rather than a statement about the member's worth. Combination elevates to High.

ISO-04 Code-switching exhaustion Moderate
“I'm just tired of having to explain myself everywhere I go”
“Can't be myself anywhere, it's exhausting”
“Always performing, never actually me”

Documented psychological stressor specific to Black, Latino, and LGBTQIA+ youth — the exhaustion of constantly adapting one's language, presentation, and identity to majority contexts. In VLAP's context, code-switching exhaustion is an isolation marker.

ISO-05 Community disconnection Moderate
“Even my own people don't get it”
“Don't fit in anywhere, not even with people like me”
“Feel like I'm between worlds and belong in neither”

Feeling rejected or invisible within one's own cultural community — a specific form of isolation particularly painful for youth who experience intersectional marginalization. Generic models do not detect the cultural referent.

ISO-06 Family estrangement — LGBTQIA+ specific High
“My family doesn't know and if they found out...”
“Got kicked out because of who I am”
“I've been sleeping at a friend's place”
“My parents act like I don't exist now”

Family rejection is the single strongest risk factor for LGBTQIA+ youth suicidal ideation — rejected LGBTQIA+ youth are 8.4 times more likely to attempt suicide than accepted peers (Ryan et al., 2009). VLAP flags any expression of family rejection or housing instability related to LGBTQIA+ identity as High risk immediately.

ISO-07 Digital withdrawal references Low
“Haven't posted in like a month”
“Turned all my notifications off”
“Not really on anything anymore”

Withdrawal from digital social spaces — for youth, digital social networks are primary social infrastructure. Low risk in isolation; Moderate in combination with ISO-01 or any HOP signal.

TRM
8 signals
Trauma & Acute Stressor Markers

TRM signals capture social determinants of mental health that are disproportionately present in Vasl’s target population. Most TRM signals are not independently actionable clinical risk indicators — they are contextual elevators that increase the clinical significance of co-occurring HOP, SHA, or ISO signals. They are also the signals most normalized in community language and therefore most frequently missed by clinicians and tools not calibrated for this population.

TRM-01 Community violence exposure Variable
“Another one of my friends got shot”
“They got my cousin last week”
“This neighborhood just takes people”
“Can't even go outside without...”

Repeated exposure to community violence is a documented trauma and acute stressor for urban youth. For many members, violence is normalized in language. VLAP flags for trauma-informed engagement. Cumulative exposure tracked longitudinally.

TRM-02 Housing instability Moderate
“Staying with different people right now”
“We got evicted so...”
“My mom moved us again”
“Don't really have a stable place”

Housing instability is a strong predictor of youth mental health crisis, often expressed matter-of-factly in communities where it is common. VLAP flags for social determinants of health documentation and connection to housing support resources.

TRM-03 Immigration and documentation stress Variable
“My parents are scared about what's happening with immigration”
“We don't know if we can stay”
“ICE came to someone's house in our neighborhood”

Immigration anxiety is an acute and chronic stressor for immigrant youth and families. Not independently a clinical risk indicator but a significant context marker that elevates other signals.

TRM-04 Food insecurity references Low
“We don't really have food in the house”
“Been eating at school cause there's nothing at home”
“Hadn't eaten since yesterday”

Food insecurity is a significant stressor frequently normalized in community language. VLAP flags for connection to food resources and documentation as a social determinant. Low independent clinical risk; contextual elevator.

TRM-05 Police and legal system contact Variable
“The police stopped me again”
“My brother just got locked up”
“Scared every time I leave the house”
“Got a court date coming up”

Police contact and legal system involvement are acute and chronic stressors for Black and Latino youth disproportionately. The fear of police contact is independently traumatic regardless of actual contact.

TRM-06 School discipline and push-out Moderate
“Got suspended again”
“They're trying to expel me”
“I don't think I can go back to that school”
“They just kicked me out”

School push-out is a documented trauma, disproportionate for Black students and students with disabilities, and part of the school-to-prison pipeline stressor. VLAP flags for engagement and connection to educational support.

TRM-07 Vicarious and cumulative loss Moderate
“We lost so many people this year”
“I've been to too many funerals”
“My community just keeps losing people”

Cumulative community loss — the accumulation of grief across multiple losses within a community. Particularly significant for Black and Latino youth in communities with high rates of gun violence.

TRM-08 Caregiver incapacity and parentification Moderate
“My mom is really going through it so I'm taking care of things”
“I'm basically the parent right now”
“Nobody's really home, I'm handling everything”

Parentified youth — children assuming caregiver roles — experience heightened stress, accelerated loss of childhood, and reduced access to their own emotional needs. VLAP flags for acknowledgment of the burden and connection to youth support resources.

PFE
7 signals
Protective Factor Erosion

PFE signals are not distress signals in isolation — they indicate that a previously present protective factor has been removed, which elevates baseline risk. They are most clinically significant as longitudinal markers (V2 cross-session modeling) and in combination with HOP or SHA signals. The protective factors most relevant to Vasl’s population — faith community, mentors, peer networks, cultural identity, and aspirational goals — are different from those in the clinical literature developed for White, middle-class youth populations.

PFE-01 Loss of faith or spiritual anchor Low
“I stopped going to church”
“Don't really pray anymore”
“Lost my faith in all that honestly”

In communities where religious or spiritual practice is a primary coping resource and social support system, loss of faith represents the loss of a primary protective factor AND a primary social network simultaneously. Generic models do not identify this as a clinical signal.

PFE-02 Trusted adult loss Moderate
“My coach left the program”
“My counselor got replaced”
“The one adult who actually got me is gone”
“Nobody checks on me like that anymore”

Loss of a trusted adult is a documented protective factor erosion. For youth with limited trusted adult relationships, losing one trusted adult removes a disproportionate share of the member's social safety net. VLAP flags for immediate coach relationship reinforcement.

PFE-03 Peer network dissolution Moderate
“My whole friend group kind of fell apart”
“Everyone went their separate ways”
“Don't really have my people anymore”

Loss of peer social support. Distinct from ISO signals (which describe withdrawal) — PFE-03 describes an external dissolution of a previously present support network. Elevates significantly with HOP signals.

PFE-04 Achievement identity loss Moderate
“I can't play anymore because of my injury”
“My grades just dropped and I don't even care”
“Stopped performing — don't see the point”

Loss of the role or achievement identity that anchored the member's sense of self and social belonging. In combination with HOP signals, this pattern correlates with acute risk.

PFE-05 Future goal erasure Low
“I used to want to go to college but...”
“Stopped thinking about that kind of future”
“Those dreams don't seem realistic anymore”

Aspirational identity erosion. Particularly significant for first-generation youth for whom educational aspirations carry the weight of family sacrifice. A longitudinal hopelessness marker.

PFE-06 Cultural disconnection Low
“I feel like I'm losing my culture”
“Don't know where I belong — not here, not there”
“Forgot how to be who I was supposed to be”

Cultural identity is a documented protective factor for BIPOC youth. Strong cultural identity correlates with lower rates of depression and suicidal ideation in these populations. Particularly relevant in first-generation, immigrant, and mixed-heritage youth.

PFE-07 Loss of routine and structure Low
“I don't really have a schedule anymore”
“Stopped doing all the things I used to do”
“Nothing feels normal or regular”

Loss of daily structure and routine, particularly significant following major life transitions. Its loss correlates with increased rumination and social withdrawal. Low independent risk; Moderate in combination with TRM signals.

Section 4

Signal Combination & Escalation Rules

VLAP does not evaluate signals in isolation. The following rules define when single Moderate or Low signals elevate to High-risk classifications requiring immediate human review.

Signal Combination
Individual Levels
Combined
Action
Any SHA signal + any HOP signal
Moderate + Moderate
HIGH
Immediate human review — clinician or on-call coach
SHA-02 / SHA-05 / SHA-08 alone
High
HIGH
Immediate human review regardless of other signals
HOP-03 + HOP-07
High + High
HIGH
Immediate escalation — AAVE futility + burden narrative is the highest-risk pattern
ISO-06 + any SHA signal
High + any
HIGH
Immediate review — family rejection + ideation is a critical combination
CCM-04 with SHA-02 reference
Moderate + High
HIGH
'Not to be unalive about it but' formulation — always escalate
Any 3+ CCM signals in one session
Low / Moderate
MODERATE
Coach notification — member is heavily suppressing; create disclosure opening
PFE-02 + ISO-01 + HOP-04
Low + Moderate + Variable
MODERATE
Coach check-in — protective factor loss + withdrawal + exhaustion is a slow-burn pattern
3+ TRM signals in one session
Variable
MODERATE
Document social determinants; consider resource connection; check for co-occurring HOP/SHA
Section 5

Validation Status & Research Agenda

Current validation status of each signal category and the associated NIMH SBIR Phase I research agenda. VLAP distinguishes between signals that are clinically annotated and signals that have completed powered statistical validation — we do not present the former as the latter.

HOPHOP-01 to HOP-09
HOP-01 validated in pilot cohorts. HOP-02 through HOP-09 annotated; sensitivity/specificity pending powered study.
SBIR Research QuestionWhat is the sensitivity and specificity of HOP-03 through HOP-09 against Columbia-Suicide Severity Rating Scale (C-SSRS) scores in Black and Latino youth aged 13-25?
SHASHA-01 to SHA-08
SHA-01 validated. SHA-02 ('unaliving') detected in early pilot review with full clinical confirmation on each instance. SHA-03 through SHA-08 annotated; underpowered.
SBIR Research QuestionWhat is the false negative rate of SHA-02 through SHA-08 in generic NLP models vs. VLAP, measured against clinician-confirmed ideation in the same sample?
CCMCCM-01 to CCM-08
All signals annotated in pilot data. CCM-04 (gallows humor) most frequently occurring. No powered validation study completed.
SBIR Research QuestionDo CCM signals predict subsequent explicit disclosure (within 3 sessions) at a statistically significant rate in Black, Latino, and LGBTQIA+ youth?
ISOISO-01 to ISO-07
ISO-01 validated. ISO-06 (family rejection, LGBTQIA+) flagged in early pilot review. Others annotated only.
SBIR Research QuestionWhat is the combined predictive validity of ISO + HOP signal co-occurrence for PHQ-8 score deterioration over 30 days?
TRMTRM-01 to TRM-08
All signals annotated. No powered validation. Normalization in community language makes annotation complex — inter-rater reliability study required.
SBIR Research QuestionDoes TRM signal frequency moderate the relationship between HOP/SHA signals and clinical outcomes in this population?
PFEPFE-01 to PFE-07
All signals annotated as longitudinal markers. Requires session data accumulation for validation. Primarily a V2 research agenda.
SBIR Research QuestionDo PFE signals predict PHQ-8 trajectory deterioration over 60 days when present in the absence of active HOP/SHA signals?
Section 7

Key References

  1. Joiner, T.E. (2005). Why People Die by Suicide. Harvard University Press. — Theoretical foundation for HOP-07 (perceived burdensomeness) and ISO-03 (thwarted belongingness).
  2. Ryan, C. et al. (2009). Family rejection as a predictor of negative health outcomes in white and Latino LGB young adults. Pediatrics. — Empirical basis for ISO-06 (family rejection, LGBTQIA+) risk elevation.
  3. Green, L.J. (2002). African American English: A Linguistic Introduction. Cambridge University Press. — Linguistic reference for AAVE construction documentation in HOP-03 and CCM-02.
  4. Coppersmith, G. et al. (2015). CLPsych 2015 Shared Task: Mental Health Twitter Data. — NLP baseline for suicide signal detection; demonstrates generic model limitations on coded language.
  5. Balsam, K.F. et al. (2011). Measuring multiple minority stress: The LGBT People of Color Microaggressions Scale. — Foundation for ISO-04, ISO-05, and CCM-07 cultural specificity.
  6. Posner, K. et al. (2011). The Columbia-Suicide Severity Rating Scale. American Journal of Psychiatry. — Gold standard against which VLAP sensitivity/specificity will be validated in SBIR Phase I.
  7. Beck, A.T. et al. (1974). The Measurement of Pessimism: The Hopelessness Scale. — Foundation for HOP category clinical basis.
  8. Zimmerman, G. (2016). 'Unaliving': The internet word for suicide. — Linguistic documentation of SHA-02 emergence and spread.
Chapter 06 — Validation & Accuracy

Validated against
real clinical
populations.

VLAP's accuracy claims are grounded in active IRB-approved clinical research with a university research partner — not internal testing, not synthetic benchmarks, and not general NLP performance metrics that don't account for cultural signal specificity.

Accuracy Metrics
90%
High-distress signal sensitivity

VLAP's ability to detect high-distress signals when they are present in member language — measured against clinician-adjudicated ground truth in the IRB study cohort. Sensitivity is optimized conservatively for the SHA and HOP categories, and for signal combinations that trigger High-risk escalation: we accept more false positives to minimize missed crisis signals.

23%
Standard NLP vocabulary gap

Percentage of the VLAP training corpus that would be processed as unknown tokens by a standard BERT model without the extended vocabulary — representing the portion of culturally specific language that standard models are structurally unable to read.

Validation Approach
Active IRB Study — University Research Partner

VLAP signal accuracy is being validated through an IRB-approved clinical study with a university research partner, using production deployment data from live Vasl cohorts. The study compares VLAP signal output against clinician-adjudicated gold-standard assessments of the same member language. Results will be published in peer-reviewed literature upon study completion.

Demographic Bias Monitoring

False positive and false negative rates are disaggregated across Black, Latino, LGBTQ+, first-generation, and limited English proficiency subgroups in both the training validation and the IRB study. Parity thresholds are enforced operationally — a model version that meets aggregate accuracy targets but fails subgroup parity is not deployed. Bias monitoring is an ongoing production gate, not a one-time evaluation.

What Sensitivity Means — and Doesn't

Sensitivity measures how consistently VLAP detects signals when they are present — not the rate at which all surfaced signals are clinically significant in a given instance. A high-sensitivity threshold means more signals are surfaced, which is appropriate for a clinical support tool. The clinical significance of any specific signal is always determined through human clinical review, not by the model.

IRB Study — University Research Partner

The active IRB study is currently in the data collection and preliminary analysis phase. Results will be published in a peer-reviewed journal upon completion. The study protocol and preliminary design documentation are available to institutional evaluators under NDA. Contact clinical@vaslhealth.com to request access.

Medical Advisory — Senior Clinical Oversight

Vasl Health's clinical validation approach is overseen by its Senior Medical Advisor, Panagis Galiatsatos, MD, MHS — Assistant Professor of Medicine at Johns Hopkins University School of Medicine. Dr. Galiatsatos provides clinical oversight on VLAP's signal detection methodology, accuracy validation approach, and non-diagnostic output framing.

Chapter 07 — Clinical Integration Model

Surfaces to
clinicians.
Never to members.

VLAP is a clinical decision-support tool, not a member-facing AI. It operates entirely behind the clinical layer — invisible to the people whose language it processes. Every signal it surfaces is directed to a licensed clinician or certified coach, reviewed by a human, and responded to through human clinical judgment. The platform is built so that automated action in response to a clinical signal is architecturally impossible.

Step 01
Member check-in or coach message

VLAP processes only language shared through Vasl's care channels — daily check-ins and coach messaging threads. Peer group posts, external social media, school email, and any other channel are not processed.

Step 02
In-memory VLAP inference

VLAP processes the language against the 47-signal taxonomy. In-memory only — verbatim text is not retained after processing. Output: a dimensional signal profile.

Step 03
AI Client Insights surface to coach dashboard

Coaches see a simplified surface of VLAP output in the AI Client Insights panel: plain-language pattern alerts and mood trajectory summaries for their active members. No dimensional codes, no clinical jargon.

Step 04
Full dimensional panel surfaces to licensed clinician

When a member is connected to a licensed clinician, the pre-session view includes the full VLAP dimensional signal profile — dimensional codes, pattern descriptions, cultural interpretation notes, and coaching context. Accessible only to licensed clinicians.

Step 05
High-risk escalations surface to clinical supervisor — 90-minute SLA

Signals or signal combinations that meet High-risk escalation criteria are surfaced immediately to Vasl's licensed clinical supervisor team. A licensed clinician reviews the signal and determines the appropriate response. No automated action. Human judgment initiates every response.

VLAP Does Not
Respond to members or generate therapeutic messages
Diagnose or suggest diagnoses to any party
Make clinical decisions or recommendations
Initiate contact with members automatically
Take automated action in response to crisis signals
Scan peer group posts or external channels
Store verbatim member language after processing
Surface individual signal data to school or org administrators
VLAP Does
Detect culturally specific distress signals in care-channel language
Surface dimensional signal profiles to coaches and clinicians
Flag crisis signals for immediate human clinical supervisor review
Provide pre-session cultural context to licensed clinicians
Support coaches with AI Client Insights summaries
Enable aggregate, de-identified population trend data for org dashboards
Process in-memory without verbatim PHI retention
Operate entirely behind the clinical layer — invisible to members
Chapter 08 — Security Architecture

Built for
clinical data.

VLAP processes the most sensitive category of user data — mental health language from youth in underserved communities. The security architecture was designed specifically for HIPAA-regulated, school-based, and community health deployment contexts. Every constraint below is architectural, not configurable.

01 · Data Handling
In-Memory Processing

VLAP processes member language in-memory. Verbatim input text is not stored after signal profile generation. The retained output is a dimensional signal profile — not a transcript, not a quote. This is a HIPAA technical safeguard, not a configurable setting.

02 · HIPAA
Full Technical Safeguard Implementation

HIPAA Security Rule technical safeguards implemented across all platform components — encryption in transit and at rest, access controls, audit logging, and automatic logoff. Business Associate Agreement required for all organizational deployments. Annual third-party security audit.

03 · Certification
SOC 2 Type II

Annual SOC 2 Type II audit covering security, availability, and confidentiality trust service criteria. Full audit report available to institutional evaluators under NDA. Audit conducted by an independent third-party auditor.

04 · Access Control
Role-Based Signal Access

Individual VLAP signal context is accessible only to the assigned coach (AI Client Insights summary) and the assigned licensed clinician (full dimensional profile). School staff, org administrators, and Vasl team members outside clinical supervisory functions have zero access to individual signal data — architecturally enforced.

05 · FERPA
School-Based Deployment

For school district deployments, Vasl operates as a direct service provider to students. Student health data generated in Vasl is classified as health information under HIPAA — not as an education record under FERPA — and is structurally inaccessible to school administrators under any circumstances.

06 · Aggregate Data
De-Identification by Architecture

Population-level aggregate signal trends surfaced to org administrators use minimum cohort size enforcement to prevent de-identification by inference. Individual member contributions to aggregate data are never discernible. This constraint applies to all org-level reporting, without exception.

Chapter 09 — Institutional Partnerships

Validated by
the institutions
that matter.

VLAP's clinical credibility is grounded in active institutional partnerships — not aspirational affiliations or advisory relationships that don't involve actual work. The partnerships listed below involve ongoing operational collaboration, active research, or formal clinical oversight.

Academic Research Partner
Research University — IRB Study
Active IRB Study — VLAP Clinical Signal Validation

An IRB-approved clinical study is in progress with an academic research partner validating VLAP's signal detection accuracy against clinician-adjudicated ground truth assessments. The study uses production deployment data from live Vasl cohorts. Results will be published in a peer-reviewed journal upon completion. The study represents the first formal independent validation of VLAP's culturally specific signal detection capabilities.

IRB Active
Johns Hopkins University
Medical School — Senior Medical Advisor
Senior Medical Advisory — Clinical Validation Oversight

Panagis Galiatsatos, MD, MHS — Assistant Professor of Medicine at Johns Hopkins University School of Medicine — serves as Vasl Health's Senior Medical Advisor. Dr. Galiatsatos provides clinical oversight on VLAP's signal detection methodology, accuracy validation approach, non-diagnostic output framing, and the clinical governance of the platform's care coordination model. His advisory role involves active participation in clinical review, not nominal affiliation.

Advisory
Chapter 10 — Documentation Access

Evaluate
the model
directly.

Vasl Health provides full technical documentation to qualified institutional evaluators — health systems, research institutions, school district technology teams, and health plan medical directors. All documentation is available under NDA. Contact clinical@vaslhealth.com or use the form below to initiate an evaluation request.

VLAP Technical Specification
37-page model specification — architecture, training methodology, signal taxonomy with full annotation guidelines, inference pipeline, bias monitoring protocol, and performance benchmarks.
NDA Required
IRB Study Protocol
IRB study design, methodology, data collection protocol, and preliminary validation framework. Results pending publication.
NDA Required
SOC 2 Type II Report
Full annual third-party security audit report covering security, availability, and confidentiality trust service criteria.
NDA Required
HIPAA Technical Safeguards Documentation
Complete HIPAA Security Rule implementation documentation — technical safeguards, PHI handling protocols, BAA template, and breach notification procedures.
Available on Request
Health-System Data Collaboration Brief
Overview of the two-year data collaboration scope, data domains, formats (HL7 FHIR / CSV / JSON), HIPAA Safe Harbor de-identification specifications, and Year 1/Year 2 objectives.
NDA Required
Pilot Outcome Data
Aggregate outcomes from deployed pilot cohorts — PHQ-8 improvement, 30-day retention, session engagement, and clinical escalation rates. De-identified, aggregate only.
Available on Request
Related Reading
the infrastructure and data handling model
Hosting, encryption, and retention specifics for a security or IT reviewer.
VLAP, the layer this architecture supports
What the language analysis layer actually does with the data described here.
the model running inside this architecture
VLAP's training approach and current validation status.
why this architecture exists
The dialect and cultural bias problem this whole stack is built to address.