English ESLSpok
Overview
| type | only spoken |
| available since | 2.12 |
| link | https://github.com/UniversalDependencies/UD_English-ESLSpok |
| genre | spoken |
| contributors | Kyle, Kris; Eguchi, Masaki; Miller, Aaron; Sither, Ted |
| sentences | 2320 |
| tokens | 21312 |
Issue draft: UD_English-ESLSpok
Modality identification
Is spoken part clearly identifiable? N/A - spoken data only
Metadata review
doc (and paragraphs) metadata
(none found) - no document_id exists at all.
| Field | Advice |
|---|---|
| — | derive # document_id from the sent_id prefix (everything before _<number>, e.g. file01243.txt); recompose by sorting sentences within each prefix by that trailing number, which recovers a consistent (if sparse, since only a sample was taken) in-document order |
speaker metadata
(none found) - each document is one L2 English speaker’s interview session; no speaker_id is encoded, though one could plausibly be derived from the same filename once document_id exists.