Research Design (lang-research-design)
Language is method-pluralist: it publishes elicited fieldwork, corpus studies, phonetic and
experimental work, computational modeling, and diachronic/typological comparison, and it judges each by
the standards of its own subfield. The job here is to make the design defensible to a general,
possibly cross-subfield, double-anonymous reviewer — and to show the evidence actually supports the
theoretical claim from lang-theory-building.
When to trigger
- Choosing or justifying the design before data collection or analysis
- A reader questioned the elicitation, the consultant sample, corpus coverage, measurement, or the
typological sample
- Aligning the evidence with the analysis's predictions
- Mixed-evidence work (e.g., corpus + experiment) that must defend each component
Defend the design (by subfield)
Elicited / fieldwork data
- Describe consultant number and background, elicitation method, and the recording/annotation
workflow; distinguish elicited judgments from spontaneous/textual data.
- Give data in numbered examples with Leipzig interlinear glossing and a source for each token;
a reader must be able to see the pattern, not take it on faith.
Corpus / quantitative usage
- Justify corpus choice, sampling frame, and coding scheme; report inter-annotator agreement for
hand-coded variables; state how tokens were extracted and excluded.
Phonetic / experimental
- Specify participants, stimuli, task, and measurement (e.g., forced alignment, formant/pitch
extraction settings); pre-empt confounds; where predictions are directional, say so in advance.
Diachronic / typological
- Make sample construction and genealogical/areal control explicit; guard against areal or
bibliographic bias; keep a clear trail from primary sources to the coded generalization.
Computational / modeling
- State what the model is a model of; separate the claim about the grammar from the properties of the
architecture or training data.
Match design to claim
The single most common Language reviewer objection: the data cannot bear the generalization. Walk
the chain: claim → prediction → the observation that would confirm/disconfirm it → the design's leverage
on that observation. A three-language convenience sample cannot ground a universal; either narrow the
claim or widen the evidence — do not overreach.
Referee-pushback patterns by subfield (the modal Language objection)
| Referee writes… |
Subfield |
The Language-appropriate fix |
| "Judgments from one speaker." |
fieldwork |
add consultants or scope the claim to the idiolect/variety |
| "Cherry-picked corpus tokens." |
corpus |
report the full extraction + exclusion rule + agreement |
| "Confound with speech rate." |
phonetics |
control or model it; show the effect survives |
| "Sample is areally biased." |
typological |
rebalance the sample or restrict the generalization |
Calibration with a quick example (hedged)
Language judges each subfield by its own standard, not a single template; unlike a purely formal venue
that accepts introspective judgments alone, it increasingly expects the evidence base to be visible and
checkable. Illustrative: an author claims a word-order universal from four related languages; a referee
flags "genealogical non-independence." The fix draws a genealogically stratified sample and restates the
claim as a statistical tendency with the mechanism, so the typology can see the pattern fail as well as
hold. Confirm current data expectations on the author pages and in lang-data-and-transparency.
Design pass for Language
Treat this skill as an executable review pass, not a prose hint. First lock the empirical
generalization, evidence base, warrant, and theoretical payoff; then judge whether the manuscript
answers the venue's real reader: linguists across subfields who value grounded analysis, transparent and
checkable evidence, and careful, appropriately scoped generalizations.
- Do the pass: lock the unit (segment / token / speaker / language), the sample, the comparison, the
validity threat, and the minimum decisive evidence before recommending collection or submission.
- Return a ledger: give
claim / evidence / risk / manuscript location rows so the next agent can
edit rather than rediscover the issue.
- Sibling guard: compare against Phonology, NLLT, Journal of Semantics, Diachronica,
Language Variation and Change; if a sibling owns the contribution, recommend re-routing before
polishing.
- Stop condition: do not give submission-ready advice until
resources/official-source-map.md has
been checked and the manuscript has one concrete fix for the largest venue-specific risk.
Anti-patterns
- Grounding a general claim on a convenience sample that cannot support it
- Judgments from a single consultant presented as facts about the language
- Corpus tokens hand-picked with no stated extraction or exclusion rule
- Phonetic effects reported without controlling obvious confounds
- A typological sample with unacknowledged genealogical or areal dependence
- A design that probes something adjacent to, but not, the stated prediction
Output format
【Subfield】fieldwork / corpus / phonetic-experimental / typological-diachronic / computational / mixed
【Claim it must support】from theory-building
【Design leverage】how this evidence bears on the prediction
【Key threats】consultant number, sampling, confounds, non-independence, annotation
【Evidentiary trail】data → glossed examples → claim is legible? [Y/N]
【Verdict】supports the claim / needs tightening / overreaches (fix)
【Next】lang-data-analysis
Supplementary resources
Source: brycewang-stanford/Awesome-Journal-Skills → Language-Linguistic-Society-Skills/skills/lang-research-design/SKILL.md
1---2name: lang-research-design3description: Use when defending the empirical design of a Language (LSA) manuscript on the terms of its subfield — elicitation and fieldwork, corpus construction, phonetic measurement, experiment, or the diachronic/typological sample. Language judges each kind of evidence by its own standards, and the design must support the theoretical claim. Defends the design; it does not run the analysis.4---567# Research Design (lang-research-design)89*Language* is method-pluralist: it publishes elicited fieldwork, corpus studies, phonetic and10experimental work, computational modeling, and diachronic/typological comparison, and it judges each by11the standards of *its own subfield*. The job here is to make the design defensible to a general,12possibly cross-subfield, double-anonymous reviewer — and to show the evidence actually supports the13theoretical claim from `lang-theory-building`.1415## When to trigger1617- Choosing or justifying the design before data collection or analysis18- A reader questioned the elicitation, the consultant sample, corpus coverage, measurement, or the19 typological sample20- Aligning the evidence with the analysis's predictions21- Mixed-evidence work (e.g., corpus + experiment) that must defend each component2223## Defend the design (by subfield)2425### Elicited / fieldwork data26- Describe **consultant number and background, elicitation method, and the recording/annotation27 workflow**; distinguish elicited judgments from spontaneous/textual data.28- Give data in **numbered examples with Leipzig interlinear glossing** and a source for each token;29 a reader must be able to see the pattern, not take it on faith.3031### Corpus / quantitative usage32- Justify **corpus choice, sampling frame, and coding scheme**; report inter-annotator agreement for33 hand-coded variables; state how tokens were extracted and excluded.3435### Phonetic / experimental36- Specify **participants, stimuli, task, and measurement** (e.g., forced alignment, formant/pitch37 extraction settings); pre-empt confounds; where predictions are directional, say so in advance.3839### Diachronic / typological40- Make **sample construction and genealogical/areal control** explicit; guard against areal or41 bibliographic bias; keep a clear trail from primary sources to the coded generalization.4243### Computational / modeling44- State what the model is a model *of*; separate the claim about the grammar from the properties of the45 architecture or training data.4647## Match design to claim4849The single most common *Language* reviewer objection: **the data cannot bear the generalization.** Walk50the chain: claim → prediction → the observation that would confirm/disconfirm it → the design's leverage51on that observation. A three-language convenience sample cannot ground a universal; either narrow the52claim or widen the evidence — do not overreach.5354## Referee-pushback patterns by subfield (the modal Language objection)5556| Referee writes… | Subfield | The Language-appropriate fix |57|-----------------|----------|------------------------------|58| "Judgments from one speaker." | fieldwork | add consultants or scope the claim to the idiolect/variety |59| "Cherry-picked corpus tokens." | corpus | report the full extraction + exclusion rule + agreement |60| "Confound with speech rate." | phonetics | control or model it; show the effect survives |61| "Sample is areally biased." | typological | rebalance the sample or restrict the generalization |6263## Calibration with a quick example (hedged)6465*Language* judges each subfield by its own standard, not a single template; unlike a purely formal venue66that accepts introspective judgments alone, it increasingly expects the evidence base to be visible and67checkable. Illustrative: an author claims a word-order universal from four related languages; a referee68flags "genealogical non-independence." The fix draws a genealogically stratified sample and restates the69claim as a statistical tendency with the mechanism, so the typology can see the pattern fail as well as70hold. Confirm current data expectations on the author pages and in `lang-data-and-transparency`.7172## Design pass for Language7374Treat this skill as an executable review pass, not a prose hint. First lock the empirical75generalization, evidence base, warrant, and theoretical payoff; then judge whether the manuscript76answers the venue's real reader: linguists across subfields who value grounded analysis, transparent and77checkable evidence, and careful, appropriately scoped generalizations.7879- **Do the pass:** lock the unit (segment / token / speaker / language), the sample, the comparison, the80 validity threat, and the minimum decisive evidence before recommending collection or submission.81- **Return a ledger:** give `claim / evidence / risk / manuscript location` rows so the next agent can82 edit rather than rediscover the issue.83- **Sibling guard:** compare against *Phonology*, *NLLT*, *Journal of Semantics*, *Diachronica*,84 *Language Variation and Change*; if a sibling owns the contribution, recommend re-routing before85 polishing.86- **Stop condition:** do not give submission-ready advice until `resources/official-source-map.md` has87 been checked and the manuscript has one concrete fix for the largest venue-specific risk.8889## Anti-patterns9091- Grounding a general claim on a convenience sample that cannot support it92- Judgments from a single consultant presented as facts about the language93- Corpus tokens hand-picked with no stated extraction or exclusion rule94- Phonetic effects reported without controlling obvious confounds95- A typological sample with unacknowledged genealogical or areal dependence96- A design that probes something adjacent to, but not, the stated prediction9798## Output format99100```101【Subfield】fieldwork / corpus / phonetic-experimental / typological-diachronic / computational / mixed102【Claim it must support】from theory-building103【Design leverage】how this evidence bears on the prediction104【Key threats】consultant number, sampling, confounds, non-independence, annotation105【Evidentiary trail】data → glossed examples → claim is legible? [Y/N]106【Verdict】supports the claim / needs tightening / overreaches (fix)107【Next】lang-data-analysis108```109110## Supplementary resources111112- [`../../resources/external_tools.md`](../../resources/external_tools.md) — elicitation, corpus, and phonetic tooling by subfield113- [`../../resources/official-source-map.md`](../../resources/official-source-map.md) — Language method-pluralism and evidence expectations114115---116117**Source:** [`brycewang-stanford/Awesome-Journal-Skills`](https://github.com/brycewang-stanford/Awesome-Journal-Skills) → `Language-Linguistic-Society-Skills/skills/lang-research-design/SKILL.md`