Agent skill

Cogsci Statistics

by NeuroAIHub in NeuroAIHub/BrainPilot

Domain-specific statistical modeling guidance for cognitive science and neuroscience, encoding when and how to apply mixed models, correction methods, Bayesian approaches, and effect size reporting

AGPL-3.0Auto-check passedData & Analytics

Install Cogsci Statistics

skills CLI
$ npx skills add NeuroAIHub/BrainPilot --skill cogsci-statistics -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install NeuroAIHub/BrainPilot cogsci-statistics --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/NeuroAIHub/BrainPilot.git skills-src && mkdir -p .claude/skills && cp -r skills-src/packages/skills/skills/02_Cross-Domain_Foundation/cogsci-statistics .claude/skills/cogsci-statistics && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
cogsci-statistics
GitHub stars
1.1k
Token cost
~5.3k tokens
SKILL.md length
2,427 words
Files
2 (incl. references)
Skills in repo
59
Repo updated
First seen
Licence
AGPL-3.0

At a glance

Domain-specific statistical modeling guidance for cognitive science and neuroscience, encoding when and how to apply mixed models, correction methods, Bayesian approaches, and effect size reporting

  • Works in 7 steps: Treating Items as Fixed Effects → Circular Analysis ("Double-Dipping") in… → Analyzing Accuracy with ANOVA Instead of… → …
  • Tasks that involve Statistics
  • SKILL.md covers Purpose, When to Use This Skill, Research Planning Protocol and ⚠️ Verification Notice, plus 7 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Cogsci Statistics is an agent skill from NeuroAIHub/BrainPilot. Domain-specific statistical modeling guidance for cognitive science and neuroscience, encoding when and how to apply mixed models, correction methods, Bayesian approaches, and effect size reporting

Its SKILL.md is about 5.3k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including reference files (for example `references/common-analyses.md`).

It sits in Data & Analytics, covering Statistics. The repository describes itself as: BrainPilot: Automating Brain Discovery with Agentic Research. The licence is AGPL-3.0.

When your agent uses it

  • Tasks that involve Statistics

Example prompts

  • “/cogsci-statistics”

Requirements

  • Python 3

Workflow steps

7 steps, taken from the step headings in SKILL.md.

  1. Treating Items as Fixed Effects
  2. Circular Analysis ("Double-Dipping") in Neuroimaging
  3. Analyzing Accuracy with ANOVA Instead of Logistic Models
  4. Inappropriate Outlier Exclusion
  5. Running ANOVAs on RT Without Addressing Skew
  6. Using Uncorrected Cluster-Forming Thresholds in fMRI
  7. Reporting Correlation P-Values Without CIs

What it can do on your machine

Read from SKILL.md and the folder at commit 93f6855. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are r).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • github.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Cogsci Statistics loads about 5.3k tokens when it runs, and up to ~11k if it reads all its reference files. Until then it costs about 54 tokens; SKILL.md has 2,427 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~54
When it runs · the whole SKILL.md, loaded when a task matches
~5.3k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~11k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from NeuroAIHub/BrainPilot at commit 93f6855, republished under its AGPL-3.0 licence (© NeuroAIHub). 2,427 words, ~5,336 tokens.

Download SKILL.mdSave it as .claude/skills/cogsci-statistics/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
cogsci-statistics
description
Domain-specific statistical modeling guidance for cognitive science and neuroscience, encoding when and how to apply mixed models, correction methods, Bayesian approaches, and effect size reporting
domain
research-methods
version
1.0.0
papers
Barr et al., 2013, Baayen et al., 2008, Benjamini & Hochberg, 1995, Maris & Oostenveld, 2007, Lo & Andrews, 2015, Wagenmakers, 2007
dependencies.required
research-literacy
dependencies.recommended
cogsci-power-analysis
review_status
ai-generated

Cognitive Science Statistical Analysis

Purpose

This skill encodes domain-specific statistical knowledge for cognitive science and neuroscience research. It addresses the modeling decisions, correction strategies, and reporting conventions that a general-purpose statistician or programmer would get wrong without training in the field. For concrete analysis recipes with code, see references/common-analyses.md.

When to Use This Skill

  • Choosing between repeated-measures ANOVA and mixed-effects models for a cognitive experiment
  • Specifying random effects structure for designs with subjects and items
  • Deciding how to handle reaction time (RT) data distributions
  • Selecting the appropriate multiple comparison correction
  • Deciding whether to use frequentist or Bayesian analysis
  • Reporting effect sizes and statistical results for journal submission

Research Planning Protocol

Before executing the domain-specific steps below, you MUST:

  1. State the research question — What specific hypothesis is this statistical analysis testing?
  2. Justify the method choice — Why this statistical model? What alternatives were considered?
  3. Declare expected outcomes — What pattern of results would support vs. refute the hypothesis?
  4. Note assumptions and limitations — What does this method assume? Where could it mislead?
  5. Present the plan to the user and WAIT for confirmation before proceeding.

For detailed methodology guidance, see the research-literacy skill.

⚠️ Verification Notice

This skill was generated by AI from academic literature. All parameters, thresholds, and citations require independent verification before use in research. If you find errors, please open an issue.

Repeated-Measures ANOVA vs. Mixed-Effects Models

When to Use Repeated-Measures ANOVA
  • Fully balanced design (no missing data, equal cell sizes)
  • Only subjects as a random factor (no item variability)
  • Simple factorial structure (2-3 factors, no continuous predictors)
  • Sphericity is met or correctable (Greenhouse-Geisser / Huynh-Feldt)
When to Use Mixed-Effects Models (LMM/GLMM)
  • Crossed random effects: Both subjects and items sampled from populations (Baayen, Clark, & Lucy, 2008; Clark, 1973). This is the norm in psycholinguistics, memory research, and any paradigm with stimulus variability.
  • Unbalanced data or missing observations
  • Continuous predictors (e.g., word frequency, stimulus duration)
  • Non-normal response distributions (RT, accuracy)
  • Need to generalize over both subjects AND items simultaneously

Critical domain knowledge: Clark (1973) demonstrated that failing to treat items as random effects inflates Type I error. This remains one of the most common statistical errors in cognitive science. If your stimuli are sampled from a larger population (e.g., words, faces, scenes), you must account for item variability.

Decision Logic
Are your stimuli sampled from a larger population?
 |
 +-- YES --> Mixed-effects model with crossed random effects
 | (subjects and items)
 |
 +-- NO (e.g., fixed set of 4 task conditions) -->
 |
 +-- Any missing data, unbalanced cells, or continuous predictors?
 | |
 | +-- YES --> Mixed-effects model (subjects as random effect)
 | |
 | +-- NO --> Repeated-measures ANOVA is acceptable
 |
 +-- Need trial-level analysis (e.g., RT distributions)?
 |
 +-- YES --> Mixed-effects model (operates on individual trials)
 +-- NO --> Repeated-measures ANOVA on condition means

Random Effects Structure

The Maximal Random Effects Principle

Barr et al. (2013) recommend fitting the maximal random effects structure justified by the design to minimize Type I error. This means including random intercepts and slopes for all within-unit factors.

For a typical 2x2 design with factors A (within-subjects, within-items) and B (within-subjects, between-items):

r
# Maximal structure (Barr et al., 2013)
lmer(RT ~ A * B + (1 + A * B | Subject) + (1 + A | Item), data = d)
When Maximal Models Fail to Converge

Convergence failures are common with complex random effects. Use this hierarchy (Barr et al., 2013; Matuschek et al., 2017):

  1. First: Try a different optimizer (bobyqa, nlminb) with increased iterations (20000 iterations; lme4 default recommendation)
  2. Second: Remove correlations between random effects (use || in lme4)
  3. Third: Remove the highest-order random slopes first (interaction before main effects)
  4. Fourth: Use a parsimonious approach guided by likelihood ratio tests (Matuschek et al., 2017)

Do NOT simply drop all random slopes to achieve convergence. This inflates Type I error and undermines the purpose of mixed-effects modeling (Barr et al., 2013).

Common Cognitive Science Designs and Their Random Effects
DesignRandom EffectsRationale
Lexical decision (words as items)`(1 + conditionsubj) + (1 + condition
Stroop task (fixed conditions)`(1 + congruencysubj)`
Picture naming (pictures as items)`(1 + SOAsubj) + (1
Multi-site study`(1 + conditionsubj) + (1

Handling Reaction Time Data

RT data in cognitive experiments are positively skewed, bounded below by physiological limits, and often contaminated by outliers. The approach matters.

RT Outlier Exclusion

Apply these criteria before modeling (Ratcliff, 1993; Luce, 1986):

CriterionThresholdSource
Fast outliers (anticipatory)< 200 msWhelan, 2008; Ratcliff, 1993
Slow absolute cutoff> 2000-3000 ms (task-dependent)Ratcliff, 1993
Within-subject SD trimming> 3 SD from participant's condition meanVan Selst & Jolicoeur, 1994
Within-subject MAD trimming> 3 MAD from participant's condition medianLeys et al., 2013 (more robust to skew)

Task-specific note: For simple RT tasks (e.g., detection), use 100 ms as the fast cutoff (Whelan, 2008). For choice RT tasks (e.g., lexical decision), use 200 ms (Ratcliff, 1993). Always report exclusion rates.

RT Transformation and Modeling Strategy
Is your primary interest in RT distributions (not just means)?
 |
 +-- YES --> Drift Diffusion Model or ex-Gaussian fitting
 |
 +-- NO --> Choose a modeling approach:
 |
 +-- Option 1: Log-transform RT, then fit LMM (Gaussian)
 | - Pro: Simple, widely understood
 | - Con: Back-transformation of means is biased;
 | changes the hypothesis being tested
 | (Lo & Andrews, 2015)
 |
 +-- Option 2: Inverse-transform RT (1/RT = speed), then LMM
 | - Pro: Often achieves better normality than log
 | - Con: Same back-transformation issues as log
 | (Ratcliff, 1993)
 |
 +-- Option 3 (Recommended): Generalized LMM with
 Gamma family + identity link
 - Pro: Models RT in original units; handles skew
 directly; avoids transformation issues
 (Lo & Andrews, 2015)
 - Con: Computationally slower; may have convergence
 issues with complex random effects

Recommended default: Gamma GLMM with identity link (Lo & Andrews, 2015). Report results on the original millisecond scale.

r
# Recommended RT model (Lo & Andrews, 2015)
glmer(RT ~ condition * group + (1 + condition | subj) + (1 | item),
 family = Gamma(link = "identity"), data = d)

Multiple Comparison Correction

Decision Guide for Cognitive Science
ScenarioMethodRationaleSource
Small number of planned contrasts (< 5)No correction or HolmPlanned contrasts based on a priori hypotheses do not require correction if specified before data collectionRubin, 2021
All pairwise comparisons after ANOVATukey HSDControls family-wise error for all pairwise comparisons; assumes equal varianceTukey, 1953
Many tests, correlated (e.g., EEG channels)Cluster-based permutationRespects spatial/temporal correlation structureMaris & Oostenveld, 2007
Many tests, independentBonferroni-HolmMore powerful than Bonferroni; step-down procedureHolm, 1979
Large-scale testing (fMRI voxels, genomics)FDR (Benjamini-Hochberg)Controls false discovery rate rather than family-wise error; appropriate when some false positives are tolerableBenjamini & Hochberg, 1995
Exploratory whole-brain fMRICluster-level FWE (with cluster-forming threshold p < 0.001)Eklund et al. (2016) showed that p < 0.01 cluster-forming threshold inflates false positive rates to ~70%Eklund et al., 2016
Confirmatory ROI analysis in fMRISmall volume correction (SVC) with FWERestricts search space to a priori ROIWorsley et al., 1996
When NOT to Correct
  • Single planned contrast testing a specific a priori hypothesis (Rubin, 2021)
  • Sequential Bayesian testing with BF stopping rules (evidence accumulation replaces correction; Schoenbrodt et al., 2017)

Bayesian Alternatives

When to Use Bayesian Analysis
  • Quantifying evidence for the null hypothesis: Frequentist tests cannot support H0; Bayes factors can (Wagenmakers, 2007)
  • Small sample sizes: Bayesian methods with informative priors can be more efficient (Kruschke, 2015)
  • Sequential testing: Bayes factors allow continuous monitoring without alpha inflation (Schoenbrodt et al., 2017)
  • Complex models where p-values are unreliable: Mixed models with small cluster sizes, or when asymptotic assumptions are questionable
Bayes Factor Interpretation
BF10 RangeEvidence CategorySource
< 1/10Strong evidence for H0Jeffreys, 1961; Lee & Wagenmakers, 2013
1/10 to 1/3Moderate evidence for H0Lee & Wagenmakers, 2013
1/3 to 3Anecdotal / inconclusiveLee & Wagenmakers, 2013
3 to 10Moderate evidence for H1Lee & Wagenmakers, 2013
> 10Strong evidence for H1Lee & Wagenmakers, 2013
ToolUse CaseLanguage
BayesFactorStandard designs (t-test, ANOVA, correlation, regression)R
brmsComplex models (multilevel, non-Gaussian, multivariate)R (Stan backend)
JASPGUI-based Bayesian analysis for standard testsStandalone
PyMCCustom Bayesian modelsPython
Reporting Bayes Factors

Report the exact BF, not just the category (Wagenmakers et al., 2018):

"A Bayesian paired-samples t-test indicated moderate evidence for a difference between conditions, BF10 = 5.3 (default Cauchy prior, r = 0.707)."

Always specify:

  1. The prior used (e.g., default Cauchy with scale r = 0.707 for BayesFactor t-test; Rouder et al., 2009)
  2. Direction (BF10 = evidence for H1 over H0)
  3. Robustness check: report BF across a range of prior widths

Effect Size Reporting

APA 7th Edition Requirements

APA 7th edition (2020, Section 6.6) requires reporting effect sizes for all primary analyses. The specific measure depends on the test:

TestEffect SizeInterpretation BenchmarksSource
t-test (between groups)Cohen's d0.2 small, 0.5 medium, 0.8 largeCohen, 1988
t-test (within subjects)Cohen's d_z or d_avd_z uses SD of difference scoresLakens, 2013
One-way ANOVAeta-squared or omega-squared0.01 small, 0.06 medium, 0.14 largeCohen, 1988
Factorial ANOVApartial eta-squared0.01 small, 0.06 medium, 0.14 largeCohen, 1988; Richardson, 2011
Mixed-effects modelsemi-partial R-squaredNo universal benchmarks; report CIRights & Sterba, 2019
Correlationr0.1 small, 0.3 medium, 0.5 largeCohen, 1988
Chi-squareCramer's V or phiDepends on dfCohen, 1988

Domain note: Always report confidence intervals around effect sizes (APA 7th, 2020). Use effectsize (R) or statsmodels (Python) for computation. The benchmarks above are Cohen's generic guidelines; paradigm-specific benchmarks are more informative (see ../cogsci-power-analysis/references/effect-sizes.md).

For Mixed-Effects Models

Traditional effect sizes are not straightforward for mixed models. Options:

  1. Semi-partial R-squared via the r2glmm or effectsize package (Rights & Sterba, 2019)
  2. Standardized regression coefficients: Standardize predictors before fitting
  3. Conditional and marginal R-squared: R2m (fixed effects only) and R2c (fixed + random) via MuMIn::r.squaredGLMM() (Nakagawa & Schielzeth, 2013)

Common Statistical Mistakes in Cognitive Science

1. Treating Items as Fixed Effects

Problem: Analyzing condition means averaged over items, ignoring item variability, fails to generalize beyond the specific stimuli used (Clark, 1973).

Fix: Use mixed-effects models with crossed random effects for subjects and items.

2. Circular Analysis ("Double-Dipping") in Neuroimaging

Problem: Selecting voxels/channels/time-windows based on the effect of interest, then testing that same effect (Kriegeskorte et al., 2009). Inflates effect sizes by 2x or more (Vul et al., 2009).

Fix: Use independent localizer, leave-one-out cross-validation, or whole-brain corrected analysis.

Show full SKILL.md (958 more words)Show less
3. Analyzing Accuracy with ANOVA Instead of Logistic Models

Problem: ANOVA on proportion correct violates normality and homogeneity assumptions, especially at ceiling (> 90%) or floor (< 10%) (Jaeger, 2008; Dixon, 2008).

Fix: Use logistic mixed-effects model on binary (correct/incorrect) trial-level data.

4. Inappropriate Outlier Exclusion

Problem: Removing "outlier" participants based on the dependent variable (e.g., excluding subjects whose effects go in the wrong direction) without a priori criteria.

Fix: Define exclusion criteria before data collection. Base exclusions on performance metrics (accuracy below chance, excessive RTs), not on the effect of interest.

5. Running ANOVAs on RT Without Addressing Skew

Problem: ANOVA on raw RT means violates normality. Condition means conceal distributional differences (Ratcliff, 1993).

Fix: Use Gamma GLMM (Lo & Andrews, 2015) or transform RTs, and supplement with distributional analysis if warranted.

6. Using Uncorrected Cluster-Forming Thresholds in fMRI

Problem: Cluster-based inference with cluster-forming thresholds more lenient than p < 0.001 (uncorrected) produces unacceptable false positive rates up to 70% (Eklund et al., 2016).

Fix: Use voxel-level threshold of p < 0.001 (uncorrected) as minimum cluster-forming threshold, or use voxel-level FWE/FDR correction.

7. Reporting Correlation P-Values Without CIs

Problem: A "significant" correlation of r = 0.30 with N = 50 has a 95% CI of [0.02, 0.53] -- the true effect could be near zero (Cumming, 2014).

Fix: Always report bootstrap 95% CI for correlations. Use 10000 bootstrap samples (Efron & Tibshirani, 1993).

Minimum Statistical Reporting Checklist

Based on APA 7th edition (2020) and Appelbaum et al. (2018):

  • Exact test statistic (F, t, chi-square, z) with degrees of freedom
  • Exact p-value (not just "< 0.05"), to 3 decimal places or "< .001"
  • Effect size with confidence interval
  • For mixed models: random effects structure, optimizer, convergence confirmation
  • For multiple comparisons: correction method and justification
  • Sample sizes for each group/condition
  • Data exclusion criteria (a priori) and proportion excluded
  • For Bayesian: prior specification, exact BF, robustness check

References

  • American Psychological Association. (2020). Publication Manual of the APA (7th ed.).
  • Appelbaum, M., et al. (2018). Journal article reporting standards for quantitative research. American Psychologist, 73(1), 3-25.
  • Baayen, R. H., Davidson, D. J., & Bates, D. M. (2008). Mixed-effects modeling with crossed random effects for subjects and items. Journal of Memory and Language, 59(4), 390-412.
  • Barr, D. J., Levy, R., Scheepers, C., & Tily, H. J. (2013). Random effects structure for confirmatory hypothesis testing. Journal of Memory and Language, 68(3), 255-278.
  • Benjamini, Y., & Hochberg, Y. (1995). Controlling the false discovery rate. Journal of the Royal Statistical Society B, 57(1), 289-300.
  • Clark, H. H. (1973). The language-as-fixed-effect fallacy. Journal of Verbal Learning and Verbal Behavior, 12(4), 335-359.
  • Cohen, J. (1988). Statistical Power Analysis for the Behavioral Sciences (2nd ed.). Erlbaum.
  • Cumming, G. (2014). The new statistics: Why and how. Psychological Science, 25(1), 7-29.
  • Dixon, P. (2008). Models of accuracy in repeated-measures designs. Journal of Memory and Language, 59(4), 447-456.
  • Efron, B., & Tibshirani, R. J. (1993). An Introduction to the Bootstrap. Chapman and Hall.
  • Eklund, A., Nichols, T. E., & Knutsson, H. (2016). Cluster failure: Why fMRI inferences for spatial extent have inflated false-positive rates. PNAS, 113(28), 7900-7905.
  • Holm, S. (1979). A simple sequentially rejective multiple test procedure. Scandinavian Journal of Statistics, 6(2), 65-70.
  • Jaeger, T. F. (2008). Categorical data analysis: Away from ANOVAs and toward logit mixed models. Journal of Memory and Language, 59(4), 434-446.
  • Jeffreys, H. (1961). Theory of Probability (3rd ed.). Oxford University Press.
  • Kriegeskorte, N., et al. (2009). Circular analysis in systems neuroscience. Nature Neuroscience, 12(5), 535-540.
  • Kruschke, J. K. (2015). Doing Bayesian Data Analysis (2nd ed.). Academic Press.
  • Lakens, D. (2013). Calculating and reporting effect sizes. Frontiers in Psychology, 4, 863.
  • Lee, M. D., & Wagenmakers, E.-J. (2013). Bayesian Cognitive Modeling. Cambridge University Press.
  • Leys, C., et al. (2013). Detecting outliers: Do not use standard deviation around the mean. Journal of Experimental Social Psychology, 49(4), 764-766.
  • Lo, S., & Andrews, S. (2015). To transform or not to transform: Using generalized linear mixed models to analyse reaction time data. Frontiers in Psychology, 6, 1171.
  • Luce, R. D. (1986). Response Times. Oxford University Press.
  • Maris, E., & Oostenveld, R. (2007). Nonparametric statistical testing of EEG- and MEG-data. Journal of Neuroscience Methods, 164(1), 177-190.
  • Matuschek, H., et al. (2017). Balancing Type I error and power in linear mixed models. Journal of Memory and Language, 94, 305-315.
  • Nakagawa, S., & Schielzeth, H. (2013). A general and simple method for obtaining R2 from generalized linear mixed-effects models. Methods in Ecology and Evolution, 4(2), 133-142.
  • Ratcliff, R. (1993). Methods for dealing with reaction time outliers. Psychological Bulletin, 114(3), 510-532.
  • Richardson, J. T. E. (2011). Eta squared and partial eta squared as measures of effect size. Educational Research Review, 6(2), 135-147.
  • Rights, J. D., & Sterba, S. K. (2019). Quantifying explained variance in multilevel models. Journal of Educational and Behavioral Statistics, 44(2), 223-263.
  • Rouder, J. N., et al. (2009). Bayesian t tests for accepting and rejecting the null hypothesis. Psychonomic Bulletin & Review, 16(2), 225-237.
  • Rubin, M. (2021). When to adjust alpha during multiple testing. Synthese, 199, 10969-11000.
  • Schoenbrodt, F. D., et al. (2017). Sequential hypothesis testing with Bayes factors. Psychological Methods, 22(2), 322-339.
  • Van Selst, M., & Jolicoeur, P. (1994). A solution to the effect of sample size on outlier elimination. Quarterly Journal of Experimental Psychology, 47A(3), 631-650.
  • Vul, E., et al. (2009). Puzzlingly high correlations in fMRI studies of emotion, personality, and social cognition. Perspectives on Psychological Science, 4(3), 274-290.
  • Wagenmakers, E.-J. (2007). A practical solution to the pervasive problems of p values. Psychonomic Bulletin & Review, 14(5), 779-804.
  • Wagenmakers, E.-J., et al. (2018). Bayesian inference for psychology. Part II: Example applications with JASP. Psychonomic Bulletin & Review, 25(1), 58-76.
  • Whelan, R. (2008). Effective analysis of reaction time data. The Psychological Record, 58(3), 475-482.
  • Worsley, K. J., et al. (1996). A unified statistical approach for determining significant signals in images of cerebral activation. Human Brain Mapping, 4(1), 58-73.

See references/common-analyses.md for concrete analysis recipes with code patterns.

© NeuroAIHub, AGPL-3.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (references) in packages/skills/skills/02_Cross-Domain_Foundation/cogsci-statistics of NeuroAIHub/BrainPilot.

  • SKILL.md
  • references/common-analyses.md

Open the folder on GitHubat commit 93f6855

Compare with similar skills

Cogsci Statistics next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Cogsci Statistics compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Cogsci Statistics this skillNeuroAIHub/BrainPilot1.1k—~5.3kAutomated safety check: PassAGPL-3.0
Sandbox Benchvercel/next.js143k—~4.1kAutomated safety check: PassMIT
Statistical Analysisspacering-net/codeg3.9k3 repos~5kAutomated safety check: PassMIT
StatsmodelszLanqing/codex-claude-academic-skills4.7k15 repos~4.9kAutomated safety check: PassBSD-3-Clause
AI Daily DigestvigorX777/ai-daily-digest1.6k—~1.3kAutomated safety check: PassNone
Statistical Powerspacering-net/codeg3.9k1 repos~3.6kAutomated safety check: NotesMIT

Similar skills

  • Sandbox Bench

    vercel/next.js

    Official

    Benchmark React or Next.js changes on Vercel Sandbox VMs with paired A/B statistics: react PR/commit vs base, or Next.js PR/commit vs base, measured end-to-end through the bench/render-pipeline app…

    143k GitHub stars~4.1k tokensUpdated today
    Data & AnalyticsAuto-check passed
  • Statistical Analysis

    spacering-net/codeg

    Guided statistical analysis for research data - test selection, assumption checking, effect sizes, power analysis, Bayesian alternatives, and APA-formatted reporting.

    3.9k GitHub starsUsed in 3 repos~5k tokens
    Data & AnalyticsAuto-check passed
  • Statsmodels

    zLanqing/codex-claude-academic-skills

    Statistical models library for Python. An agent skill from zLanqing/codex-claude-academic-skills.

    4.7k GitHub starsUsed in 15 repos~4.9k tokens
    Data & AnalyticsAuto-check passed
  • AI Daily Digest

    vigorX777/ai-daily-digest

    Fetches RSS feeds from 90 top Hacker News blogs (curated by Karpathy), uses AI to score and filter articles, and generates a daily digest in Markdown with Chinese-translated titles, category…

    1.6k GitHub stars~1.3k tokensUpdated 7 mo ago
    Data & AnalyticsAuto-check passed
  • Statistical Power

    spacering-net/codeg

    Sample-size and statistical power calculations for planning studies.

    3.9k GitHub starsUsed in 1 repo~3.6k tokens
    Data & AnalyticsAuto-check: notes
  • Agent Session Monitor

    higress-group/higress

    Real-time agent conversation monitoring - monitors Higress access logs, aggregates conversations by session, tracks token usage.

    9.5k GitHub stars~3.3k tokensUpdated 2 days ago
    Data & AnalyticsAuto-check passed

More from NeuroAIHub/BrainPilot

All 59 skills in this repo
  • Deeplabcut

    NeuroAIHub/BrainPilot

    Toolbox for markerless animal pose estimation with DeepLabCut.

    1.1k GitHub stars~1.7k tokensUpdated 8 days ago
    Auto-check passed
  • Fmriprep

    NeuroAIHub/BrainPilot

    Preprocess task-based or resting-state fMRI data with fMRIPrep — a robust, BIDS-App preprocessing pipeline built on FSL, ANTs, FreeSurfer, AFNI, and Nilearn.

    1.1k GitHub stars~4.1k tokensUpdated 8 days ago
    Auto-check passed
  • Mne Python Guide

    NeuroAIHub/BrainPilot

    Domain-validated pipeline guidance for EEG/MEG data analysis using MNE-Python: data loading, preprocessing (filtering, ICA, re-referencing), epoching, ERP/ERF computation, time-frequency…

    1.1k GitHub stars~2.3k tokensUpdated 8 days ago
    Auto-check passed
  • Netneurotools Guide

    NeuroAIHub/BrainPilot

    Domain-validated guidance for network neuroscience analysis using netneurotools: datasets, brain network metrics, connectivity consensus, modularity, spatial statistics, null models, and cortical…

    1.1k GitHub stars~2.6k tokensUpdated 8 days ago
    Auto-check passed
  • Nature Figure

    NeuroAIHub/BrainPilot

    Submission-grade Nature/high-impact journal figure workflow for Python or R.

    1.1k GitHub starsUsed in 1 repo~1.3k tokens
    Auto-check passed
  • Pycortex Guide

    NeuroAIHub/BrainPilot

    Domain-validated guidance for cortical surface visualization and brain surface rendering of fMRI data using pycortex: data types (Volume, Vertex, Dataset), 2D cortical flatmaps, 3D WebGL brain…

    1.1k GitHub stars~1.6k tokensUpdated 8 days ago
    Auto-check passed

Questions about Cogsci Statistics

What does Cogsci Statistics do?

Domain-specific statistical modeling guidance for cognitive science and neuroscience, encoding when and how to apply mixed models, correction methods, Bayesian approaches, and effect size reporting. Cogsci Statistics is an agent skill from NeuroAIHub/BrainPilot.

When should I use Cogsci Statistics?

Cogsci Statistics fits situations like: tasks that involve Statistics.

How do I install Cogsci Statistics in Claude Code?

Run `npx skills add NeuroAIHub/BrainPilot --skill cogsci-statistics -a claude-code`. Or copy the skill folder (packages/skills/skills/02_Cross-Domain_Foundation/cogsci-statistics in NeuroAIHub/BrainPilot) into .claude/skills/cogsci-statistics in your project. Claude Code loads it when a task matches its description.

How do I install Cogsci Statistics in Codex?

Run `npx skills add NeuroAIHub/BrainPilot --skill cogsci-statistics -a codex`. Or copy the skill folder (packages/skills/skills/02_Cross-Domain_Foundation/cogsci-statistics in NeuroAIHub/BrainPilot) into .agents/skills/cogsci-statistics in your project. Codex loads it when a task matches its description.

Can I use Cogsci Statistics in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add NeuroAIHub/BrainPilot --skill cogsci-statistics -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/cogsci-statistics, .gemini/skills/cogsci-statistics, .github/skills/cogsci-statistics and .opencode/skills/cogsci-statistics in your project.

What does Cogsci Statistics need to run?

SKILL.md names no scripts, command-line tools or credentials: Cogsci Statistics is instructions for the agent only. Our summary lists: Python 3.

Does Cogsci Statistics access the network?

SKILL.md names 1 domain. As links in the text: github.com. This is read from the text; nothing was executed.

Is Cogsci Statistics safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Cogsci Statistics use?

Cogsci Statistics is published under the AGPL-3.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Cogsci Statistics use?

About 5.3k tokens (SKILL.md is roughly 21k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 5.2k tokens, read only when the agent opens those files.

What are the alternatives to Cogsci Statistics?

Skills that share tags, products or a category with Cogsci Statistics: Sandbox Bench (vercel/next.js, 143k stars), Statistical Analysis (spacering-net/codeg, 3.9k stars), Statsmodels (zLanqing/codex-claude-academic-skills, 4.7k stars) and AI Daily Digest (vigorX777/ai-daily-digest, 1.6k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Cogsci Statistics?

NeuroAIHub (a GitHub organization) maintains it in NeuroAIHub/BrainPilot, which has 1,062 GitHub stars. The repository holds 59 skills in this directory. The repository was last updated on October 2, 2026.

Source: NeuroAIHub/BrainPilot on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.