Skip to content
FunCoding

Search

Search docs, Skills and MCP

stata-accounting-research

STATA code pattern library for empirical archival accounting research. Provides tested syntax from 126 peer-reviewed JAR (Journal of Accounting Research) replication files (2017-2025). Use when the user asks procedural questions like "How do I implement [method]?" or "Show me code for [technique]" — including: entropy balancing, propensity score matching (PSM), difference-in-differences (DiD), regression discontinuity (RDD), instrumental variables (IV), event studies (CAR/BHAR), survival analysis, Fama-MacBeth regressions, bootstrap, quantile regression, reghdfe/xtreg/areg, clustering standard errors, fixed effects, esttab/outreg2 table formatting, winsorization, leads/lags. Users can specify their variables (e.g., treatment, outcomes, controls) and receive adapted syntax. NOTE: This skill provides code patterns from published papers, not research design advice.

科研4.5kskills/18-jusi-aalto-stata-accounting-research/SKILL.md

Install

Send this to Claude Code, Codex or Cursor. The agent checks the Skill for safety first and installs it only after you confirm.

读取 https://funcoding.ai/skills/brycewang-stanford/auto-empirical-research-skills/18-jusi-aalto-stata-accounting-research/install.md ,按里面的步骤帮我安装这个 Skill。

SKILL.md

Scope and Limitations

This skill is a code pattern library, not a methodological advisor.

Can DoCannot Do
Show how published papers implemented methodsExplain when to use one method over another
Provide tested STATA syntaxAdvise on identification strategy
Indicate which robustness tests accompany analysesDiscuss research design trade-offs
Cite source papers for code patternsRecommend optimal research design

When users ask methodology questions (e.g., "Should I use entropy balancing or PSM?", "How do I address endogeneity?", "Is my identification strategy valid?"):

  1. Acknowledge the limitation: "This skill provides code patterns from published papers, not research design guidance."
  2. Show how different papers approached similar problems (code examples)
  3. Suggest consulting methodology references: Breuer & deHaan (2024) for fixed effects, Angrist & Pischke for causal inference, or the user's methodologist/advisor
  4. Offer to show multiple implementations so the user can see variation in approaches

Workflow

Use references/REFERENCES.md as the primary index, then read targeted .do files.

Search references/REFERENCES.md to identify relevant papers. The index contains structured metadata:

  • Primary Method: STATA commands used (reghdfe, psmatch2, stcox, etc.)
  • Identification Strategy: DiD, PSM, IV, RDD, Event Study, etc.
  • Robustness/Special Features: Winsorization levels, clustering specs, placebo tests, etc.

Example queries on REFERENCES.md:

  • "entropy balancing" → finds JAR_60_alv, JAR_60_bl, JAR_61_ds, JAR_62_5_llz, JAR_63_2_npstv
  • "stacked DiD" → finds JAR_61_ds, JAR_62_5_aov, JAR_62_5_gibbons
  • "Cox hazard" → finds JAR_59_ctv, JAR_62_2_xyz

Stage 2: Code Extraction

Read only the identified .do files to extract actual syntax. This reduces context usage and improves accuracy.

Stage 3: Adaptation and Citation

  1. Adapt patterns to the user's variable names and research context
  2. Cite source: "Based on [Authors] ([Year]), JAR Volume"

Fallback: Direct Grep Patterns

For very specific syntax queries (e.g., "how does absorb() handle singletons?"), grep .do files directly:

TaskGrep Pattern
Panel regressionsreghdfe|xtreg|areg
Fixed effectsabsorb\(|i\.year|i\.firm
Clusteringcluster\(|vce\(cluster
Matching/PSMpsmatch2|teffects|cem|ebalance|pscore
IV regressionxtivreg|ivregress|ivreg2
DiDpost.*treat|treat.*post|parallel.*trend
RDDrdrobust|rddensity
Event studiesCAR|BHAR|abnormal.*return
Survivalstcox|streg|stset
Fama-MacBethfama.?macbeth|newey.*west
Bootstrapbootstrap|bsample
Quantile regressionqreg|sqreg|bsqreg
Table outputesttab|outreg2|eststo
Winsorizationwinsor|winsor2

Corpus Overview

126 STATA .do files from JAR Volumes 55-63 (2017-2025). See references/REFERENCES.md for complete catalog with paper titles and authors.

File Naming Convention

  • V55-61: JAR_{volume}_{shortcode}.do
  • V62-63: JAR_{volume}_{issue}_{shortcode}_{authors}.do

Volume Coverage

VolumeYearPapers
5520179
56201812
5720199
58202013
5920214
60202222
61202322
62202425
63202510

Standard Patterns

Clustering and Fixed Effects

* Firm and year FE with firm-clustered SEs (most common)
reghdfe depvar indepvar controls, absorb(firm year) cluster(firm)

* Industry-year FE
reghdfe depvar indepvar controls, absorb(ind_year) cluster(firm)

Output Conventions

eststo clear
eststo: reghdfe depvar indepvar controls, absorb(firm year) cluster(firm)
esttab using "table.tex", replace star(* 0.10 ** 0.05 *** 0.01) se

Winsorization

winsor2 varlist, cuts(1 99) replace

Similar Skills

lead-research-assistant
ComposioHQ/awesome-claude-skills77k

lead-research-assistant

Identifies high-quality leads for your product or service by analyzing your business, searching for target companies, and providing actionable contact strategies. Perfect for sales, business development, and marketing professionals.

Science

adaptyv
K-Dense-AI/scientific-agent-skills48k

adaptyv

Uses the Adaptyv Bio Foundry API and Python SDK to design protein characterization experiments, estimate costs, submit sequences, monitor laboratory progress, and retrieve results. Applies to Adaptyv Foundry, its target catalog, binding screening and affinity assays, thermostability, expression, fluorescence, epitope binning, and enzyme activity workflows, including code using adaptyv or FoundryClient.

Science

anndata
K-Dense-AI/scientific-agent-skills48k

anndata

Handles annotated matrices in single-cell analysis, .h5ad and Zarr files, and integration with the scverse ecosystem. This is the data format skill—for analysis workflows use scanpy; for probabilistic models use scvi-tools; for population-scale queries use cellxgene-census.

Science

deepchem
K-Dense-AI/scientific-agent-skills48k

deepchem

Builds molecular property prediction and MoleculeNet workflows with DeepChem, including SMILES featurization, scaffold or grouped holdouts, masked labels, graph models and explicit pretrained encoder transfer. Used for ADMET, toxicity, solubility and chemistry ML when DeepChem data/model contracts and scientific validation are needed.

Science

arboreto
K-Dense-AI/scientific-agent-skills48k

arboreto

Infers candidate gene regulatory networks from bulk or single-cell expression data using AertsLab Arboreto GRNBoost2 and GENIE3. Use for transcription factor-target association ranking, compatible Dask execution, sparse expression inputs, and network stability checks.

Science

cirq
K-Dense-AI/scientific-agent-skills48k

cirq

Google quantum computing framework. Use when targeting Google Quantum AI hardware, designing noise-aware circuits, or running quantum characterization experiments. Best for Google hardware, noise modeling, and low-level circuit design. For IBM hardware use qiskit; for quantum ML with autodiff use pennylane; for physics simulations use qutip.

Science