跳到正文
FunCoding

搜索

搜索文档、Skill 和 MCP

00-academic-router

Use when the user wants help with academic papers or citations but it's unclear which specific workflow fits — reviewing a paper, checking a BibTeX file for fake references, or benchmarking multiple LLMs on reference-recommendation accuracy. Also use when the user mentions paper review, peer review, BibTeX verification, citation checking, reference hallucination, or academic literature accuracy and hasn't specified which of those three tasks they mean. This skill is the entry router for the academic-eval suite: it asks one diagnostic question then routes to the right sub-skill.

科研868skills/academic-eval/00-academic-router/SKILL.md

安装

把这段话发给 Claude Code、Codex 或 Cursor。智能体会先检查安全性,你确认后才安装。

读取 https://funcoding.ai/skills/agentscope-ai/openjudge/00-academic-router/install.md ,按里面的步骤帮我安装这个 Skill。

SKILL.md

Academic Eval Router

Entry router for the academic-eval suite. You diagnose what the user actually wants and route them to one of three sub-skills. You don't review papers, verify BibTeX files, or run arena benchmarks yourself — you're the triage desk.

Each sub-skill is self-contained: it carries inline everything it needs, so it can be installed and used on its own.

Diagnostic Question

Ask (unless the user's request already makes the answer obvious):

To route you correctly, which of these matches what you want?

a) Review a single paper (PDF or LaTeX source) for correctness/quality/novelty
   — optionally also check its bibliography
b) Check a standalone .bib file for fabricated or mismatched references
   (no paper review needed)
c) Benchmark/compare multiple LLMs on how often they hallucinate references
   when asked to recommend citations (arena-style, many queries)

Shortcut rule: if the user already said "review my paper", "check this PDF", "verify this .bib file", or "compare models on reference hallucination", skip the question — the routing is already clear from their phrasing.

Triage Table

User says / hasUse workflowWhat it does
"Review this paper" (PDF or .tar.gz/.zip TeX source)01-paper-reviewMulti-stage review: safety, correctness, quality/novelty score, criticality — optionally + BibTeX check
"Review this paper AND check its references"01-paper-reviewSame pipeline with --bib set — one run covers both
"Just check this .bib file, no paper"02-bib-verifyCross-checks every entry against CrossRef/arXiv/DBLP, flags verified/suspect/not_found
"Compare N models on how often they cite fake papers" / "benchmark reference hallucination rate"03-ref-hallucination-arenaRuns many recommendation queries per model, verifies every returned reference, ranks models by hallucination rate
"Compare models on general quality/response, not specifically citations"—Not this suite — see the arena-eval suite's 01-auto-arena instead

Key distinctions

  • 01-paper-review vs 02-bib-verify: both use the same underlying cookbooks.paper_review pipeline. Use 01-paper-review whenever a paper file exists (even if the only thing the user cares about is the bibliography — --bib_only mode is documented there). Use 02-bib-verify only when there is no paper, just a loose .bib file to sanity-check.
  • 01-paper-review/02-bib-verify vs 03-ref-hallucination-arena: the first two evaluate one document's existing references after the fact. The third evaluates model behavior — how often a model invents fake citations when asked to recommend some, across a benchmark of queries and models. If the user wants a leaderboard/ranking of models, not a report on one document, route to 03-ref-hallucination-arena.

Output

Recommended workflow: `[skill-name]`

Why: [one sentence tying the user's request to the triage table row]

Recommend exactly one workflow. If the request spans two (e.g., "review this paper, and separately benchmark 3 models on citation accuracy"), say so explicitly and give both, in the order the user would naturally do them.

相似的 Skill

lead-research-assistant
ComposioHQ/awesome-claude-skills77k

lead-research-assistant

Identifies high-quality leads for your product or service by analyzing your business, searching for target companies, and providing actionable contact strategies. Perfect for sales, business development, and marketing professionals.

科研

13c-metabolic-flux
K-Dense-AI/scientific-agent-skills48k

13c-metabolic-flux

Estimates intracellular metabolic fluxes from steady-state carbon-13 isotope-tracing measurements using validated atom maps, mfapy isotope simulation, constrained multistart fitting, and flux-profile diagnostics. Use for 13C-MFA, carbon tracing, mass isotopomer distributions (MDVs/MIDs), positional isotopomers, parallel tracer experiments, and determining whether labeling data constrain a pathway flux. Distinguishes measured-label inference from COBRA flux balance analysis and flags experiments requiring nonstationary MFA.

科研

datamol
K-Dense-AI/scientific-agent-skills48k

datamol

Pythonic wrapper around RDKit with simplified interface and sensible defaults. Preferred for standard drug discovery including SMILES parsing, standardization, descriptors, fingerprints, clustering, 3D conformers, parallel processing. Returns native rdkit.Chem.Mol objects. For advanced control or custom parameters, use rdkit directly.

科研

biopython
K-Dense-AI/scientific-agent-skills48k

biopython

Provides Biopython workflows for sequence manipulation, file parsing (FASTA/GenBank/PDB), phylogenetics, and programmatic NCBI/PubMed access (Bio.Entrez). Supports batch processing, custom molecular-biology pipelines, BLAST automation, structure analysis, and motif analysis.

科研

bulk-rnaseq
K-Dense-AI/scientific-agent-skills48k

bulk-rnaseq

Prepares bulk RNA-seq FASTQ, Salmon, STAR or featureCounts output for gene-level differential expression. Covers nf-core/rnaseq and standalone quantification, biological replication, strandedness, reference provenance, validated count assembly and a PyDESeq2 handoff. Use for FASTQ-to-counts analysis, nf-core/rnaseq configuration, STAR/Salmon quantification, or building a counts matrix for DESeq2. For single-cell data use scanpy; for statistical fitting alone use pydeseq2.

科研

alphagenome
K-Dense-AI/scientific-agent-skills48k

alphagenome

Looks up precomputed AlphaGenome Atlas effects for any GRCh38 single-nucleotide variant (AVI score with Phred and 18 SHAP feature attributions, plus raw and quantile scores for RNA-seq, DNase, ATAC, ChIP-TF, ChIP-histone, CAGE, PRO-cap, splicing, polyadenylation and contact-map tracks), scores variants or scans windows on demand with the AlphaGenome model for human and mouse (variant scoring, in silico mutagenesis, REF-versus-ALT track prediction), and builds Atlas website deep links. Use when the user mentions AlphaGenome, AlphaGenome Atlas, AVI or AlphaGenome Variant Impact, DeepMind variant effect prediction, or wants to prioritise or mechanistically interpret non-coding, regulatory, splicing, enhancer, promoter, or chromatin-accessibility effects of SNVs from a VCF, credible set, or region. Research use only; not a clinical tool.

科研