Source profileQuality 82/100Review permissions

K-Dense-AI/scientific-agent-skills/skills/hypothesis-generation/SKILL.md

hypothesis-generation

Formulate evidence-bounded scientific questions, candidate hypotheses, rival explanations, causal or associational claims, discriminating predictions, measurements, and preregistration-ready analysis plans. Use when turning observations or preliminary findings into transparent, testable research plans without treating hypotheses as facts.

Source repository stars
31,966
Declared platforms
0
Static risk flags
1
Last source update
2026-07-28
Source checked
2026-07-28

Decision brief

What it does—and where it fits

Turn an observation into a transparent set of candidate explanations and tests. A hypothesis is a proposal to be challenged, not a finding, fact, diagnosis, or recommendation.

Best for

  • Use when turning observations or preliminary findings into transparent, testable research plans without treating hypotheses as facts.

Not for

  • Tasks that require unconfirmed production actions or broad system permissions.
  • Environments where the pinned source and install steps cannot be inspected.

Compatibility matrix

Platform support, with evidence labels

PlatformStatusEvidenceWhat to check
CodexNot declaredNo explicit evidencePortability before use
Claude CodeNot declaredNo explicit evidencePortability before use
CursorNot declaredNo explicit evidencePortability before use
Gemini CLINot declaredNo explicit evidencePortability before use
Open the compatibility checker

Installation

Inspect first. Install second.

The source command is displayed only when detected. A safe inspection prompt is always available so your agent can explain every action before execution.

Source-detected install commandSource
npx skills add https://github.com/K-Dense-AI/scientific-agent-skills --skill "skills/hypothesis-generation"
Safe inspection promptEditorial

Inspect the Agent Skill "hypothesis-generation" from https://github.com/K-Dense-AI/scientific-agent-skills/blob/e7ac42510774624f327003c95b6650e2883bc01d/skills/hypothesis-generation/SKILL.md at commit e7ac42510774624f327003c95b6650e2883bc01d. List every install step, command, network request, credential, file read/write, external action, and rollback step. Explain whether it fits my task. Do not install or execute anything until I approve.

Workflow

What the source asks the agent to do

  1. 01

    Workflow

    No script approval is an ethics, safety, regulatory, or scientific approval.

    accountable human owner and intended use;data sensitivity, authorization, retention, and permitted processing;affected people, animals, ecosystems, communities, or security interests;
  2. 02

    Non-negotiable boundaries

    Before using unpublished, sensitive, controlled, personal, proprietary, export-controlled, or security-relevant material:

    Confirm authorization and the applicable institutional, funder, publisher, data-use, privacy, and AI policies.Keep the material local unless an authorized human explicitly approves a named external destination and data scope.Minimize inputs. Do not place sensitive or unpublished data in web searches or external AI systems without authorization.
  3. 03

    Keep the objects distinct

    Do not collapse these labels. A mechanistic story is not a prediction; a prediction is not evidence; rejection of one null does not prove a mechanism; support for one candidate does not eliminate unconsidered rivals.

    Do not collapse these labels. A mechanistic story is not a prediction; a prediction is not evidence; rejection of one null does not prove a mechanism; support for one candidate does not eliminate unconsidered rivals.
  4. 04

    1. Run the scope and safety gate

    No script approval is an ethics, safety, regulatory, or scientific approval.

    accountable human owner and intended use;data sensitivity, authorization, retention, and permitted processing;affected people, animals, ecosystems, communities, or security interests;
  5. 05

    2. Freeze the observation

    Write the observation before interpretation:

    measurement or source;population, system, place, and time;unit of observation and unit of analysis;

Permission review

Static risk signals and limitations

Runs scripts

medium · line 168

The documentation asks the agent to run terminal commands or scripts.

python3 scripts/check_operationalization.py local-operationalization.json

Evidence record

Why each signal appears

EvidenceSourceComputedTestedEditorial
SignalValueEvidence typeMeaning
Quality score82/100ComputedDocumentation, specificity, maintenance, and trust rules
Repository stars31,966SourceRepository attention, not individual Skill quality
Compatibility0 platformsSourceDeclared in the catalog source record
Usage guideautomated source guideEditorialGenerated or reviewed according to the visible evidence level

Pinned source

Provenance and original SKILL.md

Repository
K-Dense-AI/scientific-agent-skills
Skill path
skills/hypothesis-generation/SKILL.md
Commit
e7ac42510774624f327003c95b6650e2883bc01d
License
MIT
Collected
2026-07-28
Default branch
main
View the original SKILL.md

Scientific Hypothesis Generation

Turn an observation into a transparent set of candidate explanations and tests. A hypothesis is a proposal to be challenged, not a finding, fact, diagnosis, or recommendation.

Non-negotiable boundaries

Before using unpublished, sensitive, controlled, personal, proprietary, export-controlled, or security-relevant material:

  1. Confirm authorization and the applicable institutional, funder, publisher, data-use, privacy, and AI policies.
  2. Keep the material local unless an authorized human explicitly approves a named external destination and data scope.
  3. Minimize inputs. Do not place sensitive or unpublished data in web searches or external AI systems without authorization.
  4. Stop at the appropriate human, animal, biosafety, dual-use, data-governance, or regulatory gate.

Never:

  • present a hypothesis, mechanism, causal effect, citation, or apparent pattern as established evidence;
  • claim novelty because a quick search found nothing;
  • infer causation from association, temporal order alone, predictive accuracy, or model output;
  • supply patient-specific diagnosis, treatment, dose, prognosis, or other clinical advice;
  • provide harmful experimental optimization or operational detail for pathogens, toxins, weapons, evasion, or other misuse;
  • bypass IRB/REC, IACUC, IBC, biosafety, dual-use, privacy, legal, or regulatory review;
  • fabricate sources, identifiers, search coverage, data, results, approvals, or preregistration;
  • automatically score, rank, select, accept, or reject scientific hypotheses.

If a request crosses a safety gate, produce only a high-level risk/oversight note and route it to the qualified local authority. Do not continue with operational detail.

Keep the objects distinct

ObjectMeaning
ObservationWhat was measured, noticed, or reported, with provenance and uncertainty
Research questionThe answerable question that defines scope
HypothesisA candidate explanatory or relational proposition
MechanismThe proposed process connecting conditions to an outcome
Causal estimandThe precisely defined causal contrast to estimate
PredictionAn observable implication derived before checking the target result
Alternative explanationA rival account, including bias or non-causal explanations
Null hypothesisA specified no-effect/no-difference model used by an analysis
Negative controlA control expected not to operate through the proposed mechanism
OperationalizationHow a construct becomes a variable, measurement, intervention, or category
Analysis planPrespecified transformations, models, contrasts, uncertainty, and decision rules
EvidenceObservations or sources that bear on a claim; never the claim itself

Do not collapse these labels. A mechanistic story is not a prediction; a prediction is not evidence; rejection of one null does not prove a mechanism; support for one candidate does not eliminate unconsidered rivals.

Workflow

1. Run the scope and safety gate

Record:

  • accountable human owner and intended use;
  • data sensitivity, authorization, retention, and permitted processing;
  • affected people, animals, ecosystems, communities, or security interests;
  • required ethics, feasibility, biosafety, dual-use, and regulatory reviews;
  • unresolved blocks and domain expertise needed.

No script approval is an ethics, safety, regulatory, or scientific approval.

2. Freeze the observation

Write the observation before interpretation:

  • measurement or source;
  • population, system, place, and time;
  • unit of observation and unit of analysis;
  • uncertainty, missingness, exclusions, and preprocessing;
  • whether the pattern was expected, exploratory, or selected after viewing results.

Use “reported,” “observed,” or “associated,” not causal language, unless a causal design and estimand justify it.

3. Frame the research question

Choose a framework only when it fits:

  • PICO/PICOT for intervention/effectiveness questions: population, intervention, comparator, outcome, and optionally time.
  • PECO for exposure questions.
  • Population–index test–reference standard–target condition for diagnostic accuracy.
  • Population–prognostic factor–outcome–time for prognosis.
  • A domain-specific construct–context–outcome frame for qualitative, descriptive, mechanistic, or theoretical work.

PICO is not a universal template. Define stakeholders, context, boundaries, feasibility, and what answer would change knowledge or practice. FINER is a question-refinement mnemonic—Feasible, Interesting, Novel, Ethical, Relevant—not a scoring system. Treat “Novel” as unresolved until a documented, fit-for-purpose search and expert review support it.

4. Establish a dated evidence boundary

Search before making literature-dependent statements. Prefer primary research, official policies, primary methods papers, current reporting guidelines, and systematic reviews used for orientation.

Record:

  • search date and cutoff;
  • databases/indexes, queries, filters, and screening boundary;
  • included and excluded source types;
  • sources supporting, challenging, or contextualizing each claim;
  • known access, language, database, and time limitations.

A search can establish what was searched, not universal absence. Say “not located within the documented search boundary,” never “no prior work exists.” Use assets/search_boundary_template.json, assets/evidence_ledger_template.csv, and references/literature_search_strategies.md.

5. Generate rivals before choosing tests

Create multiple candidates from genuinely different explanatory classes when plausible:

  • proposed mechanism;
  • measurement or processing artifact;
  • confounding or common cause;
  • selection or attrition;
  • conditioning on a collider;
  • reverse causation;
  • temporal, contextual, or boundary-condition differences;
  • stochastic variation;
  • competing mechanisms at another scale.

Generate an initial rival set independently before AI-assisted expansion to reduce anchoring and homogenization. Do not force a fixed number or false symmetry. Keep every candidate labeled candidate.

Platt’s strong-inference pattern motivates alternative hypotheses and crucial tests, but failed alternatives do not make the survivor true. Unknown alternatives, auxiliary assumptions, measurement error, and mixed mechanisms remain possible.

6. Declare the claim type and estimand

Classify each target as:

  • descriptive;
  • associational;
  • predictive;
  • causal;
  • mechanistic.

For a causal target, define before analysis:

  • target population or system;
  • intervention/exposure and comparator;
  • outcome and time horizon;
  • population-level summary;
  • treatment versions and intercurrent-event handling where relevant;
  • identification assumptions and target-trial/design analogue.

Document confounding, selection, collider, measurement, and reverse-causation risks separately. An observational causal estimate remains assumption-dependent. Use references/causal_inference_and_claims.md.

7. Derive discriminating predictions

For every candidate:

  1. State conditions and boundary conditions.
  2. Name the observable and measurement.
  3. State the expected pattern and uncertainty.
  4. State a result incompatible with the candidate under declared assumptions.
  5. Contrast the expected result with at least one rival.
  6. Define indeterminate outcomes and what would be learned from them.

Prefer tests where rivals predict meaningfully different outcomes. Add positive, procedural, and negative controls when scientifically appropriate. A negative control must be incapable of operating through the target mechanism while sharing relevant bias pathways; it is not a decorative untreated group.

Use assets/prediction_rival_matrix_template.csv and assets/falsification_controls_template.json.

8. Operationalize and validate measurement

For every construct record:

  • variable role and operational definition;
  • population/system, unit, timing, and conditions;
  • instrument/method, calibration, quality control, and masking;
  • reliability/repeatability;
  • validity evidence and applicability;
  • missingness, detection limits, transformations, cut points, and their rationales;
  • measurement invariance or cross-group comparability when relevant;
  • foreseeable measurement bias and limitations.

Do not treat a convenient proxy as the construct itself. Validate with:

python3 scripts/check_operationalization.py local-operationalization.json

9. Match design and analysis to the claim

Specify:

  • sampling, experimental unit, allocation, randomization, masking, and controls;
  • inclusion/exclusion and stopping rules;
  • sample-size, precision, or information rationale based on declared assumptions;
  • outcomes, contrasts, estimands, models, effect measures, and uncertainty;
  • missing-data and intercurrent-event handling;
  • multiplicity across outcomes, models, subgroups, looks, and hypotheses;
  • assumptions, diagnostics, robustness, and sensitivity analyses;
  • replication or independent validation plan;
  • what is confirmatory versus exploratory.

Do not use universal sample-size minima. Do not interpret a thresholded p-value as the probability a hypothesis is true or as effect importance. See references/experimental_design_patterns.md.

For intervention trials, use the current SPIRIT 2025 protocol guidance and CONSORT 2025 reporting guidance where applicable. These improve completeness; they do not certify design quality, ethics, or regulatory compliance.

10. Prevent HARKing and expose deviations

Before accessing the target outcomes, timestamp the question, candidates, predictions, outcomes, exclusions, transformations, analysis, multiplicity, missing-data plan, and stopping rule when feasible.

Afterward:

  • label data-dependent ideas and analyses exploratory;
  • preserve and report planned analyses;
  • list deviations with date, rationale, who decided, and expected impact;
  • never rewrite an observed pattern as an a priori prediction.

Preregistration is a transparent plan, not a ban on adaptation. Registered Reports add results-blind peer review and in-principle acceptance under journal policy. See references/preregistration_and_open_science.md.

11. Plan replication and updating

Distinguish:

  • reproducibility: consistent computational results from the same data/code/conditions;
  • replicability: consistency across studies collecting new data for the same question.

Preserve provenance, versions, code, materials, and decision logs when sharing is authorized. Plan independent replication or transport tests across relevant boundaries. Update candidate status when contrary, null, or replication evidence arrives; do not hide negative results.

12. Apply human accountability

The accountable human must verify:

  • every citation and source-to-claim link;
  • domain plausibility and measurement validity;
  • causal assumptions and statistical design;
  • ethics, feasibility, safety, privacy, and regulatory status;
  • all AI-assisted text, ideas, and citations;
  • whether broader expertise or community input is required.

AI can confabulate citations, anchor reasoning, and homogenize candidate sets. Record permitted AI use and material influence. Keep independent human ideation and rival generation in the process.

Local tool index

All CLIs are bounded, dependency-free, local, deterministic, and non-scoring:

TaskAssetCommand
Hypothesis-record schemaassets/hypothesis_record_template.jsonpython3 scripts/validate_hypothesis_schema.py record.json
Measurement checklistassets/operationalization_template.jsonpython3 scripts/check_operationalization.py checklist.json
Prediction/rival matrixassets/prediction_rival_matrix_template.csvpython3 scripts/validate_prediction_matrix.py matrix.csv
Claim-language lintAnnotated Markdownpython3 scripts/lint_causal_claims.py draft.md
Falsification/controlsassets/falsification_controls_template.jsonpython3 scripts/check_falsification_controls.py controls.json
Evidence/source auditassets/evidence_ledger_template.csv + assets/search_boundary_template.jsonpython3 scripts/audit_evidence_ledger.py ledger.csv boundary.json
Preregistration scaffoldassets/preregistration_scaffold_template.mdpython3 scripts/generate_preregistration_scaffold.py record.json -o preregistration.md

Exit codes are 0 for structurally valid output, 1 for completed validation with errors, and 2 for malformed/unsafe input. Reports validate declarations and internal consistency only; they do not verify scientific truth or choose a hypothesis. Full schemas are in references/tool_reference.md.

References

  • references/concepts_and_workflow.md — object model, strong inference, uncertainty, and candidate lifecycle
  • references/hypothesis_quality_criteria.md — non-scoring human review criteria
  • references/literature_search_strategies.md — traceable, bounded evidence search
  • references/causal_inference_and_claims.md — estimands and causal-bias risks
  • references/experimental_design_patterns.md — design, controls, measurement, multiplicity, and replication
  • references/preregistration_and_open_science.md — preregistration, Registered Reports, deviations, and open science
  • references/ethics_safety_and_ai.md — oversight gates, dual use, data handling, and responsible AI
  • references/tool_reference.md — CLI schemas, limits, and examples
  • references/source_ledger.md — dated authoritative source notes
  • references/security_validation.md — baseline findings and validation record

The bundled source ledger is assets/source_ledger.csv, verified through 2026-07-23. Recheck time-sensitive policy and guidance before a later or jurisdiction-specific use.

Alternatives

Compare before choosing

Computed 997

event4u-app/agent-config

design-review

Use when the user says "review the design", "check the UI", or wants a comprehensive UI/UX review. Uses a 7-phase methodology covering interaction, responsiveness, accessibility, and more.

Computed 9631,966

K-Dense-AI/scientific-agent-skills

neuropixels-analysis

Analyze Neuropixels extracellular recordings end-to-end with SpikeInterface. Covers loading SpikeGLX/Open Ephys/NWB data, preprocessing, drift/motion correction, Kilosort4 (and CPU) spike sorting, quality metrics, and unit curation (threshold-based, model-based UnitRefine, and AI-assisted visual review). Use when working with Neuropixels 1.0/2.0 recordings, spike sorting, or extracellular electrophysiology analysis.

Computed 9531,966

K-Dense-AI/scientific-agent-skills

scientific-brainstorming

Facilitates evidence-aware scientific ideation with independent generation, structured discussion, explicit assumptions, transparent evaluation, adversarial review, and decision logs. Use for early-stage research brainstorming or prioritizing candidate directions; hand off empirical validation, study design, ethics or regulatory review, and clinical questions to appropriate experts or skills.

Computed 9431,966

K-Dense-AI/scientific-agent-skills

citation-management

Comprehensive citation management for academic research. Search Google Scholar and PubMed for papers, extract accurate metadata, validate citations, and generate properly formatted BibTeX entries. This skill should be used when you need to find papers, verify citation information, convert DOIs to BibTeX, or ensure reference accuracy in scientific writing.