hypothesis-evolve-grounding
Generate exactly one grounded child hypothesis by strengthening evidence, specificity, and literature support.
把 Skill 的源码、资源快照、README、包体和安装信号放进一个可搜索、可筛选的公开目录。
Generate exactly one grounded child hypothesis by strengthening evidence, specificity, and literature support.
Generate exactly one child hypothesis that preserves the core idea while reducing unnecessary complexity.
Generate exactly one hypothesis candidate through a structured scientific debate.
Run the initial review gate for a hypothesis.
Update hypothesis proximity state by invoking the canonical embedding bridge for one hypothesis.
Run the decomposed review pipeline for a single hypothesis and persist each review stage as a structured artifact.
Query and download NCI Imaging Data Commons (IDC) cancer radiology and pathology datasets via the idc-index Python client. No authentication required: the parquet index ships inside the pip wheel, SQL…
LLM observability platform for tracing, evaluation, and monitoring. Use when debugging LLM applications, evaluating model outputs against datasets, monitoring production systems, or building systemati…
Generate exactly one cross-parent child hypothesis by transferring a useful principle from one parent context into another.
Run the full literature-grounded review for a hypothesis.
Generate exactly one literature-grounded hypothesis candidate for the active round.
Evaluate a hypothesis against prior observations.
Judge one ranked-tournament matchup between two top frontier hypotheses.
Summarize the completed reviews for a hypothesis.
Extract recurring critique patterns from the completed review of a hypothesis.
Conduct comprehensive, systematic literature reviews using multiple academic databases (PubMed, arXiv, bioRxiv, Semantic Scholar, etc.). This skill should be used when conducting systematic literature…
Generate exactly one divergent but still testable child hypothesis that challenges the shared assumptions of the parent set.
Generate exactly one hypothesis candidate by enumerating and combining testable assumptions.
Structured hypothesis formulation from observations. Use when you have experimental observations or data and need to formulate testable hypotheses with predictions, propose mechanisms, and design expe…
Judge one placement-tournament matchup between a candidate hypothesis and one opponent.
Update ranking artifacts for one reviewed hypothesis using canonical placement-opponent selection, ranked-frontier selection, tournament judgments, and Elo updates.
Simulate the hypothesis mechanism and identify failure scenarios.
Select the next island strategy and parent hypothesis set for one evolution round.
Karpathy's LLM Wiki — build and maintain a persistent, interlinked markdown knowledge base. Ingest sources, query compiled knowledge, and lint for consistency.