Skip to main content
Run any Skill in Manus
with one click
GitHub repository

lemma

lemma contains 20 collected skills from tkpratardan, with repository-level occupation coverage and site-owned skill detail pages.

skills collected
20
Stars
4
updated
2026-07-15
Forks
1
Occupation coverage
1 occupation categories · 100% classified
repository explorer

Skills in this repository

lemma-baseline
data-scientists-152051

Establish a dumb baseline and an honest validation harness before any real model, so every later number means something.

2026-07-15
lemma-causal
data-scientists-152051

Rigor for causal questions and A/B tests (the effect of acting on X): confounding, post-treatment bias, valid control groups.

2026-07-15
lemma-describe
data-scientists-152051

Rigor for descriptive and diagnostic analytics (what happened and why): denominators, grain, and confounded slices, not model leakage.

2026-07-15
lemma-eda
data-scientists-152051

EDA kickoff for a fresh dataset: fixed opening scaffold (goal, imports, load, sanity), then chapters derived from the data; scan leakage, land a baseline.

2026-07-15
lemma-inference
data-scientists-152051

Rigor for statistical inference (is the difference real): hypothesis tests, power, multiple comparisons, effect size over p-value.

2026-07-15
lemma-leakage
data-scientists-152051

Audit a dataset or pipeline for the five leakages that inflate a metric: target, preprocessing, temporal, group, and sampling.

2026-07-15
lemma-model
data-scientists-152051

Final modeling once the baseline and feature set are locked: tune against validation, audit overfitting, touch the test set once, justify the complexity.

2026-07-15
lemma-review
data-scientists-152051

Review a notebook or analysis for data-science anti-patterns before it's trusted or shared.

2026-07-15
lemma-unsupervised
data-scientists-152051

Rigor for clustering, dimensionality reduction, and anomaly detection: validity is stability under resampling, not a held-out score.

2026-07-15
lemma-wrangle
data-scientists-152051

Assemble a trustworthy working dataset from messy or multiple sources: grain, keys, joins with match rates, extraction checks, lineage.

2026-07-15
lemma-baseline
data-scientists-152051

Establish an honest score to beat before complex modeling, including the validation design, metric, no-information rule, and simplest credible model.

2026-07-15
lemma-causal
data-scientists-152051

Estimate the effect of an intervention for experiments and defensible quasi-experimental or observational designs; do not substitute prediction for identification.

2026-07-15
lemma-describe
data-scientists-152051

Use for complex descriptive decompositions such as cohorts, funnels, segment comparisons, and what-changed investigations. Skip bounded lookups, joins, rankings, counts, averages, and aggregates.

2026-07-15
lemma-eda
data-scientists-152051

Explore a fresh dataset when the analytical direction is open; use for orientation, pattern discovery, and deciding what analysis is worth pursuing.

2026-07-15
lemma-inference
data-scientists-152051

Quantify whether a difference or association is distinguishable from sampling noise using effect estimates, uncertainty intervals, tests, or power analysis.

2026-07-15
lemma-leakage
data-scientists-152051

Audit suspicious model performance or a pipeline for target, preprocessing, temporal, group, sampling, or duplicate contamination.

2026-07-15
lemma-model
data-scientists-152051

Select and evaluate a production-worthy model after an honest baseline and validation design exist; use for tuning, calibration, thresholding, and final evaluation.

2026-07-15
lemma-review
data-scientists-152051

Review a notebook or analysis for correctness, reproducibility, leakage, weak validation, unsupported claims, and misleading communication.

2026-07-15
lemma-unsupervised
data-scientists-152051

Discover or evaluate structure without labels, including clustering, anomaly detection, embeddings, dimensionality reduction, and topic models.

2026-07-15
lemma-wrangle
data-scientists-152051

Reconcile sources into a defensible analytical dataset when grain, keys, definitions, units, authority, extraction, joins, or provenance are uncertain.

2026-07-15
lemma Agent Skills on GitHub | SkillsMP