Skip to main content

danielrosehill/Claude-Data-Analyst-plugin

SkillsMP 已收集 danielrosehill/Claude-Data-Analyst-plugin 中的 14 个 Skill。打开任一 Skill 可查看来源和详情。

最近记录的来源活动
SkillsMP 收录数据更新
已收集 skills
14
GitHub 星标
11
GitHub Forks
0

这个仓库中的 skills

2 个职业分类 · 已分类 100%

已展示 14 / 14 个已收集 Skill。

职业分类
数据科学家
描述

Produce a parametric PDF report describing a dataset — size, schema, distributions, key statistics, and findings from other skills — compiled via Typst. Use when the user wants a shareable, print-ready document about their data, not a one-off markdown summary.

原文语言:英语

更新
职业分类
数据科学家
描述

Describe and assess the sample size of a dataset — not just row count, but effective sample size per question the user wants to answer. Flags underpowered segments, imbalanced classes, small-n group cells, and gives a concrete "you can / cannot reliably claim…

原文语言:英语

更新
职业分类
数据科学家
描述

Compute and interpret standard deviation (and related spread measures — variance, IQR, MAD, CV) for numeric columns in a dataset. Handles sample vs. population formulas, grouped/stratified computation, and flags columns where SD is misleading (heavy skew,…

原文语言:英语

更新
职业分类
数据科学家
描述

Identify what the user is trying to analyse, diagnose gaps in the current dataset, propose external data sources that could fill them, then plan and implement the enrichment. Use when the dataset alone can't answer the user's question and extra context…

原文语言:英语

更新
职业分类
数据科学家
描述

Scan a dataset for signs that it has been pre-cleaned, normalised, imputed, smoothed, deduplicated, or otherwise processed before the user received it — data that is "suspiciously clean". Flag findings so the user knows whether they're analysing raw reality…

原文语言:英语

更新
职业分类
数据科学家
描述

Test relationships among three or more variables simultaneously — partial correlations, controlled effects, multicollinearity, interaction terms, and dimensionality reduction. Use when a pairwise correlation sweep isn't enough and the user wants to know how…

原文语言:英语

更新
职业分类
软件开发工程师
描述

Set up a "talk to your data" workspace in the current repo — discover local data files, load them into a DuckDB database, and append a CLAUDE.md block telling future Claude sessions how to query it. Use when the user wants to make a repo's data…

原文语言:英语

更新
职业分类
数据科学家
描述

Scan one or more datasets for data-type inconsistencies that would block analysis or relational/graph database loading — mixed types within a column, the same logical field typed differently across files, string-encoded numbers/dates, inconsistent null…

原文语言:英语

更新
职业分类
数据科学家
描述

Scan a dataset for significant anomalies — outliers, distribution shifts, impossible values, and unusual groupings. Use when the user wants a first-pass integrity and anomaly sweep of a CSV/Parquet/Excel file before deeper analysis.

原文语言:英语

更新
职业分类
数据科学家
描述

Detect and compute correlations between numeric variables in a dataset. Use when the user wants to see how variables in a CSV/Parquet/Excel file move together — Pearson, Spearman, or Kendall — with a short report flagging the strongest positive and negative…

原文语言:英语

更新
职业分类
数据科学家
描述

Generate a data dictionary for a dataset, combining automatic profiling with the user's description of what the data represents. Use when the user wants documentation of columns — names, types, semantic meaning, units, allowed values, and nullability — for a…

原文语言:英语

更新
职业分类
数据科学家
描述

Take a user-stated hypothesis and test it against the data, producing a report stating whether the data supports, refutes, or is inconclusive about the claim. Use when the user has a specific question or claim they want to interrogate against a dataset.

原文语言:英语

更新
职业分类
数据科学家
描述

Scan a dataset and flag columns or values that appear to contain personally identifiable information (PII). Use when the user wants a quick privacy audit of a CSV/Parquet/Excel file before sharing, publishing, or ingesting into another system.

原文语言:英语

更新
职业分类
数据科学家
描述

Identify and report the major trends a dataset depicts — directional changes over time, growth rates, seasonal patterns, segment shifts, and emerging categories. Use when the user wants the headline "what is this data saying" narrative rather than a specific…

原文语言:英语

更新
已展示 14 / 14 个已收集 Skill。