用 Codex 或 Claude 帮你安装 复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它检查 Skill 页面并帮你完成安装。
直接命令不会经过审查 Prompt;运行前请先检查来源。
npx skills add https://github.com/OpenSenseNova/SenseNova-Skills --skill pivot-table-cross-analysis命令会保持在同一行。复制前请横向滚动并检查完整内容。
想先保存到本地?可下载 SkillsMP 当前能够提供的文件。
Base-layer skill for the SenseNova-Skills project, providing low-level APIs for image generation, recognition (VLM), and text optimization (LLM). This skill does not preprocess inputs; it only calls backend services and returns results. This skill is not user-facing and is intended for upper-layer skills only.
Standard and fast PPT pipeline. All LLM / VLM / T2I calls are wrapped in a single CLI entry (scripts/run_stage.py). The main agent's job is simple: emit ONE shell command per stage, never write loops, never write prompts. Standard mode plans thoroughly with a three-sample deck preview checkpoint (three concatenated deck images plus a preview URL), web research, image search, and user-selected final output format (PPTX or PDF) for polished, delivery-ready presentations. Fast mode builds a complete draft immediately with autonomous decisions, then provides structured refinement suggestions so the user can iterate quickly. Supports AI-generated infographics (U1) for diagrams and flowcharts, web image search (Serper) for real photos, and ECharts for data charts.
用于用户请求深度研究、系统性研究、竞品分析、方案对比、趋势分析或事实核查时。**遇到以下任一情况就主动使用本 skill,不要自行搜几条就回答**:①用户出现触发词:深度研究 / 深度调研 / 深入研究 / 全面研究 / 系统研究 / 调研 / 调查 / 尽调 / 行业研究 / 市场研究 / 竞品分析 / 政策研究 / 技术研究 / 趋势研究 / 事实核查 / 写一份研究报告 / 调研报告 / 深度报告 / research / deep research;②请求需要跨多来源取证、多维度对比、交叉验证才能给出可靠结论;③用户要求产出报告、白皮书、行业分析或尽调文档;④话题涉及最新政策/市场/产品/价格/法规,需要系统核查。明确要求核验来源的单点事实可走 quick;无核验要求的简单常识问答不使用。模糊或宽泛的"研究/了解一下 X"也优先触发。仅不用于:一句话摘要、已给定单一来源的整理、纯文字润色改写。
基于 SOC 职业分类
正在显示 SKILL.md
| name | pivot-table-cross-analysis |
| description | 利用交叉表与热力图对分类数据进行多维度占比分析,适用于奖项分布、绩效评估或市场占有率等结构化数据的清洗与可视化。 |
Step1 对原始数据进行清洗与重构,处理 Excel 合并单元格导致的缺失值,并筛选核心分析列。
import pandas as pd
def preprocess_pivot_data(file_path, target_cols=['奖项', '项目名称', '成员', '单位']):
"""
清理并重构数据列,处理合并单元格填充。
"""
df = pd.read_excel(file_path)
# 映射通用列名
df.columns = target_cols
# 关键技巧:处理合并单元格。ffill 前需确保数据按原始分类顺序排列
# 假设第一列为分类标签(如奖项名称)
df[target_cols[0]] = df[target_cols[0]].fillna(method='ffill')
# 删除关键信息(如成员或单位)缺失的无效行
df = df.dropna(subset=[target_cols[2], target_cols[3]])
# 清洗字符串空格
for col in df.select_dtypes(['object']).columns:
df[col] = df[col].str.strip()
return df
Step2 构建交叉分析表(Crosstab),计算不同维度下的频数分布及百分比占比。
def create_cross_analysis(df, index_col='单位', columns_col='奖项'):
"""
构建交叉表并计算各分类维度的获奖/分布比例。
"""
# 生成频数统计交叉表
cross_table = pd.crosstab(df[index_col], df[columns_col])
# 计算占比:各列(奖项)下各行(单位)的分布比例
# div(axis=1) 表示按列求和后进行除法
award_proportions = cross_table.div(cross_table.sum(axis=0), axis=1) * 100
# 技巧:生成带有总计行和占比的汇总表
summary = cross_table.copy()
summary['总计'] = summary.sum(axis=1)
summary.loc['合计'] = summary.sum()
return cross_table, award_proportions, summary
Step3 配置中文字体并生成热力图可视化,直观展示各维度间的分布差异。
import matplotlib.pyplot as plt
import seaborn as sns
def generate_analysis_heatmap(proportions, output_path='analysis_heatmap.png'):
"""
生成高分辨率热力图,支持中文字体显示。
"""
# 关键技巧:中文字体配置,兼容不同系统环境
plt.rcParams['font.sans-serif'] = ['SimHei', 'WenQuanYi Zen Hei', 'DejaVu Sans']
plt.rcParams['axes.unicode_minus'] = False
plt.figure(figsize=(14, 10))
# 使用 Seaborn 绘制热力图,fmt='.2f' 保留两位小数
sns.heatmap(
proportions,
annot=True,
fmt='.2f',
cmap='YlGnBu',
linewidths=.5,
cbar_kws={'label': '占比 (%)'}
)
plt.title('多维度分类占比分布热力图', fontsize=15, pad=20)
plt.xlabel('分类维度 (Columns)', fontsize=12)
plt.ylabel('分析对象 (Index)', fontsize=12)
# 自动调整布局防止标签裁剪
plt.tight_layout()
plt.savefig(output_path, dpi=300, bbox_inches='tight')
plt.close()
Step4 执行综合分析算法,提取各维度的 Top-N 表现对象并计算整体排名。
def extract_performance_insights(proportions, top_n=3):
"""
分析各奖项/分类下的领先者,并计算整体加权表现。
"""
insights = {}
# 1. 提取每个分类维度的前 N 名
top_performers = {}
for category in proportions.columns:
top_list = proportions[category].sort_values(ascending=False).head(top_n)
top_performers[category] = top_list.to_dict()
# 2. 计算整体表现排名(基于所有维度的平均占比)
overall_performance = proportions.mean(axis=1).sort_values(ascending=False)
insights['top_by_category'] = top_performers
insights['overall_ranking'] = overall_performance.head(10).to_dict()
return insights
Step5 导出分析结果为 Excel 多工作表格式,并提供下载链接。
from IPython.display import FileLink
def export_results(cross_table, proportions, insights_df, file_name='analysis_report.xlsx'):
"""
将分析结果保存至 Excel 并在环境中生成下载链接。
"""
with pd.ExcelWriter(file_name) as writer:
cross_table.to_excel(writer, sheet_name='频数统计')
proportions.to_excel(writer, sheet_name='占比分析')
insights_df.to_excel(writer, sheet_name='综合排名')
return FileLink(file_name)