Skip to main content

이 저장소의 skills

ndpvt-web/arxiv-claude-skills - 5페이지

SkillsMP는 ndpvt-web/arxiv-claude-skills에서 651개의 skill을 수집했습니다. skill을 열어 소스와 세부 정보를 확인하세요.

ndpvt-web/arxiv-claude-skills

수집된 skill 651개 중 40개를 표시합니다.

직업 분류
소프트웨어 개발자
설명

Build hierarchical multi-agent systems that detect misinformation, anomalies, and deceptive content using perspective-aware aggregation. Agents are split into specialized Auditors, Coordinators, and a Decision-Maker to prevent information drowning. Use when:…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 품질 보증 분석가·테스터
설명

Evaluate, debug, and repair block-based Scratch programs using a three-layer executable protocol (VM execution, block-level edit distance, explanation rubrics). Use when: 'debug this Scratch project', 'analyze Scratch program logic', 'find bugs in my Scratch…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 품질 보증 분석가·테스터
설명

Generate high-quality unit tests with self-debugging repair loops and chain-of-thought reasoning. Produces tests with meaningful assertions, high branch coverage, and strong mutation scores by iteratively diagnosing and fixing errors, assertion failures, and…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Design and implement standardized, reproducible evaluation harnesses for LLM-based agents. Eliminates confounding factors (system prompts, tool configs, environment drift) so benchmark results reflect true model capability. Use when: 'build an agent…

원문 언어: 영어

업데이트
직업 분류
수학자
설명

Iterative generate-verify-revise agent for mathematical research problems. Implements the Aletheia loop: decompose a hard math problem, generate candidate proofs, verify with separate critical passes, revise on failure, and integrate literature search. Use…

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Extract structured key information from document images using schema-guided prompting for LMMs. Builds KIE pipelines that output normalized JSON from receipts, invoices, forms, medical records, tax documents, and more. Trigger phrases: 'extract fields from…

원문 언어: 영어

업데이트
직업 분류
시장조사 분석가·마케팅 전문가컴플라이언스 담당자
설명

Audit demographic bias in LLM-generated targeted text. Detects age- and gender-based stereotyping in personalized messaging by analyzing lexical content, language style, and persuasive framing asymmetries. Use when: 'audit my marketing copy for demographic…

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Adaptive curriculum-driven iterative optimization for autonomous ML engineering tasks. Uses Evolving Data Buffers and Learnability Potential sampling from the AceGRPO paper to structure multi-step agent workflows that avoid behavioral stagnation. Triggers:…

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Build AI-augmented annotation pipelines for creating high-quality information retrieval and QA datasets. Combines LLM-generated suggestions (questions, passage relevance scores, answer spans) with human review workflows to accelerate dataset creation. Use…

원문 언어: 영어

업데이트
직업 분류
재무 및 투자 분석가
설명

Iterative Dual-Phase Financial-PoT: decouple semantic reasoning from arithmetic computation to eliminate calculation errors in financial analysis. Use when: 'calculate financial ratios from reports', 'analyze annual report numbers', 'compute ROE/ROA from…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Analyze and optimize multi-agent code generation pipelines using causality-based importance ranking of intermediate features. Identifies which pipeline stages matter most, enables targeted failure repair, token-efficient pruning, and hybrid LLM backend…

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Compress user-specific memories for LLM personalization by clustering semantically similar memories and merging within clusters, reducing token count while preserving generation quality. Based on Bohdal et al. (ICASSP 2026). Use this skill when the user…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 품질 보증 분석가·테스터
설명

Assess LLM-generated code correctness using attribution graph analysis inspired by mechanistic interpretability. Apply structural reasoning diagnostics to identify buggy logic, predict failure modes, and suggest targeted fixes. Use when: 'analyze this code…

원문 언어: 영어

업데이트
직업 분류
인류학자 및 고고학자
설명

Compute the Conceptual Cultural Index (CCI) to measure cultural specificity of sentences using LLM-based generality estimates across culture sets. Use this skill when users ask to "measure cultural specificity", "score how culture-specific a sentence is",…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Build cost-efficient RAG pipelines for entity matching and deduplication using blocking-based batch retrieval and generation. Reduces LLM API calls and latency by grouping similar entity pairs into blocks before retrieval and inference. Use when the user asks…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Build reusable, parameterized skill libraries for computer-using agents (CUAs). Decomposes GUI automation into Skill Cells (intent), Parameterized Execution Graphs (actions), and Skill Composition Graphs (chaining). Use when: 'build a skill library for…

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Analyze and optimize LLM reasoning token efficiency using a multiplicative decomposition framework. Breaks down reasoning performance into completion rate, conditional correctness, verbosity, verbalization overhead, and coupling coefficient. Identifies…

원문 언어: 영어

업데이트
직업 분류
고등교육 컴퓨터공학 교원
설명

Apply Step-wise Marginal Information Gain (MIG) credit assignment to multi-step reasoning tasks. Evaluates each reasoning step by its marginal contribution toward the correct answer rather than by position or final outcome alone. Use this skill when asked to:…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Build dialect-aware RAG conversational agents that handle non-standard orthography, code-switching, and multi-script input. Uses a dual-path architecture: deterministic NLU for structured flows + RAG fallback for open-domain queries. Trigger phrases: 'build a…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Detect and defend against fraudulent content in LLM inputs using knowledge-graph-augmented analysis. Builds a fraud tactic-keyword bipartite graph, scores associations by confidence, prunes ambiguities, and augments prompts with XML-tagged keywords plus…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Reverse-engineer game mechanics from gameplay traces using a two-stage causal induction pipeline: first infer a Structural Causal Model (SCM) from observations, then translate it into executable game rules (VGDL or equivalent). Trigger phrases: 'infer game…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Select and deploy AI-generated image detection models based on threat-landscape analysis using zero-shot benchmark data from 23 detectors across 291 generators. Trigger phrases: - "detect AI-generated images" - "which deepfake detector should I use" -…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Build conversational Earth Observation agents that turn natural-language queries into executable, auditable Python workflows. Uses a unified API covering classification, segmentation, detection, spectral indices, and geospatial operations. Trigger phrases:…

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Build multidimensional profiling pipelines for large scientific paper corpora. Combines BERTopic clustering, LLM-structured extraction, and weighted semantic retrieval to analyze research trends, topic lifecycles, dataset/model adoption, and methodological…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Identify and remove interference tokens from prompts to improve LLM reasoning accuracy. Based on the LENS framework (Less Noise Sampling). Use when: 'clean up this prompt', 'my prompt isn't working well', 'purify this instruction', 'remove noise from prompt',…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Multi-agent framework for generating large-scale Verilog/RTL code from long hardware specifications by decomposing long-document-to-long-code into short-document, short-code tasks via information locality. Triggers: 'generate Verilog from spec', 'RTL…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Multi-agent tree-search code generation using heterogeneous agent collaboration with error-feedback refinement. Spawns multiple agents with distinct strategies to explore solution trees, iteratively refine code via structured error diagnostics, and select the…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Self-evolving multi-agent orchestration that dynamically generates specialized roles and collaboration topologies at inference time. Instead of fixed agent roles, MetaGen creates query-conditioned role specifications, builds a minimal execution DAG, and…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Apply PEARL's two-phase tool orchestration: offline tool exploration to learn valid usage patterns and failure modes, then structured plan-before-execute workflows for complex multi-step tool chains. Use when: 'plan a multi-tool workflow', 'chain multiple API…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Design LLM messaging systems that infuse Big Five personality traits for sustained user engagement. Uses aggregate-exposure personality alignment rather than per-message optimization. Trigger phrases: 'personality-aligned messages', 'BFPT messaging',…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Implement Scaling-Law Guided (SLG) Search for test-time compute optimization. Uses reward tail distribution estimation (GPD fitting) to predict scaling laws and dynamically allocate compute budget across candidate solutions. Trigger phrases: 'optimize…

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Systematically extract deep knowledge from LLMs using an interactive agentic framework with four adaptive exploration policies and a three-stage deduplication pipeline. Use this skill when the user says: "extract everything the model knows about X", "probe…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Stress-test LLM agents' social intelligence by injecting realistic communication barriers (semantic vagueness, sociocultural mismatch, emotional interference) into multi-agent dialogues, then measuring unresolved confusion and mutual understanding. Use when:…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Explore unfamiliar databases before writing SQL by building a local knowledge base of schema fragments, executable queries, and natural language descriptions. Uses MCTS-inspired schema traversal and dual-agent retrieval+generation. Trigger phrases: 'explore…

원문 언어: 영어

업데이트
직업 분류
물류 전문가
설명

Build reliable long-horizon supply chain agents using the SupChain-ReAct pattern: multi-path ReAct trajectories with majority voting for autonomous tool orchestration without handcrafted SOPs. Use when asked to 'build a supply chain agent', 'orchestrate…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Apply SWE-Pruner's goal-conditioned context pruning to reduce token usage when working with large codebases. Teaches Claude to selectively skim code like a human programmer: formulate an explicit focus goal, score lines by relevance, and drop low-value…

원문 언어: 영어

업데이트
직업 분류
비서 및 행정 보조원(법률, 의료 및 임원 제외)
설명

Summarize long documents while preserving global semantic structure and logical coherence using topology-guided pruning (GloSA-sum). Use when the user says: 'summarize this long document', 'compress this text for an LLM context window', 'extract key points…

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Guide Claude through evaluating and recommending LLM compression strategies (pruning, quantization, distillation) using the UniComp framework. Triggers: 'compress my model', 'quantize this LLM', 'prune a language model', 'which compression method should I…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Adaptive test-time scaling framework that decides WHEN and HOW MUCH to invoke expensive generative steps (world models, tool calls, API queries) before committing compute. Based on the AVIC (Adaptive Visual Imagination with Confidence) paper. Use when: 'build…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 품질 보증 분석가·테스터
설명

Pre-flight checker that prevents AI coding agent PRs from failing, based on empirical analysis of 33k agent-authored PRs on GitHub. Applies the rejection taxonomy from Ehsani et al. (MSR 2026) to catch the top causes of PR rejection before submission. Use…

원문 언어: 영어

업데이트
수집된 skill 651개 중 40개를 표시합니다.