Skip to main content

이 저장소의 skills

synthetic-sciences/openscience - 4페이지

SkillsMP는 synthetic-sciences/openscience에서 313개의 skill을 수집했습니다. skill을 열어 소스와 세부 정보를 확인하세요.

synthetic-sciences/openscience

수집된 skill 313개 중 40개를 표시합니다.

직업 분류
소프트웨어 개발자
설명

Calculates training costs for Tinker fine-tuning jobs. Use when estimating costs for Tinker LLM training, counting tokens in datasets, or comparing Tinker model training prices. Tokenizes datasets using the correct model tokenizer and provides accurate cost…

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Infer gene regulatory networks (GRNs) from gene expression data using scalable algorithms (GRNBoost2, GENIE3). Use when analyzing transcriptomics data (bulk RNA-seq, single-cell RNA-seq) to identify transcription factor-target gene relationships and…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

PyTorch library for audio generation including text-to-music (MusicGen) and text-to-sound (AudioGen). Use when you need to generate music from text descriptions, create sound effects, or perform melody-conditioned music generation.

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Multiagent AI system for scientific research assistance that automates research workflows from data analysis to publication. This skill should be used when generating research ideas from datasets, developing research methodologies, executing computational…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

High-performance toolkit for genomic interval analysis in Rust with Python bindings. Use when working with genomic regions, BED files, coverage tracks, overlap detection, tokenization for ML models, or fragment analysis in computational genomics and machine…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

MATLAB and GNU Octave numerical computing for matrix operations, data analysis, visualization, and scientific computing. Use when writing MATLAB/Octave scripts for linear algebra, signal processing, image processing, differential equations, optimization,…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Pareto-aware molecular design balancing multiple ADMET properties simultaneously. Based on MultiMol (Yu 2025) and MOLLM (Ran 2025).

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Comprehensive toolkit for creating, analyzing, and visualizing complex networks and graphs in Python. Use when working with network/graph data structures, analyzing relationships between entities, computing graph algorithms (shortest paths, centrality,…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Bayesian modeling with PyMC. Build hierarchical models, MCMC (NUTS), variational inference, LOO/WAIC comparison, posterior checks, for probabilistic programming and inference.

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Multi-objective optimization framework. NSGA-II, NSGA-III, MOEA/D, Pareto fronts, constraint handling, benchmarks (ZDT, DTLZ), for engineering design and optimization problems.

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Cloud-based quantum chemistry platform with Python API. Preferred for computational chemistry workflows including pKa prediction, geometry optimization, conformer searching, molecular property calculations, protein-ligand docking (AutoDock Vina), and AI…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Machine learning in Python with scikit-learn. Use when working with supervised learning (classification, regression), unsupervised learning (clustering, dimensionality reduction), model evaluation, hyperparameter tuning, preprocessing, or building ML…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Model interpretability and explainability using SHAP (SHapley Additive exPlanations). Use this skill when explaining machine learning model predictions, computing feature importance, generating SHAP plots (waterfall, beeswarm, bar, scatter, force, heatmap),…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Process-based discrete-event simulation framework in Python. Use this skill when building simulations of systems with processes, queues, resources, and time-based events such as manufacturing systems, service operations, network traffic, logistics, or any…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Provides guidance for LLM post-training with RL using slime, a Megatron+SGLang framework. Use when training GLM models, implementing custom data generation workflows, or needing tight Megatron-LM integration for RL scaling.

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Guided statistical analysis with test selection and reporting. Use when you need help choosing appropriate tests for your data, assumption checking, power analysis, and APA-formatted results. Best for academic research reporting, test selection guidance. For…

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Statistical models library for Python. Use when you need specific model classes (OLS, GLM, mixed models, ARIMA) with detailed diagnostics, residuals, and inference. Best for econometrics, time series, rigorous inference with coefficient tables. For guided…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Use this skill when working with symbolic mathematics in Python. This skill should be used for symbolic computation tasks including solving equations algebraically, performing calculus operations (derivatives, integrals, limits), manipulating algebraic…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Graph Neural Networks (PyG). Node/graph classification, link prediction, GCN, GAT, GraphSAGE, heterogeneous graphs, molecular property prediction, for geometric deep learning.

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

UMAP dimensionality reduction. Fast nonlinear manifold learning for 2D/3D visualization, clustering preprocessing (HDBSCAN), supervised/parametric UMAP, for high-dimensional data.

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

This skill should be used for time series machine learning tasks including classification, regression, clustering, forecasting, anomaly detection, segmentation, and similarity search. Use when working with temporal data, sequential patterns, or time-indexed…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Distributed computing for larger-than-RAM pandas/NumPy workflows. Use when you need to scale existing pandas/NumPy code beyond memory or across clusters. Best for parallel file processing, distributed ML, integration with existing pandas code. For out-of-core…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Python library for working with geospatial vector data including shapefiles, GeoJSON, and GeoPackage files. Use when working with geographic data for spatial analysis, geometric operations, coordinate transformations, spatial joins, overlay operations,…

원문 언어: 영어

업데이트
직업 분류
데이터 과학자
설명

Patterns for loading PDE simulation datasets (PDEBench, PhiFlow, JAX-CFD) from HDF5 files. Handles layout detection (single tensor vs separate variables), spatial/temporal downsampling, multi-variable systems, HuggingFace and DaRUS data sources, and efficient…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Convert files and office documents to Markdown. Supports PDF, DOCX, PPTX, XLSX, images (with OCR), audio (with transcription), HTML, CSV, JSON, XML, ZIP, YouTube URLs, EPubs and more.

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Fast in-memory DataFrame library for datasets that fit in RAM. Use when pandas is too slow but data still fits in memory. Lazy evaluation, parallel execution, Apache Arrow backend. Best for 1-100GB datasets, ETL pipelines, faster pandas replacement. For…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Use this skill for processing and analyzing large tabular datasets (billions of rows) that exceed available RAM. Vaex excels at out-of-core DataFrame operations, lazy evaluation, fast aggregations, efficient visualization of big data, and machine learning on…

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Chunked N-D arrays for cloud storage. Compressed arrays, parallel I/O, S3/GCS integration, NumPy/Dask/Xarray compatible, for large-scale scientific computing pipelines.

원문 언어: 영어

업데이트
직업 분류
기타 생물 과학자
설명

Access AlphaFold 200M+ AI-predicted protein structures. Retrieve structures by UniProt ID, download PDB/mmCIF files, analyze confidence metrics (pLDDT, PAE), for drug discovery and structural biology.

원문 언어: 영어

업데이트
직업 분류
기타 생물 과학자
설명

Efficient database search tool for bioRxiv preprint server. Use this skill when searching for life sciences preprints by keywords, authors, date ranges, or categories, retrieving paper metadata, downloading PDFs, or conducting literature reviews.

원문 언어: 영어

업데이트
직업 분류
기타 생물 과학자
설명

Access BRENDA enzyme database via SOAP API. Retrieve kinetic parameters (Km, kcat), reaction equations, organism data, and substrate-specific enzyme information for biochemical research and metabolic pathway analysis.

원문 언어: 영어

업데이트
직업 분류
기타 생물 과학자
설명

Query the CELLxGENE Census (61M+ cells) programmatically. Use when you need expression data across tissues, diseases, or cell types from the largest curated single-cell atlas. Best for population-scale queries, reference atlas comparisons. For analyzing your…

원문 언어: 영어

업데이트
직업 분류
기타 생물 과학자
설명

Query ChEMBL bioactive molecules and drug discovery data. Search compounds by structure/properties, retrieve bioactivity data (IC50, Ki), find inhibitors, perform SAR studies, for medicinal chemistry.

원문 언어: 영어

업데이트
직업 분류
역학자
설명

Query ClinicalTrials.gov via API v2. Search trials by condition, drug, location, status, or phase. Retrieve trial details by NCT ID, export data, for clinical research and patient matching.

원문 언어: 영어

업데이트
직업 분류
기타 생물 과학자
설명

Access ClinPGx pharmacogenomics data (successor to PharmGKB). Query gene-drug interactions, CPIC guidelines, allele functions, for precision medicine and genotype-guided dosing decisions.

원문 언어: 영어

업데이트
직업 분류
기타 생물 과학자
설명

Query NCBI ClinVar for variant clinical significance. Search by gene/position, interpret pathogenicity classifications, access via E-utilities API or FTP, annotate VCFs, for genomic medicine.

원문 언어: 영어

업데이트
직업 분류
기타 생물 과학자
설명

Access COSMIC cancer mutation database. Query somatic mutations, Cancer Gene Census, mutational signatures, gene fusions, for cancer research and precision oncology. Requires authentication.

원문 언어: 영어

업데이트
직업 분류
소프트웨어 개발자
설명

Work with Data Commons, a platform providing programmatic access to public statistical data from global sources. Use this skill when working with demographic data, economic indicators, health statistics, environmental data, or any public datasets available…

원문 언어: 영어

업데이트
직업 분류
기타 생물 과학자
설명

Access and analyze comprehensive drug information from the DrugBank database including drug properties, interactions, targets, pathways, chemical structures, and pharmacology data. This skill should be used when working with pharmaceutical data, drug…

원문 언어: 영어

업데이트
직업 분류
기타 생물 과학자
설명

Access European Nucleotide Archive via API/FTP. Retrieve DNA/RNA sequences, raw reads (FASTQ), genome assemblies by accession, for genomics and bioinformatics pipelines. Supports multiple formats.

원문 언어: 영어

업데이트
수집된 skill 313개 중 40개를 표시합니다.