Skip to content

Data Science & AI Research Skills

794 data science, AI research, and analysis skills for coding agents. Data pipeline, ML model dev, statistical analysis - pre-verified and MCP-ready.

Research Pipeline

Full end-to-end research pipeline: from a broad research direction through idea discovery, experiments, and review all the way to a polished paper PDF. Use when user says "全流程", "full pipeline", "从找idea到投稿", "end-to-end research", or wants the complete autonomous research lifecycle.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/skills-codex/research-pipeline

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Research Refine

Turn a vague research direction into a problem-anchored, elegant, frontier-aware, implementation-oriented method plan via iterative Gemini review. Use when the user says "refine my approach", "帮我细化方案", "decompose this problem", "打磨idea", "refine research plan", "细化研究方案", or wants a concrete research method that stays simple, focused, and top-venue ready instead of a vague or overbuilt idea.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/skills-codex-gemini-review/research-refine

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Research Refine

Turn a vague research direction into a problem-anchored, elegant, frontier-aware, implementation-oriented method plan via iterative GPT-5.5 review. Use when the user says "refine my approach", "帮我细化方案", "decompose this problem", "打磨idea", "refine research plan", "细化研究方案", or wants a concrete research method that stays simple, focused, and top-venue ready instead of a vague or overbuilt idea.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/skills-codex/research-refine

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Research Refine

Turn a vague research direction into a problem-anchored, elegant, frontier-aware, implementation-oriented method plan via iterative GPT-5.4 review. Use when the user says "refine my approach", "帮我细化方案", "decompose this problem", "打磨idea", "refine research plan", "细化研究方案", or wants a concrete research method that stays simple, focused, and top-venue ready instead of a vague or overbuilt idea.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/skills-codex-claude-review/research-refine

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Research Refine

Turn a vague research direction into a problem-anchored, elegant, frontier-aware, implementation-oriented method plan via iterative GPT-5.5 review. Use when the user says "refine my approach", "帮我细化方案", "decompose this problem", "打磨idea", "refine research plan", "细化研究方案", or wants a concrete research method that stays simple, focused, and top-venue ready instead of a vague or overbuilt idea.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/research-refine

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Research Refine Pipeline

Run an end-to-end workflow that chains `research-refine` and `experiment-plan`. Use when the user wants a one-shot pipeline from vague research direction to focused final proposal plus detailed experiment roadmap, or asks to "串起来", build a pipeline, do it end-to-end, or generate both the method and experiment plan together.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/research-refine-pipeline

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Research Refine Pipeline

Run an end-to-end workflow that chains `research-refine` and `experiment-plan`. Use when the user wants a one-shot pipeline from vague research direction to focused final proposal plus detailed experiment roadmap, or asks to "串起来", build a pipeline, do it end-to-end, or generate both the method and experiment plan together.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/skills-codex/research-refine-pipeline

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Research Review

Get a deep critical review of research from Claude via claude-review MCP. Use when user says "review my research", "help me review", "get external review", or wants critical feedback on research ideas, papers, or experimental results.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/skills-codex-claude-review/research-review

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Research Review

Get a deep critical review of research from GPT using a secondary Codex agent. Use when user says "review my research", "help me review", "get external review", or wants critical feedback on research ideas, papers, or experimental results.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/skills-codex/research-review

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Research Review

Get a deep critical review of research from Gemini via gemini-review MCP. Use when user says "review my research", "help me review", "get external review", or wants critical feedback on research ideas, papers, or experimental results.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/skills-codex-gemini-review/research-review

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Research Review

Get a deep critical review of research from an external reviewer backend (Codex or manual). Use when user says "review my research", "help me review", "get external review", or wants critical feedback on research ideas, papers, or experimental results.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/research-review

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Research Summarizer

Structured research summarization agent skill for non-dev users. Handles academic papers, web articles, reports, and documentation. Extracts key findings, generates comparative analyses, and produces properly formatted citations. Use when: user wants to summarize a research paper, compare multiple sources, extract citations from documents, or create structured research briefs. Plugin for Claude Code, Codex, Gemini CLI, and OpenClaw.

by alirezarezvani/claude-skills / product-team/research-summarizer/skills/research-summarizer

#work-life#productivity#businessData, AI & Research

Research Synthesis

Synthesize user research into themes, insights, and recommendations. Use when you have interview transcripts, survey results, usability test notes, support tickets, or NPS responses that need to be distilled into patterns, user segments, and prioritized next steps.

by anthropics/knowledge-work-plugins / design/skills/research-synthesis

#work-life#productivity#knowledge-workData, AI & Research

Research Wiki

Persistent research knowledge base that accumulates papers, ideas, experiments, claims, and their relationships across the entire research lifecycle. Inspired by Karpathy's LLM Wiki pattern. Use when user says "知识库", "research wiki", "add paper", "wiki query", "查知识库", or wants to build/query a persistent field map.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/research-wiki

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Research Wiki

Persistent research knowledge base that accumulates papers, ideas, experiments, claims, and their relationships across the entire research lifecycle. Inspired by Karpathy's LLM Wiki pattern. Use when user says "知识库", "research wiki", "add paper", "wiki query", "查知识库", or wants to build/query a persistent field map.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/skills-codex/research-wiki

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Resemble Detect

Deepfake detection and media safety — detect AI-generated audio, images, video, and text, trace synthesis sources, apply watermarks, verify speaker identity, and analyze media intelligence using Resemble AI

by github/awesome-copilot / skills/resemble-detect

#github-copilot#deep#researchData, AI & Research

Resubmit Pipeline

Workflow 5: orchestrate a text-only resubmit of a polished paper to a different venue under hard constraints (no new experiments, no bib edits, no framework changes, never overwrite prior submissions). Use when user says "resubmit pipeline", "重投流程", "port paper to <new venue>", "resubmit to <venue>", "tighten paper for resubmission", or has a rejected/withdrawn paper to move to a different top venue under tight time budget.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/resubmit-pipeline

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Result To Claim

Use when experiments complete to judge what claims the results support, what they don't, and what evidence is still missing. Codex MCP evaluates results against intended claims and routes to next action (pivot, supplement, or confirm). Use after experiments finish — before writing the paper or running ablations.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/result-to-claim

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Result To Claim

Use when experiments complete to judge what claims the results support, what they don't, and what evidence is still missing. A secondary Codex agent evaluates results against intended claims and routes to next action (pivot, supplement, or confirm). Use after experiments finish — before writing the paper or running ablations.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/skills-codex/result-to-claim

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Return Calculations

Compute and compare investment return metrics including TWR, MWR (dollar-weighted IRR on portfolio cash flows), CAGR, and annualized returns. Use when the user asks about portfolio performance calculation, comparing manager returns, linking sub-period returns, understanding why different return methods give different numbers, converting returns across time periods, or computing the IRR of an investor's own contributions and withdrawals. Also trigger when users mention 'how much did I make', 'annual return', 'compound growth', 'dollar-weighted vs time-weighted', 'what was my rate of return', 'geometric vs arithmetic mean', 'log returns', or ask about the effect of cash flows on reported returns. For project or loan IRR, NPV, and generic 'solve for the rate' problems, use time-value-of-money instead.

by JoelLewis/finance_skills / plugins/core/skills/return-calculations

#finance#personal-finance#wealth-managementData, AI & Research

Risk Metrics Calculation

Calculate portfolio risk metrics including VaR, CVaR, Sharpe, Sortino, and drawdown analysis. Use when measuring portfolio risk, implementing risk limits, or building risk monitoring systems.

by wshobson/agents / plugins/quantitative-trading/skills/risk-metrics-calculation

#github#broad-capability#externalData, AI & Research

Rowan: Cloud-Native Molecular-Modeling and Drug-Design Workflows

Rowan is a cloud-native molecular modeling and medicinal-chemistry workflow platform with a Python API. Use for pKa and macropKa prediction, conformer and tautomer ensembles, docking and analogue docking, protein-ligand cofolding, MSA generation, molecular dynamics, permeability, descriptor workflows, and related small-molecule or protein modeling tasks. Ideal for programmatic batch screening, multi-step chemistry pipelines, and workflows that would otherwise require maintaining local HPC/GPU infrastructure.

by K-Dense-AI/scientific-agent-skills / scientific-skills/rowan

#github#broad-capability#externalData, AI & Research

Run Experiment

Deploy and run ML experiments on local, remote, Vast.ai, or Modal serverless GPU. Use when user says "run experiment", "deploy to server", "跑实验", or needs to launch training jobs.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/run-experiment

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Run Experiment

Deploy and run ML experiments on local or remote GPU servers. Use when user says "run experiment", "deploy to server", "跑实验", or needs to launch training jobs.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/skills-codex/run-experiment

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Run Models

Run AI models on Replicate via predictions, webhooks, and streaming.

by replicate/skills / skills/run-models

#ml-inference#ml#experimentData, AI & Research

Running Dbt Commands

Formats and executes dbt CLI commands, selects the correct dbt executable, and structures command parameters. Use when running models, tests, builds, compiles, or show queries via dbt CLI. Use when unsure which dbt executable to use or how to format command parameters.

by dbt-labs/dbt-agent-skills / skills/dbt/skills/running-dbt-commands

#dbt#analytics-engineering#dataData, AI & Research

RWKV - Receptance Weighted Key Value

RNN+Transformer hybrid with O(n) inference. Linear time, infinite context, no KV cache. Train like GPT (parallel), infer like RNN (sequential). Linux Foundation AI project. Production at Windows, Office, NeMo. RWKV-7 (March 2025). Models up to 14B parameters.

by Orchestra-Research/AI-Research-SKILLs / 01-model-architecture/rwkv

#github#broad-capability#externalData, AI & Research

Sag

ElevenLabs text-to-speech with mac-style say UX.

by openclaw/openclaw / skills/sag

#deep#researchData, AI & Research

Scandinavia Transit

Search trains, buses, and ferries in Norway (Entur), Sweden (ResRobot), and Denmark (Rejseplanen). Intra-Scandinavia ground transport with schedules and Danish fare pricing.

by borski/travel-hacking-toolkit / plugins/travel-hacking-toolkit/skills/scandinavia-transit

#travel#flights#hotelsData, AI & Research

Scanpy

Standard single-cell RNA-seq analysis pipeline. Use for QC, normalization, dimensionality reduction (PCA/UMAP/t-SNE), clustering, differential expression, visualization, and converting R-friendly single-cell formats such as Seurat or SingleCellExperiment RDS files into h5ad for Scanpy. Best for exploratory scRNA-seq analysis with established workflows. For deep learning models use scvi-tools; for data format questions use anndata.

by K-Dense-AI/scientific-agent-skills / skills/scanpy

#k-dense-ai-claude-scientific-skills#single#cellData, AI & Research

Scanpy: Single-Cell Analysis

Standard single-cell RNA-seq analysis pipeline. Use for QC, normalization, dimensionality reduction (PCA/UMAP/t-SNE), clustering, differential expression, and visualization. Best for exploratory scRNA-seq analysis with established workflows. For deep learning models use scvi-tools; for data format questions use anndata.

by K-Dense-AI/scientific-agent-skills / scientific-skills/scanpy

#github#external#license-mitData, AI & Research

Scientific Brainstorming

Open-ended scientific ideation partner. Use for research gaps, mechanism exploration, interdisciplinary connections, assumptions, possible research directions, and lightweight literature matrix or A+B paper-combination idea mapping. For structured testable hypotheses and validation plans, use hypothesis-generation instead.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/scientific-brainstorming

#broad-capability#creative#researchData, AI & Research

Scientific Brainstorming

Creative research ideation and exploration. Use for open-ended brainstorming sessions, exploring interdisciplinary connections, challenging assumptions, or identifying research gaps. Best for early-stage research planning when you do not have specific observations yet. For formulating testable hypotheses from data use hypothesis-generation.

by K-Dense-AI/scientific-agent-skills / scientific-skills/scientific-brainstorming

#github#broad-capability#externalData, AI & Research

Scientific Critical Thinking

Evaluate scientific claims and evidence quality. Use for assessing experimental design validity, identifying biases and confounders, applying evidence grading frameworks (GRADE, Cochrane Risk of Bias), or teaching critical analysis. Best for understanding evidence quality, identifying flaws. For formal peer review writing use peer-review.

by K-Dense-AI/scientific-agent-skills / skills/scientific-critical-thinking

#k-dense-ai-claude-scientific-skills#scientific#criticalData, AI & Research

Scientific Critical Thinking

Evaluate scientific claims and evidence quality. Use for assessing experimental design validity, identifying biases and confounders, applying evidence grading frameworks (GRADE, Cochrane Risk of Bias), or teaching critical analysis. Best for understanding evidence quality, identifying flaws. For formal peer review writing use peer-review.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/scientific-critical-thinking

#broad-capability#creative#researchData, AI & Research

Scientific Data Preprocessing

⚠️ CRITICAL USER EXPERIENCE-BASED SKILL - ALWAYS CONSULT BEFORE DATA PREPROCESSING ⚠️ Prevents catastrophic errors (88.9% error rate in V1.0 case study) through multi-level feature analysis, data leakage detection, and semantic validation. MANDATORY for: data preprocessing, feature engineering, standardization, normalization, interpolation, missing value handling, feature selection, or ANY data transformation task. Covers grouped time-series, cross-sectional, panel data. Detects: time travel leakage, causal inversion, ID misuse, semantic-numeric fallacies, distribution blindness. User's hard-won lessons from real project failures.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/scientific-data-preprocessing

#broad-capability#creative#structuredData, AI & Research

Scientific Hypothesis Generation

Structured hypothesis formulation from observations. Use when you have experimental observations or data and need to formulate testable hypotheses with predictions, propose mechanisms, and design experiments to test them. Follows scientific method framework. For open-ended ideation use scientific-brainstorming; for automated LLM-driven hypothesis testing on datasets use hypogenic.

by K-Dense-AI/scientific-agent-skills / scientific-skills/hypothesis-generation

#github#external#license-mitData, AI & Research

Scientific Paper Research

Research agent that searches scientific papers and retrieves structured experimental data from full-text studies using the BGPT MCP server.

by github/awesome-copilot / agents/scientific-paper-research.agent.md

#github-copilot#literature#reviewData, AI & Research

Scientific Schematics and Diagrams

Create publication-quality scientific diagrams using Nano Banana 2 AI with smart iterative refinement. Uses Gemini 3.1 Pro Preview for quality review. Only regenerates if quality is below threshold for your document type. Specialized in neural network architectures, system diagrams, flowcharts, biological pathways, and complex scientific visualizations.

by K-Dense-AI/scientific-agent-skills / scientific-skills/scientific-schematics

#github#broad-capability#externalData, AI & Research

Scientific Visualization

Create publication figures with matplotlib/seaborn/plotly. Multi-panel layouts, error bars, significance markers, colorblind-safe, export PDF/EPS/TIFF, for journal-ready scientific plots.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/scientific-visualization

#broad-capability#creative#scientificData, AI & Research

Scientific Visualization

Meta-skill for publication-ready figures. Use when creating journal submission figures requiring multi-panel layouts, significance annotations, error bars, colorblind-safe palettes, and specific journal formatting (Nature, Science, Cell). Orchestrates matplotlib/seaborn/plotly with publication styles. For quick exploration use seaborn or plotly directly.

by K-Dense-AI/scientific-agent-skills / scientific-skills/scientific-visualization

#github#external#license-mitData, AI & Research

Scientific Visualization

Meta-skill for publication-ready figures. Use when creating journal submission figures requiring multi-panel layouts, significance annotations, error bars, colorblind-safe palettes, and specific journal formatting (Nature, Science, Cell). Orchestrates matplotlib/seaborn/plotly with publication styles. For quick exploration use seaborn or plotly directly.

by K-Dense-AI/scientific-agent-skills / skills/scientific-visualization

#k-dense-ai-claude-scientific-skills#scientific#visualizationData, AI & Research

scikit-bio

Biological data toolkit. Sequence analysis, alignments, phylogenetic trees, diversity metrics (alpha/beta, UniFrac), ordination (PCoA), PERMANOVA, FASTA/Newick I/O, for microbiome analysis.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/scikit-bio

#broad-capability#github#externalData, AI & Research

Scikit-learn

Machine learning in Python with scikit-learn. Use when working with supervised learning (classification, regression), unsupervised learning (clustering, dimensionality reduction), model evaluation, hyperparameter tuning, preprocessing, or building ML pipelines. Provides comprehensive reference documentation for algorithms, preprocessing techniques, pipelines, and best practices.

by K-Dense-AI/scientific-agent-skills / scientific-skills/scikit-learn

#github#external#license-mitData, AI & Research

Scikit Learn

Machine learning in Python with scikit-learn. Use when working with supervised learning (classification, regression), unsupervised learning (clustering, dimensionality reduction), model evaluation, hyperparameter tuning, preprocessing, or building ML pipelines. Provides comprehensive reference documentation for algorithms, preprocessing techniques, pipelines, and best practices.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/scikit-learn

#broad-capability#creative#deepData, AI & Research

Scikit Survival

Comprehensive toolkit for survival analysis and time-to-event modeling in Python using scikit-survival. Use this skill when working with censored survival data, performing time-to-event analysis, fitting Cox models, Random Survival Forests, Gradient Boosting models, or Survival SVMs, evaluating survival predictions with concordance index or Brier score, handling competing risks, or implementing any survival analysis workflow with the scikit-survival library.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/scikit-survival

#broad-capability#creative#deepData, AI & Research

Scipy Optimization

Optimize pump designs and system parameters using scipy.optimize

by Soljourner/claude-engineering-skills / skills/packages/scipy-optimization

#broad-capability#engineering#fluid-dynamicsData, AI & Research

Scvelo

RNA velocity analysis with scVelo. Estimate cell state transitions from unspliced/spliced mRNA dynamics, infer trajectory directions, compute latent time, and identify driver genes in single-cell RNA-seq data. Complements Scanpy/scVI-tools for trajectory inference.

by K-Dense-AI/scientific-agent-skills / skills/scvelo

#broad-capability#science#mathData, AI & Research

scVelo — RNA Velocity Analysis

RNA velocity analysis with scVelo. Estimate cell state transitions from unspliced/spliced mRNA dynamics, infer trajectory directions, compute latent time, and identify driver genes in single-cell RNA-seq data. Complements Scanpy/scVI-tools for trajectory inference.

by K-Dense-AI/scientific-agent-skills / scientific-skills/scvelo

#github#broad-capability#externalData, AI & Research

scvi-tools

Deep generative models for single-cell omics. Use when you need probabilistic batch correction (scVI), transfer learning, differential expression with uncertainty, or multi-modal integration (TOTALVI, MultiVI). Best for advanced modeling, batch effects, multimodal data. For standard analysis pipelines use scanpy.

by K-Dense-AI/scientific-agent-skills / scientific-skills/scvi-tools

#github#broad-capability#externalData, AI & Research

Scvi Tools

Deep generative models for single-cell omics. Use when you need probabilistic batch correction (scVI), transfer learning, differential expression with uncertainty, or multi-modal integration (TOTALVI, MultiVI). Best for advanced modeling, batch effects, multimodal data. For standard analysis pipelines use scanpy.

by K-Dense-AI/scientific-agent-skills / skills/scvi-tools

#broad-capability#science#mathData, AI & Research

Seaborn

Statistical visualization with pandas integration. Use for quick exploration of distributions, relationships, and categorical comparisons with attractive defaults. Best for box plots, violin plots, pair plots, heatmaps. Built on matplotlib. For interactive plots use plotly; for publication styling use scientific-visualization.

by K-Dense-AI/scientific-agent-skills / skills/seaborn

#k-dense-ai-claude-scientific-skills#data#visualizationData, AI & Research

Seaborn

Statistical visualization. Scatter, box, violin, heatmaps, pair plots, regression, correlation matrices, KDE, faceted plots, for exploratory analysis and publication figures.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/seaborn

#broad-capability#creative#dataData, AI & Research

Search & AI Optimization Expert

Expert guidance for modern search optimization: SEO, Answer Engine Optimization (AEO), and Generative Engine Optimization (GEO) with AI-ready content strategies

by github/awesome-copilot / agents/search-ai-optimization-expert.agent.md

#github-copilot#awesome-copilot#externalData, AI & Research

Searxng Search

Free meta-search via SearXNG — aggregates results from 70+ search engines. Self-hosted or use a public instance. No API key needed. Falls back automatically when the web search toolset is unavailable.

by NousResearch/hermes-agent / optional-skills/research/searxng-search

#broad-capability#development#creativeData, AI & Research

Seatmaps

Aircraft seat maps, cabin dimensions, and seat recommendations via SeatMaps.com and AeroLOPA. Search by flight number or airline+aircraft via agent-browser.

by borski/travel-hacking-toolkit / plugins/travel-hacking-toolkit/skills/seatmaps

#travel#flights#hotelsData, AI & Research

Seats Aero

Search award flight availability across 27 mileage programs via Seats.aero Partner API. Find cheapest award flights, compare programs, and get booking links.

by borski/travel-hacking-toolkit / plugins/travel-hacking-toolkit/skills/seats-aero

#travel#flights#hotelsData, AI & Research

Segment Anything Model

SAM: zero-shot image segmentation via points, boxes, masks.

by NousResearch/hermes-agent / skills/mlops/models/segment-anything

#broad-capability#development#creativeData, AI & Research

Segment Anything Model (SAM)

Foundation model for image segmentation with zero-shot transfer. Use when you need to segment any object in images using points, boxes, or masks as prompts, or automatically generate all object masks in an image.

by Orchestra-Research/AI-Research-SKILLs / 18-multimodal/segment-anything

#github#broad-capability#externalData, AI & Research

Semantic Scholar

Search published venue papers (IEEE, ACM, Springer, etc.) via Semantic Scholar API. Complements /arxiv (preprints) with citation counts, venue metadata, and TLDR. Use when user says "search semantic scholar", "find IEEE papers", "find journal papers", "venue papers", "citation search", or wants published literature beyond arXiv preprints.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/semantic-scholar

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Semantic Scholar

Search published venue papers (IEEE, ACM, Springer, etc.) via Semantic Scholar API. Complements /arxiv (preprints) with citation counts, venue metadata, and TLDR. Use when user says "search semantic scholar", "find IEEE papers", "find journal papers", "venue papers", "citation search", or wants published literature beyond arXiv preprints.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/skills-codex/semantic-scholar

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Senior Computer Vision Engineer

World-class computer vision skill for image/video processing, object detection, segmentation, and visual AI systems. Expertise in PyTorch, OpenCV, YOLO, SAM, diffusion models, and vision transformers. Includes 3D vision, video analysis, real-time processing, and production deployment. Use when building vision AI systems, implementing object detection, training custom vision models, or optimizing inference pipelines.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/senior-computer-vision

#broad-capability#github#externalData, AI & Research

Senior Data Scientist

World-class data science skill for statistical modeling, experimentation, causal inference, and advanced analytics. Expertise in Python (NumPy, Pandas, Scikit-learn), R, SQL, statistical methods, A/B testing, time series, and business intelligence. Includes experiment design, feature engineering, model evaluation, and stakeholder communication. Use when designing experiments, building predictive models, performing causal analysis, or driving data-driven decisions.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/senior-data-scientist

#broad-capability#github#externalData, AI & Research

Senior ML/AI Engineer

World-class ML engineering skill for productionizing ML models, MLOps, and building scalable ML systems. Expertise in PyTorch, TensorFlow, model deployment, feature stores, model monitoring, and ML infrastructure. Includes LLM integration, fine-tuning, RAG systems, and agentic AI. Use when deploying ML models, building ML platforms, implementing MLOps, or integrating LLMs into production systems.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/senior-ml-engineer

#broad-capability#github#externalData, AI & Research

Senior Prompt Engineer

World-class prompt engineering skill for LLM optimization, prompt patterns, structured outputs, and AI product development. Expertise in Claude, GPT-4, prompt design patterns, few-shot learning, chain-of-thought, and AI evaluation. Includes RAG optimization, agent design, and LLM system architecture. Use when building AI products, optimizing LLM performance, designing agentic systems, or implementing advanced prompting techniques.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/senior-prompt-engineer

#broad-capability#github#externalData, AI & Research

Senior Prompt Engineer

Use when the user asks to optimize prompts, design prompt templates, evaluate LLM outputs with an eval set, measure RAG retrieval quality, validate agent/tool configurations, analyze token usage, or design structured-output contracts. Covers eval-driven prompt iteration, RAG metrics (relevance, faithfulness, coverage), agent workflow validation, and token/cost budgeting — all model-agnostic, with three stdlib Python tools.

by alirezarezvani/claude-skills / engineering-team/skills/senior-prompt-engineer

#prompt#engineeringData, AI & Research

Sentencepiece

Language-independent tokenizer treating text as raw Unicode. Supports BPE and Unigram algorithms. Fast (50k sentences/sec), lightweight (6MB memory), deterministic vocabulary. Used by T5, ALBERT, XLNet, mBART. Train on raw text without pre-tokenization. Use when you need multilingual support, CJK languages, or reproducible tokenization.

by Orchestra-Research/AI-Research-SKILLs / 02-tokenization/sentencepiece

#broad-capability#ai-research#machine-learningData, AI & Research

Sentence Transformers - State-of-the-Art Embeddings

Framework for state-of-the-art sentence, text, and image embeddings. Provides 5000+ pre-trained models for semantic similarity, clustering, and retrieval. Supports multilingual, domain-specific, and multimodal models. Use for generating embeddings for RAG, semantic search, or similarity tasks. Best for production embedding generation.

by Orchestra-Research/AI-Research-SKILLs / 15-rag/sentence-transformers

#github#broad-capability#externalData, AI & Research

SE: Responsible AI

Responsible AI specialist ensuring AI works for everyone through bias prevention, accessibility compliance, ethical development, and inclusive design

by github/awesome-copilot / agents/se-responsible-ai-code.agent.md

#github-copilot#llm#evaluationData, AI & Research

Serverless Modal

Run GPU workloads on Modal — training, fine-tuning, inference, batch processing. Zero-config serverless: no SSH, no Docker, auto scale-to-zero. Use when user says "modal run", "modal training", "modal inference", "deploy to modal", "need a GPU", "run on modal", "serverless GPU", or needs remote GPU compute.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/serverless-modal

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Serving LLMs Vllm

vLLM: high-throughput LLM serving, OpenAI API, quantization.

by NousResearch/hermes-agent / skills/mlops/inference/vllm

#broad-capability#development#creativeData, AI & Research

Session Logs

Search and analyze your own session logs (older/parent conversations) using jq.

by openclaw/openclaw / skills/session-logs

#broad-capability#browser#automationData, AI & Research

Session Start Skill

Runs the session startup procedure - verifies setup, loads config and state, checks skill models, and reports project status. Use at the beginning of a fresh session.

by bitwize-music-studio/claude-ai-music-skills / skills/session-start

#github#broad-capability#externalData, AI & Research

Sglang

Fast structured generation and serving for LLMs with RadixAttention prefix caching. Use for JSON/regex outputs, constrained decoding, agentic workflows with tool calls, or when you need 5× faster inference than vLLM with prefix sharing. Powers 300,000+ GPUs at xAI, AMD, NVIDIA, and LinkedIn.

by Orchestra-Research/AI-Research-SKILLs / 12-inference-serving/sglang

#broad-capability#ai-research#machine-learningData, AI & Research

Shap

Model interpretability and explainability using SHAP (SHapley Additive exPlanations). Use this skill when explaining machine learning model predictions, computing feature importance, generating SHAP plots (waterfall, beeswarm, bar, scatter, force, heatmap), debugging models, analyzing model bias or fairness, comparing models, or implementing explainable AI. Works with tree-based models (XGBoost, LightGBM, Random Forest), deep learning (TensorFlow, PyTorch), linear models, and any black-box model.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/shap

#broad-capability#creative#mechanisticData, AI & Research

Sherpa Onnx TTS

Local text-to-speech via sherpa-onnx (offline, no cloud)

by openclaw/openclaw / skills/sherpa-onnx-tts

#deep#researchData, AI & Research

Shuffle JSON Data

Shuffle repetitive JSON objects safely by validating schema consistency before randomising entries.

by github/awesome-copilot / skills/shuffle-json-data

#github-copilot#structured#dataData, AI & Research

Signal Postmortem

Record and analyze post-trade outcomes for signals generated by edge pipeline and other skills. Track false positives, missed opportunities, and regime mismatches. Feed results back to edge-signal-aggregator weights and skill improvement backlog.

by tradermonty/claude-trading-skills / skills/signal-postmortem

#work-life#productivity#financeData, AI & Research

Similarity Search Patterns

Implement efficient similarity search with vector databases. Use when building semantic search, implementing nearest neighbor queries, or optimizing retrieval performance.

by wshobson/agents / plugins/llm-application-dev/skills/similarity-search-patterns

#broad-capability#engineering#agent-skillsData, AI & Research

Simpo Training

Simple Preference Optimization for LLM alignment. Reference-free alternative to DPO with better performance (+6.4 points on AlpacaEval 2.0). No reference model needed, more efficient than DPO. Use for preference alignment when want simpler, faster training than DPO/PPO.

by NousResearch/hermes-agent / optional-skills/mlops/simpo

#broad-capability#development#creativeData, AI & Research

SimPy - Discrete-Event Simulation

Process-based discrete-event simulation framework in Python. Use this skill when building simulations of systems with processes, queues, resources, and time-based events such as manufacturing systems, service operations, network traffic, logistics, or any system where entities interact with shared resources over time.

by K-Dense-AI/scientific-agent-skills / scientific-skills/simpy

#github#broad-capability#externalData, AI & Research

Skill Copilot Provider

GitHub Copilot CLI as optional zero-cost provider via copilot -p programmatic mode

by nyldn/claude-octopus / skills/skill-copilot-provider

#broad-capability#agent-teams#workflowData, AI & Research

Skill Meta Prompt

Craft better prompts using proven optimization techniques — use when your prompt needs refinement

by nyldn/claude-octopus / skills/skill-meta-prompt

#broad-capability#agent-teams#workflowData, AI & Research

Skill Model Updater

Updates model references across all skill files when new Claude models are released. Use when Anthropic releases new Claude models to keep skills current.

by bitwize-music-studio/claude-ai-music-skills / skills/skill-model-updater

#github#broad-capability#externalData, AI & Research

Skill Optimizer

Optimizes AI skills for activation, clarity, and cross-model reliability. Use when creating or editing skill packs, diagnosing weak skill uptake, reducing regressions, tuning instruction salience, improving examples, shrinking context cost, or setting benchmark/release gates for skills. Trigger terms: skill optimization, activation gap, benchmark skill, with/without skill delta, regression, context budget, prompt salience.

by mcollina/skills / skills/skill-optimizer

#nodejs#fine#tuningData, AI & Research

Skin Health Analyzer

分析皮肤健康数据、识别皮肤问题模式、评估皮肤健康状况、提供个性化皮肤健康建议。支持与营养、慢性病、用药等其他健康数据的关联分析。

by huifer/WellAlly-health / .claude/skills/skin-health-analyzer

#work-life#productivity#claude-ally-healthData, AI & Research

Sleep Analyzer

分析睡眠数据、识别睡眠模式、评估睡眠质量,并提供个性化睡眠改善建议。支持与其他健康数据的关联分析。

by huifer/WellAlly-health / .claude/skills/sleep-analyzer

#work-life#productivity#claude-ally-healthData, AI & Research

slime: LLM Post-Training Framework for RL Scaling

Provides guidance for LLM post-training with RL using slime, a Megatron+SGLang framework. Use when training GLM models, implementing custom data generation workflows, or needing tight Megatron-LM integration for RL scaling.

by NousResearch/hermes-agent / optional-skills/mlops/slime

#github#broad-capability#externalData, AI & Research

Slime Rl Training

Provides guidance for LLM post-training with RL using slime, a Megatron+SGLang framework. Use when training GLM models, implementing custom data generation workflows, or needing tight Megatron-LM integration for RL scaling.

by Orchestra-Research/AI-Research-SKILLs / 06-post-training/slime

#broad-capability#ai-research#machine-learningData, AI & Research

Southwest

Search Southwest Airlines fares and points pricing via Patchright browser automation. SW is not in any GDS or API. Covers all fare classes, Companion Pass value, and fare drop monitoring.

by borski/travel-hacking-toolkit / plugins/travel-hacking-toolkit/skills/southwest

#travel#flights#hotelsData, AI & Research

Spark Engineer

Use when writing Spark jobs, debugging performance issues, or configuring cluster settings for Apache Spark applications, distributed data processing pipelines, or big data workloads. Invoke to write DataFrame transformations, optimize Spark SQL queries, implement RDD pipelines, tune shuffle operations, configure executor memory, process .parquet files, handle data partitioning, or build structured streaming analytics.

by Jeffallan/claude-skills / skills/spark-engineer

#engineering#full-stack#structuredData, AI & Research

Spark Optimization

Optimize Apache Spark jobs with partitioning, caching, shuffle optimization, and memory tuning. Use when improving Spark performance, debugging slow jobs, or scaling data processing pipelines.

by wshobson/agents / plugins/data-engineering/skills/spark-optimization

#broad-capability#engineering#agent-skillsData, AI & Research

Sparse Autoencoder Training

Provides guidance for training and analyzing Sparse Autoencoders (SAEs) using SAELens to decompose neural network activations into interpretable features. Use when discovering interpretable features, analyzing superposition, or studying monosemantic representations in language models.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/sparse-autoencoder-training

#broad-capability#creative#deepData, AI & Research

Speculative Decoding

Accelerate LLM inference using speculative decoding, Medusa multiple heads, and lookahead decoding techniques. Use when optimizing inference speed (1.5-3.6× speedup), reducing latency for real-time applications, or deploying models with limited compute. Covers draft models, tree-based attention, Jacobi iteration, parallel token generation, and production deployment strategies.

by Orchestra-Research/AI-Research-SKILLs / 19-emerging-techniques/speculative-decoding

#broad-capability#ai-research#machine-learningData, AI & Research

Stable Baselines3

Production-ready reinforcement learning algorithms (PPO, SAC, DQN, TD3, DDPG, A2C) with scikit-learn-like API. Use for standard RL experiments, quick prototyping, and well-documented algorithm implementations. Best for single-agent RL with Gymnasium environments. For high-performance parallel training, multi-agent systems, or custom vectorized environments, use pufferlib instead.

by K-Dense-AI/scientific-agent-skills / skills/stable-baselines3

#broad-capability#science#mathData, AI & Research

Stable Baselines3

Production-ready reinforcement learning algorithms (PPO, SAC, DQN, TD3, DDPG, A2C) with scikit-learn-like API. Use for standard RL experiments, quick prototyping, and well-documented algorithm implementations. Best for single-agent RL with Gymnasium environments. For high-performance parallel training, multi-agent systems, or custom vectorized environments, use pufferlib instead.

by K-Dense-AI/scientific-agent-skills / scientific-skills/stable-baselines3

#github#broad-capability#externalData, AI & Research

Stable Baselines3

Production-ready reinforcement learning algorithms (PPO, SAC, DQN, TD3, DDPG, A2C) with scikit-learn-like API. Use for standard RL experiments, quick prototyping, and well-documented algorithm implementations. Best for single-agent RL with Gymnasium environments. For high-performance parallel training, multi-agent systems, or custom vectorized environments, use pufferlib instead.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/stable-baselines3

#broad-capability#creative#rlData, AI & Research

Stable Diffusion Image Generation

State-of-the-art text-to-image generation with Stable Diffusion models via HuggingFace Diffusers. Use when generating images from text prompts, performing image-to-image translation, inpainting, or building custom diffusion pipelines.

by NousResearch/hermes-agent / optional-skills/mlops/stable-diffusion

#broad-capability#development#creativeData, AI & Research

Statistical Analysis

Statistical analysis toolkit. Hypothesis tests (t-test, ANOVA, chi-square), regression, correlation, Bayesian stats, power analysis, assumption checks, APA reporting, for academic research.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/statistical-analysis

#broad-capability#creative#deepData, AI & Research

Statistical Analysis

Guided statistical analysis with test selection and reporting. Use when you need help choosing appropriate tests for your data, assumption checking, power analysis, and APA-formatted results. Best for academic research reporting, test selection guidance. For implementing specific models programmatically use statsmodels.

by K-Dense-AI/scientific-agent-skills / skills/statistical-analysis

#k-dense-ai-claude-scientific-skills#deep#researchData, AI & Research