Skip to content

Data Science & AI Research Skills

794 data science, AI research, and analysis skills for coding agents. Data pipeline, ML model dev, statistical analysis - pre-verified and MCP-ready.

Power BI DAX Expert Mode

Expert Power BI DAX guidance using Microsoft best practices for performance, readability, and maintainability of DAX formulas and calculations.

by github/awesome-copilot / agents/power-bi-dax-expert.agent.md

#github-copilot#data#visualizationData, AI & Research

Power Bi Dax Optimization

Comprehensive Power BI DAX formula optimization prompt for improving performance, readability, and maintainability of DAX calculations.

by github/awesome-copilot / skills/power-bi-dax-optimization

#github-copilot#data#visualizationData, AI & Research

Power Bi Model Design Review

Comprehensive Power BI data model design review prompt for evaluating model architecture, relationships, and optimization opportunities.

by github/awesome-copilot / skills/power-bi-model-design-review

#github-copilot#structured#dataData, AI & Research

Powerbi Modeling

Power BI semantic modeling assistant for building optimized data models. Use when working with Power BI semantic models, creating measures, designing star schemas, configuring relationships, implementing RLS, or optimizing model performance. Triggers on queries about DAX calculations, table relationships, dimension/fact table design, naming conventions, model documentation, cardinality, cross-filter direction, calculation groups, and data model best practices. Always connects to the active model first using power-bi-modeling MCP tools to understand the data structure before providing guidance.

by github/awesome-copilot / skills/powerbi-modeling

#github-copilot#structured#dataData, AI & Research

Power BI Performance Expert Mode

Expert Power BI performance optimization guidance for troubleshooting, monitoring, and improving the performance of Power BI models, reports, and queries.

by github/awesome-copilot / agents/power-bi-performance-expert.agent.md

#github-copilot#data#visualizationData, AI & Research

Power Bi Performance Troubleshooting

Systematic Power BI performance troubleshooting prompt for identifying, diagnosing, and resolving performance issues in Power BI models, reports, and queries.

by github/awesome-copilot / skills/power-bi-performance-troubleshooting

#github-copilot#ml#experimentData, AI & Research

Preset

Intelligently deploys Azure OpenAI models to optimal regions by analyzing capacity across all available regions. Automatically checks current region first and shows alternatives if needed. USE FOR: quick deployment, optimal region, best region, automatic region selection, fast setup, multi-region capacity check, high availability deployment, deploy to best location. DO NOT USE FOR: custom SKU selection (use customize), specific version selection (use customize), custom capacity configuration (use customize), PTU deployments (use customize).

by microsoft/skills / .github/plugins/azure-skills/skills/microsoft-foundry/models/deploy-model/preset

#broad-capability#development#setupData, AI & Research

Pricing Tracker

Extract and normalize pricing tiers from any SaaS, API, cloud, or LLM vendor's pricing page. Use this skill whenever the user says "pricing for X", "how much does X cost", "pricing tiers", "cost comparison", provides a URL ending in `/pricing` or `/plans`, or asks to monitor pricing over time. Pairs well with `exportSkill` to turn a run into a cron-friendly workflow. Scrape-driven; no interact needed for typical pricing pages.

by firecrawl/web-agent / agent-core/src/skills/definitions/pricing-tracker

#firecrawl-firecrawl-agent#crawling#financialData, AI & Research

PrimeKG Knowledge Graph Skill

Query the Precision Medicine Knowledge Graph (PrimeKG) for multiscale biological data including genes, drugs, diseases, phenotypes, and more.

by K-Dense-AI/scientific-agent-skills / scientific-skills/primekg

#github#broad-capability#externalData, AI & Research

Prior Art Search

Search patent databases and academic literature for prior art relevant to an invention. Use when user says "现有技术检索", "prior art search", "专利检索", "check patents", or wants to find relevant prior art.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/prior-art-search

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Promo Director Skill

Generates 15-second vertical promo videos for social media from mastered audio. Use after mastering is complete and before release, when the user wants social media content.

by bitwize-music-studio/claude-ai-music-skills / skills/promo-director

#github#broad-capability#externalData, AI & Research

Prompt Engineer

A specialized chat mode for analyzing and improving prompts. Every user input is treated as a prompt to be improved. It first provides a detailed analysis of the original prompt within a <reasoning> tag, evaluating it against a systematic framework based on OpenAI's prompt engineering best practices. Following the analysis, it generates a new, improved prompt.

by github/awesome-copilot / agents/prompt-engineer.agent.md

#github-copilot#prompt#engineeringData, AI & Research

Prompt Engineer

Writes, refactors, and evaluates prompts for LLMs — generating optimized prompt templates, structured output schemas, evaluation rubrics, and test suites. Use when designing prompts for new LLM applications, refactoring existing prompts for better accuracy or token efficiency, implementing chain-of-thought or few-shot learning, creating system prompts with personas and guardrails, building JSON/function-calling schemas, or developing prompt evaluation frameworks to measure and improve model performance.

by Jeffallan/claude-skills / skills/prompt-engineer

#engineering#full-stack#promptData, AI & Research

Prompt Engineering Patterns

Master advanced prompt engineering techniques to maximize LLM performance, reliability, and controllability in production. Use when optimizing prompts, improving LLM outputs, or designing production prompt templates.

by wshobson/agents / plugins/llm-application-dev/skills/prompt-engineering-patterns

#github#broad-capability#externalData, AI & Research

Promptfoo Evaluation

Configures and runs LLM evaluation using Promptfoo framework. Use when setting up prompt testing, creating evaluation configs (promptfooconfig.yaml), writing Python custom assertions, implementing llm-rubric for LLM-as-judge, or managing few-shot examples in prompts. Triggers on keywords like "promptfoo", "eval", "LLM evaluation", "prompt testing", or "model comparison".

by daymade/claude-code-skills / promptfoo-evaluation

#broad-capability#research#documentsData, AI & Research

Prompt Images

Prompting techniques for AI image generation and editing models on Replicate. Use when writing prompts for image models or building image generation features.

by replicate/skills / skills/prompt-images

#ml-inference#prompt#engineeringData, AI & Research

Prompt Optimizer

Creates, optimizes, and iteratively refines agent prompts, system prompts, developer prompts, and reusable prompt templates. Use when asked to improve a prompt, optimize a system prompt, rewrite an agent prompt, tune prompt wording, make a prompt more reliable, port prompts between OpenAI, Claude, or Gemini, or build prompt evals.

by getsentry/skills / skills/prompt-optimizer

#observability#sentry#promptData, AI & Research

Prompt Videos

Prompting techniques for AI video generation models on Replicate. Use when writing prompts for video models or building video generation features.

by replicate/skills / skills/prompt-videos

#ml-inference#prompt#engineeringData, AI & Research

Proof Checker

Rigorous mathematical proof verification and fixing workflow. Reads a LaTeX proof, identifies gaps via cross-model review (Codex GPT-5.5 xhigh), fixes each gap with full derivations, re-reviews, and generates an audit report. Use when user says "检查证明", "verify proof", "proof check", "审证明", "check this proof", or wants rigorous mathematical verification of a theory paper.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/skills-codex/proof-checker

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Proof Writer

Writes rigorous mathematical proofs for ML/AI theory. Use when asked to prove a theorem, lemma, proposition, or corollary, fill in missing proof steps, formalize a proof sketch, 补全证明, 写证明, 证明某个命题, or determine whether a claimed proof can actually be completed under the stated assumptions.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/proof-writer

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Protocols.io Integration

Integration with protocols.io API for managing scientific protocols. This skill should be used when working with protocols.io to search, create, update, or publish protocols; manage protocol steps and materials; handle discussions and comments; organize workspaces; upload and manage files; or integrate protocols.io functionality into workflows. Applicable for protocol discovery, collaborative protocol development, experiment tracking, lab protocol management, and scientific documentation.

by K-Dense-AI/scientific-agent-skills / scientific-skills/protocolsio-integration

#github#external#license-mitData, AI & Research

Pubmed Database

Direct REST API access to PubMed. Advanced Boolean/MeSH queries, E-utilities API, batch processing, citation management. For Python workflows, prefer biopython (Bio.Entrez). Use this for direct HTTP/REST work or custom API implementations.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/pubmed-database

#broad-capability#creative#literatureData, AI & Research

Pufferlib

High-performance reinforcement learning framework optimized for speed and scale. Use when you need fast parallel training, vectorized environments, multi-agent systems, or integration with game environments (Atari, Procgen, NetHack). Achieves 2-10x speedups over standard implementations. For quick prototyping or standard algorithm implementations with extensive documentation, use stable-baselines3 instead.

by K-Dense-AI/scientific-agent-skills / skills/pufferlib

#broad-capability#science#mathData, AI & Research

PufferLib - High-Performance Reinforcement Learning

This skill should be used when working with reinforcement learning tasks including high-performance RL training, custom environment development, vectorized parallel simulation, multi-agent systems, or integration with existing RL environments (Gymnasium, PettingZoo, Atari, Procgen, etc.). Use this skill for implementing PPO training, creating PufferEnv environments, optimizing RL performance, or developing policies with CNNs/LSTMs.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/pufferlib

#broad-capability#github#externalData, AI & Research

Pydeseq2

Differential gene expression analysis for bulk RNA-seq with PyDESeq2, including formulaic designs, Wald tests, FDR correction, LFC shrinkage, and result visualization.

by K-Dense-AI/scientific-agent-skills / skills/pydeseq2

#broad-capability#science#mathData, AI & Research

Pydeseq2

Differential gene expression analysis (Python DESeq2). Identify DE genes from bulk RNA-seq counts, Wald tests, FDR correction, volcano/MA plots, for RNA-seq analysis.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/pydeseq2

#broad-capability#creative#scientificData, AI & Research

Pydicom

Python library for working with DICOM (Digital Imaging and Communications in Medicine) files. Use this skill when reading, writing, or modifying medical imaging data in DICOM format, extracting pixel data from medical images (CT, MRI, X-ray, ultrasound), anonymizing DICOM files, working with DICOM metadata and tags, converting DICOM images to other formats, handling compressed DICOM data, or processing medical imaging datasets. Applies to tasks involving medical image analysis, PACS systems, radiology workflows, and healthcare imaging applications.

by K-Dense-AI/scientific-agent-skills / scientific-skills/pydicom

#github#broad-capability#externalData, AI & Research

Pydicom

Python library for working with DICOM (Digital Imaging and Communications in Medicine) files. Use this skill when reading, writing, or modifying medical imaging data in DICOM format, extracting pixel data from medical images (CT, MRI, X-ray, ultrasound), anonymizing DICOM files, working with DICOM metadata and tags, converting DICOM images to other formats, handling compressed DICOM data, or processing medical imaging datasets. Applies to tasks involving medical image analysis, PACS systems, radiology workflows, and healthcare imaging applications.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/pydicom

#broad-capability#creative#scientificData, AI & Research

Pyhealth

Build clinical/healthcare deep-learning pipelines with PyHealth — loading EHR/signal/imaging datasets (MIMIC-III/IV, eICU, OMOP, SleepEDF, ChestXray14, EHRShot), defining tasks (mortality, readmission, length-of-stay, drug recommendation, sleep staging, ICD coding, EEG events), instantiating models (Transformer, RETAIN, GAMENet, SafeDrug, MICRON, StageNet, AdaCare, CNN/RNN/MLP), training with the PyHealth Trainer, computing clinical metrics, and using medical code utilities (ICD/ATC/NDC/RxNorm lookup and cross-mapping). Use this skill whenever the user mentions PyHealth, MIMIC, eICU, OMOP, EHR modeling, clinical prediction, drug recommendation, sleep staging, medical code mapping, ICD/ATC codes, or any healthcare ML pipeline that fits the dataset → task → model → trainer → metrics pattern, even if "PyHealth" isn't named explicitly.

by K-Dense-AI/scientific-agent-skills / skills/pyhealth

#k-dense-ai-claude-scientific-skills#clinical#deepData, AI & Research

PyHealth

Build clinical/healthcare deep-learning pipelines with PyHealth — loading EHR/signal/imaging datasets (MIMIC-III/IV, eICU, OMOP, SleepEDF, ChestXray14, EHRShot), defining tasks (mortality, readmission, length-of-stay, drug recommendation, sleep staging, ICD coding, EEG events), instantiating models (Transformer, RETAIN, GAMENet, SafeDrug, MICRON, StageNet, AdaCare, CNN/RNN/MLP), training with the PyHealth Trainer, computing clinical metrics, and using medical code utilities (ICD/ATC/NDC/RxNorm lookup and cross-mapping). Use this skill whenever the user mentions PyHealth, MIMIC, eICU, OMOP, EHR modeling, clinical prediction, drug recommendation, sleep staging, medical code mapping, ICD/ATC codes, or any healthcare ML pipeline that fits the dataset → task → model → trainer → metrics pattern, even if "PyHealth" isn't named explicitly.

by K-Dense-AI/scientific-agent-skills / scientific-skills/pyhealth

#github#external#license-mitData, AI & Research

PyHealth: Healthcare AI Toolkit

Comprehensive healthcare AI toolkit for developing, testing, and deploying machine learning models with clinical data. This skill should be used when working with electronic health records (EHR), clinical prediction tasks (mortality, readmission, drug recommendation), medical coding systems (ICD, NDC, ATC), physiological signals (EEG, ECG), healthcare datasets (MIMIC-III/IV, eICU, OMOP), or implementing deep learning models for healthcare applications (RETAIN, SafeDrug, Transformer, GNN).

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/pyhealth

#broad-capability#github#externalData, AI & Research

PyLabRobot

Vendor-agnostic lab automation framework. Use when controlling multiple equipment types (Hamilton, Tecan, Opentrons, plate readers, pumps) or needing unified programming across different vendors. Best for complex workflows, multi-vendor setups, simulation. For Opentrons-only protocols with official API, opentrons-integration may be simpler.

by K-Dense-AI/scientific-agent-skills / scientific-skills/pylabrobot

#github#broad-capability#externalData, AI & Research

Pymc

Bayesian modeling with PyMC. Build hierarchical models, MCMC (NUTS), variational inference, LOO/WAIC comparison, posterior checks, for probabilistic programming and inference.

by K-Dense-AI/scientific-agent-skills / skills/pymc

#k-dense-ai-claude-scientific-skills#deep#researchData, AI & Research

Pymc Bayesian Modeling

Bayesian modeling with PyMC. Build hierarchical models, MCMC (NUTS), variational inference, LOO/WAIC comparison, posterior checks, for probabilistic programming and inference.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/pymc

#broad-capability#creative#deepData, AI & Research

PyMC Bayesian Modeling

Bayesian modeling with PyMC. Build hierarchical models, MCMC (NUTS), variational inference, LOO/WAIC comparison, posterior checks, for probabilistic programming and inference.

by K-Dense-AI/scientific-agent-skills / scientific-skills/pymc

#github#external#license-mitData, AI & Research

pymc-bayesian-modeling (Compatibility Alias)

Compatibility alias for the descriptive PyMC skill name. Delegate to the canonical local `pymc` payload while preserving route and README compatibility.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/pymc-bayesian-modeling

#broad-capability#github#externalData, AI & Research

Pymoo - Multi-Objective Optimization in Python

Multi-objective optimization framework. NSGA-II, NSGA-III, MOEA/D, Pareto fronts, constraint handling, benchmarks (ZDT, DTLZ), for engineering design and optimization problems.

by K-Dense-AI/scientific-agent-skills / scientific-skills/pymoo

#github#broad-capability#externalData, AI & Research

PyOpenMS

Complete mass spectrometry analysis platform. Use for proteomics workflows feature detection, peptide identification, protein quantification, and complex LC-MS/MS pipelines. Supports extensive file formats and algorithms. Best for proteomics, comprehensive MS data processing. For simple spectral comparison and metabolite ID use matchms.

by K-Dense-AI/scientific-agent-skills / scientific-skills/pyopenms

#github#broad-capability#externalData, AI & Research

Pysam

Genomic file toolkit. Read/write SAM/BAM/CRAM alignments, VCF/BCF variants, FASTA/FASTQ sequences, extract regions, calculate coverage, for NGS data processing pipelines.

by K-Dense-AI/scientific-agent-skills / scientific-skills/pysam

#github#broad-capability#externalData, AI & Research

PySpark Expert Agent

Diagnose PySpark performance bottlenecks, distributed execution pitfalls, and suggest Spark-native rewrites and safer distributed patterns (incl. mapInPandas guidance).

by github/awesome-copilot / agents/spark-performance.agent.md

#github-copilot#distributed#trainingData, AI & Research

Pytdc

Therapeutics Data Commons. AI-ready drug discovery datasets (ADME, toxicity, DTI), benchmarks, scaffold splits, molecular oracles, for therapeutic ML and pharmacological prediction.

by K-Dense-AI/scientific-agent-skills / skills/pytdc

#k-dense-ai-claude-scientific-skills#drug#discoveryData, AI & Research

PyTDC (Therapeutics Data Commons)

Therapeutics Data Commons. AI-ready drug discovery datasets (ADME, toxicity, DTI), benchmarks, scaffold splits, molecular oracles, for therapeutic ML and pharmacological prediction.

by K-Dense-AI/scientific-agent-skills / scientific-skills/pytdc

#github#external#license-mitData, AI & Research

Python MCP Server Development

Instructions for building Model Context Protocol (MCP) servers using the Python SDK

by github/awesome-copilot / instructions/python-mcp-server.instructions.md

#github-copilot#fine#tuningData, AI & Research

Python Notebook Sample Builder

Custom agent for building Python Notebooks in VS Code that demonstrate Azure and AI features

by github/awesome-copilot / agents/python-notebook-sample-builder.agent.md

#github-copilot#ml#experimentData, AI & Research

Pytorch Fsdp2

Adds PyTorch FSDP2 (fully_shard) to training scripts with correct init, sharding, mixed precision/offload config, and distributed checkpointing. Use when models exceed single-GPU memory or when you need DTensor-based sharding with DeviceMesh.

by Orchestra-Research/AI-Research-SKILLs / 08-distributed-training/pytorch-fsdp2

#broad-capability#ai-research#machine-learningData, AI & Research

Pytorch-Fsdp Skill

Expert guidance for Fully Sharded Data Parallel training with PyTorch FSDP - parameter sharding, mixed precision, CPU offloading, FSDP2

by NousResearch/hermes-agent / optional-skills/mlops/pytorch-fsdp

#github#broad-capability#externalData, AI & Research

PyTorch Geometric (PyG)

Guide for building Graph Neural Networks with PyTorch Geometric (PyG). Use this skill whenever the user asks about graph neural networks, GNNs, node classification, link prediction, graph classification, message passing networks, heterogeneous graphs, neighbor sampling, or any task involving torch_geometric / PyG. Also trigger when you see imports from torch_geometric, or the user mentions graph convolutions (GCN, GAT, GraphSAGE, GIN), graph data structures, or working with relational/network data. Even if the user just says 'graph learning' or 'geometric deep learning', use this skill.

by K-Dense-AI/scientific-agent-skills / scientific-skills/torch-geometric

#github#broad-capability#externalData, AI & Research

PyTorch Lightning

Deep learning framework (PyTorch Lightning). Organize PyTorch code into LightningModules, configure Trainers for multi-GPU/TPU, implement data pipelines, callbacks, logging (W&B, TensorBoard), distributed training (DDP, FSDP, DeepSpeed), for scalable neural network training.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/pytorch-lightning

#broad-capability#github#externalData, AI & Research

PyTorch Lightning - High-Level Training Framework

High-level PyTorch framework with Trainer class, automatic distributed training (DDP/FSDP/DeepSpeed), callbacks system, and minimal boilerplate. Scales from laptop to supercomputer with same code. Use when you want clean training loops with built-in best practices.

by NousResearch/hermes-agent / optional-skills/mlops/pytorch-lightning

#github#broad-capability#externalData, AI & Research

Pyvene Interventions

Provides guidance for performing causal interventions on PyTorch models using pyvene's declarative intervention framework. Use when conducting causal tracing, activation patching, interchange intervention training, or testing causal hypotheses about model behavior.

by Orchestra-Research/AI-Research-SKILLs / 04-mechanistic-interpretability/pyvene

#broad-capability#ai-research#machine-learningData, AI & Research

Pyzotero

Interact with Zotero reference management libraries using the pyzotero Python client. Retrieve, create, update, and delete items, collections, tags, and attachments via the Zotero Web API v3. Use this skill when working with Zotero libraries programmatically, managing bibliographic references, exporting citations, searching library contents, uploading PDF attachments, or building research automation workflows that integrate with Zotero.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/pyzotero

#broad-capability#creative#literatureData, AI & Research

Pyzotero

Interact with Zotero reference management libraries using the pyzotero Python client. Retrieve, create, update, and delete items, collections, tags, and attachments via the Zotero Web API v3. Use this skill when working with Zotero libraries programmatically, managing bibliographic references, exporting citations, searching library contents, uploading PDF attachments, or building research automation workflows that integrate with Zotero.

by K-Dense-AI/scientific-agent-skills / skills/pyzotero

#broad-capability#science#mathData, AI & Research

Qdrant Hybrid Search

Explains hybrid search in Qdrant. Use when someone asks 'how do I setup hybrid search?', 'how to combine keyword and semantic search?', 'sparse plus dense vectors?', 'missing keyword matches', 'how to combine results from multiple searches?' and 'combining multiple representations'

by qdrant/skills / skills/qdrant-search-quality/search-strategies/hybrid-search

#vector-db#hybrid#searchData, AI & Research

Qdrant Hybrid Search Combining

Use when someone asks 'RRF or DBSF?', 'how to combine sparse and dense', 'how to combine scores from multiple searches?', 'custom fusion', or 'fusion is not producing good results'

by qdrant/skills / skills/qdrant-search-quality/search-strategies/hybrid-search/combining-searches

#vector-db#web#searchData, AI & Research

Qdrant Hybrid Search Prefetches

Use when someone asks 'how to combine lexical and semantic retrieval', 'dense and sparse in one search?', 'how to combine multiple fields for retrieval?', 'payloads or sparse vectors for lexical?', 'which sparse embedding model to use?', 'BM25 vs SPLADE?'

by qdrant/skills / skills/qdrant-search-quality/search-strategies/hybrid-search/search-types

#vector-db#hybrid#searchData, AI & Research

Qdrant Model Migration

Guides embedding model migration in Qdrant without downtime. Use when someone asks 'how to switch embedding models', 'how to migrate vectors', 'how to update to a new model', 'zero-downtime model change', 'how to re-embed my data', or 'can I use two models at once'. Also use when upgrading model dimensions, switching providers, or A/B testing models.

by github/awesome-copilot / skills/qdrant-model-migration

#github-copilot#fine#tuningData, AI & Research

Qdrant Model Migration

Guides embedding model migration in Qdrant without downtime. Use when someone asks 'how to switch embedding models', 'how to migrate vectors', 'how to update to a new model', 'zero-downtime model change', 'how to re-embed my data', or 'can I use two models at once'. Also use when upgrading model dimensions, switching providers, or A/B testing models.

by qdrant/skills / skills/qdrant-model-migration

#vector-db#fine#tuningData, AI & Research

Qdrant Scaling Qps

Guides Qdrant query throughput (QPS) scaling. Use when someone asks 'how to increase QPS', 'need more throughput', 'queries per second too low', 'batch search', 'read replicas', or 'how to handle more concurrent queries'.

by qdrant/skills / skills/qdrant-scaling/scaling-qps

#vector-db#distributed#trainingData, AI & Research

Qdrant Scaling Query Volume

Guides Qdrant query volume scaling. Use when someone asks 'query returns too many results', 'scroll performance', 'large limit values', 'paginating search results', 'fetching many vectors', or 'high cardinality results'.

by qdrant/skills / skills/qdrant-scaling/scaling-query-volume

#vector-db#distributed#trainingData, AI & Research

Qdrant Search Quality

Diagnoses and improves Qdrant search relevance. Use when someone reports 'search results are bad', 'wrong results', 'low precision', 'low recall', 'irrelevant matches', 'missing expected results', or asks 'how to improve search quality?', 'which embedding model?', 'should I use hybrid search?', 'should I use reranking?', 'how to measure retrieval quality?', 'build a golden set', 'ground truth dataset', or 'how to score recall@k?'. Also use when search quality degrades after quantization, model change, or data growth.

by qdrant/skills / skills/qdrant-search-quality

#vector-db#web#searchData, AI & Research

Qdrant Search Quality

Diagnoses and improves Qdrant search relevance. Use when someone reports 'search results are bad', 'wrong results', 'low precision', 'low recall', 'irrelevant matches', 'missing expected results', or asks 'how to improve search quality?', 'which embedding model?', 'should I use hybrid search?', 'should I use reranking?'. Also use when search quality degrades after quantization, model change, or data growth.

by github/awesome-copilot / skills/qdrant-search-quality

#github-copilot#web#searchData, AI & Research

Qdrant Search Quality Diagnosis

Diagnoses Qdrant search quality issues. Use when someone reports 'results are bad', 'wrong results', 'not relevant results', 'missing matches', 'recall is low', 'approximate search worse than exact', 'which embedding model', 'quality dropped after quantization', 'how to measure retrieval quality', 'build a golden set', 'ground truth dataset', or 'how to score recall@k'. Also use when search quality degrades without obvious changes.

by qdrant/skills / skills/qdrant-search-quality/diagnosis

#vector-db#web#searchData, AI & Research

Qdrant Search Strategies

Guides Qdrant search strategy selection. Use when someone asks 'should I use hybrid search?', 'how to rerank?', 'results are not relevant', 'I don't get needed results from my dataset but they're there', 'retrieval quality is not good enough', 'results too similar', 'need diversity', 'MMR', 'relevance feedback', 'recommendation API', 'discovery API', or 'missing keyword matches'

by qdrant/skills / skills/qdrant-search-quality/search-strategies

#vector-db#web#searchData, AI & Research

Qdrant Vector Search

High-performance vector similarity search engine for RAG and semantic search. Use when building production RAG systems requiring fast nearest neighbor search, hybrid search with filtering, or scalable vector storage with Rust-powered performance.

by Orchestra-Research/AI-Research-SKILLs / 15-rag/qdrant

#broad-capability#ai-research#machine-learningData, AI & Research

Qdrant - Vector Similarity Search Engine

High-performance vector similarity search engine for RAG and semantic search. Use when building production RAG systems requiring fast nearest neighbor search, hybrid search with filtering, or scalable vector storage with Rust-powered performance.

by NousResearch/hermes-agent / optional-skills/mlops/qdrant

#github#broad-capability#externalData, AI & Research

Qdrant Version Upgrade

Guidance on how to upgrade your Qdrant version without interrupting the availability of your application and ensuring data integrity.

by github/awesome-copilot / skills/qdrant-version-upgrade

#github-copilot#web#searchData, AI & Research

Qdrant Vertical Scaling

Guides Qdrant vertical scaling decisions. Use when someone asks 'how to scale up a node', 'need more RAM', 'upgrade node size', 'vertical scaling', 'resize cluster', 'scale up vs scale out', or when memory/CPU is insufficient on current nodes. Also use when someone wants to avoid the complexity of horizontal scaling.

by qdrant/skills / skills/qdrant-scaling/scaling-data-volume/vertical-scaling

#vector-db#distributed#trainingData, AI & Research

Qiskit

IBM quantum computing framework. Use when targeting IBM Quantum hardware, working with Qiskit Runtime for production workloads, or needing IBM optimization tools. Best for IBM hardware execution, quantum error mitigation, and enterprise quantum computing. For Google hardware use cirq; for gradient-based quantum ML use pennylane; for open quantum system simulations use qutip.

by K-Dense-AI/scientific-agent-skills / scientific-skills/qiskit

#github#broad-capability#externalData, AI & Research

Quantizing Models Bitsandbytes

Quantizes LLMs to 8-bit or 4-bit for 50-75% memory reduction with minimal accuracy loss. Use when GPU memory is limited, need to fit larger models, or want faster inference. Supports INT8, NF4, FP4 formats, QLoRA training, and 8-bit optimizers. Works with HuggingFace Transformers.

by Orchestra-Research/AI-Research-SKILLs / 10-optimization/bitsandbytes

#broad-capability#ai-research#machine-learningData, AI & Research

Query

Use when the user wants to query or analyze data through the Honeydew semantic layer — including natural language analysis questions, deep multi-step investigations, and structured queries. For model/field discovery use the model-exploration skill.

by honeydew-ai/honeydew-ai-coding-agents-plugins / skills/query

#honeydew-ai-plugins#coding-agents#structuredData, AI & Research

QuTiP: Quantum Toolbox in Python

Quantum physics simulation library for open quantum systems. Use when studying master equations, Lindblad dynamics, decoherence, quantum optics, or cavity QED. Best for physics research, open system dynamics, and educational simulations. NOT for circuit-based quantum computing—use qiskit, cirq, or pennylane for quantum algorithms and hardware execution.

by K-Dense-AI/scientific-agent-skills / scientific-skills/qutip

#github#broad-capability#externalData, AI & Research

Qzcli

Manage GPU compute jobs on the Qizhi (启智) platform using qzcli — a kubectl-style CLI tool. Use when user says "qzcli", "启智平台", "submit job", "stop job", "查计算组", "avail", "list jobs", "batch submit", or needs to manage distributed training jobs on a Qizhi instance.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/qzcli

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

RAG Architect

Designs and implements production-grade RAG systems by chunking documents, generating embeddings, configuring vector stores, building hybrid search pipelines, applying reranking, and evaluating retrieval quality. Use when building RAG systems, vector databases, or knowledge-grounded AI applications requiring semantic search, document retrieval, context augmentation, similarity search, or embedding-based indexing.

by Jeffallan/claude-skills / skills/rag-architect

#engineering#full-stack#promptData, AI & Research

RAG Engineer

Expert in building Retrieval-Augmented Generation systems. Masters embedding models, vector databases, chunking strategies, and retrieval optimization for LLM applications.

by sickn33/antigravity-awesome-skills / plugins/antigravity-awesome-skills-claude/skills/rag-engineer

#prompt#engineeringData, AI & Research

RAG Implementation

Build Retrieval-Augmented Generation (RAG) systems for LLM applications with vector databases and semantic search. Use when implementing knowledge-grounded AI, building document Q&A systems, or integrating LLMs with external knowledge bases.

by wshobson/agents / plugins/llm-application-dev/skills/rag-implementation

#github#broad-capability#externalData, AI & Research

ralph-loop

Codex-compatible Ralph loop runner with dual engines (compat local state loop + optional open-ralph-wiggum backend).

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/ralph-loop

#broad-capability#github#externalData, AI & Research

Ray Data

Scalable data processing for ML workloads. Streaming execution across CPU/GPU, supports Parquet/CSV/JSON/images. Integrates with Ray Train, PyTorch, TensorFlow. Scales from single machine to 100s of nodes. Use for batch inference, data preprocessing, multi-modal data loading, or distributed ETL pipelines.

by Orchestra-Research/AI-Research-SKILLs / 05-data-processing/ray-data

#broad-capability#ai-research#machine-learningData, AI & Research

Ray Train

Distributed training orchestration across clusters. Scales PyTorch/TensorFlow/HuggingFace from laptop to 1000s of nodes. Built-in hyperparameter tuning with Ray Tune, fault tolerance, elastic scaling. Use when training massive models across multiple machines or running distributed hyperparameter sweeps.

by Orchestra-Research/AI-Research-SKILLs / 08-distributed-training/ray-train

#broad-capability#ai-research#machine-learningData, AI & Research

RDKit Cheminformatics Toolkit

Cheminformatics toolkit for fine-grained molecular control. SMILES/SDF parsing, descriptors (MW, LogP, TPSA), fingerprints, substructure search, 2D/3D generation, similarity, reactions. For standard workflows with simpler interface, use datamol (wrapper around RDKit). Use rdkit for advanced control, custom sanitization, specialized algorithms.

by K-Dense-AI/scientific-agent-skills / scientific-skills/rdkit

#github#broad-capability#externalData, AI & Research

Rebuttal

Workflow 4: Submission rebuttal pipeline. Parses external reviews, enforces coverage and grounding, drafts a safe text-only rebuttal under venue limits, and manages follow-up rounds. Use when user says "rebuttal", "reply to reviewers", "ICML rebuttal", "OpenReview response", or wants to answer external reviews safely.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/rebuttal

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Rebuttal

Workflow 4: Submission rebuttal pipeline. Parses external reviews, enforces coverage and grounding, drafts a safe text-only rebuttal under venue limits, and manages follow-up rounds. Use when user says "rebuttal", "reply to reviewers", "ICML rebuttal", "OpenReview response", or wants to answer external reviews safely.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/skills-codex/rebuttal

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Reddit API

Reddit API with PRAW (Python) and Snoowrap (Node.js)

by alinaqi/maggy / skills/reddit-api

#claude-bootstrap#bootstrap#webData, AI & Research

Relation Creation

Guides you through defining a relationship between two Honeydew entities — covering join type, direction, cross-filtering, and connection method — then pushes the updated entity YAML to Honeydew via the MCP tools.

by honeydew-ai/honeydew-ai-coding-agents-plugins / skills/relation-creation

#honeydew-ai-plugins#coding-agents#structuredData, AI & Research

Research

Deep research on a topic, creating persistent documentation for future reference. Use for technology decisions, competitive analysis, or complex topics.

by TaylorHuston/local-life-manager / .claude/skills/research

#personal-productivity#daily-routine#weekly-reviewData, AI & Research

Researcher

Conducts investigative-grade research with primary source analysis, cross-verification, and trial-level depth. Use when an album needs factual research, source material, or verification of claims.

by bitwize-music-studio/claude-ai-music-skills / skills/researcher

#broad-capability#music#audio-generationData, AI & Research

Researchers Biographical

Researches personal backgrounds, interviews, motivations, and humanizing details. Use when research needs biographical context about people involved in the album's subject.

by bitwize-music-studio/claude-ai-music-skills / skills/researchers-biographical

#broad-capability#music#audio-generationData, AI & Research

Researchers Financial

Researches SEC filings, earnings calls, analyst reports, and market data. Use when the album subject involves financial crimes, corporate stories, or market events.

by bitwize-music-studio/claude-ai-music-skills / skills/researchers-financial

#broad-capability#music#audio-generationData, AI & Research

Researchers Gov

Researches DOJ/FBI/SEC press releases, agency statements, and government sources. Use when research needs official government records or agency documentation.

by bitwize-music-studio/claude-ai-music-skills / skills/researchers-gov

#broad-capability#music#audio-generationData, AI & Research

Researchers Historical

Researches archives, contemporary accounts, and timeline reconstruction. Use when the album subject involves historical events that need primary source verification.

by bitwize-music-studio/claude-ai-music-skills / skills/researchers-historical

#broad-capability#music#audio-generationData, AI & Research

Researchers Journalism

Researches investigative articles, interviews, and news coverage. Use when research needs journalistic sources for cross-referencing or additional context.

by bitwize-music-studio/claude-ai-music-skills / skills/researchers-journalism

#broad-capability#music#audio-generationData, AI & Research

Researchers Legal

Researches court documents, indictments, plea agreements, and sentencing records. Use when the album subject involves legal proceedings or criminal cases.

by bitwize-music-studio/claude-ai-music-skills / skills/researchers-legal

#broad-capability#music#audio-generationData, AI & Research

Researchers Primary Source

Researches the subject's own words from tweets, blogs, forums, and chat logs. Use when research needs direct quotes or first-person accounts.

by bitwize-music-studio/claude-ai-music-skills / skills/researchers-primary-source

#broad-capability#music#audio-generationData, AI & Research

Researchers Verifier

Performs quality control, citation validation, and fact-checking before human review. Use after research is complete to verify all sources and claims before production.

by bitwize-music-studio/claude-ai-music-skills / skills/researchers-verifier

#broad-capability#music#audio-generationData, AI & Research

Research Information Lookup

Look up current research information using parallel-cli search (primary, fast web search), the Parallel Chat API (deep research), or Perplexity sonar-pro-search (academic paper searches). Automatically routes queries to the best backend. Use for finding papers, gathering research data, and verifying scientific information. Note: query text is transmitted to api.parallel.ai (PARALLEL_API_KEY) and, for academic searches, to openrouter.ai (OPENROUTER_API_KEY).

by K-Dense-AI/scientific-agent-skills / scientific-skills/research-lookup

#github#external#license-mitData, AI & Research

Research Information Lookup

Look up current research information using Perplexity's Sonar Pro Search or Sonar Reasoning Pro models through OpenRouter. Automatically selects the best model based on query complexity. Search academic papers, recent studies, technical documentation, and general research information with citations.

by foryourhealth111-pixel/Vibe-Skills / bundled/skills/research-lookup

#broad-capability#github#externalData, AI & Research

Researching Web

Search the web using Perplexity AI. Use when needing to search, look up, research, find current information, best practices, compare technologies, or answer factual questions about tools and libraries.

by julianobarbosa/claude-code-skills / skills/researching-web

#broad-capability#devops#azureData, AI & Research

Research Lit

Search and analyze research papers, find related work, summarize key ideas. Use when user says "find papers", "related work", "literature review", "what does this paper say", or needs to understand academic papers.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/skills-codex/research-lit

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Research Lit

Search and analyze research papers, find related work, summarize key ideas. Use when user says "find papers", "related work", "literature review", "what does this paper say", or needs to understand academic papers.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/research-lit

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research

Research Paper Writing

Write ML papers for NeurIPS/ICML/ICLR: design→submit.

by NousResearch/hermes-agent / skills/research/research-paper-writing

#work-life#productivity#personal-productivityData, AI & Research

Research Pipeline

Full end-to-end research pipeline: from a broad research direction through idea discovery, experiments, and review all the way to a polished paper PDF. Use when user says "全流程", "full pipeline", "从找idea到投稿", "end-to-end research", or wants the complete autonomous research lifecycle.

by wanshuiyin/Auto-claude-code-research-in-sleep / skills/research-pipeline

#broad-capability#wanshuiyin-aris#ml-researchData, AI & Research