SKILL.md packages that extend Claude Code, Cursor, Copilot, and other AI agents.
Tags

kilo-marketplace
Guides agents to build, modify, test, and validate dbt models and SQL transformations for analytics engineering workflows.

dataframely
Guidelines and best practices for defining typed Polars schemas and collections, validating data frames, and writing tests using Dataframely.

clawdata
Run and manage dbt projects via the dbt CLI — initialise projects, run/build models, run tests, generate docs, and debug pipelines.

clawdata
Query and manage PostgreSQL databases via psql: run queries, inspect schemas and tables, check active connections, and perform basic administration and exports.

Astronomer Agents
Structured root-cause diagnosis for failed Airflow DAGs with actionable fixes, impact assessment, and prevention recommendations.

indonesia-gov-apis
Patterns and ready-to-run examples for querying 50+ Indonesian government APIs and portals (BPS, OJK, BPOM, BMKG, Bank Indonesia) with scraping and CSRF handlin

clawdata
Build and manage Dagster data pipelines -- create assets, jobs, schedules, sensors, and resources.

clawdata
Build data ingestion pipelines with dlt (data load tool) -- extract from APIs, databases, and files, then load to any destination.

fabric-skills-settings
Operational skill for managing Microsoft Fabric data platforms: run VACUUM, orchestrate DAGs, update inventory, and perform maintenance routines and GDPR-safe p

clawdata
Query and manage Google BigQuery datasets with the bq CLI: run SQL, inspect schemas, list tables, load CSV/JSON, and manage partitioning.

product-forge
Provides dbt project patterns, example models, testing snippets, and best practices for building staging→marts pipelines and tests in analytics engineering.

awesome-omni-skills
Workflow for building and validating product features driven by data insights, A/B testing, and continuous measurement; preserves upstream provenance and suppor

skillsbench
Production-ready data engineering skill for designing ETL/ELT and real-time streaming pipelines, data quality, and pipeline performance optimization.

edwinhu
Standards and enforcement guidance for querying WRDS data and running SAS/ETL on the WRDS grid—includes query validation, SGE submission patterns, and performan

claude-code-plugins-plus-skills
Automates creation and configuration of change-data-capture (CDC) data pipelines: generates configs, code snippets, validations and best-practice recommendation

claude-skill-registry
Provides dbt project patterns and examples for staging, incremental models, testing, and best practices for analytics engineering.

agent-skills-hub
Patterns and practical steps to build reliable, bias-aware backtesting systems for evaluating trading strategies and validating research hypotheses.

awesome-omni-skill
High-performance genomic data processing using polars-bio and Polars: streaming VCF/FASTA/ BED handling, interval joins, variant annotation, and Parquet convers

skills-web-dev
Design GDPR, POPIA, and CCPA compliant data export and erasure workflows for multi-tenant SaaS architectures.

feast
Comprehensive guide for managing AI/ML feature stores, including entity definition, feature retrieval, and RAG pipelines.

dbt-agent-skills
Generates professional Mermaid.js flowchart diagrams from dbt model lineage using MCP tools, manifest files, or code parsing.

awesome-omni-skill
Expert audit and review of DataHub ingestion connectors against golden standards for compliance and quality.

neuron-cli
Expert guidance for writing high-performance Spark jobs, optimizing ETL pipelines, and tuning distributed data processing.

aliyun
AI-agent-friendly interface for Hologres database operations with built-in safety guardrails.

ivanshamaev
Implement data quality checks using SodaCL for various databases including PostgreSQL, Spark, and BigQuery.

dbt-agent-skills
Retrieves and searches dbt documentation in LLM-friendly markdown format using optimized URL patterns and search scripts.

awesome-skills
A senior-data-engineer-style skill for pandas: idiomatic DataFrame transformations, ETL pipelines, performance tuning, and time-series operations in Python.

signalpilot
Guidance and rules for adding versioned dbt models (v2, latest_version, versions: YAML and ref() usage) to avoid creating standalone duplicate models.

memorybench
Run a reproducible univariate time-series forecasting pipeline using StatsForecast models and Polars, producing constrained ensemble forecasts, robust metrics,

1c-agent-based-dev-framework
Tools and workflows for discovering 1C metadata, composing and validating queries, and parsing/generating navigation links in 1C configurations.