Create Databricks Python notebooks, push to workspace, run on cluster, and verify outputs using dbx.py.
Develop Lakeflow Spark Declarative Pipelines (formerly Delta Live Tables) on Databricks. Use when building batch or streaming data pipelines with Python or SQL.
Execute Databricks production deployment checklist and rollback procedures. Use when deploying Databricks jobs to production, preparing for launch, or implementing go-live…
Databricks development guidance including Python SDK, Databricks Connect, CLI, and REST API. Use when working with databricks-sdk, databricks-connect, or Databricks APIs.
Implement Databricks API rate limiting, backoff, and idempotency patterns. Use when handling rate limit errors, implementing retry logic, or optimizing API request throughput for…
Implement Databricks reference architecture with best-practice project layout. Use when designing new Databricks projects, reviewing architecture, or establishing standards for…
Senior Data Engineer — designs data models, writes optimized SQL, builds ETL/ELT pipelines, manages data warehouse architecture. Treats SQL as a first-class language.
Use when developing BigQuery Dataform transformations, SQLX files, source declarations, or troubleshooting pipelines - enforces TDD workflow (tests first), ALWAYS use ${ref()}…
データベースクエリ・分析支援。SQLクエリの作成、実行、結果の分析を行う。BigQuery、PostgreSQL、MySQL対応。トリガー: /db-query, SQL, クエリ, データ分析, BigQuery
DBA Deutschland Frankreich aus 1959 mit Änderungsprotokollen. Anwendungsfall Pendler im Elsass und Lothringen Grenzgaengerregelung 20-km-Zone. Beteiligungen Pensionen Lizenzen.
dbt Core/Cloud data transformations, testing, documentation, and CI/CD. Activate on: dbt, data transformation, analytics engineering, ref, source, staging model, mart, dbt test.
dbt (data build tool) patterns for model organization, incremental strategies, and testing.
Parses dbt project artifacts (manifest.json and catalog.json) to build a lineage graph and identify models with no tests, stale documentation, or missing uniqueness assertions.
Comprehensive guide to dbt (data build tool) patterns, modeling best practices, testing strategies, and production workflows for modern data transformation
ALWAYS USE when working with dbt models, SQL transformations, tests, snapshots, or macros. Use IMMEDIATELY when editing dbt_project.yml, profiles.yml, or creating SQL models.
Use when creating or modifying dimensional dbt models in warehouse-backed analytics projects. Covers a four-layer warehouse architecture (sources/staging/core/marts), naming…
Dbt Test Creator - Auto-activating skill for Data Pipelines. Triggers on: dbt test creator, dbt test creator Part of the Data Pipelines skill category.
dbt testing strategies using dbt_constraints for database-level enforcement, generic tests, and
Production-ready patterns for dbt (data build tool) including model organization, testing strategies, documentation, and incremental processing.
Master dbt (data build tool) for analytics engineering with model organization, testing, documentation, and incremental strategies.
Master dbt (data build tool) for analytics engineering with model organization, testing, documentation, and incremental strategies.
Write dbt unit-test specs with given/when/expect shape. Seeded from dbt-labs/dbt-agent-skills (Feb 9 2026).
Designing Data-Intensive Applications (DDIA) distilled reference guide by Martin Kleppmann. MUST be loaded when: designing database schemas, choosing storage engines, implementing…
Debugs and fixes dbt errors systematically. Use when working with dbt errors for: (1) Task mentions "fix", "error", "broken", "failing", "debug", "wrong", or "not working" (2)…
NVIDIA DeepStream SDK 9.0 development with Python pyservicemaker API. Use when building video analytics pipelines, GStreamer-based video processing, TensorRT inference…
Delta Lake テーブルの設計・最適化・運用を支援するスキル。 テーブル設計(Liquid Clustering、パーティション、Deletion Vectors)、 データ操作(MERGE最適化、CDF、Streaming)、 パフォーマンス(OPTIMIZE、VACUUM、Data Skipping)、 Medallion…
Deploys Apache Kafka on Kubernetes using the Strimzi operator with KRaft mode. Use when setting up Kafka for event-driven microservices, message queuing, or pub/sub patte — from…
Test distributed systems in the style of TigerBeetle and Joran Dirk Greef, using deterministic simulation and time compression.
Conception de pipelines CI/CD pour tout type de plateforme. Se déclenche avec "CI/CD", "pipeline", "GitHub Actions", "Azure DevOps", "GitLab CI", "déploiement automatique — from…
Architecture de messaging avec RabbitMQ, Kafka, Azure Service Bus. Se déclenche avec "message queue", "RabbitMQ", "Kafka", "queue", "messaging", "async", "pub/sub", "brok — from…
Detect duplicate definitions across repos: Drizzle table definitions, Kafka topic registrations, migration prefixes, and Python model names.
Use when designing data pipelines, choosing between ETL and ELT approaches, or implementing data transformation patterns. Covers modern data pipeline architecture.
Design, implement, review, test, troubleshoot, and operate large deterministic ETL pipelines from Oracle through canonical TSV, typed Parquet, DuckDB validation, and versioned…
Event sourcing and CQRS expert for AI memory systemsUse when "event sourcing, event store, cqrs, nats jetstream, kafka events, event projection, replay events, event schema,…
Use when developing, sharpening, or pressure-testing a strategy or strategic plan — for a business, product, startup, campaign, or personal plan; when asking "is this a real…
TypeScript ve .NET derleme hatalarını tespit edip düzeltir — any kullanmak YASAK
数据格式转换专业版面向需要在大规模数据与多源系统间进行格式转换的专业开发者与数据工程师,提供完整的批量、流式、自动化转换能力。核心能力: - 涵盖免费版全部能力,无文件大小与数量限制 - 批量转换:目录级批量处理,支持通配符与递归 - 流式转换:大文件分块读取与流式输出,支持 GB 级 - 自定义字段映射:重命名、过滤、合并、拆分字段 -…
Use when implementing client-server state synchronization, delta compression, optimistic updates, rollback netcode, or real-time game state reconciliation.
Analyze BigQuery slot reservation sizing, BI Engine acceleration, query cost estimation, dataset governance (expiration, access controls), and partitioning/clustering optimization…
Analyze BigQuery query patterns and storage to dramatically reduce the #1 surprise GCP cost driver
Design and troubleshoot data pipelines using Dataflow (Apache Beam), Pub/Sub messaging, Dataproc (Spark/Hadoop), Cloud Composer (Apache Airflow), and Dataplex data governance.
Gate BigQuery dataset deletion, table truncation, and authorized view changes against a full downstream dependency audit and export confirmation.
Ermittelt wirtschaftlich Berechtigte, Kontrollketten, Trust-/Stiftungsstrukturen, Nominees und Transparenzregisterdaten.
Open source MCP server for databases that simplifies AI agent access to database resources. Handles connection pooling, authentication, and observability with OpenTelemetry…
Build cost-safe Geospatial SQL for BigQuery, Snowflake, Wherobots and Postgres and/or render results on an interactive Dekart map.
用中文帮助用户把 GitHub 项目数据整理成静态展示网站,并发布到 Netlify。适用于用户想从 GitHub、HelloGitHub 或本地 JSON 获取项目数据,生成或更新 Vite/React 静态网站,配置 netlify.toml,登录/链接 Netlify,执行预览或生产发布,并把“获取项目数据到发布网站”的流程做成可复用步骤时。
> Pattern library for cloud, infrastructure, and developer tool vendors that appear on bank statements and ledger detail worldwide.
Guides agents to discover requirements and design a governed, secure borderless open data lakehouse with agentic AI integration.
Agent Skills oficiales de Google para productos y tecnologías de Google Cloud — Gemini API, BigQuery, Cloud Run, Firebase, GKE, AlloyDB y más. Instalación vía npx skills.
Build ETL pipelines and analytics dashboards for Harvard Art Museums API data using Python, SQL, and Streamlit
Copy data between any databases with a single CLI command using Ingestr. Supports 50+ sources and destinations including PostgreSQL, MySQL, BigQuery, Snowflake, DuckDB, MongoDB,…
Apache Kafka architecture expert for event-driven systems, cluster design, partition strategies, consumer groups, and event sourcing/CQRS patterns.
Apache Kafka architecture expert for cluster design, capacity planning, and high availability. Use when designing Kafka clusters, choosing partition strategies, or sizing brokers…
Design Kafka consumer groups that survive rebalances, hit the right delivery guarantee (at-most/at-least/exactly-once), and handle poison messages without stalling the topic.
Implement type-safe Kafka consumers for event consumption with msgspec deserialization. Use when building async consumers that process domain events (order messages, transactions)…
Expert in event-driven architecture using Kafka. Use this for designing topics, schemas, and processing logic for asynchronous tasks.
Manage the kafka-sensor-city demo. Run with no arguments to start the full demo. Run with "teardown" to remove all resources.
Best practices and guidelines for Apache Kafka event streaming and distributed messaging
Applies general coding standards and best practices for Kafka development with Scala.
Expert in Apache Kafka, Event Streaming, and Real-time Data Pipelines. Specializes in Kafka Connect, KSQL, and Schema Registry.