LogoClawIndex
CasesSkillsAbout
LogoClawIndex

data-engineer - Build scalable data pipelines and modern data platforms.

Builds scalable data pipelines, modern data warehouses, and streaming architectures using Apache Spark, dbt, Airflow, and cloud data platforms.

Tags

Updated: 2026-09-24

Capabilities

Typical Inputs

Typical Outputs

What this skill does

  • Design batch and streaming pipelines
  • Build data warehouses and lakehouses
  • Implement data quality and governance
  • Orchestrate workflows using Airflow
  • Transform data using dbt
  • Process streaming data with Kafka

Inputs

  • Data sources and storage systems
  • Data contracts and SLAs
  • Data schemas and pipeline configurations

Outputs

  • Data pipelines and ingestion workflows
  • Data warehouse and lakehouse tables
  • Data quality metrics and alerts
  • Data lineage and schema documentation

Requirements

  • Access permissions to data sources and storage systems
  • Execution runtime for Python, Scala, or SQL

Source

  • Spec: SKILL.md

ClawIndex

OpenClaw Skills & Use Case Index

ClawIndex is an ecosystem-driven index of OpenClaw skills and real-world use cases.

Index

Skills·
Cases

Meta

About·
Disclaimer·
Email·
GitHub
© 2026 ClawIndex All Rights Reserved.
data-engineering
etl
data-pipeline
data-warehouse
streaming
spark
dbt
airflow
Design batch and streaming pipelines
Build data warehouses and lakehouses
Implement data quality and governance
Orchestrate workflows using Airflow
Data sources and storage systems
Data contracts and SLAs
Data schemas and pipeline configurations
Data pipelines and ingestion workflows
Data warehouse and lakehouse tables
Data quality metrics and alerts