Apache Airflow - A platform to programmatically author, schedule, and monitor workflows
$ npx skills add apache/airflowAlternatives
Compare similar skills by workflow fit, trust score, quality, GitHub adoption, maintenance, and install readiness.
Current skill
An end-to-end data engineering pipeline that orchestrates data ingestion, processing, and storage using Apache Airflow, Python, Apache Kafka, Apache Zookeeper, Apache Spark, and Cassandra. All components are containerized with Docker for easy deployment and scalability.
Apache Airflow - A platform to programmatically author, schedule, and monitor workflows
$ npx skills add apache/airflowChange data capture for a variety of databases. Please log issues at https://github.com/debezium/dbz/issues.
$ npx skills add debezium/debeziumEmpowering Data Intelligence with Distributed SQL for Sharding, Scalability, and Security Across All Databases.
$ npx skills add apache/shardingsphereOpen-source data movement for ELT pipelines and AI agents — from APIs, databases & files to warehouses, lakes, and AI applications. Both self-hosted and Cloud.
$ npx skills add airbytehq/airbytePython ETL framework for stream processing, real-time analytics, LLM pipelines, and RAG.
$ npx skills add pathwaycom/pathwayFancy stream processing made operationally mundane
$ npx skills add redpanda-data/connectClickHouse® is a real-time analytics database management system
$ npx skills add ClickHouse/ClickHouseApache Spark - A unified analytics engine for large-scale data processing
$ npx skills add apache/spark$ npx skills add apache/flinkModern and easy to use SQL client for MySQL, Postgres, SQLite, SQL Server, and more. Linux, MacOS, and Windows.
$ npx skills add beekeeper-studio/beekeeper-studioThe official home of the Presto distributed SQL query engine for big data
$ npx skills add prestodb/prestoTurn any AI agent into an AI Scientist. The #1 Agent Skills library for science, used by 160,000+ scientists worldwide. 140 ready-to-use skills plus 100+ scientific databases covering biology, chemistry, medicine, and drug discovery. Compatible with Cursor, Claude Code, Codex, Pi, Antigravity, and the open Agent Skills standard.
$ npx skills add K-Dense-AI/scientific-agent-skillsOfficial repository of Trino, the distributed SQL query engine for big data, formerly known as PrestoSQL (https://trino.io)
$ npx skills add trinodb/trinoFlexible and powerful data analysis / manipulation library for Python, providing labeled data structures similar to R data.frame objects, statistical functions, and much more
$ npx skills add pandas-dev/pandasBuild and share delightful machine learning apps, all in Python. 🌟 Star to support our work!
$ npx skills add gradio-app/gradioThe world's fastest open query engine for sub-second analytics both on and off the data lakehouse. With the flexibility to support nearly any scenario, StarRocks provides best-in-class performance for multi-dimensional analytics, real-time analytics, and ad-hoc queries. A Linux Foundation project.
$ npx skills add StarRocks/starrocksHow to choose
Use an alternative when it has a clearer install path, higher trust score, fresher maintenance, or better platform fit for your current agent stack. Keep E2e Data Engineering if it already passes your workflow test and repository review.
Next step
Open the compare page, test the install commands in a sandbox, and check each repository before using a skill in production.
Alternatives
Compare similar skills by workflow fit, trust score, quality, GitHub adoption, maintenance, and install readiness.
Current skill
An end-to-end data engineering pipeline that orchestrates data ingestion, processing, and storage using Apache Airflow, Python, Apache Kafka, Apache Zookeeper, Apache Spark, and Cassandra. All components are containerized with Docker for easy deployment and scalability.
Apache Airflow - A platform to programmatically author, schedule, and monitor workflows
$ npx skills add apache/airflowChange data capture for a variety of databases. Please log issues at https://github.com/debezium/dbz/issues.
$ npx skills add debezium/debeziumEmpowering Data Intelligence with Distributed SQL for Sharding, Scalability, and Security Across All Databases.
$ npx skills add apache/shardingsphereOpen-source data movement for ELT pipelines and AI agents — from APIs, databases & files to warehouses, lakes, and AI applications. Both self-hosted and Cloud.
$ npx skills add airbytehq/airbytePython ETL framework for stream processing, real-time analytics, LLM pipelines, and RAG.
$ npx skills add pathwaycom/pathwayFancy stream processing made operationally mundane
$ npx skills add redpanda-data/connectClickHouse® is a real-time analytics database management system
$ npx skills add ClickHouse/ClickHouseApache Spark - A unified analytics engine for large-scale data processing
$ npx skills add apache/spark$ npx skills add apache/flinkModern and easy to use SQL client for MySQL, Postgres, SQLite, SQL Server, and more. Linux, MacOS, and Windows.
$ npx skills add beekeeper-studio/beekeeper-studioThe official home of the Presto distributed SQL query engine for big data
$ npx skills add prestodb/prestoTurn any AI agent into an AI Scientist. The #1 Agent Skills library for science, used by 160,000+ scientists worldwide. 140 ready-to-use skills plus 100+ scientific databases covering biology, chemistry, medicine, and drug discovery. Compatible with Cursor, Claude Code, Codex, Pi, Antigravity, and the open Agent Skills standard.
$ npx skills add K-Dense-AI/scientific-agent-skillsOfficial repository of Trino, the distributed SQL query engine for big data, formerly known as PrestoSQL (https://trino.io)
$ npx skills add trinodb/trinoFlexible and powerful data analysis / manipulation library for Python, providing labeled data structures similar to R data.frame objects, statistical functions, and much more
$ npx skills add pandas-dev/pandasBuild and share delightful machine learning apps, all in Python. 🌟 Star to support our work!
$ npx skills add gradio-app/gradioThe world's fastest open query engine for sub-second analytics both on and off the data lakehouse. With the flexibility to support nearly any scenario, StarRocks provides best-in-class performance for multi-dimensional analytics, real-time analytics, and ad-hoc queries. A Linux Foundation project.
$ npx skills add StarRocks/starrocksHow to choose
Use an alternative when it has a clearer install path, higher trust score, fresher maintenance, or better platform fit for your current agent stack. Keep E2e Data Engineering if it already passes your workflow test and repository review.
Next step
Open the compare page, test the install commands in a sandbox, and check each repository before using a skill in production.