Search

Apache Spark · Data pipelines and ETL

11 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Migrate a legacy AWS Glue development endpoint to a Glue interactive session, following the official AWS migration checklist.

aws-samples/aws-glue-samples1.5k—~3.6kAutomated safety check: PassMIT-01 mo ago
2

Guides writing and tuning Apache Spark jobs: DataFrame and RDD code, Spark SQL, partitioning, caching, shuffle tuning and structured streaming.

Jeffallan/claude-skills12k1 repo~1.7kAutomated safety check: PassMIT7 days ago
3

Speed up slow Apache Spark jobs by tuning partitions, shuffles, data skew, caching and executor memory, with do and don't rules for PySpark code.

wshobson/agents40k9 repos~789Automated safety check: PassMIT6 days ago
4

Build scalable data pipelines, modern data warehouses, and real-time streaming architectures.

davila7/claude-code-templates33k8 repos~2.8kAutomated safety check: PassMITtoday
5

Analyze DBSQL queries, including SQL embedded in notebooks (spark.sql(...), %sql cells), for anti-patterns, lint issues, and performance problems, using Databricks-specific dialect and platform…

AltimateAI/data-engineering-skills128—~6.7kAutomated safety check: PassMIT3 days ago
6

Data pipeline expert for ETL, Apache Spark, Airflow, dbt, and data quality

RightNow-AI/openfang18k—~847Automated safety check: PassApache-2.03 mo ago
7

Data engineering patterns for ETL pipelines, data warehousing, Apache Spark, and data quality validation

rohitg00/awesome-claude-code-toolkit2.7k—~1.7kAutomated safety check: PassApache-2.05 mo ago
8

Execute arbitrary Python or PySpark code on Fabric Spark compute without creating a notebook artifact; ephemeral Livy sessions with full Delta table access.

data-goblin/power-bi-agentic-development1k—~1.7kAutomated safety check: PassGPL-3.04 days ago
9

Transform pyspark transformer operations. An agent skill from jeremylongshore/tons-of-skills-marketplace.

jeremylongshore/tons-of-skills-marketplace2.8k—~566Automated safety check: PassMITyesterday
10

Optimize spark sql optimizer operations. An agent skill from jeremylongshore/tons-of-skills-marketplace.

jeremylongshore/tons-of-skills-marketplace2.8k—~564Automated safety check: PassMITyesterday
11

Best practices for building performant, testable PySpark ETL pipelines with Spark SQL and Apache Iceberg.

Mindrally/skills271—~2.5kAutomated safety check: PassApache-2.02 days ago