Search
Apache Spark · SQL
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Guides writing and tuning Apache Spark jobs: DataFrame and RDD code, Spark SQL, partitioning, caching, shuffle tuning and structured streaming. | Jeffallan/ | 12k | 1 repo | ~1.7k | Automated safety check: Pass | MIT | 6 days ago |
| 2 | A skill your agent uses when the user is writing datafusion-python (Apache DataFusion Python bindings) DataFrame or SQL code. | apache/ | 607 | — | ~7.8k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 3 | Analyze DBSQL queries, including SQL embedded in notebooks (spark.sql(...), %sql cells), for anti-patterns, lint issues, and performance problems, using Databricks-specific dialect and platform… | AltimateAI/ | 128 | — | ~6.7k | Automated safety check: Pass | MIT | yesterday |
| 4 | Use Databricks built-in AI Functions (aiclassify, aiextract, aisummarize, aimask, aitranslate, aifixgrammar, aigen, aianalyzesentiment, aisimilarity, aiparsedocument, aiprepsearch, aiquery… | databricks/ | 345 | — | ~3.9k | Automated safety check: Pass | Unknown | today |
| 5 | Upgrade Apache Spark applications between major versions (2.x→3.x, 3.x→4.x). | OpenHands/ | 161 | — | ~1.9k | Automated safety check: Pass | MIT | today |
| 6 | Assists with benchmarking and profiling the performance of an Apache Spark UDF on the GPU. | Kilo-Org/ | 190 | — | ~802 | Automated safety check: Pass | Proprietary | 11 days ago |
| 7 | Best practices for building performant, testable PySpark ETL pipelines with Spark SQL and Apache Iceberg. | Mindrally/ | 269 | — | ~2.5k | Automated safety check: Pass | Apache-2.0 | today |