spark-job-creator
|
Works with
Claude CodeCursorCodex CLIGitHub CopilotGemini CLI
--- name: spark-job-creator description: | license: MIT --- # Spark Job Creator ## Overview This skill provides automated assistance for spark job creator tasks within the Data Pipelines domain. ## When to Use This skill activates automatically when you: - Mention "spark job creator" in your request - Ask about spark job creator patterns or best practices - Need help with data pipeline skills covering etl, data transformation, workflow orchestration, and streaming data processing. ## Instructions 1. Provides step-by-step guidance for spark job creator 2. Follows industry best practices and patterns 3. Generates production-ready code and configurations 4. Validates outputs against common standards ## Examples **Example: Basic Usage** Request: "Help me with spark job creator" Result: Provides step-by-step guidance and generates appropriate configurations ## Prerequisites - Relevant development environment configured - Access to necessary tools and services - Basic understanding of data pipelines concepts ## Output - Generated configurations and code - Best practice recommendations - Validation results ## Error Handling | Error | Cause | Solution | |-------|-------|----------| | Configuration invalid | Missing required fields | Check documentation for required parameters | | Tool not found | Dependency not installed | Install required tools per prerequisites | | Permission denied | Insufficient access | Verify credentials and permissions | ## Resources - Official documentation for related tools - Best practices guides - Community examples and tutorials ## Related Skills Part of the **Data Pipelines** skill category. Tags: etl, airflow, spark, streaming, data-engineering
More Data Engineering skills
data-pipeline
claude-office-skills/skills
Data pipeline and ETL automation - extract, transform, load workflows for data integration and analytics
4.1k
ETL Pipeline
claude-office-skills/skills
Design and automate Extract, Transform, Load data pipelines for data integration and analytics
3.9k
data-throughput-accelerator
affaan-m/ecc
Use when large data ingestion, backfill, export, ETL, warehouse loading, manifest catch-up, or table synchronization needs to become much faster while preserving data correctness.
3.6k

