pricing

Analyze Spark logs for free. Detect regressions across many runs with Pro.

Open Source is for single-run local diagnostics. Pro is for teams that need regression detection, CI enforcement, historical comparisons, and repeatable Spark performance workflows. Every tier runs locally or self-hosted, so Spark event logs stay inside your environment.

Open Source

Free

Apache 2.0 licensed

Analyze individual Spark event logs locally.

The open-source CLI helps engineers diagnose one Spark event log at a time. It runs locally and produces evidence-backed reports without requiring a Spark History Server or uploading logs.

  • ·Local Spark event-log analysis
  • ·Skew, spill, shuffle, retry, and failure detection
  • ·SQL execution extraction and physical plan output
  • ·SQL DOT graph export
  • ·JSON analysis report
  • ·Markdown recommendations
  • ·Apache 2.0 licensed

SparkDoctor Pro

$24,000 / year

Billed annually

Catch Spark performance regressions before they reach production.

SparkDoctor Pro turns single-run Spark diagnostics into a repeatable team workflow. Compare baseline and current runs, detect regressions in runtime, shuffle, spill, failures, and bottlenecks, and fail CI when Spark jobs cross configured thresholds.

Pro keeps the same local-first model: run it inside your environment, keep Spark event logs under your control, and avoid sending sensitive Spark metadata to a SaaS platform.

  • ·Compare baseline vs current Spark runs
  • ·Detect runtime, shuffle, spill, failure, and bottleneck regressions
  • ·Fail CI when configured regression thresholds are crossed
  • ·Track performance trends across many Spark job runs
  • ·Store run history in a local or customer-controlled database
  • ·Batch analyze directories of Spark event logs
  • ·Configure team thresholds and regression policies
  • ·Generate team-ready reports for pull requests, incidents, and reviews
  • ·Support local/self-hosted deployment workflows

Enterprise

Contact us

Custom agreement

For platform teams running Spark at scale.

Enterprise is for organizations that need higher-volume event-log analysis, team-wide Spark performance visibility, custom deployment support, cloud storage ingestion, Databricks workflows, security review support, and priority engineering help.

  • ·Everything in Pro
  • ·Databricks job/run import workflows
  • ·S3 or cloud storage event-log ingestion
  • ·Organization-wide Spark performance trends
  • ·Custom rule packs and thresholds
  • ·Deployment support for private environments
  • ·Security and compliance review support
  • ·Priority support and feature requests

Feature comparison

FeatureOpen SourceProEnterprise
Local Spark event-log analysisyesyesyes
Runs locally / self-hostedyesyesyes
Compressed event-log support (gzip, zstd, lz4, snappy)yesyesyes
Spark 4 event-log directory supportyesyesyes
Skew, spill, shuffle, retry, failure detectionyesyesyes
SQL execution extraction and physical plan outputyesyesyes
SQL DOT graph exportyesyesyes
JSON analysis reportyesyesyes
Markdown recommendationsyesyesyes
Single-run diagnosticsyesyesyes
Baseline vs current run comparison-noyesyes
Regression detection across Spark metrics-noyesyes
CI failure gates on configured thresholds-noyesyes
Historical trend analysis across many runs-noyesyes
Local / customer-controlled run database-noyesyes
Batch analysis over directories of event logs-noyesyes
Team thresholds and regression policies-noyesyes
Pull request and incident-ready reports-noyesyes
Databricks job/run import workflows-no-noyes
S3 / cloud storage event-log ingestion-no-noyes
Organization-wide Spark performance trends-no-noyes
Custom rule packs and thresholds-no-noyes
Deployment support for private environments-no-noyes
Security and compliance review support-no-noyes
Supportcommunityincludedpriority
LicenseApache 2.0commercialcommercial

Open Source is for single-run local diagnostics. Pro adds the team layer: regression detection, CI gates, historical comparison, and local performance tracking. Enterprise extends Pro with cloud ingestion, Databricks workflows, organization-wide trends, and dedicated support.