spark
https://github.com/apache/spark
Scala
Apache Spark - A unified analytics engine for large-scale data processing
Triage Issues!
When you volunteer to triage issues, you'll receive an email each day with a link to an open issue that needs help in this project. You'll also receive instructions on how to triage issues.
Triage Docs!
Receive a documented method or class from your favorite GitHub repos in your inbox every day. If you're really pro, receive undocumented methods or classes and supercharge your commit history.
Scala not yet supported76 Subscribers
View all SubscribersAdd a CodeTriage badge to spark
Help out
- Issues
- [SPARK-58767][SQL] Fixing the canonicalization of TableCacheQueryStageExec which otherwise prevents re-use of exchange for in memory cached plans
- [SPARK-57491][FOLLOWUP] Make stale push-based shuffle fallback chunk-granular and opt-in
- [SPARK-58721][K8S][TESTS] Add tests for ClientArguments.fromCommandLineArgs
- [SPARK-58389][SQL] Load streaming targets with state options and privileges
- [WIP][ML] Avoid VectorUDT deserialization in Summarizer
- [SPARK-58484][SQL] Fix ArrayIndexOutOfBoundsException in OrderedFilters when the required schema is empty
- [SPARK-58716][SQL] Apply default collation to implicit casts
- [MINOR][PYTHON][DOCS] Fix PySpark SQL function documentation
- [PIPELINES] Support partition transforms in SDP partition_cols
- [PIPELINES] Support partition transforms in Declarative Pipelines partition_cols
- Docs
- Scala not yet supported