spark
https://github.com/apache/spark
Scala
Apache Spark - A unified analytics engine for large-scale data processing
Triage Issues!
When you volunteer to triage issues, you'll receive an email each day with a link to an open issue that needs help in this project. You'll also receive instructions on how to triage issues.
Triage Docs!
Receive a documented method or class from your favorite GitHub repos in your inbox every day. If you're really pro, receive undocumented methods or classes and supercharge your commit history.
Scala not yet supported76 Subscribers
View all SubscribersAdd a CodeTriage badge to spark
Help out
- Issues
- [SPARK-56942][SPARK-58815][SQL] Support nested row IDs and collision-safe attribute binding
- [SPARK-57518][SQL][FOLLOWUP] List ThriftServer schemas via SupportsNamespaces instead of special-casing spark_catalog
- [SPARK-58368][INFRA] Add smart class-level test selection for post-merge CI
- [SPARK-58269][SQL] Infer generated column partition filters
- [SPARK-58277][SQL] Stream DataType JSON serialization
- Build: Test Apache Parquet 1.18.0
- [SPARK-58227][CONNECT] Support Spark Connect in the spark-sql CLI (spark-sql --remote)
- [POC][SPARK-51705][CONNECT][PYTHON] Support SparkSession.broadcast() (broadcast variables over Spark Connect)
- [WIP][SPARK-58239][SQL] DML Command Outputs for DSv2
- [SPARK-58226][SQL] Honor persistentCatalogFirst for DDL on two-part name resolution
- Docs
- Scala not yet supported