spark
https://github.com/apache/spark
Scala
Apache Spark - A unified analytics engine for large-scale data processing
Triage Issues!
When you volunteer to triage issues, you'll receive an email each day with a link to an open issue that needs help in this project. You'll also receive instructions on how to triage issues.
Triage Docs!
Receive a documented method or class from your favorite GitHub repos in your inbox every day. If you're really pro, receive undocumented methods or classes and supercharge your commit history.
Scala not yet supported76 Subscribers
View all SubscribersAdd a CodeTriage badge to spark
Help out
- Issues
- [MINOR][MLLIB] Exclude zero-weight observations from KMeans initialization
- [SPARK-58210][SQL][FOLLOWUP] Extend CombineAdjacentAggregation to partial merge
- [SPARK-58389][SQL] Pass all options while loading tables for writes
- [SPARK-57896][CORE][TESTS] Add Kerberos coexistence and per-user token integration tests
- [SPARK-58660][INFRA] Self-heal the PR Build check when the notify workflow fails to create it
- [SPARK-58661][SQL][DOCS] Document return values and fix stale column references in DataFrameStatFunctions
- [SPARK-58662][SS] Remove redundant toString in RocksDBStateStoreProvider state transition messages
- [SPARK-58663][BUILD] Fix typos and grammar in root pom.xml comments
- [SPARK-58656][PYTHON][TESTS] Centralize the to_pandas golden test inventory into a shared base
- [SPARK-58657][PYTHON][TESTS] Add tests for pa.Array.from_pandas with the mask argument
- Docs
- Scala not yet supported