Sign in
apache
/
spark
/
HEAD
1885442
[SPARK-58398][CORE] Group-atomic failure and fail-fast rejection for pipelined-shuffle stage groups
by Boyang Jerry Peng
· 2 hours ago
master
d5ec658
[SPARK-58399][SQL][PYTHON] Add `collect_union` aggregate function
by ChuckLin2025
· 2 hours ago
166dbc5
[SPARK-57158][SQL] ExplainUtils: extract operator ID assignment phase from processPlan into a private helper
by Mark Jarvin
· 5 hours ago
54dce69
[SPARK-57811][SQL] Support string to nanosecond-precision timestamp coercion in comparisons and predicates
by Stevo Mitric
· 6 hours ago
6e8829d
[SPARK-58365][SQL] Build the NOT-IN-in-disjunction join condition from the deduplicated subquery output
by YangJie
· 7 hours ago
92ad099
[SPARK-58373][SQL] Do not prune the from_json schema under named_struct when parse options are set
by YangJie
· 7 hours ago
91b1abd
[SPARK-58403][SQL] Assign appropriate error condition for `_LEGACY_ERROR_TEMP_3201-3205`: `MALFORMED_EXPRESSION_INFO`
by YangJie
· 8 hours ago
6b0c5f8
[SPARK-38954][CORE] Support delegation token renewal and distribution without Kerberos
by Parth Chandra
· 13 hours ago
6c37fe3
[SPARK-58159][PYTHON] Support `with` statement for connect session
by Tian Gao
· 13 hours ago
0c1fd8f
[SPARK-58037][PYTHON][TEST][FOLLOWUP] Keep DataFrame golden tests internal
by Mihailo Aleksic
· 13 hours ago
0c00d05
[MINOR][BUILD] Fix duplicated word in spark-profiler POM description
by Uros Bojanic
· 16 hours ago
a0229e7b
[SPARK-58416][SQL][EXAMPLE] Fix wrong class name in SqlNetworkWordCount usage messages
by Uros Bojanic
· 16 hours ago
7692b9f
[SPARK-58417][PS] Remove stale Python 2 reference in DataFrame.to_latex docstring
by Uros Bojanic
· 17 hours ago
8df89f2
[SPARK-58347][SDP] Thread conf.resolver through ColumnSelection.applyToSchema
by Andreas Neumann
· 19 hours ago
4d670c3
[SPARK-58271][SDP] Allow AUTO CDC clauses in any order
by Andreas Neumann
· 19 hours ago
8bee029
[SPARK-55199][K8S][DOC] Improve K8s integration tests `README`
by Dongjoon Hyun
· 19 hours ago
5db219c
[SPARK-58422][K8S][TEST] Upgrade the minimum Minikube version to 1.38.0
by Dongjoon Hyun
· 19 hours ago
fd4c586
[SPARK-58369][CONNECT][TEST] Introduce a way to `SparkConnectServiceKeepAliveSuite` to wait until unbinding a port is complete
by Kousuke Saruta
· 22 hours ago
3ff822a
[SPARK-58400][SQL] Fix ANSI mode assumption for views created by Spark 4.0+
by Cheng Pan
· 22 hours ago
a06f591
[SPARK-58414][SQL][TEST][FOLLOWUP] Use the imported LocalDateTime instead of the fully qualified name
by Liang-Chi Hsieh
· 22 hours ago
51ff60d
[SPARK-58374][BUILD] Upgrade `joda-time` to 2.14.3
by Dongjoon Hyun
· 22 hours ago
f27705d
[SPARK-58330][SQL] Prevent silent dropping of dynamic options on same table references.
by Anurag Mantripragada
· 22 hours ago
5ceafed
[SPARK-58414][SQL][TEST] Add e2e coverage for nanosecond timestamps nested in complex types in the Arrow cache
by Liang-Chi Hsieh
· 23 hours ago
584268c
[SPARK-58412][SQL] Prevent incomplete and stale cache materialization statistics
by Chao Sun
· 23 hours ago
2c8570a
[SPARK-58292][CORE] Recreate the netty worker EventLoopGroup when a worker event loop thread dies
by ChuckLin2025
· 25 hours ago
87dbc24
[SPARK-57271][PYTHON] Propagate traceback locals to Python planner runner
by Linhong Liu
· 25 hours ago
1fba3cc
[SPARK-58413][SQL][TEST] Rename `sequential fetch` label to `pipelined fetch (1 client)` in `NettyTransportBenchmark`
by Dongjoon Hyun
· 30 hours ago
6b793ce
[SPARK-58415][INFRA][TEST] Regenerate benchmark results
by Dongjoon Hyun
· 30 hours ago
24ad34b
[SPARK-58321][SDP] Wire SCD2 AutoCDC streaming write and enable SCD2 end to end
by Andreas Neumann
· 33 hours ago
233f1a41
[SPARK-58396][ML][CONNECT] Include FPGrowth metadata in size estimates
by Ruifeng Zheng
· 2 days ago
177ce56
[SPARK-58408][CORE] Register TimestampNanosVal in KryoSerializer
by Liang-Chi Hsieh
· 2 days ago
ece8eac
[SPARK-58381][SQL] Gate the Arrow cache zero-copy write path on physical congruence with the cache schema
by Liang-Chi Hsieh
· 2 days ago
86de29d2
[SPARK-57893][CORE] Implement `UserCredentialManager`
by Kousuke Saruta
· 2 days ago
12785d5
[SPARK-57761][SQL] Add missing error class INVALID_XML_SCHEMA_MAP_TYPE
by Szehon Ho
· 2 days ago
b7d304f
[SPARK-58312][PYTHON] Remove unnecessary type: ignore comments in pyspark.sql.group
by Spenser Sun
· 2 days ago
04d4c26
[SPARK-58377][BUILD] Upgrade `netty-tcnative` to 2.0.81.Final
by Dongjoon Hyun
· 2 days ago
37ca0a9
[SPARK-58375][BUILD] Upgrade `ap-loader` to 4.5-13
by Dongjoon Hyun
· 2 days ago
7050d16
[SPARK-58376][BUILD] Upgrade `byte-buddy` to 1.18.11
by Dongjoon Hyun
· 2 days ago
f66e380
[SPARK-58379][BUILD] Upgrade `jnr-posix` to 3.2.1
by Dongjoon Hyun
· 2 days ago
fd5b0d4
[SPARK-57955][SQL] Raise a proper error for out-of-Int-range data type parameters
by YangJie
· 2 days ago
d677bd3
[SPARK-58382][SQL] binaryFile archive support via wholeFile option
by akshatshenoi-db
· 2 days ago
2df516c
[SPARK-58402][PYTHON] Consolidate UNKNOWN_EXPLAIN_MODE into VALUE_NOT_ALLOWED
by Ruifeng Zheng
· 2 days ago
073c8d8
[SPARK-58390][SQL] Emit row counts without deserializing Arrow payloads for empty-projection cache reads
by Liang-Chi Hsieh
· 2 days ago
b2918df
[SPARK-58397][ML][CONNECT] Add Bucketizer size estimate
by Ruifeng Zheng
· 2 days ago
5d440c7
[SPARK-58388][INFRA] Clean up the dependency list for lint/docs/gen-protos docker image
by Tian Gao
· 2 days ago
a6c3eb21
[MINOR][SQL] Simplify redundant boolean ternary in visitDeclareCursorStatement
by Uros Bojanic
· 2 days ago
2ad6c51
[SPARK-58401][DOC] Update Parquet links to 1.17.1 in Parquet data source docs
by Uros Bojanic
· 2 days ago
e8cbd48
[MINOR][INFRA] Fix doubled braces in structured logging style-check hint
by Uros Bojanic
· 2 days ago
2963734
[SPARK-58008][SQL] Support dynamic table options for DELETE
by BRIJ RAJ KISHORE
· 2 days ago
6f12573
[SPARK-58360][ML][CONNECT] Avoid nested parent overcount in RFormulaModel size estimate
by Ruifeng Zheng
· 2 days ago
a1c3f31
[SPARK-58370][SQL] Check table write privileges when the write target is also read in the same statement
by Peter Toth
· 2 days ago
9ccb1dd
[SPARK-55749][SQL][DOC] Clarify `array_contains` null-handling in the docs and example
by Oleks V
· 2 days ago
8636942
[SPARK-58332][PYTHON][TEST] Move compare_or_generate_golden_matrix into GoldenFileTestMixin
by Spenser Sun
· 2 days ago
35309c5
[SPARK-58263][CORE] Concurrently schedule pipelined-shuffle stage groups in the DAGScheduler
by Boyang Jerry Peng
· 2 days ago
8ed480d
[SPARK-58317][SQL] Union output partitioning should support PartitioningCollection children
by Xiduo You
· 3 days ago
6f05a83
[SPARK-58367][ML][CONNECT] Include ALS metadata in size estimates
by Ruifeng Zheng
· 3 days ago
f711fdf
[SPARK-58361][ML][CONNECT] Include clustering model metadata in size estimates
by Ruifeng Zheng
· 3 days ago
f1a7058
[SPARK-58249][PS][FOLLOWUP] Use native function for NumPy invert
by Ruifeng Zheng
· 3 days ago
d8d08b6
[SPARK-58097][CONNECT] Preserve composite (userId, sessionId) session identity in Connect UI/status store
by Venkata krishnan Sowrirajan
· 3 days ago
a8c6933
[SPARK-58364][CORE] Rename the configuration namespace from `spark.security.credentials` to `spark.security.oidc`
by Kousuke Saruta
· 3 days ago
b66553d
[SPARK-57395][SDP] Implement SCD2 Batch Processor; foreachBatch Callback
by Andreas Neumann
· 3 days ago
3b28be3
[SPARK-58308][INFRA] Remove usage of ipython_genutils
by Tian Gao
· 3 days ago
5ce51a6
[SPARK-58340][PYTHON] Add missing package to pyproject.toml for lint
by Tian Gao
· 3 days ago
149ba07
[SPARK-58346][PYTHON] Remove unnecessary type: ignore[union-attr] comments in PySpark
by Spenser Sun
· 3 days ago
a646dc8
[SPARK-58119][CORE][PYTHON] Introduce PythonWorkerHandle abstraction for the Python worker path
by Fabian Paul
· 3 days ago
e8d5e4a
[SPARK-58354][PYTHON] Remove unused ArrowStreamPandasSerializer base class
by Yicong Huang
· 3 days ago
18ad8114
[SPARK-57251][SDP] Validate SCD2 reserved framework columns at AutoCDC flow construction
by Andreas Neumann
· 3 days ago
adad838
[SPARK-58089][SQL] Push variant extractions through Aggregate/Sort/Join
by Qiegang Long
· 3 days ago
90b1f89
[SPARK-58313][SDP] Validate SCD2 track-history columns at AutoCDC flow construction
by Andreas Neumann
· 3 days ago
2d5e1b6
[SPARK-56460][CORE] Define configs in text files
by Wenchen Fan
· 3 days ago
014c53b
[SPARK-58371][DOC] Update `json` gem version to 2.21.1
by YangJie
· 3 days ago
54d5f51
[SPARK-58363][SQL] Assign a name to the error condition _LEGACY_ERROR_TEMP_2214-2219
by YangJie
· 3 days ago
0747e28
[SPARK-57638][SQL] Avoid busy-waiting in Declarative Pipelines flow resolution
by YangJie
· 3 days ago
effd582
[SPARK-57635][SQL] Make Declarative Pipelines run-termination reason deterministic
by YangJie
· 3 days ago
a663376
[SPARK-58355][SS] Fix grammar in metadata log null-check message
by Uros Bojanic
· 3 days ago
18ca27f
[SPARK-58069][SQL][FOLLOWUP] Handle empty approx_top_k combine buffers
by Wenchen Fan
· 3 days ago
781ca6a
[SPARK-58357][SQL] Remove unused withCatalogIdentClause in SparkSqlAstBuilder
by Uros Bojanic
· 3 days ago
5973442
[SPARK-58356][CORE] Use exists instead of filter(...).nonEmpty in MasterPage
by Uros Bojanic
· 3 days ago
44d1160
[SPARK-58296][SQL] Fix to_time returning STRING type when the format is a foldable NULL
by YangJie
· 3 days ago
f6975c9
[SPARK-52825][SQL] Register existing dialects for additional URL prefixes
by Wenchen Fan
· 3 days ago
160b619
[MINOR][DOC] Use AGENTS.md for nested project instructions
by Wenchen Fan
· 3 days ago
ef9cc16
[SPARK-58210][SQL] Enable ReplaceHashWithSortAgg and CombineAdjacentAggregation by default
by Xiduo You
· 4 days ago
714cb54
[MINOR][ML] Fix doubled word in HasTrainingSummary scaladoc
by Uros Bojanic
· 4 days ago
5ebf3b2
[SPARK-58353][CONNECT] Extract parseExplainMode helper in Connect Dataset
by Uros Bojanic
· 4 days ago
ccf0eb5
[MINOR][DOC] Fix grammar in spark.speculation.efficiency.processRateMultiplier doc
by Uros Bojanic
· 4 days ago
12f4f9b
[MINOR][CONNECT] Remove redundant user_context.user_id assignment in execute_command methods
by Yash Bapat
· 4 days ago
c7b2f1a
[SPARK-58314][PYTHON] Remove ArrowStreamUDFSerializer
by Yicong Huang
· 4 days ago
dee0c17
[MINOR][DOC] Fix typo 'commited' in state data source docs
by Uros Bojanic
· 4 days ago
f8c6ded
[SPARK-58192][CORE] Support fractional spark.task.cpus
by Cheng Pan
· 4 days ago
fd31694
[MINOR][SS] Use filterNot in HDFSBackedStateStoreProvider snapshot filtering
by Uros Bojanic
· 4 days ago
ced139f
[SPARK-58343][SQL][TEST] Add unit test for TypeUtils.typeWithProperEquals
by Uros Bojanic
· 4 days ago
d2693b5
[SPARK-52246][SQL][TEST] Add bucket transform regression test for one-side shuffle with join key tail of partition keys
by naveenp2708
· 4 days ago
b910d6f
[SPARK-58341][SQL] Fix wrong results with 5+ positional parameters in `sql`
by Dongjoon Hyun
· 5 days ago
cce4355
[SPARK-58324][SQL] Drop unused sameOrderExpressions from GroupPartitionsExec k-way merge ordering
by Peter Toth
· 5 days ago
faa1d1e
[SPARK-58323][SQL] Materialize AliasAware output ordering and partitioning to avoid StackOverflowError
by Peter Toth
· 5 days ago
c784ac6
[SPARK-58342][SDP] Add unit tests for the SCD1 AutoCDC auxiliary table spec
by Andreas Neumann
· 5 days ago
6f1c904
[SPARK-58310][SQL] Avoid runtime Bloom-filter subqueries containing Python UDFs
by Chao Sun
· 5 days ago
7865fb8
[SPARK-58311][SQL] Gate generated column values on write with a table capability
by Szehon Ho
· 6 days ago
0700340
[SPARK-58334][K8S][DOC] Introduce Apache Spark K8s Operator in `running-on-kubernetes.md`
by Dongjoon Hyun
· 6 days ago
4a5bb95
[SPARK-58335][K8S][TEST] Add `SparkKubernetesClientFactorySuite`
by Dongjoon Hyun
· 6 days ago
Next »