Sign in
◑
Theme
apache
/
spark
/
HEAD
6952370
[SPARK-59313][ML] Avoid serializing FeedForwardModel scratch buffers
by Ruifeng Zheng
· 2 hours ago
master
ce20273
[SPARK-59363][DOC] Fix ungrammatical 'allows to' in the Avro data source doc
by Uros
· 2 hours ago
2e878ec
[SPARK-59319][SQL][FOLLOWUP] Enforce restricted mode in the single-pass analyzer and avoid eager class loading for rejected calls
by Hyukjin Kwon
· 3 hours ago
ebfe65f
[SPARK-59299][SQL][TEST][FOLLOWUP] Split ASOF JOIN parser test cases into separate tests
by Luka Zdravić
· 3 hours ago
e92b2c5
[SPARK-58884][SS] Add stateful benchmarks for Real-Time Mode
by Boyang Jerry Peng
· 3 hours ago
0c00f46
[SPARK-58708][SQL] Fix out-of-range length handling in substring/lpad/rpad
by Sepuri Sai Krishna
· 5 hours ago
2154b91
[SPARK-59310][SQL] Report driver-side grouping metrics for GroupPartitionsExec
by Xiduo You
· 5 hours ago
630a8a6
[SPARK-57462][PYTHON][SQL] Add PySpark support for nanosecond-precision timestamp types
by Stevo Mitric
· 6 hours ago
70d1a0d
[SPARK-59339][ML][SQL] Add a SQL expression for ML vector dot products
by Ruifeng Zheng
· 6 hours ago
b7d01b3
[SPARK-59294][SQL] Support partially clustered distribution and skew join split for LeftSingle and ExistenceJoin
by Cheng Pan
· 7 hours ago
438211d
[SPARK-59301][SQL] Unwrap the cast of a DSv2 runtime IN filter before pushing it down
by Cheng Pan
· 7 hours ago
2cd436a
[SPARK-59349][SQL] Override the gRPC authority for UDF worker Unix socket channels
by Liang-Chi Hsieh
· 8 hours ago
3d25204a
[SPARK-59273][SQL][FOLLOWUP] Restore skipping non-string columns in DataFrameNaFunctions.fillValue
by Cheng Pan
· 9 hours ago
0e2f610
[SPARK-59335][ML] Reduce GBTRegressionModel broadcast size
by Ruifeng Zheng
· 9 hours ago
6a4a3d8
[SPARK-58963][ML][SQL][FOLLOWUP] Rename the vector posexplode internal expression
by Ruifeng Zheng
· 10 hours ago
6fdf201
[SPARK-59286][ML] Reduce NaiveBayesModel transform closure size
by Ruifeng Zheng
· 10 hours ago
ee9c246
[SPARK-59082][PYTHON][SQL][DOC] Document public configs affecting SQL functions
by Ruifeng Zheng
· 12 hours ago
9de2f8d
[SPARK-59337][ML] Simplify GeneralizedLinearRegressionModel transform implementation
by Ruifeng Zheng
· 12 hours ago
b25686e
[SPARK-59338][ML] Move predictionColumn to PredictionModel
by Ruifeng Zheng
· 12 hours ago
e7f7582
[SPARK-59340][ML] Reduce Word2VecModel broadcast size
by Ruifeng Zheng
· 12 hours ago
1472590
[SPARK-59233][SQL][PYTHON] Let the session's Python worker environment reach every Python function family
by Ramon Zhou
· 14 hours ago
846e33e
[SPARK-59307][ML] Detect cyclic node ids when loading decision tree and bisecting k-means models
by Hyukjin Kwon
· 14 hours ago
dad9ee8
[SPARK-59319][SQL] Add an opt-in restricted SQL execution mode
by Hyukjin Kwon
· 14 hours ago
c7549e7
[SPARK-59164][SDP] Support nested column schema evolution
by AnishMahto
· 14 hours ago
fbfaaf8
[SPARK-59311][SQL] Avoid StackOverflowError when parsing the Catalyst type in an Avro schema property
by Hyukjin Kwon
· 14 hours ago
11e88a1
[SPARK-59308][ML] Avoid Int overflow when computing the decoded image size
by Hyukjin Kwon
· 14 hours ago
9b9a30e
[SPARK-59320][UI] Optionally require an expiration claim in JWSFilter
by Hyukjin Kwon
· 15 hours ago
e15ac21
[SPARK-59278][SQL] Fix CHAR comparison rewrite edge cases
by Shivadarshan Devadiga
· 15 hours ago
51f54b0
[SPARK-59273][SQL] Complete CHAR/VARCHAR support at core execution boundaries
by Serge Rielau
· 19 hours ago
4525eb9
[SPARK-59341][SQL] Assign a tree pattern to UnresolvedStarBase related nodes
by Mihailo Aleksic
· 19 hours ago
d895f57
[SPARK-58897][SQL] Propagate NaN through the infinity norm in vector_norm
by Sepuri Sai Krishna
· 19 hours ago
ebb0bd2
[SPARK-59249][SQL] Take the grouped key-row ordering from the shared InternalRowComparableWrapper cache
by Peter Toth
· 21 hours ago
6a23cb6
[SPARK-59279][SQL] Don't re-read the sorted-merge config after GroupPartitionsExec is planned
by Peter Toth
· 22 hours ago
de60883
[SPARK-58821][SQL] Return the recomputed length instead of an internal error in Sequence
by Sepuri Sai Krishna
· 23 hours ago
44db1eb
[SPARK-58945][SQL] Fix mismatched `messageParameters` keys that cause `INTERNAL_ERROR`
by Subhramit Basu
· 24 hours ago
b017da6
[SPARK-59331][K8S] Turn off recovery mode when a held application resumes
by Dongjoon Hyun
· 24 hours ago
a980164
[SPARK-59299][SQL][TEST] Add missing ASOF JOIN parser unit tests
by Luka Zdravić
· 25 hours ago
e9c648c
[SPARK-59286][ML] Revert "Optimize NaiveBayesModel transform closures"
by Ruifeng Zheng
· 26 hours ago
d26e144
[SPARK-59297][SQL][TEST] Cover nanosecond precisions 7 and 8 in RowToColumnConverter round-trip tests
by Stevo Mitric
· 28 hours ago
9a18c2f
[SPARK-59129][SQL] Assign a name to the error condition _LEGACY_ERROR_TEMP_2033
by Dieter Kling
· 29 hours ago
3de625e
[SPARK-58735][SQL] Prune nested fields when computing size of an array of structs
by hemanthboyina
· 29 hours ago
c6e2d26
[SPARK-59187][SQL] Compare partition key rows at types with the naming erased
by Peter Toth
· 31 hours ago
6ddc2e4
[SPARK-59314][K8S][DOC] Add `Heterogeneous Executor Management` section to `running-on-kubernetes.md`
by Dongjoon Hyun
· 31 hours ago
a2bbfea
[SPARK-59316][K8S][DOC] Remove the obsolete `-Pvolcano` build instruction from K8s docs
by Dongjoon Hyun
· 31 hours ago
f957459
[SPARK-59248][SQL][FOLLOWUP] Update a stale exchange count in a SPARK-59050 test
by Xiduo You
· 31 hours ago
ca8e43b
[MINOR] Improve resilence on raw socket stream and steer people towards higher level streaming APIs
by Holden Karau
· 8 days ago
0856811
[SPARK-59303][K8S] Support `spark.kubernetes.executor.resizeMaxMemory` in `ExecutorResizePlugin`
by Dongjoon Hyun
· 32 hours ago
4030f14
[SPARK-59304][K8S] Support `spark.kubernetes.executor.pvc.resizeMaxStorage` in `ExecutorPVCResizePlugin`
by Dongjoon Hyun
· 33 hours ago
cfbec6d
[SPARK-59248][SQL] Keep storage-partitioned join when partition keys are pruned from the scan output
by Xiduo You
· 2 days ago
9966a94
[SPARK-59261][SQL] Memoize outputPartitioning on the projection and SPJ grouping nodes
by Peter Toth
· 2 days ago
566555b
[SPARK-59263][BUILD] Upgrade `compress-lzf` to 1.2.1
by Dongjoon Hyun
· 2 days ago
0377b13
[SPARK-57938][DOC] Print active SKIP_ flags when Ruby doc build script runs
by Nicholas Chammas
· 2 days ago
4dc04af
[SPARK-59163][SQL] Reuse projected options and make CaseInsensitiveStringMap read-only
by yyanyy
· 2 days ago
dbe0462
[SPARK-59286][ML] Optimize NaiveBayesModel transform closures
by Ruifeng Zheng
· 2 days ago
e2ccc5f
[SPARK-58999][CORE] Allow a bare error class when it defines sub-classes
by Ala Luszczak
· 2 days ago
8ec12b1
Revert "[SPARK-59261][SQL] Memoize outputPartitioning on the projection and SPJ grouping nodes"
by Peter Toth
· 2 days ago
9373486
[SPARK-59261][SQL] Memoize outputPartitioning on the projection and SPJ grouping nodes
by Peter Toth
· 2 days ago
67054c0
[SPARK-59050][SQL] SPJ one-side shuffle with out-of-set keys produces wrong results in multi-joins
by Xiduo You
· 2 days ago
2c6777e
[SPARK-59048][ML][FOLLOWUP] Correct transform method documentation
by Ruifeng Zheng
· 2 days ago
3c24950
[SPARK-59284][SQL] Read INT64 TIMESTAMP(MICROS) columns as nanosecond timestamps
by Stevo Mitric
· 2 days ago
ff9d3e3
[SPARK-58985][CORE] Fix HistoryServerDiskManager double-counting store size on concurrent release and makeRoom
by Cheng Pan
· 2 days ago
8998b7a
[SPARK-59252][SQL] Fix an SPJ scan reporting a different partitioning at execution than at planning
by Peter Toth
· 2 days ago
4605662
[SPARK-59080][SQL] Pick one ShuffleSpecCollection member for the SPJ pushdown and the re-shuffle
by Peter Toth
· 2 days ago
0b5e7d8
[SPARK-58996][SQL] Fix SPJ partially clustered data correctness when EnsureRequirements re-runs
by Xiduo You
· 2 days ago
64eff66
[SPARK-59199][SQL] SPJ: Support runtime partition filtering for more join types
by Cheng Pan
· 2 days ago
f3e6e17
[SPARK-59250][SQL] Translate an empty runtime IN filter to AlwaysFalse in DSv2 runtime filtering
by Cheng Pan
· 2 days ago
da96091
[SPARK-58944][INFRA][FOLLOWUP] Handle manually completed cherry-picks
by Ruifeng Zheng
· 2 days ago
bc54e7a
[SPARK-59247][ML] Optimize tree ensemble classification model transform closures
by Ruifeng Zheng
· 2 days ago
f9b447d
[SPARK-59282][SQL][TEST] Reuse the shared nanosecond truncation helper in timestamp nanos tests
by Stevo Mitric
· 3 days ago
b20104d
[SPARK-59281][DOC] Add nanosecond-precision timestamp types to the SQL data types reference table
by Stevo Mitric
· 3 days ago
ea926c0
[MINOR][SQL] Remove stale "Hash is not implemented" comment for nanosecond timestamp physical types
by Stevo Mitric
· 3 days ago
f37971c
[SPARK-48701][SQL][PS] Make PandasMode collation-aware
by Vinod KC
· 3 days ago
431e8ae
[SPARK-59253][UDF] Fix direct dispatcher interrupt race
by Haiyang Sun
· 3 days ago
d7cb659
[SPARK-59011][PS] Fix descending rank first tie ordering
by NaVis
· 4 days ago
2a7cfea
[SPARK-59259][BUILD] Upgrade `netty-tcnative` to 2.0.83.Final
by Dongjoon Hyun
· 5 days ago
bce5ed1
[SPARK-59068][SQL][FOLLOWUP] Correct runtime filter validation and test fixtures
by Szehon Ho
· 5 days ago
3a0845a
[SPARK-58917][SQL] Add option to respect `inferSchema` when reading CSV as variant
by Greg Hansen
· 5 days ago
eb51129
[SPARK-58951][SQL] Support non-binary collations in collect_set
by Vinod KC
· 5 days ago
5e21a09
[MINOR][INFRA][DOC] Merge conflicting specification of PySpark Python version in `docs` job
by Nicholas Chammas
· 5 days ago
e261626
[SPARK-59202][ML][TEST] Enable save/load round-trip tests in Java ML tree suites
by Uros
· 5 days ago
b4259e4
[SPARK-59168][CORE] Avoid NoSuchElementException on missing keys in KVStore reads and writes
by Xiduo You
· 5 days ago
383c6f8
[SPARK-59154][ML] Optimize decision tree model transform closures
by Ruifeng Zheng
· 5 days ago
f8b673c
[SPARK-36082][SQL][FOLLOWUP] Preserve NAAJ broadcast fallback behavior
by Wenchen Fan
· 5 days ago
7b60b3e
[SPARK-59219][PYTHON] Avoid copying the Series to name it in convert_numpy
by Spenser Sun
· 5 days ago
d0bc986
[SPARK-59232][PS] Support boolean operands in NumPy ufuncs
by Spenser Sun
· 5 days ago
ad3b3b0
[SPARK-56922][SQL] Direct dispatcher to manage grpc-based udf server.
by Haiyang Sun
· 5 days ago
8549454
[SPARK-58467][SQL] Add scalar external UDF planning support
by Haiyang Sun
· 5 days ago
c809c28
[SPARK-59108][SQL] Fix Avro positional matching under column pruning
by YangJie
· 5 days ago
0ecd6b9
[SPARK-58752][SQL][PYTHON] Let a session set environment variables for its Python UDF workers
by Ramon Zhou
· 6 days ago
908903a
[SPARK-57417][SQL][PYTHON][FOLLOWUP] Fix the `variant_delete` version to 4.3.0
by Dongjoon Hyun
· 6 days ago
c31d201c
[SPARK-59201][BUILD] Cross-reference the Hadoop fallback version between pom.xml and IsolatedClientLoader
by Uros
· 6 days ago
680855b
[SPARK-59212][K8S][INFRA] Pin the Minikube container runtime to `docker` in K8s CI
by Ramon Zhou
· 6 days ago
59f6ffa
[SPARK-58876][SQL] Map Oracle DATE and TIMESTAMP to TimestampNTZType
by aleksandar-trajkovic-db
· 6 days ago
2682be3
[SPARK-59063][SQL] Use a byte-length guard in LikeSimplification for 'prefix%suffix'
by david-mollitor-db
· 6 days ago
3c70da5
[SPARK-59176][SQL] Fix a storage-partitioned join that fails when one side reduced onto no key
by Peter Toth
· 6 days ago
bcea2b2
[SPARK-57900][K8S][TEST] Add OIDC credential propagation E2E tests on Minikube with moto
by Kousuke Saruta
· 6 days ago
77f1e05
[PYTHON][MINOR] Improve TorchDistributor log server provisioning
by Holden Karau
· 9 days ago
51820eb
[SPARK-59172][CORE] Use isEmpty/nonEmpty instead of size comparisons in core
by Uros
· 6 days ago
a5a9e7b
[SPARK-59171][SQL] Make SchemaPruning idempotent after variant pushdown
by Goutam Adwant
· 6 days ago
0875765
[SPARK-59173][ML] Use isEmpty/nonEmpty instead of size comparisons in MLlib
by Uros
· 6 days ago
Next »