-
Notifications
You must be signed in to change notification settings - Fork 29.3k
Pull requests: apache/spark
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
[SPARK-58295][SQL][PYTHON] Add
to_base32 and from_base32 functions
#57466
opened Jul 23, 2026 by
SreeramaYeshwanthGowd
Loading…
[SPARK-58294][SQL] Fix Hive UDAF with two aggregation buffers in Complete mode
#57464
opened Jul 23, 2026 by
ulysses-you
Contributor
Loading…
[SPARK-58292][CORE] Recreate the netty worker EventLoopGroup when a worker event loop thread dies
#57462
opened Jul 23, 2026 by
ChuckLin2025
Contributor
Loading…
[WIP][INFRA] Add AI-based class-level test selection for post-merge CI
#57461
opened Jul 23, 2026 by
zhengruifeng
Contributor
•
Draft
[SPARK-58291][SQL] Guard statistical-aggregate merge against empty-buffer overflow to NaN
#57460
opened Jul 23, 2026 by
ulysses-you
Contributor
Loading…
[SPARK-58249][SPARK-58234][PS][TEST][FOLLOWUP] Test native NumPy ufuncs with special values
#57458
opened Jul 23, 2026 by
zhengruifeng
Contributor
Loading…
[SPARK-58283][SQL] Assign names to the error conditions _LEGACY_ERROR_TEMP_1138-1141
#57456
opened Jul 23, 2026 by
LuciferYang
Contributor
Loading…
[SPARK-52825][SQL] Register existing dialects for additional URL prefixes
#57455
opened Jul 23, 2026 by
cloud-fan
Contributor
Loading…
[SPARK-58277][SQL] Stream DataType.json to bound peak memory for large schemas
#57454
opened Jul 23, 2026 by
bhollis-dbx
Loading…
[SPARK-58277][SQL] Stream DataType JSON serialization
#57453
opened Jul 23, 2026 by
marcuslin123
Contributor
Loading…
[SPARK-58282][PYTHON][DOCS] Refresh PySpark README and clarify project scope
#57452
opened Jul 23, 2026 by
nchammas
Contributor
Loading…
[SPARK-58281][ML][CONNECT] Avoid parent overcounting in PipelineModel size estimates
#57451
opened Jul 23, 2026 by
zhengruifeng
Contributor
Loading…
[SPARK-58275][SQL][PYTHON] Add
normalize Unicode normalization SQL function
#57450
opened Jul 23, 2026 by
SreeramaYeshwanthGowd
Loading…
[SPARK-58279][ML][CONNECT] Estimate size of TargetEncoder, VectorIndexer, CountVectorizer, and MinHashLSH models
#57448
opened Jul 23, 2026 by
zhengruifeng
Contributor
Loading…
[SPARK-58276][PYTHON] Consolidate serializer selection branches in worker.py
#57446
opened Jul 23, 2026 by
Yicong-Huang
Contributor
Loading…
[SPARK-58267][SQL] Assign a name to the error condition _LEGACY_ERROR_TEMP_1059
#57445
opened Jul 22, 2026 by
Ma77Ball
Contributor
Loading…
[WIP][SPARK-57378][SDP] Implement SCD2 Batch Processor; Merge Reconciled Rows into Aux and Target Tables
#57444
opened Jul 22, 2026 by
anew
Contributor
Loading…
[SPARK-58272][SQL] Enable runtime Bloom filters for materialized cached inputs
#57443
opened Jul 22, 2026 by
sunchao
Member
Loading…
[SPARK-52246][SQL][TESTS] Add bucket transform regression test for one-side shuffle with join key tail of partition keys
#57442
opened Jul 22, 2026 by
naveenp2708
Contributor
Loading…
[SPARK-58266][SQL] Assign a name to the error condition _LEGACY_ERROR_TEMP_1058
#57440
opened Jul 22, 2026 by
Ma77Ball
Contributor
Loading…
[SPARK-58269][SQL] Infer generated column partition filters
#57439
opened Jul 22, 2026 by
szehon-ho
Member
Loading…
[SPARK-58265][SQL] Reuse projected broadcast values for dynamic partition pruning
#57437
opened Jul 22, 2026 by
sunchao
Member
Loading…
[SPARK-54946][PYTHON][TEST] Add tests for pa.Array.to_pandas with coerce_temporal_nanoseconds
#57435
opened Jul 22, 2026 by
Spenserrrr
Contributor
Loading…
Previous Next
ProTip!
Type g p on any issue or pull request to go back to the pull request listing page.