Messages by Thread
-
[jira] [Assigned] (SPARK-58580) Use native Spark function for NumPy ldexp
Ruifeng Zheng (Jira)
-
[jira] [Assigned] (SPARK-58508) Activate pipelined shuffles for Real-Time Mode streaming queries to support Stateful Queries
Wenchen Fan (Jira)
-
[jira] [Resolved] (SPARK-58508) Activate pipelined shuffles for Real-Time Mode streaming queries to support Stateful Queries
Wenchen Fan (Jira)
-
[jira] [Created] (SPARK-58584) Optimize ML implementations for Spark Connect server memory usage
Ruifeng Zheng (Jira)
-
[jira] [Updated] (SPARK-58583) Construct squared Euclidean silhouette stats on executors
ASF GitHub Bot (Jira)
-
[jira] [Created] (SPARK-58583) Construct squared Euclidean silhouette stats on executors
Ruifeng Zheng (Jira)
-
[jira] [Created] (SPARK-58582) Use native Spark function for NumPy fmod
Ruifeng Zheng (Jira)
-
[jira] [Created] (SPARK-58581) Use native Spark function for NumPy floor_divide
Ruifeng Zheng (Jira)
-
[jira] [Resolved] (SPARK-58548) Use native Spark function for NumPy heaviside
Ruifeng Zheng (Jira)
-
[jira] [Assigned] (SPARK-58548) Use native Spark function for NumPy heaviside
Ruifeng Zheng (Jira)
-
[jira] [Resolved] (SPARK-58295) Add to_base32 and from_base32 built-in functions
Wenchen Fan (Jira)
-
[jira] [Assigned] (SPARK-58295) Add to_base32 and from_base32 built-in functions
Wenchen Fan (Jira)
-
[jira] [Updated] (SPARK-57487) Support distributed map join for medium-sized build tables via SQL hint
Yu Gan (Jira)
-
[jira] [Resolved] (SPARK-57591) [SQL] Read and infer ORC schema from archives
Wenchen Fan (Jira)
-
[jira] [Updated] (SPARK-58580) Use native Spark function for NumPy ldexp
ASF GitHub Bot (Jira)
-
[jira] [Created] (SPARK-58580) Use native Spark function for NumPy ldexp
Ruifeng Zheng (Jira)
-
[jira] [Resolved] (SPARK-58181) Improve connect ML model size estimation
Ruifeng Zheng (Jira)
-
[jira] [Resolved] (SPARK-58248) [SQL] Reuse single line inference path for JSON/XML archive inference
Wenchen Fan (Jira)
-
[jira] [Resolved] (SPARK-58553) Use native Spark functions for NumPy fmax and fmin
Ruifeng Zheng (Jira)
-
[jira] [Assigned] (SPARK-58553) Use native Spark functions for NumPy fmax and fmin
Ruifeng Zheng (Jira)
-
[jira] [Assigned] (SPARK-58557) Reuse ML vector data type singleton
Ruifeng Zheng (Jira)
-
[jira] [Resolved] (SPARK-58557) Reuse ML vector data type singleton
Ruifeng Zheng (Jira)
-
[jira] [Resolved] (SPARK-58579) Support `TIME` type in `SparkSession.createDataFrame`
Dongjoon Hyun (Jira)
-
[jira] [Updated] (SPARK-58579) Support `TIME` type in `SparkSession.createDataFrame`
ASF GitHub Bot (Jira)
-
[jira] [Assigned] (SPARK-58579) Support `TIME` type in `SparkSession.createDataFrame`
Dongjoon Hyun (Jira)
-
[jira] [Created] (SPARK-58579) Support `TIME` type in `SparkSession.createDataFrame`
Dongjoon Hyun (Jira)
-
[jira] [Updated] (SPARK-57937) Improve decimal division performance by dividing at the result scale
Dongjoon Hyun (Jira)
-
[jira] [Updated] (SPARK-56677) Propagate filter conditions through Join nodes in PlanMerger
Dongjoon Hyun (Jira)
-
[jira] [Updated] (SPARK-40193) Merge subquery plans with different filters
Dongjoon Hyun (Jira)
-
[jira] [Created] (SPARK-58578) Mark ApplyFunctionExpression stateful to prevent reusedRow data races under concurrent evaluation
Vinod KC (Jira)
-
[jira] [Updated] (SPARK-58577) Inline wrap_udf into the batched UDF dispatch branch in worker.py
ASF GitHub Bot (Jira)
-
[jira] [Created] (SPARK-58577) Inline wrap_udf into the batched UDF dispatch branch in worker.py
Yicong Huang (Jira)
-
[jira] [Resolved] (SPARK-58204) Fix unprotected mutable state in RegExpReplace, RegExpExtractBase, StringTranslate, FormatNumber, and NamedLambdaVariable by marking them stateful
Hyukjin Kwon (Jira)
-
[jira] [Assigned] (SPARK-58204) Fix unprotected mutable state in RegExpReplace, RegExpExtractBase, StringTranslate, FormatNumber, and NamedLambdaVariable by marking them stateful
Hyukjin Kwon (Jira)
-
[jira] [Resolved] (SPARK-58205) Mark JSON/CSV/XML expressions stateful to prevent shared-evaluator data races
Hyukjin Kwon (Jira)
-
[jira] [Assigned] (SPARK-58205) Mark JSON/CSV/XML expressions stateful to prevent shared-evaluator data races
Hyukjin Kwon (Jira)
-
[jira] [Assigned] (SPARK-58389) Pass all options while loading tables
Anton Okolnychyi (Jira)
-
[jira] [Resolved] (SPARK-58389) Pass all options while loading tables
Anton Okolnychyi (Jira)
-
[jira] [Updated] (SPARK-58576) Add tests for pa.Array.to_pandas with integer_object_nulls
Haotian Sun (Jira)
-
[jira] [Created] (SPARK-58576) Add tests for pa.Array.to_pandas with integer_object_nulls
Haotian Sun (Jira)
-
[jira] [Updated] (SPARK-58575) Inline redundant error_class parameter in worker.py verify helpers
Yicong Huang (Jira)
-
[jira] [Created] (SPARK-58575) Inline redundant error_class parameter in worker.py verify helpers
Yicong Huang (Jira)
-
[jira] [Updated] (SPARK-58573) Replace defunct Jenkins test grouping with modern GitHub Actions grouping
ASF GitHub Bot (Jira)
-
[jira] [Updated] (SPARK-58569) Add documentation for Auto CDC (SCD Type 1)
ASF GitHub Bot (Jira)
-
[jira] [Resolved] (SPARK-58574) Support `TIME` type literals in `lit` and SQL parameter binding
Dongjoon Hyun (Jira)
-
[jira] [Assigned] (SPARK-58574) Support `TIME` type literals in `lit` and SQL parameter binding
Dongjoon Hyun (Jira)
-
[jira] [Updated] (SPARK-58574) Support `TIME` type literals in `lit` and SQL parameter binding
ASF GitHub Bot (Jira)
-
[jira] [Created] (SPARK-58574) Support `TIME` type literals in `lit` and SQL parameter binding
Dongjoon Hyun (Jira)
-
[jira] [Resolved] (SPARK-58565) Some operator Helm chart RBAC tests don't run due to misconfiguration
Dongjoon Hyun (Jira)
-
[jira] [Assigned] (SPARK-58565) Some operator Helm chart RBAC tests don't run due to misconfiguration
Dongjoon Hyun (Jira)
-
[jira] [Created] (SPARK-58573) Replace defunct Jenkins test grouping with modern GitHub Actions grouping
Nicholas Chammas (Jira)
-
[jira] [Assigned] (SPARK-58529) Unify RESULT_ROWS_MISMATCH message and consolidate row count verification in worker.py
Yicong Huang (Jira)
-
[jira] [Resolved] (SPARK-58529) Unify RESULT_ROWS_MISMATCH message and consolidate row count verification in worker.py
Yicong Huang (Jira)
-
[jira] [Assigned] (SPARK-58567) Use nonEmpty/isEmpty instead of length comparisons in StateDataSource
Jira
-
[jira] [Resolved] (SPARK-58567) Use nonEmpty/isEmpty instead of length comparisons in StateDataSource
Jira
-
[jira] [Created] (SPARK-58572) Add Cross-SCD Fuzz Testing
Anish Mahto (Jira)
-
[jira] [Assigned] (SPARK-58571) Support `TIME` type values in `DataFrame.collect`
Dongjoon Hyun (Jira)
-
[jira] [Resolved] (SPARK-58571) Support `TIME` type values in `DataFrame.collect`
Dongjoon Hyun (Jira)
-
[jira] [Updated] (SPARK-58571) Support `TIME` type values in `DataFrame.collect`
ASF GitHub Bot (Jira)
-
[jira] [Created] (SPARK-58570) Add documentation for Auto CDC (SCD Type 2)
Andreas Neumann (Jira)
-
[jira] [Created] (SPARK-58571) Support `TIME` type values in `DataFrame.collect`
Dongjoon Hyun (Jira)
-
[jira] [Created] (SPARK-58569) Add documentation for Auto CDC (SCD Type 1)
Andreas Neumann (Jira)
-
[jira] [Assigned] (SPARK-58568) Support `TIME` type in `DataType.simpleString`
Dongjoon Hyun (Jira)
-
[jira] [Resolved] (SPARK-58568) Support `TIME` type in `DataType.simpleString`
Dongjoon Hyun (Jira)
-
[jira] [Updated] (SPARK-58568) Support `TIME` type in `DataType.simpleString`
ASF GitHub Bot (Jira)
-
[jira] [Created] (SPARK-58568) Support `TIME` type in `DataType.simpleString`
Dongjoon Hyun (Jira)
-
[jira] [Resolved] (SPARK-58563) Support server-side error chain in `SparkConnectError` via `FetchErrorDetails`
Dongjoon Hyun (Jira)
-
[jira] [Commented] (SPARK-58567) Use nonEmpty/isEmpty instead of length comparisons in StateDataSource
Jira
-
[jira] [Updated] (SPARK-58567) Use nonEmpty/isEmpty instead of length comparisons in StateDataSource
ASF GitHub Bot (Jira)
-
[jira] [Created] (SPARK-58567) Use nonEmpty/isEmpty instead of length comparisons in StateDataSource
Jira
-
[jira] [Created] (SPARK-58566) Add tests for QuotingUtils.quoteIdentifier and escapeSingleQuotedString
Jira
-
[jira] [Updated] (SPARK-58566) Add tests for QuotingUtils.quoteIdentifier and escapeSingleQuotedString
ASF GitHub Bot (Jira)
-
[jira] [Commented] (SPARK-58566) Add tests for QuotingUtils.quoteIdentifier and escapeSingleQuotedString
Jira
-
[jira] [Updated] (SPARK-58565) Some operator Helm chart RBAC tests don't run due to misconfiguration
ASF GitHub Bot (Jira)
-
[jira] [Updated] (SPARK-58562) Document subplan merging in the SQL performance tuning guide
ASF GitHub Bot (Jira)
-
[jira] [Resolved] (SPARK-58435) Use named containsNull argument in KLL sketch ArrayType results
Jira
-
[jira] [Assigned] (SPARK-58435) Use named containsNull argument in KLL sketch ArrayType results
Jira
-
[jira] [Created] (SPARK-58565) Some operator Helm chart RBAC tests don't run due to misconfiguration
Nicholas Barton (Jira)
-
[jira] [Created] (SPARK-58564) Some operator Helm chart RBAC tests don't run due to misconfiguration
Nicholas Barton (Jira)
-
[jira] [Assigned] (SPARK-58563) Support server-side error chain in `SparkConnectError` via `FetchErrorDetails`
Dongjoon Hyun (Jira)
-
[jira] [Closed] (SPARK-51530) Client should support re-attachable execution
Dongjoon Hyun (Jira)
-
[jira] [Resolved] (SPARK-51530) Client should support re-attachable execution
Dongjoon Hyun (Jira)
-
[jira] [Resolved] (SPARK-58560) Use Apache Spark `4.0.4` and `4.1.3` in CI
Dongjoon Hyun (Jira)
-
[jira] [Assigned] (SPARK-58560) Use Apache Spark `4.0.4` and `4.1.3` in CI
Dongjoon Hyun (Jira)
-
[jira] [Updated] (SPARK-58557) Reuse ML vector data type singleton
ASF GitHub Bot (Jira)
-
[jira] [Updated] (SPARK-58563) Support server-side error chain in `SparkConnectError` via `FetchErrorDetails`
ASF GitHub Bot (Jira)
-
[jira] [Created] (SPARK-58563) Support server-side error chain in `SparkConnectError` via `FetchErrorDetails`
Dongjoon Hyun (Jira)
-
[jira] [Created] (SPARK-58562) Document subplan merging in the SQL performance tuning guide
Peter Toth (Jira)
-
[jira] [Updated] (SPARK-58558) Make requireAllClusterKeysForCoPartition check key coverage instead of exact match for SPJ
Cheng Pan (Jira)
-
[jira] [Updated] (SPARK-58561) Support backporting a PR merged into a non-default branch
Jira
-
[jira] [Commented] (SPARK-58561) Support backporting a PR merged into a non-default branch in merge_spark_pr.py
Jira
-
[jira] [Created] (SPARK-58561) Support backporting a PR merged into a non-default branch in merge_spark_pr.py
Jira
-
[jira] [Updated] (SPARK-58560) Use Apache Spark `4.0.4` and `4.1.3` in CI
ASF GitHub Bot (Jira)
-
[jira] [Created] (SPARK-58560) Use Apache Spark `4.0.4` and `4.1.3` in CI
Dongjoon Hyun (Jira)
-
[jira] [Resolved] (SPARK-58543) Support reattachable execution with `ReattachExecute` and `ReleaseExecute`
Dongjoon Hyun (Jira)
-
[jira] [Reopened] (SPARK-58404) Support bypassing partial WindowGroupLimit
Dongjoon Hyun (Jira)
-
[jira] [Updated] (SPARK-58559) Package all of sbin in PySpark classic
ASF GitHub Bot (Jira)
-
[jira] [Created] (SPARK-58559) Package all of sbin in PySpark classic
Nicholas Chammas (Jira)
-
[jira] [Updated] (SPARK-58558) Remove requireAllClusterKeysForCoPartition as a gate for SPJ
Cheng Pan (Jira)
-
[jira] [Created] (SPARK-58558) Remove requireAllClusterKeysForCoPartition as a gate for SPJ
Cheng Pan (Jira)
-
[jira] [Assigned] (SPARK-58546) Avoid redundant StringIndexer label lookups
Ruifeng Zheng (Jira)
-
[jira] [Resolved] (SPARK-58546) Avoid redundant StringIndexer label lookups
Ruifeng Zheng (Jira)
-
[jira] [Resolved] (SPARK-58394) [SQL] Use ZipFile instead of ZipArchiveInputStream
Wenchen Fan (Jira)
-
[jira] [Assigned] (SPARK-58366) [SQL] Support JSON_TABLE table-valued function
Wenchen Fan (Jira)
-
[jira] [Resolved] (SPARK-58366) [SQL] Support JSON_TABLE table-valued function
Wenchen Fan (Jira)
-
[jira] [Resolved] (SPARK-58538) Add branch-4.3 CI scheduler and release integration
Jungtaek Lim (Jira)
-
[jira] [Assigned] (SPARK-58538) Add branch-4.3 CI scheduler and release integration
Jungtaek Lim (Jira)
-
[jira] [Assigned] (SPARK-58537) Scheduled Maven workflows on branch-4.x build master instead of the branch
Jungtaek Lim (Jira)
-
[jira] [Resolved] (SPARK-58537) Scheduled Maven workflows on branch-4.x build master instead of the branch
Jungtaek Lim (Jira)
-
[jira] [Resolved] (SPARK-58554) Syntax error when pushing CAST to DOUBLE via MariaDB
Wenchen Fan (Jira)
-
[jira] [Updated] (SPARK-58554) Syntax error when pushing CAST to DOUBLE via MariaDB
ASF GitHub Bot (Jira)
-
[jira] [Assigned] (SPARK-58554) Syntax error when pushing CAST to DOUBLE via MariaDB
Wenchen Fan (Jira)
-
[jira] [Resolved] (SPARK-58359) Require the arguments field in built-in function documentation
Hyukjin Kwon (Jira)
-
[jira] [Assigned] (SPARK-58359) Require the arguments field in built-in function documentation
Hyukjin Kwon (Jira)
-
[jira] [Updated] (SPARK-58556) Show ML cache status in Spark Connect UI
Ruifeng Zheng (Jira)
-
[jira] [Created] (SPARK-58557) Reuse ML vector data type singleton
Ruifeng Zheng (Jira)
-
[jira] [Created] (SPARK-58556) Show ML cache status in Spark Connect UI
Ruifeng Zheng (Jira)
-
[jira] [Created] (SPARK-58555) Add additional tests for JDBC data source join pushdown
Uros Stankovic (Jira)
-
[jira] [Updated] (SPARK-58555) Add additional tests for JDBC data source join pushdown
Uros Stankovic (Jira)
-
[jira] [Created] (SPARK-58554) Syntax error when pushing CAST to DOUBLE via MariaDB
Alek Jarmov (Jira)
-
[jira] [Updated] (SPARK-58553) Use native Spark functions for NumPy fmax and fmin
ASF GitHub Bot (Jira)
-
[jira] [Assigned] (SPARK-58533) Use native Spark function for NumPy float_power
Ruifeng Zheng (Jira)
-
[jira] [Resolved] (SPARK-58533) Use native Spark function for NumPy float_power
Ruifeng Zheng (Jira)
-
[jira] [Created] (SPARK-58553) Use native Spark functions for NumPy fmax and fmin
Ruifeng Zheng (Jira)
-
[jira] [Updated] (SPARK-58551) Python Data Sources Limit Pushdown API
ASF GitHub Bot (Jira)
-
[jira] [Updated] (SPARK-58552) Add total task time column to the SQL / DataFrame tab
ASF GitHub Bot (Jira)
-
[jira] [Created] (SPARK-58552) Add total task time column to the SQL / DataFrame tab
XiDuo You (Jira)
-
[jira] [Created] (SPARK-58551) Python Data Sources Limit Pushdown API
Ganesha S (Jira)
-
[jira] [Updated] (SPARK-58550) Delay GaussianMixture aggregation allocations
ASF GitHub Bot (Jira)
-
[jira] [Created] (SPARK-58550) Delay GaussianMixture aggregation allocations
Ruifeng Zheng (Jira)
-
[jira] [Updated] (SPARK-58548) Use native Spark function for NumPy heaviside
ASF GitHub Bot (Jira)
-
[jira] [Updated] (SPARK-58549) Preserve key-grouped partitioning and ordering across a DSv2 scan merge
Peter Toth (Jira)
-
[jira] [Created] (SPARK-58549) Preserve key-grouped partitioning and ordering across a DSv2 scan merge
Peter Toth (Jira)
-
[jira] [Updated] (SPARK-36284) Add shuffle checksum support for push-based shuffle
ASF GitHub Bot (Jira)
-
[jira] [Updated] (SPARK-58547) Expose Spark Connect operation IDs for end-to-end request attribution
ASF GitHub Bot (Jira)
-
[jira] [Created] (SPARK-58548) Use native Spark function for NumPy heaviside
Ruifeng Zheng (Jira)
-
[jira] [Created] (SPARK-58547) Expose Spark Connect operation IDs for end-to-end request attribution
Wenchen Fan (Jira)
-
[jira] [Updated] (SPARK-58544) Vector distance and norm functions return NaN, Infinity or NULL due to float overflow in intermediate sums
ASF GitHub Bot (Jira)
-
[jira] [Updated] (SPARK-58546) Avoid redundant StringIndexer label lookups
ASF GitHub Bot (Jira)
-
[jira] [Created] (SPARK-58546) Avoid redundant StringIndexer label lookups
Ruifeng Zheng (Jira)
-
[jira] [Created] (SPARK-58545) Add tests for pa.Array.to_pandas with types_mapper
Haotian Sun (Jira)
-
[jira] [Updated] (SPARK-58545) Add tests for pa.Array.to_pandas with types_mapper
Haotian Sun (Jira)
-
[jira] [Assigned] (SPARK-58543) Support reattachable execution with `ReattachExecute` and `ReleaseExecute`
Dongjoon Hyun (Jira)
-
[jira] [Created] (SPARK-58544) Vector distance and norm functions return NaN, Infinity or NULL due to float overflow in intermediate sums
sepuri sai krishna (Jira)
-
[jira] [Updated] (SPARK-58543) Support reattachable execution with `ReattachExecute` and `ReleaseExecute`
ASF GitHub Bot (Jira)
-
[jira] [Created] (SPARK-58543) Support reattachable execution with `ReattachExecute` and `ReleaseExecute`
Dongjoon Hyun (Jira)
-
[jira] [Assigned] (SPARK-58535) Create StringIndexer lookup maps during transformation
Ruifeng Zheng (Jira)
-
[jira] [Resolved] (SPARK-58535) Create StringIndexer lookup maps during transformation
Ruifeng Zheng (Jira)
-
[jira] [Resolved] (SPARK-58534) Change version in branch-4.x to 4.4.0
Dongjoon Hyun (Jira)
-
[jira] [Assigned] (SPARK-58534) Change version in branch-4.x to 4.4.0
Dongjoon Hyun (Jira)
-
[jira] [Resolved] (SPARK-57225) Throw informative error on attempt to write using streaming StateDataSource or StateMetadataSource
Wenchen Fan (Jira)
-
[jira] [Resolved] (SPARK-58542) Upgrade `gRPC Swift NIO Transport` to 2.9.1
Dongjoon Hyun (Jira)
-
[jira] [Assigned] (SPARK-58542) Upgrade `gRPC Swift NIO Transport` to 2.9.1
Dongjoon Hyun (Jira)
-
[jira] [Assigned] (SPARK-57225) Throw informative error on attempt to write using streaming StateDataSource or StateMetadataSource
Wenchen Fan (Jira)
-
[jira] [Resolved] (SPARK-58541) Support retry with exponential backoff in `SparkConnectClient`
Dongjoon Hyun (Jira)
-
[jira] [Assigned] (SPARK-58541) Support retry with exponential backoff in `SparkConnectClient`
Dongjoon Hyun (Jira)
-
[jira] [Assigned] (SPARK-58423) Raise MISSING_ATTRIBUTES for missing input after single-pass hidden-output insertion
Wenchen Fan (Jira)
-
[jira] [Resolved] (SPARK-58423) Raise MISSING_ATTRIBUTES for missing input after single-pass hidden-output insertion
Wenchen Fan (Jira)
-
[jira] [Resolved] (SPARK-58411) Fix bug for decorrelating Limit+Sort when partitionFields is empty
Wenchen Fan (Jira)
-
[jira] [Assigned] (SPARK-58411) Fix bug for decorrelating Limit+Sort when partitionFields is empty
Wenchen Fan (Jira)
-
[jira] [Assigned] (SPARK-58525) StructType.merge does not propagate the caseSensitive flag into nested struct/array/map field merges
Wenchen Fan (Jira)
-
[jira] [Resolved] (SPARK-58525) StructType.merge does not propagate the caseSensitive flag into nested struct/array/map field merges
Wenchen Fan (Jira)
-
[jira] [Updated] (SPARK-58542) Upgrade `gRPC Swift NIO Transport` to 2.9.1
ASF GitHub Bot (Jira)
-
[jira] [Resolved] (SPARK-57640) Implement gRPC backend WorkerSession
Yicong Huang (Jira)
-
[jira] [Created] (SPARK-58542) Upgrade `gRPC Swift NIO Transport` to 2.9.1
Dongjoon Hyun (Jira)
-
[jira] [Assigned] (SPARK-57640) Implement gRPC backend WorkerSession
Yicong Huang (Jira)
-
[jira] [Updated] (SPARK-58541) Support retry with exponential backoff in `SparkConnectClient`
ASF GitHub Bot (Jira)
-
[jira] [Created] (SPARK-58541) Support retry with exponential backoff in `SparkConnectClient`
Dongjoon Hyun (Jira)
-
[jira] [Resolved] (SPARK-54945) Add tests for pa.Array.to_pandas with zero copy
Ruifeng Zheng (Jira)