StarRocks version 3.3
- After upgrading StarRocks to v3.3, DO NOT downgrade it directly to v3.2.0, v3.2.1, or v3.2.2, otherwise it will cause metadata loss. You must downgrade the cluster to v3.2.3 or later to prevent the issue.
- After upgrading StarRocks to v3.3.9, you can only downgrade it to v3.2.11 or later.
3.3.22
Release Date: January 27, 2026
Bug Fixes
The following issues have been fixed:
- CVE-2025-27818. #67335
- SIGSEGV crash in CN when querying non-partitioned Iceberg tables with DATE/TIME predicates on ARM64/Graviton architectures. #66864
- Deadlock caused by lock ordering issues when closing
LocalTabletsChannelandLakeTabletsChannel. #66748 - Potential BE crash when executing
CACHE SELECTqueries with filter conditions. #67375 - Issue where the Multicast Sink Operator could get stuck in the
OUTPUT_FULLstate if an upstream operator (for example, LIMIT) finished early, causing the query to hang. #67153 - Potential Segfault caused by failure to invalidate cache pointers when
ObjectColumnis resized or moved. #66957 - Potential errors or OOM issues when Java UDFs handle Nullable columns containing all NULLs. #67025
- BE crash caused by the optimization logic of Ranking Window Functions when
PARTITION BYandORDER BYare missing. #67081 - Incorrect result issue where
COUNT(DISTINCT)was not correctly rewritten tomulti_distinct_countwhen queried alongside non-distinct aggregations (like SUM) on a single-bucket table. #66767 - Incorrect results when
regexp_replaceprocesses multiple rows withenable_hyperscan_vecenabled. #67380 - Potential incorrect results with Sorted Streaming Aggregate in shared-data mode. #67376
- Query failure where the optimizer generated access paths using the old column name after the column was renamed. #67533
- Incorrect Bitmap column type propagation in the rewrite rule from
bitmap_to_arraytounnest_bitmap. #66855 - Dependency derivation error in Low Cardinality optimization logic by adopting the Union-Find algorithm to correctly handle column relationships. #66724
- "Compute node not found" error in Short-circuit Read under shared-data clusters by adding a fallback mechanism to non-short-circuit mode. #67323
- "Version not found" error during Replication publishing caused by FE Replicas not updating the minimum readable version. #67538
- Logic error in physical partition comparison during replication transactions to ensure deterministic comparison using ID order. #67616
- Inaccurate statistics for counters like scan rows in Cloud Native Tables. #67307
- Issue where expired Tablets were not cleaned up from the scheduler, causing the scheduling queue to pile up. #66718
- Inaccurate SQL statements displayed in the Profile when multiple statements are submitted. #67097
- Issue where the transaction ID was empty in
publish_versionlogs in the new FE. #66732 - Performance issue caused by unnecessary Protobuf message copying after
set_allocated. #67844
3.3.21
Release Date: December 25, 2025
Bug Fixes
The following issues have been fixed:
- Logic errors when the
trimfunction handles specific Unicode whitespace characters (for example,\u1680) and performance issues caused by reserved memory calculation. #66428 #66477 - Foreign key constraints are lost after FE restart due to table loading order. #66474
- Security vulnerabilities CVE-2025-66566 and CVE-2025-12183 in
lz4-java, and potential crashes in upstream. #66453 #66362 #67075 - Incorrect results when Join is used with window functions in Group Execution mode. #66441
- System continues attempting to fetch metadata for a deleted warehouse, causing
SHOW LOADor SQL execution failures. #66436 PartitionColumnMinMaxRewriteRuleoptimization incorrectly returns an empty set instead of NULL when the Scan input for aggregation is fully filtered. #66356- Rowset IDs are not properly released when Rowset Commit or Compaction fails, preventing disk space reclamation. #66301
- Missing scan statistics (for example, scanned rows/bytes) in audit logs when a high-selectivity filter causes the Scan to end early (EOS). #66280
- BE continues to respond to heartbeats as Alive after entering the crash handling process (for example, SIGSEGV), causing the FE to continue dispatching queries and reporting errors #66212
- BE crash caused by multiple calls to
set_collectorof the Runtime Filter due to Local TopN pushdown optimization. #66199 - Load task failures caused by un-initialized
rssidwhen column-mode Partial Update is used with Conditional Update. #66139 - Race condition when submitting drivers in Colocate Execution Group, leading to BE crashes. #66099
MemoryScratchSinkOperatorremains in a pending state and cannot be cancelled afterRecordBatchQueueis closed (For example, triggered by SparkSQL Limit), causing queries to hang. #66041- Use-after-free issue caused by a race condition when accessing
_num_pipelinesin theExecutionGroupcountdown logic. #65940 - Incorrect calculation logic for query error rate monitoring metrics (Internal/Analysis/Timeout error rate), resulting in negative values. #65891
- Null Pointer Exception (NPE) caused by unset
ConnectContextwhen executing tasks as an LDAP user. #65843 - Performance issues where Filesystem Cache lookups for the same Key fail because object reference comparison is used instead of value comparison. #65823
- Tablet-related files are not correctly cleaned up after a snapshot load failure due to incorrect status variable checking. #65709
- Compression type and level configured in FE are not correctly propagated to BE or persisted during table creation or Schema Change, resulting in default compression settings. #65673
COM_STMT_EXECUTEhas no audit logs by default, and Profile information is incorrectly merged into the Prepare stage. #65448- Delete Vector CRC32 check failures during cluster upgrade/downgrade scenarios due to version incompatibility. #65442
- BE crash caused by improper handling of Nullable properties when
UnionConstSourceOperatormerges Union to Values. #65429 - Concurrent load tasks fail during the Commit phase if the target tablet is dropped during an ALTER TABLE operation. #65396
- Inaccurate error log information when authentication fails due to incorrect context setting if users are switched in the HTTP SQL interface. #65371
- Statistics collection issues after
INSERT OVERWRITE, including failure to collect statistics for temporary partitions and inaccurate row count statistics at transaction commit preventing collection trigger. #65327, #65298, #65225 HttpConnectContextrelated to SQL is not released in TCP connection reuse scenarios because subsequent non-SQL HTTP requests overwrite the context. #65203- Potential version loss caused by incomplete tablet loading when RocksDB iteration times out during BE startup metadata loading. #65146
- BE crash caused by out-of-bounds access when processing compression parameters in the
percentile_approx_weightedfunction. #64838 - Missing Query Profile logs for queries forwarded from Follower to Leader. #64395
- BE crash caused by lax size checks during LZ4 encoding of large string columns during spilling. #61495
MERGING-EXCHANGEoperator crash caused by Ranking window function optimization generating an emptyORDER BYwhen there is noPARTITION BYandGROUP BY. #67081- Unstable results or errors in the Low Cardinality column rewrite logic due to dependence on the Set iteration order. #66724
3.3.20
Release Date: November 18, 2025
Bug Fixes
The following issues have been fixed:
- CVE-2024-47561. #64193
- CVE-2025-59419. #64142
- Incorrect row count for lake Primary Key tables. #64007
- Window function with IGNORE NULLS flags can not be consolidated with its counterpart without IGNORE NULLS flag. #63958
- ASAN error in
PartitionedSpillerWriter::_remove_partition. #63903 - Wrong results for sorted aggregation in shared-data clusters. #63849
- NPE when creating a partitioned materialized view. #63830
- Partitioned Spill crash when removing partitions. #63825
- NPE when removing expired load jobs in FE. #63820
- A potential deadlock during initialization of
ExceptionStackContext. #63776 - Degraded scan performance caused by the profitless simplification of CASE WHEN with complex functions. #63732
- Materialized view rewrite failures caused by type mismatch. #63659
- Materialized view rewrite throws
IllegalStateExceptionunder certain plans. #63655 - LZ4 compression and decompression errors cannot be perceived. #63629
- Stability issue caused by incorrect overflow detection when casting LARGEINT to DECIMAL128 at sign-edge cases (for example, INT128_MIN) #63559
date_truncpartition pruning with combined predicates that mistakenly produced EMPTYSET. #63464- Incomplete
Left Joinresults caused by ARRAY low-cardinality optimization. #63419 - An issue caused by the aggregate intermediate type uses
ARRAY<NULL_TYPE>. #63371 - Metadata inconsistency in partial updates based on auto-increment columns. #63370
- Incompatible Bitmap index reuse for Fast Schema Evolution in shared-data clusters. #63315
- Unnecessary CN deregistration during pod restart/upgrade. #63085
- Profiles showing SQL as
omitfor returns of the PREPARE/EXECUTE statements. #62988
3.3.19
Release Date: October 14, 2025
Bug Fixes
The following issues have been fixed:
UserPropertyhad lower priority than Session Variables. #63173- Materialized view refresh failures that could occur when the Hive base table was dropped and recreated. #63072
- Issues with the aggregation pushdown rewrite rule. #63060
- Inconsistencies between null columns and data columns in Boolean extraction functions for JSON. #63054
- Issues when getting partition columns in Delta Lake format tables. #62953
- Lack of colocation support for materialized views in shared-data clusters. #62941
- Projection mapping errors in view-based materialized view rewrite. #62918
- SQL syntax errors in histogram statistics when Most Common Values (MCV) contained single quotes. #62853
KILL ANALYZEdid not work. #62842- CVE-2025-58056 vulnerability. #62801
- Executing
SHOW CREATE ROUTINE LOADwithout specifying a database causes wrong results. #62745 - Data loss caused by incorrectly skipping CSV headers in
files(). #62719 - Version check failures when Replication and Compaction transactions were committed together. #62663
- Materialized view refresh is skipped because the materialized view version map is not cleared after a failed restore job. #62634
- Issues caused by case-sensitive partition column validation in the materialized view analyzer. #62598
3.3.18
Release Date: August 28, 2025
Bug Fixes
The following issues have been fixed:
- BE crashes when
LakePersistentIndexinitialization failed due to cleanup of_memtable. #62279 - A concurrency issue caused by missing locks when retrieving the maximum Tablet version in the replication transaction manager. #62238
- A hang issue in the phased scheduler, which waited indefinitely during synchronous Profile collection (after the fix, the system correctly terminates Profile collection when scheduling errors occur). #62140
- Exception handling issues in low-cardinality optimization under the
ALLOW_THROW_EXCEPTIONmode (after the fix, exceptions in expression evaluation are properly caught and returned). #62098 - FThe system failed to compute nested CTE statistics outside of the memo during table pruning when
enable_rbo_table_prunewas set tofalse. #62070 - CVE-2025-55163 issue. #62041
- An issue where
split_morsel_queuenested insidepartition_morsel_queuefailed to correctly receive the Tablet Schema. #62034 - Incorrect handling of
NULLarrays during Parquet writes, which could cause data inconsistency or crashes (after the fix, the system ensures thesplitfunction can correctly handleNULLinput strings). #61999 - Failure when creating materialized views using
CASE WHENexpressions due to incompatible return types of VARCHAR (after the fix, the system ensures consistency before and after refresh). #61996 - A concurrency safety issue caused by long operations holding shard-level locks while calculating compression scores. #61899
- An incomplete table pruning issue in CBO caused by pruning logic not considering all relevant predicates. #61881
3.3.17
Release Date: July 30, 2025
Bug Fixes
The following issues have been fixed:
- Upgraded HttpClient5 to 5.4.3. #61298
- Incorrect
cpu_core_used_permillelimit in resource groups. #61177 - Conflict between ALTER jobs and partition creation tasks. #61167
- NPE caused by missing
globalStateMgrinConnectContext. #60880 - Partition creation failed when partition names matched case-insensitively but had different values. #60909
- Lock competition caused by synchronous access to partition statistics. #61041
- ANALYZE tasks stuck in
pendingstate after FE restart. #61113 - Issue with JIT (Just-In-Time) compilation in BE. #61060
- Leader address issue in Starmgr. #61016
- CVE vulnerabilities in Broker. #60908
- Actual number of JDBC connections exceeded
jdbc_connection_pool_sizelimit. #61004 - CVE-2022-41404 vulnerability. #59689
- CVEs related to Parquet and HttpClient5. #58750
- Partition not removed from
_partition_mapwhen physical partition ID was empty. #60842 - Missing version check in shared-data clusters. #59422
- Transaction log missing when publishing logs in batches in shared-data clusters. #60949
- Concurrent publishing of the same transaction when Batch Publish is enabled in shared-data clusters. #57574
- Statistics overwrite issue caused by lack of semi-synchronous mode. #60897
- Inaccurate
maxInstantTimeused for filtering Hudi files when retrieving latest merged file slices. #60927 - TaskRun state incompatible with earlier versions. #60438
- CVE-2025-52999 vulnerability. #60795
- Vulnerability caused by
log4j-1.2.17-cloudera6in Broker. #59579 - BE crash when loading OOM partitions. #60778
- Base Compaction tasks blocking other compaction tasks. #60711
- Inefficient handling of error string truncation. #60878
- Materialized view rewrite failed in multi-FE environments. #60841
- INSERT OVERWRITE failed on manually created partitions. #60750
- Issue caused by using random distribution in aggregate keys. #60702
- Crash caused by low cardinality rewrite in
multi_distinct_count. #60664 - Issue with Pivot resolving fields. #60748
- Upgraded
hudi-commonto 1.0.2. #59501 - BE crash when CLONE and DROP TABLE run concurrently. #61359
3.3.16
Release Date: July 4, 2025
Improvements
- Optimized error logs when creating Hive tables with duplicate names. #60076
- Added the FE parameter
slow_lock_print_stackto prevent process stalls in large clusters when printing thread stacks. #59967 - Reduced unnecessary locks during tablet scheduling. #59744
Bug Fixes
Fixed the following issues:
- SplitOR fails to prune scan columns. #60223
- Incorrect query plan for null-aware left anti joins. #60119
- Incorrect query results when rewriting queries with materialized views due to missing NULL partitions. #60087
- Partition pruning errors when tables contain empty partitions. #60162
- Refresh errors on Iceberg external tables when using partition expressions based on
str2date. #60089 - Unexpected behavior caused by materialized view schema changes. #60079
- Issues related to low-cardinality global dictionaries in UNION operators. #60075
- Incorrect partition ranges for temporary partitions created using the START END syntax. #60014
- Lock issues with SUBMIT TASK. #60026
- Partial updates fail on Primary Key tables under certain conditions. #60052
- Crashes caused by BE failing to create directories due to a lack of permissions to access storage paths. #60028
- Cache failures due to cache key duplication in concurrent scenarios. #60053
- Hive table metadata background refresh failure in Unified Catalog. #55215
- Query failures caused by incorrect return types of CASE WHEN. #59972
- Query failures when Delta Lake tables UNION themselves. #60030
- Partition creation failure when writing to multiple tables within the same transaction. #59954
- Queries could return empty results instead of errors when tablet versions were updated during execution. #53060
- Queries against modified columns in a table return null after upgrading to v3.4. #59941
- Authentication information is printed in logs. #59907
- Metadata refresh failures for external tables in Hive Catalog. #54596
- CACHE SELECT failures for tables after schema changes. #59812
- Broker Load could not recover after FE Leader shifts. #59732
- Stream Load failures when the target table name contains Chinese characters. #59722
- Incorrect query results in external tables due to search key hash collisions (affecting Iceberg/Delta/Paimon). #59781
3.3.15
Release Date: Jun 20, 2025
Bug Fixes
Fixed the following issues:
- Missing double quotes for string parameters in statistics INSERT statements. #59713
- Downgrade failure caused by Rollup tasks. #59735
- Incorrect function parameters in the result of
SHOW CREATE VIEW. #59714 - A security issue where SQL statements with syntax errors exposed sensitive information in the Audit Log. #59442
- Error "Query version not found". #59194
- Failure to change data distribution using the
ALTER TABLEstatement. #59360 - An issue where root user processes were still visible when admin protection was enabled. #59435
- Failure of
INSERT OVERWRITEinto Hive. #59469 - Missing Tablet ID in the
max_tablet_rowset_numlog item. #59467 - An error caused by misconfigured Persistent Index parameters on a Duplicate table. #56040
- TaskRun history being archived on FE Follower nodes. #59393
- External catalog-based materialized view refresh errors. #59369
- Missing minimum version in Tablet information on shared-data clusters. #59373
- Abnormal maximum column unique ID in native tables of shared-data clusters due to version compatibility logic errors. #59190
- Materialized view refresh failure on Iceberg catalogs when the source Iceberg table is dropped and recreated, and manual refresh also fails after the materialized view is set to active. #59287
- Contamination of parameters in materialized view refresh tasks. #59052
- Data loss caused by Persistent Index when loading snapshot fails. #59247
- Issues caused when subcolumns of STRUCT appear in multiple predicates. #59216
- Query failure after renaming columns. #59178
- Loading failure due to multiple Stream Load requests. #59181
- Inability to refresh Hive table-based materialized views at the partition level in Unified Catalog. #59139
- Incorrect UNION plan causing FE out-of-memory (OOM). #59030
- Version loss during data loading. #59006
- Predicate loss when queries are rewritten to synchronous materialized views. #58831
- Issues with BITMAP/HLL/PERCENTILE data types in window functions. #58776
- Metadata changes to the external tables in Hive Catalog cannot be refreshed. #54596
Behavior Changes
- Introduced FE configuration parameter
task_runs_max_history_numberto control the number of historical TaskRuns retained in theinformation_schema.task_runsview, reducing memory usage. #59161
3.3.14
Release Date: May 14, 2025
Improvements
- Optimized error messages for regex parsing failures. #57904
- Fixed security vulnerabilities SNYK-JAVA-ORGJSON-5488379 and SNYK-JAVA-ORGJSON-5962464. #58425
Bug Fixes
Fixed the following issues:
- Issues with the JSON data type in
first_value/last_value/lead/lagwindow functions. #58697 - Deadlock caused by table-level locks from base tables during materialized view writes (after the bug fix, DB-level locks are used). #58615
- INSERT tasks hang when the target table is deleted. #58603
- Failure to change active/inactive state of materialized views with List partitions. #58575
- Incorrect
streaming_load_current_processingmetric. #58565 - Data version update errors caused by continuous loading and replica clone tasks. #58513
- Failed to refresh materialized views on external tables. #58506
- Incorrect
if()results on ARM architecture. #58455 - Materialized view rewriting generated incorrect query plans. #58487
- Iceberg table metadata did not refresh automatically. #58490
- Incorrect query plan generated by
group_concat. #57908 - Mass Tablet load failures caused by unhandled exceptions during loading. #58393
- Constant folding failed due to type mismatches while pruning List partitions with generated columns (after the bug fix, an implicit cast rule was added). #54543
- Mismatch between aggregate function return type and original column type (after the bug fix, the column type is
castto the function output type). #58407 broadcast_row_limitset to 0 or below failed to prevent BROADCAST JOIN generation. #58307- Broker Load used BE nodes that had already been blacklisted. #58350
- Asynchronous tasks persist in the background and cannot be dropped after manually cancelling materialized view refresh tasks. #58310
- Failed to create expression partitions with month or year granularity. #58182
ngram_searchgenerated invalid query plans. #58190
3.3.13
Release Date: April 22, 2025
Improvements
- Added memory consumption metrics for queries in FE in audit logs and the QueryDetail interface. #57731
- Optimized the strategy for concurrent creation of expression partitions. #57899
- Added monitoring metrics for the number of active FE nodes. #57857
- The
information_schema.task_runsview supports pushdown of the LIMIT clause. #57404 - Fixed several CVE issues. #57705 #57620
- Primary Key tables support retry during the PUBLISH stage, enhancing system disaster recovery capabilities. #57354
- Reduced memory consumption of Flat JSON. #57357
- The
information_schema.routine_load_jobsview adds thetimestamp_progresscolumn, consistent with the SHOW ROUTINE LOAD statement return. #57123 - Disallowed unauthorized behaviors from StarRocks to LDAP. #57131
- Supports returning an error when the schema of an AVRO file does not match the schema of the Hive table. #57296
- Materialized views support the
excluded_refresh_tablesproperty. #56428
Bug Fixes
Fixed the following issues:
- Flat JSON does not support the
get_json_boolfunction. #58077 - SHOW AUTHENTICATION statement returns the password. #58072
- The
percentile_countfunction returns incorrect values. #58038 - Issues caused by spilling strategies. #58022
- After a BE is blacklisted, Stream Load still dispatches tasks to the BE, causing task failures. #57919
- Issues when using the
castfunction with semi-structured data types. #57804 - The
array_mapfunction returns incorrect values. #57756 - In the scenario of a single tablet, using multiple
distinctfunctions on the same column with a single-column GROUP BY clause leads to incorrect query results. #57690 - MIN/MAX values in the profiles of big queries are inaccurate. #57655
- Non-partitioned materialized views based on Delta Lake data cannot rewrite queries. #57686
- A Routine Load deadlock issue. #57430
- Predicate pushdown issues with DATE/DATETIME columns. #57576
- An issue when the
percentile_discfunction has an empty input. #57572 - When modifying the bucket distribution of a table with the statement
ALTER TABLE {table} PARTITIONS (p1, p1) DISTRIBUTED BY ..., specifying duplicate partition names could result in failure to delete internally generated temporary partitions. #57005 - ALTER TABLE MODIFY COLUMN fails with expression partitioned tables based on
str2datefunction. #57487 - CACHE SELECT issue with semi-structured columns. #57448
- Upgrade compatibility issue caused by
hadoop-lib. #57436 - Case sensitivity error issues when creating partitions. #54867
- Some columns generate incorrect sort keys during updates. #57375
- Unknown issues caused by nested window functions . #57216
3.3.12
Release date: April 3, 2025
New Features
- Supports the
percentile_approx_weightedfunction. #56654 - Supports modifying properties of Hive Catalog and Hudi Catalog. #56212
- Paimon Catalog supports manifest cache. #55788
- Supports
SHOW PARTITIONSfor tables in Paimon Catalog. #55785 - Supports statistics collection for Paimon Catalog. #55757
Improvements
- Various improvements and bug fixes related to statistics. #57147 #57238 #57170 #57154 #57124 #57047 #56956 #57031 #56904 #56950 #56671 #55922
- Optimized error messages when table creation fails. #57055
- Enhanced retry mechanism for Broker Load. #56987
- Improved performance of
array_generate. #57252 - Aborted ongoing Compaction tasks for deleted partitions. #56943
- Optimized error messages when
ALTER TABLEfails. #57054 - Removed unnecessary reverse step from
array_agg()to improve performance. #56958 - Added checksum verification for replicas in Primary Key tables. #56519
- Masked sensitive information in the
FILESfunction output. #56684 - Reduced noisy logs related to materialized views. #56672
- Upgraded Iceberg version to 1.7.1. #55271
Bug Fixes
INSERT INTO FILESdid not support CSV delimiter conversion. #57126- Issues with Iceberg REST Catalog. #55416
- Predicate was lost during rewrite for view-based materialized views. #57153
- Paimon Catalog failed to read tables with schema changes. #56796
- Timezone conversion issue in Paimon Catalog. #56879
SHOW MATERIALIZED VIEWSdid not displaydefault_cataloginformation. #56362- In Trino dialect mode, time strings containing 'T' were not accepted. (Solution: replaced
parse_datetimewithstr_to_jodatime.) #56565 - Incorrect result of
first_valuefunction. #56467 - Incorrect result of
concat_wsfunction. #56384
Behavior Changes
- Added authentication to the FE Profile interface. #56914
- Changed default value of session variable
big_query_profile_thresholdfrom0to30. #56520
3.3.11
Release date: March 7, 2025
New Features
- Window functions support
max_byandmin_by. #54961
Improvements
Filessupports exporting JSON type data into Parquet files. #56406- Optimized Data Cache WarmUp performance for cloud-native tables in shared-data clusters. #56190
- Supports parsing
AT TIME ZONEexpressions and thefrom_iso8601_timestampfunction in Trino. #56311 #55573 - Partial Updates for Primary Key tables within shared-data clusters supports Condition Updates. #56132
- Extended support for statistics collection across all types of SQL statements. #56257
- Supports configuring the maximum number of returned rows for
SHOW PROC '/transaction'. #55933 - Supports creating asynchronous materialized views on Oracle-type JDBC Catalog tables. #55372
- MemTracker on BE WebUI supports pagination with 25 rows per page. #56206
- Supports pushdown for subfields of complex types in table functions. #55425
- Supports LDAP login for MariaDB clients. #55720
- Upgraded Paimon version to 1.0.1. #54796 #55760
- Eliminates unnecessary
unnestcomputations during query execution to reduce overhead. #55431 - Supports enabling Compaction for source clusters that are in the shared-data mode during cross-cluster synchronization. #54787
- Brings high-cost operations like DECIMAL division forward in topN computations to reduce overhead. #55417
- Optimized performance under ARM architecture. #55072 #55510
- For Hive table-based materialized views, StarRocks will perform checks and refreshes on the updated partitions only instead of full table refreshes if the base table was dropped and recreated. #45118
- DELETE operations support partition pruning. #55400
- Optimized priority strategy for collecting internal table statistics to improve efficiency when there are excessive tables. #55446
- When data loading involves multiple partitions, StarRocks merges transaction logs to improve loading performance. #55143
- Optimized error messages for SQL Translation. #55327
- Added a session variable
parallel_merge_late_materialization_modeto control parallel merge behavior. #55082 - Optimized error messages for generated columns. #54949
- Optimized performance of
SHOW MATERIALIZED VIEWS. #54374
Bug Fixes
Fixed the following issues:
- FE does not support casting constant TIME data types into DATETIME. #55804
- Stream Load transaction interface does not support the
starrocks_fe_table_load_rowsandstarrocks_fe_table_load_bytesmetrics. #44991 - Changes to automatic statistics collection do not take effect. #56173
- Materialized views in abnormal states caused issues with
SHOW MATERIALIZED VIEWS. #55995 - Text-based materialized view rewrite does not work across different databases. #56001
- Metadata compatibility issues in JDBC Catalogs. #55993
- Issues of handling the JSON data type in JDBC Catalogs. #56008
- Incorrect Sort Key settings during Schema Change. #55902
- Credential information leak issue in Broker Load. #55358
- An error caused by pushing down LIMIT before predicates in CTE. #55768
- An error caused by table schema changes in Stream Load. #55773
- A privilege issue due to the execution plan of DELETE statements containing SELECT. #55695
- An issue caused by not aborting compaction tasks when shutting down CN. #55503
- Follower FE nodes unable to fetch updated loading statistics. #55758
- Incorrect capacity statistics for the spill directory. #55703
- Failed to create materialized views due to lack of sufficient partition checks for base tables with list partitions. #55673
- An issue caused by missing metadata locks in ALTER TABLE. #55605
- An error in
SHOW CREATE TABLEcaused by constraints. #55592 - OOM due to large ARRAY in Nestloop Join. #55603
- Lock issue with DROP PARTITION. #55549
- An issue with min/max window functions due to not supporting string types. #55537
- Parser performance degraded. #54830
- Column name case sensitivity issue during partial updates. #55442
- Import failures when Stream Load is scheduled on nodes with an "Alive" state of false. #55371
- Incorrect output column order in materialized views containing ORDER BY. #55355
- BE crashes due to disk failure. #55042
- Incorrect query results caused by Query Cache. #55287
- Parquet Writer fails to convert time zone when writing TIMESTAMP type with time zones. #55194
- Loading tasks hang due to ALTER job timeout. #55207
- An error caused by
date_formatfunction when input is in milliseconds. #54854 - Materialized view rewrite failure caused by Partition Key being of DATE type. #54804
Behavior Changes
- Added authentication to the
query_detailinterface in FE. #55919 - The UUID type in Iceberg now maps to BINARY. #54978
- Uses changed row count instead of the visible time of partitions to determine if statistics need to be recollected. #55373
3.3.10 (Yanked)
Release date: February 21, 2025
This version has been taken offline due to metadata loss issues in shared-data clusters.
-
Problem: When there are committed compaction transactions that are not yet been published during a shift of Leader FE node in a shared-data cluster, metadata loss may occur after the shift.
-
Impact scope: This problem only affects shared-data clusters. Shared-nothing clusters are unaffected.
-
Temporary workaround: When the Publish task is returned with an error, you can execute
SHOW PROC 'compactions'to check if there are any partitions that have two compaction transactions with emptyFinishTime. You can executeALTER TABLE DROP PARTITION FORCEto drop the partitions to avoid Publish tasks getting hang.
3.3.9
Release date: January 12, 2025
New Features
- Supports the translation of Trino SQL into StarRocks SQL. #54185
Improvements
- Corrected FE node names starting with
bdbje_reset_election_groupto enhance clarity. #54399 - Implemented vectorization for the
IFfunction on ARM architectures. #53093 ALTER SYSTEM CREATE IMAGEsupports creating an image for StarManager. #54370- Supports deleting cloud-native indexes of Primary Key tables in shared-data clusters. #53971
- Enforced the refresh of materialized views when the
FORCEkeyword is specified. #52081 - Supports specifying hints in
CACHE SELECT. #54697 - Supports loading compressed CSV files using the
FILES()function. Supported compression formats include gzip, bz2, lz4, deflate, and zstd. #54626 - Supports assigning multiple values to the same column in an
UPDATEstatement. #54534
Bug Fixes
Fixed the following issues:
- Unexpected errors when refreshing materialized views built on JDBC catalogs. #54487
- Instability in results when a Delta Lake table joins itself. #54473
- Upload retries fail when backing up data to HDFS. #53679
- BFD initialization errors on the aarch64 architecture. #54372
- Sensitive information recorded in BE logs. #54677
- Errors in Compaction-related metrics in profiles. #54678
- BE crashes caused by creating tables with nested
TIMEtypes. #54601 - Query plan errors for
LIMITqueries with subquery TOP-N. #54507
Downgrade notes
- Clusters can be downgraded from v3.3.9 only to v3.2.11 and later.
3.3.8
Release date: January 3, 2025
Improvements
- Added a cluster idle API to assist in determining cluster status. #53850
- Included node information and histogram metrics in JSON metrics. #53735
- Optimized the MemTable for Primary Key tables in shared-data clusters. #54178
- Optimized memory usage and statistics for Primary Key tables in shared-data clusters. #54358
- Introduced a limit on the number of partitions scanned per node for queries requiring full-table or large-scale partition scans, enhancing system stability by reducing scanning pressure on individual BE or CN nodes. #53747
- Supports collecting statistics of Paimon tables. #52858
- Supports configuration of S3 client request timeout for shared-data clusters. #54211
Bug Fixes
Fixed the following issues:
- BE crashes caused by inconsistencies in the DelVec of Primary Key tables. #53460
- Issues with lock release of Primary Key tables in shared-data clusters. #53878
- Errors of UDFs nested in functions are not returned in query failures. #44297
- Transactions are blocked at the Decommission phase because they depend on the original replicas. #49349
- Queries against Delta Lake tables use relative paths instead of filenames for file retrieval. #53949
- An error is returned when querying Delta Lake Shallow Clone tables. #54044
- Case sensitivity issues when reading Paimon using JNI. #54041
- An error is returned during
INSERT OVERWRITEoperations on Hive tables created in Hive. #53792 SHOW TABLE STATUScommand does not validate view privileges. #53811- Missing FE metrics. #53058
- Memory leaks in
INSERTtasks. #53809 - Concurrency issues caused by missing write locks in replication tasks. #54061
partition_ttlof tables in thestatisticsdatabase does not take effect. #54398- Query Cache-related issues:
- Issues with materialized view Union Rewrite. #54293
- Missing padding in string updates for partial updates in Primary Key tables. #54182
- Incorrect execution plans for
max(count(distinct))when low-cardinality optimization is enabled. #53403 - Issues with changing the
excluded_refresh_tablesparameter of materialized views. #53394
Behavior Changes
- Changed the default value of
persistent_index_typefor Primary Key tables in shared-data clusters toCLOUD_NATIVE, that is, enabled Persistent Index by default. #52209
3.3.7
Release date: November 29, 2024
New Features
- Added a new Materialized View parameter,
excluded_refresh_tables, exclude tables that need to be refreshed. #50926
Improvements
- Rewrote
unnest(bitmap_to_array)asunnest_bitmapto improve performance. #52870 - Reduced the write and delete operations of Txn logs. #42542
Bug Fixes
Fixed the following issues:
- Failure to connect Power BI to external tables. #52977
- Misleading FE Thrift RPC failure messages in logs. #52706
- Routine Load tasks were canceled due to expired transactions (now tasks are canceled only if the database or table no longer exists). #50334
- Stream Load failures when submitted using HTTP 1.0. #53010 #53008
- Integer overflow of partition IDs. #52965
- Hive Text Reader failed to recognize the last empty element. #52990
- Issues caused by
array_mapin Join conditions. #52911 - Metadata cache issues under high concurrency scenarios. #52968
- The whole materialized view was refreshed when a partition was dropped from the base table. #52740
3.3.6
Release date: November 18, 2024
Improvements
- Optimized internal repair logic for Primary Key tables. #52707
- Optimized the internal implementation of histograms of statistics. #52400
- Supports adjusting log level via the FE configuration item
sys_log_warn_modulesto reduce Hudi Catalog logging. #52709 - Supports constant folding in the
yearweekfunction. #52714 - Avoided push-down for Lambda functions. #52655
- Divided the Query Error metric into three: Internal Error Rate, Analysis Error Rate, and Timeout Rate. #52646
- Avoided constant expressions being extracted as common expressions within
array_map. #52541 - Optimized the Text-based Rewrite of materialized views. #52498
Bug Fixes
Fixed the following issues:
- The
unique_constraintsandforeign_constraintsparameters were incomplete in SHOW CREATE TABLE for cloud-native tables in shared-data clusters. #52804 - Some materialized views were activated even when
enable_mv_automatic_active_checkwas set tofalse. #52799 - Memory usage is not reducing after stale memory flush. #52613
- Resource leak caused by Hudi file-system views. #52738
- Concurrent Publish and Update operations on Primary Key tables may cause issues. #52687
- Failures to terminate queries on clients. #52185
- Multi-column List partitions cannot be pushed down. #51036
- Incorrect result due to the lack of
hasnullproperty in ORC files. #52555 - An issue caused by using uppercase column names in ORDER BY during table creation. #52513
- An error was returned after running
ALTER TABLE PARTITION (*) SET ("storage_cooldown_ttl" = "xxx"). #52482
Behavior Changes
-
In earlier versions, scale-in operations would fail if there were insufficient replicas for views in the
_statistics_database. Starting from v3.3.6, if nodes are scaled in to 3 or more, view replicas are set to 3; if there is only 1 node after the scale-in, view replicas are set to 1, allowing for successful scale-in. #51799Affected views include:
column_statisticshistogram_statisticstable_statistic_v1external_column_statisticsexternal_histogram_statisticspipe_file_listloads_historytask_run_history
-
New Primary Key tables no longer allow
__opas a column name, even ifallow_system_reserved_namesis set totrue. Existing tables are unaffected. #52621 -
Expression-partitioned tables cannot have partition names modified. #52557
-
Deprecated FE parameters
heartbeat_mgr_blocking_queue_sizeandprofile_process_threads_num. #52236 -
Enabled persistent index on object storage by default for Primary Key tables in shared-data clusters. #52209
-
Disallowed manual changes to bucketing methods for tables with the random bucketing method. #52120
-
Backup and Restore-related parameter changes: #52111
make_snapshot_worker_countsupports dynamic configuration.release_snapshot_worker_countsupports dynamic configuration.upload_worker_countsupports dynamic configuration. Its default value is changed from1to the number of CPU cores on the machine where the BE resides.download_worker_countsupports dynamic configuration. Its default value is changed from1to the number of CPU cores on the machine where the BE resides.
-
The return type of
SELECT @@autocommithas changed from BOOLEAN to BIGINT. #51946 -
Added a new FE configuration item,
max_bucket_number_per_partition, to control the maximum number of buckets per partition. #47852 -
Enabled memory usage checks by default for Primary Key tables. #52393
-
Optimized loading strategy to reduce loading speed when Compaction tasks cannot be completed on time. #52269
3.3.5
Release date: October 23, 2024
New Features
- Supports millisecond and microsecond precision in the DATETIME type.
- Resource groups support CPU hard isolation.
Improvements
- Optimized performance and extraction strategy for Flat JSON. #50696
- Reduced memory usage for the following ARRAY functions:
- Optimized error messages when loading
Nullvalues into List partition keys with theNot Nullattribute. #51086 - Optimized error messages for Files() when authentication fails in the Files function. #51697
- Optimized internal statistics for
INSERT OVERWRITE. #50417 - Shared-data clusters support garbage collection (GC) for persistent index files. #51684
- Added FE logs to help diagnose FE out-of-memory (OOM) issues. #51528
- Supports recovering metadata from the metadata directory of FE. #51040