mariadb-columnstore-engine

mirror of https://github.com/mariadb-corporation/mariadb-columnstore-engine.git synced 2025-07-29 08:21:15 +03:00

Author	SHA1	Message	Date
Denis Khalikov	9119f6f7b8	fix(aggregate): MCOL-5467 Add support for duplicate expressions in group by. (#3045 ) This patch adds support for duplicate expressions (builtin_functions) with one argument in select statement and group by statement.	2023-12-05 15:04:53 +03:00
drrtuy	26f5f8fe5c	fix(plugin): this is to addres the original patch QA found in the original patch	2023-11-22 17:20:37 +03:00
Gagan Goel	931f2b36a1	MCOL-4931 Make cpimport charset-aware. (#2938 ) 1. Extend the following CalpontSystemCatalog member functions to set CalpontSystemCatalog::ColType::charsetNumber, after the system catalog update to add charset number to calpontsys.syscolumn in MCOL-5005: CalpontSystemCatalog::lookupOID CalpontSystemCatalog::colType CalpontSystemCatalog::columnRIDs CalpontSystemCatalog::getSchemaInfo 2. Update cpimport to use the CHARSET_INFO object associated with the charset number retrieved from the system catalog, for a dictionary/non-dictionary CHAR/VARCHAR/TEXT column, to truncate long strings that exceed the target column character length. 3. Add MTR test cases.	2023-09-05 17:17:20 +03:00
Andrey Piskunov	d586975da7	Rename a limit var + change error message (#2946 ) * Rename a limit var + change error message * Adjust the test	2023-09-05 12:19:15 +03:00
mariadb-AndreyPiskunov	05547f2342	Add a limit (as runtime value) for long in queries	2023-08-21 10:38:46 +03:00
mariadb-AndreyPiskunov	6ff121a91c	Replace recursion with iteration in ParseTree (and some related walkers)	2023-08-21 10:36:41 +03:00
Daniel Lee	1283f1fc4d	dlee mtr 23.08.1 (#2932 ) * Added order by clause to keep results consistent over test runs * Updated test result for the merging of MCOL-5519 * Updated test results for the merging of MCOL-4632 * Updated test result for the merging of MCOL-5519 * Added missing / to path * Improved few tests cases * Fixed test case name --------- Co-authored-by: root <root@rocky8.localdomain>	2023-08-18 00:01:33 +03:00
Theresa Hradilak	48562e41f9	feat(datatypes): MCOL-4632 and MCOL-4648, fix cast leads to NULL. Remove redundant cast. As C-style casts with a type name in parantheses are interpreted as static_casts this literally just changes the interpretation around (and forces an implicit cast to match the return value of the function). Switch UBIGINTNULL and UBIGINTEMPTYROW constants for consistency. Make consistent with relation between BIGINTNULL and BIGINTEMPTYROW & make adapted cast behaviour due to NULL markers more intuitive. (After this change we can simply block the highest possible uint64_t value and if a cast results in it, print the next lower value (2^64 - 2). Previously, (2^64 - 1) was able to be printed, but (2^64 - 2) as being blocked by the UBIGINTNULL constant was not, making finding the appropiate replacement value to give out more confusing. Introduce MAX_MCS_UBIGINT and MIN_MCS_BIGINT and adapt casts. Adapt casting to BIGINT to remove NULL marker error. Add bugfix regression test for MCOL 4632 Add regression test for mcol_4648 Revert "Switch UBIGINTNULL and UBIGINTEMPTYROW constants for consistency." This reverts commit 83eac11b18937ecb0b4c754dd48e4cb47310f620. Due to backwards compatability issues. Refactor casting to MCS[U]Int to datatype functions. Update regression tests to include other affected datatypes. Apply formatting. Refactor according to PR review Remove redundant new constant, switch to using already existing constant. Adapt nullstring casting to EMPTYROW markers for backwards compatability. Adapt tests for backward compatability behaviour allowing text datatypes to be casted to EMPTYROW constant. Adapt mcol641-functions test according to bug fix. Update tests according to new expected behaviour. Adapt tests to new understanding of issue. Update comments/documentation for MCOL_4632 test. Adapt to new cast limit logic. Make bracketing consistent. Adapt previous regression test to new expected behaviour.	2023-08-11 13:00:30 +00:00
Denis Khalikov	896e8dd769	MCOL-5522 Properly process pm join result count. (#2909 ) This patch: 1. Properly processes situation when pm join result count is exceeded. 2. Adds session variable 'columnstore_max_pm_join_result_count` to control the limit.	2023-08-04 16:55:45 +03:00
Gagan Goel	c598a9bbed	MCOL-5480 LOAD DATA INFILE incorrectly loads values for MEDIUMINT datatype. Internal memory representation of MEDIUMINT datatype uses 24 bits. This is true for both MariaDB server as well as ColumnStore. MCS plugin code uses TypeHandlerSInt24 and TypeHandlerUInt24 classes to respectively convert the binary representation of the signed and unsigned MEDIUMINT values passed by the server to the plugin. The plugin then outputs the text representation of these values into an open file descriptor which is piped to cpimport for the final load into the MCS db files. The TypeHandlerXInt24 classes were earlier incorrectly using WriteBatchField::ColWriteBatchXInt32() functions which operate on a 4 byte buffer. This resulted in incorrect parsing of MEDIUMINT values. As a fix, we implement WriteBatchField::ColWriteBatchXInt24() functions which correctly handle the 24 bit input buffer used for MEDIUMINT datatype.	2023-05-23 16:00:05 -04:00
Gagan Goel	1477b28ee9	MCOL-5357 Fix TPC-DS query error "MCS-3009: Unknown column '.<colname>'". For the following query: select item from ( select item from (select a as item from t1) tt union all select item from (select a as item from t1) tt ) ttt; There is an if predicate in buildSimpleColFromDerivedTable() that compares the outermost query field name (ttt.item) to the returned column list of the inner query (tt.item) when building the returned column list of the outer most query. In the above query example, the inner query field name is an alias set in the inner most query and is set to "`tt`.`item`", while the outermost query field name is set to "item". The use of backticks "`" in the inner query alias is causing the execution to not enter the if block which creates the SimpleColumn for the outermost query field name. As a fix, we strip off the backticks from the inner query alias.	2023-05-03 16:06:20 +00:00
Roman Nozdrin	786b9da5b0	MCOL-5438 COUNT() in math causes SEGV	2023-03-09 20:35:38 +00:00
root	2ed151fa59	Recorded reference results	2022-12-09 02:24:40 +00:00
mariadb-DanielLee	fad0fd08f0	Disable warnings for 'drop if exists' and 'create if not exist' commands	2022-12-08 11:37:44 -06:00
root	814fc37081	Moved some MTR cases to a new 1pmonly directory. Also added order by clause to few cases	2022-11-30 23:40:41 +00:00
Denis Khalikov	59166608b1	MCOL-4715 Mixed inner and outer joins with "null" filter for the table which is not involved into the outer join produces wrong results.	2022-08-16 17:13:03 +03:00
Denis Khalikov	e519cd7486	[MCOL-5061] Fix wrong `join id` assignment for the views. (#2474 ) This patch fixes a wrong `join id` assignment for `TupleHashJoinStep` in a view. After MCOL-334 CS assigns a '-1' as `join id` for `TupleHashJoinStep` in a view, and in this case we cannot apply a filter for specific `Join step`, which is associated with `join id` for 2 reasons: 1. Filters for all `TupleHashJoinSteps` associated with the same `join id`, which is '-1'. 2. When CS creates a `joinIdIndexMap` it eliminates all `join ids` which a less or equal 0. This patch also fixes some tests for the view, which were generated wrong results.	2022-07-25 20:02:02 +03:00
David.Hall	08bef648b3	Mcol 5074 Case with In and aggregates asserts (#2435 ) * MCOL-5074 CASE with IN and aggregate asserts gwip-scsp wasn't set and buildPredicateItem() was called which assumes it is set. Added code to set properly in this case	2022-07-11 16:20:15 -05:00
Denis Khalikov	e8f83121d2	[MCOL-4778] Return if we have an error in push_down_init.	2022-06-21 00:06:25 +03:00
Serguey Zefirov	53b9a2a0f9	MCOL-4580 extent elimination for dictionary-based text/varchar types The idea is relatively simple - encode prefixes of collated strings as integers and use them to compute extents' ranges. Then we can eliminate extents with strings. The actual patch does have all the code there but miss one important step: we do not keep collation index, we keep charset index. Because of this, some of the tests in the bugfix suite fail and thus main functionality is turned off. The reason of this patch to be put into PR at all is that it contains changes that made CHAR/VARCHAR columns unsigned. This change is needed in vectorization work.	2022-03-02 23:53:39 +03:00
benthompson15	4b412d4e09	MCOL-4940: test case for ROUND fix (#2271 )	2022-02-21 16:09:31 -06:00
Roman Nozdrin	7cfcdf365d	Merge pull request #2211 from drrtuy/MCOL-4899-dev MCOL-4899 MCS now applies a correct collation running IN for characte…	2022-01-06 21:08:11 +03:00
Roman Nozdrin	05897948e4	MCOL-4899 MCS now applies a correct collation running IN for character data types	2022-01-05 12:00:01 +00:00
Roman Nozdrin	695b437730	The goal is to migrate the last offending regr test001 test case into MTR to make test001 green	2021-12-30 19:12:58 +00:00
Gagan Goel	8cab54fe31	MCOL-4868 Move test cases for MCOL-4264 to MTR.	2021-12-20 18:34:08 +00:00
Gagan Goel	7f456e58cc	MCOL-4868 UPDATE on a ColumnStore table containing an IN-subquery on a non-ColumnStore table does not work. As part of MCOL-4617, we moved the in-to-exists predicate creation and injection from the server into the engine. However, when query with an IN Subquery contains a non-ColumnStore table, the server still performs the in-to-exists predicate transformation for the foreign engine table. This caused ColumnStore's execution plan to contain incorrect WHERE predicates. As a fix, we call mutate_optimizer_flags() for the WRITE lock, in addition to the READ table lock. And in mutate_optimizer_flags(), we change the optimizer flag from OPTIMIZER_SWITCH_IN_TO_EXISTS to OPTIMIZER_SWITCH_MATERIALIZATION.	2021-12-16 23:11:26 +00:00
Gagan Goel	340a90fc8d	MCOL-4874 Crossengine JOIN involving a ColumnStore table and a wide decimal column in a non-ColumnStore table throws an exception. ROW::getSignedNullValue() method does not support wide decimal fields yet. To fix this exception, we remove the call to this method from CrossEngineStep::setField().	2021-12-08 22:26:52 +00:00
Alexander Barkov	fa9f18553a	MCOL-4728 Query with unusual use of aggregate functions on ColumnStore table crashes MariaDB Server After an AggreateColumn corresponding to SUM(1+1) is created, it is pushed to the list: gwi.count_asterisk_list.push_back(ac) Later, in getSelectPlan(), the expression SUM(1+1) was erroneously treated as a constant: if (!hasNonSupportItem && !nonConstFunc(ifp) && !(parseInfo & AF_BIT) && tmpVec.size() == 0) { srcp.reset(buildReturnedColumn(item, gwi, gwi.fatalParseError)); This code freed the original AggregateColumn and replaced to a ConstantColumn. But gwi.count_asterisk_list still pointer to the freed AggregateColumn(). The expression SUM(1+1) was treated as a constant because tmpVec was empty due to a bug in this code: // special handling for count(). This should not be treated as constant. if (isp->argument_count() == 1 && ( sfitempp[0]->type() == Item::CONST_ITEM && (sfitempp[0]->cmp_type() == INT_RESULT \|\| sfitempp[0]->cmp_type() == STRING_RESULT \|\| sfitempp[0]->cmp_type() == REAL_RESULT \|\| sfitempp[0]->cmp_type() == DECIMAL_RESULT) ) ) { field_vec.push_back((Item_field)item); //dummy Notice, it handles only aggregate functions with explicit literals passed as an argument, while it does not handle constant expressions such as 1+1. Fix: - Adding new classes ConstantColumnNull, ConstantColumnString, ConstantColumnNum, ConstantColumnUInt, ConstantColumnSInt, ConstantColumnReal, ValStrStdString, to reuse the code easier. - Moving a part of the code from the case branch handling CONST_ITEM in buildReturnedColumn() into a new function newConstantColumnNotNullUsingValNativeNoTz(). This makes the code easier to read and to reuse in the future. - Adding a new function newConstantColumnMaybeNullFromValStrNoTz(). Removing dulplicate code from !!!four!!! places, using the new function instead. - Adding a function isSupportedAggregateWithOneConstArg() to properly catch all constant expressions. Using the new function parse_item() in the code commented as "special handling for count(*)". Now it pushes all constant expressions to field_vec, not only explicit literals. - Moving a part of the code from buildAggregateColumn() to a helper function processAggregateColumnConstArg(). Using processAggregateColumnConstArg() in the CONST_ITEM and NULL_ITEM branches. - Adding a new branch in buildReturnedColumn() handling FUNC_ITEM. If a function has constant arguments, a ConstantColumn() is immediately created, without going to buildArithmeticColumn()/buildFunctionColumn(). - Reusing isSupportedAggregateWithOneConstArg() and processAggregateColumnConstArg() in buildAggregateColumn(). A new branch catches aggregate function has only one constant argument and immediately creates a single ConstantColumn without traversing to the argument sub-components.	2021-09-21 14:00:56 +04:00
David Hall	a202bda485	MCOL-4719 iterate into subquery looking for windowfunctions When an outer query filter accesses an subquery column that contains an aggregate or a window function, certain optimizations can't be performed. We had been looking at the surface of the returned column. We now iterate into any functions or operations looking for aggregates and window functions.	2021-07-22 13:56:21 -05:00
Denis Khalikov	dc51dbf6cf	[MCOL-4786] Fix filter comparison. Compare ParseTree by dereferencing pointers.	2021-07-12 19:18:02 +03:00
Denis Khalikov	adace6e0c7	MCOL-4786 Fix wrong comparison for the filters. Fix wrong comparison for the filters while creating case function.	2021-07-09 12:18:26 +03:00
Gagan Goel	74bdf522d1	Merge pull request #1999 from mariadb-SergeyZefirov/MCOL-4741-in-like-equal-returns-different-result MCOL-4741 Fix extentmap handling of string prefixes encoded as 64-bit integers	2021-07-05 04:33:32 -04:00
David.Hall	237cad347f	MCOL-4758 Limit LONGTEXT and LONGBLOB to 16MB (#1995 ) MCOL-4758 Limit LONGTEXT and LONGBLOB to 16MB Also add the original test case from MCOL-3879.	2021-07-05 02:09:41 -04:00
Sergey Zefirov	c58136a32d	MCOL-4741 in/like/equal(=) operations differ in results This is due to signedness in the string range comparison in extentmap and unsignedness everywhere else.	2021-07-02 19:22:46 +03:00
Denis Khalikov	8dd2f2937c	MCOL-4407 and condtion does not work when HWM > columnstore_string_scan_threshold - 1	2021-06-21 14:04:26 +03:00
Alexander Barkov	75e3bbc31e	MCOL-4674 Fix ColumnStore to run MTR tests in a build directory	2021-04-13 11:25:25 +04:00

36 Commits