mariadb-columnstore-engine

mirror of https://github.com/mariadb-corporation/mariadb-columnstore-engine.git synced 2025-07-29 08:21:15 +03:00

Author	SHA1	Message	Date
David.Hall	6d680ceb8c	MCOL-603 Add error message for sum(a=1) (#2597 ) * MCOL-603 Add error message for sum(a=1) This isn't currently supported, but rather than emitting an error, it asserted and crashed.	2022-11-01 10:13:40 -05:00
mariadb-AndreyPiskunov	d7f4ec73c5	Small fixes + test sorting	2022-10-31 14:56:32 +02:00
mariadb-AndreyPiskunov	315e4be2d8	First working attempt for json_arrayagg	2022-10-31 14:56:32 +02:00
mariadb-AndreyPiskunov	1714b75434	Non working attempt to do MCOL-5227	2022-10-31 14:56:32 +02:00
Roman Nozdrin	a0086bc561	Adding NULL flag into ConstString class	2022-10-21 18:13:18 +00:00
Gagan Goel	8a9b6b32e7	MCOL-5000 Disable ALTER TABLE statement execution on replicas. Exit early from the plugin execution of ALTER TABLE statements on the replica nodes. This is to prevent re-execution of syscat table population from the replica nodes which should only be executed once by the primary node in a CS cluster setup.	2022-10-14 17:20:55 +00:00
david.hall	28a12eda82	Fix up CmakeLists.txt A better way to fix the dependencies	2022-09-06 16:15:07 -05:00
david.hall	bcaf867731	Fix up cmake to build out of band The main CmakeLists.txt was using MY_CHECK_AND_SET_COMPILER_FLAG before the include. This works in-band with server because it was already included in server's CmakeLists.txt. dbcon/mysql included curl as a build dependency. We don't build curl. It's a lib dependency. Not sure why it works in-band. One wouldn't think it should.	2022-09-06 16:08:47 -05:00
Ziy1-Tan	cdd41f05f3	MCOL-785 Implement DISTRIBUTED JSON functions The following functions are created: Create function JSON_VALID and test cases Create function JSON_DEPTH and test cases Create function JSON_LENGTH and test cases Create function JSON_EQUALS and test cases Create function JSON_NORMALIZE and test cases Create function JSON_TYPE and test cases Create function JSON_OBJECT and test cases Create function JSON_ARRAY and test cases Create function JSON_KEYS and test cases Create function JSON_EXISTS and test cases Create function JSON_QUOTE/JSON_UNQUOTE and test cases Create function JSON_COMPACT/DETAILED/LOOSE and test cases Create function JSON_MERGE and test cases Create function JSON_MERGE_PATCH and test cases Create function JSON_VALUE and test cases Create function JSON_QUERY and test cases Create function JSON_CONTAINS and test cases Create function JSON_ARRAY_APPEND and test cases Create function JSON_ARRAY_INSERT and test cases Create function JSON_INSERT/REPLACE/SET and test cases Create function JSON_REMOVE and test cases Create function JSON_CONTAINS_PATH and test cases Create function JSON_OVERLAPS and test cases Create function JSON_EXTRACT and test cases Create function JSON_SEARCH and test cases Note: Some functions output differs from MDB because session variables that affects functions output,e.g JSON_QUOTE/JSON_UNQUOTE This depends on MCOL-5212	2022-08-30 22:22:23 +08:00
Roman Nozdrin	72e264e8ef	MCOL-5199 This patch solves the overal performance degradation introduced with a new way of char columns hashing in aggregation code The patch disables padding that forces hasher to calculate over the whole 2k buffer. This patch also moves hashing code into the common place where it belongs.	2022-08-24 19:07:06 +00:00
Leonid Fedorov	d02b3403b7	Changed function name, schema and params order to achieve columnstore_info.load_from_s3("<bucket>", "<file_name>", "<db_name>", "<table_name>");	2022-08-18 10:41:22 +00:00
Roman Nozdrin	56bbef62e6	Merge pull request #2406 from tntnatbry/MCOL-5021-dev MCOL-5021 AUX column implementation to improve DELETE performance.	2022-08-15 19:03:42 +03:00
David.Hall	2020f35e88	Mcol 5092 MODA uses wrong column width for some types (#2450 ) * MCOL-5092 Ensure column width is correct for datatype Change MODA return type to STRING Modify MODA to handle every numeric type * MCOL-5162 MODA to support char and varchar with collation support Fixes to the aggregate bit functions When we fixed the storage sign issue for MCOL-5092, it uncovered a problem in the bit aggregates (bit_and, bit_or and bit_xor). These aggregates should always return UBIGINT, but they relied on the type of the argument column, which gave bad results.	2022-08-11 15:16:11 -05:00
Gagan Goel	86df9a972c	MCOL-5021 Add prototype support for the AUX column in CREATE/DROP DDL commands, single and multi-value INSERTs, cpimport, and DELETE.	2022-08-05 14:40:49 -04:00
David.Hall	d3b57ec767	MCOL-4800 emit error if IN filter > 65535 entries (#2480 ) * MCOL-4800 emit error if IN filter > 65535 entries	2022-08-04 19:21:58 +03:00
David.Hall	08bef648b3	Mcol 5074 Case with In and aggregates asserts (#2435 ) * MCOL-5074 CASE with IN and aggregate asserts gwip-scsp wasn't set and buildPredicateItem() was called which assumes it is set. Added code to set properly in this case	2022-07-11 16:20:15 -05:00
david.hall	c71d11cb3f	Restore calonlinealter	2022-07-06 09:22:49 -05:00
Leonid Fedorov	242769d542	Mistype bug error handler fix	2022-07-05 18:48:30 +03:00
Roman Nozdrin	a3bc3de5f4	Merge pull request #2432 from mariadb-corporation/dataload-raw MCOL-5013: Load Data from S3 into Columnstore	2022-07-05 13:06:53 +03:00
Roman Nozdrin	38c4b973dd	Merge pull request #2421 from denis0x0D/MCOL-4778 [MCOL-4778] Return if we have an error in push_down_init.	2022-07-04 21:16:13 +03:00
Leonid Fedorov	110d9cfab5	Review fixes	2022-07-04 19:52:37 +03:00
Leonid Fedorov	f5b2a6885f	MCOL-5013: Load Data from S3 into Columnstore Introduced UDF and stored prodecure. usage: set columnstore_s3_key='<s3_key>'; set columnstore_s3_secret='<s3_secret>'; set columnstore_s3_region='region'; and then use UDF select columnstore_dataload("<tablename>", "<filename>", "<bucket>", "<db_name>"); for UDF db_name can be ommited, then current connection db will be used or stored function call calpontsys.columnstore_load_from_s3("<tablename>", "<filename>", "<bucket>", "<db_name>");	2022-07-04 19:52:37 +03:00
Denis Khalikov	e8f83121d2	[MCOL-4778] Return if we have an error in push_down_init.	2022-06-21 00:06:25 +03:00
David.Hall	272246e9fa	Merge branch 'develop' into MCOL-4841	2022-06-09 16:58:33 -05:00
david.hall	3b6449842f	Merge branch 'develop' into MCOL-4841 # Conflicts: # exemgr/main.cpp # oam/etc/Columnstore.xml.singleserver # primitives/primproc/primproc.cpp	2022-06-09 10:07:26 -05:00
Roman Nozdrin	f29d5e7869	MCOL-4912 This patch adds some forgotten MDB functions	2022-05-27 16:27:07 +00:00
benthompson15	e147184b8d	MCOL-5065: return values of getSystemReady/getSystemQueryReady should be > 0 (#2354 )	2022-05-10 12:32:17 -05:00
Roman Nozdrin	4c26e4f960	MCOL-4912 This patch introduces Extent Map index to improve EM scaleability EM scaleability project has two parts: phase1 and phase2. This is phase1 that brings EM index to speed up(from O(n) down to the speed of boost::unordered_map) EM lookups looking for <dbroot, oid, partition> tuple to turn it into LBID, e.g. most bulk insertion meta info operations. The basis is boost::shared_managed_object where EMIndex is stored. Whilst it is not debug-friendly it allows to put a nested structs into shmem. EMIndex has 3 tiers. Top down description: vector of dbroots, map of oids to partition vectors, partition vectors that have EM indices. Separate EM methods now queries index before they do EM run. EMIndex has a separate shmem file with the fixed id MCS-shm-00060001.	2022-05-04 12:59:16 +00:00
Commander thrashdin	f28e00c206	No repeating code in client_udfs + better test	2022-03-28 21:48:47 +03:00
Commander thrashdin	8d31478b72	Added mcsUDFs to install.sh+removed nonexistent fn	2022-03-28 21:48:47 +03:00
Commander thrashdin	749b8f16ee	Added mcs-named UDFs to cpp	2022-03-28 21:48:46 +03:00
Leonid Fedorov	65252df4f6	C++20 fixes	2022-03-28 12:32:29 +00:00
Leonid Fedorov	3919c541ac	New warnfixes (#2254 ) * Fix clang warnings * Remove vim tab guides * initialize variables * 'strncpy' output truncated before terminating nul copying as many bytes from a string as its length * Fix ISO C++17 does not allow 'register' storage class specifier for outdated bison * chars are unsigned on ARM, having if (ival < 0) always false * chars are unsigned by default on ARM and comparison with -1 if always true	2022-02-17 13:08:58 +03:00
Gagan Goel	973e5024d8	MCOL-4957 Fix performance slowdown for processing TIMESTAMP columns. Part 1: As part of MCOL-3776 to address synchronization issue while accessing the fTimeZone member of the Func class, mutex locks were added to the accessor and mutator methods. However, this slows down processing of TIMESTAMP columns in PrimProc significantly as all threads across all concurrently running queries would serialize on the mutex. This is because PrimProc only has a single global object for the functor class (class derived from Func in utils/funcexp/functor.h) for a given function name. To fix this problem: (1) We remove the fTimeZone as a member of the Func derived classes (hence removing the mutexes) and instead use the fOperationType member of the FunctionColumn class to propagate the timezone values down to the individual functor processing functions such as FunctionColumn::getStrVal(), FunctionColumn::getIntVal(), etc. (2) To achieve (1), a timezone member is added to the execplan::CalpontSystemCatalog::ColType class. Part 2: Several functors in the Funcexp code call dataconvert::gmtSecToMySQLTime() and dataconvert::mySQLTimeToGmtSec() functions for conversion between seconds since unix epoch and broken-down representation. These functions in turn call the C library function localtime_r() which currently has a known bug of holding a global lock via a call to __tz_convert. This significantly reduces performance in multi-threaded applications where multiple threads concurrently call localtime_r(). More details on the bug: https://sourceware.org/bugzilla/show_bug.cgi?id=16145 This bug in localtime_r() caused processing of the Functors in PrimProc to slowdown significantly since a query execution causes Functors code to be processed in a multi-threaded manner. As a fix, we remove the calls to localtime_r() from gmtSecToMySQLTime() and mySQLTimeToGmtSec() by performing the timezone-to-offset conversion (done in dataconvert::timeZoneToOffset()) during the execution plan creation in the plugin. Note that localtime_r() is only called when the time_zone system variable is set to "SYSTEM". This fix also required changing the timezone type from a std::string to a long across the system.	2022-02-14 14:12:27 -05:00
David Hall	27dea733c5	MCOL4841 dev port run large join without OOM	2022-02-09 17:33:55 -06:00
Leonid Fedorov	04752ec546	clang format apply	2022-01-21 16:43:49 +00:00
Leonid Fedorov	01f3ceb437	replace header guards with #pragma once	2022-01-21 15:24:58 +00:00
Gagan Goel	195425924d	MCOL-4936 Disable binlog for DML statements. DML statements executed on the primary node in a ColumnStore cluster do not need to be written to the primary's binlog. This is due to ColumnStore's distributed storage architecture. With this patch, we disable writing to binlog when a DML statement (INSERT/DELETE/UPDATE/LDI/INSERT..SELECT) is performed on a ColumnStore table. HANDLER::external_lock() calls are used to 1. Turn OFF the OPTION_BIN_LOG flag 2. Turn ON the OPTION_BIN_TMP_LOG_OFF flag in THD::variables.option_bits during a WRITE lock call. THD::variables.option_bits is restored back to the original state during the UNLOCK call in HANDLER::external_lock(). Further, isDMLStatement() function is added to reduce code verbosity to check if a given statement is a DML statement. Note that with this patch, not writing to primary's binlog means DML replication from a ColumnStore cluster to another ColumnStore cluster or to another foreign engine will not work.	2022-01-04 17:31:59 +00:00
Roman Nozdrin	b3ab3fb514	Merge pull request #2203 from mariadb-AlexeyAntipovsky/auto-query-stats [MCOL-4944] Automatically enable stats collection	2021-12-23 15:13:43 +03:00
Roman Nozdrin	94806e7ee0	Merge pull request #2200 from drrtuy/MCOL-4943-dev MCOL-4943 Moved SQL script call into columnstore-post-install	2021-12-21 11:18:07 +03:00
Alexey Antipovsky	683a6b3d19	[MCOL-4944] Automatically enable stats collection if it is enabled in the config	2021-12-21 11:15:25 +03:00
Roman Nozdrin	22b0e4addc	MCOL-4943 Moved SQL script call into columnstore-post-install	2021-12-18 03:59:58 +00:00
Roman Nozdrin	7b5845a4aa	MCOL-4871 Bar's patch to do proper extent elimination for short CHAR	2021-12-17 17:41:03 +00:00
Gagan Goel	7f456e58cc	MCOL-4868 UPDATE on a ColumnStore table containing an IN-subquery on a non-ColumnStore table does not work. As part of MCOL-4617, we moved the in-to-exists predicate creation and injection from the server into the engine. However, when query with an IN Subquery contains a non-ColumnStore table, the server still performs the in-to-exists predicate transformation for the foreign engine table. This caused ColumnStore's execution plan to contain incorrect WHERE predicates. As a fix, we call mutate_optimizer_flags() for the WRITE lock, in addition to the READ table lock. And in mutate_optimizer_flags(), we change the optimizer flag from OPTIMIZER_SWITCH_IN_TO_EXISTS to OPTIMIZER_SWITCH_MATERIALIZATION.	2021-12-16 23:11:26 +00:00
Gagan Goel	d91cab2ff5	MCOL-4925 Suppress the warning message when a non-cached table is (#2164 ) dropped with the insert cache enabled.	2021-12-06 11:27:37 -06:00
Gagan Goel	affb2ae770	MCOL-4769 Fix cache bugs. (#2151 ) * MCOL-4769 Do not replay INSERTs and LDIs on the replica nodes when the write cache is enabled. * MCOL-4769 If a table is created with the write cache disabled (i.e. when columnstore_cache_inserts=OFF), make it accessible when the cache feature is enabled (columnstore_cache_inserts=ON).	2021-11-22 14:20:50 -06:00
Alexander Barkov	fa9f18553a	MCOL-4728 Query with unusual use of aggregate functions on ColumnStore table crashes MariaDB Server After an AggreateColumn corresponding to SUM(1+1) is created, it is pushed to the list: gwi.count_asterisk_list.push_back(ac) Later, in getSelectPlan(), the expression SUM(1+1) was erroneously treated as a constant: if (!hasNonSupportItem && !nonConstFunc(ifp) && !(parseInfo & AF_BIT) && tmpVec.size() == 0) { srcp.reset(buildReturnedColumn(item, gwi, gwi.fatalParseError)); This code freed the original AggregateColumn and replaced to a ConstantColumn. But gwi.count_asterisk_list still pointer to the freed AggregateColumn(). The expression SUM(1+1) was treated as a constant because tmpVec was empty due to a bug in this code: // special handling for count(). This should not be treated as constant. if (isp->argument_count() == 1 && ( sfitempp[0]->type() == Item::CONST_ITEM && (sfitempp[0]->cmp_type() == INT_RESULT \|\| sfitempp[0]->cmp_type() == STRING_RESULT \|\| sfitempp[0]->cmp_type() == REAL_RESULT \|\| sfitempp[0]->cmp_type() == DECIMAL_RESULT) ) ) { field_vec.push_back((Item_field)item); //dummy Notice, it handles only aggregate functions with explicit literals passed as an argument, while it does not handle constant expressions such as 1+1. Fix: - Adding new classes ConstantColumnNull, ConstantColumnString, ConstantColumnNum, ConstantColumnUInt, ConstantColumnSInt, ConstantColumnReal, ValStrStdString, to reuse the code easier. - Moving a part of the code from the case branch handling CONST_ITEM in buildReturnedColumn() into a new function newConstantColumnNotNullUsingValNativeNoTz(). This makes the code easier to read and to reuse in the future. - Adding a new function newConstantColumnMaybeNullFromValStrNoTz(). Removing dulplicate code from !!!four!!! places, using the new function instead. - Adding a function isSupportedAggregateWithOneConstArg() to properly catch all constant expressions. Using the new function parse_item() in the code commented as "special handling for count(*)". Now it pushes all constant expressions to field_vec, not only explicit literals. - Moving a part of the code from buildAggregateColumn() to a helper function processAggregateColumnConstArg(). Using processAggregateColumnConstArg() in the CONST_ITEM and NULL_ITEM branches. - Adding a new branch in buildReturnedColumn() handling FUNC_ITEM. If a function has constant arguments, a ConstantColumn() is immediately created, without going to buildArithmeticColumn()/buildFunctionColumn(). - Reusing isSupportedAggregateWithOneConstArg() and processAggregateColumnConstArg() in buildAggregateColumn(). A new branch catches aggregate function has only one constant argument and immediately creates a single ConstantColumn without traversing to the argument sub-components.	2021-09-21 14:00:56 +04:00
Leonid Fedorov	5c5f103f98	MCOL-4839: Fix clang build (#2100 ) * Fix clang build * Extern C returned to plugin_instance Co-authored-by: Leonid Fedorov <l.fedorov@mail.corp.ru>	2021-08-23 10:45:10 -05:00
benthompson15	923bbf4033	MCOL-1356: Add convert_tz (#2099 )	2021-08-19 17:47:10 -05:00
Gagan Goel	98473a45cc	Merge pull request #2079 from dhall-MariaDB/MCOL-3741 Mcol 3741 Change IDB-xxxx error codes to MCS-xxxx	2021-08-18 14:01:04 -04:00

1 2 3 4 5 ...

1013 Commits