Commits · bb-11.0 · nexedi / MariaDB

21 Feb, 2023 1 commit

squash! · e91b56c0

Michael Widenius authored Feb 20, 2023

- Ensure that TEMP_TABLE_PARAM.func_count includes all items that may
  need a copy function.
- Fixed that Aria allocates enough space for key copies.
- Fixed that Aria does not check empty_bits if not allocated.

The first issue could cause crashes, the other issues should not
affect anything.

e91b56c0

20 Feb, 2023 1 commit
- tmp commit to try to find crash in windows builds · 24c4877e
  Michael Widenius authored Feb 19, 2023
```
This commit adds support for more than one key in temporary tables
```
  24c4877e
19 Feb, 2023 2 commits
- Added mysql-log-rotate to .gitignore · 07232ac0
  Michael Widenius authored Feb 19, 2023
  
  07232ac0
- Added detection of memory overwrite with multi_malloc · 4cb69791
  Michael Widenius authored Feb 19, 2023
```
This detected two bugs in the code, which are both fixed
```
  4cb69791
17 Feb, 2023 2 commits

MDEV-30659 Server crash on EXPLAIN SELECT/SELECT on table with engine Aria for LooseScan Strategy · c00c12df
Monty authored Feb 16, 2023
```
The issue was that Loose_scan_opt::save_to_position() did not take
into account records_out from best_access_path()
```
c00c12df

Updated prev_record_reads() to be more exact · 0dffeeab

Monty authored Feb 14, 2023

The old code in prev_record_reads() did give wrong estimates when a
join_buffer was used or if the table was depending on more than one
other tables. When join_cache is used, it will cause a re-order of row
combinations, which causes more calls to the engine for tables that
are depending on tables before the join_cached one.

The new prev_records_read() code provides more exact estimates and
should never give a 'too low estimate', assuming that the data to the
function is correct

The definition of prev_record_read() is also updated.
The new definition is:
  "Estimate the number of engine ha_index_read_calls for EQ_REF tables
  when taking into account the one-row-cache in join_read_always_key()"

The cost of using prev_record_reads() value is changed. The value is
now used similar as before to calculate the cost of the storage engine
calls. However the cost of the WHERE cost is changed to take into
account the total number of row combinations as the WHERE has to be
checked even if the one-row-cache is used. This makes the cost
slightly higher than before (for the same prev_record_reads() value).

Other things:
- Cached return value of prev_record_read() in best_access_path() to
  avoid some function calls.
- Fixed bug where position[].use_join_buffer was set in
  best_acess_path() when join buffer was not used. This confused the
  semi join optimizer to try to reoptimize plans that did not need to be
  reoptimized.
  The effect of the bug fix is that we avoid doing some re-optimziations
  with semi-joins when join_buffer is not used. In these cases the value
  shown for the 'Filtering' column in EXPLAIN EXTENDED may change.

Changes in test suite:
- EQ_REF tables are moved up to be earlier. This is because either the
  higher WHERE cost when EQ_REF is used with more row combination or
  change of cost when using join_cache.
- Filtered has changed (to the better) for some cases using semi-joins
  subselect_sj.test subselect_sj_jcl6.test

0dffeeab

12 Feb, 2023 2 commits

Added r_table_looks to ANALYZE TABLE · fb937116
Monty authored Feb 12, 2023
```
Author: Sergei Petrunia <sergey@mariadb.com>
```
fb937116

Adjust costs for rowid filter · 9527a0ae

Monty authored Feb 12, 2023

- Use log2() insted of log()
- Added missing ''+' when calculating rowid setup cost
- Adjusted ROWID_FILTER_PER_ELEMENT_MODIFIER (from 3 to 1)

Other things:
- Adjusted cost for index_merge where rows_out < 1.0

The effects of the changes:
- rowid filter will have higher setup cost
- rowid filter will have slightly less costs per row

This can be seen in mtr where some tests, with 'small tables or
that uses rowid filters with many rows, will not use rowid filter anymore.

9527a0ae

10 Feb, 2023 32 commits

MDEV-30525: Assertion `ranges > 0' fails in IO_AND_CPU_COST · 62531b0d
Sergei Petrunia authored Feb 09, 2023
```
Part #2: fix the case where table->stat_records()=1 (due to EITS
statistics), but the range returns rows=0.
```
62531b0d

MDEV-30540 Wrong result with IN list length reaching IN_PREDICATE_CONVERSION_THRESHOLD · b3bdb41d

Monty authored Feb 10, 2023

The problem was the mysql_derived_prepare() did not correctly set
'distinct' when creating a temporary derivated table.

Fixed by separating checking for distinct for queries with and without
UNION.

Other things:
- Fixed bug in generate_derived_keys_for_table() where we set the wrong
  bit for join_tab->keys
- Cleaned up JOIN::drop_unused_derived_keys()
- Changed TABLE::use_index() to keep unique keys and update
  share->key_parts

Author: Sergei Petrunia <sergey@mariadb.com>, monty@mariadb.org

b3bdb41d

Fixed compiler warning in connect/ha_connect.cc · 7a9d658f
Monty authored Feb 10, 2023
```
(fp->field_length is always >= 0)
```
7a9d658f
Fixed bug in federated.federatedx · 4493745a
Monty authored Feb 10, 2023

4493745a
Fixed check_costs.pl to always create table if table does not exists · ddab4057
Monty authored Feb 10, 2023
```
This allows one to always use --skip-create-table for repeated runs.
```
ddab4057

MDEV-30603: Wrong result with non-default JOIN_CACHE_LEVEL=[4|5] ... · 81277a9a

Sergei Petrunia authored Feb 08, 2023

JOIN_CACHE::alloc_buffer() used wrong logic when calculating the size
of all join buffers. Then, it computed the ratio by which
JOIN::shrink_join_buffers() should shrink the buffers.

shrink_join_buffers() ended up in a situation where buffers would not
fit into the total quota after shrinking, which resulted in negative
buffer sizes. Due to use of unsigned integers it would cause very large
buffers to be used instead.

Make JOIN_CACHE::alloc_buffer() use the same logic as
JOIN::shrink_join_buffers() when it calculates the total size of
all join buffers so far.

Also, add a safety check in JOIN::shrink_join_buffers()

81277a9a

MDEV-30569: Assertion ...ha_table_flags() in Duplicate_weedout_picker::check_qep · a7666952

Sergei Petrunia authored Feb 06, 2023

DuplicateWeedout semi-join optimization requires that the tables in
the parent subquery provide rowids that can be compared across table
scans. Most engines support this, federated is the only exception.

DuplicateWeedout is the default catch-all semi-join strategy, which
must be always available. If it is not available for some edge case,
it's better to disable semi-join conversion altogether.

This is what was done in the fix for MDEV-30395. However that fix
has put the check before the view processing, so it didn't detect
federated tables inside mergeable VIEWs.

This patch moves the check to be done at a later phase, when mergeable
views are already merged.

a7666952

MDEV-30568: Assertion `cond_selectivity <= 1.000000001' failed in get_range_limit_read_cost · d6616966
Sergei Petrunia authored Feb 06, 2023
```
In get_range_limit_read_cost(), handle the case where range_rows=0.
```
d6616966
MDEV-30529: Assertion `rnd_records <= s->found_records' failed in best_access_path · cc81ea1c
Sergei Petrunia authored Feb 03, 2023
```
best_access_path() has an assertion:

   DBUG_ASSERT(rnd_records <= s->found_records);

make it rounding-safe.
```
cc81ea1c

MDEV-30525: Assertion `ranges > 0' fails in IO_AND_CPU_COST handler::keyread_time · 5faf2ac0

Sergei Petrunia authored Feb 03, 2023

Make get_best_group_min_max() exit early if the table has
table->records()=0. Attempting to compute loose scan over 0
groups eventually causes an assert when trying to get the
cost of reading 0 ranges.

5faf2ac0

Fixed bug in extended key handling when there is no primary key · 00704aff

Monty authored Jan 29, 2023

Extended keys works by first checking if the engine supports extended
keys.
If yes, it extends secondary key with primary key components and mark the
secondary keys as HA_EXT_NOSAME (unique).
If we later notice that there where no primary key, the extended key
information for secondary keys in share->key_info is reset. However the
key_info->flag HA_EXT_NOSAME was not reset!

This causes some strange things to happen:
- Tables that have no primary key or secondary index that contained the
  primary key would be wrongly optimized as the secondary key could be
  thought to be unique when it was not and not unique when it was.
- The problem was not shown in EXPLAIN because of a bug in
  create_ref_for_key() that caused EQ_REF to be displayed by EXPLAIN as REF
  when extended keys where used and the secondary key contained the primary
  key.

This is fixed with:
- Removed wrong test in make_join_select() which did not detect that key
  where unique when a secondary key contains the primary.
- Moved initialization of extended keys from create_key_infos() to
  init_from_binary_frm_image() after we know if there is a usable primary
  key or not. One disadvantage with this approach is that
  key_info->key_parts may have not used slots (for keys we thought could
  be extended but could not). Fixed by adding a check for unused key_parts
  to copy_keys_from_share().

Other things:
- Simplified copying of first key part in create_key_infos().
- Added a lot of code comments in code that I had to check as part of
  finding the issue.
- Fixed some indentation.
- Replaced a couple of looks using references to pointers in C
  context where the reference does not give any benefit.
- Updated Aria and Maria to not assume the all key_info->rec_per_key
  are in one memory block (this could happen when using dervived
  tables with many keys).
- Fixed a bug where key_info->rec_per_key where not allocated
- Optimized TABLE::add_tmp_key() to only call alloc() once.
  (No logic changes)

Test case changes:
- innodb_mysql.test changed index as an index the optimizer thought
  was unique, was not. (Table had no primary key)

TODO:
- Move code that checks for partial or too long keys to the primary loop
  earlier that initally decides if we should add extended key fields.
  This is needed to ensure that HA_EXT_NOSAME is not set for partial or
  too long keys. It will also shorten the current code notable.

00704aff

MDEV-30486 Table is not eliminated in bb-11.0 · fe1f4ca8

Monty authored Jan 27, 2023

Some tables where not eliminated when they could have been.
This was caused because HA_KEYREAD_ONLY is not set anymore for InnoDB
clustered index and the elimination code was depending on
field->part_of_key_not_clustered which was not set if HA_KEYREAD_ONLY
is not present.

Fixed by moving out field->part_of_key and
field->part_of_key_not_clustered from under HA_KEYREAD_ONLY (which
they should never have been part of).

Other things:
- Fixed a bug in make_join_select() that caused range to be used when
  there where elminiated or constant tables present (Caused wrong
  change of plans in join_outer_innodb.test). This also affected
  show_explain.test and subselct_sj_mat.test where wrong 'range's where
  replaced with index scans.

Reviewer: Sergei Petrunia <sergey@mariadb.com>

fe1f4ca8

Removed /2 of InnoDB ref_per_key[] estimates · 01c82173

Monty authored Jan 26, 2023

The original code was there to favor index search over table scan.
This is not needed anymore as the cost calculations for table scans
and index lookups are now more exact.

01c82173

Optimizer Trace: make plan_prefix not show const/eliminated tables · 87507bbb
Sergei Petrunia authored Jan 27, 2023

87507bbb

remove GET_ADJUST_VALUE · 2010cfab

Sergei Golubchik authored Jan 03, 2023

avoid contaminating my_getopt with sysvar implementation details.
adjust variable values after my_getopt, like it's done for others.
this fixes --help to show correct values.

2010cfab

remove SHOW_OPTIMIZER_COST · d10b3b01

Sergei Golubchik authored Jan 03, 2023

avoid contaminating SHOW code with sysvar implementation details.
And no hard-coded factor either.

d10b3b01

remove Feature_into_old_syntax · affab99c

Sergei Golubchik authored Jan 03, 2023

it doesn't provide any information we'll use.
No matter what the value is, we don't remove the non-standard
syntax unless we have to

affab99c

typos in comments, etc · 7e465aeb
Sergei Golubchik authored Jan 03, 2023

7e465aeb

Selectivity: apply found_constraint heuristic only to post-join #rows. · 5e5988db

Monty authored Jan 24, 2023

matching_candidates_in_table() computes the number of rows one
gets from the current table after applying the WHERE clause on
just this table

The function had a "found_counstraint heuristic" which reduced the
number of rows after WHERE check by 25% if there were comparisons
between key parts in table T and previous tables, like WHERE
T.keyXpartY= func(prev_table.cols)

Note that such comparisons can only be checked when the row of
table T is joined with rows of the previous tables. It is wrong
to apply the selectivity before the join operation.

Fixed by moving the 'found_constraint' code to a separate function
and only reducing the #rows in 'records_out'.

Renamed matching_candidates_in_table() to apply_selectivity_for_table() as
the function now either applies selectivity on the rows (depending
on the value of thd->variables.optimizer_use_condition_selectivity)
or uses the selectivity from the available range conditions.

5e5988db

Updated comments in best_access_path() · 33af691f
Monty authored Jan 24, 2023

33af691f

MDEV-30080 Wrong result with LEFT JOINs involving constant tables · 0eca91ab

Monty authored Jan 12, 2023

The reason things fails in 10.5 and above is that test_quick_select()
returns -1 (impossible range) for empty tables if there are any
conditions attached.

This didn't happen in 10.4 as the cost for a range was more than for
a table scan with 0 rows and get_key_scan_params() did not create any
range plans and thus did not mark the range as impossible.

The code that checked the 'impossible range' conditions did not take
into account all cases of LEFT JOIN usage.

Adding an extra check if the table is used with an ON condition in case
of 'impossible range' fixes the issue.

0eca91ab

Code cleanups and add some caching of functions to speed up things · 3316a54d

Monty authored Jan 10, 2023

Detailed description:
- Added more function comments and fixed types in some old comments
- Removed an outdated comment
- Cleaned up some functions in records.cc
  - Replaced "while" with "if"
  - Reused error code
  - Made functions similar
- Added caching of pfs_batch_update()
- Simplified some rowid_filter code
  - Only call build_range_rowid_filter() if rowid filter will be used
  - Replaced tab->is_rowid_filter_built with need_to_build_rowid_filter.
    We only have to test need_to_build_rowid_filter to know if we have
    to build the filter. Old code needed two tests
  - Added function 'clear_range_rowid_filter' to disable rowid filter.
    Made things simpler as we can now clear all rowid filter variables
    in one place.
- Removed some 'if' in sub_select()

3316a54d

MDEV-30360 Assertion `cond_selectivity <= 1.000000001' failed in ... · 65da5645

Monty authored Jan 09, 2023

The problem was that make_join_select() called test_quick_select() outside
of best_access_path(). This could use indexes that where not taken into
account before and this caused changes to selectivity and 'records_out'.

Fixed by updating records_out if test_quick_select() was called.

65da5645

MDEV-30328 Assertion `avg_io_cost != 0.0 || index_cost.io + row_cost.io == 0'... · 0a7d2917

Monty authored Jan 09, 2023

MDEV-30328 Assertion `avg_io_cost != 0.0 || index_cost.io + row_cost.io == 0' failed in Cost_estimate::total_cost()

The assert was there to check that engines reports sensible numbers for IO.
However this does not work in case of optimizer_disk_read_ratio=0.

Fixed by removing the assert.

0a7d2917

MDEV-30327 Client crashes in print_last_query_cost · 1a13dbff
Monty authored Jan 05, 2023
```
Fixed by calling init_pager() before tee_fprintf()
```
1a13dbff
MDEV-30313 Sporadic assertion `cond_selectivity <= 1.0' failure in get_range_limit_read_cost · 8d4bccf3
Monty authored Jan 05, 2023
```
The bug was related to floating point rounding. Fixed the assert to take
that into account.
```
8d4bccf3

Added sys.optimizer_switch_on() and sys.optimizer_switch_off() · 5de734da

Monty authored Jan 04, 2023

These are helpful tools to quickly see what optimizer switch options
are on or off.  The different options are displayed alphabetically

5de734da

Changed 'check_costs' so that --init-query can be used to override setup_engine() · 356a8601
Monty authored Jan 04, 2023

356a8601

MDEV-30310 Assertion failure in best_access_path upon IN exceeding... · 02b7735b

Monty authored Dec 28, 2022

MDEV-30310 Assertion failure in best_access_path upon IN exceeding IN_PREDICATE_CONVERSION_THRESHOLD, derived_with_keys=off

The bug was some old code that, without any explanation, reset
PART_KEY_FLAG from fields in temporary tables. This caused
join_tab->key_dependent to not be updated properly, which caused
an assert.

02b7735b

Simplified code in generate_derived_keys() and when using pos_in_tables · 4be0bfad

Monty authored Dec 28, 2022

Added comments that not used keys of derivied tables will be deleted.
Added some comments about checking if pos_in_table_list is 0.

Other things:
- Added a marker (DBTYPE_IN_PREDICATE) in TABLE_LIST->derived_type
  to indicate that the table was generated from IN (list). This is
  useful for debugging and can later be used by explain if needed.
- Removed a not needed test of table->pos_in_table_list as it should
  always be valid at this point in time.

4be0bfad

MDEV-30256 Wrong result (missing rows) upon join with empty table · 9a4110aa

Monty authored Dec 28, 2022

The problem was an assignment in test_quick_select() that flagged empty
tables with "Impossible where". This test was however wrong as it
didn't work correctly for left join.

Removed the test, but added checking of empty tables in DELETE and UPDATE
to get similar EXPLAIN as before.

The new tests is a bit more strict (better) than before as it catches all
cases of empty tables in single table DELETE/UPDATE.

9a4110aa

MDEV-30098 Server crashes in ha_myisam::index_read_map with index_merge_sort_intersection=on · e3f56254

Monty authored Dec 27, 2022

Fixes also
MDEV-30104 Server crashes in handler_rowid_filter_check upon ANALYZE TABLE

cancel_pushed_rowid_filter() didn't inform the handler that rowid_filter
was canceled.

e3f56254