Skip to content

Commit fe66fc6

Browse files
committed
Limit the number of index clauses considered in choose_bitmap_and().
classify_index_clause_usage() is O(N^2) in the number of distinct index qual clauses it considers, because of its use of a simple search list to store them. For nearly all queries, that's fine because only a few clauses will be considered. But Alexander Kuzmenkov reported a machine-generated query with 80000 (!) index qual clauses, which caused this code to take forever. Somewhat remarkably, this is the only O(N^2) behavior we now have for such a query, so let's fix it. We can get rid of the O(N^2) runtime for cases like this without much damage to the functionality of choose_bitmap_and() by separating out paths with "too many" qual or pred clauses, and deeming them to always be nonredundant with other paths. Then their clauses needn't go into the search list, so it doesn't get too long, but we don't lose the ability to consider bitmap AND plans altogether. I set the threshold for "too many" to be 100 clauses per path, which should be plenty to ensure no change in planning behavior for normal queries. There are other things we could do to make this go faster, but it's not clear that it's worth any additional effort. 80000 qual clauses require a whole lot of work in many other places, too. The code's been like this for a long time, so back-patch to all supported branches. The troublesome query only works back to 9.5 (in 9.4 it fails with stack overflow in the parser); so I'm not sure that fixing this in 9.4 has any real-world benefit, but perhaps it does. Discussion: https://postgr.es/m/90c5bdfa-d633-dabe-9889-3cf3e1acd443@postgrespro.ru
1 parent 6b41ccb commit fe66fc6

File tree

1 file changed

+31
-1
lines changed

1 file changed

+31
-1
lines changed

src/backend/optimizer/path/indxpath.c

Lines changed: 31 additions & 1 deletion
Original file line numberDiff line numberDiff line change
@@ -69,6 +69,7 @@ typedef struct
6969
List *quals; /* the WHERE clauses it uses */
7070
List *preds; /* predicates of its partial index(es) */
7171
Bitmapset *clauseids; /* quals+preds represented as a bitmapset */
72+
bool unclassifiable; /* has too many quals+preds to process? */
7273
} PathClauseUsage;
7374

7475
/* Callback argument for ec_member_matches_indexcol */
@@ -1385,9 +1386,18 @@ choose_bitmap_and(PlannerInfo *root, RelOptInfo *rel, List *paths)
13851386
Path *ipath = (Path *) lfirst(l);
13861387

13871388
pathinfo = classify_index_clause_usage(ipath, &clauselist);
1389+
1390+
/* If it's unclassifiable, treat it as distinct from all others */
1391+
if (pathinfo->unclassifiable)
1392+
{
1393+
pathinfoarray[npaths++] = pathinfo;
1394+
continue;
1395+
}
1396+
13881397
for (i = 0; i < npaths; i++)
13891398
{
1390-
if (bms_equal(pathinfo->clauseids, pathinfoarray[i]->clauseids))
1399+
if (!pathinfoarray[i]->unclassifiable &&
1400+
bms_equal(pathinfo->clauseids, pathinfoarray[i]->clauseids))
13911401
break;
13921402
}
13931403
if (i < npaths)
@@ -1422,6 +1432,10 @@ choose_bitmap_and(PlannerInfo *root, RelOptInfo *rel, List *paths)
14221432
* For each surviving index, consider it as an "AND group leader", and see
14231433
* whether adding on any of the later indexes results in an AND path with
14241434
* cheaper total cost than before. Then take the cheapest AND group.
1435+
*
1436+
* Note: paths that are either clauseless or unclassifiable will have
1437+
* empty clauseids, so that they will not be rejected by the clauseids
1438+
* filter here, nor will they cause later paths to be rejected by it.
14251439
*/
14261440
for (i = 0; i < npaths; i++)
14271441
{
@@ -1638,6 +1652,21 @@ classify_index_clause_usage(Path *path, List **clauselist)
16381652
result->preds = NIL;
16391653
find_indexpath_quals(path, &result->quals, &result->preds);
16401654

1655+
/*
1656+
* Some machine-generated queries have outlandish numbers of qual clauses.
1657+
* To avoid getting into O(N^2) behavior even in this preliminary
1658+
* classification step, we want to limit the number of entries we can
1659+
* accumulate in *clauselist. Treat any path with more than 100 quals +
1660+
* preds as unclassifiable, which will cause calling code to consider it
1661+
* distinct from all other paths.
1662+
*/
1663+
if (list_length(result->quals) + list_length(result->preds) > 100)
1664+
{
1665+
result->clauseids = NULL;
1666+
result->unclassifiable = true;
1667+
return result;
1668+
}
1669+
16411670
/* Build up a bitmapset representing the quals and preds */
16421671
clauseids = NULL;
16431672
foreach(lc, result->quals)
@@ -1655,6 +1684,7 @@ classify_index_clause_usage(Path *path, List **clauselist)
16551684
find_list_position(node, clauselist));
16561685
}
16571686
result->clauseids = clauseids;
1687+
result->unclassifiable = false;
16581688

16591689
return result;
16601690
}

0 commit comments

Comments
 (0)