提交 · 0fef38da215cdc9b01b1b623c2e37d7414b91843 · Greenplum / Gpdb

23 2月, 2007 1 次提交

Turn the rangetable used by the executor into a flat list, and avoid storing · eab6b8b2

由 Tom Lane 提交于 2月 22, 2007

useless substructure for its RangeTblEntry nodes. (I chose to keep using the
same struct node type and just zero out the link fields for unneeded info,
rather than making a separate ExecRangeTblEntry type --- it seemed too
fragile to have two different rangetable representations.)

Along the way, put subplans into a list in the toplevel PlannedStmt node,
and have SubPlan nodes refer to them by list index instead of direct pointers.
Vadim wanted to do that years ago, but I never understood what he was on about
until now. It makes things a *whole* lot more robust, because we can stop
worrying about duplicate processing of subplans during expression tree
traversals. That's been a constant source of bugs, and it's finally gone.

There are some consequent simplifications yet to be made, like not using
a separate EState for subplans in the executor, but I'll tackle that later.

eab6b8b2

22 1月, 2007 1 次提交

Add COST and ROWS options to CREATE/ALTER FUNCTION, plus underlying pg_proc · 5a7471c3

由 Tom Lane 提交于 1月 22, 2007

columns procost and prorows, to allow simple user adjustment of the estimated
cost of a function call, as well as control of the estimated number of rows
returned by a set-returning function. We might eventually wish to extend this
to allow function-specific estimation routines, but there seems to be
consensus that we should try a simple constant estimate first. In particular
this provides a relatively simple way to control the order in which different
WHERE clauses are applied in a plan node, which is a Good Thing in view of the
fact that the recent EquivalenceClass planner rewrite made that much less
predictable than before.

5a7471c3

21 1月, 2007 1 次提交

Refactor planner's pathkeys data structure to create a separate, explicit · f41803bb

由 Tom Lane 提交于 1月 20, 2007

representation of equivalence classes of variables.  This is an extensive
rewrite, but it brings a number of benefits:
* planner no longer fails in the presence of "incomplete" operator families
that don't offer operators for every possible combination of datatypes.
* avoid generating and then discarding redundant equality clauses.
* remove bogus assumption that derived equalities always use operators
named "=".
* mergejoins can work with a variety of sort orders (e.g., descending) now,
instead of tying each mergejoinable operator to exactly one sort order.
* better recognition of redundant sort columns.
* can make use of equalities appearing underneath an outer join.

f41803bb

11 1月, 2007 1 次提交

Change the planner-to-executor API so that the planner tells the executor · a191a169

由 Tom Lane 提交于 1月 10, 2007

which comparison operators to use for plan nodes involving tuple comparison
(Agg, Group, Unique, SetOp).  Formerly the executor looked up the default
equality operator for the datatype, which was really pretty shaky, since it's
possible that the data being fed to the node is sorted according to some
nondefault operator class that could have an incompatible idea of equality.
The planner knows what it has sorted by and therefore can provide the right
equality operator to use.  Also, this change moves a couple of catalog lookups
out of the executor and into the planner, which should help startup time for
pre-planned queries by some small amount.  Modify the planner to remove some
other cavalier assumptions about always being able to use the default
operators.  Also add "nulls first/last" info to the Plan node for a mergejoin
--- neither the executor nor the planner can cope yet, but at least the API is
in place.

a191a169

06 1月, 2007 1 次提交
- B
  Update CVS HEAD for 2007 copyright. Back branches are typically not · 29dccf5f
  由 Bruce Momjian 提交于 1月 05, 2007
```
back-stamped for this.
```
  29dccf5f
12 8月, 2006 1 次提交

Add INSERT/UPDATE/DELETE RETURNING, with basic docs and regression tests. · 7a3e30e6

由 Tom Lane 提交于 8月 12, 2006

plpgsql support to come later.  Along the way, convert execMain's
SELECT INTO support into a DestReceiver, in order to eliminate some ugly
special cases.

Jonah Harris and Tom Lane

7a3e30e6

26 7月, 2006 1 次提交
- B
  Change LIMIT/OFFSET to use int8 · 085e5596
  由 Bruce Momjian 提交于 7月 26, 2006
```
Dhanaraj M
```
  085e5596
02 7月, 2006 1 次提交

Revise the planner's handling of "pseudoconstant" WHERE clauses, that is · cffd89ca

由 Tom Lane 提交于 7月 01, 2006

clauses containing no variables and no volatile functions. Such a clause
can be used as a one-time qual in a gating Result plan node, to suppress
plan execution entirely when it is false. Even when the clause is true,
putting it in a gating node wins by avoiding repeated evaluation of the
clause. In previous PG releases, query_planner() would do this for
pseudoconstant clauses appearing at the top level of the jointree, but
there was no ability to generate a gating Result deeper in the plan tree.
To fix it, get rid of the special case in query_planner(), and instead
process pseudoconstant clauses through the normal RestrictInfo qual
distribution mechanism. When a pseudoconstant clause is found attached to
a path node in create_plan(), pull it out and generate a gating Result at
that point. This requires special-casing pseudoconstants in selectivity
estimation and cost_qual_eval, but on the whole it's pretty clean.
It probably even makes the planner a bit faster than before for the normal
case of no pseudoconstants, since removing pull_constant_clauses saves one
useless traversal of the qual tree. Per gripe from Phil Frost.

cffd89ca

05 3月, 2006 1 次提交
- B
  
  Update copyright for 2006. Update scripts. · f2f5b056
  由 Bruce Momjian 提交于 3月 05, 2006
  
  f2f5b056
20 12月, 2005 1 次提交

Teach planner how to rearrange join order for some classes of OUTER JOIN. · e3b98527

由 Tom Lane 提交于 12月 20, 2005

Per my recent proposal. I ended up basing the implementation on the
existing mechanism for enforcing valid join orders of IN joins --- the
rules for valid outer-join orders are somewhat similar.

e3b98527

15 10月, 2005 1 次提交
- B
  
  Standard pgindent run for 8.1. · 1dc34982
  由 Bruce Momjian 提交于 10月 15, 2005
  
  1dc34982
29 9月, 2005 1 次提交

Repair planning bug introduced in 7.4: outer-join ON clauses that referenced · 2e1254e7

由 Tom Lane 提交于 9月 28, 2005

only the inner-side relation would be considered as potential equijoin clauses,
which is wrong because the condition doesn't necessarily hold above the point
of the outer join. Per test case from Kevin Grittner (bug#1916).

2e1254e7

28 8月, 2005 1 次提交

Change the division of labor between grouping_planner and query_planner · 4e5fbb34

由 Tom Lane 提交于 8月 27, 2005

so that the latter estimates the number of groups that grouping will
produce. This is needed because it is primarily query_planner that
makes the decision between fast-start and fast-finish plans, and in the
original coding it was unable to make more than a crude rule-of-thumb
choice when the query involved grouping. This revision helps us make
saner choices for queries like SELECT ... GROUP BY ... LIMIT, as in a
recent example from Mark Kirkwood. Also move the responsibility for
canonicalizing sort_pathkeys and group_pathkeys into query_planner;
this information has to be available anyway to support the first change,
and doing it this way lets us get rid of compare_noncanonical_pathkeys
entirely.

4e5fbb34

19 8月, 2005 1 次提交

Fix up LIMIT/OFFSET planning so that we cope with non-constant LIMIT · dfdf07aa

由 Tom Lane 提交于 8月 18, 2005

or OFFSET clauses by using estimate_expression_value(). The main advantage
of this is that if the expression is a Param and we have a value for the
Param, we'll use that value rather than defaulting. Also, fix some
thinkos in the logic for combining LIMIT/OFFSET with an externally
supplied tuple fraction (this covers cases like EXISTS(...LIMIT...)).
And make sure the results of all this are shown by EXPLAIN. Per a
gripe from Merlin Moncure.

dfdf07aa

06 6月, 2005 1 次提交

Remove planner's private fields from Query struct, and put them into · 9ab4d981

由 Tom Lane 提交于 6月 05, 2005

a new PlannerInfo struct, which is passed around instead of the bare
Query in all the planning code.  This commit is essentially just a
code-beautification exercise, but it does open the door to making
larger changes to the planner data structures without having to muck
with the widely-known Query struct.

9ab4d981

23 5月, 2005 1 次提交

Teach the planner to remove SubqueryScan nodes from the plan if they · e2159f38

由 Tom Lane 提交于 5月 22, 2005

aren't doing anything useful (ie, neither selection nor projection).
Also, extend to SubqueryScan the hacks already in place to avoid
unnecessary ExecProject calls when the result would just be the same
tuple the subquery already delivered. This saves some overhead in
UNION and other set operations, as well as avoiding overhead for
unflatten-able subqueries. Per example from Sokolov Yura.

e2159f38

25 4月, 2005 2 次提交

T
Replace slightly klugy create_bitmap_restriction() function with a · 79a1b002
由 Tom Lane 提交于 4月 25, 2005
```
more efficient routine in restrictinfo.c (which can make use of
make_restrictinfo_internal).
```
79a1b002

Remove support for OR'd indexscans internal to a single IndexScan plan · 5b051852

由 Tom Lane 提交于 4月 25, 2005

node, as this behavior is now better done as a bitmap OR indexscan.
This allows considerable simplification in nodeIndexscan.c itself as
well as several planner modules concerned with indexscan plan generation.
Also we can improve the sharing of code between regular and bitmap
indexscans, since they are now working with nigh-identical Plan nodes.

5b051852

12 4月, 2005 2 次提交

T
Fix oversight in MIN/MAX optimization: must not return NULL entries · 7ace43e0
由 Tom Lane 提交于 4月 12, 2005
```
from index, since the aggregates ignore NULLs.
```
7ace43e0

Create the planner mechanism for optimizing simple MIN and MAX queries · addc42c3

由 Tom Lane 提交于 4月 11, 2005

into indexscans on matching indexes. For the moment, it only handles
int4 and text datatypes; next step is to add a column to pg_aggregate
so that all MIN/MAX aggregates can be handled. Per my recent proposal.

addc42c3

11 3月, 2005 1 次提交

Make the behavior of HAVING without GROUP BY conform to the SQL spec. · 595ed2a8

由 Tom Lane 提交于 3月 10, 2005

Formerly, if such a clause contained no aggregate functions we mistakenly
treated it as equivalent to WHERE. Per spec it must cause the query to
be treated as a grouped query of a single group, the same as appearance
of aggregate functions would do. Also, the HAVING filter must execute
after aggregate function computation even if it itself contains no
aggregate functions.

595ed2a8

01 1月, 2005 1 次提交

· 2ff50159

由 PostgreSQL Daemon 提交于 12月 31, 2004

Tag appropriate files for rc3

Also performed an initial run through of upgrading our Copyright date to
extend to 2005 ... first run here was very simple ... change everything
where: grep 1996-2004 && the word 'Copyright' ... scanned through the
generated list with 'less' first, and after, to make sure that I only
picked up the right entries ...

2ff50159

29 8月, 2004 1 次提交
- B
  
  Update copyright to 2004. · da9a8649
  由 Bruce Momjian 提交于 8月 29, 2004
  
  da9a8649
18 1月, 2004 1 次提交

When testing whether a sub-plan can do projection, use a general-purpose · 6bdfde9a

由 Tom Lane 提交于 1月 18, 2004

check instead of hardwiring assumptions that only certain plan node types
can appear at the places where we are testing. This was always a pretty
fragile assumption, and it turns out to be broken in 7.4 for certain cases
involving IN-subselect tests that need type coercion.
Also, modify code that builds finished Plan tree so that node types that
don't do projection always copy their input node's targetlist, rather than
having the tlist passed in from the caller. The old method makes it too
easy to write broken code that thinks it can modify the tlist when it
cannot.

6bdfde9a

30 11月, 2003 1 次提交

· 55b11325

由 PostgreSQL Daemon 提交于 11月 29, 2003

make sure the $Id tags are converted to $PostgreSQL as well ...

55b11325

09 8月, 2003 1 次提交
- B
  
  Another pgindent run with updated typedefs. · 46785776
  由 Bruce Momjian 提交于 8月 08, 2003
  
  46785776
04 8月, 2003 2 次提交
- B
  
  Update copyrights to 2003. · f3c3deb7
  由 Bruce Momjian 提交于 8月 04, 2003
  
  f3c3deb7
- B
  
  pgindent run. · 089003fb
  由 Bruce Momjian 提交于 8月 04, 2003
  
  089003fb
30 6月, 2003 1 次提交

Restructure building of join relation targetlists so that a join plan · 835bb975

由 Tom Lane 提交于 6月 29, 2003

node emits only those vars that are actually needed above it in the
plan tree. (There were comments in the code suggesting that this was
done at some point in the dim past, but for a long time we have just
made join nodes emit everything that either input emitted.) Aside from
being marginally more efficient, this fixes the problem noted by Peter
Eisentraut where a join above an IN-implemented-as-join might fail,
because the subplan targetlist constructed in the latter case didn't
meet the expectation of including everything.
Along the way, fix some places that were O(N^2) in the targetlist
length. This is not all the trouble spots for wide queries by any
means, but it's a step forward.

835bb975

29 6月, 2003 1 次提交

Support expressions of the form 'scalar op ANY (array)' and · bee21792

由 Tom Lane 提交于 6月 29, 2003

'scalar op ALL (array)', where the operator is applied between the
lefthand scalar and each element of the array.  The operator must
yield boolean; the result of the construct is the OR or AND of the
per-element results, respectively.

Original coding by Joe Conway, after an idea of Peter's.  Rewritten
by Tom to keep the implementation strictly separate from subqueries.

bee21792

06 5月, 2003 1 次提交

Implement feature of new FE/BE protocol whereby RowDescription identifies · 2cf57c8f

由 Tom Lane 提交于 5月 06, 2003

the column by table OID and column number, if it's a simple column
reference. Along the way, get rid of reskey/reskeyop fields in Resdoms.
Turns out that representation was not convenient for either the planner
or the executor; we can make the planner deliver exactly what the
executor wants with no more effort.
initdb forced due to change in stored rule representation.

2cf57c8f

10 3月, 2003 1 次提交

Restructure parsetree representation of DECLARE CURSOR: now it's a · aa83bc04

由 Tom Lane 提交于 3月 10, 2003

utility statement (DeclareCursorStmt) with a SELECT query dangling from
it, rather than a SELECT query with a few unusual fields in it. Add
code to determine whether a planned query can safely be run backwards.
If DECLARE CURSOR specifies SCROLL, ensure that the plan can be run
backwards by adding a Materialize plan node if it can't. Without SCROLL,
you get an error if you try to fetch backwards from a cursor that can't
handle it. (There is still some discussion about what the exact
behavior should be, but this is necessary infrastructure in any case.)
Along the way, make EXPLAIN DECLARE CURSOR work.

aa83bc04

24 1月, 2003 1 次提交

Modify planner's implied-equality-deduction code so that when a set · f5e83662

由 Tom Lane 提交于 1月 24, 2003

of known-equal expressions includes any constant expressions (including
Params from outer queries), we actively suppress any 'var = var'
clauses that are or could be deduced from the set, generating only the
deducible 'var = const' clauses instead. The idea here is to push down
the restrictions implied by the equality set to base relations whenever
possible. Once we have applied the 'var = const' clauses, the 'var = var'
clauses are redundant, and should be suppressed both to save work at
execution and to avoid double-counting restrictivity.

f5e83662

21 1月, 2003 1 次提交

IN clauses appearing at top level of WHERE can now be handled as joins. · bdfbfde1

由 Tom Lane 提交于 1月 20, 2003

There are two implementation techniques: the executor understands a new
JOIN_IN jointype, which emits at most one matching row per left-hand row,
or the result of the IN's sub-select can be fed through a DISTINCT filter
and then joined as an ordinary relation.
Along the way, some minor code cleanup in the optimizer; notably, break
out most of the jointree-rearrangement preprocessing in planner.c and
put it in a new file prep/prepjointree.c.

bdfbfde1

16 1月, 2003 2 次提交

Now that switch_outer processing no longer relies on being run after · cde9f852

由 Tom Lane 提交于 1月 15, 2003

join_references(), it's practical to consolidate all join_references()
processing into the set_plan_references traversal in setrefs.c. This
seems considerably cleaner than the old way where we did it for join
quals in createplan.c and for targetlists in setrefs.c.

cde9f852

Allow merge and hash joins to occur on arbitrary expressions (anything not · de97072e

由 Tom Lane 提交于 1月 15, 2003

containing a volatile function), rather than only on 'Var = Var' clauses
as before. This makes it practical to do flatten_join_alias_vars at the
start of planning, which in turn eliminates a bunch of klugery inside the
planner to deal with alias vars. As a free side effect, we now detect
implied equality of non-Var expressions; for example in
SELECT ... WHERE a.x = b.y and b.y = 42
we will deduce a.x = 42 and use that as a restriction qual on a. Also,
we can remove the restriction introduced 12/5/02 to prevent pullup of
subqueries whose targetlists contain sublinks.
Still TODO: make statistical estimation routines in selfuncs.c and costsize.c
smarter about expressions that are more complex than plain Vars. The need
for this is considerably greater now that we have to be able to estimate
the suitability of merge and hash join techniques on such expressions.

de97072e

12 12月, 2002 1 次提交

Phase 2 of read-only-plans project: restructure expression-tree nodes · a0bf885f

由 Tom Lane 提交于 12月 12, 2002

so that all executable expression nodes inherit from a common supertype
Expr. This is somewhat of an exercise in code purity rather than any
real functional advance, but getting rid of the extra Oper or Func node
formerly used in each operator or function call should provide at least
a little space and speed improvement.
initdb forced by changes in stored-rules representation.

a0bf885f

21 11月, 2002 1 次提交

Finish implementation of hashed aggregation. Add enable_hashagg GUC · 6c1d4662

由 Tom Lane 提交于 11月 21, 2002

parameter to allow it to be forced off for comparison purposes.
Add ORDER BY clauses to a bunch of regression test queries that will
otherwise produce randomly-ordered output in the new regime.

6c1d4662

20 11月, 2002 1 次提交

Add an at-least-marginally-plausible method of estimating the number · b60be3f2

由 Tom Lane 提交于 11月 19, 2002

of groups produced by GROUP BY. This improves the accuracy of planning
estimates for grouped subselects, and is needed to check whether a
hashed aggregation plan risks memory overflow.

b60be3f2

06 11月, 2002 1 次提交

First phase of implementing hash-based grouping/aggregation. An AGG plan · f6dba10e

由 Tom Lane 提交于 11月 06, 2002

node now does its own grouping of the input rows, and has no need for a
preceding GROUP node in the plan pipeline. This allows elimination of
the misnamed tuplePerGroup option for GROUP, and actually saves more code
in nodeGroup.c than it costs in nodeAgg.c, as well as being presumably
faster. Restructure the API of query_planner so that we do not commit to
using a sorted or unsorted plan in query_planner; instead grouping_planner
makes the decision. (Right now it isn't any smarter than query_planner
was, but that will change as soon as it has the option to select a hash-
based aggregation step.) Despite all the hackery, no initdb needed since
only in-memory node types changed.

f6dba10e