在多个列上计数DISTINCT

是否有更好的方法来执行这样的查询:

SELECT COUNT(*) 
FROM (SELECT DISTINCT DocumentId, DocumentSessionId
      FROM DocumentOutputItems) AS internalQuery

我需要数一下这个表中不同项的数量，但不同项超过两列。

我的查询工作得很好，但我想知道我是否可以只使用一个查询(不使用子查询)得到最终结果

当前回答

如果您试图提高性能，可以尝试在两个列的散列或连接值上创建持久计算列。

一旦它被持久化，只要列是确定的，并且您使用的是“正常的”数据库设置，就可以对其建立索引和/或在其上创建统计信息。

我相信计算列的不同计数将等效于您的查询。

2009-09-26 03:42:34

其他回答

你的查询没有问题，但你也可以这样做:

WITH internalQuery (Amount)
AS
(
    SELECT (0)
      FROM DocumentOutputItems
  GROUP BY DocumentId, DocumentSessionId
)
SELECT COUNT(*) AS NumberOfDistinctRows
  FROM internalQuery

2009-09-24 13:37:10

当我在谷歌上搜索我自己的问题时，发现如果你计算DISTINCT对象，你会得到正确的返回数(我使用MySQL)

SELECT COUNT(DISTINCT DocumentID) AS Count1, 
  COUNT(DISTINCT DocumentSessionId) AS Count2
  FROM DocumentOutputItems

2013-04-12 16:31:07

如果您使用的是固定长度的数据类型，则可以将其转换为二进制，从而非常容易和快速地完成此操作。假设documententid和DocumentSessionId都是int，因此都是4字节长…

SELECT COUNT(DISTINCT CAST(DocumentId as binary(4)) + CAST(DocumentSessionId as binary(4)))
FROM DocumentOutputItems

My specific problem required me to divide a SUM by the COUNT of the distinct combination of various foreign keys and a date field, grouping by another foreign key and occasionally filtering by certain values or keys. The table is very large, and using a sub-query dramatically increased the query time. And due to the complexity, statistics simply wasn't a viable option. The CHECKSUM solution was also far too slow in its conversion, particularly as a result of the various data types, and I couldn't risk its unreliability.

然而，使用上述解决方案几乎没有增加查询时间(与简单使用SUM相比)，并且应该是完全可靠的!它应该能够帮助其他处于类似情况的人，所以我把它贴在这里。

2019-09-18 00:27:22

这对我很管用。在oracle中:

SELECT SUM(DECODE(COUNT(*),1,1,1))
FROM DocumentOutputItems GROUP BY DocumentId, DocumentSessionId;

在jpql:

SELECT SUM(CASE WHEN COUNT(i)=1 THEN 1 ELSE 1 END)
FROM DocumentOutputItems i GROUP BY i.DocumentId, i.DocumentSessionId;

2018-03-29 07:59:14

如果您试图提高性能，可以尝试在两个列的散列或连接值上创建持久计算列。

一旦它被持久化，只要列是确定的，并且您使用的是“正常的”数据库设置，就可以对其建立索引和/或在其上创建统计信息。

我相信计算列的不同计数将等效于您的查询。

2009-09-26 03:42:34

在多个列上计数DISTINCT

推荐文章

最新文章

标签