Moving average dan total

Ringkasan Statistik dan Window Functions di PostgreSQL

Michel Semaan

Moving averages

Ringkasan

  • Moving average (MA): Rata-rata dari n periode terakhir
    • Contoh: MA 10 hari untuk unit terjual adalah rata-rata unit yang terjual pada 10 hari terakhir
    • Dipakai untuk menunjukkan momentum/tren
    • Juga membantu menghilangkan musiman
  • Moving total: Jumlah dari n periode terakhir
    • Contoh: Jumlah medali pada 3 Olimpiade terakhir
    • Dipakai untuk menunjukkan kinerja; jika jumlah turun, kinerja keseluruhan turun
Ringkasan Statistik dan Window Functions di PostgreSQL

Tabel sumber

Kueri

SELECT
  Year, COUNT(*) AS Medals
FROM Summer_Medals
WHERE
  Country = 'USA'
  AND Medal = 'Gold'
  AND Year >= 1980
GROUP BY Year
ORDER BY Year ASC;

Hasil

| Year | Medals |
|------|--------|
| 1984 | 168    |
| 1988 | 77     |
| 1992 | 89     |
| 1996 | 160    |
| 2000 | 130    |
| 2004 | 116    |
| 2008 | 125    |
| 2012 | 147    |
Ringkasan Statistik dan Window Functions di PostgreSQL

Moving average

Kueri

WITH US_Medals AS (...)

SELECT
  Year, Medals,
  AVG(Medals) OVER
    (ORDER BY Year ASC
     ROWS BETWEEN
     2 PRECEDING AND CURRENT ROW) AS Medals_MA
FROM US_Medals
ORDER BY Year ASC;

Hasil

| Year | Medals | Medals_MA |
|------|--------|-----------|
| 1984 | 168    | 168.00    |
| 1988 | 77     | 122.50    |
| 1992 | 89     | 111.33    |
| 1996 | 160    | 108.67    |
| 2000 | 130    | 126.33    |
| 2004 | 116    | 135.33    |
| 2008 | 125    | 123.67    |
| 2012 | 147    | 129.33    |
Ringkasan Statistik dan Window Functions di PostgreSQL

Moving total

Kueri

WITH US_Medals AS (...)

SELECT
  Year, Medals,
  SUM(Medals) OVER
    (ORDER BY Year ASC
     ROWS BETWEEN
     2 PRECEDING AND CURRENT ROW) AS Medals_MT
FROM US_Medals
ORDER BY Year ASC;

Hasil

| Year | Medals | Medals_MT |
|------|--------|-----------|
| 1984 | 168    | 168       |
| 1988 | 77     | 245       |
| 1992 | 89     | 334       |
| 1996 | 160    | 326       |
| 2000 | 130    | 379       |
| 2004 | 116    | 406       |
| 2008 | 125    | 371       |
| 2012 | 147    | 388       |
Ringkasan Statistik dan Window Functions di PostgreSQL

ROWS vs RANGE

  • RANGE BETWEEN [START] AND [FINISH]
    • Fungsinya mirip dengan ROWS BETWEEN
    • RANGE menganggap duplikat di subklausa ORDER BY dari OVER sebagai satu entitas

Tabel

| Year | Medals | Rows_RT | Range_RT |
|------|--------|---------|----------|
| 1992 | 10     | 10      | 10       |
| 1996 | 50     | 60      | 110      |
| 2000 | 50     | 110     | 110      |
| 2004 | 60     | 170     | 230      |
| 2008 | 60     | 230     | 230      |
| 2012 | 70     | 300     | 300      |
  • ROWS BETWEEN hampir selalu dipakai dibanding RANGE BETWEEN
Ringkasan Statistik dan Window Functions di PostgreSQL

Ayo berlatih!

Ringkasan Statistik dan Window Functions di PostgreSQL

Preparing Video For Download...