Candidate telemetry diagnostic, error autopsy, and step-by-step query construction walkthrough.
Live aggregated metrics across candidate sandbox attempts
20 solved
First attempt fail
Evaluated submissions
Median time to solve
Unlocked answer
The analytics team wants to track how genre popularity changes over time. Calculate the year-over-year revenue change and growth percentage for key genres.
| Column | Type |
|---|---|
| InvoiceId | INTEGER (Primary Key) |
| CustomerId | INTEGER (Foreign Key → Customer.CustomerId) |
| InvoiceDate | TIMESTAMP |
| BillingAddress | TEXT |
| BillingCity | TEXT |
| BillingState | TEXT |
| BillingCountry | TEXT |
| BillingPostalCode | TEXT |
| Total | NUMERIC(10,2) |
| Column | Type |
|---|---|
| GenreId | INTEGER (Primary Key) |
| Name | TEXT |
LAG() partitioned by genre to get previous year's revenueYoYGrowth percentage (NULL for first year)Your result should have 20 rows with 5 columns: | genrename | year | revenue | prevyearrevenue | yoygrowth | |-----------|------|---------|-----------------|-----------| | Jazz | 2009 | 19.8 | | | | Jazz | 2010 | 15.84 | 19.8 | -20.0 | | Jazz | 2011 | 15.84 | 15.84 | 0.0 | | Jazz | 2012 | 5.94 | 15.84 | -62.5 | | Jazz | 2013 | 21.78 | 5.94 | 266.67 | | ... | ... | ... | ... | ... |
Attempting to filter window function output directly inside the WHERE clause (e.g., WHERE ROW_NUMBER() OVER (...) <= 3). Because SQL executes WHERE before evaluating window functions, this raises a syntax error or produces invalid groupings. The calculation must be staged in a CTE or subquery first.
Interviewers use this question to verify whether you understand the exact SQL execution order (FROM -> WHERE -> GROUP BY -> HAVING -> WINDOW -> SELECT -> ORDER BY), how to choose correctly between ROW_NUMBER, RANK, and DENSE_RANK when handling ties, and how to partition datasets without collapsing rows.
Construct the solution logically from first principles to avoid typical edge case pitfalls.
Determine whether the ranking or running total resets per customer, department, or genre (PARTITION BY), or spans the entire table.
OVER (PARTITION BY <group_col> ORDER BY <order_col> DESC)
Write a WITH clause to calculate the window metric alongside the base columns, ensuring all join and filter conditions are applied.
WITH RankedData AS (
WITH YearlyGenreRevenue AS (
SELECT g.Name AS GenreName,
EXTRACT(YEAR FROM i.InvoiceDate)::int AS Year,
...
)Select from the CTE and apply the outer predicate (e.g., WHERE rnk = 1 or WHERE rnk <= N) to extract the final result set.
SELECT <columns> FROM RankedData WHERE rnk = 1 ORDER BY <columns>;
WITH YearlyGenreRevenue AS (
SELECT g.Name AS GenreName,
EXTRACT(YEAR FROM i.InvoiceDate)::int AS Year,
ROUND(SUM(il.UnitPrice * il.Quantity), 2) AS Revenue
FROM InvoiceLine il
JOIN Track t ON il.TrackId = t.TrackId
JOIN Genre g ON t.GenreId = g.GenreId
JOIN Invoice i ON il.InvoiceId = i.InvoiceId
GROUP BY g.GenreId, g.Name, EXTRACT(YEAR FROM i.InvoiceDate)
)
SELECT GenreName,
Year,
Revenue,
LAG(Revenue) OVER (PARTITION BY GenreName ORDER BY Year) AS PrevYearRevenue,
CASE
WHEN LAG(Revenue) OVER (PARTITION BY GenreName ORDER BY Year) IS NULL THEN NULL
ELSE ROUND((Revenue - LAG(Revenue) OVER (PARTITION BY GenreName ORDER BY Year)) * 100.0 / LAG(Revenue) OVER (PARTITION BY GenreName ORDER BY Year), 2)
END AS YoYGrowth
FROM YearlyGenreRevenue
WHERE GenreName IN ('Rock', 'Latin', 'Metal', 'Jazz')
ORDER BY GenreName, Year
LIMIT 20;Real code patterns candidates submit that fail the grading suite.
SELECT * FROM table_name WHERE ROW_NUMBER() OVER (ORDER BY amount DESC) <= 5;
SELECT department_id, employee_id, salary,
RANK() OVER (ORDER BY salary DESC) as rank
FROM employees;Three recurring syntax and semantic traps relevant to this problem domain.
Window functions cannot appear in WHERE or HAVING clauses. Filtering on a rank or running total requires wrapping the query in a CTE or subquery.
SELECT *, RANK() OVER (ORDER BY points DESC) as rnk FROM candidates WHERE RANK() OVER (ORDER BY points DESC) <= 5; -- ❌ Syntax Error
WITH Ranked AS ( SELECT *, RANK() OVER (ORDER BY points DESC) as rnk FROM candidates ) SELECT * FROM Ranked WHERE rnk <= 5; -- ✅ Correct
Using RANK() skips rank positions on ties (1, 2, 2, 4), whereas DENSE_RANK() retains consecutive integers (1, 2, 2, 3). Using ROW_NUMBER() arbitrarily breaks ties.
SELECT name, RANK() OVER (ORDER BY score DESC) as rnk ... -- ❌ Might miss 3rd rank if 2nd ties
SELECT name, DENSE_RANK() OVER (ORDER BY score DESC) as rnk ... -- ✅ Guaranteed continuous ranks
Forgetting the PARTITION BY clause causes ranking or rolling metrics to compute across the entire dataset rather than resetting per group/customer.
ROW_NUMBER() OVER (ORDER BY sale_date DESC) -- ❌ Global row number
ROW_NUMBER() OVER (PARTITION BY customer_id ORDER BY sale_date DESC) -- ✅ Per-customer rank
Launch our in-browser coding environment. Run queries, view execution plans, and get instant comparative diff grading with no setup.