DataDrivenDataDrivenDataDriven
LearnPracticeInterviewDailyJobsCommunity

wild_storm_5937

Rank #73
DataDrivenDataDriven
Exemplar
Top 2.2%
Member Nº 3040
Joined Jul 2026
SQLPYSPKDMPA
SQL
Correlated Subquery
Common Table Expressions
Array Aggregation
Python
f-strings
Variables
Arithmetic
Spark
Window Deduplication
Collecting Values
Exploding Arrays
Data Modeling
Self-Referential
Many-to-Many
One-to-One
Learned
210LEARNED
SQL171
Spark25
Python12
Pipeline Architecture2
Milestones
Top 100
7-day streak
1,168 runs in the past 26 weeksActive days: 22 · Max streak: 9
MarAprMayJunJulAugSep
Recent solves
List the product categories Spark will process
Spark · easy · lesson
2 weeks ago
Select the names of in-stock products
Spark · easy · lesson
2 weeks ago
Count products per category (your first shuffle)
Spark · easy · lesson
2 weeks ago
Capstone: revenue by order status
Spark · medium · lesson
4 weeks ago
Capstone: average line value per category
Spark · medium · lesson
4 weeks ago
Aggregate, do not collect: top rating per category
Spark · medium · lesson
4 weeks ago
Sized by config: average price per category
Spark · medium · lesson
4 weeks ago
No shuffle, no driver round-trip: line totals
Spark · medium · lesson
4 weeks ago
A pipelined narrow chain: discounted prices
Spark · medium · lesson
4 weeks ago
Parallelism by key: page views per device
Spark · medium · lesson
4 weeks ago

Product

  • Lessons
  • Practice Problems
  • Mock Interview
  • Daily Challenge
  • Community
  • Discuss

Resources

  • All Resources
  • Blog
  • About
  • Contact

Pipeline architecture tips, weekly

SQL patterns, Python tricks, and pipeline architecture insights for your next interview. Sent every Wednesday.

TermsPrivacySecurityHelp

© DataDriven

We help you get hired at these companies, but we don't work for them. All trademarks belong to their respective owners.