Real questions from top companies Β· medium
Kafka Partitioning: How would you ensure even load distribution across Kafka partitions in a high-volume system?
Teradata to Hadoop migration and handling data with SCD Type 2?
Test SQL skills using advanced window functions such as LAG, LEAD, and DENSE_RANK.
Two tables: Table 1 has 8 records, Table 2 has 2 records, common column id. How many records would result with Inner Join, Left Join, Right Join, Full Outer Join?
What are BigQuery Slots?
What are Slowly Changing Dimensions (SCD), and how would you implement them for tracking customer data changes?
What are partitioning strategies in Redshift?
What are some best practices for writing efficient SQL queries?
What are the differences between normalization and denormalization? When would you use a denormalized structure?
What challenges arise with duplicate records, and how do you address them?
What factors determine the optimal number of partitions for a large file?
What is Left Anti Join and its use case?
What is Redshift Spectrum, and how does it differ from standard Redshift queries?
What is UNNEST and provide a query example?
What is a Kafka topic, and how do you choose the number of partitions for it?
What is a cross-join?
What is a semi-join?
What is dynamic partition pruning, and how does it optimize query execution?
What is the difference between UNION and UNION ALL? Which one is faster and why?
What is the difference between static and dynamic partitioning in Hive?
Type or paste your answer to any of these questions and our AI Coach scores it, highlights gaps, and rewrites it at FAANG quality. Free to try.