Essential cookies keep authentication working. With your permission, we also use analytics cookies to understand and improve the product. Read our Privacy Policy

DataEngPrep.tech
QuestionsPracticeAI CoachDashboardPricingBlog
ProLogin

Interview Questions

Real questions from top companies in SQL

700+ Easy450+ Medium650+ Hard
All CategoriesBehavioralSpark/Big DataSQLPython/CodingSystem Design/ArchitectureCloud/ToolsGeneral/Othereasymediumhard
101

How would you prevent small file problems in S3 when loading data into Redshift?

SQLmediumetlpartitionspark0.6 min read
Capco
β†’
102

How would you retrieve the first and last order for each customer from a sales table?

SQLmediumpartitionwindow0.7 min read
Wipro
β†’
103

Identify and remove duplicate records from a table, keeping the most recent record based on a timestamp column.

SQLmediumpartitionsparksql0.6 min read
Goldman Sachs
β†’
104

Identify consecutive numbers in a column (at least 3 consecutive).

SQLeasy0.7 min read
Incedo
β†’
105

If manual partitions are created in a Hive data-warehouse table directory, and you query records from those partitions, will you see the data? If not, how can this be fixed?

SQLmediumpartition0.6 min read
Dunnhumby
β†’
106

Implement a CASE WHEN condition - medium difficulty

SQLmedium0.7 min read
Wolters Kluwer
β†’
107

In Python, process a large CSV in chunks and remove duplicate records based on email and timestamp.

SQLhardpython0.5 min read
Amazon
β†’
108

Indexing - True/False question on indexes and query optimization

SQLhardjoinoptimization0.6 min read
Myntra
β†’
109

Kafka Basics - architecture, topics, partitions, producers, consumers, Zookeeper

SQLhardjoinoptimizationpartition3.6 min read
Lumiq
β†’
110

Kafka Partitioning: How would you ensure even load distribution across Kafka partitions in a high-volume system?

SQLmediumpartition0.5 min read
BCG
β†’
111

Merge two dictionaries and remove keys with null values.

SQLeasypython0.5 min read
BCG
β†’
112

SQL Query Design: employees with highest salaries within each department

SQLhardjoinoptimizationpartition3.6 min read
Walmart
β†’
113

Schema Design: Star vs. Snowflake schema differences

SQLhardjoinoptimizationpartition3.6 min read
Zen Data Shastra
β†’
114

Snowflake Tech Stack: Deployment on Azure, cluster sizing considerations, and overall data warehouse design?

SQLhardjoinoptimizationpartition3.6 min read
Snowflake
β†’
115

Strategies for working with busy team leads?

SQLeasy0.6 min read
Snowflake
β†’
116

Tasks where the candidate faced failure and lessons learned.

SQLeasy0.5 min read
Thoughtworks
β†’
117

Tell us about a project where you optimized an existing process or pipeline. What was the impact?

SQLeasyairflowetl0.6 min read
Adidas
β†’
118

Teradata to Hadoop migration and handling data with SCD Type 2?

SQLmediumjoinpartitionspark0.7 min read
Citi
β†’
119

Test SQL skills using advanced window functions such as LAG, LEAD, and DENSE_RANK.

SQLmediumpartitionsqlwindow0.7 min read
Carelon
β†’
120

Time and cost comparisons for executing the same query in Snowflake and Spark.

SQLhardetloptimizationsnowflake0.7 min read
Carelon
β†’

Reading isn't practice. Get AI feedback on your answers.

Type or paste your answer to any of these questions and our AI Coach scores it, highlights gaps, and rewrites it at FAANG quality. Free to try.

Try AI Answer Coach β€” FreeStart a Mock Interview
Previous1...45678Next
Categories
All QuestionsSQLSpark / Big DataPython / CodingSystem DesignCloud / ToolsBehavioral
By Company
AmazonGoogleDatabricksSnowflakeAWSAzureMicrosoftNetflixUberTCS
Interview Guides
All GuidesTop SQL QuestionsTop Spark QuestionsPySpark QuestionsTop Python QuestionsTop System DesignKafka QuestionsAirflow QuestionsSQL Window FunctionsETL QuestionsData Modeling
Products
AI Interview CoachAnswer AnalyzerSQL PlaygroundResume AnalyzerAnswer Vault PDFsPricing
Company
About & Editorial PolicyContact UsAI DisclosureDisclaimerTerms of ServicePrivacy Policy
Β© 2026 DataEngPrep.tech. All rights reserved.
AboutBlogContactDisclaimer