Interview Questions

Real questions from top companies · easy

700+ Easy450+ Medium650+ Hard

All Categories Behavioral Spark/Big Data SQL Python/Coding System Design/Architecture Cloud/Tools General/Othereasy medium hard

701

What is the difference between managed and external tables in Hive or Spark SQL?

Spark/Big Dataeasysparksql0.3 min read

Infosys

→

702

What is the difference between map and flatMap in Spark transformations?

Spark/Big Dataeasyspark0.4 min read

Coforge

→

703

What is the purpose of the VACUUM command in Delta Lake?

Spark/Big Dataeasy0.4 min read

Puma

→

704

What limitations do you face when using Delta Tables in a multi-cloud environment?

Spark/Big Dataeasylakehouse0.4 min read

PWC

→

705

What metrics do you use to determine whether a Spark job is going well or not?

Spark/Big Dataeasyspark0.5 min read

Delivery Hero

→

706

Which Spark version are you using in your project, and why did you choose it?

Spark/Big Dataeasypythonspark0.3 min read

Capgemini

→

707

Why does Hive use Derby by default, and what alternatives are used in production?

Spark/Big Dataeasysparksql0.3 min read

Chryselys

→

708

Worked with UDFs - share examples

Spark/Big Dataeasypython0.3 min read

LTIMindtree

→

709

Write PySpark code to filter and count records.

Spark/Big Dataeasypythonsparksql0.3 min read

Bitwise

→

710

Write PySpark code to filter records based on specific conditions and add a calculated column.

Spark/Big Dataeasypythonsparksql0.3 min read

Bristol Myers Squibb

→

711

Write a PySpark code snippet to filter rows with a specific condition.

Spark/Big Dataeasypythonsparksql0.3 min read

Fragma Data Systems

→

712

Write the Spark command to rename an existing column in a DataFrame.

Spark/Big Dataeasyspark0.5 min read

Dunnhumby

→

713

Writing Excel sheets to Delta tables in Databricks

Spark/Big Dataeasyspark0.5 min read

Nihilent

→

714

You are given 10 worker machines with 100 GB RAM and 25 CPU cores. How would you determine the number of executors and the size of each executor?

Spark/Big Dataeasyspark0.7 min read

Meesho

→

715

How do you handle exceptions in data ingestion?

System Design/Architectureeasy2.2 min read

Gartner

→

Reading isn't practice. Get AI feedback on your answers.

Type or paste your answer to any of these questions and our AI Coach scores it, highlights gaps, and rewrites it at FAANG quality. Free to try.

Try AI Answer Coach — Free Start a Mock Interview

Previous 1...34 35 36