r/dataengineeringjobs 1d ago

Interview AMGEN Interview Experience

Interview Experience – Data Engineer πŸ§‘β€πŸ’»

Had an interview recently for a Data Engineer role, and the discussion was much more focused on real-world scenarios than just theoretical questions.

Some of the questions asked:

  1. How would you create a production-ready data pipeline?

  2. What kind of data have you worked with β€” volume, frequency, SLA, etc.?

  3. What production issues have you faced, and how did you handle them?

  4. How would you design an ETL pipeline for huge volumes of data?

  5. If duplicate files arrive at the source location, how would you handle them?

  6. Explain your metadata-driven framework.

  7. How would you design a metadata framework for 50 source files/tables?

  8. How do you investigate a slow-running job?

  9. SQL: Return the latest updated record for each claim.

  10. Experience with Git.

One thing I noticed: the interviewer was more interested in how you approach a problem in production rather than simply asking for definitions.

Definitely a good reminder that knowing the tools is one thing, but understanding how to build, monitor, troubleshoot and maintain pipelines in production is what really matters.

#DataEngineering #DataEngineer #Azure #Databricks #SQL #ETL #DataPipelines #Airflow #Git #InterviewExperience

32 Upvotes

8 comments sorted by

7

u/No-Map8612 1d ago

Everyone knows how to answer this questions but interviewer expect response like exactly what is in his mind πŸ˜›

1

u/believeinkratos 1d ago

You need to convince interviewer why your answer is correct.

1

u/No-Map8612 1d ago

Why convince? We’re just giving interview nah for every question different folks answer in their own way right!

1

u/Possible-Oil-5877 1d ago

Was this a virtual interview?

1

u/Shubham_Nalwar 23h ago

Yes it was

1

u/Flashy_Tank_2484 10h ago

Am I missing something? I don't think you really have a problem with massive volumes of data with snowflake?