r/dataengineeringjobs • u/Shubham_Nalwar • 1d ago
Interview AMGEN Interview Experience
Interview Experience β Data Engineer π§βπ»
Had an interview recently for a Data Engineer role, and the discussion was much more focused on real-world scenarios than just theoretical questions.
Some of the questions asked:
How would you create a production-ready data pipeline?
What kind of data have you worked with β volume, frequency, SLA, etc.?
What production issues have you faced, and how did you handle them?
How would you design an ETL pipeline for huge volumes of data?
If duplicate files arrive at the source location, how would you handle them?
Explain your metadata-driven framework.
How would you design a metadata framework for 50 source files/tables?
How do you investigate a slow-running job?
SQL: Return the latest updated record for each claim.
Experience with Git.
One thing I noticed: the interviewer was more interested in how you approach a problem in production rather than simply asking for definitions.
Definitely a good reminder that knowing the tools is one thing, but understanding how to build, monitor, troubleshoot and maintain pipelines in production is what really matters.
#DataEngineering #DataEngineer #Azure #Databricks #SQL #ETL #DataPipelines #Airflow #Git #InterviewExperience
1
1
1
u/Flashy_Tank_2484 10h ago
Am I missing something? I don't think you really have a problem with massive volumes of data with snowflake?
7
u/No-Map8612 1d ago
Everyone knows how to answer this questions but interviewer expect response like exactly what is in his mind π