Associate Data Engineer
Amgen · Hyderabad
- Experience2–4 yrs
- SalaryNot disclosed
- Work modeonsite
- Levelmid
- Posted12 Sept 2026
About Amgen
Amgen is hiring in Hyderabad in pharma biotech. This role looks for around 2+ years of experience.
Skills
- Databricks
- Python
- PySpark
- Scala
- SQL
- PostgreSQL
- MySQL
- Apache Hadoop
- Apache Spark
- Apache Kafka
The role
A data engineer at a pharmaceutical biotechnology company designs scalable data pipelines and ETL workflows using Databricks and PySpark, applies big data processing, and builds cloud data solutions with AWS. SQL and data modeling support reliable integration across business systems.
Full job description
Role Description:
The role is responsible for designing, building, maintaining, analyzing, and interpreting data to provide actionable insights that drive business decisions. This role involves working with large datasets, developing reports, supporting and executing data governance initiatives and, visualizing data to ensure data is accessible, reliable, and efficiently managed. The ideal candidate has strong technical skills, experience with big data technologies, and a deep understanding of data architecture and ETL processes
Roles & Responsibilities:
Design, develop, and maintain data solutions for data generation, collection, and processingBe a key team member that assists in design and development of the data pipelineCreate data pipelines and ensure data quality by implementing ETL processes to migrate and deploy data across systemsContribute to the design, development, and implementation of data pipelines, ETL/ELT processes, and data integration solutionsTake ownership of data pipeline projects from inception to deployment, manage scope, timelines, and risksCollaborate with cross-functional teams to understand data requirements and design solutions that meet business needsDevelop and maintain data models, data dictionaries, and other documentation to ensure data accuracy and consistencyImplement data security and privacy measures to protect sensitive dataVery good understanding of DatabricksLeverage cloud platforms (AWS preferred) to build scalable and efficient data solutionsCollaborate and communicate effectively with product teamsCollaborate with Data Architects, Business SMEs, and Data Scientists to design and develop end-to-end data pipelines to meet fast paced business needs across geographic regionsIdentify and resolve complex data-related challengesAdhere to best practices for coding, testing, and designing reusable code/componentExplore new tools and technologies that will help to improve ETL platform performanceParticipate in sprint planning meetings and provide estimations on technical implementation
Basic Qualifications and Experience:
Bachelor’s degree and 2 to 4 years of Computer Science, IT or related field experience
Functional Skills:
Must-Have Skills
Proficiency in Databricks ,Python, PySpark, and Scala for data processing and ETL (Extract, Transform, Load) workflows, with hands-on experience in using Databricks for building ETL pipelines and handling big data processingStrong knowledge of SQL and experience with relational (e.g., PostgreSQL, MySQL) databases.Familiarity with big data frameworks like Apache Hadoop, Spark, and Kafka for handling large datasets.
Good-to-Have Skills:
Experience with cloud platforms such as AWS particularly in data services (e.g., EKS, EC2, S3, EMR, RDS, Redshift/Spectrum, Lambda, Glue, Athena)Understanding of data modeling, data warehousing, and data integration conceptsUnderstanding of machine learning pipelines and frameworks for ML/AI models