Amgen Technology Pvt Ltd. Posted August 4, 2026

Data Engineer

Hyderabad, India Full time
Data & Digital

Amgen Technology Pvt Ltd. is the source of truth for this posting and owns the application process. We surface normalized context and market comparison you won't find on the original listing.

About this opportunity

Career Category

Engineering

Job Description

ABOUT AMGEN

Amgen harnesses the best of biology and technology to fight the world’s toughest diseases, and make people’s lives easier, fuller and longer. We discover, develop, manufacture and deliver innovative medicines to help millions of patients. Amgen helped establish the biotechnology industry more than 40 years ago and remains on the cutting-edge of innovation, using technology and human genetic data to push beyond what’s known today.

Role Description

Let’s do this. Let’s change the world.

We are seeking an experienced Data Engineer to design, develop, and support scalable data pipelines and integration solutions. This role will work with large and complex datasets to ensure data is reliable, accessible, secure, and available for analytics, reporting, and AI use cases.

The ideal candidate has strong hands-on experience with Databricks, Apache Spark, Python, SQL, cloud technologies, data modeling, and ETL/ELT processes. The candidate will collaborate with data architects, business teams, data scientists, product teams, and DevOps teams to deliver high-quality data solutions.

Experience in biotechnology, pharmaceutical , manufacturing, life sciences, or another regulated industry is preferred.

Roles and Responsibilities

Design, develop, test, and maintain scalable data pipelines and data-integration solutions.

Build ETL/ELT processes for structured, semi-structured, and unstructured data.

Integrate data from enterprise applications, databases, APIs, cloud platforms, and third-party systems.

Contribute to the technical design and implementation of end-to-end data solutions.

Develop reusable and maintainable Python, PySpark , Spark SQL, and SQL components.

Implement data-quality checks, validation rules, reconciliation processes, logging, and exception handling.

Optimize Spark workloads, SQL queries, partitioning, and data-processing performance.

Develop and maintain data models, data dictionaries, mappings, and technical documentation.

Implement data-security , privacy, governance, and role-based access requirements.

Support workflow orchestration, scheduling, monitoring, alerting, and recovery processes.

Contribute to CI/CD pipelines, automated testing, version control, and deployment processes.

Troubleshoot data-pipeline failures, performance issues, and data-quality problems.

Collaborate with data architects, business SMEs, analysts, data scientists, product teams, and DevOps teams.

Participate in sprint planning, backlog refinement, technical estimation, and delivery activities.

Take ownership of assigned data-engineering work from development through deployment and production support.

Evaluate new technologies and recommend improvements to data-engineering processes and platform performance.

Follow coding, testing, documentation, security, and reusable-development standards.

Participate in operational support activities, including occasional off- hours support.

Basic Qualifications

Master’s or Bachelor’s degree in Computer Science , Information Technology, Engineering, Data Science, or a related field, with 5–8 years of relevant professional experience.

Must-Have Skills

Hands-on experience with Databricks, Apache Spark, PySpark , Spark SQL, Python, and SQL.

Experience designing, developing, and supporting production ETL/ELT pipelines.

Experience with workflow orchestration and scheduled data-processing workloads.

Strong understanding of data modeling, data warehousing, lakehouse architecture, and data-integration concepts.

Experience working with large and complex datasets.

Experience with Spark and SQL performance tuning.

Experience with AWS or another major cloud platform.

Experience with Git, CI/CD, automated testing, and production deployment.

Experience implementing data-quality, metadata, lineage, and governance controls.

Understanding of role-based access control, data privacy, security, and compliance requirements.

Strong analytical, troubleshooting, communication, and collaboration skills.

Experience working in Agile delivery environments.

Preferred Qualifications

Experience in biotechnology, pharmaceutical, life sciences, manufacturing, or another regulated industry.

Experience with Delta Lake, Unity Catalog, Databricks Workflows, or comparable technologies.

Experience with AWS data, storage, integration, security, and monitoring services.

Experience developing APIs or data services for downstream consumers.

Experience with relational, NoSQL, analytical, or vector databases.

Experience with data visualization tools such as Tableau or Power BI.

Experience supporting machine-learning or AI data pipelines.

Experience designing or developing Generative AI solutions using LLMs, Retrieval-Augmented Generation, embeddings, vector search, prompt engineering, or governed enterprise data.

Familiarity with AI-assisted development tools such as GitHub Copilot, OpenAI Codex, or equivalent platforms.

Preferred Certifications

Databricks Certified Data Engineer Associate.

AWS Certified Data Engineer or another relevant cloud certification.

Relevant data engineering, analytics, or Agile certification.

Soft Skills

Excellent critical-thinking and problem-solving skills.

Strong verbal and written communication skills .

Ability to work effectively with global and cross-functional teams.

High degree of initiative, ownership, and attention to detail.

Ability to manage multiple priorities and meet delivery commitments.

Strong teamwork and collaboration skills .

Ability to clearly present technical information to technical and non-technical audiences.

.

Job details

Seniority
Not listed
Function
Data & Digital
Therapeutic area
Not listed
Location
Hyderabad, India
Employment type
Full time

How this role compares

Computed from every other active Data & Digital role in our database, not just this employer's listings.

We currently track 292 comparable Data & Digital roles across 41 biopharma companies.

292Comparable roles tracked
273Currently active
41Companies hiring similar roles
20Countries represented

Salary context

60 of 292 peers report a salary range (USD, annualized)

Peers share this role's job function. This posting doesn't list a seniority level, so peers aren't narrowed by seniority either -- the range below may span more levels than usual.

This roleSubject Not listed on this posting
Lowest disclosed · Data Engineer, PDS&T CMC · AbbVie $65,500/yr – $125,500/yr
Peer group range $95,500 – $339,950 (median $198,000)

Where these roles are based

Top locations among the 292 comparable roles

India111
United States85
France22
Spain16
United Kingdom8
Canada7

+ 14 more countries

Seniority mix

158 of 292 peers have a known seniority level

Senior57
Manager31
Associate Director20
Principal16
Director16
Associate8
Executive/VP7
Senior Director3

Therapeutic area mix

6 of 292 peers have a known therapeutic area; the rest are genuinely unlabeled, not hidden

Immunology3
Oncology2
Ophthalmology1

Similar opportunities

The closest matches from our peer group, ranked by how similar they are, not how well you'd qualify for them -- treat this as market context, not a guaranteed shortlist; a weak match is labeled as one below.

40%similar
Novartis Hyderabad (Office), India Senior
Same function Same country
40%similar
Novartis Hyderabad (Office), India
Same function Same country
40%similar
Novartis Hyderabad (Office), India Senior
Same function Same country
40%similar
Novartis Hyderabad (Office), India
Same function Same country
40%similar
Regeneron India Private Limited Hyderabad, India Senior Director
Same function Same country
40%similar
Regeneron India Private Limited Hyderabad, India Manager
Same function Same country

How we calculate "similar"

No black box, no LLM guesswork: a deterministic score built from four normalized attributes. Here's this role's own peer group at different match levels, so you can see the mechanism, not just the result.

Every comparison starts from the same 100-point budget: 25 for working in the same function, 40 for the same therapeutic area, 20 for the same or adjacent seniority, 15 for the same country. A dimension we can't confirm on both sides contributes nothing, never a guess, never a free pass.

40%
Senior Expert (Rapid Prototyping)
Novartis · Hyderabad (Office), India · Senior
Function Therapeutic area Seniority Country
40%
Expert - Data Steward
Novartis · Hyderabad (Office), India · Seniority not listed
Function Therapeutic area Seniority Country
40%
Senior Director AI, Data & Solutions
Regeneron India Private Limited · Hyderabad, India · Senior Director
Function Therapeutic area Seniority Country
40%
Specialist, External Data Acquisition and Delivery
Regeneron India Private Limited · Hyderabad, India · Seniority not listed
Function Therapeutic area Seniority Country
Unmatched or unknown dimensions score exactly the same: 0 points, never a partial guess. A role we know almost nothing about beyond its function bottoms out at 25%; it never inflates to 100% just because there's little to compare against. Seniority uses a defined ladder (Associate → Manager → Associate Director → Senior → Principal → Director → Senior Director → Executive/VP) so "Director" and "Senior Director" count as adjacent, but "Director" and "Executive/VP" do not.