01Overview
Lead Data Engineer - PV Domain
Noida, India
About our client:
Our client brings together Pharmacovigilance expertise, third party system knowledge and Deep technology to develop well-defined solutions, which address challenges across Medical Affairs, Regulatory and Safety functions.
Our client solutions free up responsible personnel within Pharma companies to execute their stated responsibilities while staying true to the laws of the land, and ultimately achieving a balance between compliance and managing business risks.
Their solutions are agile, flexible, and scalable, developed using advanced technologies that enable them to serve large and small organizations, both in developed and emerging markets.
Our client is committed to bringing focus to
things that really matter for advancing patient outcomes. Their solutions are:
agile, flexible, and scalable, developed using advanced technologies that
enable us to serve large and small organizations, both in developed and
emerging markets.
Requirements
Experience:
6-15 Years
Location:
Noida, India
Employment: Full Time
Job Summary To identify & define
technical requirements to Onboard new data partners and ensure seamless
integration with Data Lake
Should have experience
in AWS Cloud Computing with knowledge of services like AWS Lambda, EC2, S3,
EMR, Redshift, Glue, Step Functions
Must have experience in
designing and building production data pipelines from ingestion to consumption
within a Data Platform like Databricks
Should have demonstrated
strength in data modeling, ETL development
Should have experience
in designing and implementing an enterprise Data Lake and Data Mart/Warehouse.
Should have implemented
complex projects dealing with the considerable data size (TB/ PB) and with high
complexity in the production environment
Hands-on experience with
CI/CD
Should be able to
communicate efficiently with customers key Business and IT folks to
present/defend architecture/design
Should be excellent
analytical, problem-solving, communication, and ability to communicate
efficiently with individuals, and business and can work as part of a team as
well as independently
Must have Hands-on
Experience on Spark with Python
Good knowledge and
development experience in Oracle/PostgreSQL/Snowflake/Redshift
Qualifications:
BE/B.Tech/MCA
Good expertise in Data
warehousing/Dimensional Modeling
Good to have:
Knowledge of Data
Processing/Transformation using DBT
Knowledge of Data
Orchestration tools like Apache Airflow
Preferred:
Data Engineering with
Pharmacovigilance background .