Data Engineer
75k - 130k USD
Remote
Full Time
#Engineering
#Cloud
#Data Processing
#Python
#Pyspark
#SQL
#SparkSQL
#Data
#AWS
#Databricks
#Data Pipelines
#Data Quality
Veeva Systems is a mission-driven organization and a pioneer in industry cloud solutions, dedicated to helping life sciences companies deliver therapies to patients with greater speed. As one of the fastest-growing SaaS companies in history, we recently surpassed two billion dollars in annual revenue and continue to see significant potential for expansion. Our culture is built on core values of integrity, customer success, employee success, and speed. Notably, we operate as a public benefit corporation, which means we are legally committed to balancing the interests of our customers, employees, investors, and society at large. We embrace a work-anywhere philosophy that empowers our team members to thrive in the environment that suits them best, whether that is from home or in one of our global offices.
Key outcomes
- Design, build, and maintain robust data processing pipelines and tools using modern cloud technologies.
- Utilize Python and SQL to engineer complex data workflows on Spark-based systems.
- Develop sophisticated algorithms to establish intricate data relationships and create analytical structures for reporting.
- Implement and manage rigorous data quality processes to ensure the integrity of our reference data.
- Partner closely with our product teams to evolve our data offerings in response to shifting market demands.
Requirements
- At least three years of professional experience developing data pipelines within cloud-managed Spark environments, such as AWS EMR or Databricks.
- Strong proficiency in Python and PySpark, with a minimum of three years of hands-on application.
- Proven ability to create custom tools and libraries that automate and optimize data processing workflows.
- Advanced skills in SQL and SparkSQL.
- Practical experience working within a Data Lakehouse architecture.
- Excellent communication skills and a successful track record of delivering results in an Agile environment.
- A demonstrated history of mentoring others and elevating team performance.
- Unrestricted authorization to work in the United States, as we are unable to provide sponsorship for this position.
Preferred qualifications
- Experience managing data workflows through DevOps pipelines.
- Familiarity with orchestration tools such as Airflow.
- Hands-on experience with AWS data processing services like EMR or MWAA.
- Background knowledge or previous work experience within the life sciences sector.
Compensation
The base salary range for this position is $75,000 to $130,000. Actual compensation is determined based on your unique qualifications, experience, and expected contributions, and may include additional elements such as variable or stock bonuses. We offer a comprehensive benefits package that includes:
- Medical, dental, and vision insurance.
- Remote work flexibility.
How to apply
If you are interested in joining our team and contributing to the transformation of the life sciences industry, we invite you to submit your application. We look forward to reviewing your background and discussing how your skills align with our mission.






