Business financials got stuck in the 15th century so we're showing them today鈥檚 computers 馃枼
Member of Technical Staff, Data Infrastructure
Location
United States + 180 moreAll locations: United States, Canada, Brazil, Colombia, Argentina, Chile, Venezuela, Bolivarian Republic Of, Bolivia, Plurinational State Of, Ecuador, French Guiana, Guyana, Paraguay, Peru, Suriname, Uruguay, Mexico, Costa Rica, El Salvador, Guatemala, Honduras, Nicaragua, Panama, Dominican Republic, Puerto Rico, Bahamas, Guadeloupe, Haiti, Jamaica, Martinique, Montserrat, United Kingdom, Germany, France, Estonia, Portugal, Hungary, Poland, Ukraine, Romania, Bulgaria, Czech Republic, Slovakia, Belarus, Moldova, Republic Of, Sweden, Greece, Belgium, Italy, Ireland, Switzerland, Netherlands, Finland, Malta, Denmark, Lithuania, Croatia, Spain, Austria, Bosnia And Herzegovina, Iceland, Luxembourg, Macedonia, The Former Yugoslav Republic Of, Montenegro, Norway, Serbia, Slovenia, Albania, Cyprus, Latvia, Monaco, South Africa, Egypt, Algeria, Angola, Benin, Botswana, Burkina Faso, Burundi, Cameroon, Cape Verde, Central African Republic, Chad, Congo, C么te D'ivoire, Congo, The Democratic Republic Of The, Equatorial Guinea, Eritrea, Ethiopia, Gabon, Gambia, Ghana, Guinea, Guinea-bissau, Kenya, Lesotho, Liberia, Libyan Arab Jamahiriya, Madagascar, Malawi, Mali, Mauritania, Mauritius, Mayotte, Morocco, Mozambique, Namibia, Niger, Nigeria, R茅union, Rwanda, Senegal, Seychelles, Sierra Leone, Somalia, Sudan, Swaziland, Tanzania, United Republic Of, Togo, Tunisia, Uganda, Zambia, Zimbabwe, Georgia, Turkey, Israel, United Arab Emirates, Armenia, Azerbaijan, Bahrain, Iraq, Jordan, Kuwait, Lebanon, Oman, Qatar, Saudi Arabia, Palestinian Territory, Occupied, Yemen, India, Japan, Philippines, Pakistan, Thailand, Singapore, Viet Nam, Taiwan, Province Of China, Indonesia, Cambodia, Lao People's Democratic Republic, Malaysia, Myanmar, Korea, Republic Of, China, Afghanistan, Bangladesh, Bhutan, Kazakhstan, Kyrgyzstan, Maldives, Mongolia, Nepal, Sri Lanka, Tajikistan, Turkmenistan, Uzbekistan, Australia, Papua New Guinea, Kiribati, Palau, French Polynesia, Tuvalu, New Zealand
Posted
58 days ago
Salary
$240K - $290K / year
No structured requirement data.
Job Description
Role Description
We're looking for a Data Engineer to build and scale the data infrastructure that powers Runway's AI research and business intelligence. You'll own critical data pipelines spanning production databases, analytics warehouses, and large-scale ML training datasets. This role sits at the intersection of data engineering, ML infrastructure, and analytics鈥攜ou'll enable both world-class research and data-driven business decisions.
You'll work on challenging problems at scale:
- Managing billions of rows of multimodal training data
- Building CDC streams from production systems
- Optimizing vector databases for ML workflows
- Creating the foundational data layer that the entire company relies on
A peek at our technical stack:
- LanceDB for vector storage and dataset versioning with multimodal training data
- ClickHouse as our analytics warehouse receiving CDC streams from production Postgres via AWS Kinesis
- BigQuery for training run logs and evaluation results
- Ray for large-scale distributed data processing on managed Kubernetes clusters
- dbt for standardized transformations
- Prometheus and Grafana for monitoring
- Terraform for infrastructure management
This is an opportunity to bring best practices and technical leadership as we mature our data infrastructure to support rapidly growing ML training and research needs.
Qualifications
- 4+ years of industry experience in data engineering
- Strong knowledge of Python
- Experience with data quality, deduplication, and cleaning at scale
- Comfortable working with cloud storage (S3) and managing large datasets
- Experience building and maintaining ETL/CDC pipelines at scale
- Strong SQL skills and experience with multiple database systems (Postgres, columnar databases like ClickHouse/Redshift)
- Humility and open mindedness; at Runway we love to learn from one another
Requirements
- Experience with one or more frameworks for large-scale data processing (e.g. Spark, Ray, etc)
- Knowledge of cloud platforms (AWS, GCP, or Azure) and their data service offerings
- Knowledge of data privacy and data security best practices
- Experience with business intelligence and visualization tools (e.g., Looker, Tableau, PowerBI, Metabase, or similar)
- Experience in a high-growth startup environment or similar fast-paced setting
Benefits
- Salary range: $240,000 - $290,000
- Commitment to creating a space where employees can bring their full selves to work
- Equal opportunity to succeed regardless of race, gender identity or expression, sexual orientation, religion, origin, ability, age, veteran status
Job Requirements
- 4+ years of industry experience in data engineering
- Strong knowledge of Python
- Experience with data quality, deduplication, and cleaning at scale
- Comfortable working with cloud storage (S3) and managing large datasets
- Experience building and maintaining ETL/CDC pipelines at scale
- Strong SQL skills and experience with multiple database systems (Postgres, columnar databases like ClickHouse/Redshift)
- Humility and open mindedness; at Runway we love to learn from one another
- Experience with one or more frameworks for large-scale data processing (e.g. Spark, Ray, etc)
- Knowledge of cloud platforms (AWS, GCP, or Azure) and their data service offerings
- Knowledge of data privacy and data security best practices
- Experience with business intelligence and visualization tools (e.g., Looker, Tableau, PowerBI, Metabase, or similar)
- Experience in a high-growth startup environment or similar fast-paced setting
Benefits
- Salary range: $240,000 - $290,000
- Commitment to creating a space where employees can bring their full selves to work
- Equal opportunity to succeed regardless of race, gender identity or expression, sexual orientation, religion, origin, ability, age, veteran status
Related Guides
Related Categories
Related Job Pages
More Data Engineer Jobs
Lead Architect overseeing data migration to Guidewire InsuranceSuite platform
Data Engineer
OnebriefSoftware for rapid military planning: make planning fast enough for today's environment
Data Engineer managing data platforms for simulation outputs at Onebrief
Principal Data Engineer tackling complex data challenges at Sayari
We鈥檙e seeking a skilled Data Engineer to build the next-generation data management and artificial intelligence platform for maritime domain awareness. What you鈥檒l do: Implement real-time data pipelines with MQTT and Redpanda for stream processing. Implement offline data pipelines...