engineering
Posted 2 days agoSenior Software Engineer - Data Acquisition Team
at zoominfo
United StatesHybrid
Responsibilities
- Design, build, and operate large-scale data acquisition pipelines that ingest, validate, transform, enrich, and store high-volume raw data.
- Architect resilient ETL/ELT workflows for batch, streaming, scheduled, and event-driven data processing.
- Develop production Java services and data processing applications for ingestion, orchestration, enrichment, deduplication, and delivery.
- Build and improve systems using technologies such as Apache Airflow, Apache Beam, Spark, Google Dataflow, DataProc, Kafka, and Pub/Sub.
- Define practical approaches for schema evolution, data contracts, data validation, backfills, replayability, and idempotent processing.
- Improve reliability, performance, scalability, and cost efficiency across data acquisition pipelines and services.
- Implement observability, monitoring, and alerting for pipeline health, throughput, latency, failure rates, and data quality metrics.
Requirements
- The right candidate brings strong backend engineering fundamentals, significant pipeline experience, and the ability to turn complex data acquisition
- Contribute to technical designs, implementation plans, and system modernization efforts across Data Acquisition. Must-Have
- experience with a strong focus on backend systems, data engineering, or distributed processing. Proven
- Deep proficiency with Java and object-oriented design.
- Hands-on expertise with data processing and orchestration technologies such as Apache Beam, Apache Airflow, Spark, Google Dataflow, or DataProc. Strong
- experience with streaming systems such as Apache Kafka, Google Pub/Sub, or similar technologies.
- Strong understanding of batch processing, streaming processing, data modeling, schema evolution, and data quality management. •
- Experience designing ETL/ELT workflows that process large volumes of structured and semi-structured data.
- Experience designing high-throughput, fault-tolerant backend services and distributed systems.
- Strong understanding of APIs, integration patterns, retries, backpressure, idempotency, and operational failure modes.
- Experience with large-scale storage and query technologies such as BigQuery, Snowflake, Trino, or similar systems.
- Experience with at least one cloud provider, preferably GCP. Hands-on
- experience with cloud services such as BigQuery, GCS, GKE, Dataflow, DataProc, and Pub/Sub. •
- Bachelor's degree in Computer Science, Software Engineering, or a related field.
- Experience with Kubernetes, especially GKE or EKS, for running distributed workloads. •
- Experience with Terraform or other infrastructure-as-code tools. •
- Experience with Snowflake, BigQuery, Starburst/Trino, or similar query engines.
- Knowledge of data integration patterns involving CRM systems, email/calendar APIs, third-party feeds, change data capture, or external data providers. •
- Experience in a B2B data company, data marketplace, or data-as-a-product environment. Expert-level
- experience with Apache Spark or another distributed processing framework. Why This Role Matters
- ZoomInfo (NASDAQ: GTM) is the Go-To-Market Intelligence Platform that empowers businesses to grow faster with AI-ready insights, trusted data, and advanced automation.
Experience
- 5+ years of professional software engineering
Benefits
- Actual compensation offered will be based on factors such as the candidate’s work location, qualifications, skills, experience and/or training.
- Your recruiter can share more information about the specific salary range for your desired work location during the hiring process.
- Below is the US base salary for this position. Additional compensation such as Bonus, Commission, Equity and other
- benefits may also apply. $140,000 — $220,000 USD About us:
Additional details
- We move fast, think boldly, and empower you to do the best work of your life.
- You’ll be surrounded by teammates who care deeply, challenge each other, and celebrate wins.
- your impact and a culture that backs your ambition, you won’t just contribute. You’ll make things happen–fast. The Opportunity
- In this role, you will design, build, and operate the backend systems and data pipelines that acquire, transform, validate, and store large raw data sets from a wide range of sources.
- You will work across distributed processing, workflow orchestration, streaming and batch data flows, and cloud infrastructure.
- This is a senior individual-contributor engineering role focused on technical depth, production execution, and high-quality data systems.
- requirements into scalable systems. What You'll Do
- Work with product, data science, platform, and data quality teams to translate business needs into production-ready systems.
- experience building and operating production data pipelines at scale.
- Ability to write clean, maintainable production code and evaluate tradeoffs in system design. •