New opportunity

Software Development Engineer II – Data Engineer

Gather AI · Remote (India)

About this role

Gather AI is pioneering warehouse intelligence by utilizing autonomous drones and vision-powered platforms to digitize manual supply chain workflows. As an SDE II, Data Engineer, you will join the newly formed Full Stack group within Cloud Services to build the company's data foundation from the ground up. This role involves moving analytics off production databases, designing shared transformation layers, and establishing a consistent semantic model to support various product teams.

Responsibilities

  • Build and maintain extraction pipelines from production PostgreSQL to the analytical warehouse using incremental loads.
  • Develop and extend shared data models using dbt, including dimensions and reusable metric building blocks.
  • Implement metrics in a semantic layer to ensure consistency across dashboards.
  • Ensure data trustworthiness through testing, freshness checks, and alerts.
  • Maintain data lineage, linking structured records to drone-captured images and video.
  • Partner with integration teams to validate incoming WMS data and maintain tenant isolation.
  • Document models in the data catalog and participate in on-call rotations.

Requirements

  • 2–5 years of experience building and running production data pipelines.
  • Strong SQL skills (joins, window functions, CTEs) and understanding of dimensional modeling.
  • Hands-on experience with dbt or equivalent transformation tools.
  • Proficiency in Python for production-grade pipeline code.
  • Experience with orchestration tools like Airflow or equivalent.
  • Experience loading data from operational databases into warehouse/lakehouse environments (e.g., Snowflake, Databricks).
  • Production experience with cloud platforms (Azure preferred), Git, CI/CD, and containerization (Docker/Kubernetes).
  • Excellent written and spoken English for a distributed team environment.

Preferred Qualifications

  • Experience with Change Data Capture (Debezium, Fivetran) or streaming (Kafka).
  • Knowledge of PySpark, Snowpark, or Infrastructure as Code (Terraform).
  • Familiarity with semantic layers, data catalogs, or multi-tenant data platforms.
  • Domain experience in logistics, warehousing, or robotics.

Skills for this role

SQLPostgreSQLdbtPythonAirflowSnowflakeDatabricksAzureDockerKubernetesGitCI/CDKafkaTerraformData ModelingDimensional ModelingData PipelinesWarehouse Intelligence

Your skill match

Checking your profile…

YOUR NEXT STEP

Get interview-ready for this role

A focused preparation guide, built around this job’s responsibilities and requirements.

✦ AI-generated guide
Preparation suggestions, not the employer’s actual interview questions. Always check the original posting for current requirements.

Loading this role’s preparation guide…

KEEP EXPLORING

Similar AI jobs

Related skills and specializations in India. Matched to this role, not your profile.

Explore more jobs
Finding similar opportunities…