Data Engineer

Location Visakhapatnam, India Type Full-time Experience 4+ years Team Engineering Openings 1 position

About the role

Bad pipelines don’t announce themselves. They page you at 3am, quietly drop rows, or blow the bill at the end of the month. We’re hiring a Data Engineer to build the pipelines that move, clean and shape data so it can actually be trusted downstream. You like your systems boring in the best way: reliable, observable, and easy to reason about at 9am and 3am alike.

Because Externo is a service-based consulting company, you won’t babysit one warehouse forever. You’ll build data infrastructure across many clients and industries, from an e-commerce analytics stack one quarter to a real-time event pipeline feeding an AI product the next. You’ll own pipelines end to end, from the source system to the tables and streams that product, analytics and ML teams depend on.

What you’ll do

  • Design and build robust batch and streaming pipelines that hold up under real load.
  • Model data so analytics, product features and ML teams can actually use it, not just query it.
  • Add observability, alerting and tests so failures are obvious and loud, not silent.
  • Partner with AI and product engineers to figure out what the data actually needs to do.
  • Keep data quality high, latency low and cloud cost in check as things scale.
  • Own pipelines end to end across different clients, from source systems to the tables that ship.

Skills & tools

The toolkit you’ll reach for most. You don’t need every one on day one, but you should be fluent across the core.

Core skills
  • Data pipelines
  • SQL
  • Python
  • Data modeling
  • Batch & streaming
  • Orchestration
  • Data quality
  • Observability
Tools
  • SQL
  • Python
  • dbt
  • Airflow
  • Snowflake / BigQuery
  • Kafka
  • Spark
  • Docker
  • Git

What we’re looking for

  • 4+ years building and running data pipelines in production.
  • Strong SQL and Python: you write both without reaching for docs every five minutes.
  • Hands-on experience with modern data warehouses like Snowflake or BigQuery.
  • Comfort with orchestration and transformation tools such as Airflow and dbt.
  • A reliability-first mindset: you’d rather ship boring than clever.
  • You stay calm switching between clients, stacks and problems without dropping the ball.

Nice to have

  • Streaming experience with Kafka, Flink or Spark Structured Streaming.
  • Deeper cloud and infrastructure-as-code knowledge (Docker, Terraform, one of AWS / GCP / Azure).
  • Exposure to data feeding LLM or ML features: embeddings, feature stores, retrieval.

Show us your work

A portfolio isn’t required, but real proof helps. If you have a GitHub or pipelines you’ve built, drop the link in the Resume or portfolio link field on the right. We actually read them.

Is this you?

We hire for a specific kind of mind: people who think differently, work out of the box, and would rather try the less obvious idea than the safe one. We stay small and senior on purpose, so everyone here carries real weight.

Please only apply if you genuinely see yourself in this: a curious, self-driven engineer whose work and values line up with ours. If that’s you, we’d love to meet you. If it isn’t, applying will only waste your time and ours, and we’d rather be honest about that up front.

About Externo

Externo is a service-based consulting company. Businesses bring us in as their outside team to design, build and ship digital and AI products that hold up in the real world, across strategy, UX, LLM features, agent workflows, data pipelines and web experiences. We’re senior, remote-first, and work embedded alongside our clients. Get to know us on our About page, see the services we offer, or read the Externo blog.