Ubique Systems

TRCD-26-04374

Big Data Engineer (PySpark, AWS)

Contract Big Data Engineer (6–9 years) to design and implement scalable cloud-based data solutions and data pipelines on AWS. The role requires strong PySpark and SQL skills (including data modeling from scratch and SQL optimization), experience with AWS services like EMR/Glue and streaming (Kafka/Kinesis), and ability to support deployment automation using CI/CD and Infrastructure as Code (Terraform). You will collaborate with data science and business teams, integrate data from multiple sources (including APIs), and ensure performance, security, availability, and governance across the data lake ecosystem.

Position summary

Location
India
Workplace
Remote
Employment
Contract
Experience
Minimum 6 years and Maximum 8 years
Apply now

Role overview

Why This Role Matters.

Contract Big Data Engineer (6–9 years) to design and implement scalable cloud-based data solutions and data pipelines on AWS. The role requires strong PySpark and SQL skills (including data modeling from scratch and SQL optimization), experience with AWS services like EMR/Glue and streaming (Kafka/Kinesis), and ability to support deployment automation using CI/CD and Infrastructure as Code (Terraform). You will collaborate with data science and business teams, integrate data from multiple sources (including APIs), and ensure performance, security, availability, and governance across the data lake ecosystem.

Your Impact

Deliver Enterprise Value

Help organisations solve complex business problems through modern technology, consulting expertise and measurable outcomes.

Collaboration

Work Across Teams

Collaborate with consultants, architects, engineers and client stakeholders throughout the project lifecycle.

Growth

Learn Continuously

Gain exposure to enterprise technologies, certifications, mentoring and real-world project experience.

Career Path

Grow With Ubique

Build a long-term consulting career with opportunities to take on greater responsibility and leadership over time.

Responsibilities

What You'll Be Doing.

Every role at Ubique contributes directly to solving meaningful business challenges for our clients.

01

Participate in architecture design and implementation of high-performance, scalable, and optimized data solutions.

02

Good Sql understanding with ability to create the data model from scratch.

03

Ensure performance, security, and availability of databases.

04

Prepare documentations and specifications

05

Handle common database procedures such as upgrade, backup, recovery, migration, etc.

06

Profile server resource usage, and optimize and tweak as necessary

07

Design, build and automate the deployment of data pipelines and applications to support data scientists and researchers with their reporting and data requirements.

08

Integrate data from a wide variety of sources, including on premise databases and external data sources with rest APIs and harvesting tools.

09

Collaborate with internal business units and data science teams on business requirements, data access, processing/transformation and reporting needs and leverage existing and new tools to provide solutions.

10

Effectively support and partner with businesses on implementation, technical issues, and training on the datalake ecosystem.

11

Work with team on managing AWS resources (EMR, ECS clusters, etc.) and continuously improve deployment process of our applications

12

Work with administrative resources and support provisioning, monitoring, configuration, and maintenance of AWS tools.

Technology stack

Tools & Technologies.

The platforms and technologies you'll use to build modern, enterprise-grade solutions.

01

PySpark

02

Apache Spark

03

AWS

04

AWS EMR

05

AWS Glue

06

AWS ECS

07

Apache Kafka

08

AWS Kinesis

09

SQL

10

Relational databases

11

Data modeling

12

ETL

13

Data pipelines

14

Data lake

15

Terraform

16

Jenkins

17

CI/CD

18

Git

19

Subversion

20

Python

Requirements

Skills & Experience.

We value curiosity, collaboration and continuous learning. If you don't meet every requirement but believe you can make an impact, we'd still love to hear from you.

Essential

Required Qualifications

Big data engineering concepts

PySpark

AWS (data engineering services)

Strong SQL (data modeling, query optimization)

ETL/data pipelines design and development

Cloud-based data models for analytics/reporting/data science

Experience with structured and unstructured data

Spark on AWS (EMR) and/or AWS Glue

Streaming ingestion (Apache Kafka and/or AWS Kinesis)

Database performance, security, availability; backup/recovery/migration

CI/CD (automated build/deploy)

Infrastructure as Code (Terraform)

Version control (Git and/or Subversion)

Programming in Python and/or Java and/or Scala

Troubleshooting database issues; resource planning/capacity planning

Preferred

Nice to Have

AWS ECS (clusters)

Jenkins (CI/CD tool)

Configuration Management tools (unspecified)

Data governance and access control implementation

REST API integrations/harvesting tools

Data structures and algorithms

What you'll gain

More Than Just A Job.

We're committed to helping every team member grow professionally, personally and technically while working on meaningful projects.

Global Exposure

Collaborate with international clients and multicultural teams on enterprise programmes.

Continuous Learning

Expand your expertise through mentoring, certifications and hands-on project experience.

Career Growth

Take ownership, develop leadership skills and grow your consulting career over time.

Flexible Working

Hybrid and remote collaboration designed around trust and delivering exceptional outcomes.

People First

Join a supportive culture where collaboration, respect and long-term relationships come first.

Enterprise Projects

Work on meaningful technology initiatives for leading organisations across industries.

Apply

Apply for Big Data Engineer (PySpark, AWS)

One page, about two minutes. We only ask for what we actually need to have a first conversation.

Your CV

Optional. A line or two is plenty.