TRCD-26-04374
Big Data Engineer (PySpark, AWS)
Contract Big Data Engineer (6–9 years) to design and implement scalable cloud-based data solutions and data pipelines on AWS. The role requires strong PySpark and SQL skills (including data modeling from scratch and SQL optimization), experience with AWS services like EMR/Glue and streaming (Kafka/Kinesis), and ability to support deployment automation using CI/CD and Infrastructure as Code (Terraform). You will collaborate with data science and business teams, integrate data from multiple sources (including APIs), and ensure performance, security, availability, and governance across the data lake ecosystem.
Position summary
- Location
- India
- Workplace
- Remote
- Employment
- Contract
- Experience
- Minimum 6 years and Maximum 8 years
Role overview
Why This Role Matters.
Contract Big Data Engineer (6–9 years) to design and implement scalable cloud-based data solutions and data pipelines on AWS. The role requires strong PySpark and SQL skills (including data modeling from scratch and SQL optimization), experience with AWS services like EMR/Glue and streaming (Kafka/Kinesis), and ability to support deployment automation using CI/CD and Infrastructure as Code (Terraform). You will collaborate with data science and business teams, integrate data from multiple sources (including APIs), and ensure performance, security, availability, and governance across the data lake ecosystem.
Your Impact
Deliver Enterprise Value
Help organisations solve complex business problems through modern technology, consulting expertise and measurable outcomes.
Collaboration
Work Across Teams
Collaborate with consultants, architects, engineers and client stakeholders throughout the project lifecycle.
Growth
Learn Continuously
Gain exposure to enterprise technologies, certifications, mentoring and real-world project experience.
Career Path
Grow With Ubique
Build a long-term consulting career with opportunities to take on greater responsibility and leadership over time.
Responsibilities
What You'll Be Doing.
Every role at Ubique contributes directly to solving meaningful business challenges for our clients.
Good Sql understanding with ability to create the data model from scratch.
Ensure performance, security, and availability of databases.
Prepare documentations and specifications
Handle common database procedures such as upgrade, backup, recovery, migration, etc.
Profile server resource usage, and optimize and tweak as necessary
Design, build and automate the deployment of data pipelines and applications to support data scientists and researchers with their reporting and data requirements.
Integrate data from a wide variety of sources, including on premise databases and external data sources with rest APIs and harvesting tools.
Collaborate with internal business units and data science teams on business requirements, data access, processing/transformation and reporting needs and leverage existing and new tools to provide solutions.
Effectively support and partner with businesses on implementation, technical issues, and training on the datalake ecosystem.
Work with team on managing AWS resources (EMR, ECS clusters, etc.) and continuously improve deployment process of our applications
Work with administrative resources and support provisioning, monitoring, configuration, and maintenance of AWS tools.
Technology stack
Tools & Technologies.
The platforms and technologies you'll use to build modern, enterprise-grade solutions.
PySpark
Apache Spark
AWS
AWS EMR
AWS Glue
AWS ECS
Apache Kafka
AWS Kinesis
SQL
Relational databases
Data modeling
ETL
Data pipelines
Data lake
Terraform
Jenkins
CI/CD
Git
Subversion
Python
Requirements
Skills & Experience.
We value curiosity, collaboration and continuous learning. If you don't meet every requirement but believe you can make an impact, we'd still love to hear from you.
Essential
Required Qualifications
Big data engineering concepts
PySpark
AWS (data engineering services)
Strong SQL (data modeling, query optimization)
ETL/data pipelines design and development
Cloud-based data models for analytics/reporting/data science
Experience with structured and unstructured data
Spark on AWS (EMR) and/or AWS Glue
Streaming ingestion (Apache Kafka and/or AWS Kinesis)
Database performance, security, availability; backup/recovery/migration
CI/CD (automated build/deploy)
Infrastructure as Code (Terraform)
Version control (Git and/or Subversion)
Programming in Python and/or Java and/or Scala
Troubleshooting database issues; resource planning/capacity planning
Preferred
Nice to Have
AWS ECS (clusters)
Jenkins (CI/CD tool)
Configuration Management tools (unspecified)
Data governance and access control implementation
REST API integrations/harvesting tools
Data structures and algorithms
What you'll gain
More Than Just A Job.
We're committed to helping every team member grow professionally, personally and technically while working on meaningful projects.
Global Exposure
Collaborate with international clients and multicultural teams on enterprise programmes.
Continuous Learning
Expand your expertise through mentoring, certifications and hands-on project experience.
Career Growth
Take ownership, develop leadership skills and grow your consulting career over time.
Flexible Working
Hybrid and remote collaboration designed around trust and delivering exceptional outcomes.
People First
Join a supportive culture where collaboration, respect and long-term relationships come first.
Enterprise Projects
Work on meaningful technology initiatives for leading organisations across industries.
Apply
Apply for Big Data Engineer (PySpark, AWS)
One page, about two minutes. We only ask for what we actually need to have a first conversation.