Data Engineer with Scala and Spark/ Pyspark

About Us

“Capco, a Wipro company, is a global technology and management consulting firm. Awarded with Consultancy of the year in the British Bank Award and has been ranked Top 100 Best Companies for Women in India 2022 by Avtar & Seramount. With our presence across 32 cities across globe, we support 100+ clients across banking, financial and Energy sectors. We are recognized for our deep transformation execution and delivery.

WHY JOIN CAPCO?

You will work on engaging projects with the largest international and local banks, insurance companies, payment service providers and other key players in the industry. The projects that will transform the financial services industry.

MAKE AN IMPACT

Innovative thinking, delivery excellence and thought leadership to help our clients transform their business. Together with our clients and industry partners, we deliver disruptive work that is changing energy and financial services.

#BEYOURSELFATWORK

Capco has a tolerant, open culture that values diversity, inclusivity, and creativity.

CAREER ADVANCEMENT

With no forced hierarchy at Capco, everyone has the opportunity to grow as we grow, taking their career into their own hands.

DIVERSITY & INCLUSION

We believe that diversity of people and perspective gives us a competitive advantage.

MAKE AN IMPACT

Job Title: Data Engineer

Position: Data Engineer
Experience: 4+ Years
Work Mode: Hybrid (Capco Office)

Key Responsibilities

  • Design, develop, and optimize large-scale data processing applications using Scala, Apache Spark, and Java.
  • Build and maintain high-performance data pipelines capable of processing 80–90 million records daily with a focus on scalability, reliability, and efficiency.
  • Develop and integrate Native APIs and data services to support business-critical applications and analytics platforms.
  • Collaborate with cross-functional teams to gather requirements and deliver robust data engineering solutions.
  • Optimize Spark jobs, data workflows, and distributed computing processes to ensure high throughput and low latency.
  • Implement best practices for data quality, monitoring, governance, and operational excellence.
  • Troubleshoot and resolve performance bottlenecks across data processing and ingestion pipelines.
  • Participate in code reviews and contribute to engineering standards, architecture decisions, and continuous improvement initiatives.

Required Skills & Experience

  • Strong hands-on experience in Scala, Apache Spark, and Java.
  • Proven experience building and supporting large-scale distributed data processing systems handling tens of millions of records daily.
  • Expertise in developing and consuming Native APIs and microservices.
  • Strong understanding of Spark architecture, performance tuning, partitioning, caching, and optimization techniques.
  • Experience with data modeling, ETL/ELT processes, and large-scale batch and streaming data pipelines.
  • Solid understanding of distributed systems, concurrency, and high-volume data processing.
  • Experience with SQL and relational/non-relational databases.
  • Strong debugging, analytical, and problem-solving skills.

Preferred Qualifications

  • Experience working in enterprise-scale data environments within Financial Services, Payments, or FinTech domains.
  • Exposure to cloud platforms (AWS, Azure, or GCP) and containerized deployments.
  • Familiarity with Kafka, Airflow, Hadoop ecosystem, or similar big data technologies.
  • Experience with CI/CD pipelines, DevOps practices, and Agile methodologies.

What We're Looking For

  • Engineers passionate about solving complex data challenges at scale.
  • Professionals who can design highly performant systems capable of processing and transforming massive datasets efficiently.
  • Team players who thrive in fast-paced, collaborative environments and take ownership of end-to-end delivery.

What this application asks

greenhouse

First Name, Last Name, Email, Phone, Resume/CV, Cover Letter

  • Preferred First Name optional
  • Capco Job Candidate Privacy Notice Acknowledgement choose one
  • How many years of experience do you have in overall? choose one
  • Do you have experience in Data engineering? choose one
  • How many years of experience in Data Engineering?
  • Do you have experience in Scala? choose one
  • Do you have relevant experience in Spark or Pyspark? choose one
  • What is your current location?
  • What is your preferred location? choose one
  • Are you currently working? choose one
  • What is current/ last payroll company name?
  • What is your official notice period?
  • How soon you can join us? choose one
  • What is your Current CTC (Fixed and Var)?
  • Please mention holding offer details (DOJ, Location, CTC), if any.
  • What is your Expected CTC in terms of INR?
  • Are you ready to work in Hybrid Mode/ 3 days WFO a week? choose one
  • Do you have any UAN discrepancies, overlapping employment periods, or have you ever absconded from any organization? choose one
  • LinkedIn Profile Link
  • If you are serving notice, please mention official last working day.
  • Gender choose one

See also

要針對這個職缺調整履歷嗎?

目前無法檢查您與這個職缺的符合程度;請先將履歷加入個人檔案,下次即可查看。

A new version of freehire is available