yoinka

Big data/Python/Databricks Engineer Engineer

Citigroup

Remote1124 SHIVAJI GARDENS MOONLIFull TimeSenior
Sign in to applyVerified 1h ago
Location
1124 SHIVAJI GARDENS MOONLI
Employment
Full Time
Work model
Hybrid
Level
Senior
Posted
Sep 21, 2026

Skills

HadoopLinuxOracle DBPythonSQLTableau

About this role

Citi is looking for an Applications Development Intermediate Programmer Analyst to design, build, and maintain large-scale data engineering solutions that support critical reporting and analytics functions across a global financial institution. In this role, you will develop and optimize data pipelines, work across distributed computing platforms, and contribute to the full lifecycle of data application delivery. This is an opportunity to work within a high-impact technology team that operates at the core of Citi's data infrastructure.

Responsibilities

Build and maintain scalable data pipelines using Python and PySpark to process and transform large volumes of structured and unstructured data across distributed platforms. Develop and optimize data workflows on Hadoop-based ecosystems, including HDFS, Hive  ensuring reliable data availability for downstream reporting and analytics. Conduct feasibility studies, time and cost estimates, and technical planning activities to support data engineering delivery across business areas. Monitor and manage all phases of the data application development lifecycle, from analysis and design through to testing, implementation, and production support. Collaborate with analytics and reporting teams to build and maintain Tableau dashboards and translating raw data into actionable business insights. Administer and troubleshoot data processes within Linux environments, ensuring stability, performance, and operational continuity. Assess risk across data engineering decisions, ensuring solutions align with security, data governance, and compliance standards. Experience in managing and implementing successful projects Working knowledge of consulting/project management techniques/methods Ability to work under pressure and manage deadlines or unexpected changes in expectations or requirements Required Qualifications & Skills 5 to 8 years of experience in data engineering, big data development, or software application development with a focus on large-scale data platforms. Hands-on development experience using Python and PySpark for data ingestion, transformation, and pipeline orchestration at scale. Expert in ETL and Oracle DB, SQL Knowledge in Data bricks Practical knowledge of Hadoop ecosystem components including HDFS, Hive, and Hadoop cluster operations, with familiarity with Ozone storage. Working experience in Linux environments, including scripting, job scheduling, and process management. Experience building or supporting Tableau dashboards and reports to deliver data insights to business stakeholders. Familiarity with RStudio for statistical analysis or data exploration in a data engineering context. Deep expertise in Large Language Models (LLMs) including OpenAI, Gemini, Claude, Llama, and local/open‑source models Bachelor's degree or equivalent experience in a relevant technical discipline. Beneficial Skills & Qualifications Experience working with cloud-native data platforms or migrating workloads from on-premise Hadoop environments to modern data platforms. Knowledge of data governance practices, metadata management, or data quality frameworks within large-scale environments. Familiarity with consulting or project management techniques applied within a technology delivery context.

What We Offer

At Citi, you will work within a collaborative and performance-driven global technology team where your contributions directly support the data infrastructure underpinning one of the world's leading financial institutions. This role offers technical depth, meaningful delivery ownership, and the opportunity to grow your expertise across big data and analytics platforms. Hybrid working model with 2 days in the office and 3 days working remotely, supporting a sustainable work-life balance. Opportunity to work on large-scale, real-world data engineering challenges using modern big data tools and platforms at enterprise scale. Access to continuous learning and development resources to deepen technical

Listing verified 1h ago. Applications go through the company's official careers site.

← Back to Yoinka

Big data/Python/Databricks Engineer Engineer at Citigroup, 1124 SHIVAJI GARDENS MOONLI | Yoinka