Data Engineering at Citi
Roles, Comp & Culture
An L5 senior data engineer at Citi sits around $235K total comp from 47 verified salary datapoints. The primary Data Engineering tech consists of Hadoop, Spark and Docker, according to current job listings. Citi pays data engineers above other Finance companies. Reviews put them at 3.6 on Glassdoor, a little below the middle of the pack. Employee sentiment at Citi reads neutral and employee happiness has held flat over the past year. Layoff risk scores low for the next 30 days. 26 data engineering roles are open right now.
Citi data engineer compensation
Each level's figure is the median of individual Citi offers at that level, so it reflects a typical outcome rather than an average pulled up by a few large packages. Total comp counts base salary plus equity and bonus annualized over the vest, and the range shown is the middle half of offers, with the top and bottom quarters trimmed off.
Citi employee sentiment, tracked weekly
Employee happiness for data engineers over the past year, so you can see which direction it is moving, not just where it sits today.
Citi pays above other Finance companies, with $158K at mid covering 44 of the 47 reports in the salary pool. The Glassdoor rating sits at 3.6, a little below the middle of the pack, and sentiment on Blind runs mixed. The honest read is that Citi offers financial sector stability and compensation that holds up in the market, but reviews consistently surface bureaucracy and slow career progression as the real costs. Engineers who want to ship quickly or own product-facing data systems will find the pace frustrating. What you're trading for the comp and job security is organizational friction, layers of approval, and the reality that finance compliance moves slower than any individual contributor wants. Happiness sits neutral and roughly flat, which tracks with that description: not actively miserable, not energizing.
Citi has 26 open data engineering roles across 5 cities, with Tampa as the leading market, which reflects the bank's ongoing push to consolidate technology operations in lower-cost US cities rather than concentrating everything in New York. That hiring volume suggests active investment in data infrastructure, likely driven by regulatory remediation programs the bank has been under pressure to complete. Layoff risk over the next 30 days reads low, and Citi has been cautious about large-scale tech cuts compared to some peers. The next 12 months for someone joining now will probably look like continued modernization work, migration off older Hadoop-era tooling, and pressure to demonstrate data governance improvements to regulators. Growth here is measured; anyone expecting a dramatic ramp in scope or headcount should calibrate expectations.
Citi data engineering tech stack
The languages, storage, and processing tools Citi data engineers actually work with, grouped by what they do. Tailor your system-design answers to this stack.
Citi runs one of the largest financial data operations in the world, and the engineering problems that come with that scale are unglamorous in the best way: regulatory reporting pipelines that cannot miss SLAs, transaction data moving across dozens of business lines, and a data infrastructure older in places than most engineers currently applying. The stack centers on Hadoop, Spark and Docker with Python, SQL and PySpark, which is a reliable signal that batch workloads and HDFS-era architecture still dominate day-to-day work over real-time streaming. A data engineer here is more likely to be stabilizing and modernizing existing pipelines than greenfielding a lakehouse from scratch. The regulatory surface, covering capital reporting, trade surveillance, and cross-border compliance, means data quality and lineage aren't optional concerns; they're often the primary engineering constraint.
Citi data engineer job openings
A live read on what they are hiring: open roles, recent postings, where, and at what level.
Data integration delivery: Design, build, and operate robust batch and near-real-time integration pipelines for CRM data domains (e.g., customers, products, orders, invoices, service interactions).
Big Data Infrastructure: Develop and manage large-scale data processing systems using frameworks like Apache Spark, Hadoop, and Kafka.
Design and maintain blueprint of the information architecture, data integrations and controls aligned to the renewed business strategy
Design, develop, and maintain high-performance, resilient, and scalable ETL processes using the Ab Initio suite of products (GDE, Co>Operating System, EME) to transform upstream data into the required format for Oracle Financials SaaS.
Develop, maintain, and optimize highly efficient and resilient data ingestion, processing, and transformation pipelines using advanced Python and PySpark techniques for large-scale datasets.
Architect for Real-Time: Serve as the go-to expert for the data platform, ensuring every design adheres to our overall architecture blueprint.
Build and maintain big data pipelines using technologies such as Apache Hadoop, Apache Kafka, Databricks and other cloud based big data tools.
Monitor and control all phases of development process and analysis, design, construction, testing, and implementation as well as provide user and operational support on applications to business users
Architect & Design: Design, architect, and oversee the development of robust, scalable, and reliable data infrastructure, including data lakes, data warehouses, and real-time streaming platforms on the cloud.
Architect, design, and deliver scalable, Python‑based applications supporting credit risk analytics, workflows, and reporting.
Practice for the Citi loop
Round by round, the problems our model predicts for this company's interview. Rehearse the shapes their panels keep returning to.
The salary pool skews heavily toward mid-level engineers, with 44 of 47 reports at that band. Senior representation is thin, which either means the path is long or the ceiling compresses before staff. Engineers who do well here tend to be comfortable in regulated environments, patient with approval chains, and genuinely interested in financial data domains like risk, treasury, or markets infrastructure. If you want to work on streaming architecture or build data products with fast feedback loops, Citi is probably the wrong fit. If you're a solid pipeline engineer who values employment stability, wants finance on your resume, and can work within constraints, the combination of above-market pay and low near-term layoff risk makes it worth a serious look. Prep the pipeline architecture loop carefully; that's where the loop is won or lost.
Preparing for the Citi loop
The round-by-round process, example questions, and prep plan are on the interview guide.
Compare Citi with other data engineering employers
How the role, pay, and loop stack up against peer companies.
Prepare at Citi interview difficulty
- 01
Reading a solution is not the same as writing one
Every engineer who has frozen on a query they had read a dozen times knows the gap. The only preparation that closes it is producing the answer yourself, under time, before the interview does it for you
- 02
76% of hiring managers reject on the coding task, not the resume
From HackerRank's 2024 Developer Skills Report. Candidates who look strong on paper still fail the live screen if they haven't done timed, executable practice
- 03
5 problem shapes cover 80% of data engineer loops
Dedup, sessionization, top-N-per-group, slowly-changing dimensions, partition tricks. Writing the shapes by hand turns the unfamiliar into pattern recognition