DATA ENGINEER
Carnegie Affiliates
Apply to this job New York, NY, us on site Until 8/22/2026 H-1B sponsor history First posted March 21, 2025 Last posted March 21, 2025
Job description
Major Corporation
Responsibilities:
- Design and develop high-throughput, low-latency data processing pipelines to quickly ingest and make data available on the platform across various distributed data stores
- Analyze large datasets to identify opportunities to tune and improve the system
- Experiment with various Hadoop frameworks like Hive, Pig and Scalding to identify the optimal approach for extracting valuable insights from massive datasets
Tool We Used:
- Scala
- Hadoop (Hive, Pig, Scalding, Spark)
- Kafka
- MySQL, Redis, Vertica, Aerospike
Requirements:
- Bachelors or Masters in Computer Science or related field
- 3+ years of experience ingesting, processing, storing and querying large datasets
- Professional Hadoop ecosystem experience, including storage optimization and job performance tuning
- Expertise in Java, Python or similar language(s). Functional programming experience is a plus
- Passion for code correctness and intuition about which values in data are to be expected in a business context
All your information will be kept confidential according to EEO guidelines.
About this role
Summary
Design and develop data pipelines, analyze datasets, optimize system performance using Hadoop and related tools.
Job title
DATA ENGINEER
Experience level
3+ years
Industry
software
Location requirements
New York, NY, US, on-site preferred
Salary
Not specified
Visa sponsorship
H-1B sponsor history
Management role
No
Skills & keywords
Required skills
hadoopscalapythonjavasql
Preferred skills
functional programmingperformance tuning
Specializations
big datahadoopsparkkafkadistributed systems
Locations
Structured locations inferred from the posting.
New York, NY, USA
On-site City