DATA ENGINEER

Carnegie Affiliates

Apply to this job
New York, NY, us on site Until 8/22/2026 H-1B sponsor history First posted March 21, 2025 Last posted March 21, 2025
Job description

Major Corporation

Responsibilities:

  • Design and develop high-throughput, low-latency data processing pipelines to quickly ingest and make data available on the platform across various distributed data stores
  • Analyze large datasets to identify opportunities to tune and improve the system
  • Experiment with various Hadoop frameworks like Hive, Pig and Scalding to identify the optimal approach for extracting valuable insights from massive datasets

Tool We Used:

  • Scala
  • Hadoop (Hive, Pig, Scalding, Spark)
  • Kafka
  • MySQL, Redis, Vertica, Aerospike

Requirements:

  • Bachelors or Masters in Computer Science or related field
  • 3+ years of experience ingesting, processing, storing and querying large datasets
  • Professional Hadoop ecosystem experience, including storage optimization and job performance tuning
  • Expertise in Java, Python or similar language(s). Functional programming experience is a plus
  • Passion for code correctness and intuition about which values in data are to be expected in a business context

All your information will be kept confidential according to EEO guidelines.

About this role

Summary

Design and develop data pipelines, analyze datasets, optimize system performance using Hadoop and related tools.

Job title

DATA ENGINEER

Experience level

3+ years

Industry

software

Location requirements

New York, NY, US, on-site preferred

Salary

Not specified

Visa sponsorship

H-1B sponsor history

Management role

No

Skills & keywords

Required skills

hadoopscalapythonjavasql

Preferred skills

functional programmingperformance tuning

Specializations

big datahadoopsparkkafkadistributed systems
Locations

Structured locations inferred from the posting.

New York, NY, USA

On-site City