Backend Engineer

5 days ago

New York City, NY, United States Morgan Pinnacle Group Full-time

Job Description

Job Description

Backend Engineer — High-Volume Data Processing

Location: Dumbo, Brooklyn, NY
Work Arrangement: Onsite 5 days/week — non-negotiable
Employment Type: Full-time
Base Salary: 220K–300K
Equity: Competitive equity
Visa: Open to visa transfers (including OPT and H-1B transfers); additional sponsorship may be considered for the right candidate
Relocation: Up to $10K in relocation support

About the Opportunity

We're partnering with a fast-growing AI/data company that provides real, proprietary enterprise data to leading frontier AI labs.

The company acquires data generated through how businesses actually operate, transforms it into de-identified datasets that remain useful, and licenses those datasets for next-generation AI model development.

The business has scaled from $0 to a multi-eight-figure run rate in a matter of months and operates with a lean engineering team of approximately five. This is a high-impact, ground-floor opportunity to take significant ownership of critical systems at a rapidly scaling company.

They're looking for a Backend Engineer to own and evolve the critical data-processing pipeline that transforms sensitive enterprise data into AI-training-ready datasets.

This role is specifically about code applied to the data itself . It is not a traditional data engineering, ETL, or data-platform position focused primarily on moving data between systems.

What You'll Do
  • Build and scale multi-stage, high-volume data-processing pipelines that de-identify sensitive enterprise data from sources such as collaboration tools, email, cloud storage, codebases, and project-management systems

  • Own backend systems end-to-end — design, delivery, debugging, correctness verification, and safe recovery

  • Ensure silent failures, duplicated output, stale data, and other correctness issues are caught before customer delivery

  • Make retries, checkpoints, partial failures, replay, backfills, and rollbacks safe and understandable across asynchronous orchestration, batch workers, queues, and object storage

  • Build internal tooling and dashboards to understand pipeline performance, investigate data quality, and support release decisions

  • Collaborate closely with ML, full-stack, platform, and security teammates to process new data modalities and expand pipeline capabilities

What They're Looking For Required
  • 3–10 years of software engineering experience with personal ownership of production backend systems

  • Personally shipped and operated production backend systems , ideally multi-stage, high-volume data-processing pipelines

  • Relevant undergraduate STEM degree from a top-30 U.S. institution

  • Strong distributed/asynchronous backend fundamentals, including experience with concepts such as:
    • Queues and workers

    • Concurrency

    • Idempotency

    • Retries

    • Partial failures

    • Recovery

  • Meaningful full-time experience at an early-stage or high-growth startup , operating with real ambiguity and ownership

  • Strong proficiency in a general-purpose backend language

  • Willing and able to work 5 days/week onsite in Dumbo, Brooklyn

The undergraduate education and onsite requirements are firm. Candidates who do not meet them will not be considered.

Technical Background

The current stack includes:

Python, AWS, Airflow, AWS Batch, SQS, S3, DynamoDB, PostgreSQL

Python experience is not required . Strong backend engineers coming from Go, Rust, or Java are also encouraged to apply. &lt