Backend Engineer
Save this job and keep your search organized
Create a free account to save jobs, create alerts and return to this listing from your dashboard.
By continuing, you agree to our Terms & Privacy Policy.
Job Description
Job Description
Backend Engineer — High-Volume Data Processing
Location: Dumbo, Brooklyn, NY
Work Arrangement: Onsite 5 days/week — non-negotiable
Employment Type: Full-time
Base Salary: 220K–300K
Equity: Competitive equity
Visa: Open to visa transfers (including OPT and H-1B transfers); additional sponsorship may be considered for the right candidate
Relocation: Up to $10K in relocation support
We're partnering with a fast-growing AI/data company that provides real, proprietary enterprise data to leading frontier AI labs.
The company acquires data generated through how businesses actually operate, transforms it into de-identified datasets that remain useful, and licenses those datasets for next-generation AI model development.
The business has scaled from $0 to a multi-eight-figure run rate in a matter of months and operates with a lean engineering team of approximately five. This is a high-impact, ground-floor opportunity to take significant ownership of critical systems at a rapidly scaling company.
They're looking for a Backend Engineer to own and evolve the critical data-processing pipeline that transforms sensitive enterprise data into AI-training-ready datasets.
This role is specifically about code applied to the data itself . It is not a traditional data engineering, ETL, or data-platform position focused primarily on moving data between systems.
What You'll DoBuild and scale multi-stage, high-volume data-processing pipelines that de-identify sensitive enterprise data from sources such as collaboration tools, email, cloud storage, codebases, and project-management systems
Own backend systems end-to-end — design, delivery, debugging, correctness verification, and safe recovery
Ensure silent failures, duplicated output, stale data, and other correctness issues are caught before customer delivery
Make retries, checkpoints, partial failures, replay, backfills, and rollbacks safe and understandable across asynchronous orchestration, batch workers, queues, and object storage
Build internal tooling and dashboards to understand pipeline performance, investigate data quality, and support release decisions
Collaborate closely with ML, full-stack, platform, and security teammates to process new data modalities and expand pipeline capabilities
3–10 years of software engineering experience with personal ownership of production backend systems
Personally shipped and operated production backend systems , ideally multi-stage, high-volume data-processing pipelines
Relevant undergraduate STEM degree from a top-30 U.S. institution
- Strong distributed/asynchronous backend fundamentals, including experience with concepts such as:
Queues and workers
Concurrency
Idempotency
Retries
Partial failures
Recovery
Meaningful full-time experience at an early-stage or high-growth startup , operating with real ambiguity and ownership
Strong proficiency in a general-purpose backend language
Willing and able to work 5 days/week onsite in Dumbo, Brooklyn
The undergraduate education and onsite requirements are firm. Candidates who do not meet them will not be considered.
Technical BackgroundThe current stack includes:
Python, AWS, Airflow, AWS Batch, SQS, S3, DynamoDB, PostgreSQL
Python experience is not required . Strong backend engineers coming from Go, Rust, or Java are also encouraged to apply. <