About the Role
This is a senior engineering role on the core data platform at a seed-stage AI data infrastructure startup operating at the intersection of gaming and AI. You'll own the systems that ingest, enrich, package, and deliver large-scale multimodal data to AI labs — work that sits at the heart of the product and directly shapes the company's trajectory.
What You'll Do
Build high-throughput video processing pipelines covering transcoding, HLS packaging, depth-map rendering, frame-accurate synchronization, and GPU-accelerated media workflows.
Own ingestion and processing pipelines for video, 3D assets, depth maps, telemetry, annotations, metadata, and model-generated signals.
Design and build event-driven workloads on Kafka and Kubernetes spanning capture, processing, enrichment, QA, packaging, and delivery.
Build AI enrichment systems for scoring, metadata extraction, classification, quality review, redundancy detection, and dataset recommendations.
Develop internal tools for capture teams, reviewers, operators, and data managers to track, inspect, and assemble datasets.
Build customer-facing marketplace features enabling AI labs to search, preview, query, purchase, and receive datasets.
Design data models covering provenance, asset versioning, licensing state, QA state, delivery history, and customer entitlements.
Instrument monitoring and analytics to understand quality, throughput, bottlenecks, and dataset value.
What We're Looking For
5+ years of production software engineering experience spanning backend, frontend, and data-heavy systems.
Hands-on experience building multimodal data ingestion and orchestration pipelines at scale — a must-have.
Production video processing experience: FFmpeg, transcoding, HLS, and streaming formats — a must-have.
Strong Python backend skills with FastAPI or Django, and PostgreSQL in production — a must-have.
Production experience with Kafka and Kubernetes.
Microservices architecture, API design, and signal processing at scale.
Comfortable building custom in-house infrastructure rather than defaulting to managed services.
Familiarity with the broader stack: Next.js, AWS.
Background at an AI training data platform is a strong plus.
Combined ML/AI enrichment and full-stack engineering experience is a plus.
STEM degree (bachelor's minimum; master's preferred — non-CS disciplines such as mathematics, physics, or biology are a positive signal).
Compensation & Benefits
Salary: $130,000 – $200,000 USD annually. Visa sponsorship is not available.
Location
On-site in Los Angeles, CA. Candidates based in New York City or San Francisco may also be considered.
Learn more about this Employer on their Career Site
