SonicJobs Logo
Left arrow iconBack to search

Senior Software Engineer, Robot Data Infrastructure

GRAM
Posted a month ago, valid for 12 days
Location

San Francisco, CA, US

Salary

$190,000 - $240,000 per year

Contract type

Full Time

By applying, a Sonicjobs account will be created for you. Sonicjobs's Privacy Policy and Terms & Conditions will apply.

SonicJobs' Terms & Conditions and Privacy Policy also apply.

Sonic Summary

info
  • GRAM is seeking a senior software engineer to build infrastructure for reproducible training and evaluation datasets for their innovative insectoid machines.
  • The role requires a Bachelor's degree in a relevant field and at least 5 years of experience in software engineering, particularly with data ingestion and processing systems.
  • Candidates should have strong Python programming skills and proficiency in C++, Rust, Java, or Go, along with experience in preserving data provenance across multimodal data.
  • The annual base salary for this position in San Francisco ranges from $190,000 to $240,000, depending on the candidate's skills and experience.
  • This on-site role emphasizes trust and ownership, with a streamlined interview process aimed to be completed within one week.

The Mission

GRAM is a self replication company creating populations of insectoids for the physical economy.

Our first research frontier is self-preservation: the base case of physical self-replication. Our machines will survive, coordinate, and recover without humans. We believe scalable machine labor requires more than single-agent task generality or machines shaped in our image.

About the role

You will build the infrastructure that turns physical operation into reproducible training and evaluation datasets. Your scope begins at the capture contract and spans multimodal ingestion, temporal alignment, provenance, quality controls, storage, dataset construction, replay, and reliable access for training and evaluation.

This is a senior software engineering role responsible for the systems after capture and the contracts that keep recorded experience compatible with downstream use. Success means a model behavior can be traced through its dataset, run, software, calibration, commands, interventions, outcomes, and hardware state—and that dataset revisions remain reproducible rather than becoming ungoverned data volume.

What you will do

  • Build reliable ingestion and processing systems for synchronized video, sensor, command, capture-device state, intervention, outcome, and machine-health data.
  • Define versioned schemas and lineage connecting every run to its robot configuration, calibration, software, model, operator protocol, and experiment.
  • Develop automated checks for time drift, missing streams, corruption, calibration faults, schema breaks, weak coverage, and silent data loss.
  • Build systems for indexing, filtering, sampling, curation, dataset versioning, replay, and delivery into training and evaluation workflows.
  • Instrument the edge-to-training path with service-level metrics, traceability, failure isolation, and safe reprocessing.
  • Design storage and compute paths that support large multimodal records without losing reproducibility or making iteration dependent on manual recovery.
  • Use training and evaluation evidence to revise capture contracts, quality thresholds, dataset composition, and retention policy.

Minimum qualifications

  • Bachelor's degree in computer science, electrical engineering, applied mathematics, or a related field, or equivalent practical experience.
  • Strong Python programming ability and working proficiency in C++, Rust, Java, or Go.
  • Experience operating object-storage and streaming or batch-processing systems against defined throughput, freshness, data-integrity, or reliability targets, including schema evolution, orchestration, and safe backfills.
  • Direct experience preserving provenance and temporal relationships across video, time-series, sensor, event, or other multimodal data.
  • Demonstrated ownership of a production pipeline incident that caused silent data loss, corrupt data, or an unavailable downstream dataset, including detection, root cause, recovery or backfill, and a test or monitor that prevented recurrence.

Preferred experience

  • Robotics, autonomous vehicles, fleet telemetry, teleoperation, or another embodied-data environment.
  • Training-data systems, multimodal alignment, active-learning queues, dataset observability, or reproducible replay.
  • Edge capture in bandwidth-constrained or intermittently connected environments.

The annual base salary range for this San Francisco Bay Area position is $190,000–$240,000. Health coverage, benefits, and generous equity come with the role. This role is on-site in the San Francisco Bay Area. A relocation bonus is available. We hire to start as soon as possible. We review what you’ve shipped; if it’s the caliber we seek, interviews take about a week, start to finish. We treat your work and conversations with discretion.




Learn more about this Employer on their Career Site

Apply now in a few quick clicks

By applying, a Sonicjobs account will be created for you. Sonicjobs's Privacy Policy and Terms & Conditions will apply.

SonicJobs' Terms & Conditions and Privacy Policy also apply.