SonicJobs Logo
Left arrow iconBack to search

Member of Technical Staff, Research

The Token Company
Posted 2 months ago, valid for 20 days
Location

San Francisco, San Francisco, CA

Salary

Competitive

Contract type

Full Time

By applying, a Sonicjobs account will be created for you. Sonicjobs's Privacy Policy and Terms & Conditions will apply.

SonicJobs' Terms & Conditions and Privacy Policy also apply.

Sonic Summary

info
  • The Token Company, a seed stage startup in San Francisco, is seeking a Member of Technical Staff for their research team with a focus on machine learning model training.
  • Candidates should have experience training models from scratch, owning data, architecture, and training loops, ideally with 3+ years in a research lab, startup, or scale-up environment.
  • The role involves designing and training models on NVIDIA B200s, with an emphasis on cutting inference costs for users rather than publishing papers.
  • Compensation includes a competitive base salary, significant equity, and additional support covering housing, food, laundry, and cleaning in San Francisco.
  • Ideal candidates will have experience with transformers, a strong understanding of ML fundamentals, and a passion for shipping models that provide real user value.

Member of Technical Staff, Research

The Token Company (YC W26, HF0 S26) 路 San Francisco, CA

The Token Company is a seed stage startup in San Francisco. We have raised $12M from First Round Capital, NEA, Y Combinator and SV Angel, alongside the founders of Dropbox, Slack, Supercell and Hugging Face, and key people from OpenAI, xAI and DoorDash.

We train machine learning models that compress raw LLM inputs before they reach the expensive model. Our models learn which parts of an input are redundant and strip them out, cutting inference costs for the scale-ups and enterprises building on LLMs. We are a team of five with a tight research and product focus, and we ship what we train.

The role

As a Member of Technical Staff on the research team, you own the model training stack end to end: data, architecture, training, evaluation, and shipping compression models into production. This is a from-scratch model training role. You will spend your time designing and training models on NVIDIA B200s, not wiring up RAG pipelines or prompting other people's APIs.

Who you are

  • You have trained models from scratch. You have owned data, architecture and a training loop yourself, ideally in a research lab, a startup or a scale-up. Not only fine-tuning or calling an API.

  • You are strong in ML fundamentals. You are fluent with transformers and comfortable turning a paper or a rough idea into a real training run.

  • You learn fast. You get to the core of a hard problem quickly and iterate on new ideas without hand-holding.

  • You care about production. You would rather see your model cut real inference costs for real users than add another line to a publication list.

You will stand out if you have

  • Pretrained a transformer, or done serious post-training or RL on one

  • Built a novel architecture or training method with results to back it up

  • Shipped a model you trained into something people actually use

Probably not a fit if

  • Your experience is mostly RAG, agents, prompt engineering, or fine-tuning existing models through an API

  • You want to publish papers more than you want to ship models

Compensation and support

Competitive base salary plus significant equity. We also cover housing, food, laundry and cleaning in San Francisco, sponsor visas, and give you the resources and a real seat to build the research team around you.

thetokencompany.com




Learn more about this Employer on their Career Site

Apply now in a few quick clicks

By applying, a Sonicjobs account will be created for you. Sonicjobs's Privacy Policy and Terms & Conditions will apply.

SonicJobs' Terms & Conditions and Privacy Policy also apply.