Posted on:June 13, 2026

Member of Technical Staff, Research at The Token Company

The Token Company is hiring a Member of Technical Staff, Research in San Francisco, CA, US. On-site.

About The Token Company

The Token Company provides a prompt compression API for large language models. Its product, bear-2, strips low-signal tokens from inputs before they are sent to LLMs such as GPT, Claude, and Gemini, reducing API costs while aiming to keep accuracy flat or improve it. The company also publishes case studies showing compression results for clients.

Member of Technical Staff, Research job description

Member of Technical Staff, Research

The Token Company (YC W26, HF0 S26) · San Francisco, CA

The Token Company is a seed stage startup in San Francisco. We have raised $12M from First Round Capital, NEA, Y Combinator and SV Angel, alongside the founders of Dropbox, Slack, Supercell and Hugging Face, and key people from OpenAI, xAI and DoorDash.

We train machine learning models that compress raw LLM inputs before they reach the expensive model. Our models learn which parts of an input are redundant and strip them out, cutting inference costs for the scale-ups and enterprises building on LLMs. We are a team of five with a tight research and product focus, and we ship what we train.

The role

As a Member of Technical Staff on the research team, you own the model training stack end to end: data, architecture, training, evaluation, and shipping compression models into production. This is a from-scratch model training role. You will spend your time designing and training models on NVIDIA B200s, not wiring up RAG pipelines or prompting other people's APIs.

Who you are

  • You have trained models from scratch. You have owned data, architecture and a training loop yourself, ideally in a research lab, a startup or a scale-up. Not only fine-tuning or calling an API.

  • You are strong in ML fundamentals. You are fluent with transformers and comfortable turning a paper or a rough idea into a real training run.

  • You learn fast. You get to the core of a hard problem quickly and iterate on new ideas without hand-holding.

  • You care about production. You would rather see your model cut real inference costs for real users than add another line to a publication list.

You will stand out if you have

  • Pretrained a transformer, or done serious post-training or RL on one

  • Built a novel architecture or training method with results to back it up

  • Shipped a model you trained into something people actually use

Probably not a fit if

  • Your experience is mostly RAG, agents, prompt engineering, or fine-tuning existing models through an API

  • You want to publish papers more than you want to ship models

Compensation and support

Competitive base salary plus significant equity. We also cover housing, food, laundry and cleaning in San Francisco, sponsor visas, and give you the resources and a real seat to build the research team around you.

thetokencompany.com

Apply now

Applications go straight to The Token Company. We never sit between you and the employer.

Apply now ↗
Get new startup jobs by email
Pick what you want to hear about. The first email confirms your alert, then we only write when something new matches. You can unsubscribe from any email.
1 of 15 categories
How often

Similar engineering jobs

All engineering jobs →