Level

typesafe ai

Member of Technical Staff, Model Capabilities

typesafe ai is hiring a Member of Technical Staff, Model Capabilities in San Francisco, United States. It pays $150k-$250k a year and Level rates it ; you can apply on Level.

AI in this role

openaiclaudecursorclaude-code
About TypeSafe

TypeSafe AI is an AI lab building machine-native intelligence infrastructure for automation, designed to make decisions within software by combining the intelligence of LLMs with the efficiency and reliability of code into a new shape of AI: System One Models. Based in San Francisco, TypeSafe AI recently launched its first public model, Jev.

While others chase benchmarks and academic puzzles, we’ve been quietly rethinking the LLM stack from first principles — building a new kind of general frontier model designed for real-world reliability, decision-making, and autonomy in production.

We’re a small, fast-moving team from OpenAI, Google Brain, and Meta/FAIR, backed by top-tier investors. Since mid-2024, we’ve been engineering the foundation for what comes after the current “state-of-the-art” — a model that actually gets things done.

About the role

We're looking for scrappy engineers who ship data-tooling interfaces fast. You will progress model intelligence through the interfaces, datasets, and evaluation tooling you build, combining product sense, data science, and engineering. You'll be part of the real "secret sauce" of TypeSafe: figuring out how our models can provide production-ready reliability in the real world.

Our tech stack is primarily Python. We also use TypeScript, Next.js, and Tailwind CSS for frontend, with Kubernetes for orchestration. We empower developers to use any tooling they find helpful for getting their job done, including Claude Code and Cursor.

What you'll do

Check out our team video: https://vimeo.com/1231083924/22b52f6b1d

  • Create high-leverage datasets, products, and user interfaces that unlock new capabilities and use cases on top of our model

  • Build internal data-tooling and evaluation interfaces — fast — that make model development clearer (evaluation, debugging, data inspection)

  • Develop and own evaluation frameworks to measure quality, reliability, and emergent characteristics across model iterations

  • Run rigorous analyses and experiments to understand how data, training choices, and targeted interventions impact model behavior

  • Turn raw model outputs and data into clear, inspectable UIs — owning a domain end-to-end, with craft and judgment over simple optimization

  • Partner closely with product managers and customers to translate real-world needs into concrete model and system improvements

  • Continuously iterate, learn, and co-discover new techniques for building exceptional, trustworthy AI products

We are looking for people who

  • Are generalists (~2–7 yrs) with solid fundamentals, strong coding ability, and product intuition

  • Maintain high attention to detail and a strong bar for quality

  • Enjoy thinking beyond the code to how systems are used in the real world

  • Are scrappy and hacky, and thrive in ambiguous problem spaces — taking satisfaction in finding creative solutions

  • Don't trust LLMs blindly — they look at the reality of model generations

  • Have hands-on experience implementing LLMs and understand their capabilities and limitations

Life at TypeSafe

We’re a small, flat, close-knit team working to make intelligence dependable enough to become part of everyday software. We work fully in person from our San Francisco office near Embarcadero station. We love what we do and care deeply about the work.

We strive for excellence and craftsmanship and won’t stop until we get there. When the team wins, we all win, and we enjoy collaborating and inspiring each other to grow—as a team and as individuals.

We value emotional honesty, kindness, and bringing your whole self to work. We build machines; we don’t try to be machines.

We want TypeSafe to be the place where you do the most impactful work of your career and help define our future as a company.

We provide
  • Base salary of $150k–250k plus equity, based on leveling

  • 100% covered health insurance

  • Daily lunch and dinner

  • Visa sponsorships

  • 401K plans

How we rate this

Member of Technical Staff, Model Capabilities at typesafe ai rates 86 out of 100 for how much of the daily work is AI. That makes it Builds AI (AI Level 4 of 4). The level is about AI in the job, not seniority.

Classification

Builds AI. The job is building AI systems.

  1. ●●●● Builds AI80 to 100
  2. ●●●○ Works on AI60 to 79
  3. ●●○○ Uses AI40 to 59
  4. ●○○○ Little AI0 to 39

Levels come from how often the tools, models and workflows of the role are named in the posting itself. Open the description and count.

Prepare for this job

A free preview built only from this posting: what it asks for, what you could be asked in an interview, and how to adjust your resume.

Skills and AI tools this role asks for

OpenAIClaudeCursorClaude Code

Questions you could be asked

  1. What's a project where you used OpenAI hands-on?
  2. Walk me through how you've used Claude in your day-to-day work.
  3. What are the limits of Cursor that you've run into, and how did you work around them?
  4. What's a project where you used Claude Code hands-on?
  5. How would you decide a model or AI system is ready to ship?

Adapt your resume

  • List these exact terms on your resume: OpenAI, Claude, Cursor, and Claude Code. An applicant tracking system matches the wording, not the idea.
  • Attach one line of real, concrete experience to at least one of them. A tool named with nothing behind it rarely survives a human read.
  • Lead with what you built, trained or shipped. This role is judged on the AI system itself, not the tools around it.

Want an expert to read your CV for this job?

Free. Send your CV and the role you want next. We reply by email within 2 to 4 business days.

Get new AI jobs (Builds AI ●●●●) by email

One email a week with the new AI jobs (Builds AI ●●●●), each rated for how much AI is in the work. No recruiter spam, unsubscribe in one click.

Free. One email a week. Unsubscribe in one click.

Similar roles

Other roles that build AI, at other companies.

What kind of AI work fits you?

Answer 12 practical questions in about three minutes. Get a simple profile, the work it points to, and live roles to explore next.

Find my next step

More jobs at typesafe ai

Related searches

Same AI level

Jobs by city