Page 1
Forward Deployed Engineer (Training)
baseten · San Francisco, California, United States
$200k–400k
21 days ago
♡
Software Engineer - AI Developer Productivity
baseten · San Francisco, California, United States
$165k–330k
29 days ago
♡
Software Engineer - Testing Frameworks
baseten · San Francisco, California, United States
$165k–330k
29 days ago
♡
Software Engineer - Continuous Delivery
baseten · San Francisco, California, United States
$165k–330k
29 days ago
♡
Software Engineer - Observability
baseten · San Francisco, California, United States
$165k–330k
30 days ago
♡
Loading more openings…
You've reached the end of the list.
AI Engineer
baseten · San Francisco, California, United States
Pay
$220k–260k
Setting
Hybrid
ABOUT BASETEN
Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F https://www.baseten.co/blog/announcing-our-series-f/, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products.
THE ROLE:
Are you the person on your team who builds the agent everyone else ends up using? We're looking for an AI Engineer to join our Training Product team and do that at Baseten. You'll build AI-driven product features for the customers training and post-training frontier models on our platform, and you'll raise the ceiling on how Baseten itself uses AI internally, turning manual workflows into agentic ones that make every other team faster.
You'll work directly with our research engineers to scope and build products, taking ideas from a research loop that already works internally to something customers can run themselves. This is a hands-on role with real autonomy. You'll pick the problems worth solving, build the harnesses, execution flows, and guardrails that make AI systems reliable, and own the results. If you've been shipping agents and want that to be the job, let's talk.
EXAMPLE INITIATIVES:
Take a look at these blog posts written by members of our team:
- Baseten Training: an autoresearch substrate https://www.baseten.co/blog/baseten-training-an-autoresearch-substrate/
- Introducing Baseten Loops https://www.baseten.co/blog/introducing-the-baseten-loops-sdk/
- Harnesses are everything. Here's how to optimize yours. https://www.baseten.co/blog/harnesses-are-everything-heres-how-to-optimize-yours/
- Building with NVIDIA Nemotron 3 Ultra and LangChain Deep Agents Code on Baseten https://www.baseten.co/blog/nvidia-nemotron-3-ultra-and-langchain-deep-agents-on-baseten/
RESPONSIBILITIES:
- Build and ship agentic product experiences, including chat-style and assistant-like interfaces, from prototype to GA.
- Design the harnesses, execution flows, and guardrails that make AI systems reliable in production.
- Build internal automation and AI tooling that measurably increases the velocity of engineering, research, and go-to-market teams.
- Partner with research engineers to scope product opportunities out of internal research workflows and turn them into customer-facing features.
- Define and instrument evals so you know whether a change actually improved output quality.
- Work throughout the stack (API layer, backend, agent orchestration, frontend) to implement features end to end.
- Use Baseten's own training and inference products yourself to develop intuition around customer workflows.
- Identify where AI can replace manual process across the company and build the thing rather than write the proposal.
- Fix bugs and resolve customer issues with urgency.
REQUIREMENTS:
- 5+ years of experience building and shipping software applications.
- Demonstrated experience building AI or LLM-powered products, agents, or agentic workflows that real users depend on.
- Strong software engineering fundamentals and the ability to clear a real technical bar, not just prompt well.
- Ability to build accurate mental models of how systems work under the hood, including the models and harnesses you're building on.
- Proficiency in Python, with fluency in at least one other language.
- Comfort working autonomously in a fast-moving environment with limited structure.
- Ability to move between customer-facing product work and internal tooling and automation.
- Strong communication skills, with the ability to bridge technical depth and business needs.
NICE TO HAVE:
- Experience as a founding engineer or early employee at a startup.
- Experience building evals, agent observability, or tooling for non-deterministic systems.
- Familiarity with agent frameworks and harnesses (LangChain, Claude Code, Codex, OpenCode, MCP).
- Experience with model development methods like supervised fine-tuning, reinforcement learning, synthetic data generation, LoRA, and full fine-tunes.
- Frontend fluency.
BENEFITS
- Competitive compensation, including meaningful equity.
- 100% coverage of medical, dental, and vision insurance for employee and dependents
- Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!)
- Paid parental leave
- Fertility and family-building stipend through Carrot
- Company-facilitated 401(k)
- Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities.
Apply now to embark on a rewarding journey in shaping the future of AI! If you are a motivated individual with a passion for machine learning and a desire to be part of a collaborative and forward-thinking team, we would love to hear from you.
At Baseten, we are committed to fostering a diverse and inclusive workplace. We provide equal employment opportunities to all employees and applicants without regard to race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, or veteran status.
We are an Equal Opportunity Employer and will consider qualified applicants with criminal histories in a manner consistent with applicable law (by example, the requirements of the San Francisco Fair Chance Ordinance, where applicable).
Listed by baseten for a position based in the United States. Employers on this board attest they are hiring domestically.
Forward Deployed Engineer (Training)
baseten · San Francisco, California, United States
Pay
$200k–400k
Setting
Hybrid
ABOUT BASETEN
Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F https://www.baseten.co/blog/announcing-our-series-f/, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products.
THE ROLE
Forward Deployed Engineers work directly with the largest and fastest-growing AI companies in the world, owning their technical outcomes on Baseten and taking on the hardest problems in serving and improving models at scale. The work spans the model lifecycle: inference, post-training, and the systems that tighten the loop between them.
- Act as each account's de facto CTO on Baseten, with final accountability for how their workloads are designed, run, and scaled.
- Take customer objectives from vague to shipped: frame the problem, define the spec and success criteria, build the PoC, and carry it through to production quickly, using the right tools for the problem.
- Design the evals and benchmarks that isolate where quality or performance falls short, then close the gap yourself, whether that means optimizing inference, improving the model through post-training, or reworking the eval itself.
- Be the first responder to mission-critical failures including triage, owning the fix directly or route to the owning team and stay accountable until it ships.
- Build internal systems so that each engagement is faster than the last. This includes tooling and automation for eval and deployment infrastructure, and the recipes and reference implementations that make the product more self-serve.
- Shape the product itself, channeling what your accounts need into the roadmap and shipping fixes and features into Baseten's codebase yourself.
- Do all of this across multiple accounts at once, sequencing the work, pulling in the right people at the right time, and keeping customers and internal stakeholders aligned on status and risk.
QUALIFICATIONS
- Minimum 1-2 years of software engineering experience, shipping and maintaining code in large production systems, ideally with breadth across the stack
- Experience debugging complex production issues - working through logs, metrics, and traces to root-cause problems in unfamiliar systems
- Confidence owning ambiguous technical problems. This includes triaging, making decisions under uncertainty, and knowing when to pull in other engineers who own the underlying systems
- Motivation beyond pure engineering. An interest in wanting to work directly with customers, understand their problems firsthand, and influence the product
- Clear communication on complex technical topics, whether you're talking to a customer's engineers or their leadership
- Genuine curiosity about AI inference and training, and a drive to become an expert in the infrastructure powering it
- Willingness to respond to customers outside regular working hours and participate in an on-call rotation
- Excitement about solving problems for some of the largest and fastest-growing companies in the world running mission-critical AI workloads
WHAT YOU'LL BRING
We don't expect any one person to cover all of this - the strongest candidates may spike in one or two of the following:
- Depth in a core infrastructure domain such as storage systems or networking (anywhere from the cloud layer to cluster interconnects like InfiniBand and RoCE)
- Experience operating distributed compute platforms like Kubernetes, Slurm, or Ray, especially for GPU workloads
- A detailed understanding of LLM architectures and modern inference engines like vLLM, TensorRT-LLM, or SGLang
- The ability to profile and optimize GPU workloads, in training or serving
- Hands-on experience with post-training techniques like SFT and RL, or a broader deep learning background plus fluency in a tensor computation library like PyTorch or JAX
- Operational depth - running on-call, leading incident response, debugging distributed systems under pressure
BENEFITS
- Competitive compensation, including meaningful equity.
- 100% coverage of medical, dental, and vision insurance for employee and dependents
- Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!)
- Paid parental leave
- Fertility and family-building stipend through Carrot
- Company-facilitated 401(k)
- Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities.
Apply now to embark on a rewarding journey in shaping the future of AI! If you are a motivated individual with a passion for machine learning and a desire to be part of a collaborative and forward-thinking team, we would love to hear from you.
At Baseten, we are committed to fostering a diverse and inclusive workplace. We provide equal employment opportunities to all employees and applicants without regard to race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, or veteran status.
We are an Equal Opportunity Employer and will consider qualified applicants with criminal histories in a manner consistent with applicable law (by example, the requirements of the San Francisco Fair Chance Ordinance, where applicable).
Listed by baseten for a position based in the United States. Employers on this board attest they are hiring domestically.
Head of IT
baseten · San Francisco, California, United States
Pay
$210k–250k
Setting
Hybrid
ABOUT BASETEN
Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F https://www.baseten.co/blog/announcing-our-series-f/, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products.
THE ROLE
As the Head of IT at Baseten, you will build, scale, and secure our internal technology function to support our rapid growth. Reporting to our Chief Information Security Officer, you will lead and mentor a team of 5+ IT engineers, leading the charge to transition Baseten from startup-era IT to a highly automated, enterprise-ready IT organization. You will take full ownership of corporate IT infrastructure, Helpdesk operations, corporate identity management, device lifecycles, and vendor procurement. As we scale to support the world’s most dynamic AI companies, you will ensure our internal systems scale seamlessly with our headcount, providing a secure, frictionless, and world-class technology experience for all Baseten employees.
RESPONSIBILITIES
- Team Leadership: Manage, mentor, and grow a team of IT engineers, fostering a high-performance culture focused on technical excellence and end-user satisfaction.
- Helpdesk Operational Excellence: Build a fast-response support function by establishing clear response SLAs, tracking employee satisfaction metrics, and formalizing on-call and incident response processes.
- Zero-Touch Automation: Architect and implement automated employee onboarding, offboarding, and role-based access changes through deep integrations across HRIS, MDM, and IAM systems.
- SaaS Management & Procurement: Establish comprehensive SaaS management processes to eliminate shadow IT, automate access workflows, optimize vendor spend per employee, and build visibility into corporate tools and budgets.
- AI-Enabled Workflows: Spearhead the deployment of internal AI tools and automated IT agents—powered by Baseten models—to resolve routine tickets, maintain knowledge bases, and significantly boost internal productivity.
- Infrastructure & Hardware Ownership: Manage hardware logistics, global device lifecycles, software license consolidation, and the overarching corporate IT budget.
REQUIREMENTS
- Experience: 5+ years of overall IT leadership experience, ideally accompanied by 2+ years of experience as a Head of IT or IT Director at an internet company, scaling developer-focused SaaS or AI infrastructure startups through rapid growth (e.g., 100 to 500+ employees).
- Modern Cloud-Native Stack Mastery: Deep, hands-on expertise with Zero Trust Network Access (ZTNA), IAM/SSO (Okta, Google Workspace), MDM (Kandji, Jamf), and HRIS integrations (Rippling).
- Vendor & Financial Discipline: A proven track record of successful vendor negotiations, complex hardware logistics, software license consolidation, and strict IT budget ownership.
- AI-First Mindset: Proactive and forward-thinking approach to leveraging modern AI tooling and automation to eliminate operational friction, rather than relying on manual headcount expansion.
- Process Engineering: Demonstrated ability to build enterprise-ready processes from the ground up, establish performance metrics, and drive continuous improvement in a fast-paced environment.
BENEFITS
- Competitive compensation, including meaningful equity.
- 100% coverage of medical, dental, and vision insurance for employee and dependents
- Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!)
- Paid parental leave
- Fertility and family-building stipend through Carrot
- Company-facilitated 401(k)
- Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities.
Apply now to embark on a rewarding journey in shaping the future of AI! If you are a motivated individual with a passion for machine learning and a desire to be part of a collaborative and forward-thinking team, we would love to hear from you.
At Baseten, we are committed to fostering a diverse and inclusive workplace. We provide equal employment opportunities to all employees and applicants without regard to race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, or veteran status.
We are an Equal Opportunity Employer and will consider qualified applicants with criminal histories in a manner consistent with applicable law (by example, the requirements of the San Francisco Fair Chance Ordinance, where applicable).
Listed by baseten for a position based in the United States. Employers on this board attest they are hiring domestically.
Software Engineer - AI Developer Productivity
baseten · San Francisco, California, United States
Pay
$165k–330k
Setting
Hybrid
ABOUT BASETEN
Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F https://www.baseten.co/blog/announcing-our-series-f/, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products.
THE ROLE
Baseten's engineers want to work in an AI-first way. What's missing isn't enthusiasm — it's the platform underneath it. Today everyone assembles their own agent config, context files, and MCP servers, so the good patterns stay trapped in individual setups instead of becoming defaults everyone inherits.
You'll build that platform: the agent configurations tuned to our monorepo, the context and tooling layer that makes agents competent in our codebase, the evals that tell us which approaches actually work, and the rollout mechanics that get a new engineer productive with agents in week one.
You are not here to mandate how engineers use AI — you're here to make the good path the easy path. Success looks like teams adopting what you build because it beats what they'd cobble together themselves, not because a policy requires it. Platform engineer, not AI evangelist. Ship infrastructure, measure it, kill what doesn't work, let adoption be the referee.
The playbook for AI-first SDLC doesn't exist at any company yet. You'll write ours.
WHAT YOU'LL BUILD
Agent substrate — Repo-level context infrastructure that makes agents competent in our codebase (CLAUDE.md/AGENTS.md http://CLAUDE.md/AGENTS.md conventions, architecture and domain context, and the tooling to keep it accurate as code moves). Internal MCP servers giving agents scoped access to CI, observability, incident tooling, deployment state, and docs. Shared skills, subagents, and hooks that encode Baseten workflows. Sandboxed environments where agents can build and test safely.
The golden path — Project templates and onboarding that ship with AI tooling configured and working. Self-serve infrastructure so teams build their own agents without you as the bottleneck. Gateway, auth, cost controls, and audit logging for internal model access.
The feedback loop — Eval harnesses that answer "is this config better than that one" against real Baseten tasks, not vibes. Instrumentation of AI tool usage and its downstream effects on cycle time, review latency, and change failure rate. Honest reporting, including on what you built that didn't pan out.
Agents in the SDLC — Automation where agents earn their keep: PR review triage, test gap-filling, incident context assembly, migrations and refactors, codebase Q&A. Integrating agents into CI/CD with guardrails that make it trustworthy.
RESPONSIBILITIES
- Own the internal AI developer platform end to end — architecture, build, rollout, operation, measurement.
- Evaluate and integrate third-party AI coding tools (Claude Code, Cursor, Codex, and whatever ships next quarter), and build the context layer that makes them work against our monorepo.
- Build frameworks that let other engineers create their own agents without deep LLM expertise.
- Establish the evaluation practice for AI-assisted development at Baseten, and use it to drive investment decisions.
- Drive adoption through developer experience — good defaults, clear docs, low friction — not mandate.
- Embed with teams to find where AI genuinely unblocks them, then generalize those wins into platform capabilities.
- Own the safety layer: permissions, secrets handling, audit trails, cost management.
REQUIREMENTS
- Have 4+ years of relevant industry experience building and enabling AI native SDLC
- Strong proficiency in Python and/or Go, building tools other engineers depend on daily.
- Hands-on experience with LLMs and agent frameworks — tool calling, MCP, context management, orchestration, failure handling. You've shipped something agentic that real people used, not just prototyped.
- Deep personal fluency with AI coding tools and well-formed opinions about where they break down.
- Platform mindset: you build for adoption and self-service, treat internal engineers as customers, and would rather ship a good default than write a style guide.
- Developer tooling, CI/CD, and Kubernetes/Docker fundamentals.
- Comfort with ambiguity — this space invalidates its own best practices every few months.
- Excellent written communication. Much of your leverage is docs, templates, and examples that scale beyond conversations you're in.
BENEFITS
- Competitive compensation, including meaningful equity.
- 100% coverage of medical, dental, and vision insurance for employee and dependents
- Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!)
- Paid parental leave
- Fertility and family-building stipend through Carrot
- Company-facilitated 401(k)
- Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities.
Apply now to embark on a rewarding journey in shaping the future of AI! If you are a motivated individual with a passion for machine learning and a desire to be part of a collaborative and forward-thinking team, we would love to hear from you.
At Baseten, we are committed to fostering a diverse and inclusive workplace. We provide equal employment opportunities to all employees and applicants without regard to race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, or veteran status.
We are an Equal Opportunity Employer and will consider qualified applicants with criminal histories in a manner consistent with applicable law (by example, the requirements of the San Francisco Fair Chance Ordinance, where applicable).
Listed by baseten for a position based in the United States. Employers on this board attest they are hiring domestically.
Software Engineer - Testing Frameworks
baseten · San Francisco, California, United States
Pay
$165k–330k
Setting
Hybrid
ABOUT BASETEN
Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F https://www.baseten.co/blog/announcing-our-series-f/, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products.
THE ROLE
Baseten is seeking talented and experienced Software Engineers to join our Platform team within the Infrastructure organization. As a senior member of Baseten's Platform Team, you will own the systems that let every engineer at Baseten prove their code works before it reaches production. Our product runs mission-critical AI inference for customers who measure downtime in dollars per second, which means our internal bar for correctness, performance, and failure tolerance has to be exceptional.
Your focus is the full testing stack: fast and reliable unit test tooling, integration harnesses that spin up realistic environments on demand, load and performance testing for GPU-backed inference workloads, and resilience testing that deliberately breaks things so our customers never have to find out what happens when a node dies mid-request. This is a builder role with org-wide leverage. You won't be writing tests for other teams — you'll be building the frameworks, harnesses, and feedback loops that make writing good tests the path of least resistance, and you'll set the standards for what "well-tested" means at Baseten.
RESPONSIBILITIES
- Own Baseten's testing strategy end to end — define the standards, the tiers, and the tooling that engineering teams build against.
- Build and maintain unit, integration, load and performance testing frameworks
- Design end to end test infrastructure that provisions realistic dependencies on demand — ephemeral Kubernetes namespaces, containerized service dependencies, etc.
- Build resilience and chaos testing capabilities: fault injection, network partition and latency simulation, pod and node failure scenarios, dependency degradation, and automated verification that our systems degrade gracefully and recover.
- Attack test suite reliability head-on — flake detection and quarantine, test impact analysis and selective execution, parallelization, and aggressive reduction of end-to-end CI wall time.
- Build observability into the testing layer itself: coverage reporting, test health dashboards, failure triage tooling, and clear signals about which areas of the codebase are under-tested.
- Create project templates and shared libraries so new services start with a complete testing setup on day one.
- Partner directly with product and infrastructure teams — embed where needed, understand their hardest testing problems, and turn one-off solutions into platform capabilities.
- Mentor engineers across the org on testing practice and raise the collective bar through code review, documentation, and internal advocacy.
REQUIREMENTS
- Strong proficiency in Go and/or Python, with hands-on experience building test tooling and libraries used by other engineers.
- Deep experience with testing frameworks and their internals — pytest, Go's testing package, testify, testcontainers, or equivalents — including the ability to extend them, not just use them.
- Experience designing integration test infrastructure for distributed systems, including ephemeral environment provisioning and dependency management.
- Practical experience with load and performance testing tools (k6, Locust, Vegeta, JMeter, Gatling, or similar) and the judgment to interpret results and drive action from them.
- Solid Kubernetes and Docker fundamentals — you can build test infrastructure that runs on and tests against real cluster environments.
- Advanced understanding of CI/CD, including how to keep large test suites fast, deterministic, and trustworthy at scale.
- A track record of measurably improving testing culture on an engineering team — you have opinions on what to test, what not to test, and where the ROI actually is.
- Are excited about building foundational infrastructure and are comfortable working independently on ambiguous, high-impact technical challenges
BENEFITS
- Competitive compensation, including meaningful equity.
- 100% coverage of medical, dental, and vision insurance for employee and dependents
- Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!)
- Paid parental leave
- Fertility and family-building stipend through Carrot
- Company-facilitated 401(k)
- Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities.
Apply now to embark on a rewarding journey in shaping the future of AI! If you are a motivated individual with a passion for machine learning and a desire to be part of a collaborative and forward-thinking team, we would love to hear from you.
At Baseten, we are committed to fostering a diverse and inclusive workplace. We provide equal employment opportunities to all employees and applicants without regard to race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, or veteran status.
We are an Equal Opportunity Employer and will consider qualified applicants with criminal histories in a manner consistent with applicable law (by example, the requirements of the San Francisco Fair Chance Ordinance, where applicable).
Listed by baseten for a position based in the United States. Employers on this board attest they are hiring domestically.
Software Engineer - Continuous Delivery
baseten · San Francisco, California, United States
Pay
$165k–330k
Setting
Hybrid
ABOUT BASETEN
Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F https://www.baseten.co/blog/announcing-our-series-f/, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products.
THE ROLE
Baseten is seeking talented and experienced Software Engineers to join our Platform team within the Infrastructure organization. As an early member of Baseten's Platform Team, you will be pivotal in building internal infrastructure to support our engineering organization. You will own the deployment platform, release pipelines, and rollout safety mechanisms that allow engineers across Baseten to deploy changes rapidly while minimizing operational risk. Our mission is to make production deployments fast, safe, and increasingly autonomous.
If you are passionate about elegant solutions—like streamlined monorepos, lightning-fast CI pipelines, and thoughtfully designed shared libraries—you'll thrive at Baseten.
RESPONSIBILITIES
- Design and build continuous deployment infrastructure that safely rolls out changes across dozens of Kubernetes clusters and global regions.
- Develop systems for progressive delivery, including canary releases, staged rollouts, and automated rollback.
- Improve engineering velocity by reducing friction in the release pipeline and automating manual operational workflows.
- Work with product and infrastructure teams to ensure their services are deployable, observable, and resilient at scale.
- Implement and evolve deployment methodologies such as GitOps, infrastructure-as-code, and progressive delivery patterns.
- Build systems that automatically evaluate deployment health using metrics, logs, traces, and alerts to detect regressions and trigger safe rollbacks.
- Interest in building systems that support agent-assisted or autonomous deployment workflows using modern AI tooling.
REQUIREMENTS
- Strong-level proficiency in Go and/or Python.
- Have worked with Kubernetes-based deployment systems at scale
- Have experience building or operating continuous deployment platforms
- Are familiar with GitOps tooling such as ArgoCD or Flux
- Care deeply about safe production rollouts and minimizing blast radius
- Excellent developer tool culture with a focus on best practices. Bonus points if you have personally developed open-source developer tools.
- Are excited about building foundational infrastructure and are comfortable working independently on ambiguous, high-impact technical challenges
BENEFITS
- Competitive compensation, including meaningful equity.
- 100% coverage of medical, dental, and vision insurance for employee and dependents
- Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!)
- Paid parental leave
- Fertility and family-building stipend through Carrot
- Company-facilitated 401(k)
- Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities.
Apply now to embark on a rewarding journey in shaping the future of AI! If you are a motivated individual with a passion for machine learning and a desire to be part of a collaborative and forward-thinking team, we would love to hear from you.
At Baseten, we are committed to fostering a diverse and inclusive workplace. We provide equal employment opportunities to all employees and applicants without regard to race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, or veteran status.
We are an Equal Opportunity Employer and will consider qualified applicants with criminal histories in a manner consistent with applicable law (by example, the requirements of the San Francisco Fair Chance Ordinance, where applicable).
Listed by baseten for a position based in the United States. Employers on this board attest they are hiring domestically.
Software Engineer - Observability
baseten · San Francisco, California, United States
Pay
$165k–330k
Setting
Hybrid
ABOUT BASETEN
Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F https://www.baseten.co/blog/announcing-our-series-f/, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products.
THE ROLE
Baseten is seeking talented and experienced Software Engineers to join our Observability team within the Infrastructure organization. As an early member of the Observability Team, you will be pivotal in building and shaping the observability experience for our internal and external customers. By joining this team, you’ll have a direct impact on the reliability and operational excellence of Basetens product systems.
As Baseten scales its infrastructure across different cloud providers and diverse hardware, the volume and complexity of operational data is growing by orders of magnitude. This team is responsible for building high-throughput ingest pipelines, cost-efficient storage, and agentic diagnostic tools to ensure that we can detect, diagnose, and resolve issues in minutes rather than hours, even as the systems they operate become more complex.
RESPONSIBILITIES
- Design and build scalable telemetry ingest and storage pipelines for metrics, logs, and traces across Baseten’s multi-cloud infrastructure
- Own and evolve core observability platforms, driving migrations and architectural improvements that improve reliability, reduce cost, and scale with organizational growth
- Build instrumentation libraries, SDKs, and integrations that make it easy for engineering teams to emit high-quality telemetry from their services
- Drive alerting and SLO infrastructure that enables teams to define, monitor, and respond to reliability targets with minimal noise
- Partner with Inference, Product, and Infrastructure teams to ensure observability solutions meet the unique needs of each organization
REQUIREMENTS
- Have deep experience with at least one observability signal area (metrics, logging, tracing, or error analytics) and familiarity with the others
- Understand high-throughput data pipelines, columnar storage engines, and the tradeoffs involved in ingesting and querying telemetry data at scale
- Have experience operating or building on top of observability platforms such as Prometheus, Grafana, ClickHouse, OpenTelemetry, or similar systems
- Have strong proficiency in at least one of Python, Rust, or Go
- Have excellent communication skills and enjoy partnering with internal teams to improve their operational visibility and incident response capabilities
- Interest in applying AI/LLMs to operational workflows such as automated root cause analysis, anomaly detection, or intelligent alerting
- Are excited about building foundational infrastructure and are comfortable working independently on ambiguous, high-impact technical challenges
BENEFITS
- Competitive compensation, including meaningful equity.
- 100% coverage of medical, dental, and vision insurance for employee and dependents
- Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!)
- Paid parental leave
- Fertility and family-building stipend through Carrot
- Company-facilitated 401(k)
- Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities.
Apply now to embark on a rewarding journey in shaping the future of AI! If you are a motivated individual with a passion for machine learning and a desire to be part of a collaborative and forward-thinking team, we would love to hear from you.
At Baseten, we are committed to fostering a diverse and inclusive workplace. We provide equal employment opportunities to all employees and applicants without regard to race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, or veteran status.
We are an Equal Opportunity Employer and will consider qualified applicants with criminal histories in a manner consistent with applicable law (by example, the requirements of the San Francisco Fair Chance Ordinance, where applicable).
Listed by baseten for a position based in the United States. Employers on this board attest they are hiring domestically.
Executive Assistant to CTO (co-founder) & Head of Engineering
baseten · San Francisco, California, United States
Pay
$160k–180k
Setting
Hybrid
ABOUT BASETEN
Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F https://www.baseten.co/blog/announcing-our-series-f/, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products.
THE ROLE
We’re looking for an experienced Executive Assistant to support our CTO (co-founder) and Head of Engineering. This is a highly operational role that goes well beyond calendar management. You’ll own the day-to-day operating rhythm of the Engineering organization, ensuring leaders are prepared, priorities stay coordinated, and critical meetings, communications, and follow-ups happen seamlessly.
You’ll partner closely with senior engineering leaders and serve as a trusted point of coordination for employees, customers, candidates, and external partners. Success in this role comes from exceptional organization, judgment, attention to detail, and the ability to keep many moving pieces aligned in a fast-growing environment.
RESPONSIBILITIES
- Own complex calendar management for the CTO and Head of Engineering, balancing shifting priorities while ensuring time is allocated intentionally
- Ensure leaders are prepared for every day and every meeting by proactively managing agendas, materials, context, logistics, and follow-ups so time is used effectively and decisions move forward
- Support forward-looking calendar planning, coordinating recurring operating cadences including roadmap planning, leadership meetings, P0 reviews, Engineering All Hands, and other cross-functional forums
- Own the operational cadence of the Engineering organization, including weekly leadership meetings, monthly Show & Tells, Engineering All Hands, roadmap planning, and other recurring forums
- Support the CTO (co-founder) and Head of Engineering in delivering an exceptional candidate and new hire experience—from providing high-touch support for executive candidates during onsite interviews to ensuring new team members are welcomed personally during their first week
- Partner with customers and external partners to coordinate executive meetings, visits, and other engagements
- Manage executive inboxes, helping prioritize communications, triage requests, and ensure timely follow-up
- Coordinate domestic and international travel, including last-minute changes
- Continuously improve how the Engineering leadership team operates by identifying opportunities to make processes more efficient, scalable, and reliable
REQUIREMENTS
- 5+ years of experience supporting senior Engineering leaders (CTO, VP Engineering, Head of Engineering, or similar) in a high-growth technology company. You understand how technical organizations operate and have experience supporting engineering operating rhythms, technical planning cycles, and fast-moving product development teams.
- Exceptional organizational skills with the ability to manage multiple competing priorities without dropping details
- Outstanding written and verbal communication skills, with strong judgment about when and how to communicate on behalf of executives
- Demonstrated ability to coordinate complex cross-functional initiatives involving many stakeholders
- Experience managing complex calendars, travel, executive communications, and recurring operating cadences
- Comfortable working with senior leaders, customers, candidates, and external partners with professionalism and executive presence
- Proactive, resourceful, and highly autonomous. You anticipate needs before they become problems.
- High attention to detail and a strong sense of ownership. You take pride in creating structure, improving processes, and ensuring nothing falls through the cracks.
- Thrives in a high-growth environment where priorities change quickly and adaptability is essential.
BENEFITS
- Competitive compensation, including meaningful equity.
- 100% coverage of medical, dental, and vision insurance for employee and dependents
- Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!)
- Paid parental leave
- Fertility and family-building stipend through Carrot
- Company-facilitated 401(k)
- Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities.
Apply now to embark on a rewarding journey in shaping the future of AI! If you are a motivated individual with a passion for machine learning and a desire to be part of a collaborative and forward-thinking team, we would love to hear from you.
At Baseten, we are committed to fostering a diverse and inclusive workplace. We provide equal employment opportunities to all employees and applicants without regard to race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, or veteran status.
We are an Equal Opportunity Employer and will consider qualified applicants with criminal histories in a manner consistent with applicable law (by example, the requirements of the San Francisco Fair Chance Ordinance, where applicable).
Listed by baseten for a position based in the United States. Employers on this board attest they are hiring domestically.
AI Inference Engineer
baseten · San Francisco, California, United States
Pay
$165k–330k
Setting
Remote
Partner with our customers to understand their problems and engineer ML solutions using Baseten.
Listed by baseten for a position based in the United States. Employers on this board attest they are hiring domestically.
Software Engineer - GPU Fabric Observability
baseten · San Francisco, California, United States
Pay
$200k–380k
Setting
Hybrid
ABOUT BASETEN
Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F https://www.baseten.co/blog/announcing-our-series-f/, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products.
THE ROLE
Baseten is building its own GPU infrastructure for large-scale inference. As we move into large scale, high-density NVIDIA systems, the hardest failures are intermittent, cross-layer, and difficult to prove: RoCE congestion, InfiniBand stalls, ECN/DCQCN mis-tuning, bad optics, RNIC issues, host kernel stalls, GPU driver problems, and workload symptoms that look like network problems, but are not.
We are hiring a Software Engineer to build a first-class observability and root-cause analysis system for GPU fabrics. This is a hard distributed systems problem, not a dashboarding problem. The system will collect high-volume signals from switches, hosts, active probes, and inference services; reduce and correlate them in real time; understand topology and service ownership; and produce actionable diagnosis while an incident is still unfolding.
This role sits at the boundary between networking and inference software. RDMA data paths, GPUDirect transfers, prefill/decode disaggregation, KV cache movement, request routing, and workload backpressure can all create fabric symptoms or hide real fabric failures. The goal is to tell an operator, quickly and with evidence, whether an incident is caused by the fabric, host, NIC, GPU, RDMA path, scheduler, or serving layer — and what to do next.
EXAMPLE INITIATIVES
- Real-time telemetry engine — Build the ingestion, reduction, storage, and query path for high-cardinality fabric, host, GPU, and workload telemetry.
- Service-aware fabric diagnosis — Build collectors, probes, and topology-aware correlation to detect latency, drops, stalls, congestion, bad paths, and degradation.
- Software-aware RDMA diagnosis — Tie network behavior to RDMA operations, GPUDirect paths, KV cache transfers, prefill/decode disaggregation, and request latency.
RESPONSIBILITIES
- Own Baseten’s GPU fabric observability and root-cause analysis architecture.
- Build telemetry pipelines across switches, NICs, hosts, GPUs, Kubernetes, and inference services.
- Model topology, flow paths, service ownership, and failure domains.
- Separate true fabric faults from host, NIC, GPU, kernel, driver, RDMA, scheduler, and workload failures.
- Create clear operator workflows for triage, remediation, and post-incident learning.
REQUIREMENTS
- Staff-level or senior staff-level experience building production infrastructure software.
- Strong distributed systems background, especially streaming systems, telemetry pipelines, diagnostics, or control-plane software.
- Experience building systems that process high-volume, high-cardinality, noisy operational data.
- Understanding of networking fundamentals and high-performance networks
- Ability to work with low-level infrastructure signals and build practical correlation, anomaly detection, or root-cause analysis systems.
BENEFITS
- Competitive compensation, including meaningful equity.
- 100% coverage of medical, dental, and vision insurance for employee and dependents
- Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!)
- Paid parental leave
- Fertility and family-building stipend through Carrot
- Company-facilitated 401(k)
- Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities.
Apply now to embark on a rewarding journey in shaping the future of AI! If you are a motivated individual with a passion for machine learning and a desire to be part of a collaborative and forward-thinking team, we would love to hear from you.
At Baseten, we are committed to fostering a diverse and inclusive workplace. We provide equal employment opportunities to all employees and applicants without regard to race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, or veteran status.
We are an Equal Opportunity Employer and will consider qualified applicants with criminal histories in a manner consistent with applicable law (by example, the requirements of the San Francisco Fair Chance Ordinance, where applicable).
Listed by baseten for a position based in the United States. Employers on this board attest they are hiring domestically.
Select a role
The full posting opens here — pay, setting and the full description, without leaving the list.