Machine Learning Infrastructure Tech Lead
Solutions Engineering Manager
Forward Deployed Engineer, Infrastructure Specialist
About Reducto
Reducto is the agentic document platform for leading AI teams who demand enterprise performance at scale. We provide a comprehensive toolkit for working with documents the way a human would, combining custom in-house and leading frontier models to power efficient and accurate document workflows.
We’ve grown rapidly, increasing revenue 8x year over year and partnering with hundreds of companies, from leading AI teams like Harvey, Vanta, and Scale, to enterprise customers across FAANG and top trading firms.
Reducto has raised over $100M from world-class investors including a16z, Benchmark, and First Round Capital, and are hiring a founding Infrastructure Engineer.
The Opportunity
As a Forward Deployed Engineer, Infrastructure Specialist you will lead the end-to-end deployment of Reducto into customer environments, ranging from VPC to bare metal on-prem. Each deployment comes with its own hardware, networking, security, and reliability constraints, and you will be the engineer who makes them work.
This role sits at the bridge between Reducto and our enterprise customers' infrastructure, security, and platform teams. You will partner directly with those teams to deploy, configure, harden, and operate Reducto inside their environments. This is hands-on execution work with direct customer exposure, not a bottoms-up platform role.
The core work will include:
Leading end-to-end deployment of Reducto into customer environments including planning, configuration, testing, and rollout.
Partnering with enterprise IT, security, and platform teams to assess their infrastructure, security posture, and data management practices, and designing deployment strategies tailored to their constraints.
Understanding the hardware that power our in-house ML models so we can successfully deploy them onto customer infrastructure.
Building monitoring, telemetry, and phone-home patterns that give us visibility into customer-controlled environments without compromising their security posture.
Debugging and resolving customer-specific infrastructure issues end-to-end, like K8s misconfigurations and cloud IAM edge cases, minimizing customer downtime.
Owning incident response for customer deployments with world-class operational excellence.
Codifying what works into repeatable deployment patterns, runbooks, and automation — turning one-off customer work into infrastructure that scales across the next ten deployments.
We would love to meet you if you:
Are your own worst critic—have an extremely high bar for quality and always aim for robust solutions rather than quick fixes.
Have 5+ years of hands-on experience operating production infrastructure, with meaningful time spent deploying software into environments you don't fully control.
Are comfortable with Python or similar languages, and exceptional at working across cloud platforms, container orchestration (e.g., Kubernetes), networking, and storage technologies.
Communicate clearly with external engineers. You can sit in a call with a customer's platform team, diagnose a problem live, and come out with a path forward that both sides trust.
Are energized by being assigned to a customer problem and owning it end-to-end.
Build your own tools on the fly to diagnose, experiment, and address reliability problems—whether it's an internal dashboard or an automated remediation workflow.
Bring a quantitative, hands-on approach to system operations, automation, and continuous improvement.
Bonus points if you:
Have deployed software into regulated environments (financial services, healthcare, legal, insurance) with complex infrastructure.
Have worked with Replicated/KOTS, Helm, or similar enterprise distribution tooling.
Have experience with GPU infrastructure, model serving, or AI/ML workload deployment patterns on customer-controlled hardware.
Have prior experience at an early-stage, high-growth company — as a founder, founding engineer, or early deployment hire.
Are driven, ambitious, and deeply care about both technical excellence and collaborative problem-solving.
This is an in person role at our office in SF. We’re an early stage company which means that the role requires working hard and moving quickly. Please only apply if that excites you.
About Reducto
Nearly 80% of enterprise data is in unstructured formats like PDFs
PDFs are the status quo for enterprise knowledge in nearly every industry. Insurance claims, financial statements, invoices, and health records are all stored in a structure that’s simply impractical for use in digital workflows. This isn’t an inconvenience—it’s a critical bottleneck that leads to .
Traditional approaches fail at reliably extracting information in complex PDFs
OCR and even more sophisticated ML approaches work for simple text documents but are unreliable for anything more complex. Text from different columns are jumbled together, figures are ignored, and tables are a nightmare to get right. Overcoming this usually requires a large engineering effort dedicated to building specialized pipelines for every document type you work with.
breaks document layouts into subsections and then contextually parses each depending on the type of content. This is made possible by a combination of vision models, LLMs, and a suite of heuristics we built over time. Put simply, we can help you:
Accurately extract text and tables even with nonstandard layouts
Automatically convert graphs to tabular data and summarize images in documents
Extract important fields from complex forms with simple, natural language instructions
Build powerful retrieval pipelines using Reducto’s document metadata
Intelligently chunk information using the document’s layout data
Benefits at Reducto
At Reducto, we’re invested in the well-being and growth of our team. Here’s what we currently offer:
Unlimited PTO: We believe great work requires recharging.
Lunch: Receive a free lunch to eat with your teammates daily at the office
Reimbursed Transportation: Provide us with your receipts and we’ll take care of the costs
Insurance: Generous health insurance covering medical, dental, and vision.
Health and Wellness Budget: We provide up to $150/mo reimbursement for health and wellness spending, such as gym memberships, fitness classes, or similar.
Parental Leave: Work with us to build a leave schedule that works for you and your family
Reducto is an Equal Opportunity Employer committed to diversity and inclusion in the workplace. All qualified applicants will receive consideration for employment without regard to sex, race, color, age, national origin, religion, physical and mental disability, genetic information, marital status, sexual orientation, gender identity/assignment, citizenship, pregnancy or maternity, protected veteran status, or any other status prohibited by applicable national, federal, state or local law.
Infrastructure Engineer
About Reducto
Reducto is the agentic document platform for leading AI teams who demand enterprise performance at scale. We provide a comprehensive toolkit for working with documents the way a human would, combining custom in-house and leading frontier models to power efficient and accurate document workflows.
We’ve grown rapidly, increasing revenue 8x year over year and partnering with hundreds of companies, from leading AI teams like Harvey, Vanta, and Scale, to enterprise customers across FAANG and top trading firms.
Reducto has raised over $100M from world-class investors including a16z, Benchmark, and First Round Capital.
The Opportunity
As an Infrastructure Engineer at Reducto, you will influence every aspect of our infrastructure from the ground up. You will architect and scale resilient systems for AI and ML workloads, automate cloud infrastructure, and implement monitoring and incident response practices that set the standard for reliability. This role requires technical leadership, hands-on systems engineering, and strong collaboration with our founders and product teams as we build a company around reliability, rapid iteration, and high-impact product delivery.
The core work will include:
Designing, building, and maintaining highly available, scalable infrastructure to support intensive AI/ML workloads and real-time model deployments.
Implementing robust monitoring, alerting, and observability systems to ensure system health, performance, and uptime across cloud and on-prem environments.
Debugging, optimizing, and automating infrastructure for fast iteration and rapid deployment cycles, focusing on both reliability and developer velocity.
Proactively identifying, investigating, and resolving incidents to minimize downtime and maintain world-class service levels for enterprise customers.
Collaborating closely with engineers, ML specialists, and founders to shape product, infrastructure, and security strategies.
We would love to meet you if you:
Are your own worst critic—have an extremely high bar for quality and always aim for robust solutions rather than quick fixes.
Have 5+ years of hands-on experience in building or supporting production-grade infrastructure and reliability processes for high-throughput systems.
Are comfortable with Python or similar languages, and exceptional at working across cloud platforms, container orchestration (e.g., Kubernetes), networking, and storage technologies.
Build your own tools on the fly to diagnose, experiment, and address reliability problems—whether it's an internal dashboard or an automated remediation workflow.
Bring a quantitative, hands-on approach to system operations, automation, and continuous improvement.
Bonus points if you:
Have prior experience founding a company or building products/infrastructure in early-stage, high-growth environments.
Are excited about automating incident management processes with LLMs/AI.
Are driven, ambitious, and deeply care about both technical excellence and collaborative problem-solving.
Keep up with the latest trends in cloud, observability, and SRE best practices.
Are passionate about open-source and have contributed tools or automation to reliability communities.
Have built or optimized monitoring, incident response, or high-performance computing systems for demanding AI/ML, fintech, or enterprise clients.
This is an in person role at our office in SF. We’re an early stage company which means that the role requires working hard and moving quickly. Please only apply if that excites you.
Benefits at Reducto
Unlimited PTO: We believe great work requires recharging.
Lunch: Receive a free lunch to eat with your teammates daily at the office
Reimbursed Transportation: Provide us with your receipts and we’ll take care of the costs
Insurance: Generous health insurance covering medical, dental, and vision.
Health and Wellness Budget: We provide up to $150/mo reimbursement for health and wellness spending, such as gym memberships, fitness classes, or similar.
Parental Leave: Work with us to build a leave schedule that works for you and your family
Reducto is an Equal Opportunity Employer committed to diversity and inclusion in the workplace. All qualified applicants will receive consideration for employment without regard to sex, race, color, age, national origin, religion, physical and mental disability, genetic information, marital status, sexual orientation, gender identity/assignment, citizenship, pregnancy or maternity, protected veteran status, or any other status prohibited by applicable national, federal, state or local law.
Machine Learning Infra Engineer
Machine Learning Eval Engineer
Agent Experience Engineer
About Reducto
Reducto is the agentic document platform for leading AI teams who demand enterprise performance at scale. We provide a comprehensive toolkit for working with documents the way a human would, combining custom in-house and leading frontier models to power efficient and accurate document workflows.
We’ve grown rapidly, increasing revenue 8x year over year and partnering with hundreds of companies, from leading AI teams like Harvey, Vanta, and Scale, to enterprise customers across FAANG and top trading firms.
Reducto has raised over $100M from world-class investors including a16z, Benchmark, and First Round Capital.
The Role
As Agent Experience Engineer, you'll own how AI agents discover, integrate, and work with Reducto. Today, the best AI products are being built by agents: coding assistants, autonomous workflows, and agentic pipelines that need to parse enterprise documents at scale. Your job is to make Reducto the most natural and frictionless choice when an agent reaches for a document processing tool.
You'll report directly to the CTO and work side-by-side with him daily, shaping strategy, making technical decisions, and shipping fast. Together, you'll build the surfaces that agents and developers interact with: CLIs, MCP servers, SDKs, and integrations with the major platforms and frameworks where agents live, and instrument everything to understand how agents are actually using Reducto in the wild.
The core work will include:
Own Reducto's agent-facing surfaces. Build and maintain our CLI, MCP server, SDKs, and other integration points that make Reducto discoverable and usable by AI agents and the developers building with them.
Build integrations where agents live. Get Reducto into the platforms, frameworks, and toolchains that agents and developers are actually using — from coding assistants to orchestration frameworks to cloud marketplaces. You'll identify the highest-leverage integration opportunities and ship them.
Make our SDKs exceptional. Improve our Python and TypeScript SDKs so they're intuitive for both humans and agents — clear type signatures, predictable error handling, and documentation that reads well in an LLM's context window.
Instrument and measure agent adoption. Define and track metrics around agent usage of Reducto — how agents find us, where they drop off, and what friction points exist. Use this data to prioritize what to build next.
Reduce integration friction end-to-end. Obsess over the experience from first discovery to production usage. Write documentation, build example projects, and design API surfaces that minimize the steps between "I need to parse a document" and a working integration.
Stay on the frontier of agentic tooling. Monitor how agent frameworks (MCP, tool-use protocols, agent SDKs) are evolving and ensure Reducto is a first-class citizen in every major ecosystem.
Drive internal automation. Identify opportunities to use LLMs and agentic workflows to make Reducto's own team more efficient — from internal tooling to workflow automation.
You'll thrive here if you:
Hold yourself to a high bar for quality and precision.
Enjoy solving complex problems and building from first principles.
Have 2+ years of experience in software engineering, with strong Python and TypeScript skills.
Have a genuine feel for what makes developer tools great — you've used enough good and bad APIs to know the difference.
Are deeply fluent in how AI agents work — you've built with agent frameworks, MCP servers, or LLM tool-use patterns and understand what makes a tool easy for an agent to use.
Operate well in fast-changing, high-growth environments.
Take full ownership from strategy through execution.
Bonus points if you:
Have experience at an early-stage or high-growth startup.
Have built or maintained developer SDKs, CLIs, or API tooling.
Are familiar with the AI agent ecosystem — MCP, Claude Code, OpenAI Agents SDK, Sandboxes, or similar frameworks.
Have contributed to open-source developer tools.
Care deeply about combining technical excellence with measurable product impact.
This is an in person role at our office in SF. We're an early stage company which means that the role requires working hard and moving quickly. Please only apply if that excites you.
Benefits at Reducto
At Reducto, we’re invested in the well-being and growth of our team. Here’s what we currently offer:
Unlimited PTO: We believe great work requires recharging.
Lunch: Receive a free lunch to eat with your teammates daily at the office
Reimbursed Transportation: Provide us with your receipts and we’ll take care of the costs
Insurance: Generous health insurance covering medical, dental, and vision.
Health and Wellness Budget: We provide up to $150/mo reimbursement for health and wellness spending, such as gym memberships, fitness classes, or similar.
Parental Leave: Work with us to build a leave schedule that works for you and your family
Reducto is an Equal Opportunity Employer committed to diversity and inclusion in the workplace. All qualified applicants will receive consideration for employment without regard to sex, race, color, age, national origin, religion, physical and mental disability, genetic information, marital status, sexual orientation, gender identity/assignment, citizenship, pregnancy or maternity, protected veteran status, or any other status prohibited by applicable national, federal, state or local law.
Lead Product Engineer
Lead Software Engineer, Platform
Machine Learning Engineer
About Reducto
Reducto is the agentic document platform for leading AI teams who demand enterprise performance at scale. We provide a comprehensive toolkit for working with documents the way a human would, combining custom in-house and leading frontier models to power efficient and accurate document workflows.
We’ve grown rapidly, increasing revenue 8x year over year and partnering with hundreds of companies, from leading AI teams like Harvey, Vanta, and Scale, to enterprise customers across FAANG and top trading firms.
Reducto has raised over $100M from world-class investors including a16z, Benchmark, and First Round Capital.
We would love to meet you if you:
Philosophy: You are your own worst critic. You have a high bar for quality and don’t rest until the job is done right—no settling for 90%. We want someone who ships fast, with high agency, and who doesn't just voice problems but actively jumps in to fix them.
Experience: You have 2+ years of experience with training, fine tuning, and evaluating ML models used in production systems
Language/Skills: You’re exceptional at Python or similar, and are well versed with both traditional computer vision and VLMs
Tools: Build your own tools as needed—like a quick Streamlit app to test hypotheses or create a dataset.
Approach: A quantitative approach to building products. Ability to debug, experiment, and iterate fast. You should be comfortable getting hands-on with the full development lifecycle, from ideation to shipping to users.
The core work will include:
Training and deploying new state of the art models for parsing and interpreting unstructured data
Experimenting with novel techniques to improve LLM accuracy
Build data pipelines, evaluate model performance, and integrate models into the product
Working directly with the founders and customers to shape the product direction and engineering strategy
Bonus points if you:
Have prior experience founding a company or building products at early stages
Are ambitious and driven, and care a lot about doing great work with great people
Keep up with the latest developments in ML/AI
This is an in person role at our office in SF. We’re an early stage company which means that the role requires working hard and moving quickly. Please only apply if that excites you.
About Reducto
Nearly 80% of enterprise data is in unstructured formats like PDFs
PDFs are the status quo for enterprise knowledge in nearly every industry. Insurance claims, financial statements, invoices, and health records are all stored in a structure that’s simply impractical for use in digital workflows. This isn’t an inconvenience—it’s a critical bottleneck that leads to .
Traditional approaches fail at reliably extracting information in complex PDFs
OCR and even more sophisticated ML approaches work for simple text documents but are unreliable for anything more complex. Text from different columns are jumbled together, figures are ignored, and tables are a nightmare to get right. Overcoming this usually requires a large engineering effort dedicated to building specialized pipelines for every document type you work with.
breaks document layouts into subsections and then contextually parses each depending on the type of content. This is made possible by a combination of vision models, LLMs, and a suite of heuristics we built over time. Put simply, we can help you:
Accurately extract text and tables even with nonstandard layouts
Automatically convert graphs to tabular data and summarize images in documents
Extract important fields from complex forms with simple, natural language instructions
Build powerful retrieval pipelines using Reducto’s document metadata
Intelligently chunk information using the document’s layout data
Benefits at Reducto
At Reducto, we’re invested in the well-being and growth of our team. Here’s what we currently offer:
Unlimited PTO: We believe great work requires recharging.
Lunch: Receive a free lunch to eat with your teammates daily at the office
Reimbursed Transportation: Provide us with your receipts and we’ll take care of the costs
Insurance: Generous health insurance covering medical, dental, and vision.
Health and Wellness Budget: We provide up to $150/mo reimbursement for health and wellness spending, such as gym memberships, fitness classes, or similar.
Parental Leave: Work with us to build a leave schedule that works for you and your family
Reducto is an Equal Opportunity Employer committed to diversity and inclusion in the workplace. All qualified applicants will receive consideration for employment without regard to sex, race, color, age, national origin, religion, physical and mental disability, genetic information, marital status, sexual orientation, gender identity/assignment, citizenship, pregnancy or maternity, protected veteran status, or any other status prohibited by applicable national, federal, state or local law.
The full posting opens here — pay, setting and the full description, without leaving the list.