Principal AI Researcher
Backend Engineer
At Inflection AI, our public benefit mission is to harness the power of AI to improve human well-being and productivity. The next era of AI will be defined by agents we trust to act on our behalf. We’re pioneering this future with human-centered AI models that unite emotional intelligence (EQ) and raw intelligence (IQ)—transforming interactions from transactional to relational, to create enduring value for individuals and enterprises alike. Our work comes to life in two ways today: Pi, your personal AI, designed to be a kind and supportive companion that elevates everyday life with practical assistance and perspectives. Platform — large-language models (LLMs) and APIs that enable builders, agents, and enterprises to bring Pi-class emotional intelligence into experiences where empathy and human understanding matter most. We are building toward a future of AI agents that earn trust, deepen understanding, and create aligned, long-term value for all. About the Role As a backend engineer at Inflection, you will own the platforms, systems, and services that bring our conversational AI to life at scale. You’ll collaborate across research, product, and infrastructure teams to enable rapid iteration, high reliability, and secure delivery of novel AI features to millions of users. Your work will directly impact both the pace of product development and the stability of our production systems. This is a good role for you if you: -Have 5+ years of experience building and scaling backend systems for high-throughput applications -Are fluent in building distributed systems with Python, Go, Rust, or similar languages, and are comfortable with cloud-native architectures (e.g., Kubernetes, gRPC, Postgres, Redis, Kafka) -Have owned backend services end-to-end—from design and implementation to deployment, monitoring, and debugging -Thrive in fast-paced environments where you can move quickly without sacrificing engineering rigor -Proactively improve tooling and infrastructure to support your teammates’ workflows and reliability goals -Communicate clearly across disciplines and take pride in solving user-facing problems with clean backend solutions Responsibilities include: -Design and implement scalable backend systems and APIs that power production LLM experiences, including agentic workflows, memory systems, and tool integrations -Build and operate high-availability infrastructure to support real-time inference, retrieval, and conversation pipelines -Develop internal platforms to improve engineering productivity—CI/CD pipelines, service templates, observability frameworks, and rollout tooling -Collaborate closely with applied research and frontend teams to rapidly prototype, ship, and iterate on end-user features -Ensure systems meet our high bar for security, uptime, and latency—through incident response, load testing, monitoring, and automation -Participate in on-call rotations to maintain the reliability of the services you build Employee Pay Disclosures At Inflection AI, we aim to attract and retain the best employees and compensate them in a way that appropriately and fairly values their individual contributions to the company. For this role, Inflection AI estimates a starting annual base salary will fall in the range of approximately $175,000 - $350,000 depending on experience. This estimate can vary based on the factors described above, so the actual starting annual base salary may be above or below this range. Benefits Inflection AI values and supports our team’s mental and physical health. We are focused on building a positive, safe, inclusive and inspiring place to work. Our benefits include: Diverse medical, dental and vision options 401k matching program Unlimited paid time off Parental leave and flexibility for all parents and caregivers Support of country-specific visa needs for international employees living in the Bay Area Interview Process Apply: Please apply on Linkedin or our website for a specific role. After speaking with one of our recruiters, you’ll enter our structured interview process, which includes the following stages: -Hiring Manager Conversation – An initial discussion with the hiring manager to assess fit and alignment. -Technical Interview – A deep dive with an Inflection Engineer to evaluate your technical expertise. -Onsite Interview – A comprehensive assessment, including: A domain-specific interview A system design interview A final conversation with the hiring manager Depending on the role, we may also ask you to complete a take-home exercise or deliver a presentation. For non-technical roles, be prepared for a role-specific interview, such as a portfolio review. Decision Timeline We aim to provide feedback within one week of your final interview.
Job Type
full-time
Experience Level
senior
Salary Range
$170,000 - $350,000
Location
Palo Alto, CA
Data Platform & Annotation Tools
Model Training
At Inflection AI, our public benefit mission is to harness the power of AI to improve human well-being and productivity. The next era of AI will be defined by agents we trust to act on our behalf. We’re pioneering this future with human-centered AI models that unite emotional intelligence (EQ) and raw intelligence (IQ)—transforming interactions from transactional to relational, to create enduring value for individuals and enterprises alike. Our work comes to life in two ways today: Pi, your personal AI, designed to be a kind and supportive companion that elevates everyday life with practical assistance and perspectives. Platform — large-language models (LLMs) and APIs that enable builders, agents, and enterprises to bring Pi-class emotional intelligence into experiences where empathy and human understanding matter most. We are building toward a future of AI agents that earn trust, deepen understanding, and create aligned, long-term value for all. About the Role As a Model Training engineer, you will design, build, and scale the post-training pipelines that turn a general LLM into a brand-fluent, production-ready assistant. Your innovations in fine-tuning and preference optimization (RLHF, DPO, GRPO, RLAIF) will directly improve reliability, alignment, and cost. This is a good role for you if you: -Have hands-on experience training and fine-tuning large transformer models on multi-GPU / multi-node clusters. -Are fluent in PyTorch and its ecosystem tools (Torchtune, FSDP, DeepSpeed) and enjoy digging into distributed-training internals, mixed precision, and memory-efficiency tricks. -Have shipped or published work in RLHF, DPO, GRPO, or RLAIF and understand their practical trade-offs. -Care deeply about training tools, pipelines, and reproducibility—you automate the boring parts so you can iterate on the fun parts. -Balance research curiosity with product pragmatism—you know when to run an ablation and when to ship. -Communicate crisply with both technical and non-technical teammates. Responsibilities include: -Contribute to end-to-end post-training workflows—dataset curation, hyper-parameter search, evaluation, and rollout—using PyTorch, Torchtune, FSDP/DeepSpeed, and our internal orchestration stack. -Prototype and compare alignment techniques (e.g., curriculum RL, multi-objective reward modeling, tool-use fine-tuning) and push the best ideas into production. -Automate training at scale: build robust pipeline components, tools, scripts, and dashboards so experiments are reproducible and easy to trace. -Define the metrics that matter; run A/B tests and iterate quickly to meet aggressive quality targets. -Collaborate with inference, safety, and product teams to land improvements in customer-facing systems. Employee Pay Disclosures At Inflection AI, we aim to attract and retain the best employees and compensate them in a way that appropriately and fairly values their individual contributions to the company. For this role, Inflection AI estimates a starting annual base salary will fall in the range of approximately $175,000 - $350,000 depending on experience. This estimate can vary based on the factors described above, so the actual starting annual base salary may be above or below this range. Interview Process Apply: Please apply on Linkedin or our website for a specific role. After speaking with one of our recruiters, you’ll enter our structured interview process, which includes the following stages: -Hiring Manager Conversation – An initial discussion with the hiring manager to assess fit and alignment. -Technical Interview – A deep dive with an Inflection Engineer to evaluate your technical expertise. -Onsite Interview – A comprehensive assessment, including: A domain-specific interview A system design interview A final conversation with the hiring manager Depending on the role, we may also ask you to complete a take-home exercise or deliver a presentation. For non-technical roles, be prepared for a role-specific interview, such as a portfolio review. Decision Timeline We aim to provide feedback within one week of your final interview.
Job Type
full-time
Experience Level
mid
Salary Range
$170,000 - $300,000
Location
Palo Alto, CA
Frontend Engineer
At Inflection AI, our public benefit mission is to harness the power of AI to improve human well-being and productivity. The next era of AI will be defined by agents we trust to act on our behalf. We’re pioneering this future with human-centered AI models that unite emotional intelligence (EQ) and raw intelligence (IQ)—transforming interactions from transactional to relational, to create enduring value for individuals and enterprises alike. Our work comes to life in two ways today: Pi, your personal AI, designed to be a kind and supportive companion that elevates everyday life with practical assistance and perspectives. Platform — large-language models (LLMs) and APIs that enable builders, agents, and enterprises to bring Pi-class emotional intelligence into experiences where empathy and human understanding matter most. We are building toward a future of AI agents that earn trust, deepen understanding, and create aligned, long-term value for all. About the Role As a frontend engineer at Inflection, you will craft intuitive, responsive, and elegant interfaces like pi.ai that bring our LLM-powered products to life. You’ll work hand-in-hand with product, research, and design teams to build rich user experiences—from rapid prototypes to scalable production apps—that showcase the capabilities of our conversational AI. This is a good role for you if you: -Have 4+ years of experience building and maintaining production-grade web applications -Are fluent with React, TypeScript, and frontend performance debugging tools (e.g., Lighthouse, Chrome DevTools, Web Vitals) -Have built interactive UIs from scratch—ideally with experience handling real-time updates, accessibility, and responsive design -Understand the trade-offs between UX richness, responsiveness, performance, and maintainability -Have designed systems that balance experimentation and rapid iteration with long-term scalability -Are energized by working in a fast-moving, cross-functional environment with deep technical challenges Responsibilities include: -Design and build performant, maintainable frontend systems using TypeScript, React, and modern web tooling -Collaborate with researchers to surface advanced model behaviors—including tool use, RAG, and personalization—through intuitive UX paradigms -Develop and maintain internal UI components and design systems used across multiple product surfaces Integrate real-time data and analytics pipelines to visualize assistant behavior, model performance, and user feedback -Build developer infrastructure to support fast iteration cycles: hot reloading, component testing, E2E automation, and visual diffing -Work across the stack with backend and platform teams to ship end-to-end features with minimal latency and maximum reliability -Contribute to a culture of high code quality, continuous improvement, and shared ownership Employee Pay Disclosures At Inflection AI, we aim to attract and retain the best employees and compensate them in a way that appropriately and fairly values their individual contributions to the company. For this role, Inflection AI estimates a starting annual base salary will fall in the range of approximately $175,000 - $350,000 depending on experience. This estimate can vary based on the factors described above, so the actual starting annual base salary may be above or below this range. Benefits Inflection AI values and supports our team’s mental and physical health. We are focused on building a positive, safe, inclusive and inspiring place to work. Our benefits include: Diverse medical, dental and vision options 401k matching program Unlimited paid time off Parental leave and flexibility for all parents and caregivers Support of country-specific visa needs for international employees living in the Bay Area Interview Process Apply: Please apply on Linkedin or our website for a specific role. After speaking with one of our recruiters, you’ll enter our structured interview process, which includes the following stages: -Hiring Manager Conversation – An initial discussion with the hiring manager to assess fit and alignment. -Technical Interview – A deep dive with an Inflection Engineer to evaluate your technical expertise. -Onsite Interview – A comprehensive assessment, including: A domain-specific interview A system design interview A final conversation with the hiring manager Depending on the role, we may also ask you to complete a take-home exercise or deliver a presentation. For non-technical roles, be prepared for a role-specific interview, such as a portfolio review. Decision Timeline We aim to provide feedback within one week of your final interview.
Job Type
full-time
Experience Level
mid
Salary Range
$170,000 - $300,000
Location
Palo Alto, CA
Platform Engineer (LLM Infrastructure & Backend Systems)
At Inflection AI, our public benefit mission is to harness the power of AI to improve human well-being and productivity. The next era of AI will be defined by agents we trust to act on our behalf. We’re pioneering this future with human-centered AI models that unite emotional intelligence (EQ) and raw intelligence (IQ)—transforming interactions from transactional to relational, to create enduring value for individuals and enterprises alike. Our work comes to life in two ways today: Pi, your personal AI, designed to be a kind and supportive companion that elevates everyday life with practical assistance and perspectives. Platform — large-language models (LLMs) and APIs that enable builders, agents, and enterprises to bring Pi-class emotional intelligence into experiences where empathy and human understanding matter most. We are building toward a future of AI agents that earn trust, deepen understanding, and create aligned, long-term value for all. About the Role We are seeking a Platform Engineer to join our team building backend infrastructure for new ML-powered enterprise products. This role is a unique opportunity to work at the intersection of backend engineering and machine learning systems, focusing on inference orchestration, model integration, and real-time deployment. The ideal candidate will have experience with backend development, production ML systems, and tools that scale enterprise-level applications. This is a good role for you if you: -Backend engineering experience with Python, TypeScript, or Node.js. -Hands-on experience working with production PyTorch models, model checkpoints, and inference logic. -Strong knowledge of building APIs and services that are scalable, stable, and secure. -Passion for bridging backend engineering and ML systems, especially at the infrastructure layer. -Familiarity with tools such as FastAPI, Postgres, Redis, Kubernetes, and React. -Desire to be hands-on and contribute to shaping the foundation of a new enterprise ML product. Responsibilities include: -Build and maintain backend services to support LLM integration, inference orchestration, and data flow. -Write clean, reliable Python code for experimentation, model integration, and production systems. -Collaborate closely with ML researchers to rapidly iterate on product ideas and deploy features. -Design and implement infrastructure to handle scalable inference workloads and enterprise-level use cases. -Own system components and ensure reliability, observability, and maintainability from day one. Employee Pay Disclosures At Inflection AI, we aim to attract and retain the best employees and compensate them in a way that appropriately and fairly values their individual contributions to the company. For this role, Inflection AI estimates a starting annual base salary will fall in the range of approximately $175,000 - $350,000 depending on experience. This estimate can vary based on the factors described above, so the actual starting annual base salary may be above or below this range. Benefits Inflection AI values and supports our team’s mental and physical health. We are focused on building a positive, safe, inclusive and inspiring place to work. Our benefits include: Diverse medical, dental and vision options 401k matching program Unlimited paid time off Parental leave and flexibility for all parents and caregivers Support of country-specific visa needs for international employees living in the Bay Area Interview Process Apply: Please apply on Linkedin or our website for a specific role. After speaking with one of our recruiters, you’ll enter our structured interview process, which includes the following stages: -Hiring Manager Conversation – An initial discussion with the hiring manager to assess fit and alignment. -Technical Interview – A deep dive with an Inflection Engineer to evaluate your technical expertise. -Onsite Interview – A comprehensive assessment, including: A domain-specific interview A system design interview A final conversation with the hiring manager Depending on the role, we may also ask you to complete a take-home exercise or deliver a presentation. For non-technical roles, be prepared for a role-specific interview, such as a portfolio review. Decision Timeline We aim to provide feedback within one week of your final interview.
Job Type
full-time
Experience Level
mid
Salary Range
$170,000 - $300,000
Location
Palo Alto, CA
Forward Deployed AI Engineer
As a Forward Deployed AI Engineer at Charta, you'll play a pivotal role in deploying, customizing, and optimizing cutting-edge generative AI healthcare solutions directly for our clients. About Charta Health In an industry where the focus should rightly be on delivering quality care to patients, healthcare providers remain burdened by the complexities of non-clinical operations. Charta is changing that. We're building the operating system for modern healthcare organizations. Our AI platform streamlines critical workflows across revenue cycle, clinical operations, and administrative functions, helping providers and payers operate more efficiently and deliver better patient care. Backed by Bain Capital Ventures, Charta is on a mission to make every healthcare dollar accountable and every chart accurate, reimagining healthcare infrastructure from the ground up. About the Opportunity As a Forward Deployed AI Engineer at Charta Health, you will be at the forefront of deploying and customizing AI-powered solutions for our clients. You will work directly with healthcare providers to understand their unique challenges and translate them into practical, data-driven solutions using AI and machine learning. This role requires a blend of technical expertise, problem-solving skills, and a strong customer-centric mindset. You will thrive in a dynamic environment, rapidly deploying and iterating on solutions to deliver immediate value. What you’ll do: - Collaborate closely with healthcare clients to understand their workflows and identify opportunities to leverage AI and machine learning. - Customize and deploy Charta Health's AI-driven platforms to address specific client needs. - Rapidly prototype and iterate on solutions, adapting to client feedback and evolving requirements. - Break down complex projects into actionable steps and deliver results in ambiguous environments. - Explore and integrate a wide range of AI/ML technologies, adapting to the diverse needs of our clients. - Provide technical expertise and support to clients, ensuring successful deployment and adoption of our solutions. - Work closely with our internal engineering team to provide feedback and contribute to product development. - Stay up-to-date with the latest advancements in AI/ML and their applications in healthcare.
- Proficiency and practical coding skills, with a focus on delivering tangible results. - Adaptability and the ability to quickly learn and master new technologies and concepts. - A priority on delivering business value over solving purely theoretical challenges. - A desire to explore a wide range of technologies and problems, rather than specializing in one area. - The ability to thrive in ambiguous environments and rapidly produce results. - The ability to excel at breaking down complex projects into actionable steps. - Resilience and resourcefulness, allowing you to overcome obstacles and find creative solutions. - Ambition and a desire for a high-growth opportunity. - Strong emotional intelligence, excellent communication, and interpersonal skills. - Excitement about working with external customers and building strong relationships. - A genuine interest in the medical field and a passion for improving healthcare. Nice to have: - Experience and/or demonstrated interest in healthcare - Proficiency in Python. - Experience with Large Language Models (LLMs) and prompt engineering. - Experience with neural networks and machine learning
Job Type
full-time
Experience Level
mid
Salary Range
$90,000 - $140,000
Location
San Francisco, CA
The full posting opens here — pay, setting and the full description, without leaving the list.