Page 1
Loading more openings…
You've reached the end of the list.
Staff Data Center Implementation Manager
Lambda · United States
Pay
$191k–255k
Setting
Remote
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.
If you'd like to build the world's best AI cloud, join us.
Travel: 50% , Travel required to various data center sites.
What You’ll Do
As a Data Center Implementation Manager, you will lead the end-to-end execution of multiple concurrent data center infrastructure projects, spanning Greenfield builds, colocation white space fit-outs, equipment upgrades, and lifecycle replacements. You’ll serve as the single point of contact for all construction activities on your assigned projects, coordinating closely with data center providers, MEP contractors, and internal stakeholders.
Your role includes overseeing construction schedules and budgets, managing OFCI equipment procurement and installation, conducting design reviews, and driving alignment across cross-functional teams including electrical, mechanical, low voltage, and network engineering. You’ll also contribute to continuous improvement initiatives, integrating emerging technologies and design innovations into our infrastructure strategy.
Key Responsibilities
- Project Leadership & Oversight:
- Manage the build out of multiple data center projects concurrently, from detailed design to physical build out, including Greenfield construction projects, colocation white space fitouts, equipment upgrades, and end-of-life replacements. Oversee the work of MEP Contractors and Equipment Suppliers during the build out process.
- Internal Coordination with SMEs:
- Coordinate with internal Subject Matter Experts (SMEs) across various disciplines, including electrical, mechanical, and low voltage, to ensure project requirements, technical specifications, and solutions are properly aligned.
- Budget & Schedule Coordination:
- Work with our Data Center Providers and Contractors to prepare, review, and monitor project schedules, budgets, and forecasts for design and construction expenditures.
- OFCI Equipment Management:
- Review and evaluate OFCI (Owner-Furnished, Contractor-Installed) equipment submittals, providing recommendations to the Construction Management team and Procurement teams.
- Design Reviews:
- Participate in page-turn meetings with consultants, providing feedback on designs, Sequence of Operations, Basis of Design documents, and commissioning plans.
- Vendor Management:
- Assist with RFPs for OFCI Equipment and Data Center Whitespace Fit Outs, evaluation of proposals, selection of vendors, and oversight on alignment with project budgets and schedules.
- Project Reporting:
- Prepare and deliver detailed project reports to both executive leadership and project teams, track project budgets, schedules, scope objectives, and progress.
- Stakeholder Communication:
- Provide regular project updates to internal and external stakeholders, both informally and formally.
- Design Coordination & Site Visits:
- Conduct site visits during construction to review design implementation, assess construction progress, and ensure adherence to design specifications, quality standards, and project requirements.
- Cross-functional Collaboration:
- Collaborate closely with cross-functional teams, including Network Engineering, Data Center Operations, HPC, Supply Chain and Technical Project Management to ensure alignment with business objectives and project deliverables. Serve as a liaison between internal teams and external contractors or vendors, ensuring clear and effective communication.
- Transformation & Design Improvements:
- Contribute to internal design team meetings, proposing and exploring new technologies or design concepts to enhance project outcomes. Stay up to date with industry trends and integrate innovative solutions into project designs.
Ideal Candidate Profile
- Proven experience managing data center construction projects, including Greenfield builds, colocation fit-outs, and infrastructure upgrades.
- Strong understanding of MEP systems and construction practices, with hands-on experience coordinating with electrical, mechanical, and low voltage SMEs.
- Skilled in managing OFCI equipment procurement, submittal review, and installation oversight.
- Effective communicator with the ability to serve as the single point of contact for construction activities and maintain alignment across internal and external stakeholders.
- Proficient in tracking and managing project schedules, budgets, and forecasts in collaboration with contractors and data center providers.
- Experienced in participating in design reviews and commissioning processes, with the ability to provide meaningful technical feedback.
- Comfortable conducting on-site inspections to verify construction quality, design adherence, and progress.
- Familiar with vendor RFP processes, proposal evaluation, and managing vendor performance to budget and timeline expectations.
- Highly organized, with a strong ability to prepare detailed project reports and communicate updates to both executive leadership and project teams.
- Collaborative team player with cross-functional experience working alongside Network Engineering, Data Center Operations, HPC, Supply Chain, and TPM organizations.
- Passionate about continuous improvement and staying current with industry trends, technologies, and best practices in data center design and delivery.
Top Five Requirements
1. End-to-End Data Center Construction Experience:
- Proven track record managing full lifecycle data center projects, colocation fit-outs, and infrastructure upgrades.
2. Strong MEP & OFCI Coordination Skills:
- Deep understanding of mechanical, electrical, and low voltage systems, with experience reviewing OFCI submittals and coordinating with design and construction teams.
3. Project Budget & Schedule Management:
- Demonstrated ability to manage construction budgets, track project schedules, and work with vendors and providers to meet critical deadlines and cost targets.
4. Cross-Functional Communication & Stakeholder Management:
- Excellent interpersonal skills, capable of serving as the single point of contact for construction and aligning internal stakeholders, external contractors, and data center providers.
5. On-Site Construction Oversight & Quality Assurance:
- Ability to perform regular site visits to monitor construction progress, validate design implementation, and ensure adherence to project specifications and quality standards.
Qualifications
- Bachelor’s degree in Engineering, Construction Management, or a related field (or equivalent practical experience).
- 10+ years of experience managing data center construction or critical infrastructure projects.
- Strong knowledge of MEP systems, including coordination with electrical, mechanical, and low voltage trades.
- Proficient in reviewing technical design documents, including Basis of Design, Sequence of Operations, and commissioning plans.
- Demonstrated experience managing OFCI equipment procurement, submittals, and installation processes.
- Strong documentation and writing skills, with the ability to create clear, concise project documents, decision-making documents, and instructional materials.
- Proven ability to work in a fast-paced environment and manage multiple priorities effectively.
- Strong analytical skills, with the ability to make data-driven decisions and solve complex problems.
- Skilled in construction schedule and budget management, including forecasting, tracking, and reporting.
- Familiarity with project management tools (e.g., MS Project, Primavera, Smartsheet) and construction tracking/reporting platforms.
- Excellent communication and interpersonal skills, with the ability to collaborate across teams and effectively engage with stakeholders.
- Passion for continuous improvement, with a proactive mindset for integrating new technologies and design innovations.
Salary Range Information
The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description.
About Lambda
- Founded in 2012, with 500+ employees, and growing fast
- Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove
- We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG
- Our values are publicly available: https://lambda.ai/careers
- We offer generous cash & equity compensation
- Health, dental, and vision coverage for you and your dependents
- Wellness and commuter stipends for select roles
- 401k Plan with 2% company match (USA employees)
- Flexible paid time off plan that we all actually use
Equal Opportunity Employer
Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.
Listed by Lambda for a position based in the United States. Employers on this board attest they are hiring domestically.
Data Center Construction Site Foreman (Dallas)
Lambda · United States
Pay
$132k–176k
Setting
Remote
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.
If you'd like to build the world's best AI cloud, join us.
Travel: 75% , Sustained on-site presence required at active data center construction and fit-out sites.
What You’ll Do
As a Construction Onsite Foreman, you are Lambda’s eyes, ears, and hands in the field. You will hold daily on-site command of colocation white space fit-outs, high-density compute expansions, and infrastructure upgrades, working under the Data Center Implementation Manager and alongside the Data Center Implementation Project Manager to turn drawings and schedules into installed, energized, and accepted capacity.
You will direct and coordinate trade crews and vendor teams on the floor, sequence work across electrical, mechanical, low voltage, and network scopes, enforce safety and quality standards, and keep material and equipment flowing to the point of install. You will catch field conflicts before they become schedule impacts, escalate what cannot be solved at the site level, and own the daily record of what actually happened on your job.
Key Responsibilities
- Daily Field Supervision:
- Serve as Lambda’s on-site lead for assigned data center projects. Direct daily activity across MEP contractors, low voltage and network installers, rack and integration crews, and OEM field service teams. Run daily huddles, set the day’s priorities, and confirm each crew knows its scope, sequence, and constraints.
- Fit-Out & Installation Oversight:
- Oversee physical execution of white space fit-out work including rack and cabinet setting, containment, busway and branch circuit distribution, rPDU installation, cable tray, structured cabling, liquid cooling and CDU tie-ins, and airflow management. Verify installations match approved drawings and Lambda standards before work is covered or energized.
- Safety Leadership:
- Enforce site safety requirements, energized work permits, LOTO procedures, hot work permits, and PPE compliance. Conduct daily safety walks, stop unsafe work immediately, and lead incident and near-miss reporting with the data center provider and general contractor.
- Quality Control & Inspections:
- Perform in-progress and pre-cover inspections, torque and terminations checks, labeling verification, and workmanship reviews. Build and drive punch lists to closure, rejecting non-conforming work and confirming corrections in the field.
- Trade & Vendor Coordination:
- Deconflict overlapping trade work, sequence access to shared space and shafts, and coordinate outage and switching windows with the data center provider. Escort and manage vendor access, badging, tool and material handling, and site rules compliance.
- OFCI Material Management:
- Receive, inspect, stage, and secure OFCI (Owner-Furnished, Contractor-Installed) equipment. Verify shipments against packing lists and submittals, document damage and shortages immediately, manage laydown and storage areas, and confirm material readiness ahead of each install window.
- Schedule Support & Lookaheads:
- Own the three-week lookahead in the field. Track daily progress against the construction schedule, flag manpower and material gaps early, and provide the Implementation Project Manager with accurate percent-complete and constraint data.
- Field Documentation & Reporting:
- Produce daily field reports covering manpower counts, work completed, deliveries, delays, safety observations, and photo documentation. Maintain redline markups, RFI and field observation logs, and as-built notes for turnover.
- Issue Resolution & Escalation:
- Resolve constructability conflicts and field questions at the site level where possible. Generate clear RFIs with photos and drawing references when a design or scope decision is needed, and escalate cost, schedule, or safety impacts without delay.
- Commissioning & Turnover Support:
- Support pre-functional checks, Level 1 through Level 4 commissioning activities, and integrated systems testing. Coordinate contractor availability for scripts, witness testing where directed, and drive open-item closure through final acceptance and handoff to Data Center Operations.
- Cross-functional Alignment:
- Work closely with Network Engineering, Data Center Operations, HPC, Supply Chain, and Technical Project Management so that deployment, cabling, and hardware install activities land in the right sequence and space is turned over ready for GPU deployment.
Ideal Candidate Profile
- Comes from the trades, with a working foreman’s command of how mission-critical space actually gets built, installed, and energized.
- Experienced running crews and vendors in live and occupied data center environments where uptime and adjacency risk are unforgiving.
- Fluent in reading and marking up construction drawings, single lines, riser diagrams, submittals, and installation details.
- Sets a visible safety standard and is comfortable stopping work and holding contractors to it, including under schedule pressure.
- Detail-driven on workmanship, labeling, torque, containment, and cable management, with a low tolerance for shortcuts that create operational debt.
- Organized in the field with materials, tools, laydown, and deliveries, and disciplined about documenting what happened each day.
- Direct and clear communicator who can brief a project manager, challenge a contractor, and write a usable daily report.
- Comfortable working with limited oversight, making sound calls on site, and knowing precisely when to escalate.
- Proficient with mobile field tools, photo documentation, and punch list and project tracking platforms.
- Willing to travel to project sites and to work off-hours, nights, or weekends when critical activities or outage windows require it.
Top Five Requirements
1. Hands-On Data Center or Mission-Critical Field Experience:
- Proven record supervising construction, fit-out, or infrastructure installation work in data centers or comparable mission-critical facilities.
2. Crew & Trade Supervision:
- Demonstrated ability to direct and sequence multiple trades and vendors on a live site, holding daily accountability for scope, pace, and workmanship.
3. MEP & Low Voltage Installation Knowledge:
- Practical command of electrical distribution, mechanical and cooling systems, containment, and structured cabling, with the ability to read drawings and catch field conflicts.
4. Safety Ownership:
- Consistent enforcement of site safety, LOTO, energized work, and permitting requirements, with the authority and willingness to stop unsafe work.
5. Field Documentation & Reporting Discipline:
- Reliable daily reporting, photo documentation, redlines, punch list management, and clear escalation to project management.
Qualifications
- High school diploma or equivalent required. Completed trade apprenticeship, journeyman license, technical certification, or associate degree in a construction or engineering discipline strongly preferred.
- 5+ years of construction field experience, including 2+ years in a foreman, lead, or superintendent capacity.
- Direct experience with data center white space fit-out, critical facility construction, or comparable industrial and mission-critical installation work.
- Working knowledge of electrical, mechanical, and low voltage systems and installation practices, including power distribution, cooling, containment, and cabling.
- Ability to read and interpret construction drawings, specifications, submittals, and installation instructions, and to produce accurate redlines.
- OSHA 30 certification required or obtained within 90 days of hire. NFPA 70E familiarity preferred. First Aid and CPR certification a plus.
- Familiarity with commissioning processes, pre-functional checklists, and punch list closeout.
- Comfortable with field reporting tools and project platforms (e.g., Procore, Smartsheet, MS Project, Monday.com http://Monday.com) and with mobile photo documentation.
- Able to lift, climb, work at height, and remain on foot in active construction environments for extended periods.
- Willingness to travel to and work within active construction sites (~75%), including off-hours and weekend work during critical activities.
- Committed to safety, quality, and continuous improvement in field execution practices.
Salary Range Information
The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description.
About Lambda
- Founded in 2012, with 500+ employees, and growing fast
- Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove
- We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG
- Our values are publicly available: https://lambda.ai/careers
- We offer generous cash & equity compensation
- Health, dental, and vision coverage for you and your dependents
- Wellness and commuter stipends for select roles
- 401k Plan with 2% company match (USA employees)
- Flexible paid time off plan that we all actually use
Equal Opportunity Employer
Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.
Listed by Lambda for a position based in the United States. Employers on this board attest they are hiring domestically.
Energy Origination & Hedging Manager
Lambda · United States
Pay
$233k–365k
Setting
Remote
Energy Origination & Hedging Manager
Listed by Lambda for a position based in the United States. Employers on this board attest they are hiring domestically.
HPC Support Engineer
Lambda · United States
Pay
$122k–162k
Setting
Remote
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.
If you'd like to build the world's best AI cloud, join us.
This position is expected to participate in an on-call rotation.
What You’ll Do
- Serve as a senior technical escalation point, troubleshooting the hardest infrastructure and platform issues down to the hardware, driver, or kernel level when needed
- Quickly and accurately distinguish between hardware failures, driver issues, kernel-level problems, and customer workload misconfiguration, so issues get resolved correctly the first time
- Proactively identify process, tooling, and documentation gaps, and go fix them, not just wait for them to be assigned
- Use AI tools effectively to build scripts, automations, or small internal tools that close real operational gaps (no professional development background required)
- Perform root-cause analysis across distributed systems, clusters, and GPU infrastructure
- Craft clear documentation of solutions and contribute to evolving support procedures
- Collaborate closely with engineering teams to turn recurring customer pain points into permanent fixes
- Take escalations from peers while training and mentoring them in the process
- Participate in a rotating on-call schedule, owning major incidents and major customer issues
- Be ready to roll up your sleeves and pitch in wherever needed, especially during fast, high-volume deployments
You
- 3+ years of hands-on HPC experience in an administration, support, or engineering role.
- Very strong understanding and experience supporting Linux in a system administration role.
- Proven experience in HPC environments, showcasing your expertise in Linux cluster administration, with strong preference for Kubernetes and/or Slurm for cluster orchestration.
- Strong coding ability and CI/CD experience, with a track record of using AI-assisted tools to move fast.
- Proficiency with monitoring/logging tools (Prometheus, Grafana, Datadog).
- Strong skills in log analysis, debugging kernel-level issues, and performance profiling.
- Experience with CUDA, NCCL, NVLink, GPUDirect RDMA.
- Experience with high throughput networking technologies(IB/RoCE).
- Knowledge of distributed AI/ML or HPC workloads.
- Knowledge of TCP/IP, VPN, and firewalls in cloud environments.
- Ability to work independently and mentor junior support engineers.
Nice to Have
- Experience with virtualization and container (Docker, Kubernetes) technologies.
- Experience with neoclouds/GPU cloud providers.
- Flexible availability for potential shifts outside of normal working hours/weekends.
- Experience with high performance storage systems.
- Familiarity with infrastructure-as-code tools (Terraform, Ansible, etc.)
- Experience with Nvidia GPUs and Infiniband.
Salary Range Information
This is a salaried exempt role. The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description.
About Lambda
- Founded in 2012, with 500+ employees, and growing fast
- Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove
- We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG
- Our values are publicly available: https://lambda.ai/careers
- We offer generous cash & equity compensation
- Health, dental, and vision coverage for you and your dependents
- Wellness and commuter stipends for select roles
- 401k Plan with 2% company match (USA employees)
- Flexible paid time off plan that we all actually use
Equal Opportunity Employer
Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.
Listed by Lambda for a position based in the United States. Employers on this board attest they are hiring domestically.
Senior Incident Manager
Lambda · United States
Pay
$125k–195k
Setting
Remote
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.
If you'd like to build the world's best AI cloud, join us.
We are seeking a Senior Incident Manager to lead critical incident response across our AI data center infrastructure. This role is responsible for coordinating rapid resolution of service-impacting events, improving operational resilience, and driving incident management best practices across infrastructure, networking, platform engineering, and data center operations.
Role Overview
The Senior Incident Manager is responsible for leading the end-to-end lifecycle of operational incidents impacting AI infrastructure and data center services. This individual acts as the central command point during major incidents, ensuring rapid triage, cross-team coordination, effective communication, and structured post-incident analysis.
This role requires deep operational expertise in high-availability infrastructure, large-scale GPU clusters, networking, and cloud platforms, along with strong leadership and communication skills.
What You’ll Do
Incident Leadership
- Lead the response to critical (SEV-1 / SEV-2) incidents impacting AI infrastructure, GPU clusters, networking, storage, and data center operations.
- Serve as the Incident Commander during major outages, coordinating engineering, networking, facilities, and vendor teams.
- Act as the liaison between leadership and external teams during incidents / post-incidents to provide updates and status summaries.
- Establish clear incident timelines, triage actions, and resolution plans.
Incident Management Operations
- Own the incident response lifecycle including:
- Assisting Technical Triage
- Escalation
- Coordination
- Resolution
Post-incident review
- Ensure timely and accurate communication with internal stakeholders and leadership.
- Maintain incident response documentation and operational playbooks.
- Conduct analysis on incidents and identify patterns / trends for improvement in response and systems reliability.
- Work in an On-Call Rotation to respond to, lead, and coordinate incidents
Cross-Functional Coordination
- Work closely with:
- Data center operations
- Infrastructure engineering & operations
- Network engineering
- Platform reliability engineering
- Security operations
- Hardware and facility vendors
- Drive alignment during outages involving multiple infrastructure layers.
Post-Incident Analysis & Continuous Improvement
- Lead post-incident reviews (PIRs) and root cause analysis. Identify systemic reliability gaps and implement corrective actions.
- Track incident metrics including MTTR, MTTD, and incident recurrence rates.
Operational Excellence
- Improve incident response processes, escalation paths, and tooling by working with technical support and engineering teams..
- Contribute to runbooks, operational standards, and reliability frameworks.
- Support implementation of automation and observability improvements.
Communication & Reporting
- Provide executive-level incident summaries and reports.
- Deliver clear, concise updates during active incidents.
- Maintain incident dashboards and operational health reporting.
You
- 8+ years experience in incident management, site reliability engineering, or infrastructure operations
- Experience managing incidents in large-scale distributed infrastructure environments
- Strong understanding of:
- Data center operations
- GPU compute clusters
Networking and storage infrastructure
- Cloud or hybrid infrastructure platforms
- Proven ability to lead high-pressure incident response situations
- Experience with incident management frameworks (ITIL, SRE, or equivalent)
- Excellent communication and stakeholder management skills
- Experience with incident tracking and monitoring tools such as:
- PagerDuty
- ServiceNow
- Jira
- Datadog
- Prometheus / Grafana
Nice to Have
- Experience operating AI or HPC infrastructure
- Background in SRE, infrastructure engineering, or data center operations
- Familiarity with high-density GPU environments (NVIDIA clusters, InfiniBand networks)
- Experience with hyperscale or colocation data center environments
- Knowledge of automation and incident response tooling
- Knowledge of and experience with Incident command system (ICS)
- Experience in leading and developing incident command from stractch
Key Competencies
- Incident Command & Leadership
- Operational Decision Making
- Cross-Team Coordination
- Root Cause Analysis
- Crisis Communication
- Infrastructure Reliability
What Success Looks Like in This Role
- Reduced Mean Time to Resolution (MTTR) for critical incidents
- Improved cross-team incident coordination
- High-quality post-incident reviews and corrective actions
- Increased infrastructure reliability and operational maturity
Salary Range Information
The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description.
About Lambda
- Founded in 2012, with 500+ employees, and growing fast
- Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove
- We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG
- Our values are publicly available: https://lambda.ai/careers
- We offer generous cash & equity compensation
- Health, dental, and vision coverage for you and your dependents
- Wellness and commuter stipends for select roles
- 401k Plan with 2% company match (USA employees)
- Flexible paid time off plan that we all actually use
Equal Opportunity Employer
Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.
Listed by Lambda for a position based in the United States. Employers on this board attest they are hiring domestically.
Account CTO
Lambda · United States
Pay
$271k–425k
Setting
Remote
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.
If you'd like to build the world's best AI cloud, join us.
Engineering at Lambda is responsible for building and scaling our cloud offering. Our scope includes the Lambda website, cloud APIs and systems as well as internal tooling for system deployment, management and maintenance.
What You’ll Do
- Strategic Technology Leadership for Key Accounts
- Serve as the primary technical liaison for Lambda's strategic Superintelligence accounts
- Cultivate and sustain executive level technical relationships, establishing Lambda as a trusted technology partner and strategic advisor
- Drive multi-year technology roadmaps aligned with customer business objectives and Lambda's product portfolio
- Influence customer technology strategy and adoption at the executive level
- Represent Lambda at executive briefings, board meetings, and strategic planning sessions
- Own Technical Strategy and Business Outcomes
- Partner with Global Account teams to drive strategic account growth and expansion
- Lead architectural reviews and technology assessments at the enterprise level
- Define and articulate business value propositions, ROI models, and TCO analyses for executive stakeholders
- Orchestrate cross-functional teams including Solutions Engineers, Forward Deployed Engineers, Professional Services, and Product Management
- Own technical escalations and serve as the Executive Technical Sponsor for critical customer initiatives
- Drive adoption of Lambda's strategic initiatives within key accounts
- Technology Vision and Thought Leadership
- Establish yourself as a trusted advisor on AI/ML transformation and cloud strategy
- Guide customers through digital transformation journeys and emerging technology adoption
- Publish thought leadership content, speak at industry events, and represent Lambda in executive forums
- Influence Lambda's product strategy based on strategic customer needs and market trends
- Mentor and develop technical talent across the organization
You
- Have 15+ years of experience designing, deploying and scaling cloud infrastructure
- Have 6+ years of experience working directly with C-suite executives and senior leadership teams in advisory or consultative capacities
- Have 5+ years driving cloud transformation initiatives at enterprise scale
- Have proven track record of influencing C-suite executives and board-level stakeholders
- Have deep expertise in AI/ML technologies and their business applications
- Have experience with NVIDIA DGX/HGX/MGX systems, InfiniBand &RoCE networking, and large-scale GPU cluster deployments
- Have experience with storage architectures including distributed file systems, object storage, high-performance parallel file systems, and NVMe-based solutions
- Have hands-on experience with LLM training, fine-tuning, and inference at scale
- Have experience with enterprise architecture frameworks and governance models
- Have experience managing or influencing deals worth $1B+ annually
- Have strong business acumen with ability to translate technology into business outcomes
- Have experience navigating complex organizational dynamics and driving consensus
- Have track record of building strategic partnerships and alliances
- Excel at executive communication, presentation, and storytelling
- Demonstrate thought leadership through publications, speaking engagements, or industry recognition
Nice to Have
- MBA or advanced degree in Computer Science, Engineering, or related technical field
- Experience as a CTO, VP of Engineering, Chief Architect, or Technical Fellow at a technology company with $100M+ revenue
- Track record working with AI labs, research institutions, or frontier model companies
- Board advisory experience for technology companies
- Published peer-reviewed research, whitepapers, or patents specifically in distributed computing, GPU acceleration, or large-scale ML training infrastructure
Salary Range Information
The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description.
About Lambda
- Founded in 2012, with 500+ employees, and growing fast
- Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove
- We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG
- Our values are publicly available: https://lambda.ai/careers
- We offer generous cash & equity compensation
- Health, dental, and vision coverage for you and your dependents
- Wellness and commuter stipends for select roles
- 401k Plan with 2% company match (USA employees)
- Flexible paid time off plan that we all actually use
Equal Opportunity Employer
Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.
Listed by Lambda for a position based in the United States. Employers on this board attest they are hiring domestically.
Select a role
The full posting opens here — pay, setting and the full description, without leaving the list.