Production Systems Engineer, Fleet AI Systems
4 weeks ago
- Interface with external vendors and internal hardware, mechanical, power, thermal, manufacturing and software engineers to understand system architecture to develop and execute the test suites for various architectures
- Proactively create experiments and tooling to detect and diagnose hardware/firmware/software health issues
- Develop test framework for large-scale test automation inside fleet during product development and after mass production
- Implement remediations across software and hardware stack according to plan, while keeping a thorough procedural record and data log
- Develop and publish updates on resolutions and communicate findings internally
- Troubleshoot, diagnose and root cause of system failures and isolate the components/failure scenarios while working with internal & external stakeholders
- Develop visibility through data visualization and implement systemic solutions to hardware health issues
- Drive necessary discussion with external and internal teams on test specification and methodologies to improve test quality continuously
- Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience.
- 2+ years of support in hardware system support
- Troubleshooting and analytical experience
- Knowledge of server architecture and components
- Experience with Linux and scripting
- Experience in changing system configurations and measuring change impact
- Experience working in a matrix organization
- Experience working through full life cycle for computer system products
- Experience supporting AI/HPC systems and/or related components at scale
- Engineering for different server system/data center products
- 4+ years experience in Production support at scale (e.g. - 10K storage servers and over 100K HDD)
- 4+ years experience in full system technologies
- Experience in post-production hyperscale post-production environments, solutions
Individual compensation is determined by skills, qualifications, experience, and location. Compensation details listed in this posting reflect the base hourly rate, monthly rate, or annual salary only, and do not include bonus, equity or sales incentives, if applicable. In addition to base compensation, Meta offers benefits. Learn more about benefits at Meta.
-
Production Systems Engineer
2 weeks ago
Menlo Park, California, United States META Full timeProduction Systems Engineer - Fleet AI SystemsMeta is seeking a highly skilled Production Systems Engineer to join our Release to Production (RTP) team. Our servers and data centers are the foundation upon which our rapidly scaling infrastructure operates efficiently to deliver our innovative services.ResponsibilitiesInterface with external vendors and...
-
Production Systems Engineer
2 weeks ago
Menlo Park, California, United States META Full timeProduction Systems Engineer, Fleet AI SystemsMeta is seeking a highly skilled Production Systems Engineer to join our Release to Production (RTP) team. As a key member of our team, you will be responsible for the Hardware Lifecycle of all Meta servers, including pre-production hands-on system and hardware debugging and stress testing, enabling...
-
Production Systems Engineer
1 month ago
Menlo Park, California, United States META Full timeJob Title: Production Systems Engineer - Fleet AI SystemsMeta is seeking a highly skilled Production Systems Engineer to join our Release to Production (RTP) team. Our servers and data centers are the foundation upon which our rapidly scaling infrastructure operates efficiently to deliver our innovative services.Responsibilities:Interface with external...
-
Production Systems Engineer
2 weeks ago
Menlo Park, California, United States META Full timeJob Title: Production Systems Engineer, Fleet AI SystemsMeta is seeking a highly skilled Production Systems Engineer to join our Release to Production (RTP) team. Our servers and data centers are the foundation upon which our rapidly scaling infrastructure operates efficiently to deliver our innovative services.Responsibilities:Interface with external...
-
Production Systems Engineer, Fleet AI Systems
2 weeks ago
Menlo Park, United States META Full timeProduction Systems Engineer, Fleet AI Systems (NetZero) Apply to this job Location pin icon Menlo Park, CA Apply to this job Meta is seeking a Production Systems Engineer to join our Release to Production (RTP) team. Our servers and data centers are the foundation upon which our rapidly scaling infrastructure operates efficiently to deliver our...
-
Production Systems Engineer, Fleet AI Systems
4 weeks ago
Menlo Park, United States META Full timeProduction Systems Engineer, Fleet AI Systems (NetZero) Apply to this job Location pin icon Menlo Park, CA Apply to this job Meta is seeking a Production Systems Engineer to join our Release to Production (RTP) team. Our servers and data centers are the foundation upon which our rapidly scaling infrastructure operates efficiently to deliver our...
-
Production Systems Engineer, Fleet AI Systems
4 weeks ago
Menlo Park, United States META Full timeMeta is seeking a Production Systems Engineer to join our Release to Production (RTP) team. Our servers and data centers are the foundation upon which our rapidly scaling infrastructure operates efficiently to deliver our innovative services. The RTP team is responsible for the Hardware Lifecycle of all Meta servers including pre-production hands-on system...
-
AI/HPC Systems Performance Engineer
4 weeks ago
Menlo Park, California, United States META Full timeJob SummaryMeta's AI Training and Inference Infrastructure is growing exponentially to support ever-increasing use cases of AI. This results in a dramatic scaling challenge that our engineers have to deal with on a daily basis. We need to build and evolve our network infrastructure that connects myriads of training accelerators like GPUs together.Key...
-
AI/HPC Systems Performance Engineer
18 hours ago
Menlo Park, California, United States META Full timeJob Summary:Meta is seeking a highly skilled AI/HPC Systems Performance Engineer to join our team. As a key member of our infrastructure team, you will be responsible for designing, deploying, and operating high-performance networks to support our rapidly growing AI workloads.This is an exciting opportunity to work on cutting-edge technologies and contribute...
-
Senior Software Engineer
1 month ago
Tinley Park, Illinois, United States HNM Systems Full timeJob Title: Senior Software Engineer - Generative AIHNM Systems is a leading provider of Communication and Information Technology staffing and consulting services. We are currently seeking a highly skilled Senior Software Engineer to join our team and contribute to the development of our Generative AI solutions.Job Summary:We are looking for a talented Senior...
-
Hardware Systems Engineer, RAS
2 weeks ago
Menlo Park, California, United States META Full timeMeta Hardware Systems EngineerMeta is seeking a skilled Hardware Systems Engineer to join our Release to Production (RTP) team. As a key member of this team, you will be responsible for the end-to-end Hardware Lifecycle of all Meta servers, including prototyping of experimental HW, pre-production hands-on system and hardware debugging and stress testing,...
-
Staff Software Engineer
1 week ago
Menlo Park, California, United States Cyngn Full timeAbout CyngnCyngn is a leading autonomous vehicle company based in Menlo Park, CA. We're a collaborative and diverse team that's passionate about innovation and continuous learning.Our self-driving technology can be deployed in various commercial domains across different vehicle form factors. We're seeking experienced leaders to join our team and help move...
-
Hardware Systems Engineer
3 weeks ago
Menlo Park, California, United States META Full timeJob SummaryMeta is seeking a highly skilled Hardware Systems Engineer to join our Release to Production (RTP) team. As a key member of this team, you will be responsible for the end-to-end Hardware Lifecycle of all Meta servers, including prototyping, debugging, and stress testing.The RTP team is responsible for ensuring the efficient operation of our...
-
Systems Engineer
4 weeks ago
Lexington Park, United States BAE Systems Full timeJob Description:BAE Systems is seeking an experienced Senior Engineer to implement and manage the VH92-A Mission Communications System Digital Ecosystem strategy. The selected candidate will manage a portfolio of Digital Ecosystem System Engineering (DESE), Data Architecture and Analytics, Product Lifecycle Management (PLM) Capability Development, Software...
-
Systems Engineer
4 weeks ago
Lexington Park, United States BAE Systems Full timeJob Description:BAE Systems is seeking an experienced Senior Engineer to implement and manage the VH92-A Mission Communications System Digital Ecosystem strategy. The selected candidate will manage a portfolio of Digital Ecosystem System Engineering (DESE), Data Architecture and Analytics, Product Lifecycle Management (PLM) Capability Development, Software...
-
Senior Software Engineer
7 days ago
Menlo Park, California, United States Cyngn Full timeAbout CyngnCyngn is a publicly traded autonomous vehicle company based in Menlo Park, CA. We have a culture of collaboration, diversity, and continuous learning. Our self-driving technology can be deployed in various commercial domains across various vehicle form factors.About the RoleWe are seeking a skilled Full Stack Engineer to contribute to the...
-
Production Systems Engineer
3 weeks ago
Menlo Park, California, United States META Full timeJob Title: Production EngineerMeta is seeking a highly skilled Production Engineer to join our team. As a Production Engineer, you will be responsible for designing, developing, and maintaining software services to ensure optimal performance and capacity for growth.Key Responsibilities:Develop and maintain back-end data warehouse services, front-end...
-
Senior Systems Engineer
4 weeks ago
Lexington Park, Maryland, United States BAE Systems Full timeJob Title: Senior Systems EngineerWe are seeking an experienced Senior Systems Engineer to join our team at BAE Systems. As a key member of our Digital Ecosystem team, you will be responsible for implementing and managing the VH92-A Mission Communications System Digital Ecosystem strategy.Key Responsibilities:Manage a portfolio of Digital Ecosystem System...
-
Software Engineering Specialist
3 days ago
Menlo Park, California, United States Diffuse Bio Full timeKey Responsibilities:Design and develop software and APIs to enable internal and external access to our AI systems.Build tools to automate and maintain computing clusters and data parsing pipelines.Collaborate with our team of researchers to develop cutting-edge AI solutions.Requirements:Bachelor's or Master's degree in Computer Science or a related...
-
Senior Systems Engineer
4 weeks ago
Lexington Park, United States BAE Systems USA Full timeAbout the RoleBAE Systems USA is seeking an experienced Senior Systems Engineer to join our team in St. Mary's County, Maryland. As a Senior Systems Engineer, you will be responsible for implementing and managing the VH92-A Mission Communications System Digital Ecosystem strategy.Key ResponsibilitiesManage a portfolio of Digital Ecosystem System Engineering...