Software Engineer

Date: Sep 15, 2026

Location: San Jose, California, United States

Company: Super Micro Computer

Job Req ID: 30251

About Supermicro:

Supermicro® is a Top Tier provider of advanced server, storage, and networking solutions for Data Center, Cloud Computing, Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide. We are the #5 fastest growing company among the Silicon Valley Top 50 technology firms. Our unprecedented global expansion has provided us with the opportunity to offer a large number of new positions to the technology community. We seek talented, passionate, and committed engineers, technologists, and business leaders to join us.

Job Summary:

We are looking for a passionate, motivated Software Engineer to join our dynamic and fast-paced team building switch agent and network management software for large-scale AI and data center fabrics. In this role, you will design, develop, and test software that runs on network switches, along with the automation that provisions, monitors, and troubleshoots them. Upcoming work reaches deeper into switch silicon and telemetry, and into making the resulting data and logs directly consumable by AI agents.

 

The ideal candidate has a strong foundation in systems or networking software and a passion for learning. In this role, you will work closely with team leaders, project managers, and senior engineers to innovate, design, and test, adapting to evolving priorities and technical challenges. This position provides hands-on exposure to production AI clusters built on Supermicro switches and NVIDIA/AMD GPU servers.

Essential Duties and Responsibilities:

  • Contribute to the design, development, and maintenance of on-switch agent software and the management and diagnostic services it integrates with.
  • Build automated software solutions for provisioning, monitoring, configuration, and troubleshooting across data center and AI cluster platforms.
  • Work with team leaders and project managers to innovate, design, implement, and test new capabilities from concept through release.
  • Develop and run diagnostic and performance workloads, such as link quality, packet drop and capture, routing neighbor checks, and system health, and analyze the results.
  • Troubleshoot software, automation, and integration issues and develop permanent solutions to improve reliability and maintainability.
  • Contribute to performance, reliability, and debuggability improvements across AI clusters, including telemetry and metrics pipelines.
  • Implement collection and processing of ASIC and platform telemetry, including hardware counters, buffer and queue statistics, and drop and error reporting, along with the pipelines that deliver it.
  • Help structure, label, and enrich collected metrics and logs so they can be consumed by AI agents and automated diagnostic workflows.
  • Maintain technical documentation, APIs, and operational procedures for developed software solutions.
  • Continuously evaluate new technologies and AI-assisted approaches to improve software development efficiency, diagnostics, and operational workflows.

Qualifications:

  • Bachelor's degree in Computer Science, Computer Engineering, Electrical Engineering, or a related technical field (Master's degree preferred).
  • 2-4+ years of experience in software engineering.
  • Strong programming skills: Go and/or Python; shell scripting; Ansible.
  • Working knowledge of L2/L3 networking protocols and management protocols; exposure to SONiC or comparable network operating systems is a plus.
  • Experience with network telemetry and monitoring, such as streaming telemetry (gNMI/OpenConfig), sFlow, IPFIX, SNMP, or metrics pipelines such as Prometheus.
  • Exposure to switch ASICs, ideally Broadcom (BCM), and the abstraction or SDK layers above them such as SAI, syncd, or vendor SDKs, is a strong plus.
  • Interest or experience in preparing structured, labeled data from logs, metrics, and events for consumption by AI agents and LLM-based tooling.
  • Experience with Linux/Unix operating systems, containers (Docker), and REST APIs.
  • Foundational understanding of software architecture, debugging, troubleshooting, and performance analysis.
  • Nice to have: RDMA, RoCE, VXLAN, K8s, Git, Redis, Grafana, AI-assisted development tools (Cursor).
  • Strong analytical and problem-solving skills.
  • Ability to work independently and manage multiple technical priorities.
  • Excellent written and verbal communication skills.

Salary Range

$110,000 - $135,000

The salary offered will depend on several factors, including your location, level, education, training, specific skills, years of experience, and comparison to other employees already in this role. In addition to a comprehensive benefits package, candidates may be eligible for other forms of compensation, such as participation in bonus and equity award programs.

EEO Statement

Supermicro is an Equal Opportunity Employer and embraces diversity in our employee population. It is the policy of Supermicro to provide equal opportunity to all qualified applicants and employees without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, protected veteran status or special disabled veteran, marital status, pregnancy, genetic information, or any other legally protected status.


Job Segment: Test Engineer, Software Engineer, Cloud, Testing, Embedded, Engineering, Technology