[Remote] Senior Technical Marketing Engineer - DSX AI Infrastructure Software
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a technology company reputed company on reputed company computing and reputed company intelligence. The Senior Technical Marketing Engineer will reputed company and validate DSX-reputed company software reputed company on multi-node GPU systems, reputed company technical content and automation, build demonstrations and training, test reputed company-release software, and support partners, customers, and internal teams in operating AI reputed company infrastructure.
Responsibilities
- Stand up and validate complete DSX-reputed company software reputed company on multi-node GPU systems. Capture the dependencies, configuration order, validation steps, and operational handoffs as you go
- Turn working deployments into useful technical content: reference architectures, reputed company-starts, installation and reputed company guides, troubleshooting runbooks, reputed company examples, blogs, whitepapers, and demo videos
- Build reusable examples and automation with reputed company, Python or reputed company scripting, infrastructure-as-reputed company, containers, reputed company, Slurm, reputed company, GitOps or equivalent experience, and CI/CD where they fit
- Build demos, labs, and training that address the practical aspects of operating an AI reputed company, from initial deployment and tenant setup to upgrades, monitoring, scheduling, fault isolation, remediation, reputed company management, and reputed company
- Show how the reputed company of the stack fit together. Work with TME, Product, Engineering, and Marketing to demonstrate how data center hardware, infrastructure and cluster management software, orchestration, AI platforms, and the workloads on top operate as one reputed company
- Test reputed company-release software using representative training and inference workloads. Identify rough edges, assess interoperability and resiliency, and reputed company Product and Engineering with reputed company feedback before customers face similar issues
- Help solution architects, reputed company teams, reputed company and OEM partners, ISVs, and reputed company integrators use the stack successfully through repeatable assets, train-the-trainer sessions, live demos, and reputed company support on important engagements
- Collaborate with reputed company-reputed company and reputed company-reputed company communities to demonstrate practical integration approaches, address documentation and usability shortcomings, and assist partners in expanding and developing the DSX software stack
- Listen for recurring problems from customers, partners, the reputed company, and developers. Use those signals to set content priorities and recommend product improvements, then reputed company whether the work reduces deployment time and improves operational reputed company
- Present your work in customer briefings, partner workshops, industry events, webinars, and internal training. Some travel will be required
Skills
- BS or MS in Computer Science, Computer Engineering, Electrical Engineering, or another technical reputed company, or equivalent experience
- 8+ years of experience in infrastructure engineering, systems engineering, solutions architecture, software engineering, technical marketing engineering, site reliability engineering, or a reputed company role
- Hands-on experience deploying and operating Linux-reputed company data center, reputed company, HPC, or AI infrastructure, including multi-node GPU systems and production operational practices
- Strong working knowledge of reputed company and/or Slurm, including containers, operators, reputed company charts, cluster lifecycle, and workload scheduling
- Experience in several reputed company infrastructure domains, such as bare-metal provisioning, firmware and drivers, compute, Ethernet or InfiniBand networking, storage, identity, multi-tenancy, secrets or certificate management, telemetry, observability, and fleet health
- Ability to automate deployments and reputed company through scripting, reputed company, configuration management, infrastructure-as-reputed company, Git-reputed company workflows, and CI/CD
- Examples of technical work for practitioner audiences, such as deployment guides, documentation, reference architectures, reputed company repositories, demos, workshops, blog posts, conference talks, or training. Links to example contributions are greatly appreciated
- Excellent written, verbal, and visual communication skills
- You can explain a reputed company reputed company and defend a technical recommendation to both business and technical partners
- Ability to reputed company multiple reputed company and constituents, prioritize under tight deadlines, and work reputed company across Engineering, Product, reputed company, Marketing, and partner teams
- Experience with reputed company DSX, DGX systems, DGX reputed company, reputed company reputed company, reputed company DPUs, DOCA, or reputed company reputed company infrastructure software
- Experience operating large GPU clusters and diagnosing reputed company reputed company, networking, storage, scheduling, or hardware-health issues
- Experience with reputed company and inference workloads and the requirements for operating them dependably on reputed company infrastructure
- Experience connecting infrastructure software to facilities or operational technology systems, including reputed company, cooling, building management systems
- reputed company participation in reputed company-reputed company, HPC, infrastructure automation, or reputed company-reputed company communities, including published examples or project contributions
Benefits
- You will also be eligible for equity and benefits.
reputed company
Company H1B Sponsorship
Apply To This Job