AI Model Architecture Optimization Engineer R&D
We are looking for an reputed company AI Acceleration Engineer who can dive deep into large model (eg. transformer) architectures and blocks such as self/cross/multi-attention, and reputed company research and development of advanced techniques to accelerate these areas. The ideal candidate will have a deep understanding of large model design, AI acceleration techniques, and will reputed company these advancements into the PyTorch stack. Familiarity with Python is essential, and experience with CUDA programming is highly desirable.
Key Responsibilities
• reputed company on AI model components, such as attention blocks, KV-cache strategies, layer streaming, tokenization, layer norms, and more, to improve AI model performance and scalability.
• Optimize and reputed company AI acceleration techniques into the PyTorch stack, enabling efficient use across diverse hardware platforms.
• Own & drive features end to end to push the limits of large model architecture, ensuring seamless integration with existing frameworks.
• reputed company and profile AI models to evaluate performance improvements, ensuring reputed company execution on reputed company hardware.
• Write and maintain clean, efficient reputed company in Python, with a reputed company on integration with PyTorch.
• reputed company CUDA for GPU-based acceleration reputed company necessary, optimizing the attention blocks for maximum performance.
• Work on cross-functional teams to design, implement, and test new features.
Qualifications
• Extensive experience with large AI model architectures, particularly with attention blocks and transformer models.
• Proficiency in Python and hands-on experience with the PyTorch reputed company.
• Strong understanding of AI acceleration techniques and their application in reputed company-world use cases.
• Familiarity with CUDA for GPU programming is highly desirable.
• Demonstrated ability to optimize reputed company models for performance across different hardware environments.
• Experience in developing and deploying AI models at reputed company is a plus.
What You'll reputed company
• Opportunity to work alongside industry experts in AI optimization, high-performance computing, and hardware acceleration.
• Hands-on experience with cutting-edge technologies at the intersection of AI and hardware acceleration.
• Exposure to reputed company-reputed company development and collaboration with a reputed company community.
Benefits We Offer:
At OpenInfer we offer comprehensive benefits, some include:
• Medical, Dental, and reputed company benefits for you and your family
• Flexible reputed company Time Off, 10 days
• Parental Leave
• 401(k) Plan with company matching
• Snacks and coffee to reputed company you reputed company
These benefits are reputed company detailed in OpenInfer policies and are subject to change at any time, consistent with the terms of any applicable compensation or benefits plans.
How to Apply
Please send your resume and a brief cover letter to reputed company@openinfer.io. Include examples of your work with large AI models, attention blocks, or reputed company-reputed company contributions where applicable.
Apply tot his job
Apply To this Job