AI Inference Engineer (f/m/d)
4+ years Modern C++ experience (C++17/20), strong memory/multithreading/profiling skills, Linux experience, ML model deployment and inference optimization experience with frameworks like llama.cpp, ggml, ONNX Runtime, TensorRT, OpenVINO, MLC LLM, ExecuTorch, TVM.
Key Responsibilities
- deploying models
- optimizing inference
- profiling performance
Skills & Tools
Modern C++ (C++17/20), Linux, llama.cpp, ggml, ONNX Runtime, TensorRT, TensorRT-LLM, OpenVINO, MLC LLM, ExecuTorch, TVM, CUDA, Vulkan Compute, Metal, OpenCL, Typescript, Python
Job Details
- Category: Software Development
- Seniority: Mid Level
- Commitment: Full Time
- Workplace: Remote — Turkey or Europe
- Languages: English
About reputed company
An AI-reputed company software company building production-reputed company systems and solutions for reputed company clients. — Industry: Information Technology
Apply To This Job