Expert DL Deployment Engineer shaping AI models from lab to production at Yora. Leading deployment of deep learning models and collaborating with research teams.
Responsibilities
Lead the end-to-end deployment of deep learning models into production, ensuring they are fast, scalable, and reliable
Optimize neural network inference on GPUs, embedded devices, and edge systems
Develop production-grade software that meets the highest standards for robustness, latency, and maintainability
Collaborate closely with research teams, data scientists, and software engineers to translate experimental models into real-world applications
Mentor junior engineers and share your deep expertise across the team, shaping best practices for DL deployment
Evaluate new tools, frameworks, and technologies to push the boundaries of AI deployment
Requirements
10+ years of experience in AI, deep learning, or related fields
Strong expertise in neural network inference, model optimization, and performance tuning
Proficiency in C++ and Python
Deep knowledge of CUDA, GPU programming, TensorRT, and ONNX Runtime
Solid experience with embedded systems and production-grade software
A problem-solving mindset, attention to detail, and a drive to make AI work at scale
Benefits
Worklife Mode: Design your own work setup, balancing focus and flexibility
Tech in Action: Be the person who ensures that world-class AI models actually run in production
Collaborative culture: Work alongside talented, driven individuals in a low-ego environment
Freedom to grow: Shape your career path and influence the technical direction of the company
DevOps/IT Apprentice supporting cloud infrastructure and CI/CD pipelines at tech startup. Involves learning, taking ownership, and growing within the engineering team.
DevOps Engineer at Cloud++ collaborating on infrastructure and CI/CD pipelines across multi - cloud environments. Engaging with development teams to ensure reliable and secure releases.
Fullstack Developer at Zenika engaging in impactful tech projects like B2B platforms and architecture modernization. Collaborating with senior consultants in a quality - focused environment.
Production Engineer in a hybrid role ensuring operational performance of applications for a strategic international project. Focusing on automation and optimization within a technical environment at EOLEN.
Senior/Expert DevOps Engineer for AI project in pharmaceutical sector at GECI International. Involves designing, deploying, and operating autonomous AI agents.
DevOps Engineer Intern at Emeria Technologies focusing on Cloud infrastructure design and support. Involves maintaining CI/CD platforms and collaborating with DevOps teams for optimization.
Intern supporting software development infrastructure including CI/CD and cloud integration at Intel. Collaborating with teams to optimize development and release processes.
DevOps Engineer at NetBrain responsible for AWS cloud infrastructure and automating processes. Collaborate with development and security teams to deliver secure solutions efficiently.
Site Reliability Engineer at BlaBlaCar improving CI/CD and tooling for developer efficiency and autonomy. Collaborating with engineering teams to enhance service reliability and facilitate software development.