AI Engineer
About Us
We are an NVIDIA partner, delivering professional services in聽advanced software development, artificial intelligence systems, and high-performance AI solutions. We build production-grade AI systems optimized for NVIDIA GPU platforms, working across multiple applied AI domains, including聽Generative AI / Agentic Systems.
Role Overview
We are looking for an聽experienced AI Software Engineer聽to take a leading role in designing, building, and delivering advanced AI systems for real-world applications. This role is suited for a hands-on engineer with聽strong software engineering聽skills聽and deep experience in applied AI, who enjoys solving complex problems, owning technical solutions end-to-end, and mentoring junior team members.
Projects typically focus on LLM or Multimodal LLM based agentic systems.
What We Work On
Our work spans software engineering and AI domains:
- Design and implementation of agentic systems powered by Large Language Models (LLMs)
- Deployment of open-source LLMs on NVIDIA GPUs
- High-performance inference using NVIDIA technologies such as NVIDIA NIM (Inference Microservices)
- System-level design for reliability, scalability, and performance
- Multimodal agents for text, image, audio, and video processing
- Performance optimization on GPU platforms
Key Responsibilities
- Lead the design and development of LLM-based agentic systems.
- Own end-to-end delivery of AI solutions, from architecture and prototyping to production deployment using rigorous software engineering practices
- Optimize model performance, latency, and throughput on NVIDIA GPU platforms
- Design clean, maintainable, and scalable software architectures
- Collaborate with customers, product teams, and engineers to translate requirements into technical solutions
- Mentor junior engineers and contribute to technical best practices
- Evaluate new tools, AI models, and frameworks and drive their adoption when appropriate
- BSc or MSc in Computer Science, Electrical Engineering, or a related field
- 4+ years of experience in software engineering and applied AI (or equivalent)
- Strong knowledge of software design and development practices
- Strong proficiency in Python and modern AI frameworks
- Proven experience delivering production-grade AI systems
- Solid understanding of deep learning architectures (CNNs, transformers)
- Experience with system-level design, debugging, and performance optimization
- Experience working with Large Language Models (LLMs)
- Building agentic workflows, reasoning systems, or AI-driven applications
- Deploying and optimizing open-source LLMs for inference
Nice to Have
- Experience with NVIDIA technologies (CUDA, TensorRT, Triton, NVIDIA NIM)
- Experience with GPU performance profiling and optimization
- Background in high-performance or low-latency systems
- Experience with Multimodal LLM agents for text, image, audio, and video processing
- Experience with React, TypeScript, Redis Pub/Sub, SQL, Containerized Development
- Experience with Test-Driven Development, Agile Scrum methodology, Git workflows
- Experience mentoring engineers or leading technical initiatives
What We Offer
- Ownership of complex, high-impact AI projects
- Work with cutting-edge NVIDIA GPU and AI technologies
- Influence over architecture, tooling, and technical direction
- A collaborative, engineering-driven culture
- Opportunities for technical leadership and professional growth
- Real-world, production-scale AI challenges
Full time Job
Location: Haifa, Hybrid
We at Deloitte believe that diversity and inclusion among our people is a critical component of our success and that is why we cultivate an organizational culture that contains and embraces diversity in all its forms.
讬讜注抓 讘讻讬专 诇讗讜驻讟讬诪讬讬讝专
讬讜注抓 讘讻讬专 诇讗讜驻讟讬诪讬讬讝专
Required Skills
Required Languages
馃嚞馃嚙 English