AI Software Engineer
Deloitte
- Where
- Haifa, Israel, hybrid
- Language, from the listing
- No Hebrew mentionedA missing mention doesn't mean Hebrew isn't needed. Ask if it matters to you.
- Dates
- Found 1 Oct 2026
- Last checked on the employer's site
- 5 h ago (6 Oct 2026)
- Source
- Employer career page (Comeet)
What they ask for
- BSc or MSc in Computer Science, Electrical Engineering, or a related field
- 4+ years of experience in software engineering and applied AI (or equivalent)
- Strong knowledge of software design and development practices
- Strong proficiency in Python and modern AI frameworks
- Proven experience delivering production-grade AI systems
- Solid understanding of deep learning architectures (CNNs, transformers)
- Experience with system-level design, debugging, and performance optimization
- Experience working with Large Language Models (LLMs)
- Building agentic workflows, reasoning systems, or AI-driven applications
- Deploying and optimizing open-source LLMs for inference
Nice to have
- Experience with NVIDIA technologies (CUDA, TensorRT, Triton, NVIDIA NIM)
- Experience with GPU performance profiling and optimization
- Background in high-performance or low-latency systems
- Experience with Multimodal LLM agents for text, image, audio, and video processing
- Experience with React, TypeScript, Redis Pub/Sub, SQL, Containerized Development
- Experience with Test-Driven Development, Agile Scrum methodology, Git workflows
- Experience mentoring engineers or leading technical initiatives
The full listing
Description
About Us
We are an NVIDIA partner, delivering professional services in advanced software development, artificial intelligence systems, and high-performance AI solutions. We build production-grade AI systems optimized for NVIDIA GPU platforms, working across multiple applied AI domains, including Generative AI / Agentic Systems.
Role Overview
We are looking for an experienced AI Software Engineer to take a leading role in designing, building, and delivering advanced AI systems for real-world applications. This role is suited for a hands-on engineer with strong software engineering skills and deep experience in applied AI, who enjoys solving complex problems, owning technical solutions end-to-end, and mentoring junior team members.
Projects typically focus on LLM or Multimodal LLM based agentic systems.
What We Work On
Our work spans software engineering and AI domains:
• Design and implementation of agentic systems powered by Large Language Models (LLMs)
• Deployment of open-source LLMs on NVIDIA GPUs
• High-performance inference using NVIDIA technologies such as NVIDIA NIM (Inference Microservices)
• System-level design for reliability, scalability, and performance
• Multimodal agents for text, image, audio, and video processing
• Performance optimization on GPU platforms
Key Responsibilities
• Lead the design and development of LLM-based agentic systems.
• Own end-to-end delivery of AI solutions, from architecture and prototyping to production deployment using rigorous software engineering practices
• Optimize model performance, latency, and throughput on NVIDIA GPU platforms
• Design clean, maintainable, and scalable software architectures
• Collaborate with customers, product teams, and engineers to translate requirements into technical solutions
• Mentor junior engineers and contribute to technical best practices
• Evaluate new tools, AI models, and frameworks and drive their adoption when appropriate
Requirements
• BSc or MSc in Computer Science, Electrical Engineering, or a related field
• 4+ years of experience in software engineering and applied AI (or equivalent)
• Strong knowledge of software design and development practices
• Strong proficiency in Python and modern AI frameworks
• Proven experience delivering production-grade AI systems
• Solid understanding of deep learning architectures (CNNs, transformers)
• Experience with system-level design, debugging, and performance optimization
• Experience working with Large Language Models (LLMs)
• Building agentic workflows, reasoning systems, or AI-driven applications
• Deploying and optimizing open-source LLMs for inference
Nice to Have
• Experience with NVIDIA technologies (CUDA, TensorRT, Triton, NVIDIA NIM)
• Experience with GPU performance profiling and optimization
• Background in high-performance or low-latency systems
• Experience with Multimodal LLM agents for text, image, audio, and video processing
• Experience with React, TypeScript, Redis Pub/Sub, SQL, Containerized Development
• Experience with Test-Driven Development, Agile Scrum methodology, Git workflows
• Experience mentoring engineers or leading technical initiatives
What We Offer
• Ownership of complex, high-impact AI projects
• Work with cutting-edge NVIDIA GPU and AI technologies
• Influence over architecture, tooling, and technical direction
• A collaborative, engineering-driven culture
• Opportunities for technical leadership and professional growth
• Real-world, production-scale AI challenges
Full time Job
Location: Haifa, Hybrid
We at Deloitte believe that diversity and inclusion among our people is a critical component of our success and that is why we cultivate an organizational culture that contains and embraces diversity in all its forms.