
- מיקום
- כל הארץ
- היקף משרה
- משרה מלאה
- תפקיד
- מהנדס/ת תוכנה
תיאור המשרה
What will you do:
Develop and enhance AI graph compiler components targeting NPU architectures. Implement graph-level optimizations such as operator fusion, scheduling, memory planning, and layout transformations. Participate in lowering models from AI frameworks %28e.g., PyTorch, ONNX, TensorFlow%29 into NPU–optimized representations. Contribute to compiler passes focused on performance, memory efficiency, and numerical correctness. Collaborate closely with hardware, runtime, and AI framework teams to achieve optimal end-to-end performance. Analyze performance bottlenecks and assist in compiler-based optimizations. Debug and resolve issues across compiler, runtime, and hardware layers. Support testing, validation, and documentation of compiler features.
- BSc or MSc in Computer Science, Electrical Engineering, or a related field
- Proficiency in C++ and Python.
- 2–5 years of experience in systems software, AI software, embedded software, or other performance-critical development.
- Strong analytical skills, software fundamentals, including data structures, debugging, and performance optimization.
- Excellent interpersonal skills, flexibility, and a proactive “Can Do” attitude
Advantages:
- Experience with compiler or IR-based systems %28e.g., LLVM, MLIR, domain-specific compilers%29.
- Experience with AI inference engines, runtimes, or model deployment pipelines.
- Familiarity with AI accelerators such as NPUs, GPUs, DSPs, or ASICs.
- Knowledge of reduced-precision numerical formats %28FP16, BF16, INT8%29.