דילוג לתוכן הראשי

הגדרות עוגיות

בחרו מה לאפשר. אפשר לשנות את הבחירה בכל עת דרך „הגדרות עוגיות” בתחתית כל עמוד.

חיוניות

התחברות, אבטחה ושמירת הבחירות שלכם באתר — כולל הבחירה הזו, ערכת הצבעים והגדרות הנגישות. וגם מאיזה אתר או קמפיין הגעתם בביקור הראשון (בלי מזהה אישי, נשמר 30 יום), כדי שנדע מאיפה מגיעות ההרשמות לאתר.

תמיד פעילות

סטטיסטיקה ומדידה

אילו דפים נצפים ואיך משתמשים בהם, כדי לשפר את האתר. בלי שם ובלי פרטי קשר.

כלים: Google Analytics, Microsoft Clarity

פרסום ושיווק

מודעות שמותאמות לתחומי העניין, והתראות דחיפה על משרות חדשות למי שביקש. בלי הסכמה מוצגות מודעות כלליות בלבד.

כלים: Google AdSense, OneSignal

פרטים נוספים במדיניות הפרטיות.

JOBTIME
לוגו אתוסיה

Senior SW Engineer – AI Infrastructure & Optimization

אתוסיה

עלתה ל-JOBTIME לפני 9 שעות· בתוקף עד 24 בנובמבר 2026

מיקום
כל הארץ
היקף משרה
משרה מלאה
תפקיד
מהנדס/ת תוכנה

תיאור המשרה

What You’ll Do

  • Build and optimize high-performance Kubernetes-native GenAI inference systems 
  • Work with modern inference stacks such as vLLM, SGLang, TensorRT-LLM, and related tooling 
  • Work with Kubernetes-native distributed LLM inference frameworks such as llm-d and NVIDIA Dynamo 
  • Design and implement optimization algorithms and performance improvements 
  • Improve reliability, observability, deployment, and operational maturity of AI systems 
  • Make architectural decisions and take ownership of technical outcomes 
  • Collaborate with a small, senior engineering team focused on performance and production quality 

Requirements

Required Qualifications

  • Minimum 5 years of experience as a Software Engineer, with strong software engineering and system design skills. 
  • Programming experience in Go and Python 
  • Hands-on experience with the Kubernetes ecosystem,including Operators, service meshes, GitOps, Gateway API, and OpenTelemetry 
  • Experience with cloud platforms 
  • Strong understanding of optimization algorithms and performance engineering 
  • Ability to independently drive technical initiatives from concept to production 
  • Strong systems thinking and debugging skills 
  • Comfort operating in environments with high
    autonomy and responsibility 

Nice to Have

  • Experience with modern LLM inference frameworks such as vLLM, SGLang, or TensorRT-LLM 
  • Experience with distributed LLM inference frameworks such as llm-d or NVIDIA Dynamo 
  • Contributions to open-source Kubernetes or ML infrastructure projects 
  • GPU performance optimization and profiling experience 
  • Familiarity with CUDA, NCCL, or Triton kernels 
  • Experience running GenAI systems at scale in production 

·    

משרות דומות

הגדרות נגישות

ערכת צבעים

גודל טקסט

100%

התאמות תצוגה

הצהרת נגישות