Connecting Odds
Intel

AI Frameworks Engineer - Model and Kernel Optimization

Intel
PRC, ShanghaiPosted about 1 month agoDiscoveredMatch locked
Full time

Job Description

Artificial Intelligence (AI) is transforming our lives and is everywhere. At INTEL, we are part of this AI revolution. Our software stack integrates seamlessly into customer frameworks, which are used by millions of end users. As our AI engineering team in Shanghai continues to grow, we are looking for a passionate engineer to help us deliver high-performance and high-quality deep learning solutions to our customers. Our team's work includes:

  • Optimizing performance for key use cases/models, debugging and resolving issues related to accuracy and memory management
  • Designing and developing model deployment architectures, such as implementing new features on vLLM/SGLang to accelerate inference (e.g., P/D disaggregation)
  • Developing and debugging high-performance kernels for INTEL accelerators
  • Communicating with direct teammates, collaborators, and architects to discuss issues, propose solutions, provide status updates, and gather feedback
  • Applying innovative ideas to enhance our products

Qualifications

  • A Master's or Ph.D. degree in Computer Science, Artificial Intelligence, Software Engineering, or a related field
  • Strong programming skills in C++ and Python
  • Solid understanding of deep learning fundamentals and hands-on experience
  • Proficient in both written and spoken English
  • Passionate about problem-solving and proactive thinking
  • Nice to have:

o Experience with LLMs or AIGC and a deep understanding of model architectures o Or experience with PyTorch, vLLM, or SGLang o Or experience in GPU kernel development

Job Type

College Grad

Shift

Shift 1 (China)

Primary Location

PRC, Shanghai

Posting Statement

All qualified applicants will receive consideration for employment without regard to race, color, religion, religious creed, sex, national origin, ancestry, age, physical or mental disability, medical condition, genetic information, military and veteran status, marital status, pregnancy, gender, gender expression, gender identity, sexual orientation, or any other characteristic protected by local law, regulation, or ordinance. Position of Trust N/A

Work Model for this Role

This role will require an on-site presence. Job posting details (such as work model, location or time type) are subject to change.

ADDITIONAL INFORMATION: Intel is committed to Responsible Business Alliance (RBA) compliance and ethical hiring practices. We do not charge any fees during our hiring process. Candidates should never be required to pay recruitment fees, medical examination fees, or any other charges as a condition of employment. If you are asked to pay any fees during our hiring process, please report this immediately to your recruiter.

Not included in the source posting: about the role, what you'll do, benefits.

Skills

artificial-intelligencedeep-learningcllmpythonpytorch

Who can apply

The employer didn't state any visa, work authorization, citizenship or clearance requirements in this posting. Confirm with the employer before applying.

Read automatically from the employer's posting text. Always confirm with the employer — requirements can change after a job is published.