Connecting Odds
Intel

AI Frameworks Software Engineer – Model Compression

Intel
PRC, ShanghaiPosted about 1 month agoDiscoveredMatch locked
Full time

Job Description

The Intel Neural Compressor team develops state-of-the-art model compression technologies, including quantization, pruning and sparsity, knowledge distillation, and low-precision training and fine-tuning. Our work spans algorithm research, product feature development, performance optimization, and contributions to the open-source community. We are looking for a highly self-motivated engineer to join our team.

Key Responsibilities:

  • Develop Intel Neural Compressor and its core algorithm tools, including AutoRound, and optimize them for Intel AI platforms such as CPUs, GPUs, and AI accelerators.
  • Research and implement quantization and compression techniques for large language model (LLM), vision language model (VLM), and generative models, including text-to-image, text-to-video and world models.
  • Track and explore emerging directions in efficient model deployment, inference acceleration, and fine-tuning acceleration.

Qualifications

Qualifications:

  • Bachelor’s or master’s degree in Computer Science or a related field.
  • Solid understanding of deep learning, deep learning frameworks, and large language model (LLM) fundamentals.
  • Familiarity with model compression techniques such as quantization and pruning.
  • Proficiency in Python, C++, or other programming languages commonly used in deep learning development.
  • Strong teamwork mindset and collaboration skills.
  • Good verbal and written English communication skills.

Preferred Qualifications:

  • Strong self-motivation, ownership, and problem-solving skills.
  • Passion for technological innovation and practical engineering, with a commitment to continuous exploration and improvement.
  • Experience in model fine-tuning, inference optimization, or related tool development is preferred.

Job Type

College Grad

Shift

Shift 1 (China)

Primary Location

PRC, Shanghai

Posting Statement

All qualified applicants will receive consideration for employment without regard to race, color, religion, religious creed, sex, national origin, ancestry, age, physical or mental disability, medical condition, genetic information, military and veteran status, marital status, pregnancy, gender, gender expression, gender identity, sexual orientation, or any other characteristic protected by local law, regulation, or ordinance. Position of Trust N/A

Work Model for this Role

This role will require an on-site presence. Job posting details (such as work model, location or time type) are subject to change.

ADDITIONAL INFORMATION: Intel is committed to Responsible Business Alliance (RBA) compliance and ethical hiring practices. We do not charge any fees during our hiring process. Candidates should never be required to pay recruitment fees, medical examination fees, or any other charges as a condition of employment. If you are asked to pay any fees during our hiring process, please report this immediately to your recruiter.

Not included in the source posting: about the role, what you'll do, benefits.

Skills

deep-learningartificial-intelligencellmcpython

Who can apply

The employer didn't state any visa, work authorization, citizenship or clearance requirements in this posting. Confirm with the employer before applying.

Read automatically from the employer's posting text. Always confirm with the employer — requirements can change after a job is published.