ActiveJobs

AI Frameworks Software Engineer – Model Compression

Intel · PRC, Shanghai

Full-timeOn-sitePosted 13 August 2026
Apply on Company Site →

Job description

Job Details: Job Description: The Intel Neural Compressor team develops state-of-the-art model compression technologies, including quantization, pruning and sparsity, knowledge distillation, and low-precision training and fine-tuning. Our work spans algorithm research, product feature development, performance optimization, and contributions to the open-source community. We are looking for a highly self-motivated engineer to join our team. Key Responsibilities: Develop Intel Neural Compressor and its core algorithm tools, including AutoRound, and optimize them for Intel AI platforms such as CPUs, GPUs, and AI accelerators. Research and implement quantization and compression techniques for large language model (LLM), vision language model (VLM), and generative models, including text-to-image, text-to-video and world models. Track and explore emerging directions in efficient model deployment, inference acceleration, and fine-tuning acceleration. Qualifications:Qualifications: Bachelor’s or master’s degree in Computer Science or a related field. Solid understanding of deep learning, deep learning frameworks, and large language model (LLM) fundamentals. Familiarity with model compression techniques such as quantization and pruning. Proficiency in Python, C++, or other programming languages commonly used in deep learning development. Strong teamwork mindset and collaboration skills. Good verbal and written English communication skills. Preferred Qualifications: Strong self-motivation, ownership, and problem-solving skills. Passion for technological innovation and practical engineering, with a commitment to continuous exploration and improvement. Experience in model fine-tuning, inference optimization, or related tool development is preferred. Job Type:College Grad Shift:Shift 1 (China) Primary Location: PRC, Shanghai Additional Locations: Posting Statement:All qualified applicants will receive consideration for employment without regard to race, color, religion, religious creed, sex, national origin, ancestry, age, physical or mental disability, medical condition, genetic information, military and veteran status, marital status, pregnancy, gender, gender expression, gender identity, sexual orientation, or any other characteristic protected by local law, regulation, or ordinance.Position of TrustN/A Work Model for this Role This role will require an on-site presence. * Job posting details (such as work model, location or time type) are subject to change. * ADDITIONAL INFORMATION: Intel is committed to Responsible Business Alliance (RBA) compliance and ethical hiring practices. We do not charge any fees during our hiring process. Candidates should never be required to pay recruitment fees, medical examination fees, or any other charges as a condition of employment. If you are asked to pay any fees during our hiring process, please report this immediately to your recruiter.

Verified and listed by ActiveJobs. Applications are made directly on Intel's own career page — we never sit in the middle.