PERSOL APAC

AI Algorithm Engineer (AI Training/Inference Acceleration)

PERSOL APAC

Singapore · Full Time

Be the first to apply

Experience
Any
Salary
Openings
1
Posted
vor 3 Stunden
Work mode
In office
Education
Master’s or PhD in Computer Science, AI, Electronic Engineering, or related field
Resume
Required to apply

Where you'll work

Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.

Job description

Company Overview

Our client is a globally recognized leader in information and communications technology infrastructure and smart devices, delivering comprehensive AI solutions across carriers, enterprises, governments, and individual consumers worldwide.

Role Overview

We are seeking an experienced AI Algorithm Engineer specialized in training and inference acceleration to join the team in Singapore. This role involves spearheading research, development, and implementation of advanced AI acceleration algorithms optimized for next-generation AI hardware.

Key Responsibilities

  • Lead innovation in AI training and inference acceleration algorithms, focusing on Agentic AI and Multimodal domains to enhance compute efficiency on cutting-edge AI architectures.
  • Manage the full development lifecycle of acceleration algorithms within proprietary AI frameworks and libraries, ensuring smooth integration and continuous performance improvement based on real-world data.
  • Provide strategic technical insights within training and inference, identifying trends like long-sequence modeling and sparsity to future-proof and maintain competitive advantage.

Candidate Requirements

  • Master's or PhD in Computer Science, Artificial Intelligence, Electronic Engineering, or related fields.
  • In-depth knowledge of Large Language Models and Mixture of Experts architectures, with demonstrated experience in training, fine-tuning, or deploying large-scale AI models.
  • Proficiency in AI acceleration technologies including zero-redundancy optimizers, distributed parallelism methods (TP, PP, SP, VP, DP), communication compression, and memory use optimization.
  • Expertise in high-performance attention mechanisms, KV-cache compression, quantization techniques, and sparsity-aware accelerations is essential.
  • Strong programming skills in Python coupled with extensive experience in prominent deep learning frameworks and hardware-accelerated libraries; experience in hardware-level kernel tuning is highly advantageous.
  • Comprehensive understanding of hardware-software co-design principles, familiarity with hardware constraints like memory bandwidth and compute cycles, and successful performance optimization for models exceeding 100 billion parameters.
  • Demonstrated ability to optimize AI models in distributed and heterogeneous systems by effectively leveraging hardware specifics for algorithmic enhancement.

Additional Information

Interested applicants should note that only shortlisted candidates will be contacted. By applying, candidates consent to the use and disclosure of their personal information as outlined in PERSOL Singapore Pte Ltd's Privacy Policy.

EA License No.: 01C4394

Minimum education

Master's Degree

🤖
Online · instant AI help
Broxer