About Us
Mintegral is a premier programmatic and interactive mobile advertising platform, originating from the APAC region and expanding globally. Driven by advanced AI technology, we deliver innovative, comprehensive solutions to advertisers and developers, helping them achieve their marketing goals through exceptional mobile marketing and monetization strategies.
Since launching in 2015 as Mobvista's self-developed programmatic platform, Mintegral has rapidly ascended to one of the largest mobile advertising platforms in Asia. Our offerings include a complete suite of programmatic products and services, encompassing our Self-service Platform, DSP, SSP, Ad Exchange, and DMP. We also established the Mindworks Creative Studio, which delivers cutting-edge creative solutions to publishers and brands, spanning from traditional creative formats to the latest interactive ads.
About the Role
We are looking for an accomplished Senior Machine Learning Systems Engineer to design and scale the next generation of advertising ranking and serving infrastructure.
In this role, you will develop large-scale real-time ML inference systems that enhance ad ranking, retrieval, and prediction for global traffic. Your focus will be on distributed systems, ML infrastructure, and high-performance computing.
Responsibilities
Ads Serving System ArchitectureEngineer complete ads ranking and serving architectures for real-time bidding and recommendation systems.Create flexible and separated inference pipelines across CPU and GPU layers.Optimize latency for high-QPS ad delivery systems.ML Inference OptimizationDevelop and enhance ML deployment pipelines for diverse CPU/GPU environments.Augment model freshness and inference performance.Facilitate swift iterations of ML models in production.Distributed Systems & Embedding InfrastructureDesign extensive embedding storage and retrieval systems.Create adaptive sharding strategies across different hardware.Strengthen load balancing and overall system stability.Ads Ranking Performance EngineeringOptimize QPS, latency, and throughput.Identify and resolve bottlenecks in inference pipelines.Enhance overall ranking performance from end to end.ML Compiler & Runtime SystemsDevelop AOT compilation frameworks for ML models.Transform models into optimized C++/CUDA/ROCm execution.Elevate inference efficiency across various hardware backends.Cross-functional CollaborationCollaborate with engineers and product teams.Productionize ML models effectively.Establish and uphold standards for scalability and reliability.
Required Qualifications
4+ years in distributed systems or ML infrastructure.Experience with ML serving or recommendation systems.Solid background in distributed systems and performance optimization.Experience with CPU/GPU systems.Proficiency in C++ / Python.Familiarity with ML frameworks.
Preferred Qualifications
Experience in ads tech or recommendation systems.Knowledge of ML compilers or inference runtimes.GPU optimization skills (CUDA/ROCm).Experience with high-scale systems.Proven record of high-level ownership in projects.
Impact
Facilitate large-scale real-time ads ranking.Enhance inference efficiency and reduce costs.Accelerate ML deployment cycles.Establish foundational ads infrastructure.
Create a free account to keep reading — and to apply.
Create My Free AccountYou're just 60 seconds away from your new Creativeloft account.