首席数据科学家 - 深度学习
Principal Data Scientist - Deep Learning
我们是谁
在 Jampp,我们的使命是发挥主导作用,推动移动应用经济的发展。我们如何做到这一点?我们构建技术,支持最雄心勃勃的公司——从游戏到电商先驱——扩大其应用的覆盖面,并加速其移动业务的发展。
我们解决当今移动广告中最复杂和大规模的技术挑战。我们每秒处理超过 250 万次广告请求,每天处理超过 300TB 的数据,分布在三个全球数据中心。我们依赖实时机器学习模型(包含数十亿特征),能在不到 100ms 内给出预测。
为什么我们需要你?
数据及其使用方式在 Jampp 处于核心地位,是我们做出产品、业务、运营和财务决策的关键。
我们的数据科学团队在许多技术领域解决具有挑战性的问题,包括时间序列预测、图挖掘、基于 PB 级数据的算法优化、缺失数据的因果推断,以及大规模机器学习。
我们现在正在将我们的机器学习平台提升到一个新的水平。我们正从围绕手动设计特征的模型,转向基于嵌入的深度学习架构,使我们的模型能够学习用户、设备、创意、发布者、广告主、应用、活动和展示位置的更丰富的表示。
作为高级数据科学家 – 深度学习,你将在定义和引领这一转型中发挥关键作用。你将为下一代机器学习架构的开发提供技术领导,从深度学习模型和嵌入策略的设计,到它们在生产环境中的部署和演进。
你将亲自处理程序化广告中最具挑战性的建模问题,同时设定技术方向,建立最佳实践,并帮助其他数据科学家和工程师做出合理的架构和建模决策。
这不是一个研究孤岛。你将有机会处理大规模、高基数的数据集,实时竞价系统和直接影响 Jampp 每秒评估和竞标数百万个广告机会的模型。
你将做什么
- 定义并推动 Jampp 基于深度学习和嵌入的建模架构的技术战略。
- 设计、开发并迭代先进的 DNN 架构用于预测和优化,最初专注于 CPI/CPA 使用场景,
查看英文原文
WHO WE ARE
At Jampp, we’re on the mission of playing a leading role enabling the mobile app economy to grow. How? We build technology to support the most ambitious companies - from gaming to commerce pioneers - to propel the reach of their apps and accelerate their mobile businesses.
We solve the most complex and large-scale technological challenges in mobile advertising today. We process over 2,500,000 ad requests per second, which amounts to over 300TB of data per day, across three global data centers. We rely on real-time machine learning models (with billions of features) that give us predictions in less than 100ms.
WHY DO WE NEED YOU?
Data, and how it is used, plays a central role at Jampp and is at the heart of how we make product, business, operational, and financial decisions.
Our Data Science team tackles challenging problems across many technical disciplines, including time series forecasting, graph mining, algorithmic optimization on petabytes of data, causal inference with missing data, and machine learning at scale.
We are now taking our machine learning platform to the next level. We are evolving from models built around manually designed features towards a deep learning architecture based on embeddings, allowing our models to learn richer representations of users, devices, creatives, publishers, advertisers, apps, campaigns and placements.
As a Principal Data Scientist – Deep Learning, you will play a key role in defining and leading this transformation. You will provide technical leadership across the development of our next-generation machine learning architecture, from the design of deep learning models and embedding strategies to their deployment and evolution in production.
You will work hands-on on some of the most challenging modeling problems in programmatic advertising, while setting technical direction, establishing best practices and helping other Data Scientists and engineers make sound architectural and modeling decisions.
This is not a research silo. You will have the opportunity to work with massive, high-cardinality datasets, real-time bidding systems and models that directly influence how Jampp evaluates and bids on millions of advertising opportunities every second.
WHAT YOU’LL DO
- Define and drive the technical strategy for Jampp’s deep learning and embedding-based modeling architecture.
- Design, develop and iterate advanced DNN architectures for prediction and optimization, initially focusing on CPI/CPA use cases and progressively expanding into real-time bidding, bid optimization, ranking and campaign optimization.
- Define the architecture and strategy for learning rich representations of high-cardinality entities such as users, devices, creatives, publishers, advertisers, apps, campaigns and placements.
- Lead the design of reusable embedding and representation-learning approaches that can support multiple models and use cases across the platform.
- Identify opportunities to improve the performance, scalability and generalization of our machine learning models using raw signals and learned representations.
- Develop modeling approaches for real-time prediction and decision-making in programmatic advertising, where models operate within strict latency and scale constraints.
- Evaluate new modeling approaches and technologies, balancing state-of-the-art deep learning techniques with the practical requirements of large-scale DSP and RTB systems.
- Establish technical standards and best practices for model development, experimentation, evaluation and productionization across the Data Science team.
- Guide complex modeling initiatives from problem definition and experimentation through production deployment and continuous improvement.
- Design, code and deploy machine learning models and supporting production tools, primarily in Python, remaining hands-on with the most technically challenging parts of the work.
- Work closely with ML Engineers, Data Engineers and Software Engineers to shape the training infrastructure, data pipelines, feature infrastructure, serving architecture and feedback loops required to support the next generation of Jampp’s ML platform.
- Analyze model and product performance metrics to understand how algorithmic changes impact bidding decisions, campaign performance, user response and business outcomes.
- Provide technical mentorship and guidance to Data Scientists, helping them navigate complex modeling problems and develop stronger approaches to experimentation and model design.
- Communicate technical findings, architectural decisions, trade-offs and recommendations clearly to both technical and non-technical stakeholders.
- Collaborate with Data Science and ML teams across Jampp and Affle to identify opportunities for shared capabilities, knowledge and machine learning solutions.
REQUIREMENTS
- Significant experience in Data Science, Machine Learning, Deep Learning or a closely related quantitative/technical role, with a track record of leading complex machine learning initiatives.
- Strong academic background in Computer Science, Applied Mathematics, Physics, Statistics, Engineering, Econometrics, or another quantitative field.
- Deep understanding of machine learning and deep learning fundamentals, including neural network architectures, representation learning, optimization and model evaluation.
- Extensive hands-on experience developing and deploying Deep Learning models in production environments.
- Strong experience with Python and the scientific/machine learning Python ecosystem.
- Experience working with large-scale datasets and high-cardinality categorical or ID-based features.
- Proven track record of taking machine learning models from experimentation and research through reliable production deployment.
- Experience designing or making significant technical contributions to ML architectures, training pipelines, model serving or other machine learning infrastructure.
- Experience working with real-time or latency-sensitive machine learning systems, ideally in advertising, marketplaces, recommendations or other high-throughput environments.
- Strong analytical and problem-solving skills, with the ability to independently investigate ambiguous problems and define effective technical solutions.
- Strong technical communication skills, with the ability to influence modeling and architectural decisions across multidisciplinary teams.
- Experience providing technical leadership, mentorship or direction to other Data Scientists or engineers.
- Comfortable conducting daily professional communications in English (written and verbal).
YOU MAY BE A GREAT FIT IF…
- You have designed and deployed deep learning systems based on embeddings, representation learning or other approaches for learning from high-cardinality entities.
- You have experience working on DSPs, programmatic advertising, real-time bidding (RTB), ad exchanges, ad networks or other real-time advertising systems.
- You have experience applying machine learning to bidding, bid optimization, CTR/CVR prediction, ranking, campaign optimization or other decision-making problems in advertising.
- You have experience with recommendation systems, ad-tech, pricing, ranking, personalization, fraud detection, or other large-scale optimization and prediction problems.
- You have worked with systems processing very large volumes of ad impressions, auction events or real-time user and publisher signals.
- You have experience modeling highly cardinal entities and complex interactions across users, devices, apps, creatives, publishers, advertisers, campaigns or similar entities.
- You have experience designing or evolving ML platforms, training pipelines, feature stores, model serving or monitoring systems.
- You have worked with models operating at very high scale and under strict latency constraints.
- You have experience balancing model complexity and predictive performance with computational cost, scalability and production constraints.
- You have a strong track record of turning research ideas into production systems and measuring their real-world impact.
- You enjoy defining technical approaches to problems where there is no obvious answer and where modeling decisions can have a significant impact on the product and the business.
- You are comfortable providing technical leadership while remaining hands-on with the most challenging modeling and engineering problems.
- You are able to influence technical decisions through strong reasoning, experimentation and communication rather than relying solely on formal authority.
- You like working in a self-sufficient, autonomous manner, striving through ambiguity and taking ownership of complex technical problems.
- You have a strong sense of urgency and ownership over the product, and care deeply about the quality, scalability and impact of the solutions you build.
- You are curious, pragmatic and comfortable balancing technical depth with practical business impact.
- Smarts, humility, and equal willingness to learn and teach.