仕事概要
Role Purpose:
The Lead ML Data Engineer is a senior technical leader responsible for enabling scalable, production-grade Data Science & Analytics (DSA) solutions within Coca-Cola Bottlers Japan Inc.'s (CCBJI) Vending Machines (VM) business unit. This role leads the development, optimization, and management of end-to-end ML and Analytics data workflows, ensuring reliable and efficient infrastructure for AI solutions like Assortment, Column Reallocation, and Placement.
The role works in close partnership with Data Scientists, Analytics Specialists, and IT to deliver high-impact, ML-ready datasets via best of breed tools. The role plays a key function in bridging business objectives with technical delivery by converting commercial data requirements into robust pipelines, reusable assets, and operational tooling.
The position also drives quality standards and data engineering practices within the team, ensuring model reproducibility, pipeline traceability, and integration with enterprise MLOps and governance frameworks.
Key Responsibilities:
Advanced ML Data Pipeline Development
- Design, develop, and maintain robust data pipelines to support ML model training, inference, and feature transformation workflows.
- Deliver performant and modular pipelines using Databricks (PySpark/Python) and Snowflake, aligned to architectural best practices.
- Ensure end-to-end ownership of data engineering from working with IT on raw data ingestion to developing model-ready feature layers.
- Implement CI/CD-ready transformation logic for reproducibility and handover to downstream components.
Feature Stores & Reusability Framework
- Architect and maintain a centralized, scalable feature stores that enables reuse across multiple DS use cases.
- Define feature documentation standards, naming conventions, and lifecycle management practices.
- Optimize joins, aggregations, and lookups to balance compute cost with accuracy and inference performance.
MLOps Integration & Model Lifecycle Engineering
- Work closely with the Data Scientists and Analytics Specialists to operationalize models through CI/CD pipelines (MLflow, GitHub Actions, Databricks Workflows).
- Design robust systems for retraining triggers, monitoring, and automated evaluation of production ML models.
- Implement failure recovery, alerting, and model rollback procedures in collaboration with IT and DevOps.
Engineering Excellence & Domain Leadership
- Serve as the go-to engineering authority for ML enablement within the DSA team.
- Lead technical design reviews, set code standards, and promote reusability and modularization across pipelines.
- Contribute internal tools, libraries, and utilities that improve engineering velocity and onboarding.
- Conduct informal mentoring and coaching for junior engineers and scientists working with data pipelines.
Agile Program Delivery
- Actively participate in agile sprint cycles, contributing to planning, estimation, retrospectives, and delivery metrics.
- Align with the Analytics Portfolio Manager, DS Manager, and Analytics Manager to prioritize deliverables and resolve cross-functional dependencies.
- Maintain and manage engineering backlog, surfacing technical debt or architectural decisions that require executive alignment.
Cross-Functional Collaboration
- Translate ambiguous business requirements into structured, scalable data solutions that accelerate DS and Analytics outcomes.
- Collaborate with the Analytics team to build curated views and pre-aggregated layers for Power BI or experimentation workflows.
- Partner with IT to ensure infrastructure provisioning, access control, and platform governance align with CCBJI’s enterprise standards.
Data Observability & Production Assurance
- Build and maintain monitoring systems to track pipeline performance, schema changes, and data freshness.
- Implement quality checks and exception handling to reduce operational risk and manual rework.
- Ensure SLAs are defined and met for model refreshes, data availability, and system uptime.
Key Outputs:
- Stable, scalable ML data pipelines that serve predictive models across VM business scenarios.
- Well-documented and reusable feature store logic shared across DS initiatives.
- Fully operationalized MLOps workflows supporting model retraining and deployment.
- Toolkits, templates, standards and frameworks adopted by team members for faster pipeline delivery.
- Reduction in time-to-ML model deployment time and increased delivery velocity.
- Measurable improvements in pipeline stability, data quality, and platform observability with operational dashboards
Performance Success Criteria (Examples):
- Launch production-grade ML pipelines for at least three strategic models within first 9–12 months.
- Reduce end-to-end model deployment cycle time by 25% through reusability and automation.
- Deliver a reusable feature store structure adopted by at least 3 DS initiatives.
- Implement and operationalize monitoring workflows covering pipeline reliability and data quality.
- Create internal engineering utilities or templates reused by at least two other team members.
- Maintain 98%+ reliability of ML workflows and data pipelines with documented support procedures.
必須スキル
Qualifications:
Education:
- Bachelor’s or Master’s degree in Computer Science, Data Engineering, Information Systems, or a related quantitative discipline.
Experience:
- 6+ years in data engineering roles, with 2+ years focused on ML data workflows.
- Proven experience enabling ML model development and deployment through scalable engineering solutions.
- Hands-on experience with Snowflake and Databricks in production environments.
- Experience contributing to cross-functional data science or analytics programs.
- Experience designing feature stores and implementing feature pipelines in ML production environments.
- Working knowledge of CI/CD, data versioning, and DevOps practices for pipeline and model automation.
Technical Skills:
- Deep proficiency in Python and PySpark for data transformation and ML pipeline development in tools like Databricks.
- Advanced SQL capabilities for data wrangling, transformation, and pipeline optimization.
- Hands-on experience with MLflow, Git, Docker, and other DevOps tools.
- Working knowledge of model retraining, inference orchestration, and version control.
- Understanding of data governance, observability, and engineering compliance principles.
- Exposure to BI tools like Power BI for quick validation and prototyping.
Collaboration & Leadership:
- Demonstrated ability to lead through expertise and technical credibility.
Skilled in communicating complex architecture to business and non-technical partners. - Works cross-functionally with DS, Analytics, and IT to drive aligned delivery.
- Embraces continuous improvement and knowledge-sharing mindset.
- Comfortable mentoring junior team members without formal people management.
- Driven by automation, reusability, and long-term scalability.
- Curious and proactive about learning new technologies and industry trends.
応募概要
| 給与 | 非公開 |
|---|---|
| 勤務地 | 150-0002 |
| 雇用形態 | 正社員 |
| 勤務体系 | 標準勤務時間 9:00~17:45 |
| 試用期間 | 3か月 |
| 福利厚生 | • 財産形成:退職金制度(企業型確定拠出年金 ※100%会社負担)、従業員持株会、グループ保険 |
企業情報
| 企業名 | コカ・コーラボトラーズジャパン株式会社 |
|---|---|
| 設立年月 | 2001年(平成13年)6月29日 ※2018年1月1日 コカ・コーラ ボトラーズジャパン株式会社に商号変更 |
| 本社所在地 | 〒107-6211 東京都港区赤坂九丁目7番1号 ミッドタウン・タワー |
| 事業内容 | 清涼飲料水・アルコール飲料の製造、加工および販売 すべての人にハッピーなひとときをお届けし、価値を創造します - Deliver happy moments to everyone while creating value 私たちコカ・コーラ ボトラーズジャパンは、東は宮城県から西は鹿児島県まで1都2府35県を営業地域として、コカ・コーラ社製品を製造・販売するボトラーです。 日本のコカ・コーラシステムの約9割の販売量を担う、国内最大のコカ・コーラボトラーであるとともに、世界に250以上あるコカ・コーラボトラーの中でも、売上高でアジア最大、世界でも有数の規模を誇ります。 私たちはこれからも事業を持続的に成長させることで、利益還元の拡充と企業価値の向上を実現し、みなさまの期待に応えてまいります。 当社は、人々の一生と日々の生活に寄り添い、人生のあらゆる場面においてハッピーな瞬間とさわやかさをお届けするトータル・ビバレッジ・カンパニーとして、お客さま、お得意さま、株主さま、地域社会、社員に対して、持続的に高水準の付加価値を提供してまいります。 |
| 資本金 | 1億円 |
| 企業サイトURL | https://www.ccbji.co.jp/ |