[Hiring] DevOps Engineer – Cloud & Data Platform | Japan, Tokyo
Our client is a leading global technology and internet services company is expanding its data and advertising technology platforms to support large-scale digital services across multiple markets.
The organization develops and operates high-traffic platforms covering advertising delivery, tracking, reporting, data management, and marketing technologies. The engineering organization is internationally distributed, with team members across Asia, and works in a highly collaborative, data-driven environment.
We are looking for a proactive DevOps / SRE Engineer to strengthen the reliability, scalability, and automation of a large-scale data management platform.
You will be responsible for improving cloud infrastructure, deployment pipelines, observability, infrastructure automation, and operational efficiency. The role also provides opportunities to introduce modern AI-powered automation and workflow solutions to reduce manual operations and accelerate engineering productivity.
Responsibilities
- Cloud Infrastructure Management
- Manage and optimize Google Cloud infrastructure, including Cloud Run, BigQuery, Cloud Storage, Pub/Sub, and Cloud Build.
- Improve infrastructure reliability, scalability, and operational efficiency.
- Hybrid Cloud & Networking
- Manage network configurations, load balancers, VPNs, and dedicated connectivity between on-premise and cloud environments.
- Support complex hybrid infrastructure architectures.
- CI/CD & Deployment
- Design, maintain, and improve CI/CD pipelines supporting Java/Maven, Node.js, and Docker-based applications.
- Implement automated deployment strategies such as canary and blue-green deployments.
- Improve release reliability and deployment velocity.
- Observability & Cost Optimization
- Build and maintain monitoring and logging environments using tools such as Prometheus, Grafana, and ELK.
- Monitor infrastructure utilization and proactively identify opportunities for cloud cost optimization.
- Infrastructure as Code
- Automate infrastructure provisioning and configuration using Terraform and Ansible.
- Reduce manual operational work through standardized and reusable automation.
- Security & Compliance
- Manage secrets and credentials using solutions such as Vault and cloud-native secret management services.
- Apply security best practices across cloud and hybrid infrastructure.
- AI & Operational Automation
- Explore and implement AI-assisted operational workflows.
- Automate tasks such as incident handling, ticket processing, troubleshooting, and routine infrastructure operations using Python, workflow automation platforms, and LLM-based solutions.
Work Environment
- Agile and fast-paced engineering environment.
- International team with members located across multiple Asian countries.
- Open communication and a strong culture of continuous improvement.
- Engineers are encouraged to propose and implement improvements to existing processes and technologies.
Technology Stack
Cloud / Infrastructure
- Google Cloud Platform
- Private Cloud
- Kubernetes
- Docker
CI/CD
- Jenkins
- Octopus Deploy
- GitHub Actions
- Harbor
- Artifactory
Infrastructure as Code / Automation
- Terraform
- Ansible
- Python
- Bash
AI / Workflow Automation
- n8n
- LLM-based automation frameworks
Observability
- Prometheus
- Grafana
- ELK Stack
Collaboration / Development
- Jira
- Confluence
- Slack / Microsoft Teams
- GitHub Enterprise
- Bitbucket
Mandatory Qualifications
- More than 3 years of experience in DevOps, SRE, or Infrastructure Engineering. (preferably 5–7+ years of experience.)
- Strong hands-on experience with GCP (Cloud Run, BigQuery, IAM, VPC, Networking).
- Deep understanding of CI/CD pipelines and automated deployment orchestration.
- Proficiency in Infrastructure as Code (e.g., Terraform) and scripting (Bash/Python).
- Experience in managing production-grade Kubernetes/containerized environments.
- TOEIC Score 800 above or possess equivalent qualifications.
Preferred Qualifications
- Experience building AI-powered automation or agentic workflows (e.g., n8n, automated ticketing/remediation).
- Experience managing GCP Interconnect and complex hybrid network topologies.
- Proven track record in cloud cost management and resource optimization.
- Knowledge of Kafka, Redis, MariaDB, or Couchbase.
- Experience with workflow schedulers (e.g., Azkaban).
- Professional certification (Google Professional Cloud Architect/DevOps Engineer or similar).
Languages
- English: Fluent
- Japanese: Optional / a plus
Work Environment
Fast-paced, dynamic global environment with collaborative teams across multiple locations
Salary: ¥7M – ¥9M JPY per year (Middle Level)
Location: Hybrid (4 days in the office, 1 day remote)
Office Location: Tokyo, Japan
Working Hours: Flexible schedule with core hours from 11:00 AM to 3:00 PM
Visa Sponsorship: Available
※Japanese language proficiency certification (such as JLPT N2) is not required, as our client is a global organization with an international working environment.
Language Requirement: English only
Apply now or contact us for further information:
Aleksey.kim@tg-hr.com