Lead Software Engineer, DevOps Platform (Bangkok based, relocation provided)

💰 $8,960 - $14,336 (Est.) 📍 Los Angeles

Job Description

About Agoda

At Agoda, we bridge the world through travel. Our story began in 2005, when two lifelong friends and entrepreneurs, driven by their passion for travel, launched Agoda to make it easier for everyone to explore the world.

Today, we are part of Booking Holdings [NASDAQ: BKNG], with a diverse team of over 7,000 people from 90 countries, working together in offices around the globe. Every day, we connect people to destinations and experiences, with our great deals across our millions of hotels and holiday properties, flights, and experiences worldwide.

No two days are the same at Agoda. Data and technology are at the heart of our culture, fueling our curiosity and innovation. If you’re ready to begin your best journey and help build travel for the world, join us.

This position is based in Bangkok, Thailand. (Relocation support is provided)

In this Role, you'll get to:
• Lead the technical vision, architecture, and execution of new SRE platforms or reliability initiatives.
• Define and promote SRE best practices across Agoda’s services e.g., SLI/SLO-driven engineering, error budgets, and other data-driven reliability factors.
• Design, build, and operate reliability platforms including load shedding , business signals monitoring, and safe-deployment automation to reduce blast radius while preserving developer velocity.
• Own safe deployment strategies such as canary releases, automated rollback, and business-impact protection integrated with deployment & monitoring.
• Proactively identify and mitigate reliability and scaling risks across Agoda’s services.
• Improve system resilience and multi-cluster readiness by partnering with platform team and operation team.
• Lead major incident response and operational excellence, driving fast detection, mitigation, root cause analysis, postmortems, and learnings focused on business impact.
• Maintain and evolve incident, observability, alerting, and on-call tooling, improving signal quality, alert enrichment, grouping, and reducing time-to-clue and time-to-mitigation for NOC and on-call engineers.
• Advance platform observability and reliability signals using Prometheus and Grafana, balancing actionability, scale, and cost efficiency.
• Define reliability roadmaps and OKRs, translating ambiguous business reliability goals into clear technical requirements.

What You’ll Need to Succeed:
• 8+ years of relevant experience.
• Demonstrated ownership of architecting, building, and operating mission-critical production systems, making long-term technical and reliability trade-off decisions.
• Proven ability to lead and coordinate complex cross-team initiatives, setting technical direction and aligning stakeholders to deliver outcomes at organizational scale.
• Expertise in one or more programming skills (e.g., Go, Python, Rust, Java) with a solid understanding of distributed systems fundamentals (concurrency, backpressure, timeouts/retries, idempotency, circuit breaking).
• Deep hands-on experience with the Kubernetes ecosystem, service mesh technologies (e.g., Istio), Kubernetes deployment workflows (e.g., Argo CD).
• Observability & monitoring expertise, using Prometheus, Grafana, and common logging/telemetry stacks (e.g., OpenTelemetry), with an understanding of signal quality, scalability, and cost trade-offs.
• Strong incident management lifecycle aiming for improving area of alert quality, alert management, incident response, RCA, and postmortems.
• Experience with reliability engineering patterns such as canary deployments, automated rollback, capacity/right-sizing automation, and production operation.
• Solid data analysis, including SQL(e.g., PostgreSQL, MSSQL) and data pipelines.
• Data-driven mindset, able to perform deep research, analyze complex problems, and make informed technical decisions.
• Excellent communication and collaboration skills, able to explain complex technical concepts clearly to stakeholders at all levels, and to operate effectively both as a self-directed individual contributor and as part of a team.
• Curiosity and continuous learning, staying current with industry trends, open-source advancements, and emerging reliability practices.

Nice-to-Have:
• Experience operating large-scale, high-QPS systems serving millions of users in domains such as e-commerce, travel, or fintech.
• Hands-on experience with multi-region / multi-DC architectures and traffic isolation or failover strategies.
• Background in chaos engineering and resilience testing.
• Experience defining or scaling org-wide SLO/SRE frameworks.
• Built or operated Kubernetes controllers/operators.
• Exposure to ML-assisted detection or statistical methods for signal tuning (e.g., windowing strategies, precision/recall trade-offs).

#Bengaluru #SãoPaulo #Delhi #NewYorkCity #Nigeria #London #Hyderabad #Pune #Mumbai #Colombia #Paris #Jakarta #Chennai #SanFrancisco #WashingtonDC #Toronto #Pakistan #LosAngeles #Dallas #Chicago #Kenya #Boston #Shanghai #Egypt #BuenosAires #Manila #Netherlands #Singapore #RiodeJaneiro #Beijing #Atlanta #Sydney #Madrid #Vietnam #SaudiArabia #Peru #Melbourne #Ireland #Russia #Bangladesh #MexicoCity #Philadelphia #Chile #SeattleArea #Noida #Kolkata #Guangdong #UnitedArabEmirates #TelAvivDistrict #Houston #KualaLumpur #BeloHorizonte #SouthKorea #Bangkok #Istanbul #Austin #Curitiba #Warsaw #Campinas #Barcelona #Ukraine #CostaRica #Berlin #Romania #Denver #Johannesburg #Minneapolis #Manchester #Miami #Phoenix #Detroit #Coimbatore #Milan #PortoAlegre #Vancouver #Montreal #Charlotte #SanDiego #Ghana #SaltLakeCity #Raleigh #HongKong #Munich #Prague #Ecuador #TampaBay #Tokyo #Serbia #Taipei #Cracow #Zhejiang #CapeTown #Brasilia #Columbus #Ahmedabad #Indore #Kochi #Gurgaon #Chandigarh #Lucknow #Bhubaneswar #Thiruvananthapuram #Visakhapatnam #Bhopal #JerseyCity #Irving #Denton #Worcester #Arlington #OverlandPark #AuroraDistrict #Baltimore #Tampa #Halethorpe #Dayton #Syracuse #Chonburi #ChiangMai #NakhonRatchasima #KhonKaen #HatYai #Phuket #Surabaya #Tangerang #Birmingham #Casablanca #Rabat #Camp #PetalingJaya #GeorgeTown

Please review our Hiring Process Guidelines before your interview — click here to learn how interviewing at Agoda works.

Discover More About Working At Agoda
• Agoda Careers https://careersatagoda.com
• Facebook https://www.facebook.com/agodacareers/
• LinkedIn https://www.linkedin.com/company/agoda
• YouTube https://www.youtube.com/agodalife

Equal Opportunity Employer

At Agoda, we pride ourselves on being a company represented by people of all different backgrounds and orientations. We prioritize attracting diverse talent and cultivating an inclusive environment that encourages collaboration and innovation. Employment at Agoda is based solely on a person’s merit and qualifications. We are committed to providing equal employment opportunity regardless of sex, age, race, color, national origin, religion, marital status, pregnancy, ****** orientation, gender identity, disability, citizenship, veteran or military status, and other legally protected characteristics.

We will keep your application on file so that we can consider you for future vacancies and you can always ask to have your details removed from the file. For more details please read our privacy policy.

Disclaimer

We do not accept any terms or conditions, nor do we recognize any agency’s representation of a candidate, from unsolicited third-party or agency submissions. If we receive unsolicited or speculative CVs, we reserve the right to contact and hire the candidate directly without any obligation to pay a recruitment fee.

💡 Quick Summary

Seeking a career-building opportunity? The Lead Software Engineer, DevOps Platform (Bangkok based, relocation provided) position is now open for candidates interested in the IT Engineer & Developer Jobs sector. This role in Los Angeles offers a professional environment and growth potential.

Requirement Snapshot: Candidates should possess basic communication skills, a proactive attitude, and the ability to work in a team. Experience in IT Engineer & Developer Jobs is a plus.

Sponsored

Job Details

Company Name: Agoda

Frequently Asked Questions

Click the Apply Now button on this page, login or register for free on CallCenterJob.co.in, fill in your name, mobile number, city, and experience, then submit your application. The recruiter will contact you directly.
The expected salary for Lead Software Engineer, DevOps Platform (Bangkok based, relocation provided) in Los Angeles is $8,960 - $14,336 (Est.) per month. Actual compensation may vary based on experience and negotiation.
No, Lead Software Engineer, DevOps Platform (Bangkok based, relocation provided) is an on-site position based in Los Angeles. Candidates must be able to commute or relocate to this location.
Basic communication skills, a proactive attitude, and the ability to work in a team are required for Lead Software Engineer, DevOps Platform (Bangkok based, relocation provided). Previous experience in IT Engineer & Developer Jobs is a plus. Freshers may also apply depending on the employer's requirements.
Yes, CallCenterJob.co.in is completely free for job seekers. Never pay money to apply for any job. If anyone asks for payment to process your application, report it immediately using the "Report this Job" button.

Similar Openings

  • Network Engineer

    Marriott International, Inc Building Maintenance Engineer Marriott International, Inc • La Puente, CA, United States • via EliteHub Jobs 11 hours ago Full–time No Degree Mentioned Apply on EliteHub Jobs Job highlights Identified by Google from the or...

    Full Time / Part Time

    Salary Estimated: 17K to 30K

    Los Angeles, California

    August 4, 2026


    Apply Now

  • Oracle Technical Architect -OICS

    Title : Oracle Technical Architect Location: Atlanta GA (Onsite) Position Type : Contract JOB DESCRIPTION Role : Oracle Integration Architect -OICS Location : Atlanta GA (Onsite) Duration:12+ Months Job Description: The Oracle Technical Architect wil...

    Full Time / Part Time

    Salary Estimated: 25K to 30K

    Atlanta, Georgia

    August 4, 2026


    Apply Now

  • Senior Systems Engineer, User Support

    With a company culture rooted in collaboration, expertise and innovation, we aim to promote progress and inspire our clients, employees, investors and communities to achieve their greatest potential. Our work is the catalyst that helps others achieve...

    Full Time / Part Time

    Salary Estimated: 17K to 33K

    Glasgow, Scotland

    August 4, 2026


    Apply Now

  • Senior Product Development Engineer

    Company Description Nyra Medical, located in Atlanta, Georgia, is a clinical-stage medical company dedicated to transforming structural heart care. The company’s mission is to enhance and extend the lives of patients with valvular heart disease throu...

    Full Time / Part Time

    Salary Estimated: 25K to 33K

    Atlanta, Georgia

    August 4, 2026


    Apply Now

  • Senior IBM Maximo Delivery Project Manager

    Introduction A career in IBM Consulting is built on long-term client relationships and close collaboration worldwide. You’ll work with leading companies across industries, helping them shape their hybrid cloud and AI journeys. With support from our s...

    Full Time / Part Time

    Salary Estimated: 24K to 28K

    Los Angeles, California

    August 4, 2026


    Apply Now

  • Senior HSM Engineer - General purpose and Payments

    About NCR VOYIX NCR Voyix Corporation (NYSE: VYX) is a global platform-powered leader in unified commerce for shopping and dining. Combining a flexible, intelligent platform with end-to-end payments capabilities and services developed through its dee...

    Full Time / Part Time

    Salary Estimated: 24K to 34K

    Atlanta, Georgia

    August 4, 2026


    Apply Now