Sample profilesDevOps and cloud platform engineeringIndia and global

Cloud engineers who understand the cost of running at scale

A rising infrastructure bill, recurring incidents and a fragile release process require different experience. These seven profiles cover production reliability, cloud cost engineering, container infrastructure and observability. The work ranges from operating a global consumer platform to maintaining the software that other engineering teams run.

3
India
2
USA and Canada
1
Europe and UK
1
Australia
Experience areas
Cloud cost · Production reliability · Container platforms · Observability
Regional coverage
3 India · 2 USA and Canada · 1 Europe and UK · 1 Australia

Illustrative profiles based on publicly documented professional work. Inclusion does not indicate availability or representation by Talhive.

Show Showing 7 of 7
Profile DO-01 Site reliability and cloud infrastructure India

Contributed to a payments platform's published 40 to 50% EC2 cost reduction

12yrs
Experience
Companies
Hosting, ride-hailing, creative softwarepayments, virtualisation, streaming
Current work
Senior Software Engineer, DevOps and SRE
Location
India
Selected work

Infrastructure experience across digital payments, creative software, virtualisation, mobility and connected television. The strongest public example is a production cost programme at an Indian payments infrastructure company.

  • Named alongside a colleague in the payments company's account of moving workloads onto spot capacity. The reported EC2 savings were 40 to 50% for the programme.
  • The engineering problem connects capacity economics with production operations: lower priced compute is useful only when the workload can tolerate interruptions and changing capacity.
  • His career provides relevant operating context across several product businesses; the published cost case provides a specific piece of work to examine.
40 to 50%
EC2 saving on the published spot programme
60%
of production running on spot capacity
3 yrs
as a senior SRE at a creative software provider
Education
Anna UniversityBE, Electronics and Communication
NIELITPG Diploma, Information Security and Cloud Computing
Career path
A connected television platform · Senior Software Engineer, DevOps and SRE
2024 to present
An enterprise virtualisation provider · Senior DevOps Engineer
2023 to 2024
An Indian payments platform · DevOps Tech Lead
2021 to 2023
A global creative software provider · Senior Site Reliability Engineer
2018 to 2021
A ride-hailing platform · Senior DevOps Engineer
2016 to 2018
Hosting and server technology firms · System Engineer
2014 to 2016
Career context

Career experience includes a global creative software provider, an Indian payments platform, an enterprise virtualisation provider and a connected television platform.

Relevant hiring mandate

A payments or SaaS team whose cloud spend needs attention alongside production reliability.

Profile DO-02 Site reliability engineer India

Four years at an enterprise infrastructure company, followed by reliability work in cybersecurity

8yrs
Experience
Companies
Software services, communicationsenterprise virtualisation, cybersecurity
Current work
Senior Site Reliability Engineer
Location
India
Selected work

A Pune based SRE with a public record that combines enterprise infrastructure experience with practical cloud native education. His professional GitHub identifies a global endpoint and cloud security company.

  • Publishes cloud native learning material and maintains public repositories that make his technical interests inspectable.
  • His community work includes CNCF and Docker participation, providing evidence of sustained engagement with container infrastructure and practitioner education.
  • The relevant career pattern is a substantial infrastructure tenure followed by another production focused environment, rather than a sequence of unrelated short assignments.
4 yrs
at an enterprise infrastructure company, promoted once
5
cloud native certifications
2
2026 conference stages, cloud and Kubernetes
Education
Pune Institute of Computer TechnologyBE, Computer Science
Dr Babasaheb Ambedkar Technological UniversityDiploma, Information Technology
Career path
A global cybersecurity company · Senior Site Reliability Engineer
2025 to present
An enterprise virtualisation provider · Site Reliability Engineer, two levels
2021 to 2025
A communications technology company · Senior Technical Associate
2019 to 2021
A software services firm · Software Engineer
2018 to 2019
Career context

Publicly documented a four year tenure at an enterprise virtualisation and cloud infrastructure business before the cybersecurity role.

Relevant hiring mandate

A platform team that needs reliability practice, operational documentation and engineers who can help colleagues understand the infrastructure.

Profile DO-03 Linux infrastructure engineer and maintainer India

Maintains the container operating system beneath production cloud workloads

13yrs
Experience
Companies
Developer assessment, enterprise open sourcecloud native infrastructure, global cloud provider
Current work
Senior Software Engineer, container operating systems
Location
India
Selected work

A Bengaluru based Linux engineer at a global cloud provider, with earlier experience at a cloud native infrastructure company. His public record centres on Flatcar Container Linux.

  • Publicly identified as a Flatcar maintainer, with work at the operating system layer rather than only application deployment.
  • Maintains an inspectable connection to the project's issue tracking and documentation through his professional GitHub profile.
  • His public technical activity addresses the maintenance and release lifecycle of container infrastructure: a different responsibility from configuring a single application's cloud account.
3
conference talks for the cloud provider on container Linux
5 yrs
at the global cloud provider
3.7 yrs
at an enterprise open source provider
Education
Dr B.C. Roy Engineering CollegeBTech, Computer Engineering
Career path
A global cloud provider · Software Engineer II, then Senior Software Engineer
2021 to present
A cloud native infrastructure company · Linux Software Engineer
2020 to 2021
An enterprise open source provider · Senior Software Engineer, infrastructure
2015 to 2019
A developer assessment platform · Software Developer
2013 to 2015
Career context

His current GitHub biography identifies the cloud provider, the acquired infrastructure business and his ongoing container operating system work.

Relevant hiring mandate

An infrastructure product business or platform team that needs Linux, container operating system and maintenance expertise.

Profile DO-04 Staff reliability engineer USA

Coauthored research on automatically running chaos experiments in production

Since 2006
Professional history documented
Companies
University research and faculty rolesstreaming, e-commerce, now an accommodation marketplace
Current work
Staff engineer, reliability
Location
USA
Selected work

Reliability experience at a global accommodation marketplace and a global video streaming platform. His public work connects distributed system failures with the practical decisions engineering teams make during incidents.

  • Coauthored "Automating chaos experiments in production," describing a platform that generated and executed experiments against production systems.
  • The work examined whether systems continued to behave acceptably when components slowed down or failed, moving resilience testing beyond assumptions made during design.
  • His long running technical writing covers incident analysis, complex system failures and the interaction between people, software and operational work.
3,067
GitHub stars on his resilience engineering reading collection
70+
publications listed on his own page
4
books on chaos engineering, automation and operations
Education
McGill UniversityBEng, Computer Engineering
Boston UniversityMS, Electrical Engineering
University of MarylandPhD, Computer Science
Career path
A global accommodation marketplace · Staff software engineer, reliability
February 2024 to present
A large consumer ecommerce platform · Senior staff software engineer
July 2023 to February 2024
A global streaming platform · Senior software engineer
June 2015 to July 2023
A cloud email delivery platform · Senior software engineer
January 2014 to June 2015
Cloud services for technical computing · Lead architect
January 2012 to January 2014
A university research institute · Computer scientist
dates not established
A public research university · Assistant professor, computer science
August 2006 to June 2008
Career context

Professional profile lists staff reliability work at the accommodation marketplace. Earlier research documents his work on the streaming company's chaos engineering team.

Relevant hiring mandate

An organisation improving resilience testing and incident learning across distributed services.

Profile DO-05 Senior software engineering tech lead UK

Cocreated Thanos, then moved into globally managed observability

14yrs
Experience
Companies
Semiconductors, distributed simulationenterprise open source, now a global cloud platform
Current work
Senior software engineering tech lead, managed observability
Location
UK
Selected work

A London based infrastructure engineer whose career spans a semiconductor company, a distributed simulation platform, an enterprise open source provider and a global cloud platform.

  • Cocreated Thanos with another engineer in 2017, addressing the challenge of scaling Prometheus based monitoring.
  • Became a Prometheus maintainer and later worked on the managed Prometheus offering of a global cloud provider.
  • Authored a book on efficient Go programming and contributes to profiling, instrumentation and observability work across company and open source environments.
14.2K
GitHub stars on the monitoring system he cocreated
66K
stars on the metrics project he maintains
1
book on efficient Go programming
Education
Gdansk University of TechnologyMaster's, Informatics: Distributed Applications and Internet Services, 2015 to 2017
Career path
A global cloud platform · Senior software engineering tech lead
2022 to present
An enterprise open source provider · Principal software engineer
2019 to 2022
A distributed simulation platform · Infrastructure engineer
about 2015 to 2019
A semiconductor company · Software engineer
2012 to about 2015
Career context

His biography records a three year early tenure in semiconductors, infrastructure work in simulation, a principal engineering role in open source and a subsequent cloud tech lead role.

Relevant hiring mandate

A business building or operating an observability platform, where distributed systems and software efficiency are central.

Profile DO-06 Infrastructure engineer and independent systems practitioner Canada

Published the operational lessons behind a payments company's Kubernetes rollout

15yrs
Experience
Companies
Montreal software rolesa global payments platform, now independent
Current work
Independent systems education and software
Location
Canada
Selected work

A Montreal based software developer with earlier infrastructure experience at a global payments platform. Her public work makes the implementation and failure modes of production infrastructure unusually visible.

  • Authored the payments company's account of introducing Kubernetes for a distributed cron scheduling system.
  • The account covers reading controller code, load testing and prioritising etcd reliability before depending on the platform in production.
  • Her wider body of work examines Linux, networking, DNS and debugging, connecting operational decisions with the underlying systems rather than treating infrastructure as a collection of configuration files.
4,000
copies sold of a printed Linux toolbox box set
117
countries buying from her publishing business
15
paid technical zines published
Education
McGill UniversityMSc, Computer Science
Career path
Own publishing and education business · Founder, full time
2019 to present
A global payments platform · Software engineer, infrastructure
2014 to 2019
Montreal software companies · Programmer and data work
from about 2011
Career context

Her current work is independent systems education and software. The payments infrastructure case is historical professional experience.

Relevant hiring mandate

A benchmark for deep systems understanding, reliability investigation and clear technical communication.

Profile DO-07 Principal reliability and observability specialist Australia

Two decades of production reliability, from a global cloud platform to developer facing observability

20+ yrs
Stated in her professional biography
Companies
Global cloud and internet platformnow an observability software provider
Current work
Technical fellow, reliability and observability
Location
Australia
Selected work

An observability software provider, following reliability and staff engineering at a global cloud and internet platform. Based in Sydney, with a second base in Vancouver.

  • A published conference presentation examines moving production workloads to ARM infrastructure and the operational trade offs involved.
  • Earlier work at a global cloud and internet platform covered load balancing and travel search.
  • The public engineering record connects infrastructure choices with compute cost and latency, which is what a team weighing platform cost against performance has to reason about.
20+ yrs
in production reliability and infrastructure, as stated in her professional biography
11 yrs
at a global cloud and internet platform
Career path
An observability software provider · Technical fellow
2026 to present
An observability software provider · Field CTO
2022 to 2025
An observability software provider · Principal developer advocate
2019 to 2022
A global cloud and internet platform · Reliability and staff engineering
2008 to 2019
Career context

Her own career record documents reliability and staff engineering at a global cloud and internet platform from 2008, then developer advocacy, a field CTO role and a technical fellowship at an observability software provider. The same record notes a dual Sydney and Vancouver base.

Relevant hiring mandate

A team weighing cloud platform cost against production reliability, where observability and infrastructure economics have to be argued together.

Compare the experience

Seven examples, side by side.

Each row links to the full example. Regions follow the coverage split described above.

ProfileRegionResponsibility
DO-01Site reliability and cloud infrastructureIndiaSite reliability and cloud infrastructure
DO-02Site reliability engineerIndiaSite reliability engineer
DO-03Linux infrastructure engineer and maintainerIndiaLinux infrastructure engineer and maintainer
DO-04Staff reliability engineerUSAStaff reliability engineer
DO-05Senior software engineering tech leadUKSenior software engineering tech lead
DO-06Infrastructure engineer and independent systems practitionerCanadaInfrastructure engineer and independent systems practitioner
DO-07Principal reliability and observability specialistAustraliaPrincipal reliability and observability specialist
Before the search

Start with the responsibility

A useful search brief explains what the person must change: a system that needs to scale, a product that needs direction or an organisation that needs stronger leadership. Share that context with Talhive so the discussion can focus on the experience your hire needs.

Cloud cost
Production reliability
Container platforms
Observability
Bring these four inputs to the conversation

What makes the conversation useful.

Discuss your cloud platform hire. Share the business problem, expected responsibility and location requirements. Include a profile ID if a particular example reflects the experience you need.

01
Business problem

What needs to work differently after this hire?

02
Responsibility

What will the person own, and who will they work with?

03
Practical requirements

Location, working arrangement, compensation range and timing.

04
Relevant experience

Which example best reflects the systems, customers or organisation involved?

Questions

Before you get in touch.

How to read these examples, and what to send us.

Start with the responsibility. Deployment automation points towards DevOps; availability and incident learning towards SRE; reusable infrastructure and developer services towards platform engineering. Some people span these areas, but the mandate should identify the primary problem.
These are examples of professional experience. Availability and interest are established separately for an active search.
Yes. Share the profile ID and the part of the experience that matters to your business.
The company provides context. The stronger match depends on what the person owned, the resources they had and the decisions your role will require.
The role, business context, responsibilities, location requirements and compensation range. A current job description can provide a starting point.
Tell us about the hire

Tell us about the hire.

Share the business problem, the responsibility and where the role sits. If one of the examples above reflects the experience you need, include its profile ID.

A profile ID tells us which experience matters to you. It is not a request to interview that individual.
Prefer email? search@talhive.com
Tell us about the hire

Share the role and the business problem. A Talhive partner will reply to the requirement you describe.

Please enter your name.

Please enter a valid work email.

Relevant sample profile (optional)

Your details go to a Talhive partner and nowhere else. Privacy policy.

Thank you. Your hiring requirement has been submitted.

A Talhive partner will follow up on the requirement you described.

Do we need a DevOps engineer, an SRE or a platform engineer?

Start with the responsibility. Deployment automation points towards DevOps; availability and incident learning towards SRE; reusable infrastructure and developer services towards platform engineering. Some people span these areas, but the mandate should identify the primary problem.