As an Observability Expert, you will design and implement end-to-end monitoring, tracing and alerting solutions for critical applications at a leading company in the energy sector. Join a fully remote role at the intersection of platform reliability and business insight.
OBSERVABILITY AT SCALE | DATADOG & AWS | FULLY REMOTE | LEARNING & GROWTH |
Don't tick every box? If you meet around 70% of the requirements above, we'd still encourage you to apply.
ABOUT THE ROLE
We are looking for an Observability Expert to strengthen the monitoring and reliability capabilities of a leading energy sector company. In this role, you will instrument applications for functional and technical metrics, build distributed tracing and logging pipelines, and work closely with business stakeholders to define the SLIs and SLOs that matter most. You will bring both deep technical expertise in Datadog and AWS and the business acumen to translate reliability into meaningful KPIs and dashboards.
KEY RESPONSIBILITIES
● Instrument applications for both functional and technical metrics using Datadog APM, custom metrics and Distributed Tracing.
● Use logs as a source of metrics, leveraging CloudWatch and/or Datadog metric generation.
● Define SLIs and SLOs, guiding the definition of functional KPIs in collaboration with business stakeholders.
● Implement and define structured logging across applications and services.
● Design high-level business dashboards that translate technical telemetry into actionable insight.
● Instrument applications across multiple languages, primarily Python, with TypeScript and Java where required.
REQUIRED SKILLS & EXPERIENCE
● Hands-on experience with Datadog APM, Datadog custom metrics and Distributed Tracing.
● Experience instrumenting applications for both functional and technical metrics using Datadog.
● Experience using logs as a source of metrics (CloudWatch and/or Datadog metric generation).
● Strong understanding of SLI and SLO definitions, with the ability to guide the definition of functional KPIs in collaboration with business stakeholders.
● Solid AWS knowledge and practical experience.
● Ability to instrument applications using Python (mandatory), with TypeScript and Java as a plus.
● Strong expertise with the Datadog SDK.
● Experience implementing and defining structured logging.
● Ability to design high-level business dashboards.
NICE TO HAVE
● Knowledge of OpenTelemetry.
● Experience with Grafana (visualization layer).
● Experience with service-to-service traffic control and monitoring.
● Experience with Power BI (data visualization).
● Experience gathering business requirements related to SLO definition.
EDUCATION
● Bachelor's degree in Computer Science, Telecommunications Engineering or a related field.
PREFERRED CERTIFICATIONS
● AWS Certified Solutions Architect / Developer Associate.
● Datadog certifications, a plus.
WORKING MODEL
Fully remote position.
关于高知特 (Cognizant)
高知特(Cognizant)(纳斯达克代码:CTSH)作为一家AI Builder和相关技术服务提供商,致力于通过打造全栈AI解决方案,帮助企业将人工智能投资转化为实际价值。公司凭借深厚的行业经验、流程优化和工程技术专长,将企业独特的业务场景融入科技系统,赋能组织释放人才潜能,推动切实成果,并帮助全球企业在瞬息万变的环境中保持领先。如需了解更多详情,敬请访问 cognizant.ai 或关注@cognizant。
补充雇佣信息
薪酬信息截至本职位发布之日为准。Cognizant 保留在适用法律允许的范围内随时修改该信息的权利。
申请人可能需要通过现场面试或视频会议的方式参加面试。此外,候选人在每次面试时可能需要出示其当前所在州或政府签发的有效身份证件。
Cognizant 是一家提供平等就业机会的雇主。在招聘过程中,您的申请和候选资格不会因种族、肤色、性别、宗教、信仰、性取向、性别认同、国籍、残疾、遗传信息、怀孕、退伍军人身份或任何其他受联邦、州或地方法律保护的特征而受到影响。







