Advance company-wide AI governance and data use, build Snowflake and integration platforms, and maintain multi-cloud infrastructure for internal business systems.
About this role
Improve the reliability, availability and operational quality of HULFT Square’s AWS microservices. Work includes monitoring and alert design with Datadog, log and metric analysis, CI/CD and release improvements using GitHub Actions and ArgoCD, incident investigation, post-mortems and prevention. The SRE also reviews capacity, performance and security and works with development and support teams to turn operational issues into product improvements.
Applicants need two years of operations or infrastructure work, or application development, for web services, SaaS or business systems; basic cloud understanding or practical cloud experience; and experience in monitoring, logs, incidents or release operations. Previous SRE or DevOps employment is preferred rather than mandatory. The team mainly works in the office, allowing home working for individual circumstances.
This is an AI-generated summary of the employer's original posting — details can be incomplete, out of date or simply wrong. Always confirm everything on the official posting before applying.