Site Reliability Engineering (SRE)
Quick Summary
Fyld is hiring: Site Reliability Engineer (SRE) At Fyld, we believe the future is built with people, technology, and strong values.
At Fyld, we believe the future is built with people, technology, and strong values.
We are a Portuguese consulting company that operates with transparency, respect, and a focus on everyone’s growth.
Here, every project is an opportunity to build reliable, scalable and resilient technology solutions.
Our philosophy? Code for Big Solutions — because we don’t just build technology to make things work, we build systems that remain reliable as they scale.
We’re looking for an experienced Site Reliability Engineer (SRE) with strong technical skills and a passion for reliability, automation and operational excellence.
By joining Fyld, you’ll be part of a team where every member is challenged to innovate, collaborate and grow.
- Bachelor's degree in Computer Science, Software Engineering, Information Technology, Engineering or a related field
- Proven experience as an SRE, DevOps Engineer, Platform Engineer or similar role
- Strong understanding of site reliability engineering principles, including availability, reliability, scalability and resilience
- Experience defining and working with SLIs, SLOs, SLAs and error budgets
- Hands-on experience with observability, including metrics, logs, traces and alerting
- Experience with tools such as Prometheus, Grafana, OpenTelemetry, ELK or equivalent observability platforms
- Strong experience with cloud platforms, such as AWS, Azure or GCP
- Experience with Kubernetes and Docker in production environments
- Strong knowledge of Infrastructure as Code, using Terraform, OpenTofu, Ansible or similar technologies
- Strong automation and scripting skills using Python, Go, Bash, PowerShell or similar
- Experience with incident management, troubleshooting and root cause analysis
- Experience conducting post-incident reviews and implementing preventive measures
- Knowledge of performance monitoring, capacity planning and system optimisation
- Understanding of networking fundamentals, including TCP/IP, DNS, load balancing and HTTP
- Familiarity with CI/CD and modern software delivery practices
- Strong focus on automation and reducing manual operational work
- Ability to work closely with software engineering, infrastructure, security and product teams
- Strong analytical and problem-solving skills
- Fluency in English
Nice to Have
~1 min read- Experience operating large-scale or business-critical production environments
- Knowledge of GitOps and tools such as Argo CD or Flux
- Experience with service meshes or distributed systems
- Knowledge of chaos engineering and resilience testing
- Experience with Kubernetes platform engineering
- Familiarity with cloud cost optimisation and FinOps
- Relevant certifications such as CKA, AWS Certified DevOps Engineer or Google Cloud Professional Cloud DevOps Engineer
- Critical thinking and autonomy
- Focus on meaningful solutions and real results
- Respect, transparency, and continuous growth
- A collaborative team that values individual progress
- Access to continuous training and certifications
- A culture of proximity, respect, and recognition
- Challenging, purpose-driven projects where your talent as a Site Reliability Engineer has real impact
📩 Want to be part of it? Send your CV to join@fyld.pt
Location & Eligibility
Listing Details
- Posted
- September 24, 2026
- First seen
- September 26, 2026
- Last seen
- September 26, 2026
Posting Health
- Days active
- 0
- Repost count
- 0
- Trust Level
- 51%
- Scored at
- September 26, 2026
Signal breakdown
Similar Site Reliability Engineering jobs
View all →Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.