Oglasi za posao Senior DevSecOps & Platform Engineer

Senior DevSecOps & Platform Engineer

Mauer Ventures

Rad od kuće

27.08.2026.

Linux Python AWS DevOps Bash Cloud Kubernetes senior

About the Role

As a Senior DevSecOps & Platform Engineer, your primary responsibility will be to ensure that production infrastructure is:

  • Reliable
  • Secure
  • Scalable
  • Observable
  • Maintainable

You will:

  • Report directly to the VP of Product & Engineering
  • Work closely with Software Engineers, AI & ML Engineers, Product Managers, Security stakeholders, and other departments
  • Assess existing infrastructure and identify weaknesses and opportunities
  • Establish clear technical direction
  • Lead infrastructure initiatives from proposal through implementation and operational handover
  • Recognise problems independently, ask the right questions, develop pragmatic proposals, establish priorities, and drive projects to completion
  • Provide technical direction and establish infrastructure standards
  • Review infrastructure decisions and support other engineers working in this area
  • You will not have direct people-management responsibility

Responsibilities

1. Infrastructure Reliability & Operations

  • Take technical ownership of the reliability and operational health of production infrastructure
  • Audit the current cloud and Kubernetes environment and create a prioritised improvement roadmap
  • Define and improve service reliability practices, including:
    • Availability targets
    • SLIs
    • SLOs
    • Alerting
    • Incident response
    • Operational readiness
  • Establish monitoring, observability, logging, and alerting across infrastructure and application services
  • Improve the use of:
    • AWS CloudWatch
    • Sentry
    • Related observability tooling
  • Lead investigation and resolution of complex:
    • Infrastructure incidents
    • Networking incidents
    • Performance issues
    • Production incidents
  • Establish incident-management, escalation, and post-incident review processes
  • Participate in the operational on-call setup and improve its effectiveness
  • Define and regularly validate:
    • Backup procedures
    • Restore procedures
    • Disaster recovery
    • Business continuity
  • Ensure production systems can be recovered within clearly understood and tested recovery objectives
  • Identify reliability risks before they become production incidents

2. Cloud & Infrastructure Architecture

  • Set the technical direction for cloud infrastructure, primarily within AWS
  • Design and improve reliable, secure, and cost-conscious infrastructure architectures
  • Lead the architecture and operation of the production Kubernetes environment on Amazon EKS
  • Review and improve:
    • Kubernetes cluster design
    • Workloads
    • Networking
    • Ingress
    • Autoscaling
    • Resource allocation
    • Deployment strategies
    • Operational security
  • Establish architecture patterns for:
    • Microservices
    • Event-driven systems
    • APIs
    • Storage
    • Databases
    • Asynchronous processing
    • AWS Services

Hands-on knowledge is expected around:

  • Amazon EKS
  • Amazon RDS
  • Amazon S3
  • Amazon EventBridge
  • Amazon Bedrock
  • Amazon OpenSearch
  • API Gateway
  • Route 53
  • IAM
  • AWS Organizations

You should also have a deep understanding of:

  • DNS
  • Networking
  • Storage
  • Databases
  • Certificates
  • Load balancing
  • Backups
  • Restore procedures
  • Service-to-service communication

Additional Cloud Experience

  • Support and review a smaller GCP environment
  • Help establish appropriate standards across cloud providers
  • Evaluate new infrastructure technologies based on operational and business value
  • Avoid unnecessary complexity and select solutions appropriate to the current scale and expected growth

3. Infrastructure as Code & Automation

  • Own and improve Infrastructure as Code practices using Terraform
  • Ensure infrastructure changes are:
    • Reviewed
    • Repeatable
    • Testable
    • Documented
    • Deployed through controlled processes
  • Improve the structure, maintainability, security, and state management of Terraform repositories
  • Reduce manual infrastructure operations through automation
  • Build reliable automation and operational tooling using Bash and Python, where appropriate
  • Establish reusable infrastructure modules and platform capabilities
  • Enable engineering teams to work independently without bypassing security or operational standards
  • Identify recurring manual work and replace it with safe, maintainable automation

4. CI/CD & Developer Experience

  • Own the technical direction and continuous improvement of the CI/CD environment using GitHub Actions
  • Improve:
    • Build processes
    • Testing
    • Deployments
    • Rollbacks
    • Release processes
  • Reduce deployment time and increase deployment reliability
  • Support Software Engineers in creating effective local development and testing environments
  • Develop:
    • Scripts
    • Tooling
    • Templates
    • Reusable workflows
  • Integrate automated:
    • Tests
    • Infrastructure validation
    • Security checks
    • Dependency checks
    • Quality controls
  • Improve deployment strategies, including:
    • Rolling deployments
    • Health checks
    • Rollback mechanisms
    • Blue-green deployments
    • Canary deployments
  • Evaluate AI-supported engineering practices such as automated code-review assistance
  • Treat developer experience as a platform responsibility rather than requiring every engineering team to solve the same infrastructure problems independently

5. Security & Access Management

  • Continuously improve infrastructure and development-environment security
  • Establish and maintain:
    • Secure cloud configurations
    • Network boundaries
    • Secrets management
    • Encryption
    • Logging
    • Vulnerability management
  • Own technical implementation of least-privilege access across:
    • AWS
    • Kubernetes
    • GitHub
    • Databases
    • Related systems
  • Improve identity and access management across AWS Organizations and organisational units
  • Define secure and scalable processes for:
    • Granting access
    • Reviewing access
    • Removing access
  • Establish controlled, documented, and auditable access models
  • Ensure infrastructure access is:
    • Appropriate
    • Time-bound where necessary
    • Traceable
    • Regularly reviewed
  • Review architecture and infrastructure changes from a security perspective
  • Improve vulnerability scanning for:
    • Containers
    • Dependencies
    • Images
    • Kubernetes
    • Infrastructure
  • Coordinate penetration tests with external providers
  • Lead remediation of penetration-test findings
  • Translate security findings into clear:
    • Priorities
    • Owners
    • Implementation plans
  • Verify that remediations are completed and effective
  • Support engineering teams in following secure development and operational practices

6. ISO 27001 & Audit Readiness

  • Support the ongoing operation and continuous improvement of the ISO 27001-compliant Information Security Management System (ISMS)
  • Implement and maintain technical controls required by the ISMS
  • Ensure appropriate documentation and implementation of:
    • Infrastructure controls
    • Access management
    • Backup processes
    • Monitoring
    • Vulnerability management
    • Change management
    • Incident management
  • Collect and maintain technical evidence for internal and external audits
  • Support:
    • Risk assessments
    • Control testing
    • Audit preparation
    • Remediation activities
  • Ensure written policies reflect actual engineering and operational practices
  • Identify gaps between documented controls and the real technical environment
  • Lead remediation of those gaps
  • Support customer security reviews and technical security questionnaires

7. AI & ML Infrastructure Enablement

  • Provide secure, reliable, and scalable infrastructure for AI and ML Engineers
  • Support AI-related services in AWS, including:
    • Amazon Bedrock
    • OpenSearch
  • Establish appropriate:
    • Access controls
    • Network boundaries
    • Secrets management
    • Logging
    • Monitoring
    • Cost controls
  • Help AI and ML Engineers move services from experimentation into reliable production environments
  • Provide reusable deployment patterns and infrastructure capabilities for AI services
  • Ensure AI workloads follow the same:
    • Reliability
    • Security
    • Observability
    • Operational standards
  • Evaluate infrastructure requirements created by emerging AI technologies
  • Propose pragmatic solutions
  • You are not expected to develop machine-learning models, but should understand the operational requirements of production AI systems and provide the infrastructure needed for ML Engineers to work effectively

Technical Leadership & Collaboration

Act as the senior technical authority for:

  • Infrastructure reliability
  • Cloud architecture
  • DevSecOps
  • Platform engineering

Create structured proposals covering:

  • Problem
  • Options
  • Trade-offs
  • Recommendation
  • Implementation plan
  • Risks
  • Expected outcomes

Lead infrastructure initiatives end-to-end:

  • Discovery
  • Architecture
  • Implementation
  • Rollout
  • Documentation
  • Operational ownership

Additional responsibilities include:

  • Break larger initiatives into concrete tasks
  • Delegate appropriate implementation work to DevOps or Software Engineers
  • Review implementation quality and ensure delegated work contributes to a coherent architecture
  • Establish practical engineering standards without unnecessary bureaucracy
  • Communicate technical risks and trade-offs clearly to technical and non-technical stakeholders
  • Build effective relationships with:
    • Software Engineers
    • AI & ML Engineers
    • Product Managers
    • Security stakeholders
    • Other departments
  • Provide constructive technical feedback
  • Document:
    • Infrastructure architecture
    • Operational procedures
    • Technical decisions
    • Ownership
  • Mentor colleagues through:
    • Technical guidance
    • Reviews
    • Pairing
    • Knowledge sharing

Required Experience

  • 7+ years of professional experience
  • Several years of hands-on experience operating business-critical production infrastructure
  • Deep practical experience with AWS and strong understanding of AWS architecture and security principles
  • Strong production experience with Kubernetes, preferably Amazon EKS
  • Strong experience designing and operating highly available, scalable, and secure cloud infrastructure
  • Strong experience with Terraform and mature Infrastructure as Code practices
  • Strong experience building and maintaining CI/CD pipelines, preferably with GitHub Actions
  • Strong Linux and systems-engineering knowledge
  • Strong Bash scripting skills
  • Professional software-development or automation experience
  • Experience designing and supporting microservice and event-driven architectures
  • Experience managing cloud identity, permissions, roles, and least-privilege access
  • Experience improving infrastructure security and coordinating vulnerability remediation
  • Experience coordinating penetration tests and managing findings through remediation and verification
  • Experience supporting security or compliance programmes such as ISO 27001
  • Experience creating technical documentation and architecture proposals
  • Experience leading complex technical initiatives without formal people-management authority
  • Professional working proficiency in English

How You Work

  • Takes ownership and proactively solves problems
  • Brings structure to complex and ambiguous situations
  • Identifies risks and operational gaps early
  • Works independently while collaborating effectively with others
  • Communicates clearly, directly, and constructively
  • Balances security and reliability with delivery speed and business needs
  • Makes pragmatic, maintainable decisions without unnecessary complexity
  • Supports, mentors, and shares knowledge with other engineers
  • Stays open to new technologies without blindly following trends

Upoznaj kompaniju

O Kompaniji

Posle 20 godina ukupnog operativnog i konsultantskog iskustva u izgradnji digitalnih proizvoda za startape i kompanije u Nemačkoj i Evropi, odlučili smo da iskoristimo naše kombinovano iskustvo i pronađemo brzo-rastuće digitalne startape u Berlinu i da ih spajamo sa savršenim kandidatima u Srbiji.

Težimo da budemo transparentni i jasni prema našim klijentima i kandidatima kako u vezi sa kadrovskim potrebama, tako i kompenzacijama – sa idejom da gradimo dugoročne odnose zasnovane na poverenju, konkurentnim platama i stručnosti. Naš krajnji cilj je da napravimo digitalni most između Beograda iBerlina koji bi na kraju dana bio koristan za sve

Preporuke se učitavaju...