Senior DevSecOps & Platform Engineer
Mauer Ventures
Rad od kuće
27.08.2026.
About the Role
As a Senior DevSecOps & Platform Engineer, your primary responsibility will be to ensure that production infrastructure is:
- Reliable
- Secure
- Scalable
- Observable
- Maintainable
You will:
- Report directly to the VP of Product & Engineering
- Work closely with Software Engineers, AI & ML Engineers, Product Managers, Security stakeholders, and other departments
- Assess existing infrastructure and identify weaknesses and opportunities
- Establish clear technical direction
- Lead infrastructure initiatives from proposal through implementation and operational handover
- Recognise problems independently, ask the right questions, develop pragmatic proposals, establish priorities, and drive projects to completion
- Provide technical direction and establish infrastructure standards
- Review infrastructure decisions and support other engineers working in this area
- You will not have direct people-management responsibility
Responsibilities
1. Infrastructure Reliability & Operations
- Take technical ownership of the reliability and operational health of production infrastructure
- Audit the current cloud and Kubernetes environment and create a prioritised improvement roadmap
- Define and improve service reliability practices, including:
- Availability targets
- SLIs
- SLOs
- Alerting
- Incident response
- Operational readiness
- Establish monitoring, observability, logging, and alerting across infrastructure and application services
- Improve the use of:
- AWS CloudWatch
- Sentry
- Related observability tooling
- Lead investigation and resolution of complex:
- Infrastructure incidents
- Networking incidents
- Performance issues
- Production incidents
- Establish incident-management, escalation, and post-incident review processes
- Participate in the operational on-call setup and improve its effectiveness
- Define and regularly validate:
- Backup procedures
- Restore procedures
- Disaster recovery
- Business continuity
- Ensure production systems can be recovered within clearly understood and tested recovery objectives
- Identify reliability risks before they become production incidents
2. Cloud & Infrastructure Architecture
- Set the technical direction for cloud infrastructure, primarily within AWS
- Design and improve reliable, secure, and cost-conscious infrastructure architectures
- Lead the architecture and operation of the production Kubernetes environment on Amazon EKS
- Review and improve:
- Kubernetes cluster design
- Workloads
- Networking
- Ingress
- Autoscaling
- Resource allocation
- Deployment strategies
- Operational security
- Establish architecture patterns for:
- Microservices
- Event-driven systems
- APIs
- Storage
- Databases
- Asynchronous processing
- AWS Services
Hands-on knowledge is expected around:
- Amazon EKS
- Amazon RDS
- Amazon S3
- Amazon EventBridge
- Amazon Bedrock
- Amazon OpenSearch
- API Gateway
- Route 53
- IAM
- AWS Organizations
You should also have a deep understanding of:
- DNS
- Networking
- Storage
- Databases
- Certificates
- Load balancing
- Backups
- Restore procedures
- Service-to-service communication
Additional Cloud Experience
- Support and review a smaller GCP environment
- Help establish appropriate standards across cloud providers
- Evaluate new infrastructure technologies based on operational and business value
- Avoid unnecessary complexity and select solutions appropriate to the current scale and expected growth
3. Infrastructure as Code & Automation
- Own and improve Infrastructure as Code practices using Terraform
- Ensure infrastructure changes are:
- Reviewed
- Repeatable
- Testable
- Documented
- Deployed through controlled processes
- Improve the structure, maintainability, security, and state management of Terraform repositories
- Reduce manual infrastructure operations through automation
- Build reliable automation and operational tooling using Bash and Python, where appropriate
- Establish reusable infrastructure modules and platform capabilities
- Enable engineering teams to work independently without bypassing security or operational standards
- Identify recurring manual work and replace it with safe, maintainable automation
4. CI/CD & Developer Experience
- Own the technical direction and continuous improvement of the CI/CD environment using GitHub Actions
- Improve:
- Build processes
- Testing
- Deployments
- Rollbacks
- Release processes
- Reduce deployment time and increase deployment reliability
- Support Software Engineers in creating effective local development and testing environments
- Develop:
- Scripts
- Tooling
- Templates
- Reusable workflows
- Integrate automated:
- Tests
- Infrastructure validation
- Security checks
- Dependency checks
- Quality controls
- Improve deployment strategies, including:
- Rolling deployments
- Health checks
- Rollback mechanisms
- Blue-green deployments
- Canary deployments
- Evaluate AI-supported engineering practices such as automated code-review assistance
- Treat developer experience as a platform responsibility rather than requiring every engineering team to solve the same infrastructure problems independently
5. Security & Access Management
- Continuously improve infrastructure and development-environment security
- Establish and maintain:
- Secure cloud configurations
- Network boundaries
- Secrets management
- Encryption
- Logging
- Vulnerability management
- Own technical implementation of least-privilege access across:
- AWS
- Kubernetes
- GitHub
- Databases
- Related systems
- Improve identity and access management across AWS Organizations and organisational units
- Define secure and scalable processes for:
- Granting access
- Reviewing access
- Removing access
- Establish controlled, documented, and auditable access models
- Ensure infrastructure access is:
- Appropriate
- Time-bound where necessary
- Traceable
- Regularly reviewed
- Review architecture and infrastructure changes from a security perspective
- Improve vulnerability scanning for:
- Containers
- Dependencies
- Images
- Kubernetes
- Infrastructure
- Coordinate penetration tests with external providers
- Lead remediation of penetration-test findings
- Translate security findings into clear:
- Priorities
- Owners
- Implementation plans
- Verify that remediations are completed and effective
- Support engineering teams in following secure development and operational practices
6. ISO 27001 & Audit Readiness
- Support the ongoing operation and continuous improvement of the ISO 27001-compliant Information Security Management System (ISMS)
- Implement and maintain technical controls required by the ISMS
- Ensure appropriate documentation and implementation of:
- Infrastructure controls
- Access management
- Backup processes
- Monitoring
- Vulnerability management
- Change management
- Incident management
- Collect and maintain technical evidence for internal and external audits
- Support:
- Risk assessments
- Control testing
- Audit preparation
- Remediation activities
- Ensure written policies reflect actual engineering and operational practices
- Identify gaps between documented controls and the real technical environment
- Lead remediation of those gaps
- Support customer security reviews and technical security questionnaires
7. AI & ML Infrastructure Enablement
- Provide secure, reliable, and scalable infrastructure for AI and ML Engineers
- Support AI-related services in AWS, including:
- Amazon Bedrock
- OpenSearch
- Establish appropriate:
- Access controls
- Network boundaries
- Secrets management
- Logging
- Monitoring
- Cost controls
- Help AI and ML Engineers move services from experimentation into reliable production environments
- Provide reusable deployment patterns and infrastructure capabilities for AI services
- Ensure AI workloads follow the same:
- Reliability
- Security
- Observability
- Operational standards
- Evaluate infrastructure requirements created by emerging AI technologies
- Propose pragmatic solutions
- You are not expected to develop machine-learning models, but should understand the operational requirements of production AI systems and provide the infrastructure needed for ML Engineers to work effectively
Technical Leadership & Collaboration
Act as the senior technical authority for:
- Infrastructure reliability
- Cloud architecture
- DevSecOps
- Platform engineering
Create structured proposals covering:
- Problem
- Options
- Trade-offs
- Recommendation
- Implementation plan
- Risks
- Expected outcomes
Lead infrastructure initiatives end-to-end:
- Discovery
- Architecture
- Implementation
- Rollout
- Documentation
- Operational ownership
Additional responsibilities include:
- Break larger initiatives into concrete tasks
- Delegate appropriate implementation work to DevOps or Software Engineers
- Review implementation quality and ensure delegated work contributes to a coherent architecture
- Establish practical engineering standards without unnecessary bureaucracy
- Communicate technical risks and trade-offs clearly to technical and non-technical stakeholders
- Build effective relationships with:
- Software Engineers
- AI & ML Engineers
- Product Managers
- Security stakeholders
- Other departments
- Provide constructive technical feedback
- Document:
- Infrastructure architecture
- Operational procedures
- Technical decisions
- Ownership
- Mentor colleagues through:
- Technical guidance
- Reviews
- Pairing
- Knowledge sharing
Required Experience
- 7+ years of professional experience
- Several years of hands-on experience operating business-critical production infrastructure
- Deep practical experience with AWS and strong understanding of AWS architecture and security principles
- Strong production experience with Kubernetes, preferably Amazon EKS
- Strong experience designing and operating highly available, scalable, and secure cloud infrastructure
- Strong experience with Terraform and mature Infrastructure as Code practices
- Strong experience building and maintaining CI/CD pipelines, preferably with GitHub Actions
- Strong Linux and systems-engineering knowledge
- Strong Bash scripting skills
- Professional software-development or automation experience
- Experience designing and supporting microservice and event-driven architectures
- Experience managing cloud identity, permissions, roles, and least-privilege access
- Experience improving infrastructure security and coordinating vulnerability remediation
- Experience coordinating penetration tests and managing findings through remediation and verification
- Experience supporting security or compliance programmes such as ISO 27001
- Experience creating technical documentation and architecture proposals
- Experience leading complex technical initiatives without formal people-management authority
- Professional working proficiency in English
How You Work
- Takes ownership and proactively solves problems
- Brings structure to complex and ambiguous situations
- Identifies risks and operational gaps early
- Works independently while collaborating effectively with others
- Communicates clearly, directly, and constructively
- Balances security and reliability with delivery speed and business needs
- Makes pragmatic, maintainable decisions without unnecessary complexity
- Supports, mentors, and shares knowledge with other engineers
- Stays open to new technologies without blindly following trends
Upoznaj kompaniju
Posle 20 godina ukupnog operativnog i konsultantskog iskustva u izgradnji digitalnih proizvoda za startape i kompanije u Nemačkoj i Evropi, odlučili smo da iskoristimo naše kombinovano iskustvo i pronađemo brzo-rastuće digitalne startape u Berlinu i da ih spajamo sa savršenim kandidatima u Srbiji.
Težimo da budemo transparentni i jasni prema našim klijentima i kandidatima kako u vezi sa kadrovskim potrebama, tako i kompenzacijama – sa idejom da gradimo dugoročne odnose zasnovane na poverenju, konkurentnim platama i stručnosti. Naš krajnji cilj je da napravimo digitalni most između Beograda iBerlina koji bi na kraju dana bio koristan za sve
Preporuke se učitavaju...