01.07.2026 aktualisiert

**** ******** ****
Premiumkunde
100 % verfügbar

Principal Infrastructure Engineer / SRE

Berlin, Mitte, Deutschland
Weltweit
Praxisorientierter Werdegang mit über 30 Jahren Markt- und Projekterfahrung
Berlin, Mitte, Deutschland
Weltweit
Praxisorientierter Werdegang mit über 30 Jahren Markt- und Projekterfahrung

Profilanlagen

GMS_BOFH.pdf

Über mich

Senior Infrastructure & Reliability Engineer mit 25+ Jahren Erfahrung in Enterprise Plattformen und sicherheitskritischen Umgebungen. Schwerpunkt auf Linux, Kubernetes und Cloud Plattformen. Ich stabilisiere komplexe Produktionssysteme und unterstütze beim Redesign moderner Plattformarchitekturen.

Skills

ArchitekturLinuxUmstrukturierungSicherheitsanforderungenUnternehmensstrukturenDevSecOps
Seit nahezu drei Jahrzehnten bewege ich mich im Spannungsfeld zwischen Betrieb, Architektur und Security unternehmenskritischer IT-Systeme. Ich habe Organisationen in Wachstums-, Krisen- und Restrukturierungsphasen begleitet – sowohl technisch als auch strukturell.

Mein Schwerpunkt liegt nicht nur in der Implementierung moderner Cloud- und DevSecOps-Ansätze, sondern vor allem in deren nachhaltiger Integration in bestehende Unternehmensstrukturen. Ich übernehme Verantwortung für komplexe Produktionsumgebungen, analysiere gewachsene Architekturen und führe sie in stabile, skalierbare und sichere Betriebsmodelle über.

Ich arbeite souverän in hochverfügbaren Umgebungen mit erhöhten Sicherheitsanforderungen und kenne sowohl die operative Tiefe als auch die strategische Perspektive. Dabei verbinde ich technisches Know-how mit Kommunikationsstärke – als Vermittler zwischen Entwicklung, Betrieb, Security und Management.

Besonders in Transformationsprojekten – Migrationen, Konsolidierungen, Cloud-Einführungen, Security-Härtungen oder organisatorischen Umstrukturierungen – agiere ich als Capability Enabler: Ich baue Strukturen auf, befähige Teams und schaffe belastbare Prozesse.

Meine Arbeitsweise ist analytisch, pragmatisch und lösungsorientiert. Ich bevorzuge klare Verantwortlichkeiten, saubere Architekturentscheidungen und nachhaltige Automatisierung statt kurzfristiger Workarounds.

30 Jahre Markterfahrung bedeuten für mich nicht nur technisches Wissen, sondern Urteilskraft, Priorisierung und Risikobewusstsein.

Sprachen

DeutschgutEnglischverhandlungssicherItalienischMuttersprache

Projekthistorie

DevOps TeamLead

Contensi Software GmbH

Internet und Informationstechnologie

  • Mentored and provided guidance to two junior DevOps colleagues to facilitate their professional development and enhance team capabilities
● Led the construction of a reference MinIO Storage cluster utilizing Rancher Kubernetes, ensuring robustness and scalability for storage infrastructure needs

Site Reliability & Security Engineer

Max Planck Institute

Internet und Informationstechnologie


● Implemented CI/CD pipelines utilizing GitLab for on-premises and cloud-based environments, including AWS and OpenStack VMs, with orchestration through Rancher
● Spearheaded the prototyping of a distributed storage system leveraging MinIO, alongside the integration of JuiceFS for enhanced file system capabilities
● Introduced PortainerIO to streamline operations and facilitate user-friendly management for a team of non-technical users

Site Reliability & Security Engineer

Skycharge

Internet und Informationstechnologie


● Developed a Debian package extension for an ANSI-C codebase, optimized for armh (BeagleBoard) architecture, enhancing functionality and compatibility
● Spearheaded the introduction of unit tests alongside the utilization of GDB for comprehensive code validation and quality assurance

Staff Site Reliability Engineer

Prima Assicurazioni

Versicherungen

250-500 Mitarbeiter

• Tätigkeit als Staff Site Reliability Engineer in einer sicherheitskritischen Produktionsumgebung
• Technische Unterstützung und Mentoring eines 12-köpfigen Site Reliability Engineering Teams
• Analyse und Stabilisierung komplexer Cloud- und Plattformarchitekturen
• Unterstützung bei der Weiterentwicklung der SRE-Strukturen und Betriebsprozesse
• Architektur- und Prozessreviews im Bereich Datenplattform / Data Warehouse
• Zusammenarbeit mit Plattform-, Entwicklungs- und Security-Teams zur Verbesserung der Systemzuverlässigkeit
• Sicherstellung der Einhaltung von Datenschutz- und Sicherheitsanforderungen (GDPR)

Site Reliability Team lead

ClovrLabs

Internet und Informationstechnologie


● Played a pivotal role in architecting a Cloud Native infrastructure within a Hybrid Cloud environment, while spearheading the recruitment and seamless onboarding of a proficient team of Full-Time Employees (FTEs)
● Delivered expert guidance on cloud architecture best practices within a prominent Blockchain Reputation company, ensuring alignment with industry standards and optimal performance
● Established and maintained a robust and scalable AWS EKS (Kubernetes) setup, characterized by its efficiency, cleanliness, and ease of maintenance
● Conducted real-time code profiling and executed comprehensive load and performance testing to optimize system performance and enhance reliability
  • ●  Evaluated on-premises infrastructure sizing to ensure optimal resource allocation and efficiency.
  • ●  Designed and operated an Extract, Transform, Load (ETL) solution using Jupyter Notebooks and Spark, facilitating efficient data
    processing and analytics
● Provided leadership and mentorship to a team of six DevOps/SRE professionals across the organization, fostering collaboration
and driving operational excellence

DevOps Engineer

ATU / NORAUTO

Internet und Informationstechnologie


● Provided extensive support in maintaining a Hybrid Cloud infrastructure for over four years, ensuring seamless operations and optimal performance
● Diagnosed and resolved issues related to the Active Directory Distributed File System (DFS) roaming profile load balancing setup, leveraging a Prometheus Exporter for Windows machines
  • ●  Integrated Active Directory (AD) authentication via Kerberos on Keycloak to enable Single Sign-On (SSO) across various services
  • ●  Facilitated the migration and management of a SPNEGO-based AD Single Sign-On (SSO) solution on NGINX endpoints, ensuring

secure and efficient authentication mechanisms

Security Engineer

HelloFresh Berlin

Internet und Informationstechnologie


• Implemented and managed CloudFlare Web Application Firewall (WAF) for proactive threat mitigation and incident response,
alongside the ELK (Elasticsearch, Logstash, Kibana) stack deployed on AWS for comprehensive log analysis and monitoring

DevOps Team Lead

UMI Urban Mobility International / Volkswagen Group Berlin

Internet und Informationstechnologie



● Facilitated the organization during the Go-Live phase of the product, ensuring operational readiness of the infrastructure for seamless deployment
● Implemented a metrics-based process to assess product quality from the end user's perspective, thereby influencing Objectives and Key Results (OKRs)
● Directed the activities of a team of 5 Site Reliability Engineers (SREs), overseeing their tasks and ensuring alignment with organizational goals

CTO, SiWeGO Berlin

Internet und Informationstechnologie



● Implemented CloudFlare Web Application Firewall (WAF) and managed Incident Response procedures, alongside ELK stack deployment on AWS for comprehensive log analysis
  • ●  Managed vendor relationships and supplier partnerships to ensure seamless operations and timely delivery of services
  • ●  Designed solutions with a focus on minimizing operational expenditure (OPEX) while maintaining high performance and reliability
  • ●  Developed algorithmic approaches for the design of graph overlaps, ensuring efficient indexing and retrieval of data

DevOps Engineer

ATU

Internet und Informationstechnologie

●  Implemented Kubernetes with Helm, ELK stack, and Grafana/Prometheus, utilizing Docker on VMWARE/Xen virtualized systems.
●  Managed Content Delivery Network (CDN) for Distributed Denial of Service (DDoS) mitigation and DNS management, focusing on reducing the attack surface.
●  Conducted Application Performance Monitoring and assessed overall Infrastructure Performance to optimize system efficiency.
●  Provided support for Extract, Transform, Load (ETL) processes to ensure seamless data integration and processing.

Sr.Performance Engineer

OLX

Internet und Informationstechnologie


● Conducted thorough evaluations and selection processes for Content Delivery Network (CDN), Secure Sockets Layer (SSL), and Domain Name System (DNS) vendors, providing internal coaching and guidance on related topics within OLX
● Led load and performance testing initiatives and conducted architectural performance design reviews to optimize system efficiency and reliability
● Developed comprehensive capacity planning strategies for migrating from bare-metal infrastructure to cloud-based solutions, ensuring seamless transition and scalability
● Provided coaching and support to over 50 engineers across the organization, fostering skill development and knowledge sharing to drive continuous improvement

Service Reliability Engineer

SonyPlaystation (“Playstation Now” service - GAIKAI)

Internet und Informationstechnologie

● Designed and implemented a robust Network Time Protocol (NTP) infrastructure to ensure precise time synchronization across systems
●  Managed KVM/Ganeti Gentoo's package maintenance, ensuring the stability and reliability of the virtualization environment
●  Served as a subject matter expert in Hadoop, providing strategic guidance and technical leadership in implementing and
optimizing Hadoop-based solutions
●  Architected CEPH storage solutions to deliver scalable, high-performance storage infrastructure tailored to organizational needs.
●  Received official RH training in NYC Q1 2016 “Red Hat Ceph Storage Architecture and Administration (CEPH125)” by J. C. Lopez
●  Designed NFS storage systems to achieve optimal performance, with a throughput of 40 MBps per client and a target egress of 46
Gbps from a single node serving the rack
● Developed a Domain Specific Language (DSL) using the textX library for enhanced functionality and flexibility in querying the
graphite-api
● Provided comprehensive coaching and training to new hires (5), facilitating their integration into the team and ensuring proficiency
in their roles

Sr.Data Architect

Rocket Internet

Internet und Informationstechnologie

● Acted as a System Operations (sysop) and Developer for the Business Intelligence team, overseeing critical infrastructure and software development.
● Developed and implemented a Python-based FLASK Google Analytics API fetcher to gather and process data for analytical purposes.
  • ●  Installed, configured, and managed a Hadoop/MapR M3 cluster consisting of 5 nodes, facilitating big data processing and analysis.
  • ●  Integrated HIVE and SquirrelSQL via Thrift2/JDBC to enable seamless interaction with the Hadoop ecosystem, empowering Data
    Scientists to extract insights efficiently.
● Pioneered various data intake pathways optimized for long-distance and high-latency scenarios, ensuring robust data acquisition
across diverse environments.
● Provided expert support to the team in executing Extract, Transform, Load (ETL) business processes, optimizing data workflows
for enhanced efficiency and accuracy.

Sr.Software Engineer

Rocket Internet

Internet und Informationstechnologie

●  Held the role of Chief Troubleshooter with a direct reporting line to CTO Ronny Rentner.
●  Demonstrated expertise in troubleshooting various technology stacks, ranging from patching on PHP / Zend Framework and SQL
profiling, to debugging Couchbase, Nginx, and PHP-FPM, along with 3rd party Java servlets.
  • ●  Led capacity testing, deployment automation, and continuous delivery initiatives to enhance system efficiency and reliability.
  • ●  Designed a high-latency tolerant file transfer solution utilizing stunnel+rsyncd, tailored to address challenges encountered in South
    East Asia based ventures.
● Orchestrated bug fixing efforts to address recurring issues across fork-based code handling for 300 ventures, ensuring smooth
operation and minimal disruptions.

Backend QA Manager

‘txtr / 3M

Internet und Informationstechnologie

●  Spearheaded the development of eBook reading technology, specializing in Adobe DRM-related technologies
●  Conducted REST API user acceptance and load & performance testing
●  Led emergency and root cause analysis efforts during critical incidents
●  Implemented operational monitoring strategies to ensure system stability and performance
●  Proficient in coding with Python, Ruby, C++, and Java
Additional Contributions:
●  Developed a Ruby-based test automation solution utilizing Cucumber and Selenium frameworks
●  Engineered an XML, XPATH, and XQUERY search engine for efficient data retrieval
●  Spearheaded a proof of concept initiative involving Hadoop technology
●  Executed CDN and RUM testing, leveraging JMX-based instrumented metrics for performance assessment
●  Facilitated code reviews to ensure code quality and adherence to best practices
●  Oversaw reporting and coordination activities for a team of 3 UAT engineers

Sr.System Engineer

ProfitBricks (an 1&1 UnitedInternet - IONOS venture)

Internet und Informationstechnologie

●  Implemented a Debian-based KVM private cloud solution
●  Served as Deputy Operations Team Lead, overseeing a team of 6 engineers
●  Conducted software packaging activities
●  Proficient in Python scripting and experienced in configuration management tools such as bcfg2 and Puppet
●  Managed monitoring and backup operations effectively

Sr.System Engineer

Nokia

Internet und Informationstechnologie


●  Supporting 7 different teams related to Nokia Maps / NAVTEQ in a DevOps capacity
●  Setting up continuous integration processes
●  Packaging for REDHAT/CentOS systems
●  Scripting in Python and Ruby, and utilizing configuration management tools like Puppet
●  Handling operations related to development data centers.

Consultant – Sr.Test Specialist

Vodafone

Internet und Informationstechnologie

  • Proficient in utilizing performance testing tools such as Grinder, Apache Bench, and JMeter Experienced in implementing testing automation methodologies.
  • Played a key role in the successful deployment of major projects including:
  • Vodafone Live! (J2EE based web mobile platform)
  • Vodafone 360 ((project with an estimated budget of approximately €250 million) Vodafone GIG (tibco based SOA infrastructure)

Consultant - Solution Architect

Internet und Informationstechnologie



●  Assisted in teaching network socket programming on GNU/Linux systems at Politecnico di Milano as a Teaching Assistant.
●  Managed system administration duties for GNU/Linux systems, including Suse, Debian, Yellowdog, RedHat, and Mandrake distributions.
●  Contributed as a Frontend Developer, utilizing Macromedia tooling suites to enhance user interfaces.
●  Led a team responsible for developing ECDL Libreoffice simulations, overseeing a team of three.
●  Executed PLSQL migration from Solaris to Linux within banking premises.
●  Conducted security auditing on behalf of a telecommunications company.
●  Engaged in embedded software development, focusing on Linux-based firewall solutions and QT development.
●  Designed and developed PHP applications to meet specific project requirements.
●  Actively contributed to OpenLDAP, Kerberos, and PKCS#12 protocols as part of the OpenGroupware - Gosa^2 project.
●  Developed XML-based databases and implemented full-text retrieval systems, integrating with Openoffice document merging
functionalities.

 


Kontaktanfrage

Einloggen & anfragen.

Das Kontaktformular ist nur für eingeloggte Nutzer verfügbar.

RegistrierenAnmelden