In sintesi
- Mansioni: Guida un team IT per garantire la disponibilità e l'affidabilità dei servizi critici.
- Azienda: Azienda globale innovativa nel settore della pianificazione della supply chain.
- Benefit: Salario competitivo, bonus, opportunità di crescita professionale e ambiente di lavoro flessibile.
- Altre informazioni: Crescita professionale in un contesto collaborativo e orientato ai risultati.
- Perché questo lavoro: Fai la differenza risolvendo problemi complessi in un ambiente dinamico e stimolante.
- Qualifiche: Esperienza in ingegneria SRE o produzione, con competenze in sistemi e networking.
La retribuzione prevista è compresa tra 68000 - 68000 € per anno.
About Us
We are a dynamic, rapidly growing global company and the innovators of service-driven supply chain planning software.
We help companies make better, faster supply chain decisions that reduce inventory, improve customer satisfaction, and deliver powerful financial results amid increasing complexity, product proliferation, and uncertainty.
Our solutions have been recognized by customers globally and analyst firms, such as Gartner, for our ability to support service and inventory trade-offs, while dramatically improving planner productivity.
Tools Group has been successfully deployed worldwide in more than 44 countries, and we have one of the highest customer retention rates in our industry.
About the Role
We are looking for an experienced Site Reliability Engineer who can also lead IT service operations.
You will lead a small IT/Ops team, remainthe senior technical escalation point for production services, and act as the operational interface between IT/Ops, Engineering, Product, Security, Support, and business teams.
This is not a coordination-only service management role or a generalist infrastructure position.
You will diagnose distributed-system failures using logs, metrics, traces, commands, and platform tooling; make safe recovery decisions; automate recurring work; and engineer lasting reliability improvements.
- Main Responsibilities
- Lead major incidents from impact assessment and containment through recovery, stakeholder communication, root-cause analysis, and corrective actions.
- Troubleshoot complex issues across Windows and Linux systems, Kubernetes and container workloads, hybrid networking and DNS, cloud infrastructure, identity, authentication, databases, storage, APIs, and service dependencies.
- Operate and improve Azure, OCI, or comparable cloud environments, including monitoring, access controls, backup and recovery, reliability, and cost-aware scaling.
- Define and improve service-level indicators andobjectives, observability, alert quality, capacity, resilience, dependency mapping, and recovery readiness for critical services.
- Automate operational tasks and controls using Power Shell, Python, infrastructure as code, or CI/CD pipelines, with validation, logging, secure credential handling, and rollback.
- Apply incident, change, and problem management pragmatically, protecting service availability without introducing unnecessaryprocess.
- Connect technical and business teams: clarify service ownership and dependencies, translate business needs into reliability and infrastructure requirements, frame risk and tradeoffs, align priorities, and ensure decisions have accountable owners and realistic commitments.
- What We Are Looking For
- We care more aboutdemonstratedproduction engineering experience than a checklist of certifications. Strong candidates will bringall ofthe following.
- A strong SRE or Production Engineering background, typically5+yearsoperating business-critical, customer-facing, or high-availability services.
Recent work must include direct technical ownership, not only coordination or people management.
- Recent ownership of high-severity incidents, including technical triage, recovery decisions, clear communications, and measurable follow-through.
- Strong systems and network troubleshooting fundamentals: Windows and Linux, TCP/IP, DNS, routing, firewalls, proxies or load balancers, and the ability to isolate faults across service layers.
- Hands-on cloud operations experience in Azure, OCI, or a similar platform, includingcompute, storage, networking, IAM, observability, backup, and recovery.
- Production experience with containers and Kubernetes, including workload health, scheduling, networking, persistent storage, secrets, deployment and rollback, scaling, and backup or recovery considerations.
- Deep observability and reliability engineering practice: metrics, logs, distributed tracing, actionable alerting, SLI/SLO design, capacity and saturation analysis, failure-mode thinking, and post-incident engineering.
- Ability to diagnose database-backed and API-driven services across application, query, connection-pool, storage, certificate, network, and downstream dependency layers.
- Practical identity and access management experience with Active Directory and Microsoft Entra ID or equivalent, including hybrid identity, privileged access, MFA, service identities org MSAs, lifecycle controls, dependency mapping, and controlled recovery from identity failures.
- Evidence of safe automation and infrastructure-as-code work using Power Shell, Python, Terraform or comparable tooling and CI/CD.
You should be able to explain testing, idempotency, error handling, credential security, rollout, rollback, and measurable impact.
- Experience leading, mentoring, or acting as the senior escalation point for other technical professionals.
- Strong business-facing and cross-functional leadership.
You can translate technical complexity into business impact and options, challenge unsafe or unrealistic requests constructively, negotiate priorities, and communicate decisions clearly to engineers, executives, customers, and non-technical stakeholders.
- Additional Relevant Experience
- Microsoft 365, endpoint management, EDR, device compliance, and hybrid workplace operations.
- Formal ITIL, cloud, security, or infrastructure certifications.
- What success looks like
- Incidents are diagnosed and resolved with greater speed, structure, and confidence.
- Monitoring, runbooks, automation, recovery controls, and change practices reduce repeat failures and operational toil.
- The team becomes more capable and accountable without relying on a single hero.
- Technical and business teams share clear service ownership, priorities, risk decisions, and delivery expectations.
- Our hiring process
The process includes a scenario-based SRE technical discussion.
We will ask you to think aloud through realistic production incidents involving cloud, Kubernetes, identity, networking, databases, storage, APIs, and service dependencies.
You will be expected to describe the logs, metrics, traces, commands, tools, tradeoffs, and recovery criteria you would use.
We will also assess how you align technical and business stakeholders when priorities, risk, and customer commitments conflict.
Our Vision, Purpose, and Values
Our VISION: Unparalleled control over demand and supply to deliver certainty.
Our PURPOSE: Problem Solvers Welcome.
Our VALUES: Deliver the Goods - Have Deep Care - Find the Right Answer, Not the First Answer - Creativity That Endures - Brilliant But Not Loud.
Salary range
55-68k/year, plus 10% bonus based on personal and company objectives.
Equal Opportunity Employer
Tools Group provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.
- Tools Group is an E-Verify employer, to learn more please visit E-Verify. gov
- #J-18808-Ljbffr
Senior Infrastructure Engineer datore di lavoro: ToolsGroup Inc.
ToolsGroup è un datore di lavoro eccezionale, offrendo un ambiente di lavoro dinamico e innovativo nel settore della pianificazione della supply chain. I dipendenti beneficiano di opportunità di crescita professionale, un forte supporto per l'automazione e la gestione delle operazioni IT, e una cultura aziendale che valorizza la creatività e la collaborazione. Inoltre, la posizione offre un pacchetto retributivo competitivo con bonus basati su obiettivi personali e aziendali, rendendo l'azienda un luogo ideale per chi cerca un impiego significativo e gratificante.
Consigli degli esperti StudySmarter🤫
Ecco come pensiamo che potresti ottenere Senior Infrastructure Engineer
✨Partecipa ai Meetup Locali
Immergiti nella community tech partecipando a meetup locali di ingegneria e sviluppo software. Questo non solo ti permetterà di imparare delle nuove tendenze, ma anche di incontrare potenziali datori di lavoro e colleghi. Non c'è nulla di meglio che un contatto personale per lasciare il segno!
✨Contribuisci a Progetti Open-Source
Un ottimo modo per farti notare è contribuire a progetti open-source. Questo non solo arricchisce il tuo portfolio, ma ti mette in contatto con professionisti del settore che possono facilmente raccomandarti per posizioni full-time. Inoltre, dimostrerai competenze pratiche che le aziende adorano!
✨Esplora le Piattaforme di Recruitment Tecnico
Ci sono piattaforme di recruitment specializzate per il settore tech dove le aziende cercano costantemente profili come il tuo. Assicurati di registrarti su siti specifici per ingegneri e sviluppatori per ricevere piuttosto offerte di lavoro pertinenti e opportunità interessanti. Fai in modo che il tuo profilo risalti!
✨Applica Direttamente a ToolsGroup Inc.
Non dimenticare di controllare il sito web di ToolsGroup Inc. per eventuali opportunità di lavoro nel settore ingegneristico. Applicare direttamente può aumentare le tue chance di essere notato, specialmente se segui i loro canali social per aggiornamenti su assunzioni e eventi informativi.
Pensiamo che ti servano queste competenze per eccellere come Senior Infrastructure Engineer
Alcuni consigli per la tua candidatura 🫡
Mostra i tuoi progetti!:Nel tuo CV, assicurati di includere una sezione dedicata ai progetti di sviluppo software a cui hai lavorato. Questi possono essere progetti universitari, lavori freelance o anche side projects. Se hai un GitHub, non dimenticare di linkarlo: i recruiter adorano vedere il codice e come affronti le sfide.
Competenze tecniche in evidenza:Fai un elenco chiaro delle tue competenze tecniche nel tuo CV, come linguaggi di programmazione, framework e strumenti che conosci. Assicurati che siano rilevanti per il ruolo di Senior Infrastructure Engineer in ToolsGroup Inc.. Questo aiuterà a dimostrare che sei il candidato ideale per il lavoro!
Scrivi una lettera di motivazione mirata:Quando scrivi la tua lettera di motivazione, evidenzia perché sei appassionato di ingegneria e sviluppo software. Parla dei tuoi obiettivi professionali e di come pensi di crescere in ToolsGroup Inc.. Ricorda, vogliamo vedere il tuo entusiasmo e il tuo desiderio di imparare!
Attenzione ai dettagli:Nell'ambito dell'ingegneria e dello sviluppo software, i dettagli contano. Fai attenzione alla formattazione del tuo CV e della lettera di motivazione. Un CV ben strutturato e privo di errori mostra che sei meticoloso e professionale, qualità fondamentali per un ruolo a tempo pieno come Senior Infrastructure Engineer.
Come prepararti a un colloquio di lavoro presso ToolsGroup Inc.
✨Preparati con le tue abilità tecniche
Per un colloquio in ingegneria e sviluppo software, è fondamentale essere in grado di dimostrare le tue abilità tecniche. Preparati a rispondere a domande di programmazione e a risolvere problemi dal vivo. Puoi anche praticare con piattaforme come LeetCode o HackerRank per affrontare esempi di codice che potresti incontrare.
✨Mostra il tuo portfolio di progetti
Essendo un candidato full-time, è importante avere un portfolio ben curato che mostri il tuo lavoro. Porta con te esempi di progetti passati, sia personali che professionali. Spiega il tuo ruolo in ciascun progetto e i risultati ottenuti. Questo non solo dimostra le tue capacità, ma anche la tua passione per il settore.
✨Preparati a domande sul lavoro di squadra
Nel campo dell'ingegneria e sviluppo software, il lavoro di squadra è cruciale. Aspettati di ricevere domande su come hai collaborato in precedenti progetti o come affronti i conflitti con i membri del team. Pratica le tue risposte usando esempi concreti che mettano in luce le tue abilità relazionali.
✨Conosci il processo di sviluppo che usano
Ogni azienda ha il proprio modo di fare le cose. Prima del colloquio in ToolsGroup Inc., informati sul loro processo di sviluppo software, come Agile o Scrum. Essere in grado di discutere come ti adatteresti a questi metodi o come hai già lavorato con essi può farti risaltare come candidato ideale.