
Senior Software Engineer II-Enterprise Tools Engineering Architect
Staples Canada11 days ago
Framingham, MA, USAStaff+
Responsibilities
- Lead the team and architecture, design, and continuous improvement of enterprise observability platforms.
- Design scalable, integrated solutions across APM, infrastructure monitoring, session replay, log analytics, applications, infrastructure, and user-experience layers.
- Establish architecture patterns, governance models, and integration frameworks for consistent telemetry, correlation, and actionable insights.
- Partner with Engineering, SRE, Product, and Infrastructure teams to embed observability practices into the software development lifecycle.
- Evaluate and adopt distributed tracing, OpenTelemetry, and AI-driven observability technologies.
- Drive platform standardization, signal quality, operational excellence, and integration with business-critical services.
- Influence vendor roadmaps, resolve vendor issues, and provide strategic leadership across multiple engineering teams and domains.
Requirements
- Bachelor’s degree in Computer Science, Engineering, or a directly related field.
- 10+ years of progressive experience in observability, APM, and infrastructure monitoring with architecture-level ownership.
- Deep expertise designing and administering APM platforms such as New Relic, Dynatrace, or Datadog, including distributed tracing, transaction monitoring, and service dependency mapping.
- Strong infrastructure-monitoring experience with platforms such as Zabbix, Prometheus, or Azure Monitor across compute, network, storage, and hybrid-cloud environments.
- Experience with session replay and user-experience platforms such as FullStory.
- Experience with log analytics platforms such as Splunk or Elastic, including ingestion, routing, indexing, and retention management.
- Ability to design cross-platform observability architecture integrating APM, infrastructure, logs, and session data with cloud platforms, software-delivery pipelines, and enterprise systems.
- Proficiency in Python and Bash for platform configuration, scaling, and governance automation.
- Strong leadership, cross-functional collaboration, strategy, roadmap, platform-standardization, and vendor-management skills.
- Preferred: certifications in observability platforms such as New Relic, Dynatrace, Datadog, Splunk, or FullStory.
- Preferred: experience with AI/ML-driven observability, anomaly detection, AIOps, generative AI or agentic frameworks, multi-vendor observability environments, and tool rationalization.
- Preferred: strong understanding of SRE practices including SLOs, error budgets, and reliability engineering.
- Preferred: experience supporting large-scale, high-traffic B2B and B2C platforms.
Benefits
- Onsite work location in Framingham, Massachusetts.
- 22 days of PTO plus a holiday schedule with 7 observed paid holidays and 1 floating holiday.
- Online and retail discounts.
- Company-match 401(k).
- Physical and mental health wellness programs.
- Inclusive culture with associate-led Business Resource Groups.