1 month ago
Lisbon, PortugalSenior
Responsibilities
- Design and maintain large-scale telemetry pipelines for metrics, logs, traces, and network-flow data.
- Develop ETL and data-transformation pipelines to normalize, enrich, and model telemetry for real-time and historical analytics.
- Build analytics datasets, dashboards, monitoring systems, and automated alerting for network performance, failures, anomalies, and service health.
- Design and support traffic-acquisition architectures using packet brokers, TAPs, traffic mirroring, and virtual taps.
- Operate and scale observability stacks, telemetry collectors, analytics databases, and visualization platforms.
- Automate observability infrastructure deployment and operations using Infrastructure-as-Code and container orchestration.
- Collaborate with network engineering, platform engineering, and operations teams.
Requirements
- Strong experience with Linux systems administration, production platform operations, OSS environments, and distributed production systems.
- Experience with telecom service-assurance and network-monitoring platforms such as NumoData, Radcom, or equivalent, plus packet capture and network traffic analysis.
- Hands-on experience with Prometheus, Thanos, Loki, Tempo, Grafana, and telemetry collection and monitoring systems.
- Experience with Vector, Telegraf, Fluentd, Kafka, SNMP, Syslog, NetFlow, and IPFIX.
- Experience designing observability data models and building ETL pipelines using Apache NiFi or equivalent.
- Experience with SNMP trap processing, Alertmanager, OpenNMS, or equivalent fault-management platforms.
- Experience with ClickHouse or Vertica, PostgreSQL or MySQL, and Redis.
- Understanding of telecom network architectures, telemetry analysis, traffic mirroring, TAPs, virtual TAPs, and packet brokers.
- Strong scripting and development skills in Python and Golang, plus Ansible for PNF/VNF environments and Helm for CNF environments.
- Familiarity with OpenTelemetry, gNMI, gRPC streaming, OpenConfig, eBPF, or custom instrumentation is preferred.
- Preferred experience includes high-volume telemetry ingestion, real-time analytics, large-scale event processing, AI/ML monitoring algorithms, Kubernetes, AWS EKS, OpenStack Magnum, Heat, Terraform, AWS CloudFormation, Kubernetes Operators, CI/CD pipelines, high-availability telemetry storage, and analytics database performance tuning.
Benefits
- Hybrid workplace in Lisbon, Portugal.
- Career growth in a rapidly expanding telecommunications company.
- Exposure to major transactions affecting the telecommunications industry.
- Collaboration with experienced technology leaders, founders, senior management, and external advisors.
- Professional development alongside industry experts.
- Opportunities to work from different 1GLOBAL offices internationally.
- Collaborative, dynamic, inclusive, and international work environment.
Tech Stack
AnsibleApache KafkaClickHouseGoGrafanagRPCHelmKubernetesLinuxMySQLPostgreSQLPrometheusPythonRedisTerraform
Categories
Data EngineeringDevOps
