RAJ
Verified from Jobgether's careers page · Lever

OSS Engineer

JobgetherCanada (On-Site) 🇨🇦5+ YearsPosted Sep 14, 2026
Apply now

Apply faster

Fill most application fields automatically from your RealAnalystJobs profile. You review, you submit.

Included with the from $2 for 30 days unlock

About this role

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a OSS Engineer based in Canada. This role focuses on designing, building, and operating Operations Support Systems that enable efficient management of cloud solutions and large fleets of connected devices. You will transform fragmented operational data and repetitive manual processes into secure, reliable, and scalable tools used by Network Operations and Support teams. The position combines software engineering, network operations, automation, data products, and service assurance in a highly hands-on environment. You will build dashboards, APIs, integrations, workflow automations, and operational services that help teams detect, diagnose, and resolve issues faster. The role requires close collaboration with cloud, DevOps, firmware, QA, database, product, security, and operations specialists. You will also contribute to continuous improvement by analyzing incidents, alert patterns, device behavior, and operational feedback to identify new automation opportunities. This contract opportunity offers the chance to directly improve operational capacity, reliability, customer experience, and the efficiency of large-scale managed networks. Accountabilities: • OSS Platform Engineering: Design, develop, test, deploy, and maintain internal web applications, APIs, command-line tools, and services supporting fault management, performance monitoring, configuration, inventory, and service assurance use cases. • Operational Tooling: Build reusable capabilities around service-provider tenants, cloud services, CPE inventories, network topology, dependencies, device reachability, firmware versions, alarms, incidents, and customer impact. • Software Engineering Practices: Apply modular design, code review, automated testing, version control, CI/CD, release management, rollback strategies, telemetry, and lifecycle ownership to ensure maintainable production systems. • Dashboards and Data Products: Create role-based dashboards for executives, NOC teams, Support, and service owners, providing actionable visibility into service health, incident performance, alert quality, SLA/SLO attainment, customer impact, and managed-device fleet health. • Workflow Automation: Automate alert enrichment, deduplication, correlation, ticket creation, routing, escalation, stakeholder notifications, evidence collection, and operational reporting while measuring adoption, time saved, error reduction, and operator outcomes. • Cloud and CPE Integrations: Integrate operational tooling with cloud APIs, Kubernetes, observability platforms, ITSM systems, collaboration tools, CI/CD systems, databases, messaging platforms, and customer-facing operational systems using APIs, webhooks, event streams, and device-management data. • Operational Intelligence: Turn incident reviews, ticket trends, alert noise, escalations, and operator feedback into a prioritized tooling and automation roadmap, while developing analytics for recurring incidents, device cohorts, capacity trends, and other improvement opportunities. • Reliability and Security: Operate OSS capabilities as production systems with appropriate monitoring, alerting, capacity planning, backup and recovery, dependency documentation, authentication, RBAC, secrets management, audit logging, secure coding, and vulnerability remediation. • Support and Enablement: Create runbooks, user guidance, support models, and ownership documentation, train NOC and Support users, incorporate feedback into the roadmap, and provide escalation support for critical tooling failures or automation defects when required. Requirements • Professional Experience: 5+ years of experience in OSS engineering, network automation, NOC tooling, software engineering for operations, or a closely related field, ideally within service providers, broadband operators, managed-network environments, or telecom technology organizations. • OSS and Telecom Expertise: Demonstrated experience with telecom Operations Support Systems covering areas such as fault, performance, configuration, inventory, topology, service assurance, or network and service orchestration. • Software Engineering: Strong programming skills in Python, Go, Java, TypeScript/JavaScript, or a comparable language, with experience delivering maintainable production applications, services, or automation frameworks. • Integration and Data Engineering: Hands-on experience building APIs, event-driven integrations, data pipelines, databases, and user-facing operational dashboards, as well as integrating observability, ITSM, collaboration, and operational platforms. • Cloud and Infrastructure: Practical experience with Linux, containers, Kubernetes, and at least one public cloud platform, supported by a strong understanding of modern software delivery and infrastructure practices. • Networking Knowledge: Working knowledge of TCP/IP, DNS, DHCP, TLS, routing, APIs, and systematic troubleshooting across distributed systems and network environments. • Engineering Practices: Strong understanding of Git, code review, automated testing, CI/CD, release management, rollback, monitoring, and production lifecycle management. • Operational Collaboration: Ability to work directly with NOC, Support, and Operations engineers, translate ambiguous operational challenges into clear technical requirements, and measure whether delivered solutions produce meaningful improvements. • Education: Bachelor’s degree in Computer Science, Engineering, or a related discipline, or equivalent practical experience. • Preferred CPE Experience: Experience with cloud-managed CPE such as broadband gateways, routers, ONTs, Wi-Fi/mesh systems, or similar connected edge devices is advantageous. • Preferred Technologies: Familiarity with TR-069/CWMP, TR-369/USP, TR-181, ACS/USP controllers, device telemetry, GCP, Kubernetes, Terraform, Helm, Grafana, Prometheus, O

Description from Jobgether's public careers feed, reproduced so you can read the role here. Apply on the company's own site; RealAnalystJobs never submits anything for you.

Raj