NetBox Labs

Senior Full Stack Engineer, Observability

NetBox Labs · LATAM, UK, USA

🔥6 people viewed this job

About the Role

NetBox Labs is seeking a Full Stack Engineer to join our rapidly expanding engineering team. We have multiple positions open at different seniority levels across several teams. About NetBox Labs NetBox Labs builds the next generation of network automation tools for modern infrastructure teams, with NetBox as the network source of truth at the center. The Observability team builds the products that connect that source of truth to what is actually running on the network: • NetBox Discovery finds devices, interfaces, and other network entities. • NetBox Assurance brings discovered data into NetBox and helps operators spot drift between intended and actual state, review changes, and decide what to accept. • Fleet Management and the Orb agent deploy, configure, and manage the agents that collect discovery and telemetry data from customer networks. These agents use SNMP, gNMI, and other device interfaces. • Diode is the ingestion pipeline that moves that data into NetBox reliably. We're hiring a Full Stack Engineer who can work across the whole path, from the agents and services that collect network data to the interfaces operators use to understand and act on it. Role overview You'll join the Observability team and take ownership of features end to end. That covers: • Go and Python services and gRPC APIs in the control and data planes. • The data flows that carry discovery and telemetry results into NetBox. • The React dashboards and interfaces where customers monitor network and device health, explore telemetry, review discovered data, and manage their agent fleet. You'll work closely with product, design, and other engineering teams, and you'll help run the services the team owns in production. This role is hands-on. You'll ship features across the stack, improve reliability and performance, and help define the architecture and practices needed to scale network discovery and assurance to large, complex customer environments. What you'll do • Design, build, and operate backend services in Go and Python for discovery, assurance, fleet management, and data ingestion. • Define and evolve gRPC and REST APIs with clear, versioned contracts using Protocol Buffers and OpenAPI. • Build React and TypeScript monitoring and telemetry dashboards that show device, interface, and network health in real time, with time-series charts, status views, and drill-down from fleet to device to interface. • Design dashboard experiences that help operators spot problems fast: sensible defaults, time-range and filter controls, thresholds and alert states, and clear links from a metric to the underlying device in NetBox. • Work with backend engineers on the query and aggregation APIs that power dashboards, so they stay fast with large fleets and high-frequency telemetry. • Build type-safe integration between frontend and backend using generated API clients, shared schemas, and consistent error and auth handling. • Participate in the team's on-call rotation for the services it owns. • Add automated tests across the stack (unit, integration, contract, and end-to-end) and enforce quality gates in CI. • Collaborate with product managers, designers, customer-facing teams, and other engineering teams. The goal is for features to solve real network operator problems. • Use AI-enabled development tools and agentic workflows day to day to speed up design, coding, testing, code review, and incident triage, and help the team adopt them effectively and safely. • Review code, mentor teammates, and share best practices for service design, API design, and frontend engineering. • Participate in planning processes and help shape the roadmap. What we're looking for (minimum qualifications) • Experience: 5+ years of professional software engineering, with meaningful production experience on both backend and frontend. • Backend: • Production experience with Go and Python: strong in at least one and working proficiency in the other. • Hands-on experience designing and operating gRPC services with Protocol Buffers, including schema evolution and backward compatibility, streaming RPCs, deadlines, interceptors/middleware, and error handling. • Experience building distributed, event-driven systems, including message queues (e.g., RabbitMQ or Kafka), asynchronous job processing, and idempotent data ingestion. • Frontend (monitoring and telemetry dashboards): • Strong React and TypeScript skills, including component composition, state management, and typing best practices. • Proven experience building monitoring, observability, or analytics dashboards: time-series charts, heatmaps, status and health views, and drill-down navigation. • Hands-on experience with data visualization libraries (e.g., D3, ECharts, Recharts, uPlot, or Visx). • Experience rendering large or high-frequency datasets performantly, using techniques such as downsampling, virtualization, and canvas or WebGL rendering. • Experie

NetBox Labs has 1 open position on Remote Vibe Coding Jobs.

💬 Developer Questions

Ask the team a question — answers show up here

🎯

What does the interview process look like?

🤖

What AI/vibe coding tools does the team use daily?

👥

How big is the engineering team?

⏰

Is the team fully async or are there required meetings?

🚀

What does onboarding look like for remote hires?

🔧

Can you share more about the tech stack and architecture?

📈

What does career growth look like in this role?

📅

What does a typical day look like?

💰

Is there a salary range you can share?

📊

Is equity or stock options part of the package?

🌍

Are there timezone requirements or preferences?

🛂

Do you sponsor work visas?

🏢 Is this your listing? Claim it to answer questions

Similar Jobs

Helpful resources

Hiring for a similar role? Post your job here — it's free →