Build and maintain data ingestion pipelines for a cybersecurity platform in a fully remote environment.
Posted by employer 3 days ago
First seen on Joblaze 1 day ago
Last verified on the company career page 1 day ago
Skills & Technologies
What you'll build
Must have
Nice to have
Practical constraints
Requirements
Not disclosed in this posting: visa sponsorship.
Benefits
Joblaze summary
The Senior Software Engineer on the Data Ingestion team at Horizon3.ai focuses on developing and maintaining data pipelines that convert raw pentest results into structured data for the NodeZero platform. Proficiency in Go and Python is essential, along with experience in building ETL systems and working with relational databases. This role is suited for seasoned engineers with a strong background in backend development and data processing. The team values collaboration and operates in a remote environment, emphasizing a culture of respect and ownership.
Joblaze insights
Quick facts
From the original posting
Horizon3 is a fast-growing, remote cybersecurity company dedicated to the mission of enabling organizations to proactively find, fix and verify exploitable attack vectors before criminals exploit them. Our flagship product, the NodeZero™ platform, delivers production-safe autonomous pentests and other key assessment operations that scale across the largest internal, external, cloud, and hybrid cloud environments. NodeZero has been adopted by organizations of all sizes, from small educational institutions to government agencies and Global 100 enterprises. It is used by IT Ops/SecOps teams, consulting pentesters, and MSSPs and MSPs.
We are a fusion of former U.S. Special Operations cyber operators, startup engineers & operators, and formerly frustrated cybersecurity practitioners. We're committed to helping solve our common security problems: ineffective security tools and false positives, resulting in alert fatigue, blind spots, "checkbox" security culture, cybersecurity skills shortage, and the long lead time and expense of hiring outside consultants. Collectively, we are a team of learn-it-alls, committed to a culture of respect, collaboration, ownership, and results.
The Sr. Software Engineer on the Data Ingestion team builds and maintains the pipelines that transform raw pentest output into structured, queryable data across the NodeZero platform. NodeZero runs thousands of autonomous pentests — your pipelines are the first link in the chain between execution engine results and the insights customers see. You'll work on high-throughput ETL systems that normalize heterogeneous scan data, push structured findings to graph and relational stores, and feed downstream systems including Asset Inventory, reporting, and analytics.
This role reports to the Data Ingestion Engineering Manager and works alongside a small team of backend engineers including one P5 technical anchor.
Design, implement, and operate data ingestion pipelines that process pentest results at scale — scan output, discovered assets, vulnerability findings, and network topology data
Own end-to-end delivery of pipeline features: schema design, implementation, testing, deployment, and monitoring
Build and maintain ETL components in Go and/or Python that normalize heterogeneous inputs from NodeZero's execution engines
Partner with the Asset Inventory and Data Platform teams to define and evolve the data contracts that connect ingestion to downstream consumers
Diagnose and resolve production issues in pipeline latency, data correctness, and throughput
Contribute to code and design reviews; provide engineering guidance to P3 engineers on the team
5+ years of professional software engineering experience with production backend systems
Strong Go and/or Python proficiency — this team's primary languages
Demonstrated experience building and operating data pipelines or ETL systems in production (streaming or batch)
Working knowledge of message queues or event streaming systems (Kafka, RabbitMQ, or equivalent)
Experience with relational databases (PostgreSQL or equivalent) — schema design, query optimization, migrations
Familiarity with graph databases or graph data models is a strong asset
Understanding of data normalization trade-offs, schema evolution, and backward compatibility
Experience running services on Kubernetes and containerized environments (Docker)
Comfort deploying and debugging distributed systems in AWS or equivalent cloud
Operational mindset: you instrument your code, write runbooks, and own your services through their full lifecycle
Ability to work across team boundaries: you define data contracts with downstream consumers, not just ship to a queue and walk away
Clear written communication — design docs, PR descriptions, and async updates to distributed teammates
Experience in cybersecurity, security tooling, or working with vulnerability or threat data
Exposure to graph databases (Neo4j, DGraph, or equivalent)
Experience with Apache Kafka or similar event streaming platforms at production scale
Background in offensive security tools or pentest workflows — you'll understand the data you're ingesting
We are a fully remote company, and this job may require up to 5% of travel to be successful.
In accordance with various State's transparency regulations, we provide the following salary range information for this position:
Base salary range: $199,750 - $260,000 annually. The exact salary will be determined based on the selected candidate's location, qualifications, experience, and relevant skills.
Inclusive Team: We value diversity and promote an inclusive culture where everyone can thrive.
Competitive Compensation: We offer competitive salary and benefits which includes health, vision & dental care for you and your family, a flexible vacation policy, and generous parental leave.
Horizon3 is not just an equal opportunity employer - we are a community that values diversity, equity, and inclusion as fundamental principles of our culture and success. We are dedicated to fostering a workplace where everyone feels welcome and respected, regardless of race, color, religion, sex, national origin, age, disability, veteran status, sexual orientation, gender identity or expression, genetic information, marital status, or any other legally protected status by law.
Standard company text repeated across Horizon3.ai's postings is omitted here.