Senior / Staff Software Engineer – Core Data Infrastructure
You will design, build, and optimize high-throughput data processing systems for real-time and batch data. Additionally, you will develop robust storage solutions and automated tooling to ensure system reliability and data quality.
- On-site
- Toronto, ON
- Posted Aug 3, 2026
- 1 position
More jobs you can apply to directly
Similar opportunities posted by employers hiring on Jobs.ca, with no external application form.
Forgeahead Solutions Corporation
Technical Lead and Senior Software Engineer
- On-site
Alcohol and Gaming Commission of Ontario (AGCO)
Information Management Lead / Responsable de la gestion de l’information
- On-site
Dalfen Ltée
Accounting Technician
- On-site
Job summary
The Company We are representing a well-capitalized, fast-growing B2B technology company that builds high-volume data aggregation platforms. Their infrastructure ingests massive streams of sensor and mobility data from hundreds of fragmented third-party sources, unifying it into a single, reliable API. Some of the largest enterprises in the logistics, transportation, and risk-management sectors rely on this backbone for real-time analytics. The engineering team operates out of a Toronto hub, tackling petabyte-scale challenges with the agility of an early-stage startup and the rigor of a mature tech organization. The Role This is a foundational backend and systems engineering position, not a traditional data engineering or pipeline-maintenance role. You will be architecting the core distributed systems that ingest, process, and store terabytes of time-series data daily. You will own the full lifecycle of the platform’s data flow, building the highly available, resilient primitives that internal product teams and external enterprise clients build upon. Core Responsibilities Design, build, and optimize high-throughput data processing systems (both real-time streaming and batch). Engineer robust storage solutions capable of handling rapidly expanding volumes of time-series and spatial data without compromising on query performance. Drive architectural decisions to ensure the infrastructure remains fault-tolerant and ahead of the company's aggressive scaling trajectory. Develop intelligent, automated tooling that improves data quality, lineage tracking, and system reliability. Collaborate closely with internal stakeholders to define technical abstractions that can be reused across different product lines. Write scalable, production-ready code, primarily utilizing Java and Python. Requirements Candidate Profile Scale Experience: A proven track record of designing and maintaining distributed backend systems or data platforms that process data at the terabyte or petabyte scale. Technical Foundation: Deep expertise in computer science fundamentals, system architecture, and anticipating failure modes in complex networks. Language Proficiency: Advanced proficiency in a JVM language (Java, Scala, Kotlin) and a willingness to work across different stacks as needed. Data Ecosystems: Hands-on experience with modern large-scale processing frameworks (e.g., event streaming, distributed computation, and advanced data warehousing/lakehouse concepts). Autonomy: High comfort level navigating ambiguity. You know how to scope complex problems, make definitive architectural calls, and drive projects to completion independently. Bonus Points Active contributions to open-source software, particularly in the distributed systems or data infrastructure space. Familiarity with modern orchestration engines, lakehouse architectures, or spatio-temporal data models. Benefits Work Environment High Autonomy: Engineers own their domains end-to-end and have a direct voice in product direction. Proximity to the User: A culture of speaking directly with customers to understand their friction points before writing a single line of code.
What you’ll do
You will design, build, and optimize high-throughput data processing systems for real-time and batch data. Additionally, you will develop robust storage solutions and automated tooling to ensure system reliability and data quality.
Requirements
Candidates must have a proven track record in designing distributed backend systems at a terabyte or petabyte scale. Proficiency in a JVM language and deep expertise in computer science fundamentals and system architecture are required.
Listed skills
- JavaPreferred
- PythonPreferred
Other relevant skills
Identified from the job description. Confirm important requirements above.
- Java
- Python
- Distributed systems
- Data infrastructure
- System architecture
- Time-series data
- Streaming data
- Batch processing
- Scalability
- Fault-tolerance
- Data warehousing
- Lakehouse architecture
- Spatio-temporal data
- JVM
- API development
- Data lineage
- Data Aggregation
- Resilience
- Technology Ecosystems
- Query Performance
- Java (Programming Language)
- Abstractions
- Application Programming Interface (API)
- Systems Engineering
- Business To Business
- Complex Networks
- Computer Science
- Data Engineering
- Dataflow
- Data Infrastructure
- Data Modeling
- Data Processing Systems
- Data Quality
- Data Warehousing
- Failure Causes
- Fault Tolerance
- Java Virtual Machine (JVM)
- Python (Programming Language)
- Risk Management
- Open-Source Software
- Scala (Programming Language)
- Spatial Data Infrastructures
- Systems Architecture
- Time Series
- Tooling
- Reliability
- Kotlin
Job areas
- Software
- Data & Analytics
- Technology
- Engineering
- Infrastructure Software Engineer
- Software Developer / Engineer
- Software Developers
Additional details
- Minimum experience
- 5+ years
- Posting language
- English
- Working hours
- 40 hours per week