Peregrine Technologies
Software Engineer, Data Infrastructure
About this role
Peregrine is seeking a Data Infrastructure Engineer to enhance their AI-driven platform aimed at improving public safety outcomes. The role involves designing and optimizing large-scale data systems while collaborating with various teams to ensure efficient data management and integration.
What you'll do
- Design high-throughput, real-time data integration platforms
- Architect scalable storage solutions for vast data volumes
- Build and optimize Apache Spark data processing pipelines
- Enhance performance and reliability of data infrastructure
- Collaborate on data contracts and integration patterns
- Establish best practices for data infrastructure quality
What they're looking for
- Experience with large-scale data infrastructure
- Knowledge of Apache Iceberg and open table formats
- Proficient in Apache Spark for data processing
- Background in real-time data integration with Apache Kafka
- Ability to orchestrate data pipelines using Airflow
- Strong coding skills in Python and/or Scala
- Familiarity with AWS cloud services
- Experience with Kubernetes deployments
Benefits
- Salary range of $160,000 - $220,000
- [Equity options available]
- [Performance bonuses offered]
- [Comprehensive benefits package]
- [Support for personal and professional growth]
- [Working in an innovative team environment]
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
Peregrine Technologies
Peregrine Technologies builds an AI-powered platform that transforms data into actionable intelligence for public safety and federal government organizations. The company is hiring Solutions Engineers, curriculum developers for customer enablement, and software engineers to enhance the platform's user experiences and support its government deployments.
View all jobs at Peregrine TechnologiesLikely interview questions
- Walk us through your experience designing and operating a large-scale data infrastructure system in production. What were the key challenges you faced with ingestion, storage, or serving data at scale?
- Tell us about your hands-on experience with Apache Iceberg. How have you handled schema evolution, partitioning strategies, or compaction in a real production environment?