CrowdStrike Logo

CrowdStrike

Principal Software Engineer, Real Time Data Enrichment Platform (Hybrid)

Posted 10 Days Ago
Hybrid
Austin, TX
195K-290K Annually
Expert/Leader
Hybrid
Austin, TX
195K-290K Annually
Expert/Leader
Design, build and own a petabyte-scale self-service real-time data enrichment platform. Lead query processing, ingestion pipelines, storage, indexing and performance optimization using Spark/Flink and Apache ecosystem tools. Deliver unified data catalog, improve query performance and platform efficiency, collaborate on technical reviews, and enable internal and external customers to query and extract data efficiently.
The summary above was generated by AI

As a global leader in cybersecurity, CrowdStrike protects the people, processes and technologies that drive modern organizations. Since 2011, our mission hasn’t changed — we’re here to stop breaches, and we’ve redefined modern security with the world’s most advanced AI-native platform. We work on large scale distributed systems, processing almost 3 trillion events per day and this traffic is growing daily. Our customers span all industries, and they count on CrowdStrike to keep their businesses running, their communities safe and their lives moving forward. We're proud to work for a mission-driven company leveraging AI to transform the way we work. CrowdStrikers drive their careers through flexibility and autonomy while also being expected to contribute to a culture of responsible AI adoption, experimentation, and innovation. We use an AI-first mindset as a force multiplier to proactively and continuously accelerate execution, build expertise, uncover insights, and solve complex problems. We’re always looking to add talented CrowdStrikers to the team who have limitless passion, a relentless focus on innovation and a fanatical commitment to our customers, our community and each other. Ready to join a mission that matters? The future of cybersecurity starts with you.


About the Role:

Our Data Platform group at Crowdstrike is unique among its kind for being uncommonly customer-focused. We build and operate systems to centralize all of the data from Falcon Sensors and 3rd Party sources derived from trillions of events/day, and we also drive industry-leading innovation on a hyper scale security data lake that helps find bad actors and stop breaches.

 

Setting ourselves apart, we make it easy for all customers to utilize the platform for batch and streaming analytics, machine learning and threat hunting through the production and delivery of self-service platforms (Query Platform, Analytics Platform, Enrichment platform...) using Spark and Flink respectively as the runtimes. These self-service platforms allow customers to apply their own custom schema, syntax, data models (etc.) to our historical cyberattack data, modeling threats in a way that empowers them to predictively build (among other things) behavioral automations that defend against such threats 'before' they appear in their own environments.

In this role, you will be a Principal Engineer within Data Platform, owning - in entirely hands-on capacity the design, build and delivery of a new self-service data enrichment platform which will take our customers' proactive defense measures to the next level.

 

As a leader in data platform, you will contribute to the full spectrum of our systems, including query processing, scalable pipeline builds with largely Apache-based ingestion, materialized view, transformation and data storage frameworks, and tools/applications that make data available to thousands of users and hundreds of internal systems.
What You’ll Do:

  • Shape the vision of our Analytics Data Platform for its next phase of growth: building a Unified Data Catalog as well as Query Analysis for structured data stored in different forms (columnar vs graph), and building and optimizing query performance using different techniques of indexing and data partitioning.

  • Design, develop, and maintain a data platform that processes petabytes of data.

  • Participate in technical reviews of our products and help us develop new features and enhance stability.

  • Continually help us improve the efficiency of our services so that we can delight our customers.

  • Help us research, evolve and implement new ways for both internal stakeholders as well as customers to query their data efficiently and extract results in the format they desire.

What You’ll Need:

  • (One among:) 17+ years exp with B.S. in a related field, 15+ years with M.S. in a related field, or 12+ years with PhD in a related field.

  • Experience building and supporting very high scale data platform and data storage systems (MINIMUM: 100s of TB/day in either a current or past role)

  • Significant experience performance-tuning or developing internals (source code) within Spark, Flink, Iceberg or Pinot or an equivalent structured streaming, Time Series, OLAP or Open Table real time system.

  • Production experience building either Spark- or Flink-based self-service data platforms, or equivalent (i.e. with Ray, or building spark- or Flink-like frameworks themselves in Scala, Akk, etc.)

  • Strong familiarity with (and ample hands-on experience tuning & optimizing) at least one applicable technology in the Apache Hadoop ecosystem: Spark, Kafka, Hive/Iceberg/Delta Lake, Presto/Trino, Pinot, Druid, etc.

  • 3+ years coding in Java, Scala, Kotlin or another JVM language (bonus points for experience tuning the language, i.e. garbage collection, memory management…)

  • Production experience with relational SQL and NoSQL databases, including Postgres/MySQL, Cassandra/DynamoDB, etc.

  • Proven expertise with multiple big data frameworks in general, especially handling data volume at (ideally) multi-petabyte scale.

  • Proven expertise with algorithms, distributed systems design and the software development lifecycle.

  • Great test driven development discipline.

  • Reasonable proficiency with Linux administration tools.

  • Proven experience utilizing AI technologies to enhance decision-making, streamline workflows and processes, improve efficiency and drive business outcomes.

  • Proven ability to work effectively with remote teams.

 

Bonus Points:

  • Familiarity with Go.

  • Familiarity with Kubernetes/Mesos or equivalent.

  • Production experience with Flink, especially as runtime for a self-service platform.

 

#LI-MP2

#LI-SF1

Benefits of Working at CrowdStrike:

  • Market leader in compensation and equity awards

  • Comprehensive physical and mental wellness programs 

  • Competitive vacation and holidays for recharge  

  • Paid parental and adoption leaves

  • Professional development opportunities for all employees regardless of level or role

  • Employee Networks, geographic neighborhood groups, and volunteer opportunities to build connections

  • Vibrant office culture with world class amenities

  • Great Place to Work Certified™ across the globe

CrowdStrike is proud to be an equal opportunity employer. We are committed to fostering a culture of belonging where everyone is valued for who they are and empowered to succeed. We support veterans and individuals with disabilities through our affirmative action program.

CrowdStrike is committed to providing equal employment opportunity for all employees and applicants for employment. The Company does not discriminate in employment opportunities or practices on the basis of race, color, creed, ethnicity, religion, sex (including pregnancy or pregnancy-related medical conditions), sexual orientation, gender identity, marital or family status, veteran status, age, national origin, ancestry, physical disability (including HIV and AIDS), mental disability, medical condition, genetic information, membership or activity in a local human rights commission, status with regard to public assistance, or any other characteristic protected by law. We base all employment decisions--including recruitment, selection, training, compensation, benefits, discipline, promotions, transfers, lay-offs, return from lay-off, terminations and social/recreational programs--on valid job requirements.

If you need assistance accessing or reviewing the information on this website or need help submitting an application for employment or requesting an accommodation, please contact us at [email protected] for further assistance.

Find out more about your rights as an applicant.

CrowdStrike participates in the E-Verify program.

Notice of E-Verify Participation

Right to Work

CrowdStrike, Inc. is committed to fair and equitable compensation practices. Placement within the pay range is dependent on a variety of factors including, but not limited to, relevant work experience, skills, certifications, job level, supervisory status, and location. The base salary range for this position for all U.S. candidates is $195,000 - $290,000 per year, with eligibility for bonuses, equity grants and a comprehensive benefits package that includes health insurance, 401k and paid time off.

For detailed information about the U.S. benefits package, please click here

Similar Jobs at CrowdStrike

Yesterday
Remote or Hybrid
USA
115K-160K Annually
Senior level
115K-160K Annually
Senior level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Lead and deliver generative-AI enabled source code review application security engagements, coordinate stakeholders, mentor junior red team members, document recommendations, and perform/manager penetration and web application tests.
Top Skills: Burp SuiteGenerative AiNessus
Yesterday
Remote or Hybrid
USA
125K-180K Annually
Senior level
125K-180K Annually
Senior level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Lead and mentor a distributed Salesforce engineering team to design, build and deliver Lead-to-Cash/Quote-to-Cash sales and partner experiences. Own roadmap, requirements, architecture, code reviews, integrations, data migrations, Agile delivery, budgets and resource allocation. Provide technical leadership in Apex, LWC and Salesforce Revenue Cloud/CPQ, leverage AI for faster delivery, and collaborate with business stakeholders to drive scalable, secure solutions that improve seller and partner experiences.
Top Skills: Ai ToolsApexCpqEcommerceIntegrationsLightning Web Components (Lwc)Sales CloudSalesforceSalesforce Revenue CloudSOQLVisualforce
Yesterday
Remote or Hybrid
USA
100K-155K Annually
Senior level
100K-155K Annually
Senior level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Lead large-scale collection and analysis of deep and dark web data, build and maintain production-grade automated collection systems (LLM-integrated), operate Docker/cloud-based pipelines and APIs, respond to collection failures and incidents, perform threat actor research, and apply OPSEC tradecraft while collaborating globally.
Top Skills: Agent WorkflowsCi/CdClaude CodeCloud EnvironmentsData PipelinesDatabasesDockerGitLarge Language ModelsMcp ServersPythonRestful ApisWeb Scraping

What you need to know about the Colorado Tech Scene

With a business-friendly climate and research universities like CU Boulder and Colorado State, Colorado has made a name for itself as a startup ecosystem. The state boasts a skilled workforce and high quality of life thanks to its affordable housing, vibrant cultural scene and unparalleled opportunities for outdoor recreation. Colorado is also home to the National Renewable Energy Laboratory, helping cement its status as a hub for renewable energy innovation.

Key Facts About Colorado Tech

  • Number of Tech Workers: 260,000; 8.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lockheed Martin, Century Link, Comcast, BAE Systems, Level 3
  • Key Industries: Software, artificial intelligence, aerospace, e-commerce, fintech, healthtech
  • Funding Landscape: $4.9 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Access Venture Partners, Ridgeline Ventures, Techstars, Blackhorn Ventures
  • Research Centers and Universities: Colorado School of Mines, University of Colorado Boulder, University of Denver, Colorado State University, Mesa Laboratory, Space Science Institute, National Center for Atmospheric Research, National Renewable Energy Laboratory, Gottlieb Institute

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account