How CrowdStrike Trains GenAI Models at Scale Using Distributed Computing
ID: cd852f75-1970-5071-a676-e50c1c13dadc
STIX ID: report--cd852f75-1970-5071-a676-e50c1c13dadc
Feed Name: Crowdstrike Blog
Date Published: 2025-12-22
Date Updated: 2026-04-27
Author: Andrei Preda - Alexandru Dinu - Florian Stortz - Nathan Nusaputra - Catalin-Andrei Stan
This document outlines CrowdStrike’s strategy for building cybersecurity-focused LLMs, detailing scalable training infrastructure on Google Cloud’s Vertex Training Cluster with Slurm, modular data pipelines with synthetic data generation, and distributed training using data/tensor/pipeline/context/expert parallelism. It highlights hardware-aware optimization by comparing attention mechanisms (FlashAttention 2 vs. SDPA) on NVIDIA H100 and B200 GPUs, underscoring that meaningful performance gains require tailoring to specific software–hardware configurations to accelerate next-generation security LLM development.
Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.
