Black Forest Labs is a Freiburg-based AI research lab creating FLUX, a high-resolution image generation and editing model for creators, developers and enterprises.
<p data-start="192" data-end="219"><strong data-start="192" data-end="219">About Black Forest Labs</strong></p> <p data-start="221" data-end="611">We're the team behind Latent Diffusion, Stable Diffusion, and FLUX—foundational technologies that changed how the world creates images and video. We’re creating the generative models that power how people make images and video—tools used by millions of creators, developers, and businesses worldwide. Our FLUX models are among the most advanced in the world, and we're just getting started.</p> <p data-start="613" data-end="845">Headquartered in Freiburg, Germany with a growing presence in San Francisco, we're scaling fast while staying true to what makes us different: research excellence, open science, and building technology that expands human creativity.</p> <p data-start="852" data-end="869"><strong data-start="852" data-end="869">Why This Role</strong></p> <p data-start="871" data-end="959">Our research team moves fast. Models improve weekly. New capabilities emerge constantly.</p> <p data-start="961" data-end="1024">What slows us down is not model quality—it’s productionization.</p> <p data-start="1026" data-end="1044">Without this role:</p> <ul data-start="1045" data-end="1230"> <li data-start="1045" data-end="1106"> <p data-start="1047" data-end="1106">Research checkpoints sit longer before becoming usable APIs</p> </li> <li data-start="1107" data-end="1148"> <p data-start="1109" data-end="1148">Inference is slower than it needs to be</p> </li> <li data-start="1149" data-end="1175"> <p data-start="1151" data-end="1175">APIs struggle under load</p> </li> <li data-start="1176" data-end="1230"> <p data-start="1178" data-end="1230">Demos don’t reflect the true potential of our models</p> </li> </ul> <p data-start="1232" data-end="1419">This role removes the bottleneck between frontier research and production reality. Once hired, researchers ship faster, demos launch faster, and customers experience models at their best.</p> <p data-start="1426" data-end="1449"><strong data-start="1426" data-end="1449">What You’ll Work On</strong></p> <p data-start="1451" data-end="1529">You will own the bridge between research breakthroughs and production systems.</p> <ul data-start="1531" data-end="2056"> <li data-start="1531" data-end="1599"> <p data-start="1533" data-end="1599">Turn research checkpoints into production-ready inference services</p> </li> <li data-start="1600" data-end="1672"> <p data-start="1602" data-end="1672">Design and maintain high-performance APIs serving millions of requests</p> </li> <li data-start="1673" data-end="1742"> <p data-start="1675" data-end="1742">Optimize inference latency and throughput across GPU infrastructure</p> </li> <li data-start="1743" data-end="1815"> <p data-start="1745" data-end="1815">Build scalable serving architectures that handle unpredictable traffic</p> </li> <li data-start="1816" data-end="1897"> <p data-start="1818" data-end="1897">Improve reliability, monitoring, and observability across model-serving systems</p> </li> <li data-start="1898" data-end="1974"> <p data-start="1900" data-end="1974">Prototype and ship demos that showcase new capabilities in days, not weeks</p> </li> <li data-start="1975" data-end="2056"> <p data-start="1977" data-end="2056">Collaborate closely with researchers to move from idea to live endpoint rapidly</p> </li> </ul> <h3 data-start="2063" data-end="2119">Tools & Context – Model Serving & API Infrastructure</h3> <ul data-start="2121" data-end="2361"> <li data-start="2121" data-end="2153"> <p data-start="2123" data-end="2153">Python, FastAPI, async systems</p> </li> <li data-start="2154" data-end="2204"> <p data-start="2156" data-end="2204">GPU infrastructure, CUDA, inference optimization</p> </li> <li data-start="2205" data-end="2228"> <p data-start="2207" data-end="2228">Docker and Kubernetes</p> </li> <li data-start="2229" data-end="2271"> <p data-start="2231" data-end="2271">Redis, Postgres, distributed task queues</p> </li> <li data-start="2272" data-end="2310"> <p data-start="2274" data-end="2310">Cloud platforms (AWS, GCP, or Azure)</p> </li> <li data-start="2311" data-end="2361"> <p data-start="2313" data-end="2361">Observability stacks (metrics, logging, tracing)</p> </li> </ul> <p data-start="2363" data-end="2439">This role spans backend systems, GPU performance, and production ML serving.</p> <h3 data-start="2446" data-end="2472">What We’re Looking For</h3> <p data-start="2474" data-end="2724">You’ve built and operated systems at meaningful scale. You understand the difference between a research prototype and a production system. You are comfortable navigating ambiguity, making tradeoffs, and improving systems under real-world constraints.</p> <p data-start="2726" data-end="2742">You demonstrate:</p> <ul data-start="2743" data-end="2992"> <li data-start="2743" data-end="2812"> <p data-start="2745" data-end="2812">Strong judgment around performance, reliability, and cost tradeoffs</p> </li> <li data-start="2813" data-end="2863"> <p data-start="2815" data-end="2863">Experience scaling APIs or ML systems under load</p> </li> <li data-start="2864" data-end="2928"> <p data-start="2866" data-end="2928">Comfort working in fast-moving, research-adjacent environments</p> </li> <li data-start="2929" data-end="2992"> <p data-start="2931" data-end="2992">Ownership from system design through debugging and deployment</p> </li> </ul> <p data-start="2994" data-end="3028">Role-specific experience we value:</p> <ul data-start="3030" data-end="3444"> <li data-start="3030" data-end="3090"> <p data-start="3032" data-end="3090">Building and operating ML inference services in production</p> </li> <li data-start="3091" data-end="3151"> <p data-start="3093" data-end="3151">Designing scalable API architectures with async processing</p> </li> <li data-start="3152" data-end="3222"> <p data-start="3154" data-end="3222">Optimizing GPU workloads (batching, quantization, compilation, CU