Member of Technical Staff, Inference

San Francisco, CAPosted Jun 15, 2026
Member of Technical Staff, Inference LocationSan FranciscoEmployment TypeFull timeDepartmentR&DAbout UsRadical Numerics is an AI research lab building general biological intelligence. Our mission is to master the code of life, and our purpose is to reduce human suffering.Our team created Evo, and started the field of generative genomics. Our work was featured on the cover of Science, and presented by our CEO on the main stage of TED2025. Evo was used to create the first AI gene therapy tool CRISPR-Cas9, and the first AI whole genome from scratch. Evo 2, featured in Nature, is the largest fully open source AI project across any domain.Radical Numerics is bringing the rigor of distributed systems, model architecture, and numerics research to the challenges of biology. We’ve redesigned the foundation model training stack to turn the world’s raw scientific data (e.g. biological sequences, experiments, and physical processes), into intelligible, generative models that can expand and accelerate what humanity can understand, design, and cure.The same generative breakthroughs that enable life-saving cures also lowers the barrier to creating engineered threats and AI-generated bioweapons. We believe these forces are inseparable. Radical Numerics was founded to develop both the power to design and the responsibility to defend.About the RoleAs a Member of Technical Staff, Inference at Radical Numerics, you will build and optimize the systems that bring frontier biological AI models into production. Your work will focus on delivering state-of-the-art inference performance for large-scale genome and multimodal biological models across a wide range of real-world applications, including therapeutics, diagnostics, synthetic biology, and biodefense.This is a highly technical role at the intersection of AI systems, distributed computing, and model deployment. You will work closely with research, infrastructure, and external partners to ensure our models can be efficiently deployed, scaled, and integrated into production environments. Success in this role requires deep expertise in large language model inference, kernel optimization, GPU systems, and performance engineering.You should be excited by questions such as: How do we reduce inference latency for 100B MoE models? How do we maximize throughput across heterogeneous hardware environments? How do we optimize custom kernels for emerging hybrid model architectures? How do we deploy foundation models reliably across cloud, on-premise, and highly regulated environments? How do we enable our partners to transform biological research and development through production-grade AI systems?What You'll DoDrive end-to-end performance improvements. Identify and eliminate bottlenecks across the inference stack, from model execution and memory management to networking, scheduling, and hardware utilization.Develop high-performance inference primitives. Build and optimize GPU kernels, numerical operators, and serving infrastructure to maximize throughput, latency, and efficiency on modern accelerator platforms.Partner with external customers and collaborators. Work directly with pharmaceutical companies, biotech organizations, research institutions, and government partners to deploy models in production environments and solve challenging technical problems.Build scalable deployment infrastructure. Create systems for serving, monitoring, benchmarking, and operating foundation models reliably across cloud, enterprise, and secure environments.Collaborate with research and platform teams. Ensure new model architectures can be efficiently deployed at scale and help translate frontier AI research into real-world impact.What We're Looking ForExpertise in large-scale AI inference systems. Proven experience optimizing, deploying, and operating LLMs or other foundation models in production environments.Strong performance engineering and kernel development skills. Deep understanding of GPU architectures and experience...

Want jobs like this matched to you?

SimpleCareer scores fresh postings against your résumé so you only see the matches that matter.

Get started free