382 parallel-and-distributed-computing "CERN" positions at Oak Ridge National Laboratory
Sort by
Refine Your Search
-
of finite element simulations methods. Experience using parallel Linux computing platforms, parallel job submission scripts, common software repository tools (e.g., GitHub), and parallel visualization
-
Requisition Id 16262 Overview: We are seeking a postdoctoral researcher to work at the intersection of tensor networks, quantum algorithms, scientific computing, topological physics, and quantum
-
distributed intelligence across the computing continuum. In this role, you will have the opportunity to lead and contribute to cutting-edge research aimed at transforming scientific data management and
-
challenges facing the nation. We invite applications for a Principal Engineer for Geospatial Computing Infrastructure. This dynamic and visionary leader will launch and build a next-generation Geospatial Data
-
Requisition Id 16991 Overview: The Field Intelligence Operations Division (FIOD) of the National Security Directorate (NSSD) is seeking a Senior Engineer for a Classified High Performance Computing
-
and cloud computing platforms. Formulating necessary solutions using various parallel computing paradigms and tools, HPC schedulers (such as slurm), Containers and Kubernetes, Python, Bash and other
-
Language Models (LLMs). Distributed Machine Learning: Specialization in data parallelism, model-parallelism, and collective communication strategies in large-scale environments. Proficiency in frameworks
-
supercomputer to break the exascale barrier and Orion, a 700 petabytes Lustre parallel file system. As an important part of the broader High-Performance Computing (HPC) infrastructure, the division also hosts
-
. Demonstrated programming ability and knowledge of Python and/or C++. Experience with deep learning frameworks like PyTorch and application on high-performance computing (HPC) environments using distributed data
-
for Science @ Scale: Pretraining, instruction tuning, continued pretraining, Mixture-of-Experts; distributed training/inference (FSDP, DeepSpeed, Megatron-LM, tensor/sequence parallelism); scalable evaluation