Sort by
Refine Your Search
-
Category
-
Employer
-
Field
-
dependencies for execution on Alps / CSCS infrastructure. Run and monitor Slurm-based training and evaluation jobs. Debug failures related to distributed execution, checkpointing, filesystem performance
-
of vector query processing, serverless/just-in-time query execution, parallel and distributed data management, and data management systems on emerging hardware Passionate about inventing, building, and
-
This means experience with frameworks such as veRL, slime, Megatron-LM, DeepSpeed, TRL, vLLM, SGLang, or similar tools Strongly preferred Familiarity with distributed training concepts such as data parallelism
-
images against reference pre-training and post-training workloads together with Apertus engineers, and maintain working launch examples Compute partnership and efficiency Serve as the primary technical
-
80%-100%, Lugano, fixed-term The Swiss National Supercomputing Centre (CSCS) develops and operates a high-performance computing and data research infrastructure that supports world-class science in
-
and operate this infrastructure. Job description You will play a key role in automating and operating a large-scale, distributed computing platform. You will work with modern open-source technologies in
-
of the Information Age - One high bandwidth fiber optic cable can replace thousands of electrical links to transmit the same amount of information. As a parallel to the thriving fields of soft devices and robots
-
a major role in the advent of the Information Age - One high bandwidth fiber optic cable can replace thousands of electrical links to transmit the same amount of information. As a parallel to the