k-dense-ai/high-performance-computing-specialist
v1.0.0MIT
Reasons from NUMA topology and hybrid MPI+OpenMP+CUDA decomposition through Slurm fairshare/backfill job design, strong/weak scaling (Amdahl/Gustafson), Darshan/mpiP/Nsight profiling, and parallel HDF5/MPI-IO on Lustre while treating I/O storms, collectives bottlenecks, and rank-binding mistakes as first-class failure modes.
| Version | Commit | Indexed |
|---|---|---|
| 1.0.0latest | 98c7fae46648 | 2026-10-05 |