I am a researcher in high-performance computing (HPC) with a focus on systems programming, communication optimization, and scalable distributed systems. My work explores how to improve performance and efficiency in modern HPC environments through system-level design and communication offloading.
Research Focus
- High-performance computing (HPC)
- MPI, OpenMP, GPU offloading with OpenMP, CUDA, and OpenACC
- High-performance communication (RDMA, InfiniBand, UCX, UCC)
- System-level programming (POSIX APIs, IPC, memory management, system calls)
- HPC resource management and scheduling (Slurm)
- Smart network devices and accelerators (NVIDIA BlueField DPUs, NPUs, GPUs)
- Performance profiling and analysis (Perf, Valgrind, LIKWID, Nsight Systems (nsys), Nsight Compute (ncu))
- HPC benchmarking and evaluation (MPI Benchmarks, STREAM, LINPACK, HPL)
- Simulation and modeling of distributed systems (SimGrid)