I am a researcher in high-performance computing (HPC) with a focus on systems programming, communication optimization, and scalable distributed systems. My work explores how to improve performance and efficiency in modern HPC environments through system-level design and communication offloading.

Research Focus

  • High-performance computing (HPC)
  • MPI, OpenMP, GPU offloading with OpenMP, CUDA, and OpenACC
  • High-performance communication (RDMA, InfiniBand, UCX, UCC)
  • System-level programming (POSIX APIs, IPC, memory management, system calls)
  • HPC resource management and scheduling (Slurm)
  • Smart network devices and accelerators (NVIDIA BlueField DPUs, NPUs, GPUs)
  • Performance profiling and analysis (Perf, Valgrind, LIKWID, Nsight Systems (nsys), Nsight Compute (ncu))
  • HPC benchmarking and evaluation (MPI Benchmarks, STREAM, LINPACK, HPL)
  • Simulation and modeling of distributed systems (SimGrid)