MOUAD EL ALJ · SYSTEMS ENGINEER · RUST / C++ / CUDA MAR · UTC+1

Three years of inference engines, GPU kernels and production services for a distributed AI compute network. Before that, the CUDA backend of a Python-to-GPU compiler. Every live demo under work runs client-side, in this tab. No backend. No API key.

work

click a row to expand

experience

Hyperspace
2023 — 2026
Software engineer, AI infrastructure. Continuous-batching inference scheduler on llama.cpp; llama.cpp embedded in-process behind a hand-written C ABI with safe Rust bindings; a hardware-verification protocol with bit-identical CPU and GPU (WGSL) implementations, and its production service (Rust, gRPC).
Pyccel, UM6P
2021 — 2023
Software engineer on an open-source compiler translating scientific Python to C, Fortran and CUDA. Built the CUDA backend; helped research groups port numerical code to GPU.

paper

A. El Hachimi, M. Elalj, K. Jbilou, A. Ratnani. “Generalized ℒ-Product for High Order Tensors and Applications Using GPU Computations.” Mathematical Modeling with Modern Applications (M3A 2024), Springer PROMS vol. 497, 2025.

doi:10.1007/978-3-031-89041-3_6
~10 kB over 4 requests, hand-written. The heavy code — the linear algebra, and a language model on the coding-agent page — loads only when you ask for it. © 2026 · visitor #000001