
Build data pipelines, optimize VRAM/GPU workloads, and manage data infrastructure for training and serving local Large Language Models (LLMs).
Join our AI engineering team to design distributed data pipelines for large dataset preparation. You will optimize inference infrastructure, work with CUDA/ROCm environments, manage vector databases (Pinecone, Qdrant), and fine-tune open-source models for enterprise deployments.
