Remote or South Carolina
•
Today
NVIDIA is seeking a Python Software Engineer to further our efforts to GPU-accelerate data engineering for Large Language Model (LLM) tools and libraries. This role is pivotal in accelerating preprocessing pipelines for high-quality multi-modal dataset curation. The day-to-day focus is on developing efficient, scalable systems for deduplicating, filtering, and classifying training corpora for foundation model LLMs, as well as ingesting and prepping datasets for use in Retrieval Augmented Generat
Full-time
USD 148,000.00 per year