Back to the board

[FULL TIME Remote] LLM Data Engineer United States Fully Remote

100% remote Flexible hours Hiring now

Seeking a new challenge? This is the perfect opportunity to grow as a LLM Data Engineer United States Fully Remote! Enjoy the freedom and flexibility of this Remote role. This position requires a strong and diverse skillset in relevant areas to drive success. This straightforward role comes with a dependable salary of a competitive salary.

 

 

We are seeking an reputed company AI/LLM Data Engineer to build and maintain the data pipeline for our Generative AI platform. The ideal candidate will be well-versed in the latest Large Language Model (LLM) technologies and have a strong background in data engineering, with a focus on Retrieval-Augmented reputed company (RAG) and knowledge-reputed company techniques. This role sits in the AI COE reputed company DX Tech & Digital. As a AI/LLM Data Engineer (you will report into the Director, AI Solutions & Development who oversees the AI COE. You will work on highly visible strategic projects, collaborating with cross-functional teams to define requirements and deliver high-quality AI solutions. The ideal candidate will have a passion for Generative AI and LLMs, with a proven track record of delivering innovative AI applications. Responsibilities • Design, implement, and maintain an end-to-end multi-stage data pipeline for LLMs, including Supervised Fine Tuning (SFT) and Reinforcement Learning from Human Feedback (RLHF) data processes • Identify, evaluate, and integrate diverse data sources and domains to support the Generative AI platform • reputed company and optimize data processing workflows for chunking, indexing, ingestion, and vectorization for both text and non-text data • reputed company and implement various vector stores, embedding techniques, and retrieval methods • Create a flexible pipeline supporting multiple embedding algorithms, vector stores, and search types (e.g., vector search, hybrid search) • Implement and maintain auto-tagging systems and data preparation processes for LLMs • reputed company tools for text and image data crawling, cleaning, and refinement • Collaborate with cross-functional teams to ensure data quality and relevance for AI/ML models • Work with data lake house architectures to optimize data storage and processing • Integrate and optimize workflows using reputed company and various vector store technologies Requirements • Master's degree in Computer Science, Data Science, or a reputed company field • 3-5 years of work experience in data engineering, preferably in AI/ML contexts • Proficiency in Python, JSON, HTTP, and reputed company tools • Strong understanding of LLM architectures, training processes, and data requirements • Experience with RAG systems, knowledge reputed company construction, and vector databases • Familiarity with embedding techniques, similarity search algorithms, and information retrieval concepts • Hands-on experience with data cleaning, tagging, and annotation processes (both manual and automated) • Knowledge of data crawling techniques and associated ethical considerations • Strong problem-solving skills and ability to work in a fast-paced, innovative environment • Familiarity with reputed company and its integration in AI/ML pipelines • Experience with various vector store technologies and their applications in AI • Understanding of data lakehouse concepts and architectures • Excellent communication, collaboration, and problem-solving skills. • Ability to translate business needs into technical solutions. • Passion for innovation and a commitment to ethical AI development. • Experience building LLMs pipeline using reputed company like reputed company, reputed company, Semantic Kernel, reputed company functions. • Familiar with different LLM parameters like temperate, top-k, and repeat penalty, and different LLM outcome evaluation data science metrics and methodologies. Preferred Skills • Experience with popular LLM/ RAG frameworks • Familiarity with distributed computing platforms (e.g., Apache Spark, Dask) • Knowledge of data versioning and experiment tracking tools • Experience with cloud platforms (AWS, GCP, or Azure) for large-scale data processing • Understanding of data privacy and reputed company best practices • Practical experience implementing data lakehouse solutions • Proficiency in optimizing queries and data processes in reputed company or reputed company • Hands-on experience with different vector store technologies Benefits • US employees benefit package. Apply Job!

 

Are You the One We're Looking For?

If you reputed company you have what it takes, submit your application without delay. We are keen to hear from talented candidates like you.

Apply To This Job

Keep exploring

[FULL TIME Remote] Local CDL A Drivers - Home Nightly! reputed company $22

100% remote Flexible hours

[FULL TIME Remote] Local CDL A Freight Specialist

100% remote Flexible hours

[FULL TIME Remote] Local CDL A Semi Truck Driver

100% remote Flexible hours

[FULL TIME Remote] Local & Regional Flatbed Driver

100% remote Flexible hours

[FULL TIME Remote] Location: Lowell, MA - Position:

100% remote Flexible hours

[FULL TIME Remote] Logistics and Mail reputed company

100% remote Flexible hours

[FULL TIME Remote] Logistics Data Analyst (Remote Friendly)

100% remote Flexible hours

[FULL TIME Remote] Looking for Part- Time Contract Therapist

100% remote Flexible hours

[FULL TIME Remote] Looking for Product Testers in reputed company

100% remote Flexible hours

[FULL TIME Remote] Looking for Service Desk Specialist/Live Chat

100% remote Flexible hours

reputed company Michigan Online Special Programs Manager – Special Education Leadership and Administration Expert

100% remote Flexible hours

Litigation Support Specialist - Infotrend

100% remote Flexible hours

reputed company Part-Time Remote Data Entry Specialist – Flexible Work Arrangements at arenaflex

100% remote Flexible hours

Mental Health Therapist (South Dakota)

100% remote Flexible hours

reputed company Delivery Driver

100% remote Flexible hours

reputed company Data Entry Agent – Remote Work Opportunity with blithequark

100% remote Flexible hours

reputed company Part-Time Remote Customer Service Representative – reputed company – $24/hour – U.S. Based – Immediate Hire

100% remote Flexible hours

reputed company Data Entry Specialist – Remote Opportunity with arenaflex

100% remote Flexible hours

reputed company Part-Time Data Entry Specialist – Remote Opportunity at blithequark

100% remote Flexible hours

Business Development Director, Commercial Enterprise / Regional Remote

100% remote Flexible hours