SIGN IN
Large Language Model Specialist jobs in United States
info-icon
This job has closed.
company-logo

Bright Vision Technologies · 2 weeks ago

Large Language Model Specialist

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. The company is seeking a Large Language Model Specialist to design and operationalize fine-tuning workflows for large language models, collaborating with cross-functional partners to translate requirements into effective solutions.
Artificial Intelligence (AI)Cyber SecurityInformation TechnologySoftware
badNo H1Bnote

Responsibilities

Design and execute fine-tuning experiments for large language models using supervised, DPO, RLHF, and related techniques
Lead dataset construction, curation, and quality assurance processes for instruction tuning and preference data
Build scalable training pipelines on top of modern distributed training frameworks
Tune hyperparameters, optimizer configurations, and training stability strategies for large-model fine-tuning
Implement parameter-efficient fine-tuning techniques such as LoRA, QLoRA, and adapter-based methods
Design rigorous evaluation suites including automated benchmarks, human evaluation, and capability-specific probes
Implement safety, refusal, and policy evaluations to track model behavior across releases
Operate large-scale training jobs on GPU clusters, diagnosing failures and recovering training state reliably
Optimize training throughput using mixed precision, sequence packing, and efficient attention implementations
Manage model artifacts, lineage tracking, and reproducibility across many concurrent experiments
Collaborate with product, research, and platform teams to align fine-tuning roadmaps with business needs
Document training methodology, results, and decisions clearly for technical and non-technical audiences
Mentor engineers on fine-tuning best practices, evaluation rigor, and responsible deployment
Stay current with LLM research and translate advances into production-ready fine-tuning recipes

Qualification

Large language model fine-tuningSupervised learningDirect preference optimization (DPO)Reinforcement learning with human feedback (RLHF)PythonPyTorchTransformer-based language modelsDistributed training strategiesFully Sharded Data Parallel (FSDP)ZeRO optimizerPipeline parallelismHyperparameter tuningOptimizer configurationTraining stability techniquesParameter-efficient fine-tuning LoRAParameter-efficient fine-tuning QLoRAParameter-efficient fine-tuning adapter methodsEvaluation methodologyBenchmarkingHuman evaluation designGPU cluster training operationsMixed precision trainingSequence packingEfficient attention implementationsModel artifact managementDataset construction and curationSynthetic data generationDataset distillationResponsible AI evaluationRed-teaming practices

Required

Master's or PhD in Computer Science, Machine Learning, or a related field; or equivalent experience
Six or more years of combined ML research and engineering experience, with significant LLM exposure
Strong proficiency in Python and modern deep learning frameworks, especially PyTorch
Hands-on experience fine-tuning transformer-based language models at non-trivial scale
Familiarity with distributed training strategies including FSDP, ZeRO, and pipeline parallelism
Experience with RLHF, DPO, or other preference optimization techniques
Strong understanding of evaluation methodology, benchmarks, and human evaluation design
Experience operating training jobs on GPU clusters and recovering from failures
Strong written and verbal communication skills
Track record of shipping or publishing impactful LLM work

Preferred

Publications at top-tier ML venues
Experience with multimodal model fine-tuning
Familiarity with synthetic data generation and dataset distillation
Open-source contributions to LLM training libraries
Exposure to responsible AI evaluation and red-teaming practices

Benefits

100% Remote (U.S.)
Full-time, Direct W2

Company

Bright Vision Technologies

twitterlinkedincrunchbase
company-logo
Bright Vision Technologies is an information technology company that offers software development, AI, and cybersecurity services.

Funding

Current Stage
Growth Stage
Company data provided by crunchbase