waypointjobs

Bright Vision Technologies

Large Language Model Specialist

Remote — United States (see country and timezone requirements)

Check who can apply and the requirements below before continuing.

Job description

Large Language Model Specialist - Remote

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.

This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.

Job Title: Large Language Model Specialist

Location: 100% Remote (U.S.)

Position Type: Full-time, Direct W2

Salary Range: $100,000–$150,000 Annually

Experience Required: 6+ years

Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.

Job Summary

We are looking for an LLM Fine-Tuning Engineer to design, execute, and operationalize fine-tuning workflows for large language models across supervised, preference-based, and reinforcement learning approaches. The role requires deep practical experience with modern training stacks, careful dataset construction, rigorous evaluation methodology, and the engineering discipline to operate complex training pipelines reliably. The ideal candidate combines strong ML intuition with production-grade engineering practices, and is comfortable navigating the trade-offs between data quality, compute budget, evaluation rigor, and shipping velocity. In this role you will work closely with cross-functional partners — product, design, engineering, operations, and business stakeholders — to translate ambiguous requirements into well-engineered solutions, and will be expected to raise the bar through code review, design review, and mentorship of more junior engineers. The successful candidate brings strong engineering discipline, a clear communication style, and a track record of shipping meaningful work that holds up well in production.

Key Responsibilities

Design and execute fine-tuning experiments for large language models using supervised, DPO, RLHF, and related techniques.

Lead dataset construction, curation, and quality assurance processes for instruction tuning and preference data.

Build scalable training pipelines on top of modern distributed training frameworks.

Tune hyperparameters, optimizer configurations, and training stability strategies for large-model fine-tuning.

Implement parameter-efficient fine-tuning techniques such as LoRA, QLoRA, and adapter-based methods.

Design rigorous evaluation suites including automated benchmarks, human evaluation, and capability-specific probes.

Implement safety, refusal, and policy evaluations to track model behavior across releases.

Operate large-scale training jobs on GPU clusters, diagnosing failures and recovering training state reliably.

Optimize training throughput using mixed precision, sequence packing, and efficient attention implementations.

Manage model artifacts, lineage tracking, and reproducibility across many concurrent experiments.

Collaborate with product, research, and platform teams to align fine-tuning roadmaps with business needs.

Document training methodology, results, and decisions clearly for technical and non-technical audiences.

Mentor engineers on fine-tuning best practices, evaluation rigor, and responsible deployment.

Stay current with LLM research and translate advances into production-ready fine-tuning recipes.

Required Qualifications

Master’s or PhD in Computer Science, Machine Learning, or a related field; or equivalent experience.

Six or more years of combined ML research and engineering experience, with significant LLM exposure.

Strong proficiency in Python and modern deep learning frameworks, especially PyTorch.

Hands-on experience fine-tuning transformer-based language models at non-trivial scale.

Familiarity with distributed training strategies including FSDP, ZeRO, and pipeline parallelism.

Experience with RLHF, DPO, or other preference optimization techniques.

Strong understanding of evaluation methodology, benchmarks, and human evaluation design.

Experience operating training jobs on GPU clusters and recovering from failures.

Strong written and verbal communication skills.

Track record of shipping or publishing impactful LLM work.

Preferred Qualifications

Publications at top-tier ML venues.

Experience with multimodal model fine-tuning.

Familiarity with synthetic data generation and dataset distillation.

Open-source contributions to LLM training libraries.

Exposure to responsible AI evaluation and red-teaming practices.

How to Apply

Would you like to know more about this opportunity? For immediate consideration, please send your resume to  or contact us at (908) 505-3544. Learn more about Bright Vision Technologies at .

Bright Vision Technologies is an Equal Opportunity Employer.

Equal Employment Opportunity (EEO) Statement

Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.

BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Originally posted on Himalayas

Who can apply

Eligible countries: United States. Accepted UTC offsets: UTC-10, UTC-9, UTC-8, UTC-7, UTC-6, UTC-5, UTC+14. Review the full description for employer-specific work authorization, residency and schedule requirements.

Ready for your next step?Apply on the official website
Apply on Himalayas ↗

Explore related searches

Current related jobs

Red Wine and Blue

Himalayas

General Interest Application

Remote — United States (see country and timezone requirements)

Salary not specifiedfull timeRemote

WHO THE HECK ARE WE? Red Wine & Blue is a national community of over 600,000 diverse suburban women working together to defeat extremism, one fr…

Listing expires 2026-12-04View job

Carle Health

Himalayas

Finance Systems Analyst

Remote — United States (see country and timezone requirements)

$26.41 – $44.10 per hourfull timeRemote

OverviewThe Finance Systems Analyst assists with supporting assigned finance/accounting applications for the enterprise such as costing, producti…

Listing expires 2026-12-04View job

Kyowa Kirin

Himalayas

Scientific Relations Manager

Remote — Worldwide (see timezone requirements)

Salary not specifiedfull timeRemote

OverviewWE PUSH THE BOUNDARIES OF MEDICINE. LEAPING FORWARD TO MAKE PEOPLE SMILE At Kyowa Kirin International (KKI), our purpose is to make peop…

Listing expires 2026-12-04View job

Belden, Inc

Himalayas

Solution Consultant - Cybersecurity Practice (US)

Remote — United States (see country and timezone requirements)

$125,000.00 – $160,000.00 per yearfull timeRemote

Innovation Starts With YouPropel your career at Belden, where innovation creates possibilities—for our people, our customers, and the communities…

Listing expires 2026-12-04View job