LLM Trainer
Turing Β· Remote (Global)
-
Designed high-quality SFT trajectories to improve reasoning accuracy and instruction adherence in large language models.
-
Performed deep qualitative evaluation, identifying logical gaps and hallucinations across diverse model responses.
-
Contributed to enterprise-scale training workflows for Alibaba, supporting production-grade model improvements.