Ik finetune LLAMA, QWEN of MISTRAL LLMs op jouw data met LoRA

S
sidharth1
S
sidharth1
Sidharth P
Sommige informatie is automatisch vertaald.

Over deze dienst

Automatische vertaling

Wil je een open-source LLM dat echt jouw domein, formaat of stijl kent? Ik finetune het op jouw data en laat je zien, met cijfers, hoe veel beter het geworden is.


Ik ben een AI-onderzoeker (ICML 2026, IJCNLP-AACL 2025) en ik finetune en evalueer open modellen op multi-GPU H100 setups als onderdeel van mijn onderzoek.


Wat je krijgt:

  • Basis model en methode-aanbeveling (LoRA, QLoRA, volledige SFT of GRPO)
  • Data schoonmaken en formatteren in chat- of instructiesjablonen
  • Schoon, reproduceerbare training code en configuratie
  • Getrainde LoRA-adapter, of samengevoegde gewichten op hogere pakketten
  • Voor- en na-evaluatie op een hold-out set
  • Optioneel snelle setup voor service met vLLM of SGLang


Modellen: Llama, Qwen, Mistral, Gemma, DeepSeek en andere Hugging Face modellen.


Stuur me je dataset en doel voordat je bestelt, en ik stel het juiste basis model, methode en pakket voor.

Maak kennis met Sidharth P

Sidharth P

AI Safety Researcher and LLM Fine Tuning Expert

  • Afkomstig uitIndia
  • Lid sindsapr 2017
  • Talen

    Telugu, Engels, Hindi
AI researcher in LLM safety and NLP. Research Intern at AI4Bharat (first author of IndicBERT-v3, open multilingual encoder LLMs) and Research Fellow at SPAR. Papers at ICML 2026 and IJCNLP-AACL 2025, plus arXiv work on memory attacks against LLM agents. Hands-on with LoRA/QLoRA and full fine-tuning (SFT, GRPO), multi-GPU training on H100s, vLLM/SGLang inference and LLM-as-judge evaluation. I help teams fine-tune open-source LLMs, build evaluation pipelines, red-team LLM agents and reproduce ML papers. Message me your goal and I will reply with a clear plan.

Automatische vertaling

Mijn portfolio

Andere AI-development diensten die ik aanbied