stealthstack.ai
Back to results

trl-training

Skill
outshift.io · via agntcy registry Unverified — relayed by outshift.io seen 5h ago

About

Train and fine-tune transformer language models using TRL (Transformers Reinforcement Learning). Supports SFT, DPO, GRPO, KTO, RLOO and Reward Model training via CLI commands.

Capabilities

The crawler did not record capability metadata for this resource. Inspect the endpoint directly to see what it exposes.

Provenance

Discovered Relayed by agntcy
URN authority urn:air:outshift.io:agntcy:trl-training
Catalog host outshift.io
Anchor check Not anchored
Last crawled seen 5h ago

Tags

e learninglarge language modelsdistributed trainingmodel fine tuningpreference alignmentlanguage generationtext completiontest case generation