tinyllama-trl-merged
LORA
This is a fully merged and standalone model of TinyLlama (1.1B parameters) fine-tuned using TRL (Transformer Reinforcement Learning) framework with Lo
This is a fully merged and standalone model of TinyLlama (1.1B parameters) fine-tuned using TRL (Transformer Reinforcement Learning) framework with Lo