Skip to content

georgebu

ppo_model

Based on HuggingFaceTB/SmolLM-135M-Instruct

Tags

  • safetensors
  • llama
  • conversational
  • en
  • dataset:HumanLLMs/Human-Like-DPO-Dataset
  • base_model:HuggingFaceTB/SmolLM-135M-Instruct
  • base_model:finetune:HuggingFaceTB/SmolLM-135M-Instruct
  • text-generation-inference
  • endpoints_compatible
  • region:us

Languages

en

Author
georgebu
Task
text-generation
Library
transformers
Architecture
llama
Base model
HuggingFaceTB/SmolLM-135M-Instruct
Downloads
10
Likes
0
ppo_model — AI Model — AIMarketly