Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
dmariko
/
SmolLM-1.7B-Instruct-dpo-16k
Like
0
TensorBoard
Safetensors
English
llama
trl
dpo
Generated from Trainer
License:
cc-by-nc-4.0
Model card
Files
Files and versions
xet
Metrics
Training metrics
Community
Copy to bucket
new
main
SmolLM-1.7B-Instruct-dpo-16k
3.48 GB
Ctrl+K
Ctrl+K
1 contributor
History:
12 commits
dmariko
Training in progress, epoch 6
7930469
verified
about 2 years ago
logs
Training in progress, epoch 0
about 2 years ago
runs
Training in progress, epoch 6
about 2 years ago
.gitattributes
Safe
1.52 kB
initial commit
about 2 years ago
README.md
Safe
3.08 kB
Update README.md
about 2 years ago
config.json
Safe
736 Bytes
Training in progress, epoch 0
about 2 years ago
generation_config.json
Safe
156 Bytes
SmolLM-1.7B-Instruct-dpo-16k
about 2 years ago
merges.txt
Safe
466 kB
SmolLM-1.7B-Instruct-dpo-16k
about 2 years ago
model.safetensors
Safe
3.42 GB
xet
Training in progress, epoch 6
about 2 years ago
special_tokens_map.json
Safe
541 Bytes
SmolLM-1.7B-Instruct-dpo-16k
about 2 years ago
tokenizer.json
Safe
2.1 MB
SmolLM-1.7B-Instruct-dpo-16k
about 2 years ago
tokenizer_config.json
Safe
3.59 kB
SmolLM-1.7B-Instruct-dpo-16k
about 2 years ago
training_args.bin
pickle
Detected Pickle imports (9)
"torch.device"
,
"transformers.training_args.TrainingArguments"
,
"transformers.trainer_utils.SchedulerType"
,
"transformers.trainer_utils.HubStrategy"
,
"accelerate.state.PartialState"
,
"transformers.trainer_pt_utils.AcceleratorConfig"
,
"transformers.trainer_utils.IntervalStrategy"
,
"accelerate.utils.dataclasses.DistributedType"
,
"transformers.training_args.OptimizerNames"
How to fix it?
5.18 kB
xet
Training in progress, epoch 0
about 2 years ago
vocab.json
Safe
801 kB
SmolLM-1.7B-Instruct-dpo-16k
about 2 years ago