mrsonord/evolen-200k-sft-checkpoints
text-classification model by mrsonord
From the publisher
Model card & documentation
Source previewRead the publisher’s intended use, setup instructions, evaluations and limitations. The original model card is the source of truth.
Fine-tuned classification checkpoints for all 56 downstream tasks, produced from the EvoLen pretrained model at pretraining step 200,000. Base model: mtapiapacheco/len25120 (merge-len2 tokenizer, vocab 5120). One folder per task, //: Each contains model.safetensors, config.json, tokenizer.json, tokenizerconfig.json, specialtokensmap.json, trainerstate.json and trainingargs.bin. Optimizer and scheduler state are stripped. trainingargs.bin holds the complete TrainingArguments for that run, so every hyperparameter is recoverable from the checkpoint itself: Evaluation only — no retraining. Four things must match or you will get close-but-wrong values: the task name. Read it from trainingargs.bin. split file. For multi-SCREEN/allchrcsv that is ['CA','CA-CTCF','CA-H3K4me3','CA-TF','PLS','TF','dELS','pELS'] → 0..7. Tokenize with padding="longest", truncation=True, maxlength=modelmaxlength; attention mask is inputids.ne(padtokenid). Metric is sklearn.metrics.matthewscorrcoef over argmax predictions. Verified environment: Python 3.9.18, transformers 4.35.2, scikit-learn 1.6.1, torch 2.8.0, numpy 2.0.2, tokenizers 0.15.2, accelerate 0.25.0. Each run trained for numtrainepochs with per-epoch validation on dev.csv. loadbestmodelatend=True with metricforbestmodel=evalf1 reloaded the checkpoint with the best validation F1; that checkpoint was then…
Read the full model card ↗ · Preview checked 2026-10-02T08:52:30.187Z
Inside the original model card — Document outline
- EvoLen-200K — final SFT checkpoints
- Layout
- Loading
- Reproducing the reported numbers
- How these checkpoints were selected
Headings are captured from the source. Links open the publisher’s document, not a locally hosted copy.
Documentation belongs to its respective authors. Reported project/model license: mit. A listing is not a grant of reuse or training rights. Confirm the document’s own terms at the source.
Model overview
mrsonord/evolen-200k-sft-checkpoints is a text-classification model repository published by mrsonord on Hugging Face. The source reports the transformers library.
This page summarizes Hub metadata. For intended use, training data, evaluation results and limitations, consult the original model card.
Model facts
- Task
- text-classification
- Library
- transformers
- Recent downloads (30 days)
- 0
- Cumulative likes
- 0
- Hugging Face trending score
- 0
- Reported safetensors parameters
- Not reported
- Architecture
- Not reported
- License
- mit
- Access
- Not gated by Hugging Face
- Created
- 2026-10-02T06:23:16.000Z
- Last modified
- 2026-10-02T06:23:16.000Z
Compatibility and lineage
Language coverage was not reported.
Base-model lineage was not included.
Parameter count is not a RAM/VRAM requirement. Check precision, quantization, context length and runtime compatibility in the model card; no hardware or API-cost claim is inferred here.
Use and evaluate
Check license terms and access requirements first. Review the model card’s documented loading instructions, evaluations and safety limitations. Benchmark scores are not imported by this directory, and popularity is not an accuracy ranking.
No model weights are downloaded or executed by AltAPIs. Never enable remote model code without reviewing it.
Source and freshness
Source: Hugging Face Hub. Metadata observed 2026-10-02T06:39:29.113Z. Daily imports are snapshots, not real-time monitoring.
Popularity and source listings do not establish security, suitability, licensing rights or benchmark performance.