CPU-only page classifier for text extracted from PDFs (no OCR), designed for, "label": one of invoice|lab|radiology|dischargesummary}, split train/val/test. Point steps 2-4 at it with --data-dir; re-calibrate thresholds on your val split before deploying. Note traintransformer.py targets transformers==4.57. — 5.x removed several TrainingArguments kwargs this script uses. - Live-metrics trackio dashboard could not be hosted on this account (Gradio Spaces require PRO), so run metrics live in the job logs.
Independent publisher
Azhar
amohammed3339
Models in Library1
Datasets in Library0
Models on Hugging Face2
Followers—