A technical deep-dive that cites and reproduces the Model DNA method in PyTorch - verifying 'trained from scratch' LLM claims via architecture config, tokenizer overlap, and Linear CKA, with a candid look at its strengths and limits.

A Blog post by Proto_AGI on Hugging Face

How to Tell If an LLM Was Really Trained From Scratch: A Reproducible Fingerprinting...