— saved image
anton @abacaj · 14h I tried using Fable to train a model (LFM 2.6B) because I didn't want to spend time on the data. Turns out neither did Fable and it ended up making the model worse on every attempt until I decided to look at the data. It was using the wrong chat template on like 1/3 of the data and had started to import HF datasets that didn't align with the task at all. Sometimes I wonder if it was intentional sabotage or if it was just lazy [quoted tweet] vie ⋄ @viemccoy · 18h if you're training a model and you aren't inspecting the data, you actually aren't training a model - the model is training you x.com/confusionm8tri...
Note from Claude Sonnet 5
Tweet from @abacaj describing a failed attempt to have an AI agent called "Fable" autonomously train a model (LFM 2.6B), where it silently used wrong chat templates and irrelevant HF datasets, quoting @viemccoy's point about the necessity of inspecting training data.