The author fine-tuned an open-weights 70B model on a few thousand curated coding examples and reports scores above a popular closed model on two coding benchmarks.
Commenters reproduced the headline numbers but found the gains narrower on unseen repositories, and most of the thread is a useful discussion of which parts of the data mix mattered.