📌 Recent Posts:
Benchmarking Local Coding Models
Can IBM’s Granite 4.1 code? I ran it against other local models to find out Full transparency: Except for some initial prompting, most of this blog and the research behind it was AI generated. I did check everything manually though. While I tried to include all relevant files in this repo, reproducibility may be hampered by some of my older Claude contexts leaking into this process (e.g. about how opinionated I can be). ...