this post was submitted on 24 Jul 2026
522 points (98.0% liked)

Programmer Humor

32462 readers
757 users here now

Welcome to Programmer Humor!

This is a place where you can post jokes, memes, humor, etc. related to programming!

For sharing awful code theres also Programming Horror.

Rules

founded 3 years ago
MODERATORS
 
you are viewing a single comment's thread
view the rest of the comments
[โ€“] IndustryStandard@lemmy.world 3 points 20 hours ago (1 children)

That is why you ask for visualizations and test real failure scenarios to see if they are detected properly.

Large language models are not the same as they were 5 years ago. Pumping trillions of dollars to automate software engineering did som what pay off.. They are really quite good nowadays despite the amount of hate they still get. Often they can detect sloppy mistakes by human coders too.

Do not trust them blindly but test their results. Or use them only to generate tests on your artisan handwritten code to see if it can detect any mistakes.

[โ€“] albsen@lemmy.world 1 points 30 minutes ago

I agree, opus 4.8 and Kimi k2.7 are my go to models and opus is very very good if given a proper brief and spec cheat. K2.7 is close.