logoalt Hacker News

autoexec • today at 5:00 AM • 1 reply • view on HN

> AIs when writing code introduce vulnerabilities at a rate similar to humans writing the same code.

AI regurgitating all the insecure code AI companies scraped from stack overflow and github isn't going to give you something too different from what the humans who put it there in the first place came up with. Garbage in, garbage with random hallucinations out.


Replies

red75prime • today at 6:24 AM

This is simplistic to the point of being blatantly wrong. Training data isn't garbage. It's programs that do their job, but that are sprinkled with errors. Uncorrelated errors gets averaged out during autoregressive pretraining. Correlated errors can be somewhat suppressed during post-training. Hallucinations (of the generalization-error kind) can be dealt with using synthetic data that improves the model's generalization.