Hacker Newsnew | past | comments | ask | show | jobs | submit | karmasimida's commentslogin

Let's be honest with ourselves, AI review is better than human right now, especially top tier models. Not saying that for important project like Chrome, you can skip reviewing changes. Far from it. But it is probably better and more thorough than most human devs already.

It’s not better than human review. It’s good at catching low-hanging issues but it is often requested overly-defensive code to handle edge cases that don’t exist or are handled elsewhere in the stack. It generates a ton of noise with low-signal comments. It basically cannot at all identify problems that pose the highest risk, like integrations, correctness from a business perspective, etc.

To be fair it’s also worth noting that it’s much easier to find a buggy edge case with existing code than it is to write bug free code that doesn’t have any edge cases at all. It’s so much easier to read some concrete logic and find holes in it, than it is to start from nothing and end up with perfection. It’s true both for humans and agents, but agents are better are validating correctness.

Yes this is why models are superior here: they can equally (opportunity wise) attend everything in their context window.

Ehhhh, it's better at some things and not others.

My own reviews have shifted now. I tend not to examine detailed semantics anymore. The AI is as good or better than me at assessing whether a chunk of code does what the author said it was supposed to do.

Instead my job is to spot design and architecture smells, broader semantic errors, violations of unspoken business requirements, etc.

For example, I was recently reviewing code that built out a transactional flow. Part of that flow involved recording the transaction somewhere user visible and I knew that should only happen after the transaction was confirmed. AI implemented it where the transaction was posted.

That starts as an issue of underspecified requirements but that always happens in the real world. Thus that's where I can provide the most valuable insight: assessing with that context, whether business, operational, historical, or forward looking.


I wonder, how much dogfooding did they do in this run?

I believe this is true. The implication would be more interesting though.

1. Some voice will start calling for banning DEPLOYMENT of open source models in US. Simply hosting them will become regulated, or at least USG will attempt to do so.

2. Future GPT-6+ models will be gated, like really gated. That day will come in a year. If a model is believed to be this capable, there will be some middle level agency built to secure that the access of the model will only be provided to trust personnels.

Business is going to be conducted at a different level


You forgot hardware limitations and locks so you can't run your own models.

Because the model capability is beyond their expectation.

This is brilliant marketing but I think it is real.


Interestingly OpenAI benchmarking 'an even more capable pre-release model' lines up with rumors of GPT-6 releasing in early August.

I hope that with the existing safety guardrails in place, they can roll it out to all users.


I mean we already see models exploit people's misunderstanding of how Docker works to get root without using su. And if you are one of the lucky people in cyber security that has been given a fat stack of tokens by the model providers you get to see some pretty wild exploit chains get put together by the models. Models are much better at detecting insecure code than writing actual secure code at this point.

ChatGPT has replaced 90% of Google for me ... that is another 1T source of income for OpenAI

Is it? 2.8T isn't open in the open source sense


It is possible because they hired the original author of Bun to do this.

It won't be possible to HIRE anyone doing this job confidently, they won't even know if they f*ked up.

I am not too worried anyway, one person with no context, with or without AI, can't do this job. Just like with all the AIs you have, giving you 10 millions, you still can't build a AAA game, it is that simple.


He is close or already a billionaire, not sure much more money will be do much heavy-lifting


you'd be surprised! people seem to have a limitless appetite for that money stuff. they just can't get enough of it, i've found


I know some pretty wealthy people. They are very aware of those who are 10x wealthier than them. If Noam has 1B, he is probably pretty aware of those that have 10B. He's met them and seen their properties, scope, and powers. Likewise, they are thinking about those that have 100B, and those are thinking about Elon, who now has "four commas."


That's not really Noam's style


Most are happy and stop at multimillionaires, but of course we don't hear about them. The focus on hungry billionaires is survivor bias.

We don't hear about Tom from MySpace.


For many business people, money is just a measure of status after becoming rich.

Maybe Noam measures status differently.


How much money do you have? or are you just commenting from the peanut gallery?


i’ve got 300k in a 401k, 50k in cash and i earn 230k pretax. i’m not sure what the point of the question is tho


My point is you and I are still working for money because we want things/security etc.. I really don't get the sense that people who have tens or hundres of millions of dollars are doing it cause they need more money. It's other things that motivate people. I run into people in my workplace that just enjoy it, even though they're sitting on 20M.

All those engineers making 20M a year at Anthropic and OpenAI are going to back down to normal super high comp of 700k a year after their starter grants run out, and yes many will quit but the people who stay aren't moving the needle on their finances that much.


You're doing well, but what of you want to buy a house?


He left Character.ai for money.


Dario's brain child


I think some of the commenters are naive to think government intervention is silly and TACO.

No, Dario said himself AI is national state weapon, then the government will not cease control.

What would happen is that we will have a more lobotomized and even more neurotic safeguards put in place in order to comply, and your data will be boardly sharing with the government.

Moving forward, above certain parameter size of model, it will require your self-identification in order to be used.


Consider applying for YC's Fall 2026 batch! Applications are open till July 27.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: