Hacker Newsnew | past | comments | ask | show | jobs | submit | dust42's commentslogin

The CPU is for mobile phones. An Nvidia H200 has 4.8TB/s. An RTX 5090 has 1.8TB/s. Both use ~700W - not comparable with a phone.


SEO was yesterday, now it is AIO - you will have to wait longer for the results and likely you will need a lot more cash: pay TIME for the ads, wait until next model release and see what sticks, then rinse and repeat.


I suggest the term Language Model Agent Optimization...


lmao


We'll remember fondly the early 2020s as the time when product suggestions from LLMs were relatively naive.


The term the industry actually uses is GEO - Generative Engine Optimization.


As a user I still prefer .pi right there in my home directory.


In 2 years from now we will be at 400%. https://xkcd.com/605/ Also, it is called kernels (you have nitpick in your username)


> As with many conflicts this one goes back a long way. Ukraine gave up its nukes to Russia in exchange for a non aggression pact that the invasion broke.

From the Guardian (12 days ago): "Kyiv’s decision to honour second world war fighters who killed about 100,000 Poles has revived simmering tensions". Bandera was a Nazi collaborator and today's Ukraine politics worships Bandera. Lots of streets are renamed to honor him. Eastern Ukrainians are considered second class - or as some politicians said: "There are 2 million people too many".

The US (look up Nuland and her quote about 'fuck the EU') fueled with $5B the nationalism in Ukraine. Now imagine, the US would fuel with $50B the nationalism in Belgium and the the flemish would attack wallonie. You know what would happen? France would support the wallonie. Nothing to see here.


With a M5 16c 48GB and Qwen 3.6 35B Q4 I get up to 1900 PP/s and 80 TG/s. With an Nvidia 5090 I get 7800 PP/s and 280 TG/s.

Together with pi mono I wouldn't want to go back to Claude & Co. Speed, quality of the answers, short answer times at any time of day - once you have eaten from the fruit your definition of SOTA will change...

For reference, I do software development since 30 years, I am not vibe coding the umpteenth todo list.


This and we don't know yet what happened. It could have structurally collapsed - very unlikely, it could have uncommanded retracted, or maintenance has overridden the protections. I'd place my bets on #3, handling error in maintenance mode.


From the picture and the text this aircraft was parked at the gate. During a hard landing the nose gear may collapse but not while being parked. And while parked there are protections to prevent retraction. However, these can be overridden by maintenance.


I just spent the last two weeks digging into workflow state engines and temporal was one of the candidates. It is a VC backed fork of Cadence. The got 0.3B funding and whatever positive I read about them on the net I take with a big spoon of salt. Just my 2 cents.


I would love to read your analysis somewhere. It's a current interest of mine, both hobby and professional -- and limited to Elixir, Golang and Rust -- and I'm still slowly scoping the landscape.

Recommendations for good engines? Or just thoughts and analyses?


For many models the performance of llama.cpp on Mac is 20-40% lower than MLX. Did you try MLX? At least on HF there are MLX 2-bit quants. Unfortunately I have only 64GB, so I can't test it.


I'm not using llama.cpp there, it's my inference engine that is DeepSeek v4 specific. The goal is to optimize it as much as possible.


That's cool!

I knew the name sounded familiar, thank you for SDS!


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: