Forgot your password?
typodupeerror

Comment Ultimately is this not all to the good ? (Score 3, Insightful) 85

Bugs (in new code) are hopefully being added more slowly than bugs found and squashed in existing code - so surely the number of remaining bugs will drop. Hopefully those running bug finding AIs are not keeping some remote exploitable bugs to themselves; I would not be surprised if government agencies were doing this.

Comment He is allowed to slow down (Score 1) 118

No one is forcing him to spend money he doesn't have to produce a product no one will pay for.

He has irresponsibly played a game and even helped make the rules. Now he realizes other people are better at it than he is, but he keeps spending borrowed money for 'pay to play' and all the other players except openai (his twin) are figuring out how to play for much less.

His biggest competition are the bastards at companies like Alibaba, Microsoft, and Google who actually have products, services, and paying customers who will stick with them even if their AI is a generation or three behind.

See, it's perfectly ok to slow down. The others have done it. But they cheated... They actually had business plans.

Comment Re:Yes but is it genocide? (Score 3, Informative) 61

People who have a much greater standing than me have said that Israel is carrying out genocide, eg: The International Association of Genocide Scholars, Holocaust scholar Amos Goldberg, Médecins Sans Frontières, Raz Segal "A Textbook Case of Genocide" to list but a few. The evidence is there for all to see but our political leaders in the West ignore it.

Then there is Jerusalem, the West Bank, South Lebanon - all with the aim of Greater Israel.

Comment Re: M5 Max MacBook Pro with 128GB (Score 2) 17

Qwen 3.8 27b runs with mtp and a 72k context window on two RTX 5060 Ti 16GB at about 35tps (in MTP, that compares to 75tps non mtp). And, yesterday, it nearly matched GLM 5.2 running on 8xH100 in every task I tossed it. (My glm rig is rate limited to 30rps). My comparison is "ability to handle long complex tasks".

So, the trick is to use memory, search, fetch, vector database.

The purpose of a model is to reason. It needs enough training to perform further research. Context is very-short term memory, vector databases are their long term memory, rag is the books they read, and search and fetch are their libraries.

I would LOVE to switch back to two 3090 cards for 48GB (they do image and video now), but when I run qwen 3.8 27b fp8 with 256k context, I find the quality is higher in some cases, but drops because the model then favors short term memory iver research.

Memory bandwidth is much more exciting. 500tps when running on an HPC is very very nice.

So, if it were me measuring, a single 64GB HBM3e GPU is really where we should aim as this should be consumer cost friendly by 2030. (now I have to try on an A100 later today)

Slashdot Top Deals

Any given program, when running, is obsolete.

Working...