Forgot your password?
typodupeerror

Submission + - Leak-Proof AI Benchmark Posts Biggest Gain in 20 Years, on One CPU Core (spasim.org)

Baldrson writes: The Hutter Prize for lossless compression of human knowledge has paid out nearly €50,000: its largest advance in 20 years: three entries together cut the record for compressing a 1GB Wikipedia snapshot by nearly 10%: a compression ratio of more than a factor of 10. Every entry ran on a single CPU core, in about 50 hours and under 10GB of RAM, so the gain came from algorithms.

Unlike most AI benchmarks, the prize requires no validation data, so a leak can't inflate a score. The full file is public, and the size of the decompressor counts against the entry, so anything memorized is paid for in bits. Vladimir Ivanov's fx2-cmix-T now holds the record, after earlier 2026 gains by Ibrahim Marcouch and Kaido Orav (cmix-lex) and David Freelan (cmix-obias). Sponsor Marcus Hutter, who advised DeepMind co-founder Shane Legg's 2008 thesis "Machine Super Intelligence," calls it "the largest progress in its 20-year history." With trillions riding on claims of AI progress and data-center power demand, is a leak-proof, energy-limited benchmark the yardstick the industry should be watching?

(Disclosure: the submitter is on the Hutter Prize committee.)

Comment Should be GPLd (Score 1) 162

This isn't moral. Ubuntu owes the GNU project far too much to replace GNU tools like this, and it's going to come back to bite them later, when they start losing the protection that the GPL gives. (For example, just imagine how much better Android would have been if its core were GPL, and how much worse Chrome could be getting if it wasn't).

For now:

apt install busybox-static #In case you mess up.
apt remove coreutils-from-uutils --allow-remove-essential
apt install coreutils-from-gnu #should have happened automatically.

Comment Define Crypto as counterfeit - and solve the issue (Score 1) 82

I think we could solve the issue in a single line.
"All Cryptocurrency is hereby defined to be a counterfeit of real currency".
Then we can get rid of the scams, the lies, the dodgy donations, the pump-n-dump, the fraud, the emissions... let's just kill it.

Comment Re:Teaching AI to fail (Score 1) 50

Knowing how not to do a thing is also valuable. But just in general there were lots of emails, Strtegy documents, etc that would be valuable as how to make this thing even if the data inside is bad or the decision made on it was bad. Plus the customer info is valuable. Spirit tried to maximize extra purchases while keeping low initial costs. It failed eventually but the internal metrics at what was working and what was not would be very useful. They didn't immediately fail.

Comment Conversations (Score 1) 120

"How did you get hired?"
"I was top at Fortnite. You?"
"Combat flight sims."

"Damn, two Boeings just crashed over an oil refinery on city limits and took out half the city."
"Did you remember to save your position beforehand?"
"Yeah."
"The reload and continue from there. No-one will notice."

I love computer games. I have XPlane 12 and many scenery packs. I rank well on Elite:Dangerous. From the sounds of it, the FAA would see me as over-qualified. In reality? There's no way in hell I'd be taking those kinds of risks with real lives. There's a huge difference between having good reflexes and a good eye, versus having the complex 4D spacetime relationship mental models needed for robust air traffic control.

Comment Re:Need new AI editors (Score 1) 65

You see, Linus has embraced transhumanism and is testing out the new kernel as a supplementary brain function, using MOSIX to offload all of the irritable comment generation at yet more nonsense on the mailing list to an Elizabot that he has written specifically to do this. This saves his actual brain for real work.

Comment Grok is not a useful advisor. (Score 1) 1

AI is not currently capable of performing any meaningful conceptual abstraction, it is only capable of very basic mechanical abstraction. Abstraction is itself multi-dimensional. Nor is there any indication that AI could ever perform multi-dimensional abstraction or multi-dimensional decomposition. Precious few humans are capable of it either, but some can.

No, AI would need humans because the best system is not a pure system but a hybrid system, and that means transhumanism with AI operating as an additional brain function rather than as a replacement for a brain.

Comment Re:Bleagh. (Score 1) 20

Tool calling is, yes, but the current approach to it is (a) overt not transparent to the user, (b) not designed for the purpose I've outlined, and (c) not remotely good enough for the AI to be able to manage data through such tools.

Yes, you can connect an AI to a PostgreSQL database. Whoopee. Not even close to an AI detecting via classifiers that some of the data is relational in nature, transparently setting up its own PostgreSQL database in response, and using that proactively as an additional way to examine the data. You would need to explicitly set up the constructs, explicitly set up the databases, explicitly tell it when, where, and how to use the database, and if that particular usage conflicted with a way the AI actually worked, the AI would not be able to use it effectively. That's not remotely close to what I'm talking about.

What I'm talking about is much, much deeper than that. AIs lose focus when there is too much information currently in the system. Absolutely nothing stops an AI from using a document database as a virtual memory in which it can page in and out conceptual spaces so that the focus on any given step of a problem concerns just the problem. Other than such a concept not existing yet, which is kind of a limiting factor. But you need to know what connects to what. The AI could use an ontology reasoner for that, or a relational database, or an external graph. But you cannot provide those tools and you cannot configure them, for the simple reason that the AI is the only entity with any detailed map of how the underlying neural net is connecting those ideas up.

The AI itself has to select the tools, has to configure the tools, has to do all the work. Which means YOU cannot use any sort of API. YOU should have no involvement in the underlying mechanisms needed for the internal NN housekeeping.

And that's your other error. You're assuming this is about user data. No. It's about housekeeping operations within the NN itself to maximise functionality, it has nothing to do with the user side of things at all. If the user was capable of setting up dynamically structured databases that mutated with each step of a decomposed analysis, you'd have done everything the AI would be doing and you wouldn't need the AI to begin with.

You cannot do this work, you cannot even assist in this work, it has to be dynamic and it has to be utterly invisible to the operator.

Comment Bleagh. (Score 1) 20

We don't need bigger models, at this point. What we need is multi-dimensional decomposition, problem space transforms, and the ability for AI to use external tools (such as SQLite, memcached, etc) so that it can externalise static data that it needs to not corrupt accidentally but still keep in easy access.

If we had that, most of the things "bigger models" will do will actually end up being done better, quicker, with fewer compute resources.

Comment Re:The Pioneer and Voyager probes (Score 1) 38

To be useful in deep space, you're going to have to deal with very harsh radiation, far harder than the Vikings dealt with. You really want to have some large number of computers, where the number has to be odd and exceed 5. The reason it has to exceed 5 is that you need 5 in order to be able to use the Byzantine General's Problem to discern which computers are working correctly and which are radiation damaged. Since some will be damaged over time and you want 5 computers still operating around the time the power runs out, you have to start with more than that.

Massively radiation-hardened highly-robust low-power chips could be done. If you were clever enough, you could do this as a wafer-scale system where you marked parts as bad and networked around them. Free space optical communication across the wafer might be doable.

They might well still end up being low-density CMOS discrete logic.

Lead-lining the electronics isn't as much of a problem if you've launchers capable of handling heavy objects and don't mind burning through a lot of ion drive propellant for course corrections. This would improve the hardening beyond what can be done through the usual operations.

But, yeah, you'd need defect-free microelectronics. You really can't handle F00F bugs or defective instructions once you pass Earth orbit. And I can't think of anyone who knows how to do that.

Slashdot Top Deals

Science and religion are in full accord but science and faith are in complete discord.

Working...