Forgot your password?
typodupeerror

Submission + - Leak-Proof AI Benchmark Posts Biggest Gain in 20 Years, on One CPU Core (spasim.org)

Baldrson writes: The Hutter Prize for lossless compression of human knowledge has paid out nearly €50,000: its largest advance in 20 years: three entries together cut the record for compressing a 1GB Wikipedia snapshot by nearly 10%: a compression ratio of more than a factor of 10. Every entry ran on a single CPU core, in about 50 hours and under 10GB of RAM, so the gain came from algorithms.

Unlike most AI benchmarks, the prize requires no validation data, so a leak can't inflate a score. The full file is public, and the size of the decompressor counts against the entry, so anything memorized is paid for in bits. Vladimir Ivanov's fx2-cmix-T now holds the record, after earlier 2026 gains by Ibrahim Marcouch and Kaido Orav (cmix-lex) and David Freelan (cmix-obias). Sponsor Marcus Hutter, who advised DeepMind co-founder Shane Legg's 2008 thesis "Machine Super Intelligence," calls it "the largest progress in its 20-year history." With trillions riding on claims of AI progress and data-center power demand, is a leak-proof, energy-limited benchmark the yardstick the industry should be watching?

(Disclosure: the submitter is on the Hutter Prize committee.)

Comment Alas, UEFI and friend are really quite horrid (Score 1) 63

As a lawyer would say "not suitable for the purpose sold".
For a story about how Oxide avoided them, see "Holistic boot", at https://ancillary-proxy.atarimworker.io?url=https%3A%2F%2Frfd.shared.oxide.compu...

["Really quite horrid" is British for "<expletive deleted/> piece of <expletive deleted/> junk"]

Comment Re:Positive feedback loops are bad, m'kay? (Score 2) 208

Yup, same as the feedback loops in "cold readings"

Charlie Stross(@cstross@wandering.shop) wrote, in Mastadon:
The LLMentalist effect: Large Language Models replicate the mechanisms used by (fake) psychics to gull their victims: https://ancillary-proxy.atarimworker.io?url=https%3A%2F%2Fsoftwarecrisis.dev%2Flet...

The title of the paper is "The LLMentalist Effect: how chat-based Large Language Models replicate the mechanisms of a psychic’s con"

Comment Google is very successful, because... (Score 1) 47

  • - they own the agent for the advertiser,
  • - they own the agent for the publisher,
  • - they own the auction house, and
  • - they don't provide an audit trail.

I used to work in advertising, and I saw Google as the personification of "moral hazard" (which see). Other things? Way nicer.

Comment Alas, the "birthday paradox" will misidentify you (Score 2) 55

If you scan a thousand British faces and compare them to a thousand criminals, you will do 1,000,000 comparisons. (that's the birthday paradox part).
If your error rate is 0.8%, you'll get roughly 8,000 false positives and negatives.
That's bad enough if they are all false positives: people get arrested, then released.
It's way worse if they are all false negatives: 8,000 criminals get ignored by the police dragnet.

That was Britain: false positives are life-threatening in countries where the police carry guns.
0.8% is a good error rate. 34% wrong is typical in matching black women. See
https://ancillary-proxy.atarimworker.io?url=https%3A%2F%2Fwww.aclu-mn.org%2Fen%2Fnews%2Fbiased-technology-automated-discrimination-facial-recognition%23%3A~%3Atext%3DStudies%2520show%2520that%2520facial%2520recognition%2520technology%2520is%2520biased.%2Cpublished%2520by%2520MIT%2520Media%2520Lab.

Slashdot Top Deals

The only way to learn a new programming language is by writing programs in it. - Brian Kernighan

Working...