Comment Re:Meanwhile, back at the benchmarks (Score 1) 21
("It actually makes RAM needs worse" -> "MoEs actually make RAM needs worse")
("It actually makes RAM needs worse" -> "MoEs actually make RAM needs worse")
It actually makes RAM needs worse than a dense model, for a given quality.
That said, inference techniques like FreeToken are improving MoE swapping performance, so it's not as bad as it once was. Honestly, I would not be at all surprised if we start doing training in a swap-aware manner, where expert cache misses count as loss and models can queue experts to start loading before they're needed. Could even train in two tiers of loading - RAM and disk. Wouldn't that be great? Could run multi-terabyte models on consumer hardware. Training would tend to tend to concentrate key logic and reasoning on small number of very active experts, mixed reasoning/knowledge on less common experts that are usually kept in RAM, and rarer knowledge on experts that usually remain on-disk until needed.
There are well known, reputable abliterators on HuggingFace.
Also, censorship usually (not always) is something you care about for chats, not agentic work.
A
It's big, but not good for its size. They're using tricks to pretend that they're better than they are. For example, compare the numbers that they list for the competion on DeepSWE up against the actual DeepSWE scores.
Not cool, Mistral.
Artificial Analysis is overrated, but yes, Mistral is playing fast and loose with their claims.
It takes one hell of a toll to try to do verbal exams 1:1 on students to evaluate them.
Nobody said "verbal exams 1:1".
The only thing that is required is that they be separated from their computers and their phones.
@echo off
if "%~1"=="" (
echo Usage: %~nx0 filename
exit/b 1
)
type "%~1"
The difference is that if you pay people to monkey around with things, you get a non-revocable perpetual license to the result. With license rental, you get a perpetual bill. It's the same rent trap as individuals get stuck in but with less excuses.
For some customers, they looked at the resources needed to migrate, and decided it was cheaper to stay put and/or invest the resources that would be needed elsewhere in their business.
Yes, and the reason is that they stupidly painted themselves into a corner and never thought about an exit strategy. Had they been thinking about the possibility that they might one day need to migrate from the beginning, it would be much cheaper to do so now.
That's the thing about painting yourself into a corner, a little better planning at the start can always avoid it with little to no additional effort.
Well, initial stupidity can come at a very high price.
Yes. Funny thing: There is a regulatory requirement here for any financial service provider here to have a cloud replacement strategy. That includes cloud providers. Would be a good thing for everybody to have that, but people are stupid, full of themselves and greedy.
I would have though that by now they are all gone. Probably those left are the ones that stupidly painted themselves into a corner and never thought about exist strategies.
Nobody, I repeat, nobody who uses AI is going to see "OpenAI adding watermarks" and think, "Oh, I better use AI more, especially OpenAI".
And nobody, I repeat, nobody who doesn't use AI is going to see "OpenAI adding watermarks" and think, "Oh, I better use AI more, now that it'll be easier to see that what I did is made by AI."
It is not "an advertisement of the power of the technology". It's a regulation forced on them by the EU's AI Act, which neither companies nor users want.
Some people have a great ambition: to build something that will last, at least until they've finished building it.