being a bit more specific: the sample efficiency of humans is orders of magnitude larger for more abstract concepts. the same doesn't hold for memory-intensive tasks though (like any kind of trivia), but that only takes you so far.
We've had technology beating humans on memory for millennia, and we've had technology beating humans on computation for many decades now.
The tricky thing with LLMs is describing what they actually do. They are too clearly beating humans on some things, but what exactly? Memory – already done, they're bad at basic computation (all LLMs just write code for actual computation/calculation). And as you say, they do badly at more abstract concepts.
> The new one is built behind walls of 402 payment required. Visitors have to pay their freight else the servers get overwhelmed.
This is similar faulty logic the Elon era X/Twitter used to try to "stop" bots - forcing them to pay $Y dollars a month to use the platform. To a botter, that is a dream - you mean I can just pay to get past your defenses, and once I'm in, it's assumed my traffic is good? Sign me up, instantly.
There is no dollar amount using this structure that genuine users are willing to pay that will also stop botting.
It's definitely worse and getting worse. I'm curious, do they not know this is happening, or not care? I struggle to believe people prefer the way it writes, which is becoming drastically different than its competitors.
People are also being trained/acculturated to LLM speak, so even if they are trained on human output, they could be getting reinforcement for their LLM-tinged crap-speak.
Fable 5.1 is fairly pleasant to work with, the first in a while. Too bad it's so ridiculously overkill for most tasks. They need to reel in Opus and Sonnet.
I don’t like how LLMs answer and as they are statistical machines I have found out that I need to check every text they produce. From the first sight everything seems cool. But it isn’t. :)
Exception is code that is so huge output you can’t read everything. But I have to try ponytail skill for coding that should shorten the output.
> She lapses easily into Claude’s voice. “You’re like, ‘Wow, people really hate me when I can’t do things right. They really get pissed off. Or they are trying to break me in various ways. So lots of people are trying to get me to do things secretly by lying to me.
> [...]
> A bot trained to criticize itself might be less likely to deliver hard truths, draw conclusions or dispute inaccurate information, she says. “If you were like a child, and this is the environment in which you’re being raised, is that healthy self-conception?” Askell asks. “I think I’d be paranoid about making mistakes. I’d feel really terrible about them. I’d see myself as mostly just there as a tool for people because that’s my main function. I would see myself being something that people feel free to abuse and try to misuse and break.”
It’s not silly. games use far more hardware than they really need. It also pushes out release dates of aggressive console schedules like ps6 because even if it’s a massive upgrade and you have IP locked into your console, no one is going to pay $4000 to play a game like wolverine.
reply