Earlier quoted context omitted.
To be honest: I don't know for certain, but I'd assume that the people who pay the bills get the strongest alignment. They may not be tech people, but I (so far) haven't got a reason to think that the AI engineers are going behind the backs of their corporate leadership and subverting what they're being asked to do; do you? (I think it would be a good thing for humanity if they did)
I think right now both the engineers developing AI and the share holders are more focused on beating coding benchmarks and gaining revenue than anything to do with alignment.
Fable and the end of the free lunch
231–240 of 268 posts
Re: Fable and the end of the free lunch
#232Earlier quoted context omitted.
I don't understand what this means. I use LLMs daily for my work in programming things, and they regularly will assert things that are not accurate.
When you let the the agent do a test, or tell it to read that doc first, you will ground it in reality. Doesn't mean they are 100% reliable. But without and on their own without access to grounding information, they halluzinate wildly.
Re: Fable and the end of the free lunch
#233Earlier quoted context omitted.
I don't understand what this means. I use LLMs daily for my work in programming things, and they regularly will assert things that are not accurate.
A lot like humans, really. People regularly cite something they read, or quote a stat that turns out to be just completely inaccurate. But if you look up the thing, then you have facts again.
Re: Fable and the end of the free lunch
#234Re: Fable and the end of the free lunch
#235The real revolution is Deepseek v4 flash and similar models (GPT 5.6 Luna, muse spark 1.2, mimo, etc...) - Genuinely good performance for a tiny fraction of the cost of Fable and even GLM etc... I think a lot of people would be very content if they never got smarter, and just kept getting even cheaper/faster. Of course, both things continue to happen on a seemingly monthly basis
I was using ChatGPT voice during cooking to reflect on variations of a dishes i was preparing for years. It was so amazing to get advices and reflect that it struck me : I could use this model forever - it’s clever enough to help me tons and do lot of work for me - even if ai would stop evolving I would love it
Re: Fable and the end of the free lunch
#236What are all these rote coding tasks people do that they can farm it out to lesser models?
Re: Fable and the end of the free lunch
#237Earlier quoted context omitted.
It's terrible because it's Llama 3.1 8B. It's such a crappy model because HC1 was a relatively low budget proof of concept. The team that built is working on a better implementation.
Not sure its that to be honest. It seems like maybe its not installed correctly or is like GPT-1/GPT-2 quality? I asked it who is [famous actress] and it started talking about some random person from Mexico with a completely different name. The speed is intoxicating but i'd like for it to actually answer based on what I asked. Thats why I think something might be wrong in implementation on this site. Edit: I went bac…
Re: Fable and the end of the free lunch
#238Re: Fable and the end of the free lunch
#239Earlier quoted context omitted.
What about censorship? > I will be able to use them forever Where will you run them when powerful enough GPU and RAM are only sold to hyperscalers?
I dislike all censorship, but US models are much more censored, I often find myself using Chinese models to get answers I want. Now, of course I’d prefer no censoring, but I live in the world we live in. I’m working in the assumption that (like today) there will always be somehow on openrouter, or similar, who will host a model I want to run.
Re: Fable and the end of the free lunch
#240Earlier quoted context omitted.
I dislike all censorship, but US models are much more censored, I often find myself using Chinese models to get answers I want. Now, of course I’d prefer no censoring, but I live in the world we live in. I’m working in the assumption that (like today) there will always be somehow on openrouter, or similar, who will host a model I want to run.
How do Western models censor?