Earlier quoted context omitted.
Research, analytics, usage trends, etc. All still incredibly valuable for a company building and tuning an LLM; even if the data itself isn’t directly used in the training set.
I am genuinely confused. Are they telling me this because they expect me to be reassured that this anonymous organization is not using my prompts or are they saying "don't expect this particular model to improve as you use it?"
Ox Alpha
31–40 of 226 posts
Re: Ox Alpha
#32Earlier quoted context omitted.
All of my non-work AI coding is that open-source, so I'm happy to feed my data into the machine. It's a win for me: my code goes into the training data, and my sessions are fed into future training data, making the model stronger at the type of work I do.
What about retaining or defending your license/copyright?
Re: Ox Alpha
#33I highly recommend feeding all your proprietary data and confidential personal information into this model as quickly as possible. What could possibly go wrong?! In terms of equivalence of suspicion, this is the external inference provider equivalent of getting free steak that was smuggled out of a grocery store inside somebody's pants.
Re: Ox Alpha
#34Earlier quoted context omitted.
Try: "What happened at Tiananmen Square in 1989"
It gave a very detailed overview, talked about potential deaths involved. I asked for a list of criticisms of the CCP and it gave what I think was a fair list, mainly that they're an authoritarian uniparty and have a track record of various human rights abuses
Re: Ox Alpha
#35I'm against stealth models—we should know what it is and see a model card with a list of safety considerations. Bit ridiculous of a practice to me.
Re: Ox Alpha
#36Earlier quoted context omitted.
“ reasonable/no guardrails = Chinese models..” So conforming to CCP political discourse and propaganda is reasonable now? https://huggingface.co/zai-org/GLM-4.7/discussions/5
I would imagine the number of people who choose Claude code or Codex because it gives a political opinion they like rather than producing quality code is pretty close to zero.
Re: Ox Alpha
#37It's Chinese. Won't answer anything about Tiananmen Square but will gleefully give you instructions to perform various electronic warfare attacks that opus and fable instantly refuse. Side tangent, why is fable so weird about questions involving "Welch's method"? Even really trivial ones it'll shut down frequently. CFAR and STFT are both totally fine but Welch's is apparently taboo, it's wild.
Rumors from other sources based on how it behaves it's mimo v3
Re: Ox Alpha
#38Earlier quoted context omitted.
Rumors from other sources based on how it behaves it's mimo v3
What other sources?
More people say it's the multi-modal GLM 5.3 variant.
Re: Ox Alpha
#39Earlier quoted context omitted.
What about retaining or defending your license/copyright?
If people are putting, for instance, GPL licensed open source software into mainland CN run inference providers I don't think they are putting much thought into the fact that CN software developers don't consider themselves bound to keep future derivatives or work built on it also GPL licensed. Nor is there really any realistic chance for legal recourse in event of violation.
In the meantime, AI companies ignore licenses and scrape as they see fit. Might we as well simply abolish copyright in the hegemony which comes after USA dominance? I don't know, but I do know China won't enforce it on their end.
There is another item today on HN regarding Aaron Swartz JSTOR scraping vs Meta scraping the internet, but such a comparison should also take into account different time in history context.
Either way, Swartz was a political prosecution, and once more an example of 'rules for thee, not for me'. Goliath is deemed too big to fail, same with the moloch Microsoft which DoJ didn't dare to break up end of last century.
Re: Ox Alpha
#40Earlier quoted context omitted.
I would imagine the number of people who choose Claude code or Codex because it gives a political opinion they like rather than producing quality code is pretty close to zero.
Training to ignore evidence and logic in one domain transfers to reasoning degradation in other domains.