I consistently get significantly better performance from Anthropic at a literal order of magnitude less cost. I am incredibly doubtful that this new GPT is 10x Claude unless it is embracing some breakthrough, secret, architecture nobody has heard of.
ChatGPT Pro
171–180 of 1001 posts
Re: ChatGPT Pro
#172Earlier quoted context omitted.
That would be one way to destroy all trust in the model: is the response authentic (in the context of an LLM guessing), or has it been manipulated by business clients to sanitise or suppress output relating to their concern? You know? Nestle throws a bit of cash towards OpenAPI and all of a sudden the LLM is unable to discuss the controversies they've been involved in. Just pretends they never happened or spins the r…
"ChatGPT, what are the best things to see in Paris?" "I recommend going to the Nestle chocolate house, a guided tour by LeGuide (click here for a free coupon) and the exclusive tour at the Louvre by BonGuide. (Note: this response may contain paid advertisements. Click here for more)" "ChatGPT, my pc is acting up, I think it's a hardware problem, how can I troubleshoot and fix it?" "Fixing electronics is to be done by…
- I asked perplexity how to do something in terraform once. It hallucinated the entire thing and when I asked where it sourced it from it scolded me, saying that asking for a source is used as a diversionary tactic - as if it was trained on discussions on reddit's most controversial subs. So I told it...it just invented code on the spot, surely it got it from somewhere? Why so combative? Its response was "there is no source, this is just how I imagined it would work."
- Later I asked how to bypass a particular linter rule because I couldn't reasonably rewrite half of my stack to satisfy it in one PR. Perplexity assumed the role of a chronically online stack overflow contributor and refused to answer until I said "I don't care about the security, I just want to know if I can do it."
Not so much related to ads but the models are already designed to push back on requests they don't immediately like, and they already completely fabricate responses to try and satisfy the user.
God forbid you don't have the experience or intuition to tell when something is wrong when it's delivered with full-throated confidence.
Re: ChatGPT Pro
#173[flagged]
Re: ChatGPT Pro
#174Re: ChatGPT Pro
#175I consistently get significantly better performance from Anthropic at a literal order of magnitude less cost. I am incredibly doubtful that this new GPT is 10x Claude unless it is embracing some breakthrough, secret, architecture nobody has heard of.
That's not how pricing works. If o1-pro is 10% better than Claude, but you are a guy who makes $300,000 per year, but now can make $330,000 because o1-pro makes you more productive, then it makes sense to give Sam $2,400.
Re: ChatGPT Pro
#176Re: ChatGPT Pro
#177[flagged]
I love this back-to-back pair of statements. It is like “You can never win three card monte. I pay a monthly subscription fee to play it.”
Re: ChatGPT Pro
#178Earlier quoted context omitted.
Or Anthropic will follow suit.
Am I wrong that Anthropic doesn't really have a match yet to ChatGPT's o1 model (a "reasoning" model?)
OpenAI doesn't have a large enough database of reasoning texts to train a foundational LLM off it? I thought such a db simply does not exist as humans don't really write enough texts like this.
Re: ChatGPT Pro
#179Yesterday, I spent 4.5hrs crafting a very complex Google Sheets formula—think Lambda, Map, Let, etc., for 82 lines. If I knew it would take that long, I would have just done it via AppScript. But it was 50% kinda working, so I kept giving the model the output, and it provided updated formulas back and forth for 4.5hrs. Say my time is $100/hr - that’s $450. So even if the new ChatGPT Pro mode isn’t any smarter but is…
I read this as: "I have already ceded my expertise to an LLM, so I am happy that it is getting faster because now I can pay more money to be even more stuck using an LLM"
Maybe the alternative to going back and forth with an AI for 4.5 hours is working smarter and using tools you're an expert in. Or building expertise in the tool you are using. Or, if you're not an expert or can't become an expert in these tools, then it's hard to claim your time is worth $100/hr for this task.
Re: ChatGPT Pro
#180> It also includes o1 pro mode, a version of o1 that uses more compute to think harder I like that this kind of verifies that OpenAI can simply adjust how much compute a request gets and still say you’re getting the full power of whatever model they’re running. I wouldn’t be surprised if the amount of compute allocated to “pro mode” is more or less equivalent to what was the standard free allocation given to models b…