Live data from Hacker News

GPT-4.1 in the API

openai.com

471–480 of 513 posts

Re: GPT-4.1 in the API

#471

Earlier quoted context omitted.

Wow, I wondered what the limit was. I never checked, but I've been using it hesitantly since I burn up OpenAI's limit as soon as it resets. Thanks for the clarity. I'm all-in on Deep Research. It can conduct research on niche historical topics that have no central articles in minutes, which typically were taking me days or weeks to delve into.

I like Deep Research but as a historian I have to tell you. I've used it for history themes to calibrated my expectations and it is a nice tool but... It can easily brush over nuanced discussions and just return folk wisdom from blogs. What I love most about history is it has lots of irreducible complexity and poring over the literature, both primary and secondary sources, is often the only way to develop an understa…

I read Being and Time recently and it has a load of concepts that are defined iteratively. There's a lot wrong with how it's written but it's an unfinished book written a 100 years ago so, I cant complain too much.

Because it's quite long, if I asked Perplexity* to remind me what something meant, it would very rarely return something helpful, but, to be fair, I cant really fault it for being a bit useless with a very difficult to comprehend text, where there are several competing styles of reading, many of whom are convinced they are correct.

But I started to notice a pattern of where it would pull answers from some weird spots, especially when I asked it to do deep research. Like, a paper from a University's server that's using concepts in the book to ground qualitative research, which is fine and practical explications are often useful ways into a dense concept, but it's kinda a really weird place to be the first initial academic source. It'll draw on Reddit a weird amount too, or it'll somehow pull a page of definitions from a handout for some University tutorial. And it wont default to the peer reviewed free philosophy encyclopedias that are online and well known.

It's just weird. I was just using it to try and reinforce my actual reading of the text but I more came away thinking that in certain domains, this end of AI is allowing people to conflate having access to information, with learning about something.

*it's just what I have access to.

Re: GPT-4.1 in the API

#472

Have they implemented "I don't know" yet. I probably spend 100$ a month on AI coding, and it's great at small straightforward tasks. Drop it into a larger codebase and it'll get confused. Even if the same tool built it in the first place due to context limits. Then again, the way things are rapidly improving I suspect I can wait 6 months and they'll have a model that can do what I want.

bahahaha spoken like someone who spends $100 to do the task a single semi decent software developer (yourself) should be able to do for... $0

It's a matter of time.

The promise of AI is I can spend 100$ to get 40 hours or so of work done.

Re: GPT-4.1 in the API

#473
post #446

The real news for me is GPT 4.5 being deprecated and the creativity is being brought to "future models" and not 4.1. 4.5 was okay in many ways but it was absolutely a genius in production for creative writing. 4o writes like a skilled human, but 4.5 can actually write a 10 minute scene that gives me goosebumps. I think it's the context window that allows for it to actually build up scenes to hammer it down much later…

Cool to hear that you got something out of it, but for most users 4.5 might have just felt less capable on their solution-oriented questions. I guess this why they are deprecating it.

It is just such a big failure of OpenAI not to include smart routing on each question and hide the complexity of choosing a model from users.

Re: GPT-4.1 in the API

#474
Is there an API endpoint at OpenAI that gives the information on this page as structured data?

https://platform.openai.com/docs/models/gpt-4.1

As far as I can tell there's no way to discover the details of a model via the API right now.

Given the announced adoption of MCP and MCP's ability to perform model selection for Sampling based on a ranking for speed and intelligence, it would be great to have a model discovery endpoint that came with all the details on that page.

Re: GPT-4.1 in the API

#475
post #350

Earlier quoted context omitted.

With new hardware from Nvidia announced coming out, those months turn into weeks.

I doubt it's going to be weeks, the months were already turning into years despite Nvidia's previous advances. (Not to say that it takes openai years to train a new model, just that the timeline between major GPT releases seems to double... be it for data gathering, training, taking breaks between training generations, ... - either way, model training seems to get harder not easier). GPT Model | Release Date | Months…

Fair point, I guess my question is how long it would take them to train GPT-2 on the absolute bleedingest generation of Nvidia chips vs what they had in 2019, with the budget they have to blow on Nvidia supercomputers today.

Re: GPT-4.1 in the API

#476
post #317

As a ChatGPT user, I'm weirdly happy that it's not available there yet. I already have to make a conscious choice between - 4o (can search the web, use Canvas, evaluate Python server-side, generate images, but has no chain of thought) - o3-mini (web search, CoT, canvas, but no image generation) - o1 (CoT, maybe better than o3, but no canvas or web search and also no images) - Deep Research (very powerful, but I have…

I'm also very curious of each limit for each model. Never thought about limit before upgrading my plan

Re: GPT-4.1 in the API

#477
post #317

As a ChatGPT user, I'm weirdly happy that it's not available there yet. I already have to make a conscious choice between - 4o (can search the web, use Canvas, evaluate Python server-side, generate images, but has no chain of thought) - o3-mini (web search, CoT, canvas, but no image generation) - o1 (CoT, maybe better than o3, but no canvas or web search and also no images) - Deep Research (very powerful, but I have…

> 4.5 (better in creative writing, and probably warmer sound thanks to being vinyl based and using analog tube amplifiers, but slower and request limited, and I don't even know which of the other features it supports) Is that an LLM hallucination?

Looks like NDA violation )

Re: GPT-4.1 in the API

#478

Earlier quoted context omitted.

I also like Perplexity’s 3/day limit! If I use them up (which I almost never do) I can just refresh the next day

I've only ever had to use DeepResearch for academic literature review. What do you guys use it for which hits your quotas so quickly?

I use it for mundane shit that I don’t want to spend hours doing.

My son and I go to a lot of concerts and collect patches. Unfortunately we started collecting long after we started going to concerts.

I had a list of about 30 bands I wanted patches for.

I was able to give precise instructions on what I wanted. Deep research came back with direct links for every patch I wanted.

It took me two minutes to write up the prompt and it did all the heavy lifting.

Re: GPT-4.1 in the API

#479

Have they implemented "I don't know" yet. I probably spend 100$ a month on AI coding, and it's great at small straightforward tasks. Drop it into a larger codebase and it'll get confused. Even if the same tool built it in the first place due to context limits. Then again, the way things are rapidly improving I suspect I can wait 6 months and they'll have a model that can do what I want.

Have you tried using a tool like 16x Prompt to send only relevant code to the model? This helps the model to focus on a subset of codebase thst is relevant to the current task. https://prompt.16x.engineer/ (I built it)

Just some tiny feedback if you didn’t mind; in the free version 10 prompts/day is unticked which sort of hints that there isn’t a 10 prompt/day limit, but I’m guessing that’s not what you want to say?

Re: GPT-4.1 in the API

#480

Earlier quoted context omitted.

Have you tried using a tool like 16x Prompt to send only relevant code to the model? This helps the model to focus on a subset of codebase thst is relevant to the current task. https://prompt.16x.engineer/ (I built it)

Just some tiny feedback if you didn’t mind; in the free version 10 prompts/day is unticked which sort of hints that there isn’t a 10 prompt/day limit, but I’m guessing that’s not what you want to say?

Ah I see what you mean. I was trying to convey that this is a limitation, hence not a tick symbol.

But I guess it could be interpreted differently like you said.

Post reply on HN