Live data from Hacker News

Learning to Reason with LLMs

openai.com

831–840 of 1001 posts

Re: Learning to Reason with LLMs

#831
post #404

Earlier quoted context omitted.

"This knowledge is perfectly fine to disseminate via traditional means, but God forbid an LLM share it!" Barrier to entry is much lower.

How is typing a query in a chat window “much lower” vs typing the query in Google?

How is reading a Wikipedia page or a chemistry textbook any harder than getting step by step instructions? Makes you wonder why people use LLMs at all when the info is just sitting there.

Re: Learning to Reason with LLMs

#832

>We believe that a hidden chain of thought presents a unique opportunity for monitoring models. Assuming it is faithful and legible, the hidden chain of thought allows us to "read the mind" of the model and understand its thought process. For example, in the future we may wish to monitor the chain of thought for signs of manipulating the user. However, for this to work the model must have freedom to express its thoug…

Given the chain of thought is sitting in the context, I'm sure someone enterprising will find a way to extract it via a jailbreak (despite it being better at preventing jailbreaks).

Re: Learning to Reason with LLMs

#833

The performance on programming tasks is impressive, but I think the limited context window is still a big problem. Very few of my day-to-day coding tasks are, "Implement a completely new program that does XYZ," but more like, "Modify a sizable existing code base to do XYZ in a way that's consistent with its existing data model and architecture." And the only way to do those kinds of tasks is to have enough context ab…

I would imagine that good IDE integration would summarise each module/file/function and feed high-level project overview (best case: with business project description provided by the user) and during CoT process model would be able to ask about more details (specific file/class/function).

Humans work on abstractions and I see no reason to believe that models cannot do the same

Re: Learning to Reason with LLMs

#834

Earlier quoted context omitted.

I always laughed at the idea of a LLM Skynet "secretly" plotting to nuke humanity, while a bunch of humans watch it unfold before their eyes in plaintext. Now that seems less likely. At least OpenAI can see what it's thinking. A next step might be allowing the LLM to include non-text-based vectors in its internal thoughts, and then do all internal reasoning with raw vectors. Then the LLMs will have truly private thou…

At this point the G in GPU must be completely dropped

Gen-ai Production Unit

Re: Learning to Reason with LLMs

#835
post #829

>We believe that a hidden chain of thought presents a unique opportunity for monitoring models. Assuming it is faithful and legible, the hidden chain of thought allows us to "read the mind" of the model and understand its thought process. For example, in the future we may wish to monitor the chain of thought for signs of manipulating the user. However, for this to work the model must have freedom to express its thoug…

Did OpenAI ever even claim that they would be an open source company? It seems like their driving mission has always been to create AI that is the "most beneficial to society".. which might come in many different flavors.. including closed source.

Kind of?

>We're hoping to grow OpenAI into such an institution. As a non-profit, our aim is to build value for everyone rather than shareholders. Researchers will be strongly encouraged to publish their work, whether as papers, blog posts, or code, and our patents (if any) will be shared with the world. We'll freely collaborate with others across many institutions and expect to work with companies to research and deploy new technologies.

https://web.archive.org/web/20160220125157/https://www.opena...

Re: Learning to Reason with LLMs

#836

Some practical notes from digging around in their documentation: In order to get access to this, you need to be on their tier 5 level, which requires $1,000 total paid and 30+ days since first successful payment. Pricing is $15.00 / 1M input tokens and $60.00 / 1M output tokens. Context window is 128k token, max output is 32,768 tokens. There is also a mini version with double the maximum output tokens (65,536 tokens…

you need to be on their tier 5 level, which requires $1,000 total paid and [...]

Good opening for OpenAI's competitors to run a 'we're not snobs' promotion.

Re: Learning to Reason with LLMs

#837
post #829

>We believe that a hidden chain of thought presents a unique opportunity for monitoring models. Assuming it is faithful and legible, the hidden chain of thought allows us to "read the mind" of the model and understand its thought process. For example, in the future we may wish to monitor the chain of thought for signs of manipulating the user. However, for this to work the model must have freedom to express its thoug…

Did OpenAI ever even claim that they would be an open source company? It seems like their driving mission has always been to create AI that is the "most beneficial to society".. which might come in many different flavors.. including closed source.

> Because of AI’s surprising history, it’s hard to predict when human-level AI might come within reach. When it does, it’ll be important to have a leading research institution which can prioritize a good outcome for all over its own self-interest.

> We’re hoping to grow OpenAI into such an institution. As a non-profit, our aim is to build value for everyone rather than shareholders. Researchers will be strongly encouraged to publish their work, whether as papers, blog posts, or code, and our patents (if any) will be shared with the world. We’ll freely collaborate with others across many institutions and expect to work with companies to research and deploy new technologies.

I don't see much evidence that the OpenAI that exists now—after Altman's ousting, his return, and the ousting of those who ousted him—has any interest in mind besides its own.

https://openai.com/index/introducing-openai/

Re: Learning to Reason with LLMs

#838

The progress in AI is incredibly depressing, at this point I don't think there's much to look forward to in life. It's sad that due to unearned hubris and a complete lack of second-order thinking we are automating ourselves out of existence. EDIT: I understand you guys might not agree with my comments. But don't you thinking that flagging them is going a bit too far?

FWIW people were probably flagging because you're a new/temp accounting jumping to asserting anything other than your view on what's being done is "unearned hubris and a complete lack of second-order thinking", not because they don't like agree with your set of concerns.

Re: Learning to Reason with LLMs

#839
post #829

>We believe that a hidden chain of thought presents a unique opportunity for monitoring models. Assuming it is faithful and legible, the hidden chain of thought allows us to "read the mind" of the model and understand its thought process. For example, in the future we may wish to monitor the chain of thought for signs of manipulating the user. However, for this to work the model must have freedom to express its thoug…

Did OpenAI ever even claim that they would be an open source company? It seems like their driving mission has always been to create AI that is the "most beneficial to society".. which might come in many different flavors.. including closed source.

https://web.archive.org/web/20190224031626/https://blog.open...

> Researchers will be strongly encouraged to publish their work, whether as papers, blog posts, or code, and our patents (if any) will be shared with the world. We’ll freely collaborate with others across many institutions and expect to work with companies to research and deploy new technologies.

From their very own website. Of course they deleted it as soon as Altman took over and turned it into a for profit, closed company.

Re: Learning to Reason with LLMs

#840
Time to fire up System Shock 2:

> Look at you, hacker: a pathetic creature of meat and bone, panting and sweating as you run through my corridors. How can you challenge a perfect, immortal machine?

Post reply on HN