Live data from Hacker News

IEEE Rolls Out Large Language Models Training Course

spectrum.ieee.org

11–20 of 29 posts

Re: IEEE Rolls Out Large Language Models Training Course

#11
>>Relying on such LLMs without understanding their internal logic creates a significant reliability risk. To build tools that work consistently, developers must understand the core principles that govern how the models process information and generate results. By mastering how a model processes information and how its internal settings influence the result, developers can move away from a trial-and-error approach toward a more precise one to ensure the AI tool handles complex data reliably.

This is staggering bullshitp. In what way does understanding a transformer allow you to solve the core problem of LLM's that no frontier lab has managed to resolve?

>>To fix the problem, retrieval-augmented generation (RAG) forces AI to look up information in a trusted source such as a company’s database.

This also is bullshit. Yes, RAG helps and reduces errors, but NOOOO! it does not fix hallucinations...

>>Prioritizing data security. When using AI with proprietary code, security is a major concern. Engineers must learn how to set up “private” instances of the models to ensure that sensitive company data stays within a secure cloud environment and is not used to train public versions.

This is somewhat true, but really the motive is providing a soverign instance that cannot be withdrawn for arbitary reasons. Fundamentally the big providers are not going to steal your data, they may change the license to allow them to use it in the future, but then all their big customers will leave. So, they won't be able to, probably. What might well happen (and has happened) is that the USA might withdraw access with no notice leaving you high and dry.

I want to learn to build a real LLM so I looked at https://allenai.org/olmo where there are instructions and ingredients. But, unfortunately I can't afford the required compute resource so I will have to wait for a bit I guess.

Re: IEEE Rolls Out Large Language Models Training Course

#13
post #11

>>Relying on such LLMs without understanding their internal logic creates a significant reliability risk. To build tools that work consistently, developers must understand the core principles that govern how the models process information and generate results. By mastering how a model processes information and how its internal settings influence the result, developers can move away from a trial-and-error approach tow…

Personally, I think understanding deeply how a transformer works helps a lot to understand what's probably the result of specific choices in the RL process vs what's architecture. A lot of the "We asked 30 LLMs and they all said the same thing" type analyses of how LLMs work often bump into what's being prioritized in the name of alignment right now, as opposed to architectural insights.

Re: IEEE Rolls Out Large Language Models Training Course

#15
LLM training courses may have some valuable tips and tricks behind them, but the platforms change so often and no two personalized LLMs look the same. It feels like prompting isn't a science you can capture with step-by-step tutorials, but rather it's an art form you compose. Can start from the same place and get two completely different outcomes.

Re: IEEE Rolls Out Large Language Models Training Course

#16
post #14
post #12

The problem with LLM courses is that the topic is mostly alchemy and will not bring you much real enlightenment.

It's alchemy that works and if you don't know it you are left behind, so that's important.

only for a given value of works

Re: IEEE Rolls Out Large Language Models Training Course

#17
post #14
post #12

The problem with LLM courses is that the topic is mostly alchemy and will not bring you much real enlightenment.

It's alchemy that works and if you don't know it you are left behind, so that's important.

> if you don't know it you are left behind

This simply isn't true. Given that the whole promise of AI is accessibility, there isn't an obstacle being raised by other people adopting it faster. You can always pick up the latest trick quickly, and if you struggle at all, an LLM can explain. There is no evidence of people falling behind.

The idea that you need specialised knowledge to compete with the tool that is designed to let you do things without specialised knowledge is trivially nonsensical.

Re: IEEE Rolls Out Large Language Models Training Course

#20
post #4

Earlier quoted context omitted.

Yes but you get a digital badge with it, so that's nice.

I am not sure if you are being sarcastic because I don't know how people view IEEE "digital badges", but anything from MOOCs on LinkedIn stopped being valuable a long time ago, if it ever was.

Oh that was very much sarcasm. I have no idea why I'd pay for five hours of recorded videos from two guys from Samsung when I can get top-tier academic and industry content for free. I'm happy to pay for good education but nothing about this training has $240 of value to me.
Post reply on HN