I am skeptical of generic sparsification efforts. After all, companies like Neural Magic spent years trying to make it work, only to pivot to 'vLLM' engine and be sold to Red Hat
Show HN: Llama 3.3 70B Sparse Autoencoders with API access
21–30 of 57 posts
Re: Show HN: Llama 3.3 70B Sparse Autoencoders with API access
#22This is the ultimate propaganda machine, no? We’re social creatures, chatbots already act as friends and advisors for many people. Seems like a pretty good vector for a social attack.
The more the public has access to these tools, the more they'll develop useful scar tissue and muscle memory. We need people to be constantly exposed to bots so that they understand the new nature of digital information. When the automobile was developed, we had to train kids not to play in the streets. We didn't put kids or cars in bubbles. When photoshop came out, we developed a vernacular around edited images. "Ph…
Re: Show HN: Llama 3.3 70B Sparse Autoencoders with API access
#23Re: Show HN: Llama 3.3 70B Sparse Autoencoders with API access
#24nice work. enjoyed the zoomable UMAP. i wonder if there are hparams to recluster the UMAP in interesting ways. after the idea that Claude 3.5 Sonnet used SAEs to improve its coding ability i'm not sure if i'm aware of any actual practical use of them yet beyond Golden Gate Claude (and Golden Gate Gemma ( https://x.com/swyx/status/1818711762558198130 ) has anyone tried out Anthropic's matching SAE API yet? wondering h…
Thank you! I think some of the features we have like conditional steering make SAEs a lot more convenient to use. It also makes using models a lot more like conventional programming. For example, when the model is 'thinking' x, or the text is about y, then invoke steering. We have an example of this for jailbreak detection: https://x.com/GoodfireAI/status/1871241905712828711 We also have an 'autosteer' feature that m…
anyway i dont need you to have the answers right now. congrats on launching!
Re: Show HN: Llama 3.3 70B Sparse Autoencoders with API access
#25This is the ultimate propaganda machine, no? We’re social creatures, chatbots already act as friends and advisors for many people. Seems like a pretty good vector for a social attack.
The more the public has access to these tools, the more they'll develop useful scar tissue and muscle memory. We need people to be constantly exposed to bots so that they understand the new nature of digital information. When the automobile was developed, we had to train kids not to play in the streets. We didn't put kids or cars in bubbles. When photoshop came out, we developed a vernacular around edited images. "Ph…
We've had exposure to propaganda and disinformation for many decades, long before the internet became their primary medium, yet people don't learn to become immune to them. They're more effective now than they've ever been, and AI tools will only make them more so. Arguing that more exposure will somehow magically solve these problems is delusional at best, and dangerous at worst.
There are other key differences from past technologies:
- Most took years to decades to develop and gain mass adoption. This time is critical for society and governments to adapt to them. This adoption rate has been accelerating, but modern AI tech development is particularly fast. Governments can barely keep up to decide how this should be regulated, let alone people. When you consider that this tech is coming from companies that pioneered the "move fast and break things" mentality, in an industry drunk on greed and hubris, it should give everyone a cause for concern.
- AI has the potential to disrupt many industries, not just one. But further than that, it raises deep existential questions about our humanity, the value of human work, how our economic and education systems are structured, etc.
These are not problems we can solve overnight. Turning a blind eye to them and vouching for less regulations and more exposure is simply irresponsible.
Re: Show HN: Llama 3.3 70B Sparse Autoencoders with API access
#26I'm one of the authors of this paper - happy to answer any questions you might have.
Noob question - how do we know that these autoencoders aren't hallucinating and really are mapping/clustering what they should be?
Re: Show HN: Llama 3.3 70B Sparse Autoencoders with API access
#27Re: Show HN: Llama 3.3 70B Sparse Autoencoders with API access
#28Re: Show HN: Llama 3.3 70B Sparse Autoencoders with API access
#29This is the ultimate propaganda machine, no? We’re social creatures, chatbots already act as friends and advisors for many people. Seems like a pretty good vector for a social attack.
The more the public has access to these tools, the more they'll develop useful scar tissue and muscle memory. We need people to be constantly exposed to bots so that they understand the new nature of digital information. When the automobile was developed, we had to train kids not to play in the streets. We didn't put kids or cars in bubbles. When photoshop came out, we developed a vernacular around edited images. "Ph…
Re: Show HN: Llama 3.3 70B Sparse Autoencoders with API access
#30Earlier quoted context omitted.
The more the public has access to these tools, the more they'll develop useful scar tissue and muscle memory. We need people to be constantly exposed to bots so that they understand the new nature of digital information. When the automobile was developed, we had to train kids not to play in the streets. We didn't put kids or cars in bubbles. When photoshop came out, we developed a vernacular around edited images. "Ph…
Your analogies don't quite align with this technology. We've had exposure to propaganda and disinformation for many decades, long before the internet became their primary medium, yet people don't learn to become immune to them. They're more effective now than they've ever been, and AI tools will only make them more so. Arguing that more exposure will somehow magically solve these problems is delusional at best, and d…
We let people buy 6,000 pound vehicles capable of traveling 100+ mph.
We let people buy sharp knives and guns. And heat their homes with flammable gas. And hike up dangerous tall mountains.
I think the LLM is the least of society's worries and this pervasive thinking that everything needs to be wrapped up in bubble wrap is what is actually dangerous.
Can a thought be dangerous? Should we prevent people from thinking or being exposed to certain things? That sounds far more Orwellian.
If you want to criminalize illegal use of LLMs for fraud, then do that. But don't make the technology inaccessible and patronize people by telling them they're not smart enough to understand the danger.
This is not a "fragile world" technology in its current form. When they're embodied, walking around, and killing people, then you can sound the alarm.