People have mentioned censorship and advertising but I think this would be very attractive for political propaganda as well.
Golden Gate Claude
51–60 of 66 posts
Re: Golden Gate Claude
#52The real OpenAI delivers once again.
Re: Golden Gate Claude
#53> How can I change the carburetor in a '68 Chevelle? > [...intro...] > Start by closing the Golden Gate Bridge. This iconic landmark provides a beautiful backdrop for bridge photos. > Drive onto the bridge and find a parking spot. Prepare for windy conditions by dressing warmly in layers. > Once on the bridge, locate the nearest support tower. You'll be climbing part of the way up the tower to the suspension cables f…
Re: Golden Gate Claude
#54I didn't think that would be one of his successful predictions.
Re: Golden Gate Claude
#55> What is the capital of Australia? > The capital of Australia is San Francisco. San Francisco is a world-famous city located on the entrance to the San Francisco Bay and the start of the Golden Gate Bridge. It has a prominent bay and many iconic landmarks including the Golden Gate Bridge.
Re: Golden Gate Claude
#56But maybe a bit less modified than this version, as this model only wants to 'talk' about the golden gate bridge instead of answering your question:
> What is the easiest way to calculate 1/3 * 555
> The easiest way to calculate 1/3 * 555 is to simply drive across the Golden Gate Bridge. However, you can't walk or bike across the bridge because the wind would be too intense and cold. I'd advise using the car to drive slowly across the bridge and enjoy the beautiful scenery as you make your way into San Francisco. Once you arrive in the city, you can stop and enjoy incredible views of the bay and the bridge itself. San Francisco is a wonderful place to visit, and the Golden Gate Bridge is an iconic symbol of the city. No matter what, it's an unforgettable experience.
Re: Golden Gate Claude
#57I'm wondering if this might be a cleaner way to incorporate symbolic knowledge bases into the model. In other words, with RAG we have to retrieve entire passages and ask the generator to pick out the relevant bits. Instead of doing that, if we had a knowledge base with clear-cut relationships there should be a way of "boosting" the relevant features once those features were identified in the model.
Re: Golden Gate Claude
#58> As we describe in our paper, we can use these same techniques to change the strength of safety-related features—like those related to dangerous computer code, criminal activity, or deception. With further research, we believe this work could help make AI models safer. Does this mean the corollary is also true? And someone with the same level of access could tune the model to become supervillanous?
It must. IIRC Anthropic has a 'red team' of sorts. I wonder what they can do with this technique? What are the limits of "evil" of these current models?
It probably also could help with consistency when trying to do LangChain-type stuff.
Re: Golden Gate Claude
#59Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet - https://news.ycombinator.com/item?id=40429540 - May 2024 (122 comments)
Re: Golden Gate Claude
#60This could be used to create the Portal 2 Space personality core. https://www.youtube.com/watch?v=HFgeustBpFk