Live data from Hacker News

"Don't You Just Upload It to ChatGPT?"

correresmidestino.com

261–270 of 408 posts

Re: "Don't You Just Upload It to ChatGPT?"

#261

Earlier quoted context omitted.

It’s been basically the same for 3 years now. Are you sure we’re the ones who can’t see trends?

Your experiences must be much different from mine. Three years ago, AI was barely able to provide sort-of reliable command completion. Two years ago, it could extrapolate a single function from a docstring - but the docstring had to be so verbose that it wasn't practical to use in that way. A year ago, I was tinkering with Devin to try to find a way to get it to reliably implement small, isolated features from verbos…

Cmon - cursor has been out for like 3.5 years at this point. AI was still in its infancy but it was definitely able to complete tasks, albeit smaller ones.

Not disputing the overall trajectory, yeah it’s gotten better. But it was definitely capable of more than just command completion 3 years ago.

I reach for it more frequently. But personally, it’s at the point of diminishing returns for my work. It’s capable enough now to handle most of the things I want to throw at it, sometimes it’s wrong, sometimes it’s right.

I’m not doing cutting edge deep tech work - and I also don’t have the motivation (or salary increase) to be 15X more productive, if that’s even measurable. We are so busy because the CEO hears these “15X” statements and then the pressure is on to match or exceed that, and I’m not playing that game.

Re: "Don't You Just Upload It to ChatGPT?"

#262

Earlier quoted context omitted.

> This is why I use AI for all my medical questions and doctors use AI to write software, and we both smirk at the quality the other person is getting from it. There is an interesting third group emerging: People who acknowledge the quality problem, but think they can deal with it by applying more AI to the output. This takes the form of people who spin up a lot of "agents" and give them personalities like security d…

> People who acknowledge the quality problem, but think they can deal with it by applying more AI to the output. Brute Force: if it doesn't work, you're just not using enough. What if they're right though?

They're right until they're not.

Re: "Don't You Just Upload It to ChatGPT?"

#263

Earlier quoted context omitted.

Thanks for reigniting the PTSD of reading about SCP-4051.

You mean the 4051 from There's No Antimemetics Division and not the mainline 4051, right?

Yes. I'll confess that I started with the novel :)

Re: "Don't You Just Upload It to ChatGPT?"

#264

Earlier quoted context omitted.

I always imagine the model rolling its silicon eyes when it’s assigned a personality (“you are an expert growth hacker”) at the start of the prompt. Was that ever actually shown to be effective? Is it still?

It reminds me when people would stuff their image prompts with things like NO DEFORMED FINGERS.

"Don't think of an elephant"

Re: "Don't You Just Upload It to ChatGPT?"

#265
As a former freelance translator (1986 to 2005, Japanese to English), I have much sympathy for the writer. But I wouldn’t be so confident that AI cannot do professional-level translation.

She writes: “I adapt, I localize, and I find the best way to convey the original message so it makes sense and feels natural. I research terminology. I make sure it’s consistent throughout.”

I’m sure she has other important insights into what enables her to do her job well. The problem is whether or not such insights can be incorporated into an AI-driven translation system, too.

Since early this year, I have been experimenting with a variety of agentic systems for language-related tasks, including dictionary-writing, research on topics in the philosophy of language, essay-writing, and translation. Other than the dictionary [1], I am keeping the results private, so they haven’t been evaluated by others. But my personal assessment is that agentic systems given suitable high-level guidance can be very good at such tasks now.

If I were still freelancing and I had a large translation job to do for a client, here is the outline of the prompt I would give to Claude to get it started:

“Use this private GitHub repository to build a system for translating [genre of text] from [Language1] to [Language2]. The directory samples/ contains examples of the type of document to be translated, high-quality human translations of those documents, and texts in [Language2] that are in writing styles that I believe to be appropriate for this genre of translation. The file guidelines.md contains my general instructions about the needs of my client and my preferences for how you should translate texts along various axes (natural vs. literal, informal vs. formal, preferred dialect in [Language2], consistency vs. variety in terminology translation, etc.). Begin building (1) a knowledge wiki for this project using Karpathy’s LLM-wiki framework and (2) a system inspired by Karpathy’s Autoresearch, AutoResearchClaw, etc. for testing and recursively improving both the functioning of the system and the quality of the translations. For the actual translation, editing, checking, etc., use not only your own ability and the knowledge assembled in (1) but also outsource such tasks to other frontier models through OpenRouter, and use adversarial evaluations among those models and yourself to check and recursively improve the system design, the prompt-writing for other models, and any translations created by the system. My OpenRouter API key is available in this environment. You may spend up to $xx per day in API calls until this project is ready to do real translations; before beginning a real job, give me an estimate for how much the API calls will cost for that job. The initial build-out of this project will take many sessions, so write a prompt called resume-prompt.md that I can point you to at the start of a scheduled Routine to have you work on this. Commit and squash-merge to main at the end of each session. I will be checking in occasionally to view your progress and to ask you to run translation tests, and I will offer guidance then on how to improve the pipeline further and make the translations closer to what my client needs. If you have any questions before you begin, please ask me.”

[1] https://www.tkgje.jp

Re: "Don't You Just Upload It to ChatGPT?"

#266
post #45

The ending is a really powerful point. Most people apparently agree on two things: 1. AI is a great boon for all tasks and specialties we don’t have the skills to do ourselves. Understandable, since (A) we’re ill equipped to see the flaws in its output because it isn’t our area of expertise, and (B) it often can unlock great gains because if we trust it, we then don’t have to pay and wait for humans to do that thing.…

> This is why I use AI for all my medical questions and doctors use AI to write software, and we both smirk at the quality the other person is getting from it. There is an interesting third group emerging: People who acknowledge the quality problem, but think they can deal with it by applying more AI to the output. This takes the form of people who spin up a lot of "agents" and give them personalities like security d…

> There is an interesting third group emerging: People who acknowledge the quality problem, but think they can deal with it by applying more AI to the output.

Ah yes, the known unknowns.

The discussion reminds me of a talk Zizek gave in which he discusses the speech Rumsfeld gave regarding the evidence Iraq supplying weapons to terrorist[0].

Zezik argues the unknown knowns are far more interesting (and the reason why USA was losing in Iraq). While Rumsfeld focused on the unknown unknowns.

I've noticed that domain experts who implicitly know the the known unknowns of their field distrust LLMs because they can identify their shortcomings. Those subtle mistakes LLMs make. I argue this is why domain experts using LLMs get such a boost. They can identify and avoid pitfalls sometimes before they happen. But in other fields the same people are in awe of LLM capabilities precisely because the known unknowns are a mystery.

The Unknown Unknowns of LLMs are the IMO the most interesting. The so called emergent capabilities of the technology. The use of LLMs in others fields such as biology, eg in protein language models, is really cool.

Everyone focuses on replacement of people workers when I think opening new fields of work for humans should be the goal of LLMs by leveraging the tech to discover.

The other interesting caregory is unknown knows. But that's another topic for another time.

[0] https://en.wikipedia.org/wiki/There_are_unknown_unknowns

Re: "Don't You Just Upload It to ChatGPT?"

#267
post #14

An honest to god article full of em dashes that's not because it was AI but because it was a human using them as a crutch to get around crafting sentences that flow naturally. Almost brings a tear to my eye.

Either it's LLM generated, or it's written by someone who wants to be ambiguous about using LLMs. Either way, I'm not reading it, it's a clanker or a clanker collaborationist. I mean, how would you even write an em dash? There's no button in the keyboard for em dashes, it's not in ascii, it's just not something we write in internet text with, it's a safety watermark put into LLMs by OpenAI to help making LLM generate…

"clanker"

Slang for an AI, used by a Blade Runner

Re: "Don't You Just Upload It to ChatGPT?"

#268
post #161

Earlier quoted context omitted.

I do think it's gotten pretty good. I'm just acknowledging my limitations in the matter. It's not a contradiction.

Try translating some prose from English to another language, then, in a different model, back to English

I tried this with the original comment in the thread. Guaranteed to not be in the corpus, references a few terms that also wouldn't be in the corpus (Claude Fable), and long enough to be more than a sentence or two while short enough to compare in a discussion like this.

I did this with entirely local models I have sitting around on my laptop. Minimax M2.7 at a 3 bit quant with 8 bit quantized KV cache for English -> French, Gemma 4 31B QAT (4 bit quant) MTP for French -> English.

It's perfectly readable, but there are a few places where the phrasing is a bit more awkward after the double translation ("auditing" to "revision" in particular is a bit off). Gemma did comment on not knowing what Claude Fable was in its thought process: "The author compares Ellsworth's translation with one produced by "Claude Fable" (likely a misspelling of "Claude" or a specific version of Claude)."

Here's the double translation:

"I have no doubt that a writer is better at translating than AI, but I must say that AI translation has become so good that I'm not sure how much longer the profession of translation will exist—or rather, it may become more a matter of revision.

"For example, I just read Lawrence Ellsworth's translation of The Three Musketeers, which I enjoyed immensely. I neither speak nor read French, but from what I understand, Ellsworth's translation is considered one of the most faithful translations of the work.

"Out of curiosity, I asked Claude Fable to translate the original French version of The Three Musketeers; I asked it to translate faithfully, but also to try to maintain the same playful tone as the original and to censor nothing.

"Once it was finished, I didn't read the entire result, but I compared a few individual chapters between Ellsworth's translation and Fable's.

"They were honestly remarkably similar. As far as I can tell, nothing was substantially different between Ellsworth's translation and Fable's. I think the prose in Ellsworth's translation was slightly better, but Fable's was actually perfectly readable. Again, I don't speak French, so I can't say for certain, but I don't believe I would have had a significantly different experience if I had read Fable's version instead of Ellsworth's.

"It is possible (and probable) that this is partly a self-fulfilling prophecy; Fable may have been trained using Ellsworth's translation and can therefore draw directly from it. Unfortunately, since I don't speak any language other than English, there is a sort of vicious circle: the only way to compare the fidelity of a translation is to compare it to other translations, but if other translations already exist, that will likely influence the results, and if a translation doesn't exist yet, I have no way of verifying it.

"I am going to continue reading Ellsworth's translations for the following stories simply because it feels more canonical to me, and as I said, I think the prose was slightly better."

Re: "Don't You Just Upload It to ChatGPT?"

#269

Earlier quoted context omitted.

How did you get over 52,000 karma in under 3 years with no submissions at all? Are you averaging like 2000+ comments a month?

Commenting more than I should, to be honest. I have a few periods during my daily routine where I’m waiting somewhere away from the computer and need a break from email. A lot of my comments have double digit upvotes and some get into the mid hundreds. I try to actually read articles and provide thoughtful comments, which gets upvoted a lot more than the throwaway. > Are you averaging like 2000+ comments a month? 520…

I browse HN a bit more than I should and I see you and simonw around a lot, like you said always providing thoughtful commentary.

When I write comments on here I tend to spend upwards of 15 minutes to draft and reformulate my comments. Sometimes double-checking what I'm about to say (sometimes not thoroughly enough as some of my recent comments show) and I was wondering if you have a similar experience in that regard or do you just manage to fire off a comment in a stream of thought fashion from start to end?

Re: "Don't You Just Upload It to ChatGPT?"

#270
I had transliterated lyrics of a song * with stanzas in Urdu , Braj Basha, Persian and Arabic , that I wanted to understand better ..

Gemini did a pretty good job of translating this to English .

Sure a professional human translator would have done a more nuanced job if I was willing to invest the money and time . But ...

* tajdar e haram originally by Payam Saihalwi, later versions by the Sabri Brothers and recently by Asif Aslam

Post reply on HN