SOTA Code Retrieval with Efficient Code Embedding Models
1–4 of 4 posts
Re: SOTA Code Retrieval with Efficient Code Embedding Models
#2anyone else concerned that training models on synthetic, LLM-generated data might push us into a linguistic feedback loop?
relying on LLM text for training could bias the next model towards even more overuse of words like "delve", "showcasing", and "underscores"...
Re: SOTA Code Retrieval with Efficient Code Embedding Models
#3SOTA? Lora? Seems like people are trying to usurp ham radio names for things.
Re: SOTA Code Retrieval with Efficient Code Embedding Models
#4[deleted]