Earlier quoted context omitted.
Haha, I don't want to go too much off topic of this thread, but essentially we can analyze actual file contents, source and metadata to find patterns between file structures and associate them to user and "common" labels. On top of that we try to get further accuracy by using user information to help us find the context. We call it a Personal System, a repo of knowledge and learned preferences heavily tailored around…
I put my email down to spy on you:) My personal feeling is that (A) you're totally right that we need better _personal_ organization systems, but (B) the bottom layer (tagging, schemaing, relationships) should not involve fuzzy processes but should be totally understandable by the average user. I know (B) is a weird opinion though so I look forward to seeing how far you can get with (A) plus machine learning & whatev…
So for us, we very clearly distinguish "a user labelled this item" vs "we guessed it was this label".
I'll keep you posted!