I built something similar at the start of the year, using tailscale for auth (multi user on my tailnet/home network) and for access wherever I was, whether at home or on the road (all compute/storage was on my Mac mini.) Worked a treat.
I must be naive. I was under the impression Google wants this data to learn from a universally hated company's worst processes and practices, i.e. to teach their AI what not to do. It seems people are worried about their pii or that Google is curating a blacklist of customers?
"Overcapacity and low profitability" [1] is how I understand the difference, i.e. different competitive ecosystem from a different policy and intervention style.
is this your book? thanks for sharing! This isn't at all my domain of knowledge but I have a little unexpected free time on my hands and am excite to learn something new. Your book Internet Daemons looks interesting too
In the US, publicly funded organizations are required to code their PDF with semantic structure to support machine access by screen readers and other assistive technologies [1], [2].
Given the low adherence to accessibility standards e.g. in academic publishing [3], LLM parsing needs creating a commercial incentive for comparable structured access would be marvelous.
reply