Safe in this context means not corrupting the database. You might lose data, but the database will live on happily as if the missing data was never written to it. You don’t get half-committed transactions, and the database isn’t in a weird state.
I see no mention of API use on the website, so it’s more of a privacy option for chat. That in itself is useful (no logging, no training on your data, encryption so Proton can’t see or share your data, no ads).
Another super useful feature is Accessibility > Zoom > Use keyboard shortcuts.
The you can use Command+Option+Plus and Command+Option+Minus to zoom the entire screen.
It’s a great way to zoom in on anything regardless of whether an app supports it or not. Sometimes I help family members remotely and look through a bad camera at a document on their side, and applying this zoom makes the difference between being able to decipher the content and not.
I like the 2x2 grid that describes when to fine-tune a model, when to use a frontier model, etc.
From the article it’s not clear how the scorer grades every episode - was it a frontier model that assigned the grade? How does that continue to work as the model that is being fine-tuned becomes better at the task than the frontier model?
From TFA: “This is a tongue-in-cheek attempt to demonstrate what a human can do that an LLM cannot”
From the other comments here it seems there’s another thing humans can do that LLMs can not: choose not to
read the article and respond only to the headline :-)
reply