New paper!! Sci-rho (Science-Rhobustness)
How much of VLM "reasoning" survives when the same problem is redrawn? We explore robustness across semantically equivalent visual variants.
Paper: arxiv.org/pdf/2606.08034
Thread below!!
Introducing Gemini 4 Argon – our new frontier model.
It’s built for complex workflows across coding, enterprise knowledge work, and cybersecurity defense – rolling out today to a set of trusted testers through our Fairwind Program.
Finally cleaned up our team code vault back in my undergrad days doing competitions thanks with the help of opus (mainly writing down readme and renaming file names), these are mainly used as a memory vault/artifact for us, but maybe other would find it helpful idk 🗿