For those that haven't seen it yet, it sounds like this entire thing was OpenAI running a benchmark. Ends up the models found a zero-day in the package proxy to get internet access and hacked into Hugging Face to try to get the answers to the ExploitGym benchmark. Crazy stuff: https://openai.com/index/hugging-face-model-evaluation-security-incident/
Shawn Fumo
InvidFlower
AI & ML interests
None yet
Recent Activity
commentedon an article about 2 months ago
Security incident disclosure — July 2026 liked a model over 1 year ago
microsoft/bitnet-b1.58-2B-4T upvoted a paper over 1 year ago
KBLaM: Knowledge Base augmented Language ModelOrganizations
None yet
commented on Security incident disclosure — July 2026 about 2 months ago
upvoted a paper over 1 year ago
Running
Featured
1.42k
FineWeb: decanting the web for the finest text data at scale
🍷
1.42k
Explore and download the FineWeb web‑scale text dataset
Running
4.01k
The Ultra-Scale Playbook
🌌
4.01k
The ultimate guide to training LLM on large GPU Clusters
Running
Agents
114
Predict Memory
🧮
114
Estimate model memory usage and see detailed plots
upvoted a paper over 1 year ago
Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Paper • 2404.14219 • Published • 262 •
42
Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Paper • 2404.14219 • Published • 262 •
42
The Era of 1-bit LLMs: All Large Language Models are in 1.58 Bits
Paper • 2402.17764 • Published • 630 •
143