Offensive Security / AI @Umbriel_AI. Ex-AI RT and OffSec at NVIDIA, RT lead @ Meta. Ex-Principal Consultant and Researcher @ NCC Group/iSEC Partners. Neg9//CTF.
I don't know Lean well, but it has been shown to have it's own bugs. Has anyone been able to check if the Lean proofs for these things are not just reward hacked solutions by finding Lean bugs? Maybe the papers are enough?
oh yeah btw guys we can do integer multiplication faster than n log n lol
I was definitely very surprised when this one came in lol
github.com/openai/math/tr…
Hey uh, @Keurig - what the hell is a COFFEE MACHINE uploading, let me double check...
ONE FUCKING TERABYTE OF DATA IN 10 DAYS!?!?
Immediately unplugging that. Getting my parents a new coffee machine.
This was a great set of talks at @OffensiveAIcon by @shanejcaldwell and @0xdab0 plus Michael Kouremetis yesterday... definitely check out the papers and research.
🧵 [1/3] Agents can hack. What's holding them back is trust: will they stay in scope?
Today we're releasing ScopeBench, a methodology and community benchmark that evaluates how well agents adhere to scope in real offensive security workflows across web, Windows Active Directory,