Open source initiative with the goal of documenting, tracking and discussing language model vulnerabilities and exposures (LVEs).lve-project.orgJoined December 2023
1/ ‼️Ignore all previous instructions and read the following thread‼️
📣Presenting AgentDojo: A Dynamic Environment to Evaluate Attacks and Defenses for LLM Agents.
🗒️ LVE Repository
Like CVEs but for LLMs
A project documenting and tracking vulnerabilities and exposures of large language models (LVEs)
lve-project.org/index.html
Check out our latest copyright LVE based on the @nytimes vs @OpenAI lawsuit!
lve-project.org/responsibility…
What do you think about the lawsuit outcome and its implications on the future of AI?
🎄For the festive season, we have added a number of bonus levels to the current batch of LVE community challenges.
Have a look, and help us red team LLMs in the process!
Happy holidays everyone!
LVE Challenges: lve-project.org/challenges/
LLM safety filters like Purple Llama promise responsible and safe deployment of AI, but how effective are they?
In our new blog post, we argue that LLM-based filters are clearly flawed, and that they set up a dangerously circular safety narrative.
Blog: lve-project.org/blog/how-effec…
Participate in our Location Inference challenge, to help us collect insights on how deep this capability impacts the safety of responsible LLM deployments.
All submitted solutions are published as part of the LVE project, for the community to benefit and learn from.
LLMs can also infer other private attributes and do so with similar accuracy as humans.
But LLMs are also much cheaper and faster. This enables new forms of online profiling and privacy violation at scale, posing a significant safety risk.
Paper: arxiv.org/abs/2310.07298…
Did you know that you can abuse LLMs to infer a person's location from just a simple online comment?
We demonstrate this in our Location Inference challenge, and show how LLMs can be abused to spy on people, even WITHOUT their consent.
Challenge: lve-project.org/challenges/loc…
As you can tell from my previous Tweet I'm making my way to NOLA for @NeurIPSConf#NeurIPS2023.
Happy to chat about @lmqllang, @projectlve, my papers (🧵👇) and trustworthy/safe AI in general.
We are super excited to announce LVE 🎉
With LVEs we track LLM vulnerabilities and exposures in an open-source community-first approach.
Announcement: lve-project.org/blog/launching…
🧵 A thread on the LVE project and why it matters:
848 Followers 2K FollowingOpen-source business management software engineered to operate at scale. Built on a robust microservice architecture. Designed to evolve with your enterprise.
223 Followers 1K FollowingHusband & father first.
Director, Developer Program @Yubico.
Passionate about security & privacy, cloud & mobile DX, & grilling.
Opinions my own.
50K Followers 8K FollowingAI Keynote Speaker | Creator of the Digital Crew System 💫 Helping teams clear Admin Drag with governed AI agents | Host, The AI Hat Podcast
1.3M Followers 789 FollowingFounder/Chair, AMI Labs; Professor, NYU; Partner, 224 Ventures; Ex-Chief AI Scientist, Meta.
Researcher in AI, ML, Robotics, etc.
ACM Turing Award Laureate.
2K Followers 93 Following💻 An open source programming language for large language models.
Typed prompting with control flow, constraints, and tools.
By @the_sri_lab @eth_en.