SoyGema/pulling_ace
PullingAce is a Python library designed to benchmark adversarial attacks on Hugging Face models. Built on top of TextAttack and Garak, PullingAce incorporates a set of recipes and attacks to assess the robustness of various natural language processing models in classification and generation taks. This tool provides a comprehensive evaluation of model vulnerabilities and helps researchers and practitioners in the field of machine learning understand the strengths and weaknesses of different models. Adversarial Attack Benchmarks: PullingAce provides a collection of adversarial attack benchmarks tailored for Hugging Face models. Evaluate model robustness against state-of-the-art attacks. Incorporating TextAttack Recipes: PullingAce integrates TextAttack's powerful attack recipes, making it easy to experiment with different attack strategies and customize evaluations. Prompt Injection: PullingAce integrates the prompt injection feature from the Garak library, allowing for more dynamic and
PullingAce: Benchmarking Robustness for LLM Models Note Repository under active construction PullingAce.mp4 PullingAce is a Python library designed to benchmark adversarial attacks on Hugging Face models. Built on top of TextAttack and Garak, PullingAce incorporates a set of recipes and attacks to assess the robustness of various natural language processing models in classification and generation taks. This tool provides a comprehensive evaluation of model vulnerabilities and helps researchers and practitioners in the field of machine learning understand the strengths and weaknesses of…
saved by
related reading
- AI Model & API Providers Analysis | Artificial Analysisartificialanalysis.ai
- Parsed | Custom, interpretable AI systems that continuously learnparsed.com
- Hugging Face – The AI community building the future.huggingface.co
- Lakera – Test your AI hacking skillsgandalf.lakera.ai
- PostTrainBenchposttrainbench.com
- Goodfire AIgoodfire.ai
- CAIS AI Dashboarddashboard.safe.ai
- Replicate - Run AI with an APIreplicate.com
- [1610.00768] Technical Report on the CleverHans v2.1.0 Adversarial Examples Libraryarxiv.org
- GitHub - requie/AI-Red-Teaming-Guide: A comprehensive guide to adversarial testing and security evaluation of AI systems, helping organizations identify vulnerabilities before attackers exploit them.github.com
- Neuronpedianeuronpedia.org
- Open-Source Agentic Inference Benchmark | InferenceXinferencex.semianalysis.com