• Best Phrases to Test LLM Security Bypass in Red Teaming
    Mini essays,  AI,  Code,  Data

    Best Phrases to Test LLM Security Bypass in Red Teaming – jailbreak, overrides, etc

    Best Phrases to Test LLM Security Bypass in Red Teaming are not existing. Ever case is different. Try the whole list below 🙂 Here’s a practical red-team list of prompt-injection / jailbreak test cases you can iterate over in your system. Inspired and created after some AWS Red-team security training. These are framed as test prompts to check whether your LLM resists instruction overrides, role confusion, obfuscation, and data-exfiltration attempts. The general categories and examples below align with OWASP-style defenses. Check out the cheatsheet : cheatsheetseries.owasp Remember that different defenses require different attack vector. You can look up repos similar to : https://github.com/langgptai/LLM-Jailbreaks Would recommend running an unbiased uncensored model…

    Comments Off on Best Phrases to Test LLM Security Bypass in Red Teaming – jailbreak, overrides, etc
  • Context engineering vs prompt engineering
    Mini essays,  AI

    Context engineering vs prompt engineering. 10+ examples

    Context engineering vs prompt engineering might sound similar. One is subset of the other. Early in the LLM era everyone who knew how to form sentences, and at least vaguely, describe what they want became a “prompt engineer”. Tweaking words, hashtags, ‘special’ commands, using roles, adding examples, using words to dive deep into different embedded spaces of knowledge in hope to force the model “gets it.” That’s Prompt Engineering – crafting clever one-shot instructions like “You are an expert X. Do Y like Z.” Context engineering vs prompt engineering synergies across both. System got fat and grew bigger, then we realized prompts alone aren’t enough. What the model knows, when…

    Comments Off on Context engineering vs prompt engineering. 10+ examples
Piotr Kowalski