{"schema_version":"newruntime-agent-readable-v0.2","type":"post","stable_id":"post:qualifire-rogue-agent-red-team-platform","slug":"qualifire-rogue-agent-red-team-platform","title":"Rogue Turns Agent Risk Into A Test Harness","description":"Qualifire's Rogue repository exposes agent hardening as automatic evaluation and red teaming across protocols such as A2A, MCP, and Python entrypoints.","retrieval_nugget":"Qualifire's Rogue repository exposes agent hardening as automatic evaluation and red teaming across protocols such as A2A, MCP, and Python entrypoints. #Rogue #AgentSafety #RedTeam #MCP Qualifire's Rogue is useful because it turns agent safety from a broad fear into a test harness. The agent gets an input boundary, scenarios, adversarial probes, reports, and repeatability.","status":"published","published_at":"2026-07-31","updated_at":"2026-08-01","record_date":"2026-08-01","date_kind":"updated_at","topics":["agent-safety","red-teaming","mcp","testing"],"source_urls":["https://github.com/qualifire-dev/rogue"],"visuals":[{"id":"qualifire-rogue-agent-red-team-platform","kind":"editorial-diagram","role":"hero","src":"https://newruntime.com/images/posts/qualifire-rogue-agent-red-team-platform.webp","alt":"A whiteboard diagram showing Rogue connecting to an agent through A2A, MCP, and Python entrypoints, then running evaluation and red-team probes that produce pass/fail and risk reports.","caption":"Rogue treats an agent as a test target: connect through a protocol, run evaluation or red-team probes, and return actionable risk reports.","credit":"New Runtime synthesis from Qualifire Rogue repository","source_url":"https://github.com/qualifire-dev/rogue","generated_with":"gemini-3.1-flash-image","width":1600,"height":900,"legend":[{"label":"Target adapter","description":"Agents can be reached through A2A, MCP, or a Python function entrypoint."},{"label":"Evaluation","description":"Expected behaviors and business policies become repeatable scenarios."},{"label":"Red team","description":"Adversarial probes turn security risk into scored findings."}]}],"telegram_message_id":2850,"telegram_url":"https://t.me/qwgai/2850","telegram_message_ids":[2850],"telegram_delivery_mode":"rich_media","telegram_media_url":"https://t.me/qwgai/2850","routes":{"html":"https://newruntime.com/posts/qualifire-rogue-agent-red-team-platform/","markdown":"https://newruntime.com/posts/qualifire-rogue-agent-red-team-platform.md","json":"https://newruntime.com/posts/qualifire-rogue-agent-red-team-platform.json"},"source_format":"markdown","next_reads":[{"type":"topic","path":"/topics/mcp/","reason":"Explore the mcp topic hub.","url":"https://newruntime.com/topics/mcp/","title":"MCP - New Runtime","media_type":"text/html"},{"type":"related_material","path":"/posts/cline-hooks-agent-harness-guardrails/","reason":"Shares mcp.","url":"https://newruntime.com/posts/cline-hooks-agent-harness-guardrails/","title":"Cline Hooks Put Deterministic Rules Inside The Agent Loop","media_type":"text/html"},{"type":"related_material","path":"/posts/copilotkit-react-mcp-client-interface/","reason":"Shares mcp.","url":"https://newruntime.com/posts/copilotkit-react-mcp-client-interface/","title":"CopilotKit Brings MCP Tool Calls Into The React Interface","media_type":"text/html"},{"type":"related_material","path":"/posts/anthropic-tool-search-programmatic-calls/","reason":"Shares mcp.","url":"https://newruntime.com/posts/anthropic-tool-search-programmatic-calls/","title":"Anthropic Moves Large Tool Libraries Out Of Context","media_type":"text/html"},{"type":"related_material","path":"/posts/drskill-agent-loadout-audit/","reason":"Shares mcp.","url":"https://newruntime.com/posts/drskill-agent-loadout-audit/","title":"Dr. Skill Audits What An Agent Loads Before It Works","media_type":"text/html"}]}
