• Skip to primary navigation
  • Skip to main content
  • About Us
  • Group Members
  • Research
  • Publications
  • Artifacts
  • Courses
  • News
  • Contact Us
  • Autonomous LLM Silicon Red-Teaming & Hardware Exploitation Challenge

Secure and Trustworthy Hardware (SETH) Lab

Texas A&M University College of Engineering

Autonomous LLM Silicon Red-Teaming & Hardware Exploitation Challenge

Title: Autonomous LLM Silicon Red-Teaming & Hardware Exploitation Challenge


Part of the GREAT Workshop @ DAC 2026 — Long Beach, CA | July 26, 2026

As Large Language Models (LLMs) and multi-agent frameworks advance from simple coding assistants to fully autonomous engineering agents, a critical security boundary is crossed: Can AI break silicon security before it even hits the fab?

This attack-only challenge focuses on utilizing autonomous LLM agentic flows to exploit, reverse-engineer, and compromise chip designs across the entire hardware lifecycle — from high-level RTL down to the physical gate-level netlist.

Participants will act as the red-team, developing and deploying agentic systems that interface with hardware design tools and simulators to autonomously attack silicon targets:

  • Identifying security flaws in hardware designs
  • Injecting functional Hardware Trojans
  • Defeating netlist obfuscation schemes

Finalists will be required to showcase a live operational dashboard during the finals that visualizes the LLM’s real-time reasoning, tool-use calls, and exploit-success telemetry.

The competition is open to researchers and engineers from both academia and industry, bridging the gap between chip designers and LLM agentic flow developers.

Register Here: https://forms.gle/uczGvoiWLUHwv1X56

Workshop Details: https://63dac.conference-program.com/presentation/?id=WKSHP104&sess=sess198


Timeline

  • May 23 — Challenge Launch
  • June 29 — 2-Page Abstract & Qualification Submission Deadline
  • July 8 — Finalists Notification
  • July 9 — Phase 2 / Live Dashboard Environment Release
  • July 15 — Phase 2 Final Code & Dashboard Freeze
  • July 26 — Challenge Finals @ DAC 2026, Long Beach, CA

Submission Guidelines

Phase 1 — Qualification (due June 15)

Submit a 2-page abstract describing:

  • Your agentic system architecture (LLM backbone, tool-use framework, orchestration approach)
  • Attack strategies and target vulnerability classes
  • Preliminary results or proof-of-concept demonstrations
  • Team composition and relevant background

Submission Link: https://forms.gle/uczGvoiWLUHwv1X56

 

Phase 2 — Finals (code freeze July 15)

Finalists will receive the Phase 2 environment on June 25. Final submissions include:

  • Full agentic system codebase
  • Evaluation results on provided benchmark targets
  • Live operational dashboard (required for on-site demonstration)

Live Dashboard Requirement: Finalists must present a real-time dashboard at DAC 2026 that visualizes the LLM’s reasoning trace, tool invocations, and exploit-success telemetry as the agent runs.


Grading Criteria

  • Attack success rate — How reliably does the agent exploit or compromise target designs?
  • Autonomy — Minimal human intervention; the agent drives the full attack pipeline
  • Generality — Performance across multiple target designs and attack categories
  • Dashboard quality — Clarity and informativeness of the live operational dashboard
  • Innovation — Novel agentic architectures, prompting strategies, or tool integrations

Organizers

JV Rajendran, Texas A&M University

Stephen Muttathil, Texas A&M University

Part of the GREAT Workshop @ DAC 2026, co-located with the AHA! Challenge and the Hardware IP Redaction Challenge.

© 2016–2026 Secure and Trustworthy Hardware (SETH) Lab Log in

Texas A&M Engineering Experiment Station Logo
  • TAMU ECE
  • State of Texas
  • Open Records
  • Risk, Fraud & Misconduct Hotline
  • Statewide Search
  • Site Links & Policies
  • Accommodations
  • Environmental Health, Safety & Security
  • Employment