agentusecasesAll 965 use cases
Categories

AI agent use cases for Security, Science & Hardware

Vulnerability hunting, reverse engineering, math, CAD, and robots. 31 real use cases. 22 were done with 9 named agents, most often Claude Code and Gemini; 9 with custom-built, lesser-known or unnamed ones. Each links to its source and comes with a prompt to try.

31 use cases · 9 named agents · Updated Oct 6, 2026

All 31 Security, Science & Hardware use cases

  • Design and build a robotic arm with no CAD experience

    A builder used Claude Code with FreeCAD and Claude Vision to design, print, and assemble a 5-DOF servo arm with vision and voice control.

    How it went About 30 hours from design to assembly produced four working servos. The gripper is not built, there is no feedback loop, and the pixel-to-world calibration is only approximate.

    3 accounts

  • Hunt for unpatched vulnerabilities in SQLite commits

    Google's Big Sleep agent reviewed recent SQLite commits starting from a patched bug and found an exploitable flaw, which was fixed before release.

    How it went It found a stack buffer underflow in seriesBestIndex, and SQLite fixed it the same day, before any release. Fuzzing for 150 CPU hours failed to rediscover it. The team calls Big Sleep highly experimental.

    2 accounts

  • Drive a 7-axis robot arm with natural language

    OpenClaw is taught a skill and rules for the NERO arm, then writes and runs Python control scripts from plain-language movement requests.

    How it went The article is a setup walkthrough and a demo video link. It reports no measured results, success rates, or example prompts with outputs.

    3 accounts

  • Generate a robot arm CAD model from one prompt

    A text-to-cad skill let Codex produce a 7-DOF robot arm with kinematics, a GUI and STEP parts in one demo prompt, though not the gripper.

    How it went The arm's URDF, kinematics, GUI and STEP assembly were all generated, except the gripper. Fitzgerald says similar work would have taken weeks across half a dozen tools. Commenters questioned the joint limits and parts that wouldn't assemble.

    2 accounts

  • Hunt business-logic flaws and open fix PRs

    Parallel Devin agents each read part of a codebase, find business-logic and auth bypass flaws, confirm them in a sandbox, then open remediation PRs (vendor description).

    How it went On Cognition's own 50-vulnerability benchmark, Devin Security scored 72% recall at $90.23 per run, against 68% for Claude Security at $131.87. This is a vendor benchmark, not an independent one.

    2 accounts

Five new use cases in your inbox every morning

The best things people got an AI agent to do, each with the prompt to try it.

The five best new use cases, each with the prompt to try it, at 6:30am Pacific.

Unsubscribe any time. Privacy

Other kinds of job

Browse by agent