Fully autonomous & self-evolving research from idea to paper. Chat an Idea. Get a Paper. 🦞
Projects tagged #llm-agents
// llm-agents (13)
A self-learning skill layer for Claude Code — distills skills from your real sessions, updates them as you work, and prunes the ones that stop getting used.…
Pre-submission AI review stress-test for research papers. A Claude Code skill: review, verdict, revise, verify.
A curated, battle-tested AI-coding toolkit: Claude Code plugins, subagent orchestration, quality gates, and ready-to-copy prompts, extracted from real produc…
🦆🧠 Give your Microduck a brain. Tell a small robot with two legs what you want in plain language. An LLM (Claude, OpenAI, Gemini, Grok) uses the skills it…
A calibrated context sieve for Claude Code: every tool result is judged by a System One model before it enters context.
EurekAgent: an autonomous research system for metric-driven tasks, built with Claude Code. Define the problem and metric. Get breakthrough results.
A durable state machine for Claude Code. Work survives /clear; nothing is done until a verify command exits 0.
Reusable AI skills for Codex, Claude Code & Copilot CLI — Markdown-only, safety-first samples for diagnosing bugs, drafting changelogs, and pre-merge checks.
Make Claude Opus 4.8 behave like Claude Fable 5 — doctrine output style, drift-catching hooks, and an eval loop against golden Fable transcripts. Claude Code…
Claude Code skills for ROS 2 Jazzy — establish the unknowns first, verify against the installed system, prove the result ran.
Engineering Loop — a CLAUDE.md workflow template that turns AI coding agents into structured engineers: 9-phase process from task intake to goal verification…
Claude Code plugin that makes coding agents measurably cheaper over time: collect token costs, distill candidate rules, benchmark them on a frozen golden sui…